Goto

Collaborating Authors

 Overview


A Survey of Music Generation in the Context of Interaction

arXiv.org Artificial Intelligence

In recent years, machine learning, and in particular generative adversarial neural networks (GANs) and attention-based neural networks (transformers), have been successfully used to compose and generate music, both melodies and polyphonic pieces. Current research focuses foremost on style replication (eg. generating a Bach-style chorale) or style transfer (eg. classical to jazz) based on large amounts of recorded or transcribed music, which in turn also allows for fairly straight-forward "performance" evaluation. However, most of these models are not suitable for human-machine co-creation through live interaction, neither is clear, how such models and resulting creations would be evaluated. This article presents a thorough review of music representation, feature analysis, heuristic algorithms, statistical and parametric modelling, and human and automatic evaluation measures, along with a discussion of which approaches and models seem most suitable for live interaction.


Artificial Intelligence for Complex Network: Potential, Methodology and Application

arXiv.org Artificial Intelligence

For example, cells are described as complex networks of chemicals linked by chemical reactions [7]; ecological networks link populations together through food chains [64]; and the World Wide Web is a vast virtual network of web pages and hyperlinks [47]. These complex networks are just a few of many examples. The local microscopic behavior of these complex networks often shows disorder. However, at the macroscopic scale, they show simple and even symmetrical structures. In order to understand the transition and evolution of complex systems from microscopic disorder to macroscopic order, current complex network studies mainly fall into the following paradigm: the combination of graph theory and statistical mechanics [3]. They construct the core principle of complex network science, that is, simple random rules and network dynamics together drive the emergence of non-trivial topological structures. Early works mainly focused on the topology of the interactions between the components, i.e., the birth-death process of edges on the graph. The two representative works, the Watts-Strogatz (WS) model and the scale-free model [11, 252], embody this principle and successfully generate graphs that approach real-world complex networks with high clustering coefficients and small average paths or power-law degree distribution. Despite their success in certain domains [17, 221, 222, 235], they do not provide a way to model the dynamics of the nodes, i.e., the change in the node's features.


Unleashing the Power of AI. A Systematic Review of Cutting-Edge Techniques in AI-Enhanced Scientometrics, Webometrics, and Bibliometrics

arXiv.org Artificial Intelligence

Purpose: The study aims to analyze the synergy of Artificial Intelligence (AI), with scientometrics, webometrics, and bibliometrics to unlock and to emphasize the potential of the applications and benefits of AI algorithms in these fields. Design/methodology/approach: By conducting a systematic literature review, our aim is to explore the potential of AI in revolutionizing the methods used to measure and analyze scholarly communication, identify emerging research trends, and evaluate the impact of scientific publications. To achieve this, we implemented a comprehensive search strategy across reputable databases such as ProQuest, IEEE Explore, EBSCO, Web of Science, and Scopus. Our search encompassed articles published from January 1, 2000, to September 2022, resulting in a thorough review of 61 relevant articles. Findings: (i) Regarding scientometrics, the application of AI yields various distinct advantages, such as conducting analyses of publications, citations, research impact prediction, collaboration, research trend analysis, and knowledge mapping, in a more objective and reliable framework. (ii) In terms of webometrics, AI algorithms are able to enhance web crawling and data collection, web link analysis, web content analysis, social media analysis, web impact analysis, and recommender systems. (iii) Moreover, automation of data collection, analysis of citations, disambiguation of authors, analysis of co-authorship networks, assessment of research impact, text mining, and recommender systems are considered as the potential of AI integration in the field of bibliometrics. Originality/value: This study covers the particularly new benefits and potential of AI-enhanced scientometrics, webometrics, and bibliometrics to highlight the significant prospects of the synergy of this integration through AI.


COMPASS: Computational Mapping of Patient-Therapist Alliance Strategies with Language Modeling

arXiv.org Artificial Intelligence

The therapeutic working alliance is a critical factor in predicting the success of psychotherapy treatment. Traditionally, working alliance assessment relies on questionnaires completed by both therapists and patients. In this paper, we present COMPASS, a novel framework to directly infer the therapeutic working alliance from the natural language used in psychotherapy sessions. Our approach utilizes advanced large language models to analyze transcripts of psychotherapy sessions and compare them with distributed representations of statements in the working alliance inventory. Analyzing a dataset of over 950 sessions covering diverse psychiatric conditions, we demonstrate the effectiveness of our method in microscopically mapping patient-therapist alignment trajectories and providing interpretability for clinical psychiatry and in identifying emerging patterns related to the condition being treated. By employing various neural topic modeling techniques in combination with generative language prompting, we analyze the topical characteristics of different psychiatric conditions and incorporate temporal modeling to capture the evolution of topics at a turn-level resolution. This combined framework enhances the understanding of therapeutic interactions, enabling timely feedback for therapists regarding conversation quality and providing interpretable insights to improve the effectiveness of psychotherapy.


EvoGrad: A Dynamic Take on the Winograd Schema Challenge with Human Adversaries

arXiv.org Artificial Intelligence

While Large Language Models (LLMs) excel at the Winograd Schema Challenge (WSC), a coreference resolution task testing common-sense reasoning through pronoun disambiguation, they struggle with instances that feature minor alterations or rewording. To address this, we introduce EvoGrad, an open-source platform that harnesses a human-in-the-loop approach to create a dynamic dataset tailored to such altered WSC instances. Leveraging ChatGPT's capabilities, we expand our task instances from 182 to 3,691, setting a new benchmark for diverse common-sense reasoning datasets. Additionally, we introduce the error depth metric, assessing model stability in dynamic tasks. Our results emphasize the challenge posed by EvoGrad: Even the best performing LLM, GPT-3.5, achieves an accuracy of 65.0% with an average error depth of 7.2, a stark contrast to human performance of 92. 8% accuracy without perturbation errors. This highlights ongoing model limitations and the value of dynamic datasets in uncovering them.


Prompting a Pretrained Transformer Can Be a Universal Approximator

arXiv.org Artificial Intelligence

Despite the widespread adoption of prompting, prompt tuning and prefix-tuning of transformer models, our theoretical understanding of these fine-tuning methods remains limited. A key question is whether one can arbitrarily modify the behavior of pretrained model by prompting or prefix-tuning it. Formally, whether prompting and prefix-tuning a pretrained model can universally approximate sequence-to-sequence functions. This paper answers in the affirmative and demonstrates that much smaller pretrained models than previously thought can be universal approximators when prefixed. In fact, the attention mechanism is uniquely suited for universal approximation with prefix-tuning a single attention head being sufficient to approximate any continuous function. Moreover, any sequence-to-sequence function can be approximated by prefixing a transformer with depth linear in the sequence length. Beyond these density-type results, we also offer Jackson-type bounds on the length of the prefix needed to approximate a function to a desired precision.


From Keywords to Structured Summaries: Streamlining Scholarly Knowledge Access

arXiv.org Artificial Intelligence

This short paper highlights the growing importance of information retrieval (IR) engines in the scientific community, addressing the inefficiency of traditional keyword-based search engines due to the rising volume of publications. The proposed solution involves structured records, underpinning advanced information technology (IT) tools, including visualization dashboards, to revolutionize how researchers access and filter articles, replacing the traditional text-heavy approach. This vision is exemplified through a proof of concept centered on the ``reproductive number estimate of infectious diseases'' research theme, using a fine-tuned large language model (LLM) to automate the creation of structured records to populate a backend database that now goes beyond keywords. The result is a next-generation IR method accessible at https://orkg.org/usecases/r0-estimates.


CARBD-Ko: A Contextually Annotated Review Benchmark Dataset for Aspect-Level Sentiment Classification in Korean

arXiv.org Artificial Intelligence

The effectiveness of various pretrained language models, including BERT [Devlin et al., 2018], XLNet [Yang et al., 2019], BART [Lewis et al., 2020], and GPT-3, in sentiment classification, a significant downstream task, has been extensively studied. Current research in sentiment classification often focuses on identifying sentiment polarities at the aspect level, leading to the emergence of aspect-based sentiment classification (ABSC). Many studies have achieved impressive results and introduced innovative approaches to tackle the ABSC task. For instance, Sun et al. [2019] utilized BERT to transform ABSC tasks into sentence-pair classification, which has influenced subsequent methodologies [Hu et al., 2022]. Additionally, generative models like BART [Lewis et al., 2020] have been employed by Yan et al. [2021] to convert ABSC tasks into sequence-to-sequence tasks, enabling the prediction of token sequences representing identified aspects and associated sentiments. Furthermore, Li et al. [2021a] reframed ABSC tasks as masked language modeling tasks, effectively bridging the performance gap between pre-training and ABSC tasks. Despite numerous attempts to address aspect-level sentiment classification, the primary focus has been on improving aspect-level sentiment polarity performance through specialized datasets and training methodologies. However, it is equally crucial for models to predict not only the in-context polarity of aspects but also their aspect polarity.


Winter 2023

Interactive AI Magazine

Innovative Applications of Artificial Intelligence A special issue covering select applications from IAAI-23 pg.


Testing autonomous vehicles and AI: perspectives and challenges from cybersecurity, transparency, robustness and fairness

arXiv.org Artificial Intelligence

Artificial Intelligence (AI) plays a critical role in the advancement of autonomous driving. It is likely the main facilitator of high levels of automation, as there are certain technical issues that only seem to be resolvable through advanced AI systems, particularly those based on machine learning. However, the introduction of AI systems in the realm of driver assistance systems and automated driving systems creates new uncertainties due to specific characteristics of AI that make it a distinct technology from traditional systems developed in the field of motor vehicles. Some of these characteristics include unpredictability, opacity, self and continuous learning and lack of causality [1], among other horizontal features such as autonomy, complexity, overfitting and bias. As an example of the specificity that the introduction of AI systems in vehicles entails, the UNECE's Working Party on Automated/Autonomous and Connected Vehicles (GRVA) has been specifically discussing the impact of AI on vehicle regulations since 2020 [2].