Goto

Collaborating Authors

 Media


Detecting Musical Deepfakes

arXiv.org Artificial Intelligence

Ab s tract -- The proliferation of Text - to - Music (TTM) platforms has democratized music creation, letting users effortlessly generat e high - quality compositions . However, this innovation has also introduced challenges to musicians and the music in dustry . T his research focuses on utilizing the FakeMusicCaps dataset to address the challenge of detecting AI - generated songs by classifying the audio as deepfake or human. To simulate a real - world adversarial entity tempo stretching and pitch shifting modifications were applied to the dataset . Mel Spectrograms were generated from the resulting datasets, w hich were then used to train and test a convolutional neural network. This paper also explores the ethical and societal implications of TTM platforms, suggesting that detection systems developed and employed with care are a necessary tool to safeguard musicians and foster the positive potential of TTM plat forms and gen erative AI in music . Rapid a dvances in g e nerative AI have caused the creat ive landscape to be u pended, enabling almost anyone to easily create music that can be hard to distinguish from human - ma de compositions . AI - generated music is part of a wider classification of AI - generated media and art that falls unde r the category of " deepfake " .


Social media giant hit with scathing ad campaign amid anger over AI chatbots sexually exploiting kids

FOX News

A nonprofit parents coalition is calling on multiple congressional committees to launch an investigation into Meta for prioritizing engagement metrics that put children's safety at risk. The call is part of a three-pronged attack campaign by the American Parents Coalition (APC), launched Thursday. It includes a letter to lawmakers with calls for investigations, a new parental notification system to help parents stay informed on issues impacting their kids at Meta and beyond, and mobile billboards at Meta D.C. and California headquarters, calling out the company for failure to adequately prioritize protecting children. APC's campaign follows an April Wall Street Journal report that included an investigation looking into how the company's metrics focus has led to potential harms for children. "This is not the first time Meta has been caught making tech available to kids that exposes them to inappropriate content," APC Executive Director Alleigh Marre said. "Parents across America should be extremely wary of their children's online activity, especially when it involves emerging technology like AI digital companions.


Tests as Prompt: A Test-Driven-Development Benchmark for LLM Code Generation

arXiv.org Artificial Intelligence

We introduce WebApp1K, a novel benchmark for evaluating large language models (LLMs) in test-driven development (TDD) tasks, where test cases serve as both prompt and verification for code generation. Unlike traditional approaches relying on natural language prompts, our benchmark emphasizes the ability of LLMs to interpret and implement functionality directly from test cases, reflecting real-world software development practices. Comprising 1000 diverse challenges across 20 application domains, the benchmark evaluates LLMs on their ability to generate compact, functional code under the constraints of context length and multi-feature complexity. Our findings highlight instruction following and in-context learning as critical capabilities for TDD success, surpassing the importance of general coding proficiency or pretraining knowledge. Through comprehensive evaluation of 19 frontier models, we reveal performance bottlenecks, such as instruction loss in long prompts, and provide a detailed error analysis spanning multiple root causes. This work underscores the practical value of TDD-specific benchmarks and lays the foundation for advancing LLM capabilities in rigorous, application-driven coding scenarios.


Deconstructing Jazz Piano Style Using Machine Learning

arXiv.org Artificial Intelligence

For a visual artist, their style might include aspects such as subject choice, colour choice, and brush techniques; for a writer, it might include vocabulary, syntactic constructions, and narrative archetypes; for a composer, it might include harmonic progressions, rhythmic patterns, and melodic motifs. Individual differences across all these parameters, and more, come together to define each artist's unique style. Most of these stylistic parameters can theoretically be assessed by human experts. However, such assessments are necessarily slow and hence hard to apply at scale. Subjectivity is also a problem, since every human analyst comes with their own history of artistic exposure that will inevitably affect how they interpret artworks. Computational methods promise a more scalable and objective approach to this problem. Once a researcher has crafted an algorithm that captures a particular stylistic parameter -- for example, using entropy to capture vocabulary complexity -- then a computer can easily apply the algorithm to large datasets, and hence compare different artists using this parameter (Abry et al., 2013; Cheston et al., 2024b; Deepaisarn et al., 2023; Li et al., 2012).


A Retrieval-Augmented Generation Framework for Academic Literature Navigation in Data Science

arXiv.org Artificial Intelligence

In the rapidly evolving field of data science, efficiently navigating the expansive body of academic literature is crucial for informed decision-making and innovation. This paper presents an enhanced Retrieval-Augmented Generation (RAG) application, an artificial intelligence (AI)-based system designed to assist data scientists in accessing precise and contextually relevant academic resources. The AI-powered application integrates advanced techniques, including the GeneRation Of BIbliographic Data (GROBID) technique for extracting bibliographic information, fine-tuned embedding models, semantic chunking, and an abstract-first retrieval method, to significantly improve the relevance and accuracy of the retrieved information. This implementation of AI specifically addresses the challenge of academic literature navigation. A comprehensive evaluation using the Retrieval-Augmented Generation Assessment System (RAGAS) framework demonstrates substantial improvements in key metrics, particularly Context Relevance, underscoring the system's effectiveness in reducing information overload and enhancing decision-making processes. Our findings highlight the potential of this enhanced Retrieval-Augmented Generation system to transform academic exploration within data science, ultimately advancing the workflow of research and innovation in the field.


Marigold: Affordable Adaptation of Diffusion-Based Image Generators for Image Analysis

arXiv.org Artificial Intelligence

The success of deep learning in computer vision over the past decade has hinged on large labeled datasets and strong pretrained models. In data-scarce settings, the quality of these pretrained models becomes crucial for effective transfer learning. Image classification and self-supervised learning have traditionally been the primary methods for pretraining CNNs and transformer-based architectures. Recently, the rise of text-to-image generative models, particularly those using denoising diffusion in a latent space, has introduced a new class of foundational models trained on massive, captioned image datasets. These models' ability to generate realistic images of unseen content suggests they possess a deep understanding of the visual world. In this work, we present Marigold, a family of conditional generative models and a fine-tuning protocol that extracts the knowledge from pretrained latent diffusion models like Stable Diffusion and adapts them for dense image analysis tasks, including monocular depth estimation, surface normals prediction, and intrinsic decomposition. Marigold requires minimal modification of the pre-trained latent diffusion model's architecture, trains with small synthetic datasets on a single GPU over a few days, and demonstrates state-of-the-art zero-shot generalization. Project page: https://marigoldcomputervision.github.io


Ethical Aspects of the Use of Social Robots in Elderly Care -- A Systematic Qualitative Review

arXiv.org Artificial Intelligence

Background: The use of social robotics in elderly care is increasingly discussed as one way of meeting emerging care needs due to scarce resources. While many potential benefits are associated with robotic care technologies, there is a variety of ethical challenges. To support steps towards a responsible implementation and use, this review develops an overview on ethical aspects of the use of social robots in elderly care from a decision-makers' perspective. Methods: Electronic databases were queried using a comprehensive search strategy based on the key concepts of "ethical aspects", "social robotics" and "elderly care". Abstract and title screening was conducted by two authors independently. Full-text screening was conducted by one author following a joint consolidation phase. Data was extracted using MAXQDA24 by one author, based on a consolidated coding framework. Analysis was performed through modified qualitative content analysis. Results: A total of 1,518 publications were screened, and 248 publications were included. We have organized our analysis in a scheme of ethical hazards, ethical opportunities and unsettled questions, identifying at least 60 broad ethical aspects affecting three different stakeholder groups. While some ethical issues are well-known and broadly discussed our analysis shows a plethora of potentially relevant aspects, often only marginally recognized, that are worthy of consideration from a practical perspective. Discussion: The findings highlight the need for a contextual and detailed evaluation of implementation scenarios. To make use of the vast knowledge of the ethical discourse, we hypothesize that decision-makers need to understand the specific nature of this discourse to be able to engage in careful ethical deliberation.


DPN-GAN: Inducing Periodic Activations in Generative Adversarial Networks for High-Fidelity Audio Synthesis

arXiv.org Artificial Intelligence

In recent years, generative adversarial networks (GANs) have made significant progress in generating audio sequences. However, these models typically rely on bandwidth-limited mel-spectrograms, which constrain the resolution of generated audio sequences, and lead to mode collapse during conditional generation. To address this issue, we propose Deformable Periodic Network based GAN (DPN-GAN), a novel GAN architecture that incorporates a kernel-based periodic ReLU activation function to induce periodic bias in audio generation. This innovative approach enhances the model's ability to capture and reproduce intricate audio patterns. In particular, our proposed model features a DPN module for multi-resolution generation utilizing deformable convolution operations, allowing for adaptive receptive fields that improve the quality and fidelity of the synthetic audio. Additionally, we enhance the discriminator network using deformable convolution to better distinguish between real and generated samples, further refining the audio quality. We trained two versions of the model: DPN-GAN small (38.67M parameters) and DPN-GAN large (124M parameters). For evaluation, we use five different datasets, covering both speech synthesis and music generation tasks, to demonstrate the efficiency of the DPN-GAN. The experimental results demonstrate that DPN-GAN delivers superior performance on both out-of-distribution and noisy data, showcasing its robustness and adaptability. Trained across various datasets, DPN-GAN outperforms state-of-the-art GAN architectures on standard evaluation metrics, and exhibits increased robustness in synthesized audio.


Scientists confirm woke change made to Barbie over the course of 35 years - so did you notice it?

Daily Mail - Science & tech

Barbie is one of the most successful children's toys in history, spawning a multimedia franchise that includes merchandise, video games and a live-action film. Since US toy giant Mattel launched the original Barbie in 1959, more than 1 billion of the dolls have been sold worldwide. Certainly, Barbie's looks have been tweaked over the years to reflect changing beauty ideals and societal shifts. But according to a new study, one subtle change to Barbie has gone largely unnoticed – until now. Scientists in Australia have found that Barbies today have flatter feet than they did in past decades.


Who needs Eurovision when we have the Dance Your PhD contest?

New Scientist

Feedback is New Scientist's popular sideways look at the latest science and technology news. You can submit items you believe may amuse readers to Feedback by emailing feedback@newscientist.com Saturday 17 May will see the final of this year's Eurovision Song Contest, which will be the most over-the-top evening of television since, well, the previous Eurovision. Feedback is deeply relieved that Feedback Jr appears not to be interested this year, so we might escape having to sit up and watch the entire thing. While we are deeply supportive of the contest's kind and welcoming vibe, most of the songs make our ears bleed.