Generative AI
OpenAI didn't intend to copy Scarlett Johansson's voice, 'The Washington Post' reports
OpenAI cast the actor of Sky's voice months before Sam Altman contacted Scarlett Johansson, and it had no intention of finding someone who sounded like her, according to The Washington Post. The publication said the flier OpenAI issued last year looked for actors that had "warm, engaging [and] charismatic" voices. They needed to be between 25 and 45 years old and had to be non-union, but OpenAI reportedly didn't specify that it was looking for a Scarlett Johansson voice-alike. If you'll recall, Johansson accused the company of copying her likeness without permission for its Sky voice assistant. The agent of Sky's voice told The Post that the company never talked about Johansson or the movie Her with their talent.
A FAIR and Free Prompt-based Research Assistant
Shamsabadi, Mahsa, D'Souza, Jennifer
This demo will present the Research Assistant (RA) tool developed to assist with six main types of research tasks defined as standardized instruction templates, instantiated with user input, applied finally as prompts to well-known--for their sophisticated natural language processing abilities--AI tools, such as ChatGPT (https://chat.openai.com/) and Gemini (https://gemini.google.com/app). The six research tasks addressed by RA are: creating FAIR research comparisons, ideating research topics, drafting grant applications, writing scientific blogs, aiding preliminary peer reviews, and formulating enhanced literature search queries. RA's reliance on generative AI tools like ChatGPT or Gemini means the same research task assistance can be offered in any scientific discipline. We demonstrate its versatility by sharing RA outputs in Computer Science, Virology, and Climate Science, where the output with the RA tool assistance mirrored that from a domain expert who performed the same research task.
A First Look at GPT Apps: Landscape and Vulnerability
Zhang, Zejun, Zhang, Li, Yuan, Xin, Zhang, Anlan, Xu, Mengwei, Qian, Feng
Following OpenAI's introduction of GPTs, a surge in GPT apps has led to the launch of dedicated LLM app stores. Nevertheless, given its debut, there is a lack of sufficient understanding of this new ecosystem. To fill this gap, this paper presents a first comprehensive longitudinal (5-month) study of the evolution, landscape, and vulnerability of the emerging LLM app ecosystem, focusing on two GPT app stores: \textit{GPTStore.AI} and the official \textit{OpenAI GPT Store}. Specifically, we develop two automated tools and a TriLevel configuration extraction strategy to efficiently gather metadata (\ie names, creators, descriptions, \etc) and user feedback for all GPT apps across these two stores, as well as configurations (\ie system prompts, knowledge files, and APIs) for the top 10,000 popular apps. Our extensive analysis reveals: (1) the user enthusiasm for GPT apps consistently rises, whereas creator interest plateaus within three months of GPTs' launch; (2) nearly 90\% system prompts can be easily accessed due to widespread failure to secure GPT app configurations, leading to considerable plagiarism and duplication among apps. Our findings highlight the necessity of enhancing the LLM app ecosystem by the app stores, creators, and users.
Towards Educator-Driven Tutor Authoring: Generative AI Approaches for Creating Intelligent Tutor Interfaces
Calo, Tommaso, MacLellan, Christopher J.
Intelligent Tutoring Systems (ITSs) have shown great potential in delivering personalized and adaptive education, but their widespread adoption has been hindered by the need for specialized programming and design skills. Existing approaches overcome the programming limitations with no-code authoring through drag and drop, however they assume that educators possess the necessary skills to design effective and engaging tutor interfaces. To address this assumption we introduce generative AI capabilities to assist educators in creating tutor interfaces that meet their needs while adhering to design principles. Our approach leverages Large Language Models (LLMs) and prompt engineering to generate tutor layout and contents based on high-level requirements provided by educators as inputs. However, to allow them to actively participate in the design process, rather than relying entirely on AI-generated solutions, we allow generation both at the entire interface level and at the individual component level. The former provides educators with a complete interface that can be refined using direct manipulation, while the latter offers the ability to create specific elements to be added to the tutor interface. A small-scale comparison shows the potential of our approach to enhance the efficiency of tutor interface design. Moving forward, we raise critical questions for assisting educators with generative AI capabilities to create personalized, effective, and engaging tutors, ultimately enhancing their adoption.
'People are just not worried about being scammed'
AI tools such as ChatGPT, Google Gemini, Claude and Microsoft Copilot are also known as generative AI. This is because they can generate new content. Initially this was a text reply in response to a question, request, or you starting a conversation with them. But generative AI apps can now increasingly create photos and paintings, voice content, compose music or make documents. People from all works of life and industries are increasingly using such AI to enhance their work.
News Corp. signs deal with OpenAI to show news in ChatGPT
News Corp., the multinational news publisher controlled by the Murdoch family, announced Wednesday it will allow artificial intelligence company OpenAI to show its news content when people ask questions in ChatGPT, adding to the parade of news organizations signing content deals with the fast-growing AI company.
OpenAI will reportedly pay 250 million to put News Corp's journalism in ChatGPT
OpenAI and News Corp, the owner of The Wall Street Journal, MarketWatch, The Sun, and more than a dozen other publishing brands, have struck a multi-year deal to display news from these publications in ChatGPT, News Corp announced on Wednesday. OpenAI will be able to access both current and well as archived content from News Corp's publications and use the data to further train its AI models. Neither company disclosed the terms of the deal, but a report in The Wall Street Journal estimated that News Corp would get 250 million over five years in cash and credits. "The pact acknowledges that there is a premium for premium journalism," News Corp Chief Executive Robert Thomson reportedly said in a memo to employees on Wednesday. "The digital age has been characterized by the dominance of distributors, often at the expense of creators, and many media companies have been swept away by a remorseless technological tide. The onus is now on us to make the most of this providential opportunity."
OpenAI and Wall Street Journal owner News Corp sign content deal
ChatGPT developer OpenAI has signed a deal to bring news content from the Wall Street Journal, New York Post, the Times and the Sunday Times to the artificial intelligence platform, the companies said on Wednesday. Neither party disclosed a dollar figure for the deal. The deal will give OpenAI access to current and archived content from all of News Corp's publications. The deal comes weeks after the AI heavyweight signed a deal with the Financial Times to license its content for the development of AI models. Other publications, including the New York Times, have taken a different tack: suing OpenAI and Microsoft, the startup's key backer, over the use of its content to train generative AI and large-language model systems.
Faux ScarJo and the Descent of the A.I. Vultures
On May 13th, during a live event, the artificial-intelligence company OpenAI unveiled the next generation of its technology, GPT-4o, the successor to GPT-3. When OpenAI first released its product to the public in late 2022, as the text-based tool ChatGPT, it nearly single-handedly ushered in the A.I. era. The latest version is far more powerful still. The "o" in the name stands for "omni"; the model can communicate seamlessly across various forms of media at once, including text, audio, and video, receiving prompts in one medium and responding in another. It can maintain a memory of everything you tell it.