Large Language Model
Targeted Augmentation for Low-Resource Event Extraction
Addressing the challenge of low-resource information extraction remains an ongoing issue due to the inherent information scarcity within limited training examples. Existing data augmentation methods, considered potential solutions, struggle to strike a balance between weak augmentation (e.g., synonym augmentation) and drastic augmentation (e.g., conditional generation without proper guidance). This paper introduces a novel paradigm that employs targeted augmentation and back validation to produce augmented examples with enhanced diversity, polarity, accuracy, and coherence. Extensive experimental results demonstrate the effectiveness of the proposed paradigm. Furthermore, identified limitations are discussed, shedding light on areas for future improvement.
ParallelPARC: A Scalable Pipeline for Generating Natural-Language Analogies
Sultan, Oren, Bitton, Yonatan, Yosef, Ron, Shahaf, Dafna
Analogy-making is central to human cognition, allowing us to adapt to novel situations -- an ability that current AI systems still lack. Most analogy datasets today focus on simple analogies (e.g., word analogies); datasets including complex types of analogies are typically manually curated and very small. We believe that this holds back progress in computational analogy. In this work, we design a data generation pipeline, ParallelPARC (Parallel Paragraph Creator) leveraging state-of-the-art Large Language Models (LLMs) to create complex, paragraph-based analogies, as well as distractors, both simple and challenging. We demonstrate our pipeline and create ProPara-Logy, a dataset of analogies between scientific processes. We publish a gold-set, validated by humans, and a silver-set, generated automatically. We test LLMs' and humans' analogy recognition in binary and multiple-choice settings, and found that humans outperform the best models (~13% gap) after a light supervision. We demonstrate that our silver-set is useful for training models. Lastly, we show challenging distractors confuse LLMs, but not humans. We hope our pipeline will encourage research in this emerging field.
ERATTA: Extreme RAG for Table To Answers with Large Language Models
Roychowdhury, Sohini, Krema, Marko, Mahammad, Anvar, Moore, Brian, Mukherjee, Arijit, Prakashchandra, Punit
Large language models (LLMs) with retrieval augmented-generation (RAG) have been the optimal choice for scalable generative AI solutions in the recent past. However, the choice of use-cases that incorporate RAG with LLMs have been either generic or extremely domain specific, thereby questioning the scalability and generalizability of RAG-LLM approaches. In this work, we propose a unique LLM-based system where multiple LLMs can be invoked to enable data authentication, user query routing, data retrieval and custom prompting for question answering capabilities from data tables that are highly varying and large in size. Our system is tuned to extract information from Enterprise-level data products and furnish real time responses under 10 seconds. One prompt manages user-to-data authentication followed by three prompts to route, fetch data and generate a customizable prompt natural language responses. Additionally, we propose a five metric scoring module that detects and reports hallucinations in the LLM responses. Our proposed system and scoring metrics achieve >90% confidence scores across hundreds of user queries in the sustainability, financial health and social media domains. Extensions to the proposed extreme RAG architectures can enable heterogeneous source querying using LLMs.
Why Protesters Around the World Are Demanding a Pause on AI Development
Just one week before the world's second-ever global summit on artificial intelligence, protesters of a small but growing movement called "Pause AI" demanded that the world's governments regulate AI companies and freeze the development of new cutting edge artificial intelligence models. They say that the development of these models should only be allowed to continue if companies agree to let them be thoroughly evaluated to test their safety first. Protests took place across thirteen different countries, including the U.S., the U.K, Brazil, Germany, Australia, and Norway on Monday. In London, a group of 20 or so protesters stood outside of the U.K.'s Department of Science, Innovation and Technology chanting things like "stop the race, it's not safe" and "who's future? The protestors say their goal is to get governments to regulate the companies developing frontier AI models, including OpenAI's Chat GPT. They say that companies are not taking enough precautions to make sure their AI models are safe enough to be released into the world. "[AI companies] have proven time and time again… through the way that these companies' workers are treated, with the way that they treat other people's work by literally stealing it and throwing it into their models, They have proven that they cannot be trusted," said Gideon Futerman, an Oxford undergraduate student who gave a speech at the protest. One protester, Tara Steele, a freelance writer who works on blogs and SEO content, said that she had seen the technology impact her own livelihood. "I have noticed since ChatGPT came out, the demand for freelance work has reduced dramatically," she says. "I love writing personally… I've really loved it.
ChatGPT got an upgrade to make it seem more human
OpenAI's latest model offers a more human-like conversational experience OpenAI announced its newest artificial intelligence model, called GPT-4o, which will soon power some versions of the company's ChatGPT product. The upgraded ChatGPT can swiftly respond to text, audio and video inputs from its real-time conversational partner – all while speaking with inflections and wording that convey a strong sense of emotion and personality. The company demonstrated the emotional mimicry of the new voice mode during a supposedly live OpenAI presentation, featuring both the ChatGPT mobile app and a new desktop app, on 13 May. Speaking in a female-sounding voice and responding to the name ChatGPT, the new AI's conversational capabilities seemed more akin to the personable AI voiced by Scarlett Johansson in the 2013 science fiction film Her than to the more canned and robotic responses of typical voice assistant technologies. How this moment for AI will change society forever (and how it won't) "The new GPT-4o voice-to-voice interaction more closely parallels human-human interaction," says Michelle Cohn at the University of California, Davis.
This Is the Next Smartphone Evolution
Earlier today, OpenAI announced its newest product: GPT-4o, a faster, cheaper, more powerful version of its most advanced large language model, and one that the company has deliberately positioned as the next step in "natural human-computer interaction." Running on an iPhone in what was purportedly a live demo, the program appeared able to tell a bedtime story with dramatic intonation, understand what it was "seeing" through the device's camera, and interpret a conversation between Italian and English speakers. The model--which was powering an updated version of the ChatGPT app--even exhibited something like emotion: Shown the sentence I ChatGPT handwritten on a page, it responded, "That's so sweet of you!" Although such features are not exactly new to generative AI, seeing them bundled into a single app on an iPhone was striking. Watching the presentation, I felt that I was witnessing the murder of Siri, along with that entire generation of smartphone voice assistants, at the hands of a company most people had not heard of just two years ago.
Protesters Are Fighting to Stop AI, but They're Split on How to Do It
On a side street outside the headquarters of the Department of Science, Innovation and Technology in the center of London on Monday, 20 or so protesters are getting their chants in order. When do we want it?" These protesters are part of Pause AI, a group of activists petitioning for companies to pause development of large AI models which they fear could pose a risk to the future of humanity. Other PauseAI protests are taking place across the globe: In San Francisco, New York, Berlin, Rome, Ottawa, and a handful of other cities. Their aim is to grab the attention of voters and politicians ahead of the AI Seoul Summit--a follow-up to the AI Safety Summit held in the UK in November 2023. But the loosely organized group of protesters itself is still figuring out exactly the best way to communicate its message. "The Summit didn't actually lead to meaningful regulations," says Joep Meindertsma, the founder of PauseAI. The attendees at the conference agreed to the "Bletchley Declaration," but that agreement doesn't mean much, Meindertsma says. "It's only a small first step, and what we need are binding international treaties." The group's main demand is for a pause on the training of AI systems more powerful than GPT-4--it's calling for all countries to implement this measure, but specifically calls out the United States as the home of most leading AI labs. The group also wants all UN member states to sign a treaty that sets up an international AI safety agency with responsibility for granting new deployments of AI systems and training runs of large models. Their protests are taking place on the same day as OpenAI announced a new version of ChatGPT to make the chatbot act more like a human. "We have banned technology internationally before," says Meindertsma, pointing to the Montreal Protocol, a global agreement finalized in 1987 that saw the phaseout of CFCs and other chemicals known to deplete the ozone layer. "We've got treaties that ban blinding laser weapons.
New GPT-4o AI model is faster and free for all users, OpenAI announces
OpenAI announced on Monday that it was launching its new flagship artificial intelligence model, called GPT-4o, as well as updates that included a new desktop service and advances in its voice assistant capabilities. Chief technology officer, Mira Murati, appeared on stage to a cheering crowd in the OpenAI offices, touting the new model as a step forward in AI. The new model will bring the faster, more accurate GPT-4 AI model to free users, where it was previously reserved for paid customers. "We're looking at the future of interaction between ourselves and the machines," Murati said. "We think GPT-4o is really shifting that paradigm."
OpenAI claims that its free GPT-4o model can talk, laugh, sing and see like a human
OpenAI on Monday announced GPT-4o, a brand new AI model that that the company says is one step closer to "much more natural human-computer interaction." The new model accepts any combination of text, audio and images as input and can generate an output in all three formats. It's also capable of recognizing emotion, lets you interrupt it mid-speech, and responds nearly as fast as a human being during conversations. "The special thing about GPT-4o is it beings GPT-4 level intelligence to everyone, including our free users," said OpenAI CTO Mira Murati during a live-streamed presentation. "This is the first time we're making a huge step forward when it comes to ease of use."