Goto

Collaborating Authors

 Personal


Adding guardrails to advanced chatbots

arXiv.org Artificial Intelligence

Generative AI models continue to become more powerful. The launch of ChatGPT in November 2022 has ushered in a new era of AI. ChatGPT and other similar chatbots have a range of capabilities, from answering student homework questions to creating music and art. There are already concerns that humans may be replaced by chatbots for a variety of jobs. Because of the wide spectrum of data chatbots are built on, we know that they will have human errors and human biases built into them. These biases may cause significant harm and/or inequity toward different subpopulations. To understand the strengths and weakness of chatbot responses, we present a position paper that explores different use cases of ChatGPT to determine the types of questions that are answered fairly and the types that still need improvement. We find that ChatGPT is a fair search engine for the tasks we tested; however, it has biases on both text generation and code generation. We find that ChatGPT is very sensitive to changes in the prompt, where small changes lead to different levels of fairness. This suggests that we need to immediately implement "corrections" or mitigation strategies in order to improve fairness of these systems. We suggest different strategies to improve chatbots and also advocate for an impartial review panel that has access to the model parameters to measure the levels of different types of biases and then recommends safeguards that move toward responses that are less discriminatory and more accurate.


HELP ME THINK: A Simple Prompting Strategy for Non-experts to Create Customized Content with Models

arXiv.org Artificial Intelligence

Controlling the text generated by language models and customizing the content has been a long-standing challenge. Existing prompting techniques proposed in pursuit of providing control are task-specific and lack generality; this provides overwhelming choices for non-expert users to find a suitable method for their task. The effort associated with those techniques, such as in writing examples, explanations, instructions, etc. further limits their adoption among non-expert users. In this paper, we propose a simple prompting strategy HELP ME THINK where we encourage GPT3 to help non-expert users by asking a set of relevant questions and leveraging user answers to execute the task. We demonstrate the efficacy of our technique HELP ME THINK on a variety of tasks. Specifically, we focus on tasks that are hard for average humans and require significant thinking to perform. We hope our work will encourage the development of unconventional ways to harness the power of large language models.


AI jobs with mind-blowing paychecks of $375K a year

FOX News

Harvey Castro talks about how AI cold be used in cold cases and the symbiotic relationship between AI and a detective. There's no question that artificial intelligence is changing our lives. A bot that sounds almost human can author your emails, teach you a new language, book your trips or even be your friend. Check out direct links to try those out here. One woman I spoke with on my national radio show even married her AI companion.


It Was Founded in a Denny's. Now It's Worth More Than Facebook.

Slate

Nvidia, the company that dominates the market for graphics processing units, was once known mostly in the video game world. But these days, Nvidia GPUs are also the go-to source for the massive computing power needed to run generative A.I. systems--and the recent explosion in A.I. hype has propelled the company's stock into the stratosphere. Nvidia briefly hit a trillion-dollar valuation, putting itself in league with tech giants like Alphabet and Apple and launching a bit of a frenzy in the markets. Nvidia is looking like the first big stock win of the A.I. era, and investors are salivating. On Sunday's episode of What Next: TBD, I spoke with Don Clark, a freelance reporter who specializes in the chips industry, about how Nvidia rode the A.I. revolution, became the hottest chipmaker in the world, and made the entire A.I. craze suddenly seem very real.


Challenges and Opportunities for the Design of Smart Speakers

arXiv.org Artificial Intelligence

Advances in voice technology and voice user interfaces (VUIs) -- such as Alexa, Siri, and Google Home -- have opened up the potential for many new types of interaction. However, despite the potential of these devices reflected by the growing market and body of VUI research, there is a lingering sense that the technology is still underused. In this paper, we conducted a systematic literature review of 35 papers to identify and synthesize 127 VUI design guidelines into five themes. Additionally, we conducted semi-structured interviews with 15 smart speaker users to understand their use and non-use of the technology. From the interviews, we distill four design challenges that contribute the most to non-use. Based on their (non-)use, we identify four opportunity spaces for designers to explore such as focusing on information support while multitasking (cooking, driving, childcare, etc), incorporating users' mental models for smart speakers, and integrating calm design principles.


ChatGPT: Jack of all trades, master of none

arXiv.org Artificial Intelligence

OpenAI has released the Chat Generative Pre-trained Transformer (ChatGPT) and revolutionized the approach in artificial intelligence to human-model interaction. Several publications on ChatGPT evaluation test its effectiveness on well-known natural language processing (NLP) tasks. However, the existing studies are mostly non-automated and tested on a very limited scale. In this work, we examined ChatGPT's capabilities on 25 diverse analytical NLP tasks, most of them subjective even to humans, such as sentiment analysis, emotion recognition, offensiveness, and stance detection. In contrast, the other tasks require more objective reasoning like word sense disambiguation, linguistic acceptability, and question answering. We also evaluated GPT-4 model on five selected subsets of NLP tasks. We automated ChatGPT and GPT-4 prompting process and analyzed more than 49k responses. Our comparison of its results with available State-of-the-Art (SOTA) solutions showed that the average loss in quality of the ChatGPT model was about 25% for zero-shot and few-shot evaluation. For GPT-4 model, a loss for semantic tasks is significantly lower than for ChatGPT. We showed that the more difficult the task (lower SOTA performance), the higher the ChatGPT loss. It especially refers to pragmatic NLP problems like emotion recognition. We also tested the ability to personalize ChatGPT responses for selected subjective tasks via Random Contextual Few-Shot Personalization, and we obtained significantly better user-based predictions. Additional qualitative analysis revealed a ChatGPT bias, most likely due to the rules imposed on human trainers by OpenAI. Our results provide the basis for a fundamental discussion of whether the high quality of recent predictive NLP models can indicate a tool's usefulness to society and how the learning and validation procedures for such systems should be established.


Congratulations to the #IJCAI2023 award winners

AIHub

The winners of three IJCAI awards have been announced. These three distinctions are: the Award for Research Excellence, the John McCarthy Award and the Computers and Thought Award. The Research Excellence award is given to a scientist who has carried out a program of research of consistently high quality throughout an entire career yielding several substantial results. The winner of the 2023 Award for Research Excellence is Sarit Kraus, Professor of Computer Science, Bar-Ilan University, Israel. Professor Kraus is recognized for her pioneering work on the study of interactions among self-interested agents, creating the field of automated negotiation, and developing methods for coalition formation and teamwork, both as formal models and real-world implementations.


The Morning After: Let's talk about Air Quality

Engadget

Wildfires in Canada have led to a surge in air pollution levels in the US, with New York currently having the worst air quality of any major city. There are plenty of images of N95-mask-wearing people walking down smog-blighted streets that wouldn't look out of place in many a dystopia. Many states and cities have urged folks to stay inside unless they absolutely need to leave, and they're pumping out as much Air Quality Index data as they can. But do you actually know what the Air Quality Index is, or what it's for? We've done an AQI deep dive, exploring how it works and how you can keep yourself informed and safe. And, on the subject of being safe, we've also knocked up a guide for how to make a quick-and-dirty box fan air filter.


Where do YOU think the North of England begins? Scientists create a controversial new map

Daily Mail - Science & tech

It is a debate sure to ruffle feathers, but anything beyond the Watford Gap really should be classed as the north of England, a study suggests. This is the critical line at which high street bakery Greggs, the beacon of northernness, becomes more popular than the southerners' sandwich shop of choice, Pret A Manger, an academic study has worked out using artificial intelligence. If the national consumption of steak bakes versus houmous-filled wraps and smashed avocado on toast were not convincing enough, the researchers also looked at the distribution of Morrisons and Waitrose supermarkets across England. This too put the north-south divide within two miles of the Watford Gap. Both calculations agree that Birmingham, Coventry and Leicester are technically in the north of England. But bizarrely, the Pret and Greggs dividing line shows that Cornwall is northern.


My Family's Entire Life Is Based Around Video Games. I Can't Take It Anymore.

Slate

Care and Feeding is Slate's parenting advice column. Have a question for Care and Feeding? Submit it here or post it in the Slate Parenting Facebook group. My husband is very involved with the kids. He's a good father--he does the hard parts of parenting, happily.