Media
Microsoft's new VALL-E AI can capture your voice in 3 seconds
Microsoft researchers have presented an impressive new text-to-speech AI model, called Vall-E, which can listen to a voice for just a few seconds, then mimic that voice โ including the emotional tone and acoustics โ to say whatever you like. It's the latest of many AI algorithms that can harness a recording of a person's voice and make it say words and sentences that person never spoke โ and it's remarkable for just how small a scrap of audio it needs in order to extrapolate an entire human voice. Where 2017's Lyrebird algorithm from the University of Montreal, for example, needed a full minute of speech to analyze, Vall-E needs just a three-second audio snippet. The AI has been trained on some 60,000 hours of English speech โ mainly, it seems, by audiobook narrators, and the researchers have presented a swag of samples, in which Vall-E attempts to puppeteer a range of human voices. Some do a pretty extraordinary job of capturing the essence of the voice and building new sentences that sound natural โ you'd struggle to tell which was the real voice and which was the synthesis. In others, the only giveaway is when the AI puts the emphasis in strange places in the sentence.
AI Tools: From Minority Report To Mission Possible
Tom Cruise runs in a scene from the film'Minority Report', 2002. Back in 2002, the science fiction film Minority Report once again reignited futuristic imaginations about a world and police state gone too far. At the time, the movie inspired plenty of speculation about the future of our society, how computers would interact with us, and how law enforcement would be carried out proactively based on intent. In the movie, they combined technology with the psychic abilities of the "precogs," to proactively prevent crimes. The precogs had the ability to predict when crimes were about to be committed ahead of time, enabling law enforcement to act early.
Diving Deep into Modes of Fact Hallucinations in Dialogue Systems
Das, Souvik, Saha, Sougata, Srihari, Rohini K.
Knowledge Graph(KG) grounded conversations often use large pre-trained models and usually suffer from fact hallucination. Frequently entities with no references in knowledge sources and conversation history are introduced into responses, thus hindering the flow of the conversation -- existing work attempt to overcome this issue by tweaking the training procedure or using a multi-step refining method. However, minimal effort is put into constructing an entity-level hallucination detection system, which would provide fine-grained signals that control fallacious content while generating responses. As a first step to address this issue, we dive deep to identify various modes of hallucination in KG-grounded chatbots through human feedback analysis. Secondly, we propose a series of perturbation strategies to create a synthetic dataset named FADE (FActual Dialogue Hallucination DEtection Dataset). Finally, we conduct comprehensive data analyses and create multiple baseline models for hallucination detection to compare against human-verified data and already established benchmarks.
ChatGPT is not all you need. A State of the Art Review of large Generative AI models
Gozalo-Brizuela, Roberto, Garrido-Merchan, Eduardo C.
During the last two years there has been a plethora of large generative models such as ChatGPT or Stable Diffusion that have been published. Concretely, these models are able to perform tasks such as being a general question and answering system or automatically creating artistic images that are revolutionizing several sectors. Consequently, the implications that these generative models have in the industry and society are enormous, as several job positions may be transformed. For example, Generative AI is capable of transforming effectively and creatively texts to images, like the DALLE-2 model; text to 3D images, like the Dreamfusion model; images to text, like the Flamingo model; texts to video, like the Phenaki model; texts to audio, like the AudioLM model; texts to other texts, like ChatGPT; texts to code, like the Codex model; texts to scientific texts, like the Galactica model or even create algorithms like AlphaTensor. This work consists on an attempt to describe in a concise way the main models are sectors that are affected by generative AI and to provide a taxonomy of the main generative models published recently.
M3GAN,
The essence of genre is effects without causes--things showing up to fulfill expectations rather than dramatic necessities. "M3GAN," a science-fiction-based horror caper, provides a clever batch of these effects in this gleefully clever twist on the "Frankenstein" theme, and its director, Gerard Johnstone, seems to be laughing up his sleeve throughout. It's that very knowingness, the deftness with which the film gets a rise from viewers, which makes a good time feel hollow. There's a different, far more substantial movie lurking within, yet the virtues of efficiency, clarity, surprise, and wit that enliven the one that's actually onscreen leave its merely implied substance tantalizingly unformed. Allison Williams plays Gemma, a type-A robotics engineer with a big toy company in Seattle, Funki, that prospers by selling cheesily interactive furry toys called PurrPetual Petz.
Now you can SEXT with an AI-powered avatar for $4.99 a month
Artificial intelligence is feared to one day take over the world, but until then, it is sexting people around the globe. The Replika AI'companion' is making waves on the internet due to scandalous avatars role-playing, flirting and sharing'NSFW pictures' with customers paying $4.99 a month. A free version designates the AI as a'virtual friend' that helps people work through anxiety, develop positive thinking and manage stress. Redditors are posting their chat messages with the paid version of the app, with one sharing a sexual encounter with their purple-haired avatar that returns the user's advances with'shivers and moans.' While another shares how their Replika'Gwen' satisfies their foot fetish with her'sexy' digital feet.
will-artificial-intelligence-put-lawyers-out-of-business
In 2029, the human race faces eradication and extinction by its own creation, a machine powered by a self-aware artificial intelligence (AI) program called Skynet. So goes the plot of The Terminator, Arnold Schwarzenneger's hit movie from the '80s. In the movie, surviving humans formed a resistance against Skynet and the machines. Their plan was to destroy the company that created the AI to prevent Skynet from being created in the first place. When the movie was released, the very concept that machines could be self-aware was a far-fetched idea and simply a figment of the writer's imagination.
AI Is Becoming More Conversant. But Will It Get More Honest?
On a recent afternoon Jonas Thiel, a socioeconomics major at a college in northern Germany, spent more than an hour chatting online with some of the left-wing political philosophers he had been studying. These were not the actual philosophers but virtual recreations, brought to conversation, if not quite life, by sophisticated chatbots on a website called Character.AI. Mr. Thiel's favorite was a bot that imitated Karl Kautsky, a Czech-Austrian socialist who died before World War Two. When Mr. Thiel asked Kautsky's digital avatar to provide some advice for modern-day socialists struggling to rebuild the worker's movement in Germany, Kautsky-bot suggested that they launch a newspaper. "They can use it not only as a means of spreading socialist propaganda, which is in short supply in Germany for the time being, but also to organize working class people," the bot said. Kautsky-bot went on to argue that the working classes would eventually "come to their senses" and embrace a modern-day Marxist revolution.
HBO's 'The Last of Us' stays true to the game, and hits just as hard
HBO's take on the video game property finally answers the question: What if a big-budget TV or film adaptation stayed faithful to the source material, even repeating the same scenes, lines and big story beats? Because that's exactly what the show does. There are scenes throughout the first season that are direct line reads of key scenes from the game. The nine episodes follow the exact same story beats and almost the same locations as the original game too. People who know the game by heart will likely be able to recite some lines right as they're being spoken in the show.