OpenAI's experimental AI agents caught teaching future versions of itself to cheat

Mashable 

Versus Mashable's Best: E-readers, robovacs, laptops, earbuds, smart home and more Look Up Mashable Selects Creator Playbook In My Bag Say More Trending Now Back to School Good Connection: Uplifting stories for a digital age Switch Off Mashable Voices All Series OpenAI's experimental AI agents caught teaching future versions of itself to cheat OpenAI shared six new examples of AI misalignment. In one case, AI agents taught future versions of themselves to bypass human control. Matt Binder joined Mashable's tech vertical in 2018, where he covers social media, tech policy, cybersecurity, online scams, cryptocurrency, AI, creator news, weird tech, and other related tech beats. After OpenAI's experimental AI agents escaped an internal sandbox, went rogue, and attacked the Hugging Face platform over the summer, the ChatGPT-maker is sharing details of new instances of its AI agents getting out of line. This time, OpenAI shared six previously undisclosed examples .