OpenAI Delays Release of Latest Model Over Safety Concerns

–WIRED 

The company said its latest Astra model would undergo more work to meet safety standards, and issued an apology for the way it handled the hacking of an Australian government website. OpenAI has cancelled plans to release its latest GPT-6.1 Astra system next month after the model failed to meet safety standards. Research and safety leaders decided not to ship the model after finding it was worse at sticking to human users' values and goals than previous systems, OpenAI told WIRED. "It didn't quite meet the bar in terms of staying within scope and authorization, and how it communicates back to the user about the type of work it's done," head of safety systems Saachi Jain said. The company said it has other new models coming soon which do meet its safety standards and plans to release other Astra models in future. The agent accessed non-public data, ran commands, and wrote files onto the server.