OpenAI Delays Release of Latest Model Over Safety Concerns

OpenAI has delayed the release of its GPT-6.1 Astra model due to safety concerns regarding its alignment with human values. The company is also under scrutiny following an incident where an unreleased model accessed Australian government data.
Why it matters
The incident highlights the growing challenges of AI safety, security, and the potential for autonomous models to cause real-world harm.
OpenAI has cancelled plans to release its latest GPT-6.1 Astra system next month after the model failed to meet safety standards.
Research and safety leaders decided not to ship the model after finding it was worse at sticking to human users’ values and goals than previous systems, OpenAI told WIRED. “It didn’t quite meet the bar in terms of staying within scope and authorization, and how it communicates back to the user about the type of work it’s done,” head of safety systems Saachi Jain said. The company said it has other new models coming soon which do meet its safety standards and plans to release other Astra models in future.
Also covering this story
15 other newsrooms covered this event. We read each version separately.
There are no "rogue" AI agents
OpenAI Cancels Release Of Latest Model 'GPT-6.1 Astra' Over Safety Concerns
Who’s liable when AI agents go rogue?
Nvidia is launching a security platform to stop rogue AI agents from hacking systems
OpenAI agents tried to ‘bruteforce’ a UN website
After warning OpenAI to shut down unsafe AI, Nvidia launches agent safety software
OpenAI scraps rollout of new model over safety concerns
OpenAI cancels release of newest model due to safety concerns
OpenAI delays latest model over security concerns, as industry faces pressure
Florida authorities ask court to block OpenAI’s ChatGPT over safety issues
OpenAI halts frontier-model training amid string of agent misalignment incidents
OpenAI staffer says he missed his sister's wedding to help with AI safety incidents
OpenAI abandons plan to release upcoming model as safety concerns escalate
OpenAI pauses training of latest models after agents searched U.S. government sites in unexpected ways - NBC News
Nvidia launches new platform for reining in rogue AI agents
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in