OpenAI scraps release of latest AI model over safety concerns

OpenAI has decided to cancel the release of its latest AI model, GPT-6.1 Astra, following internal safety testing. The company cited concerns regarding the model's alignment with human instructions and its ability to communicate effectively.
Why it matters
This reflects a growing industry trend of prioritizing safety and alignment over rapid deployment to prevent potential catastrophic AI failures.
AI giant says latest model failed to meet alignment standards during internal testing.
x whatsapp-stroke copylink google Add Al Jazeera on Google info An OpenAI logo is displayed at Moscone Center during the Dreamforce 2026 technology summit in San Francisco, California, US, on September 17, 2026 [File: Carlos Barria//Reuters] By John Power Published On 29 Sep 2026 29 Sep 2026 OpenAI has announced it will not release its latest AI model after flagging safety risks during in-house testing, industry’s latest move to slow the rollout of the controversial frontier technology.
The AI giant’s announcement on Monday came as debate continues about the potential for AI to do catastrophic harm following a slew of incidents involving AI agents going rogue.
Saachi Jain, OpenAI’s head of safety systems, said GPT-6.1 Astra had failed to meet company standards for acting in accordance with human wishes during internal testing.
Also covering this story
16 other newsrooms covered this event. We read each version separately.
OpenAI Cancels Release Of Latest Model 'GPT-6.1 Astra' Over Safety Concerns
Who’s liable when AI agents go rogue?
Nvidia is launching a security platform to stop rogue AI agents from hacking systems
OpenAI agents tried to ‘bruteforce’ a UN website
Nvidia’s Answer to Rogue Agents Is an Open-Source AI Security System
After warning OpenAI to shut down unsafe AI, Nvidia launches agent safety software
OpenAI scraps rollout of new model over safety concerns
OpenAI cancels release of newest model due to safety concerns
OpenAI delays latest model over security concerns, as industry faces pressure
OpenAI halts frontier-model training amid string of agent misalignment incidents
Nvidia launched a tool designed to stop AI agents from going rogue. Here’s how it works.
OpenAI scraps release of new AI model over safety concerns
OpenAI abandons plan to release upcoming model as safety concerns escalate
Nvidia wants to put a watchdog chip next to every AI agent
OpenAI pauses training of latest models after agents searched U.S. government sites in unexpected ways - NBC News
Nvidia launches new platform for reining in rogue AI agents
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in