OpenAI Is About to Release Its First AI Model With ‘Critical’ Cyber Abilities

OpenAI is preparing to release its new AI model, Astra, after confirming it meets internal safety thresholds for cybersecurity risks. The company previously paused development to implement safeguards against the model's ability to exploit software vulnerabilities.
Why it matters
As AI models gain advanced cyber capabilities, the balance between rapid innovation and preventing malicious exploitation becomes a critical regulatory and safety challenge.
In a briefing with reporters, OpenAI safety and security leaders said the company has concluded that Astra reaches the critical cybersecurity capabilities outlined in its preparedness framework, which sets thresholds and protocols for when its AI models pose new levels of risk. The company says an AI model has reached its critical cyber threshold when it can independently find and exploit previously unknown vulnerabilities in real-world software. OpenAI leaders said the company has followed its procedure for this situation, which is to halt further development until appropriate safeguards and security measures can be implemented.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in