US finalises voluntary tests for AI models’ hacking powers

The White House has finalized a voluntary framework for testing the cybersecurity risks of advanced AI models before they are released to the public. The initiative encourages major AI labs to cooperate with the government to prevent models from being used for cyberattacks.
Why it matters
As AI capabilities grow, establishing safety benchmarks is critical to preventing the weaponization of frontier models.
The White House has finalised a voluntary framework for testing whether America’s most advanced AI models can be used to hack. A White House official said the framework, ordered in June, was completed by its deadline, with talks on next steps now under way.
The tests are cybersecurity assessments, designed to gauge the offensive capabilities of frontier models before they reach the wider world. Crucially, they are voluntary, so the government is inviting the labs to take part rather than compelling them.
The framework flows from an executive order signed on 2 June, which set the deadline and the light-touch shape of the programme. It is a narrower instrument than earlier drafts, favouring cooperation over mandates.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in