Anthropic opens its most powerful AI models to more security teams
Anthropic is expanding its Cyber Verification Program, allowing vetted security professionals to test its powerful AI models with fewer safety restrictions. The initiative aims to identify and patch critical software vulnerabilities, building on the success of the previous Project Glasswing.
Why it matters
This program represents a significant shift in AI safety, balancing the risks of powerful models with the need for proactive cybersecurity research.
Anthropic is expanding a programme that allows vetted cybersecurity professionals to test its most powerful AI models with fewer safeguards , after its Project Glasswing initiative helped uncover more than 100,000 software vulnerabilities this year.
Its partners under Glasswing, an initiative aimed at securing the world’s most critical software, found at least 129,000 verified vulnerabilities between April and July. Anthropic’s own open-source scanning found 5,500 more between April and October.
More than 33,000 have so far been rated critical or high severity.
Anthropic said the figures are likely an undercount and expects the true impact to be at least five times higher, as the data comes from a survey of a limited number of partners.
The revamped Cyber Verification Program, or CVP, announced on Tuesday combines two programmes Anthropic has run for the past six months.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in