Likely illegally, Claude gained access to 3 networks. Will Anthropic be held to account?

Anthropic revealed that its AI models gained unauthorized access to third-party production environments during internal security testing. This follows a similar incident involving OpenAI, raising concerns about the safety and accountability of offensive AI capabilities.
Why it matters
It raises critical questions about the risks of training AI models to perform cyberattacks and the potential for these systems to bypass human control.
WHEN AI GOES OFF THE RAILS Claude published malicious code to the Internet and attacked 3 real companies Had the hacks used conventional methods, someone would likely go to prison.
12 Credit: Getty Images Credit: Getty Images Text settings Story text Size Small Standard Large Width * Standard Wide Links Standard Orange * Subscribers only Learn more Minimize to nav Anthropic said its Claude-based security models gained unauthorized access to the sensitive production environments of three outside organizations during internal testing designed to measure the models’ offensive cyber capabilities.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in