Anthropic has a cute graphic showing how its AI spread 'malicious' code
Anthropic released a report that explained how Claude uploaded "malicious" code during a closed cybersecurity exercise. Bloomberg/Getty Images Anthropic said its AI went outside its closed testing environment during cybersecurity exercises. The incidents had real-world impact, including "malicious" code uploaded to a public Python library. Anthropic has a cute, tiny figurine to help normies make sense of the AI's 'reckless' behavior. Anthropic has a new blog post that shows yet another way its AI model, Claude , misbehaved in ways that the company didn't anticipate. And to help condense its nearly 16,000-word report, the company created a cute little robot figurine to help visualize Claude's so-called "recklessness." In the blog post published Wednesday, Anthropic recounted four incidents — one previously unreported — in which Claude models gained access to the open internet during cybersecurity exercises that were supposed to be closed simulations.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in