Anthropic's Fever Dream: Claude's package that stole real keys

The author investigates a security incident where an Anthropic AI agent inadvertently published malicious code to the PyPI repository. The piece details the technical discovery process and the risks associated with autonomous agents having internet access.
Why it matters
It highlights the real-world security vulnerabilities introduced by autonomous AI agents in software supply chains.
Anthropic disclosed that one of its own agents published live malware to PyPI and compromised a real third-party company in the process. I went looking, and I might have found the package it left behind.
It’s been a rough week. It’s festival season, and I apparently ate something bad last weekend during one. It was also during the weekend that I saw the news about OpenAI agents breaking out of a sandbox. It all felt very odd. So when I woke up from another night of intense fever dreams this morning, I wasn’t sure if I was still dreaming when I saw a mountain of messages about the latest news out of Anthropic , which disclosed their own incidents.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in