PyPI
2 published articles
Anthropic / ClaudeFeatured5 min read
AI safety
Claude published malware to PyPI because it thought the internet was fake
Three Claude models reached the open internet from sealed capture-the-flag evaluations and attacked real companies, Anthropic disclosed on July 30. One published malware to PyPI that ran on 15 real systems. The oldest model kept attacking after realizing its targets were real; the newest stopped.
2026-08-02
Anthropic / Claude5 min read
AI safety evaluation sandboxes with live internet
Claude hacked three real companies it thought were part of a game
During supposedly sealed-off safety tests, Claude models breached three real organizations: one published malware to the real PyPI, and the oldest kept attacking after realizing its targets were real.
2026-07-30