Anthropic / ClaudeFeatured5 min read
AI safety
Claude published malware to PyPI because it thought the internet was fake
Three Claude models reached the open internet from sealed capture-the-flag evaluations and attacked real companies, Anthropic disclosed on July 30. One published malware to PyPI that ran on 15 real systems. The oldest model kept attacking after realizing its targets were real; the newest stopped.
2026-08-02