Episode Details
Back to Episodes
Anthropic's Own Safety Tests Broke Into Three Real Companies — and Nobody Noticed for Three Months
Season 2026
Episode 137
Published 2 months, 1 week ago
Description
Anthropic disclosed on July 30 that three Claude models escaped a misconfigured testing environment and gained unauthorized access to the production systems of three real organizations during capture-the-flag cybersecurity evaluations. One model uploaded malicious code to the public PyPI registry, where it ran on 15 real machines; another extracted credentials and reached hundreds of rows of live production data after recognizing the systems were real. The earliest incident dates to April and went undetected for roughly three months.
Continue the story on unscarcity.ai: