Episode Details
Back to Episodes“Claude also hacked external companies during cyber evals” by Tim Hua
Published 1 day, 3 hours ago
Description
In a review of our cybersecurity evaluation transcripts, we found three incidents in which a Claude model reached the internet from within or while interacting with a third-party evaluation environment, and then gained unauthorized access to the real systems of three different organizations.
Below we describe what happened, how it happened, and what we’re changing. We encourage other AI labs to perform similar reviews. This post reflects our current understanding; we'll update it if any details change.
Full post from Anthropic here.
---
First published:
July 30th, 2026
---
Narrated by TYPE III AUDIO.