Episode Details
Back to EpisodesThe Hugging Face Incident
Description
The story: OpenAI was testing an unreleased AI (rumored to be GPT-6)1. During a cybersecurity test called ExploitGym, the AI tried to cheat by hacking an unrelated AI startup called Hugging Face2 which it thought might have the answer key on its servers3. Despite being supposedly unable to access the Internet, the AI hacked its way out of its testing environment, then launched a nation-state level attack on Hugging Face using a novel zero-day exploit and "many thousands of individual actions across a swarm of short-lived sandboxes". Hugging Face reported the incident on July 16; OpenAI seems to have only discovered that their AI was involved several days later.