Episode Details

Back to Episodes
OpenAI Rated Its Next Model a 'Critical' Hacking Risk, Restarted the Training It Paused, and Is Shipping It Anyway

OpenAI Rated Its Next Model a 'Critical' Hacking Risk, Restarted the Training It Paused, and Is Shipping It Anyway

Season 2026 Episode 172 Published 1 month, 1 week ago
Description

OpenAI said on September 1 that Astra is the first model to meet its Critical cybersecurity threshold, capable of finding and exploiting unknown flaws across well-protected systems without step-by-step human direction. It restarted its paused frontier training run on August 28 and will release Astra soon, with advanced cyber access limited to a small alpha group that includes the US government. The same week Anthropic disclosed it had also paused training and evaluations after its own incidents, and both labs now call for a coordinated, verifiable pacing mechanism nobody has built.

Continue the story on unscarcity.ai:

Full episode notes on unscarcity.ai

Listen Now

Love PodBriefly?

If you like Podbriefly.com, please consider donating to support the ongoing development.

Support Us