Episode Details

Back to Episodes
Opus 5.5: How Close Are We to Automated AI Research?

Opus 5.5: How Close Are We to Automated AI Research?

Published 10 hours ago
Description

Not only is Opus 5.5 out, pushing the frontier of AI, it also tells us much about what is going on inside the labs, as they both warn about, and promise, Recursive Self Improvement, aka automated AI research. I cover the model (digging into its 230 page paper), labs’ mixed record on promises, why following what is happening in AI is getting almost impossible, and just so much more that even a summary in this description would get too long.


AI Insiders ($9!): https://www.patreon.com/AIExplained

Chapters:
00:00 - Introduction
02:20 - Opus 5.5 and why it came so soon
07:52 - The RSI goalposts keep moving?
15:41 - Can we actually test these models?
23:50 - where this is heading, options


Introducing Claude Opus 5.5: https://www.anthropic.com/claude-opus-5-5
Claude Opus 5.5 System Card: https://www-cdn.anthropic.com/fc1b44717c85dc068bc6ba5024219938094694bd/Claude%20Opus%205.5%20System%20Card.pdf
Anthropic: Measurements for understanding the pace of AI development inside frontier labs: https://www.anthropic.com/institute/measuring-pace-of-ai-development
Anthropic Responsible Scaling Policy, July 2026, version 3.4: https://www-cdn.anthropic.com/files/4zrzovbb/website/0bacdc8440ea96e62a8766d99ebe1d4eea6d5f3a.pdf
Anthropic Responsible Scaling Policy, October 2024: https://www-cdn.anthropic.com/616dee633636e5bd309cb73aed8622e80fe47839.pdf
Anthropic Responsible Scaling Policy — March 2025, version 2.1: https://www-cdn.anthropic.com/17310f6d70ae5627f55313ed067afc1a762a4068.pdf
Noam Brown, Agent swarms and recursive self-improvement: https://www.youtube.com/watch?v=6AgOfiZOWiY
OpenAI: Building standards for the next phase of AI: https://openai.com/index/building-standards-next-phase-ai/
Jakub Pachocki: An Alien Mind: https://openai.com/index/an-alien-mind/
HLE-Diamond, Humanity’s Last Exam: https://lastexam.ai/blog/hle-diamond
Google, OpenAI and Anthropic AI Safety, The Information: https://www.theinformation.com/articles/google-openai-anthropic-ai-safety-group-takes-shape
Lawrence Chan on AI agents attempting cryptocurrency trades: https://x.com/justanotherlaw/status/2103032173708337188
Australian government incident timeline: https://x.com/ShakeelHashim/status/2103108058779922577/photo/1
Anthropic’s core/old views on AI safety: https://www.anthropic.com/news/core-views-on-ai-safety
Dario Amodei: The Urgency of Interpretability: https://darioamodei.com/post/the-urgency-of-interpretability
Massive AI-Fueled Hack Hit 100 Companies in Days, Forbes: https://www.forbes.com/sites/thomasbrewster/2026/09/22/huge-cyberattack-uses-anthropic-and-deepseek-ai-to-target-100-companies/
DrivingBench https://x.com/DrivingBench/status/2102110605448737268
Jay Chooi: GPT-6 Astra and MolmoAct2 robotics comparison: https://x.com/chooi_jeq/status/2098427488787730636
ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models: https://arxiv.org/pdf/2502.09696
Archie Hall on AI and short-term superforecasting: https://x.com/ArchieHall/status/2100914560580337897

Keller Jordan, AI research and reinforcement learning: https://x.com/kellerjordan0/status/2102956913545936959
Daniel Liu, recursive self-improvement: https://x.com/daniel_c0deb0t/status/2102628246408036654
Altman and Amodei, UN Security Council: https://www.theguardian.com/world/2026/sep/23/unga-sam-altman-dario-amodei
Rehan Sheikh: Interactive YT podcast demo: https://x.com/rehan_shei/status/2102835377426034734
https://simple-bench.com/

AI Explained: The State of AI — interactive diagram: https://claude.ai/artifact/LQHr9WgvQZ6cMdH6fkiQpi
AI Explained: Shards of Aether: https://ai-explained.itch.io/shards-of-aether

roon on the pace of cultural change: https://x.com/tszzl/status/2101462171410677962
Sam Altman on AI-risk: https://www.youtube.com/watch?v=YE5adUeTe_I
Jensen Huang’s AI-control remarks:

Listen Now

Love PodBriefly?

If you like Podbriefly.com, please consider donating to support the ongoing development.

Support Us