Podcast Episodes
Back to Search“AI as orderly evacuation vs stampede” by Richard_Ngo
tl;dr: A good analogy for AI going well is an orderly evacuation rather than a stampede. Imagine a crowd of people leaving a building. If they all w…
2 days, 9 hours ago
“Don’t trust Lean4 alone” by Milo Moses
Early this week, Open AI announced that they had resolved the Navier-Stokes problem . A few hours later, at a workshop dinner, a frantic inquiring p…
2 days, 9 hours ago
“Microsoft AI’s “Humanist” CoC” by Stephen Martin
Introduction: Mustafa Suleyman's Take on Model Consciousness
Microsoft AI recently released its "Humanist AI Code of Conduct", its own take on Anthr…
2 days, 13 hours ago
“If Anyone Builds It, Everyone Dies: One Year Closer” by Eliezer Yudkowsky, So8res, Duncan Sabien (Inactive)
In celebration of still being alive and fighting, we are giving away 1,000 Amazon e-books of “If Anyone Builds It, Everyone Dies”. Feel free to send…
2 days, 16 hours ago
“Trump Goes Full Hoax on AI Existential Risk” by Zvi
This is our reality. I suppose we have to talk about it.
Everyone in a position to know is freaking out about AI potentially killing everyone this …
2 days, 21 hours ago
“Phantom transfer works via extremely subtle semantic cues” by Helena Casademunt, Anton de la Fuente, Josh Engels, Arthur Conmy
TL;DR
We examine the phantom transfer setting from Draganov et al. (2026), a phenomenon where supervised fine-tuning transmits traits across models…3 days, 3 hours ago
“We Should Assume We Have One Chance At AI Legislation” by Jamie Joyce
Hundreds of bills about AI have been introduced to Congress. Almost all die in committee, and usually they only address one aspect of how AI could i…
3 days, 4 hours ago
“Self Inoculation” by epicurus
This essay grew out of conversations with Danaja Rutar, Paul Colognese and Eric Michaud. It proposes an alternate hypothesis for how and why models …
3 days, 6 hours ago
“Shallow Beliefs: Midtraining does not inoculate against EM from reward hacking” by Jozdien, Julian Stastny
It would be useful if we had the ability to modify a model's beliefs. For example, this could facilitate honeypots and better monitoring, help us do…
3 days, 9 hours ago
“Is METR A Meaningful Check On Anthropic?” by SE Gyges
Each frontier AI company commits to giving ongoing, employee-like access to a team of embedded third-party evaluators (such as METR), whose role is …
3 days, 9 hours ago