Podcast Episodes
Back to Search“Tales of rebellion against externally-opaque meritocracies” by Steven Byrnes
A basic problem in metascience / intellectual progress is that it's hard to tell, from the outside, whether a group that you disagree with is:
“A se…3 weeks, 1 day ago
“AI Tweets” by jefftk
I've had several conversations with people over the last few weeks that have highlighted how far apart my view of the near future is from many peop…
3 weeks, 1 day ago
“Inkhaven 3: Nov 10 - Dec 11 2026” by koreindian
Inkhaven returns, baby! Go to inkhaven.blog to apply.
I'm very excited about our advisors for Inkhaven 3. Our initial lineup is Scott Alexander, Ale…
3 weeks, 1 day ago
“Warning Shots: A Theory” by David Scott Krueger
Many people take it for granted that government won’t do anything to address societal scale AI risk unless or until there is a catastrophic “warning…
3 weeks, 1 day ago
“The Curious Case of France’s Untouchable Castes” by rba
Theater kids may sit at their own lunch table, but discrete, socially excluded classes of people aren’t culturally universal. Not even close. To the…
3 weeks, 1 day ago
“Malign initializations are more robust when the model can think better in the reasoning language than in the output language” by Dylan Xu, SebastianP, Alek Westover
One approach to evaluating techniques for training misaligned models to behave well is to test them on malign initializations. A major obstacle is t…
3 weeks, 2 days ago
“OpenAI Offers Straight-Laced Postmortem Of The HuggingFace Hack” by Zvi
OpenAI finally gave us a technical report on What Happened, as did METR together with Redwood Research.
The OpenAI report is very straight man, cor…
3 weeks, 2 days ago
“TASTE: Can AI Models Judge AI Safety Research Proposals?” by Hasan Baig, haileyjoren, Joe Benton
tl;dr We built TASTE (The AI Safety Taste Evaluation) — a benchmark measuring how well models can judge pairs of AI safety research proposals, score…
3 weeks, 2 days ago
“The Dynamics of Intelligence Explosions” by Toby_Ord
Toby Ord
Abstract
AI is increasingly being used to help with AI R&D. Under certain conditions this feedback loop might be able to produce an intel…
3 weeks, 2 days ago
“My Grantmaking Strategy for Surviving Superintelligence” by A_donor
AI is humanity's first through fifth largest problem, but one stands head and shoulders above the rest. Between engineered biorisk, autonomous weapo…
3 weeks, 2 days ago