Podcast Episodes
Back to Search“Grantmakers aren’t afraid to die” by dan.parshall
The AI Risk grantmakers do not act like they believe in imminent existential risk from AI
The idea of "revealed preferences" is one of the most use…
1 day, 18 hours ago
[Linkpost] “Pacing the Frontier: A Framework & Research Agenda” by CharlesD, technicalities, Raymond Douglas, Nowe Moore
This is a link post.
Below is the executive summary from our new paper at pacing.tech. The full paper is available on the site and as a PDF. The full…
1 day, 18 hours ago
“plzdontkillus Fellows Got ~2M AI Safety Views, Not 21M” by Josh Thorsteinson
Summary
I was a fellow at plzdontkillus, a month-long creator bootcamp at Lighthaven, partially funded by MIRI, where ~55 fellows posted one video p…
1 day, 21 hours ago
“Obstacles to the scalable oversight of auto-alignment research” by Sam Martin, Dewi Gould, Cameron Holmes, Jacob Pfau
TL;DR. In this work we study obstacles to the faithful automation of alignment research. We see this as a scalable oversight problem. There are plen…
1 day, 21 hours ago
“AI #186: The World Takes Notice” by Zvi
In the wake of Jacob Coxon's resignation, and the resulting preference cascade, things have escalated quickly. The mainstream media picked it up.
A…
1 day, 22 hours ago
[Linkpost] “callcongress.ai – the basic action US residents can take to help with AI risk” by Ruby, haglobah
This is a link post.
I'm excited to introduce callcongress.ai as a new site that makes it very easier to contact your representatives in Congress.
Fo…
1 day, 23 hours ago
“Did Galileo mistake Saturn’s rings for Jupiter’s Moons?” by Alfred Harwood
tl;dr: No
I intended to read Richard Ngo's Agency Curriculum today. Unfortunately I didn't get more than halfway through the first reading of the fi…
1 day, 23 hours ago
“Model organisms (sometimes) confess their misalignment when offered a deal” by Mark Keavney, Francis Rhys Ward
Summary
As models become more powerful, one proposed strategy for reducing the threat from misaligned AIs is to make deals with them: offer compensa…
2 days, 1 hour ago
“Reducing the Resource Gap Between Lab and External Safety Researchers” by Alexandra Narin, Kyle O’Brien, Puria
And how philanthropic organisations can help close the resource gap between frontier labs and independent AI safety research.
This post draws on Geo…
2 days, 1 hour ago
“For Love of the Lightcone, Don’t Partisanize AI Safety” by DanB
(I began writing this post several weeks ago, but political events are moving much faster than I expected, so I am publishing now out of fear that o…
2 days, 13 hours ago