Podcast Episodes
Back to Search“Should Less Wrong add subtitles?” by Chris_Leong
If Less Wrong wants people to be sharing more of their intellectual output on this website, we should probably be looking at Substack since it proba…
1 month ago
“Three thoughts on civilisational handoff” by Cleo Nardo
What happens when humans put AIs in charge of civilisationally important decisions? A frontier AI company might hand over internal decisions (R&D, s…
1 month ago
“Announcing: Iliad’s New 2026 Fellowships” by David Udell, Alexander Gietelink Oldenziel, Leon Lang
Timelines are short. Given that, the sooner we can onboard people into the alignment field, the better. In that spirit, and in light of our current …
1 month ago
“Q2.5 2026 Timelines Update: Uplift and Revenue” by brendanhalstead, Daniel Kokotajlo, elifland
Tl;dr: Our timelines haven’t changed much (they got slightly shorter) but our modeling and evidence base have noticeably improved, so we feel somewh…
1 month ago
“Does DiffusionGemma do latent reasoning?” by Jan Bauer, Neel Nanda
TL;DR
Google DeepMind's recent model DiffusionGemma (DG) generates text via diffusion, meaning many diffusion steps happen before generating the fi…
1 month ago
“Learning new facts can change LLM behaviour” by Richard Juggins
TL:DR: I use synthetic document fine-tuning to train an LLM to believe that in 2027 ‘long-horizon’ frontier LLMs count as moral persons. I find the …
1 month ago
“Kimi likes causal decision theory more after RL in twin prisoner’s dilemmas” by oakhu
Some multi-agent training set-ups could make language models more sympathetic to causal decision theory (CDT), even in abstract discussion. We give …
1 month ago
“Mom’s Advice For Hosting A Class Reunion” by jenn
Pour more money and effort into them than you think is reasonable. Treasure them, because you can't actually host that many of them and keep expecti…
1 month ago
“AI #181: Astra Goes Cyber Critical” by Zvi
The hacking of HuggingFace by an internal OpenAI model, and more importantly the internal events that led to that and the fallout from it, remain th…
1 month ago
“Rerunning AI safety papers on every frontier release would be pretty easy and valuable” by Zephaniah Roe, hersheys, yix
tl;dr: Some important AI safety research is never rerun on the newest models. There are probably cases where this would be valuable and a single wel…
1 month ago