Podcast Episodes
Back to Search“Twenty Years from RSI to Takeoff: Slow Learning, Scaling Slowdown, Industrial Explosion” by Vladimir_Nesov
Industrial explosion is what will make the next-model building loops (and thus learning) with LLMs 1000 times faster by about 2050, if indeed the sl…
3 weeks, 6 days ago
“Llama will abandon a correct answer if it thinks you’re educated” by Nick Merrill
TLDR: Given this exchange:
User: Janet's ducks lay 16 eggs per day. She eats three for breakfast every morning and bakes muffins for her friends eve…
4 weeks ago
″“Farm strength” vs “breath awareness”” by jimmy
If you want to become physically strong, the default solution to this problem is to go lift weights. The idea is that you can challenge your muscles…
4 weeks ago
“When is Unlimited Optimization Catastrophic?” by Winter Cross
This post discusses research I've completed along with my colleagues Leo Cymbalista, Alfred Harwood, and Jose Faustino at Dovetail Research. Most of…
4 weeks, 1 day ago
“Selection for Selectability: Inductive Biases in Evolution and in Neural Networks” by CarolusRenniusVitellius
This post was written as part of MATS 9.1 under the mentorship of Richard Ngo, and was written during Iliad Fellowship, to all of whom my thanks.
L…
4 weeks, 1 day ago
“AI #182: Pause For Reflection” by Zvi
This was a week of quiet aftermath, an opportunity to process recent events and start to figure out the path forward.
OpenAI is attempting to turn …
4 weeks, 1 day ago
“AI Text Watermarking Is Free And Good” by Zvi
Scott Aaronson, while working at OpenAI, largely solved AI text watermarking together with Hendrik Kirchner.
Here is how his solution works, or see…
4 weeks, 1 day ago
“Evaluating Explanations of LLM Behavior In The Wild with Counterfactual Experiments” by Adam Karvonen, Euan Ong, Subhash Kantamneni, Sam Marks
TL;DR: We introduce CHIVE, an agentic pipeline that discovers unexpected LLM behaviors in the wild and explains them with counterfactual prompt edit…
4 weeks, 1 day ago
“Misaligned AI in the Bronze Age” by frmsaul
The first artificial intelligence was booted up around 4000BC in southern Iraq. It seems to have begun as something like a bank, a temple pooling gr…
4 weeks, 1 day ago
“When models identify as a swarm” by julius vidal
tldr: the word 'swarm' is associated with emergent collective intelligence, but also stupid or destructive behaviour. LLM self identity matters, so …
4 weeks, 1 day ago