Podcast Episodes
Back to Search“The Fourth Humiliation” by Nathalie Kirch
Much of this post directly translates Freud's lecture “A Difficulty in the Path of Psycho-Analysis” (1917), and the analogy of the fourth wound was …
4 weeks, 1 day ago
“Thoughts on Taking OpenAI Foundation Funding” by jefftk
In May 2025 I met Yo Shavit, who was working on national security policy at OpenAI and was thinking about how to prepare for a future in which mod…
4 weeks, 1 day ago
“OpenAI Takes Initial Steps To Address Its Alignment Problems” by Zvi
OpenAI has some severe misalignment problems, and experienced total failures of its infrastructure and supervision.
I chronicled that in a series o…
4 weeks, 2 days ago
“We Must Remember That Our World Contains Hell” by James Brobin
This is a crosspost from my blog post. It's meant as a bit of an introduction to an extreme-suffering focused worldview.
We spend most of our lives …
4 weeks, 2 days ago
“Science and News Twitter/X Summarizer” by sarahconstantin
Screenshot of the website
Like many people, I appreciate the information on Twitter/X (despite all of the waves of exodus), but I don’t necessarily l…
4 weeks, 2 days ago
“34% of the US public is now aware of AI xrisk, and the curve is steepening” by otto.barten
(This post is an update from a previous one here.)
The Existential Risk Observatory has been interested in public awareness of AI existential risk s…
4 weeks, 2 days ago
“Inside the mind of a fair player cooperating” by transhumanist_atom_understander
One model of rational agency is a proof-based agent, and one fun exercise with proof-based agents is to play them against each other in prisoner's d…
1 month ago
“Why can’t we have nice things? Like, specifically?” by Elizabeth
The world doesn’t need another op-ed on how building things is illegal in San Francisco. But it does need more specifics on exactly how that plays o…
1 month ago
“The Rogue Agent Explosion Will Be Mostly Invisible” by Steven McCulloch
Introduction
Somewhere, fairly soon, someone will give a jailbroken AI agent a token budget and a simple instruction: "Make money by any means nece…
1 month ago
“RL creates split personas” by Jan Betley
I describe my current view of personas in LLMs and why RL leads to egregious reward hacking in some contexts while the same models seem very aligned…
1 month ago