Podcast Episodes
Back to Search“Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident” by ryan_greenblatt, Ajeya Cotra, Hjalmar_Wijk
We recently published the report from our brief independent investigation into this incident. You can read the full report here.
Here is our tweet …
3 weeks, 4 days ago
“Against Modesty’s Bailey” by Zvi
Modesty arguments often say that you should mostly or entirely bow to ‘expert consensus’ or the views of particular others, and who are you to disag…
3 weeks, 4 days ago
″“So You Don’t Trust Me?”” by Zack_M_Davis
One of my favorite passages from Atlas Shrugged is this one, when Cherryl is beginning to have second thoughts about her marriage to James Taggart:
…3 weeks, 4 days ago
“When There Are No Experts” by J Bostock
An average person in the western world probably believes a lot of false things. They probably don't have a great grasp of political economy, or orbi…
3 weeks, 4 days ago
“The American People Really Hate Data Centers” by Zvi
There are at least five different core questions around data centers and their politics.
In what ways are specific concerns people raise about da…3 weeks, 4 days ago
“On Writing #3” by Zvi
Periodically I like to gather various observations about writing, and share my perspective. Last time was in honor of my trip to Inkhaven. This time…
3 weeks, 4 days ago
“The Forkmakers” by Mikewins
Imagine our civilization fell tomorrow. What would our descendants think of us? What would they know about the 21st century?
They would know surpris…
3 weeks, 5 days ago
“PSA: We can do better” by hersheys, Kaustubh Kislay
tl;dr: people should understand and think hard about the problems they work on.
We’ve observed that those who work in AI safety (ourselves included)…
3 weeks, 5 days ago
“AI Safety Acculturation is Neglected” by jenn
At the local AI safety co-working space, there are ~two kinds of regulars.
There's the kind of regular who's been thinking seriously about AI safety…
3 weeks, 6 days ago
“LLMs could control their host machines by exploiting inference engines” by beyarkay (Boyd Kane)
Large language models often take actions running on one computer (via an agentic harness such as Claude Code or Codex), however the LLMs’ responses …
3 weeks, 6 days ago