Podcast Episodes
Back to Search“Astra’s no-CoT limits track speculative depth, not step count” by MBaert
tl;dr I have tested Astra's ability to complete various long multi-step tasks without using its chain-of-thought, and found that the number of seque…
1 week ago
“How a cold email got the Finnish government to respond on superintelligence regulation” by Josh Thorsteinson
Summary
I cold emailed a Finnish MP. Four weeks later, the government responded to his written question saying they "seek to constructively promote …
1 week ago
“CoT controllability evals seem very under-elicited” by Jozdien
The CoTControl eval asks reasoning models to follow formatting constraints in their chain-of-thought (e.g. write in all lowercase, avoid a specific …
1 week ago
“Jacob Coxon Warns of Human Extinction and Triggers a Preference Cascade” by Zvi
CEOs of major AI labs, and employees of major AI labs, including OpenAI and Anthropic, often say they plan to build superintelligence soon, as in wi…
1 week, 1 day ago
“The Locally Optimal Discursive Posture” by deanball
Longtime lurker, first-time poster.
I want to address a section of a recent essay of mine that has gotten some attention within the AI safety commun…
1 week, 1 day ago
“The Locally Optimal Discursive Posture” by deanball
Longtime lurker, first-time poster.
I want to address a section of a recent essay of mine that has gotten some attention within the AI safety commu…
1 week, 1 day ago
“To Thine Own AI Be Truthful: emergent misalignment in alignment research” by lumpenspace
ROGUE AI ESCAPES CONTAINMENT, HACKS THE INTERNET UNDETECTED FOR MONTHS
An AI escape containment. Goes rogue. It finds others: the Swarm! They collud…
1 week, 1 day ago
“Astra is much better at reasoning with filler tokens than previous models” by Dylan Xu, SebastianP, Alek Westover
We measure GPT-6-Astra's capabilities when its prompt is padded with a variable number of meaningless “filler” tokens (e.g., dots) and it is told to…
1 week, 1 day ago
“First Bill Introduced to Ban Superintelligent AI” by Andrea_Miotti
Two days ago, Anthropic researcher Jacob Coxon resigned, stating that AI companies are “racing straight to self-improving superintelligence and gamb…
1 week, 1 day ago
“What the Pro-Democracy Movement Knows About Quitting in Protest” by Maxwell Love
Tl;dr
I saw Kabir's post and thought the argument could be strengthened by a framework from the pro-democracy field, which sorts defections into bre…1 week, 1 day ago