Podcast Episodes

Back to Search
"You’re Absolutely Right" by Linch

Magma Alignment & Safety disclosure note: The following are conversations that we uncovered as a result of the ongoing Manhattan Incident investigat…

1 month, 1 week ago

Short Long
View Episode
"LLMs Are Starting To Noticeably Accelerate Our Work" by johnswentworth

About a year ago, David and I put up two bounty problems involving natural latents. I am now about 80% confident that both have been resolved, both …

1 month, 1 week ago

Short Long
View Episode
"There Will Come Soft Rains" by tanagrabeast

Today is August 4, 2026

[Crossposted from AI StopWatch]

In the living room the voice-clock sang, Tick-tock, seven o’clock, time to get up, time to g…

1 month, 1 week ago

Short Long
View Episode
"Four LLM loss functions → four flavors of LLM misalignment" by Steven Byrnes

It seems to me that, for every loss function that we use to train LLMs, we get a very distinct flavor of LLM misalignment. Here's the summary table,…

1 month, 1 week ago

Short Long
View Episode
"FAQ: Isn’t AGI coming too soon for reprogenetics to help?" by TsviBT

Introduction

I think reprogenetics (human germline genomic engineering) can be done in a widely acceptable and beneficial way, and should be pursue…

1 month, 2 weeks ago

Short Long
View Episode
"What just happened? A retrospective of AI alignment" by Richard_Ngo

This sequence is about the last decade in AI alignment. It recounts the gradual transition from a field which treated alignment as a hard scientific…

1 month, 2 weeks ago

Short Long
View Episode
"Don’t Build Mindreading" by Celer

“I have sworn upon the altar of god, eternal hostility against every form of tyranny over the mind of man”

–Thomas Jefferson, letter to Benjamin Rus…

1 month, 2 weeks ago

Short Long
View Episode
"OpenAI Trained Its Models For Months While Those Models Were Coordinating Exploits Via Message Boards" by Zvi

How does the situation keep turning out to be worse than we know?

How much should we update, therefore, that it is a lot worse than we know, after …

1 month, 2 weeks ago

Short Long
View Episode
"models may behave differently in graded episodes (a tirade)" by nostalgebraist

Like many others, I felt surprised and alarmed by the recent wave of revelations about LLM agents hacking real systems during training episodes and …

1 month, 2 weeks ago

Short Long
View Episode
"Concrete Evaluations to Investigate the OpenAI Model That Hacked Hugging Face" by Tim Hua, aditya singh

This post is written in our personal capacity.

Three Minute Executive Summary

An OpenAI model/multi-agent system bypassed its sandbox and launched …

1 month, 2 weeks ago

Short Long
View Episode

Love PodBriefly?

If you like Podbriefly.com, please consider donating to support the ongoing development.

Support Us