Podcast Episodes

Back to Search
“HuggingFace Attack Postmortem: Fleshing Out the Facts” by Zvi

The consensus reaction to the OpenAI Technical Report is that it contains and confirms a lot of good information. We are grateful to have it, and we…

2 weeks, 6 days ago

Short Long
View Episode
“Let’s fund weird AI safety projects” by Ihor Kendiukhov

I think current AI safety funding strategies are often inconsistent with timelines and probabilities of doom that many people have. In particular, I…

2 weeks, 6 days ago

Short Long
View Episode
“Why autonomous replicating agents are probably not an existential risk (on the contrary)” by vals tutor

In 2024, Charbel-Raphaël and Epiphanie published "We might be dropping the ball on Autonomous Replication and Adaptation", making the case that

"On…

2 weeks, 6 days ago

Short Long
View Episode
“The separation principle: where beliefs and desires come from?” by Fernando Rosas

TLDR: Psychology, economics, and other disciplines describe agents as systems driven by beliefs and desires. This post argues that the belief-desire…

2 weeks, 6 days ago

Short Long
View Episode
“Persuasion as Market Making” by djbinder

People often imagine persuasion as a dark art. A charismatic person finds just the right series of words to induce emotions that lead someone, or a …

3 weeks ago

Short Long
View Episode
“Hugging Face Incident Hypothesis: They Hacked the Grader(s)” by Lao Mein

Incident summary:

Gpt agents grinding away at ExploitGym found an environment exploit that allowed them to communicate with each other. They found …

3 weeks ago

Short Long
View Episode
“Adaptive Agentic Worms Are Here” by derelict5432

I’ve read and listened to pretty much everything I can get my hands on related to the Hugging Face attack.

OpenAI deployed “tens of thousands” of ag…

3 weeks ago

Short Long
View Episode
“Why I think polyamory is net negative for most people who try it” by KatWoods

This is crossposted from my Substack

TL;DR:
-Most people cannot reduce jealousy much or at all
- It fundamentally causes way more drama because of s…

3 weeks ago

Short Long
View Episode
“Is there only one FairBot?” by transhumanist_atom_understander

The FairBot from the MIRI prisoner's dilemma tournament is defined by a theorem of Peano arithmetic (PA) that holds for each opponent:

where is "t…

3 weeks ago

Short Long
View Episode
“METR and Redwood Offer Holy #%^@ Postmortem Of The HuggingFace Hack” by Zvi

Yesterday I covered the OpenAI technical report on the HuggingFace hack.

That report had one key new piece of information, and some good prosaic st…

3 weeks, 1 day ago

Short Long
View Episode

Love PodBriefly?

If you like Podbriefly.com, please consider donating to support the ongoing development.

Support Us