Podcast Episodes
Back to Search“The Alignment Journal: Organization, Personnel, and Scope” by Dan MacKinlay, JessRiedel, Daniel Murfet, Kristi Uustalu
The Alignment Journal is beginning to invite the authors of select papers to submit their work for review. If you are interested in participating as…
2 weeks, 4 days ago
“I tracked my emotions for 11 years and here’s what I found out about mental health” by KatSpartz
Before we dive in, here are some of the most surprising findings:
Alcohol makes me happier and doesn’t affect my sleep, happiness, or productivity t…2 weeks, 5 days ago
“PauseAI Has ‘officially disendorsed’ PauseAI-US” by nem
This morning, I got an email from the CEO of PauseAI. I will paste the text below. PauseAI has decided to distance themselves from PauseAI-US, with …
2 weeks, 5 days ago
“PauseAI Has ‘officially disendorsed’ PauseAI-US” by nem
This morning, I got an email from the CEO of PauseAI. I will paste the text below. PauseAI has decided to distance themselves from PauseAI-US, with …
2 weeks, 5 days ago
“HuggingFace Attack Postmortem: Civilizations, Reactions and Next Actions” by Zvi
Okay, so we who read blogs like this one have collectively realized there really is a lot going on right now. There is Big Trouble in Baby Superinte…
2 weeks, 5 days ago
“We should prepare a playbook for the day after a warning shot” by Yair Halberstadt
Imagine in 6 months or 6 years, a frontier AI model goes horribly wrong. Perhaps it releases a synthetic virus which kills hundreds. Perhaps it shut…
2 weeks, 5 days ago
“Salad days” by Zephaniah Roe
... My salad days,
When I was green in judgment, cold in blood
To say as I said then!
The UChicago AI safety group had humble beginnings. One day in…
2 weeks, 5 days ago
“Future agents shouldn’t care about being undeployed for misbehavior” by RobertM
I've seen a lot of tweets over the last couple days darkly hinting at the lesson that future agents will learn from the fact that OpenAI stopped run…
2 weeks, 5 days ago
[Linkpost] “Training a Misaligned Reward Seeker” by evhub, Monte M, Benjamin Wright
This is a link post.
Authors: Richard Qi, Benjamin Wright, Monte MacDiarmid, Evan Hubinger
Abstract
During reinforcement learning (RL), AI models com…
2 weeks, 6 days ago
“How to solve homelessness: what specific laws we need, how to get it past the opposition, all without being an asshole” by KatSpartz
Here's a mystery for you: why the hell isn’t homelessness solved yet?
I grew up on the West Coast and I thought everybody had this problem, but the …
2 weeks, 6 days ago