Podcast Episodes
Back to Search“What’s Going on With OpenAI’s Messaging?” by ozziegoen
This is a quickly-written opinion piece, of what I understand about OpenAI. I first posted it to Facebook, where it had some discussion.
Some argum…
2 years, 2 months ago
“Language Models Model Us” by eggsyntax
Produced as part of the MATS Winter 2023-4 program, under the mentorship of @Jessica Rumbelow
One-sentence summary: On a dataset of human-written essa…
2 years, 2 months ago
Jaan Tallinn’s 2023 Philanthropy Overview
This is a link post.to follow up my philantropic pledge from 2020, i've updated my philanthropy page with 2023 results.
in 2023 my donations funded $4…
2 years, 2 months ago
“OpenAI: Exodus” by Zvi
Previously: OpenAI: Facts From a Weekend, OpenAI: The Battle of the Board, OpenAI: Leaks Confirm the Story, OpenAI: Altman Returns, OpenAI: The Board…
2 years, 2 months ago
DeepMind’s ”Frontier Safety Framework” is weak and unambitious
FSF blogpost. Full document (just 6 pages; you should read it). Compare to Anthropic's RSP, OpenAI's RSP ("PF"), and METR's Key Components of an RSP.…
2 years, 2 months ago
Do you believe in hundred dollar bills lying on the ground? Consider humming
Introduction.
[Reminder: I am an internet weirdo with no medical credentials]
A few months ago, I published some crude estimates of the power of nitr…
2 years, 2 months ago
Deep Honesty
Most people avoid saying literally false things, especially if those could be audited, like making up facts or credentials. The reasons for this are …
2 years, 2 months ago
On Not Pulling The Ladder Up Behind You
Epistemic Status: Musing and speculation, but I think there's a real thing here.
1.
When I was a kid, a friend of mine had a tree fort. If you've neve…
2 years, 2 months ago
Mechanistically Eliciting Latent Behaviors in Language Models
Produced as part of the MATS Winter 2024 program, under the mentorship of Alex Turner (TurnTrout).
TL,DR: I introduce a method for eliciting latent be…
2 years, 2 months ago
Ironing Out the Squiggles
Adversarial Examples: A Problem
The apparent successes of the deep learning revolution conceal a dark underbelly. It may seem that we now know how to…
2 years, 2 months ago