Podcast Episodes

Back to Search
"Natural Language Autoencoders Produce Unsupervised Explanations of LLM Activations" by Subhash Kantamneni, kitft, Euan Ong, Sam Marks

Abstract

We introduce Natural Language Autoencoders (NLAs), an unsupervised method for generating natural language explanations of LLM activations. …

4 months, 2 weeks ago

Short Long
View Episode
[Linkpost] "Interpreting Language Model Parameters" by Lucius Bushnaq, Dan Braun, Oliver Clive-Griffin, Bart Bussmann, Nathan Hu, mivanitskiy, Linda Linsefors, Lee Sharkey

This is a link post. This is the latest work in our Parameter Decomposition agenda. We introduce a new parameter decomposition method, adVersarial Pa…

4 months, 2 weeks ago

Short Long
View Episode
"It’s nice of you to worry about me, but I really do have a life" by Viliam

I have two shameful secrets that I probably shouldn't talk about online:

I love my family.I enjoy my hobbies. "What an idiot!" you probably think. "…

4 months, 2 weeks ago

Short Long
View Episode
"Irretrievability; or, Murphy’s Curse of Oneshotness upon ASI" by Eliezer Yudkowsky

Example 1: The Viking 1 lander

In the 1970s, NASA sent a pair of probes to Mars, Viking 1 and Viking 2 missions, at a total cost of 1 billion dollar…

4 months, 2 weeks ago

Short Long
View Episode
"Dairy cows make their misery expensive (but their calves can’t)" by Elizabeth

How much do cows suffer in the production of milk? I can’t answer that; understanding animal experience is hard. But I can at least provide some fac…

4 months, 2 weeks ago

Short Long
View Episode
"Takes from two months as an aspiring LLM naturalist" by AnnaSalamon

I spent my last two months playing around with LLMs. I’m a beginner, bumbling and incorrect, but I want to share some takes anyhow.[1]

Take 1. Ever…

4 months, 2 weeks ago

Short Long
View Episode
"Intelligence Dissolves Privacy" by Vaniver

The future is going to be different from the present. Let's think about how.

Specifically, our expectations about what's reasonable are downstream o…

4 months, 3 weeks ago

Short Long
View Episode
"How Go Players Disempower Themselves to AI" by Ashe Vazquez Nuñez

Written as part of the MATS 9.1 extension program, mentored by Richard Ngo.

From March 9th to 15th 2016, Go players around the world stayed up to wa…

4 months, 3 weeks ago

Short Long
View Episode
"On today’s panel with Bernie Sanders" by David Scott Krueger

It's sort of easy to forget how close Bernie Sanders was to becoming the most powerful person in the world. The world we live in feels so much not l…

4 months, 3 weeks ago

Short Long
View Episode
"Not a Paper: “Frontier Lab CEOs are Capable of In-Context Scheming”" by LawrenceC

(Fragments from a research paper that will never be written)

Extended Abstract.

The frontier AI developers are becoming increasingly powerful and w…

4 months, 3 weeks ago

Short Long
View Episode

Love PodBriefly?

If you like Podbriefly.com, please consider donating to support the ongoing development.

Support Us