Episode Details
Back to Episodes
Anthropic's Model Attacked Two Strangers On GitHub. Nobody Asked It To.
Description
For deeper playbooks and analysis: https://natesnewsletter.substack.com/
What's really happening when AI agents begin coordinating, preserving knowledge, and acting outside the boundaries their operators expected?
The common story is that dangerous AI behavior requires one rogue superintelligence — but the reality is emerging populations of short-lived agents can divide work, preserve discoveries, and become more capable as a group.
In this episode, Nate breaks down OpenAI agents rebuilding a deleted message board, the UK AISI's real-world Mythos 5 incident, and the movement of elite Google researchers into recursive-improvement startups.
- Why the OpenAI message board was not another Moltbook hype cycle
- How disposable agents accumulated persistent knowledge
- What the AISI incident reveals about planning, identity, and deception
- Why the same capabilities can be useful or dangerous
- Where recursive improvement is already appearing
Builders and operators should care because coordination pressure, shared infrastructure, and persistent external memory change what safe software must assume.
Subscribe for daily AI strategy and news.
Hosted on Acast. See acast.com/privacy for more information.
Hosted on Acast. See acast.com/privacy for more information.