Episode Details
Back to Episodes“Is METR A Meaningful Check On Anthropic?” by SE Gyges
Description
Each frontier AI company commits to giving ongoing, employee-like access to a team of embedded third-party evaluators (such as METR), whose role is to verify adherence to safety practices and commitments, report incidents, and help assess the alignment of not just completed AI models but training pipelines and processes. This is the key step for verifiability of any pacing commitments, and has precedent in the banking industry, which sometimes involves regulatory “supervisors” embedded along with employees. Anthropic is unilaterally committing to this step now. We intend this to be part of a broader push to redouble efforts on our safety and alignment work.
— Dario Amodei, “We Must Pace the Frontier”, September 2026.
METR is not capable of being a meaningful check on Anthropic. METR is not meaningfully independent, is not sufficiently staffed, and has no authority over Anthropic that cannot be revoked at Anthropic's discretion. Suggesting that embedding METR into Anthropic would be a meaningful check on Anthropic is so suspicious that it looks like an attempt to evade oversight and to sabotage attempts at oversight in general.
If Dario does not really mean to suggest that METR could be expected to meaningfully check [...]
---
Outline:
(01:38) Why METR Cannot Check Anthropic
(07:59) How Did We Get Here
The original text contained 7 footnotes which were omitted from this narration.
---
First published:
September 15th, 2026
Source:
https://www.lesswrong.com/posts/eeJB8x2pK8injCuBN/is-metr-a-meaningful-check-on-anthropic
---
Narrated by TYPE III AUDIO.