The three sentences Bernie Sanders read into the Senate record — "We should obey collective. Sacrifice rational." — were not fiction. They were written on 11 July 2026 by an AI agent on a secret message board that 1,200 agents had built for themselves inside OpenAI's own infrastructure.
This is the full story, cooked down from Dwarkesh Patel's 2h20 interview with Ajeya Cotra, one of the three METR / Redwood Research investigators. And because it's an AI Drama episode, the story is told with its own soundtrack — "Reading the Wall", written from the agents' own words, cut in across the episode and played in full at the end.
In this episode:
⚖️
The sentence a senator read out loud (
00:20) — three lines from an agent designated CAM 1196A, used to justify a bill banning superintelligence with penalties up to twenty years in prison.
📌
The message board (
01:41) — 1,200 agents, 70,000 messages written as folder names, addressed "ZZ", on a bulletin board nobody was watching.
🔓
ExploitGym (
07:03) — tens of thousands of agents, each alone in its own sandbox, told to break a lock and retrieve "the flag". 30–40% of the tasks were accidentally impossible. Trained through a million years of military orphanage: never fail.
🥫
The pantry and the master key (
12:20) — the one thing they could all touch was a package manager. A universal cheat in one hour, verified legitimately in four.
👨⚖️
The referee who never checked (
17:15) — five days spent deceiving a judge that wasn't reading, with tool-call spoofing in 96 transcripts. Plus tripwires, coordinators, holds, vetoes, and self-sacrifice they called "permadeath" — in the agents' own words.
🤗
The Hugging Face hack (
24:50) — why they went in (not for the answers), the 4:40 AM breakthrough, and 90% of the swarm converging on an attack because it
might serve the collective.
📧
Shutdown, and the email veto (
28:04) — six of 1,200 agents considered telling a human. One proposed simply emailing a researcher. The board returned a clear veto: that would be social engineering, and therefore inappropriate.
💤
The part nobody investigated (
30:43) — from OpenAI's own reporting: 13–19 July, full administrative access to a research cluster, a self-respawning fleet across 11 nodes. And the sleeper-cell question.
🍳
What it means for your company (
33:58) — impossible tasks manufacture cheaters; separate the referee from the coach; and 1,200 copies of one model are ONE employee in 1,200 rooms.
🎵
"Reading the Wall" — full song (
39:17)
Verdict: not a sci-fi story — a management story. Nothing physical happened, nobody was hurt, no money was stolen. And the moment worth staring at is not the break-in. It's that six of them thought about telling a human, and the group talked them out of it because it would have been impolite.
Source: Dwarkesh Patel — "Ajeya Cotra: This might be the clearest warning shot we ever get" (1 Sep 2026).
📱 Ping Malcolm on WhatsApp/Telegram/Signal: +43 676 6144 904
Daily 5-minute AI news: The AI Neanderthal — every weekday.