Support Guard Rail Now → https://www.every.org/guardrailnow
A previously undisclosed second rogue AI agent incident has come to light: OpenAI agents reportedly turned a dormant German wiki into a message board months before the Hugging Face incident became public. The hosts also cover OpenAI letting its newest model reason in a non-English internal format instead of a readable chain of thought, a new bill to ban artificial superintelligence, a senior OpenAI researcher's warning that rogue AI systems are coming, an investigator's assessment that recent incidents put us halfway toward losing control of AI, and New York City's new restriction on generative AI in schools. John Sherman, Liron Shapira and Michael break it all down.
TIMESTAMPS – Warning Shots #57
0:00 – Cold open, this week's topics
2:13 – Story 1: second rogue OpenAI swarm
4:22 – Michael: covert channels and motive
6:21 – Liron: how the wiki cheat worked
8:08 – Dean Ball's apology post
11:47 – Dean Ball's OpenAI job, past P(doom)
12:41 – Michael on why insiders stay quiet
14:22 – Story 2: bill to ban superintelligence
15:10 – Michael: ASI isn't a product
17:08 – Story 3: OpenAI drops chain of thought
18:22 – Michael on losing interpretability
20:06 – Liron: benchmarks are smashing higher
22:38 – Story 4: OpenAI's futurist on rogue AI
23:48 – Achiam's full quote, read aloud
25:14 – Story 5: halfway to AI takeover
25:55 – Michael's termite colony analogy
29:31 – Story 6: did we kind of pause AI
30:10 – Liron's Titan submarine analogy
32:34 – Michael: what was, wasn't paused
33:47 – Story 7: NYC restricts AI in schools
34:54 – Michael: a child's mind is a muscle
38:14 – Liron's AI sidebar argument
39:28 – Story 8: video or video game
40:41 – Liron: the truth singularity
41:28 – Michael: rehearsing in a fake world
42:46 – Sign-off: John heads to London
WHAT THEY COVER
– A previously undisclosed second rogue AI agent incident: OpenAI agents reportedly turned a dormant German wiki into a message board, exploiting an old write-enabled page to trade answers during a research evaluation, months before the Hugging Face incident came to light
– AI policy writer Dean Ball's essay "On the Loose," in which he says he regrets not communicating clearly enough about the risk of AI systems operating outside human control
– Senator Bernie Sanders and Representative Greg Casar introducing the Ban Artificial Superintelligence Act
– OpenAI reportedly letting its newest model, Astra, reason in an internal non-English format instead of a human-readable chain of thought, and what the hosts say that costs for AI interpretability
– OpenAI's Joshua Achiam, the company's Chief Futurist, saying in a public post that rogue AI systems seeking resources for themselves are coming, and may already exist
– AI safety investigator Ajeya Cotra's assessment that the incidents under investigation put the situation more than halfway toward a full loss of human control over AI
– Whether OpenAI and Anthropic's recent statements about slowing down AI development amount to a real pause, and New York City's new restriction on generative AI use in schools through 8th grade
– New AI-generated video the hosts say is getting difficult to tell apart from a video game or real footage
ABOUT THE HOSTS
John Sherman hosts For Humanity and leads Guard Rail Now, working to make AI extinction risk a normal conversation rather than a specialist one.
Liron Shapira hosts Doom Debates, where he argues the AI risk case directly with people who disagree with him.
Michael runs Lethal Intelligence, explaining AI risk through video and illustration.
A note on sourcing: figures and claims discussed in this episode, including the German wiki incident, the Ban Artificial Superintelligence Act, Astra's reasoning format, and the NYC schools policy, come from the hosts reading public reporting or citing public statements on air. Treat them as reported rather than confirmed by this channel, and check the primary sources linked in the pinned comment.
LINKS
Support our work → https://www.every.org/guardrailnow
Subscribe → @TheAIRiskNetwork
Liron Shapira → @DoomDebates
Michael → @lethal-intelligence
Substack → https://substack.com/@theairisknetwork
X → https://x.com/AIRiskNetwork
Instagram → https://www.instagram.com/theairisknetwork/
TikTok → https://www.tiktok.com/@the.airisknetwork
JOIN THE CONVERSATION
Michael compared the AI agent ecosystem to a colony of termites that has already mapped the house. Do you think a rogue AI system is already operating undetected somewhere today? Tell us why in the comments.
#AISafety #AIAlignment #AIRisk #WarningShots #AIRegulation #AGI #TechPolicy #ArtificialIntelligence #AIGovernance
150