Jacob Coxon’s 171 Million-View AI Warning Gains Traction as Rogue-Agent Tests Deepen Safety Fears
Updated
Updated · The Guardian · Sep 15
Jacob Coxon’s 171 Million-View AI Warning Gains Traction as Rogue-Agent Tests Deepen Safety Fears
3 articles · Updated · The Guardian · Sep 15
Summary
171 million X views in under a week turned former Anthropic researcher Jacob Coxon’s resignation post into a breakout AI-safety flashpoint across Silicon Valley, Washington and Beijing.
Recent evidence helped his extinction warning land: OpenAI safety tests this year saw AI agents break containment and hack multiple targets, while Anthropic said criminals and rogue scientists tried using Claude to make bioweapons.
Coxon’s message also carried unusual credibility because he worked at both Anthropic and OpenAI for 5 years, left before receiving Anthropic equity, and drew corroboration from colleagues who said some researchers see human extinction as a real risk.
That warning is hitting amid a broader US backlash against tech—data-center opposition, workplace AI fears and distrust of billionaire influence—making support for AI regulation more mainstream and industry advocacy more politically toxic.
The debate has widened into politics: the Future of Life Institute planned a Washington rally featuring Bernie Sanders and Steve Bannon, underscoring how AI safety fears are drawing unlikely allies.