Q&A with AI researcher Jacob Coxon, who quit Anthropic, on the need for industry-wide, international coordination to limit recursive self-improvement, and more
How Dangerous Is AI?Dmitri Bolt /Townhall:The Story Behind the Resignation of an Anthropic AI Researcher Just Got a Lot More ComplicatedIan Bogost /The Atlantic:Everybody Wants AI to Destroy the WorldMax Burns /MS NOW:Technology shouldn't be a threat to humanity. Congress must act.Omar Gallaga /CNET:What an Ex-Anthropic Researcher's Warning About Human Extinction Really MeansRocket Drew /The Information:What Anthropic Doomsayer Jacob Coxon SawNBC News:Anthropic safety researcher resigned with a
WiredMaxwell Zeff
Context & Ripple Effects
Coxon’s departure followed his public warning that AI companies are racing toward systems they may not control; his decision to leave the industry gave that critique a named insider rather than an abstract safety argument. He left Anthropic after four months and before his equity vested, according to Axios coverage.
The interview also arrives as Anthropic’s Alignment Science lead put the risk debate in unusually stark terms with a more-than-10% estimate of existential risk. Anthropic’s reported disruption of model-assisted research that could aid biological weapons development gives the coordination argument a concrete dual-use governance context.
First-order effects
Anthropic loses a safety researcher while Coxon uses his departure to press frontier AI labs and international actors for limits on recursive self-improvement.
Coxon’s warning moves the question from an individual lab’s safety posture to whether shared cross-border restrictions are needed for the most capable systems.
Second-order effects
Anthropic and other frontier labs face greater pressure to explain how their internal safeguards relate to enforceable, industry-wide limits rather than voluntary company-specific commitments.
Dual-use incidents such as Anthropic’s disrupted biology-related work become part of the case for operational rules that can be applied across developers.
Third-order effects
If labs and governments pursue common limits, frontier-model governance would shift toward coordination mechanisms designed to constrain capabilities that individual firms cannot credibly police alone.
The debate also sharpens a structural divide between rapid model development and safety governance: public dissent from technical staff can make the adequacy of lab-led oversight a central competitive and policy issue.
The trend: Frontier AI governance is moving from firm-level safety assurances toward demands for cross-lab and international constraints on high-risk capabilities.
Former OpenAI and Anthropic researcher Jacob Coxon, who publicly resigned from Anthropic yesterday with a stark warning about AI safety, was interviewed by Anderson Cooper on his show this evening.
This is why it's a good thing that American labs are in the lead. If Chinese labs were, their employees wouldn't be able to warn us about threats. But Chinese labs will encounter the same threats within a year, in complete silence.
A lot of otherwise smart people on Twitter seem 100% convinced AI risks are all fake and stupid and part of some marketing ploy. What is surprising is that some of these people seemingly also believe that AI's positive uses are on an incredible trajectory of increasing capability…
Shame. Such a horrible take, given the open spirit in which Chinese AI labs operate, sharing their knowledge and advancing scientific knowledge of the world. Xenophobia and ethnic stereotypes will destroy the world more surely than AI.
“I don't really care about science fiction... We need to actually talk about... what's actually happening with the agent swarms” is the most perfect encapsulation of the vibes of Q3 2026 I have seen
Not to my knowledge (not sure whether that's because similar attacks haven't happened or because they weren't disclosed). I'm also not sure how open American frontier labs would have been if HF hadn't disclosed first. We're now learning about cyberattacks from months ago that wer…
So? I know you have more followers than me even if I don't know exactly how many followers you have, that doesn't necessarily feel like a contradiction. Anyway, if you're talking specifically about cyber-attacks, then I probably agree with you (but again, feels like it's mostly b…
It's absolutely ridiculous that some people are suggesting that out of these two, the one who should be distrusted because his claims are in his supposed self-interest is Coxon
AI researcher Jacob Coxon announced his resignation from Anthropic over safety fears on Tuesday in an X post that went viral and sent shock waves through Silicon Valley and beyond. Read our interview with Coxon. 👇 https://www.wired.com/...
Jacob Coxon talks to WIRED about the “mini Manhattan project” inside Anthropic, the problem with alignment, and why AI labs have just a few years left to make their systems safe. https://www.wired.com/...
“Big A.I. companies continue to develop technology they concede might cause human extinction — arguing that if they don't develop it, somebody else will.” …
one recurring theme in the coverage of this AI Anthropic doomsday newscycle is nobody actually has any REAL ideas of what to do — just a lot of ambiguous hand-waving in a country that's too corrupt to have working regulators, by a press that doesn't realize that's a central pil…
not trying to make fun of this guy because i think he's sincere and it's reasonable to float concerns about AI, but in this case the AI better have a survival strategy for not having people around to, like, fix the pipes in the data center www.wired.com/story/anthro... [image]
I think there's an arrogance to assuming we're at the Manhattan project yet; I think we're at the era where alchemists became chemists, we're discovering mustard gas and fertilizer. Slow, painful, traumatic death, not quite nuclear Holocaust [embedded post]
I am not anti-AI but I'm getting sick of reading this doomerism and worse writing about it mostly because no one is offering any specifics I can grasp about how the “endgame” will come about. [embedded post]
WIRED spoke with the Anthropic engineer who very loudly quit over concerns about runaway AI. He thinks humanity has a “couple of years” to prepare. from @mzeff.bsky.social
The great @mzeff.bsky.social has a fantastic interview with Jacob Coxon, who quit Anthropic and shared a massively viral X post about why: — www.wired.com/story/anthro...
COXON: “.. The consensus is that the next year or two is crunch time for humanity. .. These are actually just literal quotes from my colleagues .. — .. From their perspective, this is when Anthropic and its competitors decide the fate of humanity. @wired.com — www.wired.com…