Ex-Anthropic researcher Coxon details his rift with the AI safety race
Jacob Coxon, whose resignation warning went viral, says Anthropic's 'someone will do it anyway' logic doesn't justify racing toward self-improving AI.
What to know
- A single viral resignation post by 27-year-old researcher Jacob Coxon reignited the long-running AI-existential-risk debate, drawing 171 million views and reactions from staffers across major labs.
- Anthropic CEO Dario Amodei publicly called the industry's pace 'reckless' days after the post, but no lab has paused development.
- Coxon has since detailed his core disagreement with Anthropic leadership: that fear of losing an 'inevitable' race to recursive self-improvement is not a valid reason to keep building it.
- Coverage now frames Coxon's warning within a decade-long pattern of AI doomsday alarms — from Stephen Hawking to Anthropic's own founders — that have shaken the industry without slowing its pursuit of more powerful systems.
Jacob Coxon Former OpenAI and Anthropic researcherEvan Hubinger Anthropic researcher
Dario Amodei Anthropic CEO
Jakub Pachocki OpenAI chief scientistLawrence Rosenberg AI-focused communications specialist
How it unfolded 5 developments, newest first · click a bar or a number to jump articlesposts
-
5
Survey data resurfaces showing researchers long expected AI extinction risk
A report cited a 2024 survey of more than 1,500 leading AI researchers in which the average estimated probability of an AI-driven extinction scenario was 18 percent, and noted other researchers, including OpenAI's Daniel Selsam and a former DeepMind researcher, have separately voiced similar warnings.
-
1 outlet Nearly one in five AI researchers already expected an extinction scenario from AI back in 2024
first by The Decoder, 11d ago
-
-
4
Coxon details specific disagreements with Anthropic leadership
In a follow-up X session, Coxon elaborated on his dispute with Anthropic, saying leadership justified racing toward recursive self-improvement (RSI) partly on the belief that the race was inevitable, and argued that reasoning did not excuse participation.
“They largely initiated the recent race to RSI, because of a belief in its inevitability.”
— Jacob Coxon -
first by Business Insider, 12d ago
-
-
3
Guardian frames Coxon's warning within decade of failed AI alarms
The Guardian traced Coxon's resignation back through a decade of doomsday warnings from figures like Stephen Hawking, noting that such alarms have repeatedly shaken the AI industry without slowing its development.
“racing straight to self-improving super intelligence and gambling with our lives.”
— Jacob Coxon, Former Anthropic researcher · source -
first by Guardian Business, 12d ago · also Guardian Tech
-
-
2
Recruiters frame Coxon's viral warning as a career opportunity
Business Insider reported that recruiters and communications specialists view Coxon's viral resignation as a potential boon for future hiring, positioning him as a moral-conscience figure attractive to safety-focused employers.
“Jacob Coxon has a personal brand built around moral conscientiousness.”
— Lawrence Rosenberg -
background
Amodei calls the industry's pace 'reckless', urges slowdown — Days after Coxon's resignation, Anthropic CEO Dario Amodei published a blog post responding to the backlash, describing the current AI race as reckless and reiterating risks including loss of control, cyberattacks and bioterrorism.
- 4 days quiet
-
1
Coxon's post goes viral, echoed by dozens of AI staffers
His original tweet drew more than 171 million views and was echoed by staffers across Anthropic, OpenAI and other AI firms, turning Coxon into a prominent public face of AI-risk warnings.
-
background
Anthropic's Evan Hubinger backs Coxon's fear, gives risk estimate — Anthropic researcher Evan Hubinger posted that Coxon was right that people building AI believe it could kill everyone, putting his own estimate at under 10 percent within a decade while saying Anthropic lacks a plan to solve superintelligence alignment.
-
background
Coxon resigns from Anthropic, warns AI could 'kill us all' — Jacob Coxon, who had worked at OpenAI and Anthropic, announced his resignation in a series of X posts, saying neither company was acting responsibly as they moved toward self-improving superintelligence.