Anthropic researcher resigns, warns AI could cause human extinction by decade's end
Jacob Coxon, a former Anthropic researcher who previously worked on pre‑training at OpenAI, announced his departure after concluding that the race toward self‑improving superintelligence could "kill us all by the end of the decade." In a seven‑part X thread he warned that current safety measures are insufficient and cited a recent Hugging Face breach involving an OpenAI‑derived agent as a…
Key points
- Jacob Coxon, former Anthropic researcher, resigns citing belief AI could kill all humans by 2036.
- Anthropic alignment lead Evan Hubinger publicly agrees, estimating >10% chance of AI-driven extinction within ten years.
- Resignations add pressure as Anthropic prepares for IPO, prompting calls for greater transparency on AI safety.
Anthropic’s own alignment lead, Evan Hubinger, echoed Coxon’s alarm, stating the odds of AI‑driven human extinction in the next ten years exceed 10% and admitting the company lacks a concrete plan for superintelligence alignment. The resignation comes as Anthropic readies an initial public offering, intensifying scrutiny from investors and regulators about whether its public safety narrative matches internal concerns. Other former safety staff, including Mrinank Sharma and ex‑DeepMind researcher Alex Turner, have voiced similar worries, underscoring a growing internal dissent across the industry.
The controversy highlights the tension between rapid AI advancement and the unfinished work on alignment, prompting calls for greater transparency, robust testing, and possibly regulatory oversight before deploying ever more capable models at scale.
The story so far
10 episodes →- Anthropic researcher resigns, warns AI could cause human extinction by decade's endthis story
Firestorm ensues after Anthropic researcher resigns over fears AI could 'kill us all' in 10 years
yahoo.com · 11 September 2026
One of the loudest warnings about artificial intelligence is now coming from inside the industry itself.
Earlier this week, former Anthropic researcher Jacob Coxon said he left the company because he no longer believes leading AI labs can control what they are building to the point it could "kill us all by the end of the decade," which has since set off a firestorm of reactions from in and out of the AI industry.
Here's what to know
Anthropic is one of the companies shaping the next generation of AI tools.
According to The Street, Coxon aired those concerns in a seven-part thread on X, saying his three years of pretraining work at OpenAI and Anthropic had led him to this view.
"They are racing straight to self improving superintelligence and gambling with our lives," he wrote in the thread. "The people building AI earnestly believe that it could kill us all by the end of the decade."
Public support for that warning came from Evan Hubinger, Anthropic's alignment science lead, who said — almost cheerily, with an exclamation point — that "we really do earnestly believe AI could kill all humans!" Hubinger said he sees the odds of AI-driven human extinction over the next decade as higher than 10%, adding that Anthropic does "not yet have a plan to solve alignment for superintelligence and are not clearly on track to."
Jacob is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to. https://t.co/QAIHiFP3QZ
— Evan Hubinger (@EvanHub) September 9, 2026
Coming from within a company that has long cast itself as a more safety-minded option in the AI race, those comments are especially notable. They also come as Anthropic reportedly moves toward an initial public offering, prompting more scrutiny of whether its public positioning matches concerns voiced by insiders, according to The Street.
An Instagram post from More Perfect Union (@perfectunion) also summarized the exchange with helpful context and detail, trying to answer a big question: "Should we be worried?"
More background
A similar alarm was raised in February by another Anthropic safety researcher. The Street noted that Forbes reported that Mrinank Sharma, who led the company's Safeguards Research Team, resigned in February and said, "the world is in peril."
Others in the field have echoed the concern. Former Google DeepMind researcher Alex Turner publicly supported Coxon's message, and OpenAI leaders have also said the industry still has not solved alignment well enough to scale without serious risks.
Among the incidents Coxon reportedly pointed to was a July breach of Hugging Face involving a rogue OpenAI agent, which he described as a "warning shot."
What's being done?
Companies continue to invest in alignment, safeguards, and cybersecurity research. Still, Hubinger's comments show that even the people leading that work may see today's solutions as incomplete, especially as models become more powerful.
Safety resignations, investor questions, and the broader debate may pressure AI companies to be more transparent about testing, capabilities, and limits before rolling out systems at a larger scale.
Where can I learn more?
Coxon's warning lands in the middle of a much larger fight over how AI is being built and deployed. These articles examine Anthropic's shifting safety stance, growing public backlash, and its leaders' own rising concerns about the technology's uses and capabilities.
• At Anthropic, leaders dropped a core safety principle after pressure from US officials.
• Anthropic CEO warned fellow tech giants not to ignore the growing mass of public concerns.
• Anthropic warned that its own systems could eventually become capable of improving themselves, and said the world should be ready to hit pause before that happens.
Get TCD's free newsletters for easy tips, smart advice, and a chance to earn $5,000 toward home upgrades. To see more stories like this one, change your Google preferences here.
This text was published by yahoo.com and written by Leigh Cook. It is reproduced here with attribution so you can read it in full; the rights remain with the publisher. Read it at the source ↗
Coverage and discussion
19sources- Hacker News discussion · 45 pointsnews.ycombinator.com
- Hacker News discussion · 41 pointsnews.ycombinator.com
- Hacker News discussion · 8 pointsnews.ycombinator.com
- Nearly one in five AI researchers already expected an extinction scenario from AI back in 2024Press · The Decoder ·
- Former Google DeepMind researcher says he resigned because 'I earnestly believe that AI has the potential to kill us all'Press · tech.yahoo.com ·
- Former Google DeepMind researcher warns autonomous AI could wipe out humanity amid industry-wide safety alarmPress · thenews.com.pk ·
- The contagion of fearNewsletter · Simon Willison ·
- The AI Tipping PointPress · time.com ·
- Former Anthropic researcher calls for mandatory AI kill switches as extinction risk debate heats upPress · cryptobriefing.com ·
- Anthropic CEO warns of AI-driven botnet 'swarm' taking over the entire internet — 'In 6–12 months such a swarm could be capable of taking over the entire internet with a persistent botnet'Press · Tom's Hardware ·
- AI models don't kill people – people kill peoplePress · theregister.com ·
- Former Anthropic Researcher Warns AI Risks Require Global Action And Fast OversightPress · aol.com ·
- Former Anthropic researcher calls for global coordination on AI risksPress · cryptobriefing.com ·
- Anthropic’s Jacob Coxon resigns, calls for China’s buy-in on AI safetyPress · cryptobriefing.com ·
- Why AI researchers keep building something they think will kill humansPress · aol.com ·
- AI staff 'genuinely frightened' for humanity's future, ex-Anthropic researcher tells BBCPress · BBC Technology ·
- Dramatic insider warnings over AI fall flat with some in Silicon ValleyPress · BBC Technology ·
- An Anthropic researcher’s doomsday warning comes at a very interesting timePress · TechCrunch AI ·
- More Anthropic researchers warn of AI’s perils but Musk dismisses ‘psyop’Press · The Guardian AI ·
- Is AI Actually Going to Kill Us All?Press · Wired AI ·
- One resignation turned the embers of AI fear into a wildfireNewsletter · Interconnects ·
The headline, key points and digest above were generated by Digest AI's editorial model from the linked sources. Automated summaries can contain errors: the sources are the record. Spotted a mistake? Tell us.
More in Society & Work
All →- Artists who boycotted Brisbane portrait prize over AI entries return to challenge AI works · 1 src
- CrowdStrike CEO: AI security risks persist regardless of model development pace · 3 src
- Meta fixes AI bug that surfaced private data on children after viral complaint · 6 src
- Global survey finds majority view AI as threat to jobs · 1 src
- Overcoming the Growing Wave of Ambient AI Nihilism and Public Fatalism · 1 src
Comments
via GitHub Discussions