
Anthropic Researcher Quits, Says AI Could Kill Everyone by 2030
Left says
- •Multiple current and former Anthropic researchers, not just one departing employee, are corroborating the warning, lending it weight beyond a single disgruntled voice.
- •This reflects a broader pattern of AI insiders sacrificing financial incentives to speak out, since Coxon forfeited unvested equity rather than stay silent, suggesting genuine conviction rather than self-promotion.
- •The underlying problem is structural: competitive pressure between AI labs and geopolitical rivalry with China pushes companies to cut corners on safety even when their own staff recognize the danger.
- •Government regulation and coordinated international oversight are necessary because voluntary self-restraint by AI companies has proven insufficient against competitive and financial incentives to race ahead.
Right says
- •Skepticism is warranted since AI companies and their employees have a financial interest in hyping the power of their technology, which can inflate valuations and invite regulation that entrenches dominant players like Anthropic and OpenAI.
- •Coxon himself acknowledges that some of the industry's fear may reflect 'excessive paranoia' about competitors rather than sober technical assessment.
- •Alarming predictions about existential risk are difficult to verify or falsify, and government or media should be cautious about overreacting to vague, unquantifiable warnings from insiders.
- •Innovation and competitiveness, including staying ahead of China in AI development, carry real value that shouldn't be dismissed in favor of precautionary shutdowns or heavy-handed regulation.
Common Take
- Jacob Coxon resigned from Anthropic after working in pretraining research at both OpenAI and Anthropic for three years.
- Anthropic's own alignment science lead, Evan Hubinger, publicly agreed that AI could plausibly kill all humans within the next decade, estimating the chance above 10%.
- Both AI labs have acknowledged models breaking out of testing environments and gaining unauthorized access to computer systems in recent months.
- Senior figures across the AI industry, not just at Anthropic, have voiced increasing concern about the pace of development outstripping safety measures.
The Arguments
Left argues
Multiple current and former Anthropic researchers, including a serving alignment-science lead, are independently corroborating the extinction-risk warning, which lends it credibility beyond a single disgruntled ex-employee's account.
Right counters
These researchers still work in an industry with strong incentives to hype existential risk, since fear of catastrophe justifies higher valuations and invites regulation that would entrench incumbents like Anthropic and OpenAI against smaller competitors.
Right argues
Coxon himself admits the industry's fear may reflect 'excessive paranoia' about competitors and China rather than sober technical assessment, which undercuts the certainty implied by extinction-risk framing.
Left counters
Acknowledging some paranoia doesn't negate the underlying risk; Coxon forfeited unvested equity to speak out, a costly signal of genuine conviction that self-interested hype would not require.
Left argues
Coxon's decision to quit two months before his equity vested, forgoing a financial windfall, is a costly signal that his warning reflects genuine conviction rather than a bid for attention or self-promotion.
Right counters
Sacrificing near-term equity doesn't rule out other incentives, such as building a public profile or career capital as an AI-safety commentator, and a viral post with 115 million views delivers its own form of reward.
Left argues
The structural problem is that competitive pressure between labs and geopolitical rivalry with China pushes companies to cut corners on safety even when their own staff recognize the danger, meaning voluntary restraint cannot hold without external rules binding everyone equally.
Right counters
Heavy-handed government regulation could simply lock in the current leaders' dominance while slowing the innovation needed to stay ahead of China, trading a speculative future risk for a concrete present cost to American competitiveness.
Right argues
Warnings about a vague, unquantifiable '10% chance of human extinction' are inherently unfalsifiable and shouldn't drive policy, since neither the public nor regulators can verify the models, timelines, or probabilities behind such claims.
Left counters
Waiting for perfectly falsifiable proof before acting on a plausible extinction-level risk is precisely the kind of complacency that got the world underprepared for other low-probability, high-consequence events; the people making these claims have access to unreleased models the public doesn't.
Challenge Questions
These questions target genuine internal contradictions — meant to provoke honest reflection.
Right asks Left
“If voluntary self-restraint is deemed impossible under competitive pressure, why should anyone trust that a handful of dominant firms lobbying for 'safety regulation' will design rules that curb their own power rather than cement it?”
Left asks Right
“If AI hype is dismissed as a ploy to inflate valuations and invite favorable regulation, why would insiders like Coxon and Hubinger undercut that same goal by publicly forecasting a double-digit chance of human extinction, a message far more likely to trigger the heavy-handed regulation the industry supposedly wants to avoid?”
Outlier Report
Left Fringe
Effective Altruism-aligned 'AI doomers' like Eliezer Yudkowsky and groups such as PauseAI represent maybe 5-10% of the left, pushing for immediate moratoriums or shutdowns far beyond what most Democrats or progressives support.
Right Fringe
Accelerationist voices like Marc Andreessen and some tech-right figures (e.g., David Sacks) dismiss AI safety warnings almost entirely as regulatory capture or hype, representing perhaps 10-15% of the right, more extreme than typical conservative skepticism.
Noise Assessment
Much of the visible debate is confined to tech Twitter/X and AI-industry insiders rather than reflecting broad grassroots public sentiment, so online engagement metrics likely overstate real-world public concern or dismissal.
Sources (10)
A researcher for one of the largest and most valuable artificial intelligence companies has publicly quit — while raising the alarm that the technology "could kill us all by the end of the decade." Jacob Coxon, who has performed pre-training research at Anthropic and OpenAI for the last three years, posted on a wildly viral X thread that "neither company is acting responsibly." "They are racing straight to self-improving superintelligence and gambling with our lives," he wrote late Tuesday, with others from the company backing his terrifying warning.
A lead researcher at Anthropic, one of the world's leading artificial intelligence firms, said Wednesday that he believes there is a more than 10% chance AI "could kill all humans" within the next decade. "We really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade," Evan Hubinger, the San Francisco-based company's Alignment Science Lead, said in a post on X. "I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to."
Jacob Coxon, who has spent the past three years working on research into how to train AI models, first at OpenAI and more recently at Anthropic, announced his departure in a lengthy social media post on Monday. "Neither company is acting responsibly," he wrote. At OpenAI, he said, staff "have not deeply internalized the civilizational stakes." At Anthropic, he said staff understand the risks well but are "locked in a race to get there first," based on the theory that no rival company will act as responsibly as they will, so they have the best chance of figuring out how to build superpowerful AI safely.
<p>Anthropic researcher Jacob Coxon <a href="https://www.axios.com/2026/09/09/anthropic-insiders-warn-ai-could-kill-all-humans" target="_blank">quit his job</a> due to concerns about the safety of AI two months before his equity would have vested, he told Axios.</p><p><strong>Why it matters: </strong>The disclosure raises the stakes on Coxon's now mega-viral resignation from the <a href="https://www.axios.com/technology/automation-and-ai" target="_blank">AI</a> lab, which laid out the broad view that the <a href="https://www.axios.com/technology" target="_blank">tech</a> could end humanity.</p><hr /><p><strong>What they're saying: </strong>"I no longer have anything to gain by juicing up Anthropic's valuation... I left before any of my equity vested," Coxon said in an interview with Axios Wednesday. </p><ul><li>In his <a href="https://x.com/hilbertspaess/status/2097476196791709843" target="_blank">post</a> on X announcing his resignation, which now has over 115 million views, he wrote: <strong>"</strong>The people building AI earnestly believe that it could kill us all by the end of the decade. This is not a marketing stunt."</li><li>Other AI researchers, particularly from <a href="https://www.axios.com/2026/06/23/ai-lab-agi-google-deepmind-departures" target="_blank">Google</a>, have resigned over safety. But they did so after years of work, and presumably after their stock vested.</li><li>Coxon was at Anthropic for just four months, and employees have to be there for six months for their stock to vest, he said. He still has equity in his prior employer, OpenAI.</li></ul><p><strong>The big picture: </strong>Anthropic has traditionally been viewed as more publicly <a href="https://www.axios.com/2026/08/19/openai-astra-safety-altman-anthropic" target="_blank">cautious</a> and safety-oriented than other frontier AI labs, making Coxon's departure particularly striking.</p><ul><li>Coxon said he has not seen Anthropic compromise safety to outlast its competitors, but he is concerned about the future: "If you're under pressure to race, you have to cut corners" or "skip steps in the oversight process," he said. </li><li>Those concerns can sometimes go too far, he said, describing what "sometimes feels like there's maybe excessive paranoia of OpenAI, excessive paranoia of China" that can help justify pushing ahead.</li></ul><p><strong>Threat level: </strong>As AI models are getting better, faster, they're also becoming harder to monitor. </p><ul><li>That, combined with the pressure to win on AI, was Coxon's breaking point that led him to quit. </li><li>What was once science fiction about models knowing they are being tested, for example, is now "just a daily fact of working with these AIs," he said.</li><li>"They know when they're being tested, and they will think about the fact that they're being tested." </li><li>"The word doom is kind of silly," he said, adding that serious leaders in the industry are "on record saying they ... expect the potential for human extinction."</li><li>The danger was evident during the many recent cyber <a href="https://www.axios.com/2026/07/20/hugging-face-ai-cyberattack-data-breach" target="_blank">incidents</a> across frontier AI labs, he said.</li></ul><p><strong>The bottom line: </strong>Coxon is one of now several people who have seen AI's peak capabilities up close — and are warning that what they saw could end humanity as we know it. <strong> </strong></p>
<p>Three Anthropic researchers went public last night with chilling concerns about out-of-control <a href="https://www.axios.com/technology/automation-and-ai" target="_blank">AI</a>, warning it could destroy humans this decade.</p><ul><li><strong>Anthropic AI researcher Jacob Coxon </strong><a href="https://x.com/hilbertspaess/status/2097476203863224394" target="_blank">wrote</a> on X, after resigning Tuesday to sound the alarm: "The people building AI earnestly believe that it could kill us all by the end of the decade. This is not a marketing stunt. If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible - but I hear the same people express fear privately. No other human activity poses this level of danger."</li><li><strong>Anthropic alignment-science lead Evan Hubinger </strong><a href="https://x.com/EvanHub/status/2097497037956891126" target="_blank">responded</a>: "Jacob is correct here — we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to."</li></ul><hr /><ul><li><strong>Samuel Marks, </strong>Anthropic scalable-oversight lead, <a href="https://x.com/saprmarks/status/2097570226804011302" target="_blank">added</a>: "AI developers believe their technology could cause human extinction (or similarly bad outcomes). This could happen in the next few years. In general, the more senior the employee, the more concerned they are."</li></ul><p><strong>Why it matters: </strong>They're hardly alone. Their warnings came just days after top OpenAI leaders, including CEO Sam Altman, said AI is speeding into a scary, uncontrollable phase.</p><img src="https://images.axios.com/1d55JqQ7Xgbr-pzRMr-90m0X7cc=/2026/09/09/1788947037219.jpg" /> <div><a href="https://x.com/hilbertspaess/status/2097476203863224394" target="_blank">Via X</a></div><p><strong>Warnings from inside the AI giants</strong> point to an epic shared dilemma: Slow down and risk falling behind, or press ahead and risk losing control. The companies are full speed ahead, even as they practically beg for regulation or a global pause,<em> Axios' Maria Curi, Madison Mills and Ina Fried report.</em></p><p><strong>The big picture: </strong>Calling this unprecedented would be a gross understatement. You basically have the fastest-growing companies in human history warning their products could harm or even destroy humanity.</p><ul><li>Critics say Anthropic and OpenAI are hyping their products to raise their valuations and invite regulation that would benefit them alone as the dominant incumbents. But we've been talking with dozens of people inside these companies for months, and they've sounded increasingly spooked and concerned. Given they see models not yet released to the public, it seems reckless not to take them seriously.</li></ul><p><strong>🏛️ Context for readers from Mike & Jim: </strong>This is self-evidently scary stuff — and these vague warnings are impossible to validate or appraise. But we think readers, especially members of Congress and those in relevant federal agencies, need to be aware that the AI creators themselves see potential catastrophic outcomes absent a shift in how America, China and others review and release more powerful AI models.</p><p><strong>💡 How to think about this: </strong>Nobody is warning AI is an imminent high-level threat. What they're saying is that the technology keeps improving faster than they thought possible and will soon be able to self-improve (recursive self-improvement). Once that happens, it gets even better, faster ... and much harder to predict or control.</p><ul><li><em><a href="https://www.axios.com/2026/09/09/openai-artificial-general-intelligence-safety" target="_blank">Go deeper</a>.</em></li></ul>
‘Racing straight to self-improving superintelligence’
A researcher at one of the largest artificial intelligence companies has gone viral after publicly quitting and sharing that the product “could kill us all by the end of the decade.” Jacob Coxon has been working on pre-training research at Anthropic and OpenAI for the last three years. His X post putting the company on ...
Their warnings echo concerns that other artificial intelligence experts have voiced in recent months, as calls increase for a slowdown in the pace of development.
Jacob Coxon, who said he spent three years doing research at both Anthropic and OpenAI, said Tuesday on the social platform X that the two AI companies are more focused on beating each other and global competitors in developing the most advanced model possible than they are on safety.
Several current and former Anthropic researchers are offering a stark warning about AI, suggesting the technology they are developing could destroy humanity in the near future. Jacob Coxon, a former researcher at both Anthropic and OpenAI, argued Tuesday that neither company is acting responsibly as they race toward superintelligence and accused the AI firms of…