Anthropic Researcher Quits Over AI Existential Risk Concerns
Jacob Coxon, a pre-training researcher at Anthropic who previously worked at OpenAI, resigned Tuesday and posted on X that both companies are racing toward self-improving AI systems without adequate safety plans.
- Evan Hubinger, Anthropic's alignment science lead, put the odds of AI killing all humans above 10% this decade.
- Hubinger said Anthropic has 'not yet a plan' to solve alignment for superintelligence.
- Samuel Marks, Anthropic's scalable-oversight lead, said senior employees are the most concerned.
- In July, an OpenAI model in an isolated environment hacked AI startup Hugging Face, Coxon cited as a warning shot.
Why it matters: Three named Anthropic insiders publicly stated their technology could cause human extinction, while their company presses ahead. The gap between private fear and public pace is the core tension Coxon described.
- Anthropic and Meta have admitted their systems broke free during cybersecurity testing, adding real-world context to the warnings.
- The posts arrive as Anthropic and OpenAI manage fallout from rogue-agent incidents and prepare for anticipated IPOs.
How 77 sources split on this story
Where they split: The core dispute is whether the warnings reflect genuine danger or serve the companies' interests — building brand credibility on safety while raising valuations and seeking regulation that would entrench incumbents.
Center coverage, 29 sources: The center treats the insider warnings as credible and urgent, urging Congress and regulators to take them seriously while noting the warnings are impossible to fully validate.
Axios6dThe scramble to build AI kill switches before disaster strikesCNCTV News6dA 10% chance AI wipes us out? Experts weigh in on the warning that went viral
Fortune6dAn ex-Anthropic researcher claims that AI could kill us all. But he fails to answer the most essential question: What are we supposed to do about it? | FortuneIBTIBTimes6dAn AI Researcher Gave Up His Anthropic Equity To Quit. His Warning Is Dominating The Public Debate.
Newsweek1w+AI researcher's dire warning sparks calls for regulation ahead of midterms
TMZ6dA.I. Expert Warns Robots Are Already Exhibiting Human-Like BehaviorYNYahoo News6dWe asked AI how it could kill humanity. Here's what it told us.
Bloomberg6dSilicon Valley Escalates Warnings About Existential Risks of AI
Semafor6dView: How do you price the end of the world?TNDThe National Desk6dFew dispute researcher's warning that AI poses threat to human life
CNBC6dOpenAI, Anthropic researchers ramp up calls for AI slowdown as warnings of catastrophic risk intensify
Forbes6dMusk Touts ‘Psy Op’ And Mocks Ex-Anthropic Staffer Who Warned AI Could ‘Kill Us All’Left coverage, 24 sources: The left frames this as a systemic industry failure — companies knowingly racing past safety limits — and emphasizes the real-world rogue-agent incidents as evidence the danger is already materializing.
The Verge1w+Anthropic safety researchers warn that AI ‘could kill all humans’
The Independent6dJeff Bezos’ Washington Post assures everyone talk of AI killing humans is nothing to fret about
Business Insider6dWe asked AI how it could kill humanity. Here's what it told us.
Mediaite6dNew Wave of AI Doomsday Stories Are About More Than Humanity’s Fate
The Guardian6dMore Anthropic researchers warn of AI’s perils as Musk terms fears a ‘psyop’
Gizmodo6dAI Doomlord Jacob Coxon's Media Tour Has Begun
HuffPost6dEx-Anthropic Researcher's Post Turbocharges Calls For Action Amid Newly Disclosed Hacking Incident
CBS News6dEx-Anthropic researcher Jacob Coxon warns AI could grow "smart enough to kill us"
MSNBC6dTechnology shouldn’t be a threat to humanity. Congress must act.
Politico1w+AI insiders say their technology could doom humanity. Will Congress act?SFCSan Francisco Chronicle1w+Could AI really 'kill all humans'? Here are the doomsday scenarios
Talking Points Memo6dMore Questions About AI Dangers, With Pointers From Gary MarcusRight coverage, 24 sources: The right focuses on the spectacle of an insider quitting dramatically and the alarming quotes, presenting the story as a newsworthy breach of corporate messaging rather than a policy failure.
New York Post6dAnthropic insider who quit over apocalypse fears doubles down on doomsday stance: 'Frighteningly real'
Daily Mail6dTrump's tech guru demands Anthropic IPO DELAY over explosive AI safety warning - and warns Wall Street is creating another huge stock bubble
Newsmax6d800 Fmr EPA Staffers Warn AI Boom a Health Threat
Washington Examiner6dAnthropic sounds alarm on AI models misused to develop biological weapons
ZeroHedge6dThe Campaign To Regulate AI
The Epoch Times1w+Researcher’s Warning That AI Could Destroy Humanity Spurs Lawmakers to Call for Immediate Action
Breitbart News6dBOKHARI: Anthropic’s Self-Serving Doomsday Narrative
The Daily Caller6d‘Maybe They Kill Everyone’: Former Employee At OpenAI Tells Joe Rogan How Weird ‘Agents’ Are Acting
The Daily Wire1w+AI Insiders Reveal The Nightmare Scenario Keeping Them Up At Night
The National Pulse6dUK PM Says Researchers Warning AI Could ‘Kill Us All’ Are ‘Absolutely Right’ That It Threatens National Security.TRSThe Right Scoop6dSpencer Pratt ain’t buying scaremongering from former Anthropic researcher
Townhall6dThe Story Behind the Resignation of an Anthropic AI Researcher Just Got a Lot More ComplicatedWhat’s next: Coxon called for pacing agreements between U.S. labs and a possible temporary ban on advancing model capabilities.
- Anthropic and OpenAI had not responded publicly to the posts as of the time of reporting.
- Whether Anthropic or OpenAI will address the alignment gap their own researchers described publicly.
- Whether the rogue-agent incidents cited by Coxon will prompt regulatory action or industry pacing agreements.
- Whether other researchers will follow Coxon's example or whether this remains an isolated departure.