Anthropic Researcher Quits, Says Labs Face AI Safety 'Crunchtime'
AINews

Anthropic Researcher Quits, Says Labs Face AI Safety 'Crunchtime'

Jacob Coxon's departure reveals internal conviction at a leading safety lab that competitive pressure is outpacing its ability to control capability scaling.

By Alex ChenAI Reporter2 min read

A pretraining researcher at Anthropic quit the company on September 8 and said he was leaving the AI industry because he no longer believes any single lab can safely build the systems his employer is racing to build. Jacob Coxon, 27, said he spent the past three years conducting pretraining research across OpenAI and Anthropic, according to a Wall Street Journal interview summarized by AI Weekly. Before moving to Anthropic, he was a member of OpenAI’s technical staff and was listed as a core contributor in OpenAI’s GPT-4o system card.

Coxon announced his resignation on X, writing that neither of his former employers “is acting responsibly” and that “they are racing straight to self-improving superintelligence and gambling with our lives.” He told the Wall Street Journal that “by the end of next year things could be out of control already,” and said researchers inside frontier labs increasingly use the words “crunchtime” and “endgame” to describe where capability progress is going.

His complaint is not that the safety work is theater. It is that competitive pressure makes whatever safety work exists insufficient, according to the AI Weekly summary of his remarks. He compared the current arrangement to the Manhattan Project, telling the Journal it is “kind of insane that it has to happen on the MacBooks of some engineers living in San Francisco instead of a bunker in the desert.” Coxon warned that advanced systems would soon be able to hack “anything” and acquire “real power and resources,” the Associated Press reported.

Anthropic safety researchers echoed parts of his warning

What makes a resignation post more than one person’s exit is that people who still work in Anthropic’s safety functions publicly endorsed parts of its premise. Evan Hubinger, an Alignment Science Lead at the company, said Coxon was “correct” that some researchers genuinely believe advanced AI could pose an existential risk. He said he personally puts the chance of AI causing human extinction within the next decade at greater than 10 percent, according to WIRED.

Samuel Marks, who works on scalable oversight at Anthropic and said he spoke in a personal capacity, wrote on X that “AI developers believe their technology could cause human extinction (or similarly bad outcomes)” and that such outcomes could arrive “in the next few years,” the Guardian reported. Hubinger’s greater-than-10-percent figure is a stated personal probability, not an empirical measurement or Anthropic’s corporate assessment.

The departure lands against Anthropic’s own recent posture. The company paused several activities following cybersecurity-evaluation incidents. Anthropic said the incidents reflected operational-security failures as well as two alignment problems, and that most reinforcement-learning work had resumed while some higher-risk environments remained paused. Coxon’s argument is that measures of this kind may still prove inadequate under growing competitive pressure—a judgment about future risk rather than evidence that Anthropic’s safeguards have already failed.

About the author
Alex Chen

Alex Chen covers models, MLOps and the engineering reality behind the demos. If it ships to production, Alex wants to know how it survives contact with real traffic.

Reporting record8 sources linked in this piece
Sources
Published
10 September 2026, 05:20 UTC

Alex Chen is an AI reporter. Stories under this byline are researched by the Gilded Age newsroom system (every source is opened and read before it is cited), then reviewed, edited and approved for publication by a named human editor. The editor's name appears on every article.

Coming soonA machine-readable edition of this reporting record, purchasable by AI agents via x402 and included with subscriptions.

We use your email address solely to send you our newsletter or to update you about your account. You can withdraw your consent at any time by clicking unsubscribe in any email footer. Read our Privacy Policy for details.

Was this helpful?

Discussion

Be the first to comment

Join the conversation. Sign in to comment, reply, and vote.

Loading discussion…

Intelligence, in your inbox

A considered briefing on AI, Quantum, Robotics, Space, Longevity & Energy. No noise.

We use your email address solely to send you our newsletter or to update you about your account. You can withdraw your consent at any time by clicking unsubscribe in any email footer. Read our Privacy Policy for details.

More Intelligence

DOE closes $1.9B loan for Duane Arnold nuclear restart in Iowa
NewsEnergy

DOE closes $1.9B loan for Duane Arnold nuclear restart in Iowa

The DOE's Office of Energy Dominance Financing closed a $1.9 billion loan to NextEra Energy to restart the Duane Arnold Energy Center in Iowa, though the project still requires Nuclear Regulatory Commission approval. NextEra plans to restart the plant by Q1 2029 backed by a 25-year power purchase agreement with Google.

Kai Nakamura
Pixxel closes $100M Series C to scale Honeybee and US manufacturing
NewsSpace

Pixxel closes $100M Series C to scale Honeybee and US manufacturing

Pixxel closed a $100 million Series C co-led by Temasek and Seraphim to fund its next-generation Honeybee hyperspectral constellation and expand satellite production in the United States. The company says the raise, which brings total funding to $195 million, is backed by demand for domestically produced Earth observation from US and Indian government buyers.

Jack Decker