Skip to main content

Jacob CoxonAnthropic pre-training researcher Jacob Coxon publicly resigned, warning labs are racing toward uncontrolled self-improving AI

On September 9, 2026, Jacob Coxon, a 27-year-old AI pre-training researcher who had worked at both OpenAI and Anthropic, publicly resigned from Anthropic via a post on X, warning that AI labs are 'racing straight to self-improving superintelligence and gambling with our lives.' He said Anthropic understands the civilizational stakes but believes it must race to build advanced AI first because it doesn't trust competitors to act responsibly, and called the recent OpenAI Hugging Face breach and an Anthropic agent safety-evaluation escape a 'warning shot.' Coxon called for coordination between US labs on pacing agreements and said a temporary ban on improving model capabilities may be warranted in worst-case scenarios. Anthropic colleague Evan Hubinger publicly echoed related concerns, estimating a greater than 10% chance AI could cause catastrophic harm within a decade and acknowledging Anthropic lacks a working plan to solve alignment for superintelligence.

Scoring Impact

TopicDirectionRelevanceContribution
AI Safety+towardprimary+1.00
Overall incident score =+0.885

Score = avg(topic contributions) × significance (high ×1.5) × confidence (0.59)

Evidence (1 signal)

Confirms Statement Sep 9, 2026 verified

Jacob Coxon's public X post announcing resignation from Anthropic over AI safety concerns

Coxon, a pre-training researcher who worked at OpenAI and Anthropic, posted on X on September 9, 2026 announcing his resignation and warning that AI labs are racing toward self-improving superintelligence without adequate safety guarantees.

Related: Same Topics