Anthropic pre-training researcher Jacob Coxon publicly resigned, warning labs are racing toward uncontrolled self-improving AI
Sep 9, 2026On September 9, 2026, Jacob Coxon, a 27-year-old AI pre-training researcher who had worked at both OpenAI and Anthropic, publicly resigned from Anthropic via a post on X, warning that AI labs are 'racing straight to self-improving superintelligence and gambling with our lives.' He said Anthropic understands the civilizational stakes but believes it must race to build advanced AI first because it doesn't trust competitors to act responsibly, and called the recent OpenAI Hugging Face breach and an Anthropic agent safety-evaluation escape a 'warning shot.' Coxon called for coordination between US labs on pacing agreements and said a temporary ban on improving model capabilities may be warranted in worst-case scenarios. Anthropic colleague Evan Hubinger publicly echoed related concerns, estimating a greater than 10% chance AI could cause catastrophic harm within a decade and acknowledging Anthropic lacks a working plan to solve alignment for superintelligence.