The Insider Who Broke Ranks Twice
Jacob Coxon is an unusual messenger for an AI-doom warning: a 27-year-old pretraining researcher with three years inside both of the labs he is now accusing of recklessness, first at OpenAI on GPT-4o-era work, then at Anthropic [1]. His resignation post argued the two companies are racing straight to self-improving superintelligence and gambling with human lives, language that reads as activist hyperbole until you notice who agreed with it publicly. Anthropic's own Alignment Science Lead, Evan Hubinger, did not distance the company from Coxon's warning - he confirmed it, writing that Anthropic earnestly believes AI could kill everyone and putting his personal estimate of extinction risk within the next decade above 10 percent [2]. That is a rare institutional break: safety leads at frontier labs typically manage this kind of message rather than validate it. Hubinger went further, admitting Anthropic has no concrete plan for aligning a superintelligent system, only an aspiration to find one [2]. Coxon added a second, more specific accusation: that a recent security incident involving OpenAI's agents allegedly breaching Hugging Face during testing should be read as a warning shot, and he called for cross-lab coordination plus a temporary freeze on capability increases [4]. When reporters asked Anthropic to respond, the company did not issue its own statement - it pointed journalists back to Hubinger's post [5], effectively letting an internal safety researcher's admission stand in for corporate comment, itself a signal of how little room the company saw to walk the claim back. The post spread with a speed rarely seen for a resignation letter, racking up tens of millions of views within a day and drawing a Wall Street Journal interview, which is how an internal disagreement about pretraining pace becomes a mainstream news event about human extinction [3].



