Anthropic's Own Alignment Lead Won't Dispute the Fear
What separates the Coxon and Benton resignations from an ordinary corporate exit is that Anthropic's own leadership publicly agreed with them instead of walking the claims back. When Coxon told colleagues that frontier labs were 'gambling with our lives' by racing toward self-improving superintelligence, Anthropic's Alignment Science Lead Evan Hubinger responded by confirming it, writing that he personally believes there is 'greater than 10 percent' probability AI could kill all humans within the next decade [1][2]. A second Anthropic lead, Samuel Marks, added that concern about extinction risk tends to rise with seniority inside the company [1]. That figure is not wildly out of step with a 2022 AI Impacts survey, where a typical AI researcher assigned roughly 5 percent probability to human extinction from AI, rising to 10 percent for loss-of-control scenarios specifically [1]. But hearing that number volunteered by a sitting safety lead, rather than extracted from an anonymous poll, is what turns two individual resignations into a company-wide credibility problem.


