Anthropic researcher quits warning of >10% extinction risk; alignment lead concurs

Jacob Coxon quit Anthropic in a public resignation, accusing both Anthropic and OpenAI of 'racing straight to self-improving superintelligence and gambling with our lives.' The extraordinary escalation came when Anthropic's own Alignment Science Lead, Evan Hubinger, endorsed the underlying concern — estimating a greater-than-10% chance of human extinction within a decade and conceding that no concrete plan for aligning superintelligence currently exists. Hubinger's statement drew more than 10 million views.
The episode split the community sharply. On Reddit, r/Anthropic's thread 'Another Anthropic safety researcher quits: "We may not survive this"' drew 859 upvotes and 757 comments, while a companion thread asked 'What EXACTLY do they mean when they say that AI could cause an extinction event?' (118 upvotes, 331 comments). Critics on X and Reddit pressed the obvious question: why would employees who genuinely believe in those odds continue to work at the companies driving the risk? Venture investor David Sacks went further, calling for Anthropic's rumored IPO to be suspended pending investigation.
Contextually, the resignation lands in the same week as Dario Amodei's 'Pace the Frontier' essay urging an industry slowdown, and reports of DeepMind pursuing recursive self-improvement and OpenAI's Astra hitting a 'Critical' cyber threshold. That juxtaposition — a company publicly committing to caution while a senior alignment leader says the odds of catastrophe exceed 10% — is precisely the contradiction critics seized on.
What to watch: whether other researchers follow Coxon out the door, how Anthropic's leadership responds to Hubinger's public estimate, and whether the resignation influences the regulatory and IPO conversations now swirling around the lab.