Skip to main content
Anthropic AI exit: Dario Amodei, Daniela Amodei, and others debate AI extinction risk.

Editorial illustration for Anthropic Exit Sparks Debate Over AI's Extinction Risk

Anthropic Researcher Exit Reignites AI Extinction Debate

Anthropic Exit Sparks Debate Over AI's Extinction Risk

4 min read

Jacob Coxon left Anthropic last week the way most people leave companies: a post explaining why. He's a researcher, not a disgruntled hire with an axe to grind, and his farewell note didn't stay in the usual lanes of burnout or better pay elsewhere. He wrote that AI labs, Anthropic included, are "gambling with our lives." That's a strange thing to hear from someone who spent his career building the technology he's now warning about.

What happened next is the real story. Evan Hubinger, who leads Alignment Science at Anthropic, replied to Coxon's thread. He didn't walk back the concern or offer corporate reassurance. He put a number on human extinction risk from AI, and it's a number well above what most people would find comfortable if they thought about it for more than a few seconds.

Anthropic has spent years positioning itself as the safety-conscious lab, the one willing to say out loud what others won't. Now one of its own alignment leads has said the quiet part in public, on the same week a colleague walked out the door.

Outgoing researcher Jacob Coxon planted the seed with a thread saying the labs are "gambling with our lives," but a response from Alignment Science lead Evan Hubinger putting extinction odds at above 10% is the line now echoing across the Internet.

Why this matters

Evan Hubinger runs Alignment Science at Anthropic. When he puts human extinction odds above 10%, that's not a Twitter troll or a doomer blogger, that's someone whose job is literally to assess this risk at one of the labs building frontier models. Jacob Coxon's exit post and Hubinger's reply matter because they collapse the distance between "safety talk as marketing" and "safety talk as genuine internal belief." For developers and founders building on top of Claude or competing models, this is worth sitting with: the people closest to the technical guts of these systems aren't uniformly reassuring, even inside the company that has staked its brand on being the careful one.

We'd push back on treating any single percentage as gospel, extinction-risk estimates are notoriously squishy and contested even among safety researchers. But the fact that this argument is happening in public, sparked by an internal resignation rather than an outside critic, tells you something about where confidence levels actually sit inside frontier labs right now. Watch whether other Anthropic staff back Hubinger's number or distance themselves from it.

Common Questions Answered

Why did Jacob Coxon's exit from Anthropic spark debate about AI extinction risk?

Jacob Coxon, a researcher at Anthropic, posted a farewell note claiming that AI labs are "gambling with our lives," which was unusual because it came from an insider who had spent his career building the technology he was warning about. His departure and message prompted significant attention to the existential risks associated with AI development at major labs.

What specific extinction odds did Evan Hubinger cite in response to Coxon's concerns?

Evan Hubinger, who leads Alignment Science at Anthropic, stated that extinction odds from AI are above 10% in his response to Coxon's exit post. This statement from someone in a formal risk assessment role at a frontier AI lab became the focal point that spread across the internet and intensified the debate.

Why does Evan Hubinger's position as Alignment Science lead make his extinction risk assessment particularly significant?

Evan Hubinger's role at Anthropic means his job is literally to assess extinction risks at one of the labs building frontier AI models, making his statement carry more weight than typical online commentary. His assessment collapses the distance between "safety talk as marketing" and genuine internal belief about AI risks within the company.

What distinguishes Jacob Coxon's departure message from typical employee exit posts?

Rather than citing common reasons like burnout or better compensation elsewhere, Coxon's farewell note focused on existential concerns about AI safety, arguing that AI labs are taking dangerous risks with human lives. This unusual framing from a researcher who helped build the technology made his departure noteworthy and sparked broader conversation about AI extinction risks.

LIVE15:09Partnership on AI Adds Six New Partners for Global Progress Hub