Skip to content
Tech News
← Back to articles

Anthropic researcher says AI has more than 10% chance of 'killing all humans' after colleague quits

read original more articles
Why This Matters

Two Anthropic insiders publicly voicing existential-risk concerns — one quitting outright — is a rare, credibility-laden crack in the AI industry's confident public posture. It matters because these warnings come from inside a lab that markets itself on safety, and land while Anthropic and OpenAI raise huge sums and move toward public listings. For consumers and enterprises betting on these tools, the gap between internal risk estimates and commercial urgency is now on the record.

Key Takeaways

There is more than a 10% chance that artificial intelligence could "kill all humans," an Anthropic safety researcher said on Tuesday, hours after another employee said he was quitting the company over concerns that AI labs are "gambling with our lives."

The comments underscore growing concerns among those at the heart of AI development that the technology could get out of control and pose a threat to humanity, even as Anthropic and OpenAI continue to raise large sums of money and head toward expected public listings.

Jacob Coxon, a researcher at Anthropic, said on Tuesday he resigned from the company. Coxon said neither Anthropic nor OpenAI is acting responsibly.

"They are racing straight to self-improving superintelligence and gambling with our lives," Coxon said in a post on X.

Self-improvement is the idea that AI systems can improve themselves without much human intervention. Recursive self-improvement, as it is often called, is not yet possible, but AI labs are working toward the goal.