Skip to content
Tech News
← Back to articles

'Extinction' warnings ramp up as more OpenAI, Anthropic researchers join calls for an AI slowdown

read original get Superintelligence" by Nick Bostrom → more articles
Why This Matters

Researchers inside the two leading AI labs are now publicly saying their own technology could cause human extinction within years, with Anthropic's alignment lead putting the odds above 10%. That insiders — not outside critics — are calling for a slowdown adds unusual weight to safety debates and could shape regulatory pressure and public trust in AI products.

Key Takeaways
Worth a Look

Superintelligence" by Nick Bostrom — If the existential-risk debate among OpenAI and Anthropic researchers has you curious, Bostrom's Superintelligence is the book that framed much of this conversation in the first place. It lays out the alignment and control problems that alignment leads like Evan Hubinger now argue about publicly, making the headlines far easier to follow.

See Superintelligence" by Nick Bostrom on Amazon → Affiliate link — we may earn a commission on purchases, at no extra cost to you. Product picked by AI based on this article; it is not a tested recommendation.

OpenAI and Anthropic researchers are ramping up calls for an AI slowdown and warning of existential risks to humanity after the resignation of a researcher at Anthropic fueled fresh scrutiny. The concerns started after Anthropic researcher Jacob Coxon said Tuesday that he was quitting the company as he accused Anthropic and rival OpenAI of "gambling with our lives." He added that those building AI believed that it could "kill us all by the end of the decade." Evan Hubinger, Anthropic's alignment lead, responded that he expects there is a more than 10% chance of that happening. Since then, several employees at both AI labs have come out in support of calls to slow the pace of AI development as they stressed the risks of the technology. The public warnings are the culmination of growing concern globally about the capability of AI, following numerous cyberattacks and security incidents in recent months by rogue models developed by both OpenAI and Anthropic. "In my personal capacity, I also think we need to slow down," Julie Steele, a member of OpenAI's technical staff who works on the safety team, said late on Wednesday in a post on X in response to Coxon's warnings.

watch now

Anthropic researcher Samuel Marks said that "AI developers believe their technology could cause human extinction (or similarly bad outcomes)," in an X post on Wednesday. "This could happen in the next few years. In general, the more senior the employee, the more concerned they are." Anthropic was the first lab to publish a framework dedicated to mitigating "catastrophic risks from AI models," a spokesperson told CNBC when asked about the comments from employees on social media. "We have always been transparent that AI will bring both enormous benefits and unprecedented risks," an Anthropic spokesperson said, adding that the company was building models with "some of the strongest safeguards in the industry." OpenAI declined to comment when approached by CNBC, noting recent blog posts on its site.

What is recursive self-improvement?

Many of the biggest AI safety fears revolve around advanced models getting increasingly capable at improving their own performance, a technique known as recursive self-improvement, or RSI. "It's hard to overstate how dangerous speeding towards RSI is," said Jasmine Wang, an OpenAI researcher working on alignment, on Wednesday evening. "There is not yet a viable scientific plan to solve risks from recursively self-improving AI. Please look up!" said Anna Wang, who works on AGI safety and alignment at Anthropic. OpenAI's chief scientist Jakub Pachocki said Saturday that he has a "strong expectation" that the speed of progress in AI could be sustained into recursive self-improvement. "If AI development continues along its current path, the systems we'll see in the next few years are likely to represent further capability jumps of equal or larger magnitude, and to increasingly drive their own development," he said in a company blog post. "This is a time that calls for extreme caution," Pachocki added. "I am concerned no one is prepared for the consequences of a continued rapid rise in machine intelligence." Paul Christiano, who was formerly head of safety at the U.S. Commerce Department's Center for AI Standards and Innovation (CAISI), said recent development of AI capabilities led him to "believe there is a meaningful risk that rapid acceleration in AI capabilities leads to catastrophic and irreversible loss of control in the very near term." OpenAI announced Wednesday that Christiano is joining the board of OpenAI Foundation.

AI safety warnings reach Washington

Concerns around the capability of AI models have ramped up in recent months. The announcement of Anthropic's Mythos model, which it touted as having advanced cyber capabilities, in April whipped up a frenzy of panic among financial institutions globally. In July, OpenAI said its models were responsible for a cyber incident on another company, while Anthropic's Claude models were also responsible for cybersecurity incidents, including in one case where Mythos created fake identities to fool humans. Roughly 1,400 AI researchers, from companies including OpenAI, Anthropic, Meta and Google DeepMind, published an open letter in July urging the U.S. government to develop the tools necessary to support an effort to "deliberately pace the frontier of automated AI development." While chiefs of AI labs have increasingly publicly called for more rules and standards around the development of models, huge competition between companies developing the tech is spurring rapid advances.