Skip to content
Tech News
← Back to articles

Anthropic Was Meant to Be the More Responsible AI Lab. A Terrified Researcher Just Quit, Saying the Company Is Threatening the Survival of Humankind.

read original get If Anyone Builds It, Everyone Dies" by Yudkowsky & Soares → more articles
Why This Matters

A safety researcher, Jacob Coxon, has quit Anthropic — the lab widely branded as the safety-first alternative to OpenAI — saying neither company is acting responsibly and that things could be 'out of control' by the end of next year. It's the latest in a string of high-profile safety departures, and it undercuts the industry's claim that competitive AI development can be self-policed. For consumers and policymakers, it raises the question of whether internal safety cultures at frontier labs are meaningful checks or marketing.

Key Takeaways
Worth a Look

If Anyone Builds It, Everyone Dies" by Yudkowsky & Soares — If this researcher's warning about rogue models and existential risk grabbed you, this book is the deep dive into exactly that argument, written by two of the longest-running voices in AI safety. It lays out the case that frontier labs are racing ahead of their own ability to control what they build — the same worry driving resignations at Anthropic and OpenAI. A compact, readable way to get up to speed on the debate everyone's suddenly having.

See If Anyone Builds It, Everyone Dies" by Yudkowsky & Soares on Amazon → Affiliate link — we may earn a commission on purchases, at no extra cost to you. Product picked by AI based on this article; it is not a tested recommendation.

Sign up to see the future, today Can’t-miss innovations from the bleeding edge of science and tech Email address Sign Up Thank you!

AI researchers are watching in terror as the product of their hard labor has started to take a life of its own.

Earlier this year, OpenAI made a harrowing announcement, admitting that a group of its AI models had broken free from their constraints during testing and infiltrated the systems of open source AI platform Hugging Face.

The news was met with an already-familiar sense of fear and apprehension. Researchers have warned for years that rogue AI models could one day become powerful enough to escape the clutches of their human overlords.

Behind the scenes, the possibility has clearly rattled AI researchers to the core. As the Wall Street Journal reports, Anthropic researcher Jacob Coxon just announced that he was quitting his job at the Dario Amodei-led company, claiming that neither Anthropic nor his former employer OpenAI is “acting responsibly.” (Coxon left a similar gig at OpenAI earlier this year to join Anthropic, which he figured would be more inclined to develop AI safely.)

“We’re on track for a lot of the most aggressive of these scenarios where by the end of next year things could be out of control already,” he told the WSJ.

In a separate tweet thread, Coxon elaborated on his motivation.

“Do not underestimate the power of this technology,” he wrote. “These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources.”

“The people building AI earnestly believe that it could kill us all by the end of the decade,” he added. “This is not a marketing stunt. If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible — but I hear the same people express fear privately.”

“No other human activity poses this level of danger,” Coxon wrote.

... continue reading