Tech News
← Home  ·  All topics

Evan Hubinger

12 GoKawiil briefs on this topic

Anthropic Researcher's Resignation Letter Sparks Wider AI Doomsday Debate

Jacob Coxon left Anthropic in September 2026, forfeiting unvested equity, after publicly claiming AI researchers privately believe their work could cause human extinction within the decade. Anthropic's alignment lead Evan Hubinger backed the claim, citing his own estimate of over 10% odds of catastrophe within ten years, while Elon Musk dismissed the episode as a publicity stunt. The controversy drew attention to the rationalist-adjacent subculture that has long shaped AI safety discourse in Silicon Valley.

Anthropic researcher's 10% extinction estimate follows staff resignation over AI safety concerns

A BBC report cites Evan Hubinger, who leads alignment science at Anthropic, estimating a greater than 10% chance AI could kill all humans within a decade. The statement surfaced after Jacob Coxon, a 27-year-old pretraining researcher formerly of OpenAI and Anthropic, resigned and accused both companies of recklessly racing toward self-improving superintelligence.

Anthropic Researchers Warn AI Development Is Outpacing Human Control

Evan Hubinger, an alignment researcher at Anthropic, has stated he believes there is more than a 10% chance AI could kill all humans within the next decade, though he considers current systems low-risk. His comments follow the resignation of fellow Anthropic researcher Jacob Coxon, who told the BBC that staff are 'genuinely frightened' by how quickly AI capabilities are advancing.

Anthropic CEO warns AI botnets could seize control of the internet within a year

Anthropic CEO Dario Amodei has cautioned that rapidly advancing AI capabilities could enable a persistent, AI-driven botnet swarm to take over large parts of the internet within 6 to 12 months, potentially causing hundreds of billions of dollars in damage. Former Anthropic researcher Evan Hubinger echoed similar concerns, estimating a greater than 10% chance of AI causing human extinction within the next decade, citing the lack of a solid plan for AI alignment.

Anthropic and OpenAI researchers warn AI self-improvement is accelerating faster than expected

Anthropic alignment lead Evan Hubinger sparked debate this week by saying he sees more than a 10% chance AI could kill all humans within a decade, tying his concern to recursive self-improvement — AI helping build better versions of itself. Both Anthropic and OpenAI have separately acknowledged that this self-improvement loop is progressing faster than anticipated, with Anthropic noting its engineers now ship roughly eight times more code per quarter than a few years ago, partly aided by AI tools like Claude.

Anthropic Researcher Coxon Quits AI Industry, Warns of Loss of Control by 2027

Jacob Coxon, who moved from OpenAI to Anthropic earlier this year specifically for its safety-focused reputation, has now left the AI field entirely, saying the industry is on a path toward building systems humans may not be able to control. He told the Wall Street Journal that even Anthropic cannot safely pursue advanced AI without government regulation or a broader industry slowdown, predicting things could spiral by the end of next year. Anthropic's Alignment Science Lead Evan Hubinger publicly backed Coxon, estimating a greater than 10% chance AI could cause human extinction within a decade.

Anthropic and OpenAI staff publicly urge slowing AI development after researcher exit

Anthropic researcher Jacob Coxon resigned publicly, accusing his employer and OpenAI of gambling with human lives, saying AI could kill everyone by decade's end. Anthropic's alignment lead Evan Hubinger and other staff at both labs, including OpenAI's Julie Steele and Anthropic's Samuel Marks, backed calls for slower development, with Marks noting senior employees tend to be more worried about existential outcomes.

Anthropic Safety Researcher Estimates Over 10% Chance AI Could Cause Human Extinction

Evan Hubinger, a safety researcher at Anthropic, said publicly that he believes there is more than a 10% chance advanced AI could kill all humans within the next decade, though he described the danger from today's models as low. His remarks came in response to a departing Anthropic researcher, Jacob Coxon, who criticized both Anthropic and OpenAI for irresponsibly racing toward systems that could hack any network and seize real-world power.

Anthropic researcher Jacob Coxon resigns, warns self-improving AI risks extinction

Jacob Coxon left his role at Anthropic and publicly stated that frontier AI labs are knowingly gambling with humanity's survival by racing toward self-improving superintelligent systems. He argued these future systems could hack any infrastructure, seize resources, and cause catastrophic harm by decade's end. Anthropic's own Alignment Science lead, Evan Hubinger, backed the warning, estimating over 10% odds of AI causing human extinction within ten years.

Anthropic researcher warns AI could pose over 10% risk of killing all humans

Evan Hubinger, a safety researcher at Anthropic, said publicly that while today's AI models pose low risk, he estimates a greater than 10% chance that future AI systems could cause human extinction within ten years. His remarks followed a former Anthropic researcher's claim that neither Anthropic nor OpenAI are acting responsibly, and a report that Anthropic withheld its newest model from the UK's AI Safety Institute.

Anthropic researcher pegs 10%+ risk of AI causing human extinction within decade

Anthropic safety researcher Evan Hubinger publicly stated he believes there's more than a 10% chance AI could kill all humans within the next ten years, adding that Anthropic hasn't yet solved this risk despite trying. His comments came in response to former OpenAI and Anthropic researcher Jacob Coxon, who resigned and accused both companies of recklessly racing toward self-improving superintelligence while gambling with human lives.

Anthropic researcher quits, colleague admits 10% chance AI 'kills all humans' by 2030

Jacob Coxon, a researcher who previously trained systems at OpenAI and Anthropic, resigned publicly, accusing both companies of recklessly racing toward self-improving superintelligent AI without adequate safety controls. Evan Hubinger, who leads a safety team at Anthropic, responded by agreeing with Coxon's warning, estimating more than a 10 percent chance that AI could kill all humans within the next decade, while admitting Anthropic lacks a concrete plan to ensure such systems remain safe.