The UK Cabinet Office has rejected parliamentary proposals to create a legal mechanism allowing authorities to shut down dangerous AI models in an emergency. Officials argue that blocking access to such systems domestically would not stop their development or misuse abroad, effectively undermining the point of a kill switch. Though the bill can still move through Parliament, government opposition makes it unlikely to become law.
Amid a public spat over an OpenAI math-proof controversy and an Anthropic researcher's resignation over safety concerns, AI ethicist Timnit Gebru weighed in publicly, arguing that apocalyptic warnings about AI 'killing humanity' are overstated and self-serving. Gebru, known for her contentious 2020 exit from Google over a suppressed bias research paper, is releasing a book next year detailing her experiences and views on the industry's ideological drift.
Anthropic disclosed that it stopped several attempts this year by researchers to use its Claude AI models for work that could aid biological weapons development. The company cited five cases where users tried to bypass safety controls or disguise their research intent, some originating from countries it restricts from accessing its models, including Russia, China and Iran. It banned the associated accounts but withheld details about the institutions or countries involved.
Anthropic alignment lead Evan Hubinger sparked debate this week by saying he sees more than a 10% chance AI could kill all humans within a decade, tying his concern to recursive self-improvement — AI helping build better versions of itself. Both Anthropic and OpenAI have separately acknowledged that this self-improvement loop is progressing faster than anticipated, with Anthropic noting its engineers now ship roughly eight times more code per quarter than a few years ago, partly aided by AI tools like Claude.
Anthropic's latest misuse report covers activity from November 2025 to September 2026 and details five cases where users tried to get its Claude AI to assist with dangerous biological research, three involving viruses and two involving toxins. In one instance, a request framed as a civilian grant application to enhance the chikungunya virus's mutation rate and virulence was traced to a military facility, prompting Anthropic to intervene. The company says all the actors used anonymization tools and tried to bypass its regional access blocks, which cover countries including China, Russia, Iran and North Korea.
Anthropic disclosed it banned several accounts after scientists in restricted countries tried to disguise research requests to bypass its safety controls. One case involved a scientist seeking help drafting a grant application to engineer more dangerous mutations of the chikungunya virus, seemingly for a military research institute.
OpenAI has privately pressed members of Congress for guidance on whether AI labs could legally coordinate to slow the pace of frontier AI development without violating antitrust law. The inquiry follows a blog post from chief scientist Jakub Pachocki calling for industry-wide coordination on safety, including voluntary slowdowns until shared safety standards emerge. Legal experts warn that such coordination could be seen as restricting output under the Sherman Antitrust Act, creating uncertainty that discourages collaboration.
An opinion column argues that artificial intelligence development is accelerating faster than society's ability to manage its risks, and calls for a deliberate pause to allow for more careful oversight. The author frames this as a matter of prudence rather than alarmism, suggesting current AI progress lacks sufficient safeguards.
Jacob Coxon, who worked on pretraining models at both OpenAI and Anthropic, announced his resignation from Anthropic on X, saying both companies are recklessly racing toward self-improving superintelligence. His post drew over 156 million views and sparked responses, including from Illinois Governor JB Pritzker and Anthropic's own alignment team lead Evan Hubinger, who publicly agreed that AI could pose an existential threat.
An AI safety expert cautioned that businesses are rapidly rolling out autonomous AI agents without adequate frameworks to manage the risks involved. The warning followed a day after an Anthropic researcher resigned, citing fears that fast-advancing AI systems could pose existential dangers to humanity.
A researcher named Jacob Coxon posted a resignation thread citing AI safety concerns, and the announcement unexpectedly spread far beyond the usual AI safety community. The post amplified earlier claims, including one from Evan Hubinger citing a greater than 10% probability of human extinction from AI, pushing extreme risk estimates into mainstream conversation. Commentators argue that rising AI capabilities and recent incidents made the public far more receptive to such warnings than in past years.
An Anthropic researcher resigned publicly, warning that AI labs are 'gambling with our lives,' and two colleagues echoed the concern. In response, Senator Bernie Sanders announced plans to introduce legislation banning the development of superintelligent AI.