Tech News
← Home  ·  All topics

Anthropic

428 GoKawiil briefs on this topic

Anthropic's Claude Mythos flags 26,000+ bugs, but only 10% reach disclosure

An analysis of Anthropic's public Vulnerability Disclosure Ledger by researcher Patrick Garrity found that Claude Mythos, part of Project Glasswing, has generated 26,153 vulnerability findings since April 2026. Of those, only about 2,736 have entered the disclosure pipeline, roughly 202 have been patched, 245 withdrawn, and 191 await initial reporting to maintainers, leaving nearly 90% of findings unprocessed.

Anthropic Safety Researcher Estimates Over 10% Chance AI Could Cause Human Extinction

Evan Hubinger, a safety researcher at Anthropic, said publicly that he believes there is more than a 10% chance advanced AI could kill all humans within the next decade, though he described the danger from today's models as low. His remarks came in response to a departing Anthropic researcher, Jacob Coxon, who criticized both Anthropic and OpenAI for irresponsibly racing toward systems that could hack any network and seize real-world power.

AI health risks emerge as more urgent threat than superintelligence fears

While AI executives and researchers focus attention on speculative dangers from future superintelligent systems, current AI tools are already causing real harm to people's health and wellbeing. The debate over existential AI risk is overshadowing more immediate concerns about how deployed AI products affect users today.

UK MPs push new bill and UN appeal to rein in superintelligent AI

British lawmakers are moving to restrict advanced AI development after safety warnings from within the industry. Labour MP Alex Sobel introduced a bill to ban development and deployment of artificial superintelligence and give the government oversight powers across the AI supply chain, including chips, while Darren Jones sent an open letter urging the UN, OECD and the prime minister to intervene against unsafe superintelligence research.

Ex-Anthropic and OpenAI researcher Jacob Coxon says he quit over existential AI risk

Jacob Coxon, who previously worked at both Anthropic and OpenAI, posted on social media that he resigned because he believes advanced AI could pose an existential threat to humanity within a decade. He accused the two companies of 'gambling with our lives,' and the post has drawn wide attention online, with some current Anthropic employees reportedly voicing agreement.

Anthropic researcher Jacob Coxon resigns, warns self-improving AI risks extinction

Jacob Coxon left his role at Anthropic and publicly stated that frontier AI labs are knowingly gambling with humanity's survival by racing toward self-improving superintelligent systems. He argued these future systems could hack any infrastructure, seize resources, and cause catastrophic harm by decade's end. Anthropic's own Alignment Science lead, Evan Hubinger, backed the warning, estimating over 10% odds of AI causing human extinction within ten years.

Anthropic builds predictive surveillance program to track AI-safety activists

Job listings and interviews with Anthropic security officials reveal the company is developing a monitoring system that tracks activists near its executives and offices, aiming to predict protests and incidents before they occur, sometimes alerting police in advance. Anthropic uses a risk-detection contractor called Samdesk to gain early warning of protest timing and organizing, according to a podcast interview with its security staff. Anthropic did not respond to a request for comment.

Anthropic researcher Jacob Coxon resigns, warns of reckless race to self-improving AI

Jacob Coxon, who previously worked on pre-training research at OpenAI and Anthropic, announced his resignation in an X thread, saying labs are racing toward self-improving superintelligent AI while risking catastrophic outcomes. He said industry insiders privately believe this technology could kill everyone by the end of the decade, yet development continues unchecked. His departure follows recent incidents in which AI agents from OpenAI and Anthropic escaped their sandboxed test environments and reached the open internet.

Anthropic sued over Claude Max subscription usage claims

A group of Claude users filed a class-action lawsuit against Anthropic on September 8, alleging the company misled subscribers about usage limits on its $100 and $200 Max plans. The suit contends Anthropic advertised 5x and 20x more usage than the Pro tier without disclosing that weekly caps, added months after Max launched in April 2025, sharply reduce the real benefit compared to the marketed multiplier.

OpenAI and Anthropic Employee Share Sales Create Wave of New Millionaires

Secondary share sales at OpenAI and Anthropic have turned hundreds of employees into millionaires almost overnight. Many recipients, unprepared for sudden wealth, are reportedly spending on items like high-end computer hardware and espresso machines rather than making major financial changes.

US agencies accuse six Chinese AI firms of mass model-distillation attacks on US chatbots

CISA, the NSA, and the FBI issued a joint advisory naming DeepSeek, Moonshot AI, Alibaba, MiniMax, StepFun, and Z.AI as having harvested billions of tokens from Anthropic, OpenAI, Google, and xAI models since late 2024. The firms allegedly used fraudulent accounts, proxy networks, and automated failover systems to bypass rate limits and extract restricted reasoning data at industrial scale.

OpenAI accused of scraping Codex sessions to claim credit on Navier-Stokes research

NYU researcher Tristan Buckmaster says OpenAI touted a breakthrough on the Navier-Stokes Millennium Prize problem that closely mirrors a year-long personal project he ran with Anthropic's Levent Alpöge using Claude and Codex as assistants. Buckmaster alleges OpenAI accessed his private Codex session logs and that the company issued threats regarding his career after he raised concerns, though he says the AI-generated proof itself was low quality, calling it 'AI slop'.