Anthropic researcher warns AI could pose over 10% risk of killing all humans
Evan Hubinger, a safety researcher at Anthropic, said publicly that while today's AI models pose low risk, he estimates a greater than 10% chance that future AI systems could cause human extinction within ten years. His remarks followed a former Anthropic researcher's claim that neither Anthropic nor OpenAI are acting responsibly, and a report that Anthropic withheld its newest model from the UK's AI Safety Institute.