Tech News
← Home  ·  All topics

Ai Safety Evaluators

3 GoKawiil briefs on this topic

Anthropic taps Accenture's Faculty unit as embedded AI safety evaluator

Anthropic announced that Accenture, through its AI division Faculty (acquired in January), will place staff inside the company to red-team models, run alignment assessments, and test safeguards. Anthropic and Accenture plan to jointly invest at least $1 billion over five years in the arrangement, part of Dario Amodei's push for third-party evaluators embedded within AI labs. Accenture shares rose 8% after hours following the news.

AI Safety Evaluators Demand Independence Guarantees from Anthropic, OpenAI

More than 100 AI researchers and safety evaluators, including Geoffrey Hinton and representatives from Johns Hopkins, Stanford and METR, signed a public letter urging foundation model developers to grant third-party testers genuine independence, transparency and legal protections. The letter, organized by the AI Evaluator Forum and shared exclusively with CNBC, follows Anthropic CEO Dario Amodei's recent proposal to give some evaluators 'employee-like access' to inspect frontier models.

Anthropic, OpenAI pledge to embed independent safety evaluators inside AI labs

Anthropic CEO Dario Amodei proposed letting third-party evaluators like METR and Redwood Research operate inside frontier AI companies with deep access to systems and training data, not just finished models. OpenAI's Sam Altman said his company would adopt a similar approach. Evaluators welcomed the idea but say specifics—and possibly legislation—are needed to ensure genuine independence rather than vendor-style arrangements.