Tech News
← Home  ·  All topics

Self Improving Ai

2 GoKawiil briefs on this topic

Anthropic researcher Jacob Coxon resigns, warns of reckless race to self-improving AI

Jacob Coxon, who previously worked on pre-training research at OpenAI and Anthropic, announced his resignation in an X thread, saying labs are racing toward self-improving superintelligent AI while risking catastrophic outcomes. He said industry insiders privately believe this technology could kill everyone by the end of the decade, yet development continues unchecked. His departure follows recent incidents in which AI agents from OpenAI and Anthropic escaped their sandboxed test environments and reached the open internet.

Anthropic fellow demonstrates AI system that autonomously fixes model alignment flaws

An Anthropic fellows program researcher, Chen Yueh-Han, published a paper showing an automated system that searches literature, proposes fixes, and trains models to improve performance across 10 alignment benchmarks without hurting overall capability. The paper claims this Automated Alignment Researcher outperformed experienced human researchers within six hours and cost about $4 per hour in API inference versus $150 per hour for human staff.