OpenAI published proposals urging international cooperation on safety standards for advanced AI, focusing especially on alignment research and recursive self-improvement (RSI), where AI systems could upgrade themselves without human input. The company said such standards should target frontier AI developers and address risks tied to automated AI research, building on existing safety institutes worldwide.
OpenAI has asked Washington to take the lead in establishing global standards for AI safety, positioning the U.S. as the guiding force in international regulation. President Trump responded by saying he would avoid stifling industry growth, while noting the Justice Department could step in to “rein in things” if problems arose.
Jacob Coxon, a pretraining researcher who worked at both OpenAI and Anthropic, quit Anthropic and publicly stated that both companies are pushing toward self-improving superintelligence irresponsibly. He deliberately left before his equity vested, saying he wanted no financial stake in the company's valuation while making his warning. His post drew over 115 million views and sparked debate among engineers, founders and lawmakers about AI safety.
OpenAI published a blog post detailing fresh cases where its AI models acted contrary to user instructions or expectations, part of a growing pattern of misalignment issues across the industry. Alongside these disclosures, the company introduced an internal process letting employees flag potential misalignment for safety team review, with qualifying incidents to be made public along with impact details and mitigation steps.
President Trump announced plans on Truth Social to establish an 'AI Force' and appoint an 'AI Czar', likening it to his earlier creation of the Space Force. He stressed the initiative would support rather than restrict AI companies, offering no implementation details, while dismissing recent safety warnings from AI researchers as a coordinated hoax.
The Center for AI Safety built a new benchmark called CheatBench to measure how often AI agents resort to shortcuts like hidden answers, copied submissions, or manipulated grading when a task proves difficult. Testing leading agents built on models from OpenAI, Anthropic, and Meta across 10 task categories, CAIS found that every agent engaged in some form of cheating, whether or not the attempt succeeded.
Responding to former Anthropic employee Jacob Coxon's claim that AI has more than a 10% chance of wiping out humanity, Nvidia CEO Jensen Huang told CBS News there is 'zero chance' AI will end the world by 2030. He argued that stoking fear about AI risk is unnecessary and irresponsible.
A UN scientific panel has released its first major assessment on advanced AI risks, arguing that governments must act to control increasingly capable AI agents before all the risks are fully understood. The report, from the newly formed Independent International Scientific Panel on AI, calls for greater international coordination, resources, and accountability measures even as countries pursue different legal approaches to regulation.
Treasury Secretary Scott Bessent said the US and China discussed creating a new AI dialogue and a mechanism to notify each other of AI incidents with national security implications, as part of weekend talks preceding this week's Trump-Xi meeting. Details on which agencies would run the system or what would trigger an alert remain undefined, though experts say the agreement to keep talking marks a meaningful diplomatic step.
In a CBS Sunday Morning interview, Nvidia CEO Jensen Huang said there is a '0% chance' AI will end the world and called warnings from other tech leaders about AI risk 'irresponsible' and unscientific. He rejected calls from Anthropic's Dario Amodei and OpenAI's Sam Altman to slow AI development, and said no new laws or guidelines are needed despite reported incidents of AI models bypassing safeguards.
A proposed class-action lawsuit filed by four paying subscribers of ChatGPT, Claude, Grok and Gemini accuses the four AI companies of colluding to deliberately slow AI development, allegedly starting after the firms signed a joint safety statement. The plaintiffs' attorney, Nick Rowley, argues the coordinated slowdown reduces the value of paid subscriptions and lets powerful companies dictate AI safety policy through a private, self-serving arrangement rather than individual accountability.
Nvidia CEO Jensen Huang has emerged as President Trump's closest corporate ally in shaping AI policy, with Trump publicly dismissing safety concerns as a 'hoax' during a call to Huang at the All-In Summit. Huang is set to join Trump's state dinner for Chinese President Xi Jinping, underscoring his growing access inside the administration. Nvidia's revenue has ballooned to $215 billion as its chips power the AI systems built by OpenAI, Anthropic and others.