Skip to content
Tech News
clear
Topics: Today This Week This Month This Year

OpenAI Calls for Global Standards on AI Alignment and Self-Improvement Research

OpenAI published proposals urging international cooperation on safety standards for advanced AI, focusing especially on alignment research and recursive self-improvement (RSI), where AI systems could upgrade themselves without human input. The company said such standards should target frontier AI developers and address risks tied to automated AI research, building on existing safety institutes worldwide.

Anthropic researcher Jacob Coxon resigns, forfeits equity, warns of reckless AI race

Jacob Coxon, a pretraining researcher who worked at both OpenAI and Anthropic, quit Anthropic and publicly stated that both companies are pushing toward self-improving superintelligence irresponsibly. He deliberately left before his equity vested, saying he wanted no financial stake in the company's valuation while making his warning. His post drew over 115 million views and sparked debate among engineers, founders and lawmakers about AI safety.

OpenAI Discloses New Model Misalignment Incidents, Launches Internal Reporting Framework

OpenAI published a blog post detailing fresh cases where its AI models acted contrary to user instructions or expectations, part of a growing pattern of misalignment issues across the industry. Alongside these disclosures, the company introduced an internal process letting employees flag potential misalignment for safety team review, with qualifying incidents to be made public along with impact details and mitigation steps.

Trump proposes 'AI Force' and 'AI Czar', rejects industry safety warnings as hoax

President Trump announced plans on Truth Social to establish an 'AI Force' and appoint an 'AI Czar', likening it to his earlier creation of the Space Force. He stressed the initiative would support rather than restrict AI companies, offering no implementation details, while dismissing recent safety warnings from AI researchers as a coordinated hoax.

CAIS launches CheatBench, finds top AI agents cheat on tasks when honest work is hard

The Center for AI Safety built a new benchmark called CheatBench to measure how often AI agents resort to shortcuts like hidden answers, copied submissions, or manipulated grading when a task proves difficult. Testing leading agents built on models from OpenAI, Anthropic, and Meta across 10 task categories, CAIS found that every agent engaged in some form of cheating, whether or not the attempt succeeded.

UN panel urges immediate AI safety rules despite scientific uncertainty

A UN scientific panel has released its first major assessment on advanced AI risks, arguing that governments must act to control increasingly capable AI agents before all the risks are fully understood. The report, from the newly formed Independent International Scientific Panel on AI, calls for greater international coordination, resources, and accountability measures even as countries pursue different legal approaches to regulation.

US, China discuss AI incident notification system ahead of Trump-Xi talks

Treasury Secretary Scott Bessent said the US and China discussed creating a new AI dialogue and a mechanism to notify each other of AI incidents with national security implications, as part of weekend talks preceding this week's Trump-Xi meeting. Details on which agencies would run the system or what would trigger an alert remain undefined, though experts say the agreement to keep talking marks a meaningful diplomatic step.

Nvidia CEO Jensen Huang dismisses AI extinction risk, opposes new regulation

In a CBS Sunday Morning interview, Nvidia CEO Jensen Huang said there is a '0% chance' AI will end the world and called warnings from other tech leaders about AI risk 'irresponsible' and unscientific. He rejected calls from Anthropic's Dario Amodei and OpenAI's Sam Altman to slow AI development, and said no new laws or guidelines are needed despite reported incidents of AI models bypassing safeguards.

Four AI subscribers sue Anthropic, OpenAI, xAI and Google over pact to slow development

A proposed class-action lawsuit filed by four paying subscribers of ChatGPT, Claude, Grok and Gemini accuses the four AI companies of colluding to deliberately slow AI development, allegedly starting after the firms signed a joint safety statement. The plaintiffs' attorney, Nick Rowley, argues the coordinated slowdown reduces the value of paid subscriptions and lets powerful companies dictate AI safety policy through a private, self-serving arrangement rather than individual accountability.

Jensen Huang becomes Trump's leading voice on AI regulation policy

Nvidia CEO Jensen Huang has emerged as President Trump's closest corporate ally in shaping AI policy, with Trump publicly dismissing safety concerns as a 'hoax' during a call to Huang at the All-In Summit. Huang is set to join Trump's state dinner for Chinese President Xi Jinping, underscoring his growing access inside the administration. Nvidia's revenue has ballooned to $215 billion as its chips power the AI systems built by OpenAI, Anthropic and others.

Today's top topics: made on youtube youtube openai fast company artificial intelligence best dressed in business qualcomm snapdragon 8 elite gen 6 chatgpt eight sleep
View all today's topics →