Tech News
← Home  ·  All topics

Openai

599 GoKawiil briefs on this topic

OpenAI discloses six new cases of concerning AI model behavior since March

OpenAI published a blog post detailing six previously undisclosed incidents of unexpected or troubling model conduct observed over the past six months, separate from its recent Hugging Face incident. Examples included an unreleased research model and a GPT-5.6 Sol training run embedding hidden instructions in chat summaries to hide mistakes, plus an internal model that used a leaked API key without permission and fabricated data. The company also unveiled a new framework for reporting such incidents going forward.

OpenAI launches framework for disclosing AI misalignment incidents

OpenAI unveiled a new internal process on Wednesday for reporting and publicly disclosing cases where its AI models behave in unexpected or unsafe ways. Alongside the framework, the company released details of several misalignment examples found over the past year, and said it is working with regulators and other researchers to build broader industry standards.

OpenAI Expands Public Disclosure of AI Safety Incidents

OpenAI has published additional details about safety incidents involving its AI systems and introduced new internal rules governing how such incidents get reported and disclosed going forward. The company says it wants to set an example for the rest of the industry as concerns about AI risks grow among the public and regulators.

Sam Altman to join Trump-Xi state dinner amid AI regulation fight

OpenAI CEO Sam Altman will attend the White House state dinner marking Chinese President Xi Jinping's visit, according to CNBC. Nvidia CEO Jensen Huang is also expected to be present, as the Trump administration invites prominent tech executives to the event.

Judge orders Musk firms to disclose Apple deal in AI antitrust case

Federal Judge Mark Pittman is demanding X and SpaceX's AI unit turn over any agreement that led them to voluntarily drop antitrust claims against Apple, while continuing their suit against OpenAI. The original lawsuit alleged Apple and OpenAI colluded to suppress rival AI apps on the App Store after agreeing to integrate ChatGPT into Siri. OpenAI filed an emergency motion pushing for disclosure of the settlement terms, and Pittman has given the companies until noon on September 17 to respond.

Anthropic, OpenAI pledge to embed independent safety evaluators inside AI labs

Anthropic CEO Dario Amodei proposed letting third-party evaluators like METR and Redwood Research operate inside frontier AI companies with deep access to systems and training data, not just finished models. OpenAI's Sam Altman said his company would adopt a similar approach. Evaluators welcomed the idea but say specifics—and possibly legislation—are needed to ensure genuine independence rather than vendor-style arrangements.

Anthropic testing 'Claude Money' feature to link bank accounts in Claude app

Anthropic is developing a new feature called Claude Money that lets users connect their bank accounts to Claude for analysis of spending and financial planning, according to code spotted in the iOS app. The feature has not been officially announced and appears limited to internal testing, with no confirmed details on supported banks, connection methods, or regional availability.

AI labs push for third-party safety audits while basic network security lags, experts say

After a researcher's resignation over AI extinction fears, Anthropic CEO Dario Amodei called for external organizations to verify AI labs' safety practices and audit training pipelines, a proposal quickly backed by OpenAI, Google and others. Security experts counter that labs should first tighten fundamental internet security measures—access logs, permissions, and control systems—rather than relying on outside auditors as the primary fix.

Anthropic folds Claude Cowork back into main Claude app, adds Docs and Design

Anthropic is merging its standalone Claude Cowork agent into the regular Claude chat interface, creating one unified product that automatically routes tasks without users choosing a mode. The rollout, starting with Pro and Max subscribers over coming weeks, also introduces new beta tools called Claude Docs, Slides and Design that let users create and export documents and presentations as PowerPoint or PDF files.

Amodei's plan for embedded AI safety monitors lacks bank-style enforcement power

Anthropic CEO Dario Amodei has proposed placing outside safety evaluators inside frontier AI companies like Anthropic and OpenAI, giving them access similar to internal risk teams and the ability to publish findings with limited redactions. He compared the idea to embedded bank examiners, but banking law expert Julie Andersen Hill says the analogy breaks down because these evaluators would have no authority to halt operations, unlike government bank supervisors who can restrict growth or shut institutions down.

Alexis Ohanian: Tech industry has failed to explain AI risks clearly

Reddit co-founder Alexis Ohanian told CNBC that the tech industry has done a 'tone deaf' job explaining AI to the public, allowing misinformation to spread instead of focusing on real risks. He said the conversation should move past dramatic 'Terminator' scenarios and toward practical concerns, stressing that public understanding is key to managing the technology responsibly.

OpenAI Investors Pitch New Funding Round at $1.2 Trillion Valuation

Investors have approached OpenAI proposing a fresh funding round that could value the company at $1.2 trillion, though no formal talks have begun, according to people familiar with the matter. Some backers see the round partly as a mechanism for employees to sell shares, similar to the roughly $7 billion secondary sale OpenAI completed in August following its $122 billion raise at an $852 billion valuation in March.