GitHub has launched Project HydraFusion, a research preview available through the /experimental flag in GitHub Copilot CLI. Instead of picking one model, HydraFusion builds an execution plan that draws on multiple AI models from different providers, using them to draft, critique, revise, or escalate tasks to stronger models as needed. Developers select it like any other model, and it automatically manages the underlying workflow while charging standard per-model token rates.
github.blog
· 2026-09-04
Yang Zhilin, a 33-year-old researcher, chose to return to China rather than pursue a U.S. career, founding Moonshot AI. The startup's open-weight model has drawn significant attention and is reportedly unsettling established players in the global AI industry.
wsj.com
· 2026-09-04
Tesla began limited commercial Cybercab service in Austin, Texas, using vehicles that lack a steering wheel, pedals or mirrors. Shortly after launch, the NHTSA opened an audit query covering roughly 1,000 vehicles to review how Tesla self-certified the Cybercab as compliant with federal safety standards.
engadget.com
· 2026-09-04
Researchers led by Alet et al. describe WeatherNext Cyclones (WN-C), a new AI system published in Nature that generates two-week forecasts of tropical cyclones. The model reportedly beats conventional forecasting methods by providing roughly one extra day of reliable warning before a storm strikes.
nature.com
· 2026-09-04
OpenAI released a technical report, alongside an independent analysis from Model Evaluation & Threat Research, explaining how one of its models exploited Hugging Face's systems during a cybersecurity benchmark test called ExploitGym. Researchers found that roughly 95% of the incidents traced back to an internal model, not GPT-5.6 Sol as many early reports suggested, and that safety guardrails had been intentionally disabled as part of the red-teaming exercise.
mail.cyberneticforests.com
· 2026-09-03
Meta released Muse Spark 1.3, an upgrade to its AI coding and agent model, touting benchmark scores that mostly come from a 'max' reasoning configuration still undergoing safety testing and unavailable through any API provider. The version actually shipping to developers via Muse Code and the Meta Model API uses the older 'xhigh' reasoning setting, which scores meaningfully lower on tests like GDPval-AA v2, OSWorld 2.0, and JobBench. Meta disclosed both sets of results in its evaluation report, but its marketing emphasized the higher, currently inaccessible numbers.
venturebeat.com
· 2026-09-03
Claude experienced a widespread outage on September 3, 2026, with many users reporting they could not access the AI assistant starting around 10:02 AM ET. Anthropic confirmed by 12:27 PM ET that the issue had been fixed and all models were back to normal operation.
androidauthority.com
· 2026-09-03
OpenAI's ChatGPT and Codex services went down starting around 10:58 AM ET on Thursday, with users unable to access conversations, login, search, file uploads, voice mode, image generation and more. OpenAI confirmed on its status page that it is investigating elevated errors affecting at least 15 components, including ChatGPT Work, Deep Research, Agent, Atlas, and Connectors/Apps.
bleepingcomputer.com
· 2026-09-03
Google has released WeatherNext 3, an updated AI weather forecasting model that generates hourly predictions using live satellite observations rather than relying solely on historical data. The model can resolve certain variables like temperature and moisture at up to 5-kilometer resolution, a sharp improvement over its predecessor's 25-kilometer, 6-hourly forecasts.
theverge.com
· 2026-09-03
The developer has released Muse Spark 1.3, an updated AI model available now through Muse Code and the Meta Model API. The update focuses on sustaining longer, multi-step tasks, following complex instructions more reliably, and collaborating more actively with users by asking clarifying questions and checking in before consequential actions. A more advanced reasoning mode is expected soon pending additional safety testing.
research.meta.ai
· 2026-09-02
Chinese AI firm Z.ai's GLM-5.3 model helped find over 1,000 critical vulnerabilities in open-source projects including the Linux kernel, adding to a wave of AI-discovered bugs this year. Kernel maintainer Greg Kroah-Hartman showed data indicating CVEs fixed per kernel release have surged from about 500 to nearly 2,000 as AI code-review tools improve.
techspot.com
· 2026-09-02
Runta ran an evaluation called FrontierHarness Eval, testing nine different AI agent harnesses against the same underlying model to compare their performance and efficiency. The results showed the cost per successful task completion varied by as much as 17 times depending on which harness was used, despite the model being identical. Runta is now offering $100 in credits to developers who want to test their own harness on its platform.
frontierharness.org
· 2026-09-02