Tech News
← Home  ·  All topics

Claude Opus 5

17 GoKawiil briefs on this topic

OpenAI releases GPT-6 Sol and Luna models with lower API pricing

OpenAI launched GPT-6 Sol and GPT-6 Luna, positioned as cheaper, more accurate successors to GPT-5.6 Sol and Luna, with API costs cut by 50%. OpenAI says GPT-6 Sol outperforms Claude Opus 5 at roughly 9% of its cost and matches Claude Fable 5.1 on coding tasks at lower cost, while making about half as many errors as its predecessor. Both models are rolling out today in ChatGPT Work and Codex for Pro, Plus, Business, Enterprise, and Edu users, with Luna also available to Free and Go plan users in the desktop app.

Anthropic launches Claude Opus 5.5 with faster, cheaper responses

Anthropic released Claude Opus 5.5, which it says outperforms its predecessor Opus 5 while responding over 30% faster and costing about 40% less to run. The company also says it improved the model's communication style and added new safety safeguards, and it is raising usage limits for Pro, Max, Team, and Enterprise subscribers.

AI agents design and Lean-verify new shortest-path algorithm beating published bounds

Ten Claude Opus 5.5 agents were tasked with finding a faster exact shortest-path algorithm for directed graphs with non-negative real weights and proving it correct in the Lean theorem prover. Over roughly 15 hours and 733 messages, the agents produced an algorithm called C-HD, which its creators say formally improves on previously published complexity bounds, including Dijkstra's algorithm and two recent 2025 and 2026 papers.

Anthropic Releases Claude Opus 4.5 Days After CEO's AI Slowdown Remarks

Anthropic launched a new AI model, Claude Opus 4.5, roughly ten days after its CEO publicly called for a slowdown in AI development. The company says the model performs well on internal alignment safety measures, while also stating that government policy will eventually be needed to guard against risks from AI systems that can improve themselves.

OpenAI and Anthropic launch cheaper model tiers amid open-weight competition

OpenAI added two new GPT-6 tiers, Sol and Luna, cutting API prices by 50% versus GPT-5.6 promotional rates, with Sol built for complex coding tasks and Luna aimed at high-volume document work. Anthropic released Claude Opus 5.5, a more token-efficient model it says costs about 40% less to run than Opus 5.

Claude Opus 5.5 scores 58 on Artificial Analysis Intelligence Index, priced above median

Anthropic's Claude Opus 5.5, running in Adaptive Reasoning Max Effort mode, scored 58 on the Artificial Analysis Intelligence Index against a median of 25 among comparable models. The model handles text and image input with text output and a 1M token context window, but is priced at $4.00 per 1M input tokens and $20.00 per 1M output tokens, both above the reported medians of $2.00 and $10.00. Evaluating it on the Intelligence Index generated 260M tokens and cost $8,708.20 in total.

Anthropic launches Claude Opus 5.5 with faster, cheaper performance gains

Anthropic has released Claude Opus 5.5, an upgrade to its top-tier Opus model that arrives roughly two months after the previous version, in line with the company's usual release pace. The company says it performs close to its higher-end Fable/Mythos tier while being about 30% faster and 40% cheaper per task than Opus 5, with notable gains in large-scale coding work and software optimization tasks.

Anthropic launches Claude Opus 5.5, cutting compute costs 40%

Anthropic released Claude Opus 5.5 less than two months after Opus 5, claiming performance near its top-tier Fable 5.1 model while running about 40% cheaper. The company says the model uses fewer tokens, produces less verbose answers without sacrificing accuracy, and generates output over 30% faster than its predecessor. Sonnet 5.5 and Haiku 5.5 versions are expected in coming weeks.

Anthropic launches Claude Opus 5.5, cutting inference costs 40% versus Opus 5

Anthropic released Claude Opus 5.5, the first entry in its 5.5 model family, matching the performance of Claude Fable 5.1 on most tasks while running 40% cheaper than its predecessor, Opus 5. The model underwent external evaluation by groups including Frontier Design and METR, and scored higher than any prior Anthropic model on the company's internal automated behavioral audit for alignment and safety.

Anthropic releases Claude Opus 5.5 with tighter cybersecurity safeguards

Anthropic introduced Claude Opus 5.5, a cheaper, more efficient model that reroutes risky cybersecurity requests to the weaker Opus 4.8 and flagged biology queries to Opus 5. The company says it is the top performer on its internal alignment testing and was vetted by outside evaluators Frontier Design and METR before release.

Hacktron researchers used Anthropic's Claude to breach OpenAI employee accounts

Three independent security researchers at Hacktron reportedly used Anthropic's Claude Opus models to gain access to OpenAI employee accounts within 72 hours, exploiting a HEIF image-processing flaw in Discourse, the third-party service running OpenAI's community forums. They proved access by submitting a pull request through a compromised employee's Codex account but stopped short of touching OpenAI's proprietary code in its 'Monorepo' repository.

AutoBot agent tops AssistantBench leaderboard, beats OpenAI and Anthropic on OSWorld 2.0

Autonomous Production released AutoBot, an open-source agentic harness that lets a local Mac AI system handle long, multi-step knowledge work with live voice control. The company reports AutoBot scored 18.5% higher task completion than OpenAI's published Sol Max baseline and outperformed Anthropic's Claude Opus 5 Max on the OSWorld 2.0 benchmark, while also ranking first on AssistantBench's official hidden-test leaderboard with 50.70% accuracy across 181 tasks.