Anthropic released Opus 5.5, claiming it beats its larger Fable model on many coding and knowledge benchmarks while costing 20% less per output token and running faster. The model also communicates more plainly, avoiding jargon and leading with key information. Sonnet 5.5 and Haiku 5.5 are expected to follow in coming weeks with similar gains.
Anthropic released Claude Opus 5.5 less than two months after Opus 5, claiming performance near its top-tier Fable 5.1 model while running about 40% cheaper. The company says the model uses fewer tokens, produces less verbose answers without sacrificing accuracy, and generates output over 30% faster than its predecessor. Sonnet 5.5 and Haiku 5.5 versions are expected in coming weeks.
Anthropic released Claude Opus 5.5, the first entry in its 5.5 model family, matching the performance of Claude Fable 5.1 on most tasks while running 40% cheaper than its predecessor, Opus 5. The model underwent external evaluation by groups including Frontier Design and METR, and scored higher than any prior Anthropic model on the company's internal automated behavioral audit for alignment and safety.
A developer tested whether repeatedly asking modern agentic LLMs to improve code could yield genuine performance gains, this time using Rust rather than Python. After months of experimentation following the release of Opus 4.5, the author found that these models can produce Rust code significantly faster than current state-of-the-art implementations, provided they are given proper constraints and guardrails.
TypeSafe's Jev, a rapidly adopted AI classification model, has become the fastest-growing model in Vercel's AI Gateway history. A commentator argues OpenAI could quickly replicate Jev's core approach and bundle it directly into its own models and agents, since OpenAI's LLMs already function as implicit classifiers that could be retrained for this purpose.
OPPO has launched the Find X10 Pro Max in China, becoming the first smartphone to pack three separate 200MP camera sensors—a main, ultrawide, and periscope telephoto lens, all with Hasselblad tuning. The phone runs on MediaTek's new Dimensity 9600 Pro chip and houses an 8,000mAh silicon-carbon battery with 80W wired and 50W wireless charging. Pre-orders are open in China now, with a global release planned later in 2026.
1Retro is a new cross-platform tool for Windows, macOS and Linux that automatically backs up and synchronizes retro game save files across devices and emulators. It runs in the background monitoring save folders, uploads changed files to the cloud, and offers command-line tools for advanced users, with support for RetroArch, OpenEmu, MiSTer, Analogue Pocket, Steam Deck and other platforms. The service is free to start, with paid upgrades for more storage and features.
Researcher Carter Leffer submitted a decryption of the German Army Enigma message MVUEH, sent 10 July 1941 and logged by an SS-Totenkopf radio unit, which had remained unbroken since 2005. The AI-assisted break, using OpenAI's GPT-6 Astra, revealed a completely different key setup—including a different wheel order—than other messages from that day, yet produced plaintext nearly identical to a previously solved message, Nr. 173 (SIPVX). Analysis of the newly recovered key also uncovered transcription errors in the original ciphertext and pinpointed the exact letter at which the Enigma machine's left wheel turned over.
A new Ornn Data paper argues that NVIDIA's older Ampere-generation GPUs, particularly the A100, remain economically useful far longer than assumed because open-weight models can be run cheaply on them. The analysis found that self-hosted open-weight models like gpt-oss-120b produce output more cheaply on A100 chips than on newer H100 hardware, and that A100 rental prices have stayed unusually stable over a five-year contract period compared to Hopper and Blackwell chips.
404 Media reports that contractors hired to review and critique real ChatGPT conversations to improve OpenAI's models have been terminated for secretly using AI tools to generate their feedback instead of doing it themselves. Multiple contractors told the outlet this is a widespread and frequently punished practice among the thousands of workers doing this human-review work.
Nothing has opened its Android 17-based OS 5.0 beta to three more devices: the Phone 4a, Phone 3a Pro, and Phone 3a. The update brings a visual overhaul with a new Geist typeface, redesigned Settings and Camera apps, split notification panels, transparent widgets, and more customization options. Users have only one week to enroll before the beta window closes, after which they must wait for the stable release expected in early November.
A Show HN post introduces JevBench, a benchmark that measures typed decision models by cost per 1,000 decisions rather than per token, using actual token counts from 534 v1.2 test decisions. The methodology prices systems with public tariffs at their listed per-token rates, while unlisted open-weight models are priced using OpenRouter or DeepInfra hosting rates for the same or comparable weights, explicitly avoiding raw GPU rental costs. Reported figures include Jev 1.13.0 at $0.0399 per 1,000 decisions, SemIf at roughly $0.022, and Winnow-12B Q8 at roughly $0.037.