Skip to content
Tech News
clear
Topics: Today This Week This Month This Year
151.
Writer launches a ‘super agent’ that actually gets sh*t done, outperforms OpenAI on key benchmarks (venturebeat.com)
152.
It’s Qwen’s summer: new open source Qwen3-235B-A22B-Thinking-2507 tops OpenAI, Gemini reasoning models on key benchmarks (venturebeat.com)
153.
Show HN: Llm-benchmark – Benchmarks LLM-optimized code across multiple providers (news.ycombinator.com)
154.
Show HN: OrioleDB Beta12 Features and Benchmarks (news.ycombinator.com)
155.
Cache Benchmarks (news.ycombinator.com)
156.
Moonshot AI’s Kimi K2 outperforms GPT-4 in key benchmarks — and it’s free (venturebeat.com)
157.
AI agent benchmarks are broken (news.ycombinator.com)
158.
AI Agent Benchmarks Are Broken (news.ycombinator.com)
159.
Koala: A benchmark suite for performance-oriented shell-optimization research (news.ycombinator.com)
160.
Firefox 120 to Firefox 141 Web Browser Benchmarks (news.ycombinator.com)
161.
Show HN: Arch-Router – 1.5B model for LLM routing by preferences, not benchmarks (news.ycombinator.com)
162.
New benchmarks show SteamOS outperforming Windows 11 on Lenovo's handheld PC (techspot.com)
163.
Snapdragon 8s Gen 4 benchmarks show why Nothing Phone 3 might not be truly Elite (androidauthority.com)
164.
A Chinese firm has just launched a constantly changing set of AI benchmarks (technologyreview.com)
165.
Mistral just updated its open source Small model from 3.1 to 3.2: here’s why (venturebeat.com)
166.
Gigabyte Radeon RX 9060 XT Review: Great Value Gaming (wired.com)
Today's top topics: polymarket
View all today's topics →