Skip to content
Tech News
clear
Topics: Today This Week This Month This Year
451.
Gemini 2.5 Deep Think (news.ycombinator.com)
452.
Deep Think in the Gemini app (news.ycombinator.com)
453.
Benchmarking MicroPython (news.ycombinator.com)
454.
Benchmarks in CI: Escaping the Cloud Chaos (news.ycombinator.com)
455.
Show HN: Terminal-Bench-RL: Training long-horizon terminal agents with RL (news.ycombinator.com)
456.
Writer launches a ‘super agent’ that actually gets sh*t done, outperforms OpenAI on key benchmarks (venturebeat.com)
457.
Show HN: Terminal-Bench-RL: Training Long-Horizon Terminal Agents with RL (news.ycombinator.com)
458.
Ryzen 7 9800X3D vs. Ryzen 5 7600X: CPU and GPU Scaling Benchmark (techspot.com)
459.
It’s Qwen’s summer: new open source Qwen3-235B-A22B-Thinking-2507 tops OpenAI, Gemini reasoning models on key benchmarks (venturebeat.com)
460.
VC Victor Lazarte is leaving Benchmark to launch his own firm (techcrunch.com)
461.
I wasted weeks hand optimizing assembly because I benchmarked on random data (news.ycombinator.com)
462.
VectorDB bench now support S3Vector (news.ycombinator.com)
463.
A new AI coding challenge just published its first results — and they aren’t pretty (techcrunch.com)
464.
A new AI coding challenge just published its first results – and they aren’t pretty (techcrunch.com)
465.
Show HN: Llm-benchmark – Benchmarks LLM-optimized code across multiple providers (news.ycombinator.com)
466.
Show HN: OrioleDB Beta12 Features and Benchmarks (news.ycombinator.com)
467.
Benchmark in talks to lead Series A for Greptile, valuing AI-code reviewer at $180M, sources say (techcrunch.com)
468.
Grok 4 benchmark results: Tops math, ranks second in coding (bleepingcomputer.com)
469.
AI coding tools are shifting to a surprising place: The terminal (techcrunch.com)
470.
Cache Benchmarks (news.ycombinator.com)
471.
Galaxy Z Flip 7’s Exynos 2500 benchmarked: Flagship power or foldable flop? (androidauthority.com)
472.
Show HN: DesignArena – crowdsourced benchmark for AI-generated UI/UX (news.ycombinator.com)
473.
Moonshot AI’s Kimi K2 outperforms GPT-4 in key benchmarks — and it’s free (venturebeat.com)
474.
The ChompSaw: A benchtop power tool that's safe for kids to use (news.ycombinator.com)
475.
AI agent benchmarks are broken (news.ycombinator.com)
476.
AI Agent Benchmarks Are Broken (news.ycombinator.com)
477.
The ChompSaw: A Benchtop Power Tool That's Safe for Kids to Use (news.ycombinator.com)
478.
Former Intel CEO launches a benchmark to measure AI alignment (techcrunch.com)
479.
Why Brembo uses endurance racing as a test bench for brake development (arstechnica.com)
480.
Koala: A benchmark suite for performance-oriented shell-optimization research (news.ycombinator.com)
Today's top topics: android authority artificial intelligence anthropic donald trump openai polymarket
View all today's topics →