Gemini 2.5 Deep Think
(news.ycombinator.com)
451.
452.
Deep Think in the Gemini app
(news.ycombinator.com)
453.
Benchmarking MicroPython
(news.ycombinator.com)
454.
Benchmarks in CI: Escaping the Cloud Chaos
(news.ycombinator.com)
455.
Show HN: Terminal-Bench-RL: Training long-horizon terminal agents with RL
(news.ycombinator.com)
456.
457.
Show HN: Terminal-Bench-RL: Training Long-Horizon Terminal Agents with RL
(news.ycombinator.com)
458.
459.
460.
VC Victor Lazarte is leaving Benchmark to launch his own firm
(techcrunch.com)
461.
I wasted weeks hand optimizing assembly because I benchmarked on random data
(news.ycombinator.com)
462.
VectorDB bench now support S3Vector
(news.ycombinator.com)
463.
464.
465.
Show HN: Llm-benchmark – Benchmarks LLM-optimized code across multiple providers
(news.ycombinator.com)
466.
Show HN: OrioleDB Beta12 Features and Benchmarks
(news.ycombinator.com)
467.
468.
Grok 4 benchmark results: Tops math, ranks second in coding
(bleepingcomputer.com)
469.
AI coding tools are shifting to a surprising place: The terminal
(techcrunch.com)
470.
Cache Benchmarks
(news.ycombinator.com)
471.
Galaxy Z Flip 7’s Exynos 2500 benchmarked: Flagship power or foldable flop?
(androidauthority.com)
472.
Show HN: DesignArena – crowdsourced benchmark for AI-generated UI/UX
(news.ycombinator.com)
473.
474.
The ChompSaw: A benchtop power tool that's safe for kids to use
(news.ycombinator.com)
475.
AI agent benchmarks are broken
(news.ycombinator.com)
476.
AI Agent Benchmarks Are Broken
(news.ycombinator.com)
477.
The ChompSaw: A Benchtop Power Tool That's Safe for Kids to Use
(news.ycombinator.com)
478.
Former Intel CEO launches a benchmark to measure AI alignment
(techcrunch.com)
479.
Why Brembo uses endurance racing as a test bench for brake development
(arstechnica.com)
480.
Koala: A benchmark suite for performance-oriented shell-optimization research
(news.ycombinator.com)
Today's top topics:
android authority
artificial intelligence
anthropic
donald trump
openai
polymarket