Micro-Agent: Beat Frontier Models with Collaboration Inside Model API
(news.ycombinator.com)
31.
32.
DSpark: Speculative decoding accelerates LLM inference [pdf]
(news.ycombinator.com)
33.
DeepSeek open-sources inference optimizations with 60–85% faster generation [pdf]
(news.ycombinator.com)
34.
35.
OpenAI unveils its first custom chip, built by Broadcom
(news.ycombinator.com)
36.
37.
OpenAI, Broadcom Develop Custom Chip for AI Inference
(feeds.content.dowjones.io)
38.
OpenAI unveils its first custom chip, built by Broadcom
(techcrunch.com)
39.
OpenAI’s new ‘Jalapeno’ chip is the company’s first step towards the future
(androidauthority.com)
40.
41.
OpenAI and Broadcom unveil LLM-optimized inference chip
(news.ycombinator.com)
42.
43.
OpenAI reveals its first AI processor: Jalapeño
(theverge.com)
44.
Disparate privacy risks from medical AI
(feeds.nature.com)
45.
Modal Auto Endpoints: Optimized inference you own
(news.ycombinator.com)
46.
Record type inference for dummies
(news.ycombinator.com)
47.
48.
49.
50.
51.
The Null Is Always False (Except When It Is True) (2014)
(news.ycombinator.com)
52.
54.
MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
(news.ycombinator.com)
55.
Claude AI: What's free in 2026 and what isn't?
(engadget.com)
56.
58.
Bringing Up DeepSeek-V4-Flash on AMD MI300X
(news.ycombinator.com)
59.
How is Groq raising more money?
(news.ycombinator.com)
Today's top topics:
google
pixel
longbets
longbets org
wins
www longbets
fish
incentives
right
sumner