121.
122.
Two different tricks for fast LLM inference
(news.ycombinator.com)
123.
124.
126.
As Rocks May Think
(news.ycombinator.com)
127.
128.
Waypoint-1: Real-Time Interactive Video Diffusion from Overworld
(news.ycombinator.com)
129.
130.
Three types of LLM workloads and how to serve them
(news.ycombinator.com)
131.
Weight Transfer for RL Post-Training in under 2 seconds
(news.ycombinator.com)
132.
133.
Launch HN: Tamarind Bio (YC W24) – AI Inference Provider for Drug Discovery
(news.ycombinator.com)
134.
Nvidia just admitted the general-purpose GPU era is ending
(venturebeat.com)
135.
Five Things to Know About Nvidia’s $20 Billion Licensing Deal
(feeds.content.dowjones.io)
136.
137.
138.
Post-transformer inference: 224× compression of Llama-70B with improved accuracy
(news.ycombinator.com)
139.
Vsora Jotunn-8 5nm European inference chip
(news.ycombinator.com)
140.
Principles of Vasocomputation
(news.ycombinator.com)
141.
Cloud-Native Computing Is Poised To Explode
(slashdot.org)
143.
144.
Ovi: Twin backbone cross-modal fusion for audio-video generation
(news.ycombinator.com)
145.
Ovi
(news.ycombinator.com)
146.
Elixir 1.19
(news.ycombinator.com)
147.
Cerebras systems raises $1.1B Series G
(news.ycombinator.com)
148.
Cerebras Systems Raises $1.1B Series G at $8.1B Valuation
(news.ycombinator.com)
149.
GPT-OSS Reinforcement Learning
(news.ycombinator.com)
150.
Show HN: Run Qwen3-Next-80B on 8GB GPU at 1tok/2s throughput
(news.ycombinator.com)