Skip to content
Tech News
clear
Topics: Today This Week This Month This Year
1.
Speculative Decoding in vLLM on AMD GPUs (news.ycombinator.com)
2.
DSpark: Speculative decoding accelerates LLM inference [pdf] (news.ycombinator.com)
3.
Eagle 3.1: Collaboration Between the EAGLE Team, vLLM Team, and TorchSpec Team (news.ycombinator.com)
4.
Accelerating Gemma 4: faster inference with multi-token prediction drafters (news.ycombinator.com)
Today's top topics: gemini openai google claude james webb space telescope google ai tilly norwood piers morgan flock safety license plate readers
View all today's topics →