Performance per dollar is getting faster and cheaper
(news.ycombinator.com)
1.
2.
GLM5.2 on AMD MI355X at 2626 tok/s/node at over 2x lower cost than Blackwell
(news.ycombinator.com)
3.
Boosting multimodal inference performance by >10% with a single Python dict
(news.ycombinator.com)
4.
DeepSeek-V4 on Day 0: From Fast Inference to Verified RL with SGLang and Miles
(news.ycombinator.com)
5.
Inference startup Inferact lands $150M to commercialize vLLM
(techcrunch.com)
6.
7.
GLM-4.7-Flash
(news.ycombinator.com)
8.
NVIDIA DGX Spark In-Depth Review: A New Standard for Local AI Inference
(news.ycombinator.com)
9.
Deploying DeepSeek on 96 H100 GPUs
(news.ycombinator.com)