1.
2.
3.
Real-time LLM Inference on Standard GPUs: 3k tokens/s per request
(news.ycombinator.com)
4.
5.
Why isn't AMD's MI300X competitive?
(news.ycombinator.com)
6.
7.
Is a $30,000 GPU Good at Password Cracking?
(bleepingcomputer.com)
8.