1.
2.
KVarN: Native vLLM backend for KV-cache quantization by Huawei
(news.ycombinator.com)
4.
5.
Optimize for change not application performance
(news.ycombinator.com)
6.
447 TB/cm² at zero retention energy – atomic-scale memory on fluorographane
(news.ycombinator.com)
7.
8.
9.
VLLM: Easy, Fast, and Cheap LLM Serving with PagedAttention
(news.ycombinator.com)