Understanding FlashAttention Pt 1: Personal Notes
(news.ycombinator.com)
1.
2.
3.
Thelio Mira AI Linux Workstation: 192 GB GPU Memory
(news.ycombinator.com)
4.
What happens when a GPU writes memory
(news.ycombinator.com)
5.
What happens when a GPU reads memory
(news.ycombinator.com)
6.
Attention Decode on AMD MI450 GPUs: A Gluon Kernel Optimization Guide
(news.ycombinator.com)
7.
My first impressions on ROCm and Strix Halo
(news.ycombinator.com)
8.
Nvidia says it can shrink LLM memory 20x without changing model weights
(venturebeat.com)
9.
How to Spot (and Fix) 5 Common Performance Bottlenecks in Pandas Workflows
(news.ycombinator.com)
10.
11.
GPUHammer: Rowhammer attacks on GPU memories are practical
(news.ycombinator.com)
12.
NVIDIA shares guidance to defend GDDR6 GPUs against Rowhammer attacks
(bleepingcomputer.com)
13.
14.
AMD's Pre-Zen Interconnect: Testing Trinity's Northbridge
(news.ycombinator.com)