Skip to content
Tech News
clear
Topics: Today This Week This Month This Year
1.
Nvidia details Rubin architectural optimizations for inference – improvements target better performance and efficiency from the GPU to the rack (tomshardware.com)
2.
DeepSeek-V4 on Day 0: From Fast Inference to Verified RL with SGLang and Miles (news.ycombinator.com)
3.
4-bit floating point FP4 (news.ycombinator.com)
4.
Running local models on Macs gets faster with Ollama's MLX support (arstechnica.com)
5.
Ollama is now powered by MLX on Apple Silicon in preview (news.ycombinator.com)
6.
Huawei unveils new Atlas 350 AI accelerator with 1.56 PFLOPS of FP4 compute and up to 112GB of HBM — claims 2.8x more performance than Nvidia's H20 (tomshardware.com)
7.
SVDQuant+NVFP4: 4× Smaller, 3× Faster FLUX with 16-bit Quality on Blackwell GPUs (news.ycombinator.com)
Today's top topics: apple openai anthropic nvidia google meta android authority iphone samsung ai models
View all today's topics →