Skip to content
Tech News
clear
Topics: Today This Week This Month This Year
1.
AirLLM 70B inference with single 4GB GPU (news.ycombinator.com)
2.
Microsoft laptop with Nvidia RTX Spark leaked and benchmarked before launch (techspot.com)
3.
Inflect-Micro-v2: complete voice in 9.36M parameters (news.ycombinator.com)
4.
Geekbench 7 debuts with a smarter multi-core benchmark and CUDA support (techspot.com)
5.
Geekbench 7 introduces biggest overhaul yet — real-world CPU testing, new media workloads, AI benchmarks, and CUDA support (tomshardware.com)
6.
Nvidia's first Windows on Arm GeForce driver confirms RTX Spark configurations (techspot.com)
7.
Nvidia has shipped 'hundreds of thousands of Grace standalone servers’ — GPU firm pivots messaging as CPUs take center stage in agentic data centers (tomshardware.com)
8.
Inference startup Infinity raises $15M from Touring Capital, OpenAI and Anthropic researchers (techcrunch.com)
9.
Inference startup Infinity raises $15M from Touring Capital, OpenAI and Athropic researchers (techcrunch.com)
10.
Nvidia DGX Spark as a daily driver (news.ycombinator.com)
11.
Alternative(s) to run CUDA on non-Nvidia hardware (news.ycombinator.com)
12.
A tiny London startup built a CUDA compiler that reportedly beats AMD's own tools on AMD hardware (techspot.com)
13.
Reduce GVisor Cold Starts with GPU Snapshotting (news.ycombinator.com)
14.
Zluda 6 release (run unmodified CUDA applications on non-Nvidia GPUs) (news.ycombinator.com)
15.
What happens when you run a CUDA kernel? (news.ycombinator.com)
16.
Show HN: NanoEuler – GPT-2 scale model in pure C/CUDA from scratch (news.ycombinator.com)
17.
I need your clothes, your boots, and your motorcycle (news.ycombinator.com)
18.
Show HN: cuTile Rust: Safe, data-race-free GPU kernels in Rust (news.ycombinator.com)
19.
What about OpenCL and CUDA C++ alternatives? (news.ycombinator.com)
20.
Tiny hackable CUDA language model implementation (news.ycombinator.com)
21.
Use your Nvidia GPU's VRAM as swap space on Linux (news.ycombinator.com)
22.
Microsoft debuts Surface RTX Spark Dev Box — Nvidia-powered mini-PC helps devs get ready for an agentic Windows (tomshardware.com)
23.
AI Agent Guidelines for CS336 at Stanford (news.ycombinator.com)
24.
Nvidia's long-awaited N1/N1X SoC specs leak ahead of Computex launch — N1 to feature up to 20 Arm-based cores, standard N1 equipped with 12- and 10-core configs (tomshardware.com)
25.
I put a datacenter GPU in my gaming PC (news.ycombinator.com)
26.
I Put a Datacenter GPU in My Gaming PC for £200 (news.ycombinator.com)
27.
Show HN: Tiny-vLLM – high performance LLM inference engine in C++ and CUDA (news.ycombinator.com)
28.
Show HN: KVBoost – chunk-level KV cache reuse for HuggingFace, 5–48x faster TTFT (news.ycombinator.com)
29.
Cutting inference cold starts by 40x with LP, FUSE, C/R, and CUDA-checkpoint (news.ycombinator.com)
30.
CUDA Books (news.ycombinator.com)
Today's top topics: spacex apple google openai billion anthropic android github android authority et al
View all today's topics →