Skip to content
Tech News
clear
Topics: Today This Week This Month This Year
1.
Georgi Gerganov on llama.cpp/ggml future after Nvidia acquisition of HuggingFace (news.ycombinator.com)
2.
Run Qwen3.8 27B locally: real numbers from my Mac Studio (news.ycombinator.com)
3.
DFlash 2: Keep Drafting Parallel (news.ycombinator.com)
4.
Unsloth Dynamic 3.0 GGUFs (news.ycombinator.com)
5.
Show HN: Shoehorn – Quantize any model down to run on your machine (news.ycombinator.com)
6.
Llama.cpp v0.1.0 (news.ycombinator.com)
7.
Qwen3.8-27B at 256K on a 24GB RTX PRO 4000 SFF (432 GB/s): 50 tok/s with MTP (news.ycombinator.com)
8.
llama.cpp (news.ycombinator.com)
9.
Apple Silicon and macOS VMs: Faster LLM Inference with llama.cpp (news.ycombinator.com)
10.
Apple Silicon and macOS VMs: 11–16× Faster LLM Inference with Llama.cpp (news.ycombinator.com)
11.
Meta's 'Open' Muse Glimmer Model Can Run On a Single Computer (slashdot.org)
12.
Meta's 'open source' Muse Glimmer model can run on a single computer (engadget.com)
13.
Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows (news.ycombinator.com)
14.
Meta Muse Glimmer – Open weights 30B local coding model (news.ycombinator.com)
15.
Meta Muse Glimmer – open weights 30B local coding model (news.ycombinator.com)
16.
Building a Rust Inference Engine That Matches Llama.cpp (news.ycombinator.com)
17.
Homebench – Benchmark local LLMs for speed, memory, and quality (news.ycombinator.com)
18.
Same model, same Q4_K_M label: 5.02, 5.07 and 5.27 bits per weight (news.ycombinator.com)
19.
Ditching the cloud for local AI — how I use two mini PCs to process millions of tokens a day and save money on costly API fees (tomshardware.com)
20.
Russian Spam and Profanities Are Now Plaguing the Arch Linux AUR (slashdot.org)
21.
Running local models is good now (news.ycombinator.com)
22.
How to setup a local coding agent on macOS (news.ycombinator.com)
23.
How to Setup a Local Coding Agent on macOS (news.ycombinator.com)
24.
A 10 year old Xeon is all you need (news.ycombinator.com)
25.
A 10 year old Xeon is all you need (for 26B-A4B MTP Drafters without GPU) (news.ycombinator.com)
26.
Odysseus – self-hosted AI workspace (news.ycombinator.com)
27.
Liquid AI reveals 8B-A1B MoE trained on 38T (news.ycombinator.com)
28.
Social Animus (news.ycombinator.com)
29.
A Comma and a Question Mark, Redux: Quick Terminal Helpers Using Pi (news.ycombinator.com)
30.
A Comma and a Question Mark (news.ycombinator.com)
Today's top topics: promo codes openai wordle ai safety ios 27 apple ai alignment anthropic claude google
View all today's topics →