Building a Rust Inference Engine That Matches Llama.cpp
(news.ycombinator.com)
1.
2.
Homebench – Benchmark local LLMs for speed, memory, and quality
(news.ycombinator.com)
3.
The PImpl idiom and the C++26 std:indirect type
(news.ycombinator.com)
4.
Transcribe.cpp
(news.ycombinator.com)
5.
Same model, same Q4_K_M label: 5.02, 5.07 and 5.27 bits per weight
(news.ycombinator.com)
6.
Improvements to Std:Format in C++26
(news.ycombinator.com)
7.
8.
9.
10.
Running local models is good now
(news.ycombinator.com)
11.
Orthodox C++
(news.ycombinator.com)
12.
How to setup a local coding agent on macOS
(news.ycombinator.com)
13.
How to Setup a Local Coding Agent on macOS
(news.ycombinator.com)
14.
A 10 year old Xeon is all you need
(news.ycombinator.com)
15.
A 10 year old Xeon is all you need (for 26B-A4B MTP Drafters without GPU)
(news.ycombinator.com)
16.
Odysseus – self-hosted AI workspace
(news.ycombinator.com)
17.
Liquid AI reveals 8B-A1B MoE trained on 38T
(news.ycombinator.com)
18.
Social Animus
(news.ycombinator.com)
19.
A Comma and a Question Mark, Redux: Quick Terminal Helpers Using Pi
(news.ycombinator.com)
20.
A Comma and a Question Mark
(news.ycombinator.com)
21.
22.
DeepSeek-V4-Flash means LLM steering is interesting again
(news.ycombinator.com)
23.
What's in a GGUF, besides the weights – and what's still missing?
(news.ycombinator.com)
24.
Hugging Face Packages Weaponized With a Single File Tweak
(darkreading.com)
25.
Running local models on an M4 with 24GB memory
(news.ycombinator.com)
26.
How do I deal with memory leaks? (2022)
(news.ycombinator.com)
27.
Bjarne Stroustrup: How do I deal with memory leaks? (2022)
(news.ycombinator.com)
28.
Bjarne Stroustrup: How do I deal with memory leaks?
(news.ycombinator.com)
29.
DeepSeek 4 Flash local inference engine for Metal
(news.ycombinator.com)
30.
Show HN: Adam – An embeddable cross-platform AI agent library
(news.ycombinator.com)