Shapelearn Qwen 3.8 27B (13.1 GB VRAM)
(news.ycombinator.com)
1.
2.
Benchmarking Qwen3.8 27B quantizations: 4-bit holds up, 1-bit collapses
(news.ycombinator.com)
3.
Unsloth Dynamic 3.0 GGUFs
(news.ycombinator.com)
4.
Unsloth Qwen3.8-27B GGUF files
(news.ycombinator.com)
5.
Qwen3.8-27B
(news.ycombinator.com)
6.
Building a Rust Inference Engine That Matches Llama.cpp
(news.ycombinator.com)
7.
GLM-5.2 – How to Run Locally
(news.ycombinator.com)
8.
Runing GLM-5.2 on local hardware
(news.ycombinator.com)
9.
Unsloth GLM-5.2 – How to Run Locally
(news.ycombinator.com)
10.
How to setup a local coding agent on macOS
(news.ycombinator.com)
11.
How to Setup a Local Coding Agent on macOS
(news.ycombinator.com)
12.
What's in a GGUF, besides the weights – and what's still missing?
(news.ycombinator.com)
13.
Advanced Quantization Algorithm for LLMs
(news.ycombinator.com)
14.
We got 207 tok/s with Qwen3.5-27B on an RTX 3090
(news.ycombinator.com)
15.
MDST Engine: run GGUF models in the browser with WebGPU/WASM
(news.ycombinator.com)
16.
Rust implementation of Mistral's Voxtral Mini 4B Realtime runs in your browser
(news.ycombinator.com)
17.
Show HN: Sweep, Open-weights 1.5B model for next-edit autocomplete
(news.ycombinator.com)
Today's top topics:
apple
googlebook
google
siri ai
mac studio
mac mini
gemini
ios 27
m6 chip
openai