1.
2.
3.
Smaller, faster, safer: running Kimi and GLM at scale
(news.ycombinator.com)
4.
Show HN: Fuse – statically typed functional programming language
(news.ycombinator.com)
5.
Morph (YC S23) Is Hiring Member of Technical Staff
(news.ycombinator.com)
6.
Morph (YC S23) Is Hiring Member of Technical Stuff
(news.ycombinator.com)
7.
Predictive Speculative KV Replication for Bursty LLM Inference
(news.ycombinator.com)
8.
Everyone is building LLM routers, we deprecated ours
(news.ycombinator.com)
9.
Why we write our own C and C++ inference engines
(news.ycombinator.com)
10.
Show HN: Noisegate – a differential-privacy gateway for untrusted AI agents
(news.ycombinator.com)
11.
Show HN: Local text, image, video, music and 3D from one CLI, no Python
(news.ycombinator.com)
12.
Using an open model feels surprisingly good
(news.ycombinator.com)
13.
Kimi K3 Now Available via Telnyx Inference API
(news.ycombinator.com)
14.
Hetzner is working on LLM Inference
(news.ycombinator.com)
16.
17.
18.
Homomorphically encrypted CIFAR-10 inference in 200ms
(news.ycombinator.com)
19.
20.
21.
DeepSeek cut prices 75%. The 100x problem remains
(venturebeat.com)
22.
Meta’s new AI chips will begin production in September
(techcrunch.com)
23.
SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence
(news.ycombinator.com)
24.
25.
26.
Inference Optimization for MiMo v2.5: Pushing Hybrid SWA Efficiency to the Limit
(news.ycombinator.com)
27.
GLM 5.2 and the coming AI margin collapse
(news.ycombinator.com)
28.
Show HN: Morph Reflexes – Multi-head classifiers for agent traces
(news.ycombinator.com)
29.
30.
Popping the GPU Bubble
(news.ycombinator.com)