Is Chain-of-Thought Reasoning of LLMs a Mirage? A Data Distribution Lens
(news.ycombinator.com)
241.
242.
LLMs' "simulated reasoning" abilities are a brittle mirage
(news.ycombinator.com)
243.
244.
Tesla shuts down in-house Dojo AI supercomputer project
(engadget.com)
245.
Achieving 10,000x training data reduction with high-fidelity labels
(news.ycombinator.com)
246.
247.
248.
Persona vectors: Monitoring and controlling character traits in language models
(news.ycombinator.com)
249.
250.
Show HN: Terminal-Bench-RL: Training long-horizon terminal agents with RL
(news.ycombinator.com)
251.
Show HN: Terminal-Bench-RL: Training Long-Horizon Terminal Agents with RL
(news.ycombinator.com)
252.
GLM-4.5: Reasoning, Coding, and Agentic Abililties
(news.ycombinator.com)
253.
254.
Is HR ready for AI?
(zdnet.com)
255.
How to scale RL to 10^26 FLOPs
(news.ycombinator.com)
256.
The upcoming GPT-3 moment for RL
(news.ycombinator.com)
257.
ETH Zurich and EPFL to release a LLM developed on public infrastructure
(news.ycombinator.com)
258.
259.
LLM-Ready Training Dataset for Apple's Foundation Models (iOS 26)
(news.ycombinator.com)
260.
Smollm3: Smol, multilingual, long-context reasoner LLM
(news.ycombinator.com)
261.
262.
263.
264.
Can the music industry make AI the next Napster?
(theverge.com)
265.
Did AI companies win a fight with authors? Technically
(theverge.com)
266.
Reinforcement learning, explained with a minimum of math and jargon
(news.ycombinator.com)
267.
Fault Tolerant Llama training – PyTorch blog
(news.ycombinator.com)
268.
269.