61.
62.
63.
CS234: Reinforcement Learning Winter 2025
(news.ycombinator.com)
64.
65.
Olympiad-level formal mathematical reasoning with reinforcement learning
(feeds.nature.com)
66.
GPT-OSS Reinforcement Learning
(news.ycombinator.com)
67.
Launch HN: RunRL (YC X25) – Reinforcement learning as a service
(news.ycombinator.com)
68.
GEPA optimizes LLMs without costly reinforcement learning
(venturebeat.com)
69.
GEPA: Reflective prompt evolution can outperform reinforcement learning
(news.ycombinator.com)
70.
GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning
(news.ycombinator.com)
71.
Supervised fine tuning on curated data is reinforcement learning
(news.ycombinator.com)
72.
Supervised Fine Tuning on Curated Data is Reinforcement Learning
(news.ycombinator.com)
73.
Reinforcement Learning from Human Feedback (RLHF) in Notebooks
(news.ycombinator.com)
74.
Reinforcement learning, explained with a minimum of math and jargon
(news.ycombinator.com)
Today's top topics:
anthropic
openai
apple
ios 27
siri ai
google
valve
steam frame
microsoft
fast company