Topic
Reinforcement learning
1 story on this topic, newest first.
ResearchExplainer
What "reasoning" models actually do differently
Reasoning models are trained to spend tokens thinking before they answer. Here is what that training involves, why it works on some problems and not others, and how to decide when to pay for it.
