Topic

Reinforcement learning

1 story on this topic, newest first.

  1. ResearchExplainer

    What "reasoning" models actually do differently

    Reasoning models are trained to spend tokens thinking before they answer. Here is what that training involves, why it works on some problems and not others, and how to decide when to pay for it.

    4 min read