It began as a prompting trick — appending “think step by step” improved results — and worked because a model conditions on what it has already written, so intermediate steps become scaffolding for later ones.
Reasoning models internalise this rather than waiting to be asked. Note that reasoning is not verification: checking an answer against the model’s own reasoning is not an independent check.
