Timeline milestone
Reasoning models spend more computation before answering
-
Research
Reasoning models spend more computation before answering
OpenAI introduced o1-preview, using additional inference-time computation to improve performance on multi-step problems.
The release shifted part of the scaling race from larger pre-training runs toward “test-time compute,” where a model performs more internal processing before responding. Better benchmark performance did not guarantee truthfulness or remove the need to verify outputs, but the approach influenced the next generation of frontier models.Sources & references 2 sources