OpenAI's reasoning models trained to think before answering, starting with o1.
A model trained with RL to reason at length before answering. More thinking at inference buys accuracy.