Open-weight reasoning model trained largely with reinforcement learning.
Official site ↗
An open-weight reasoning model trained largely with RL, competitive with closed models at a fraction of the reported cost.