The relationship between reasoning and performance in large language models--o3 (mini) thinks harder, not longer
arXiv:2502.15631v2 Announce Type: replace-cross Abstract: Large language models have demonstrated remarkable progress in mathematical reasoning, leveraging chain-of-thought and reinforcement learning.