Efficient Reinforcement Finetuning via Adaptive Curriculum Learning
arXiv:2504.05520v4 Announce Type: replace Abstract: Reinforcement finetuning (RFT) has shown great potential for enhancing the mathematical reasoning capabilities of large language models (LLMs), but