Taming the Thinker: Conditional Entropy Shaping for Adaptive LLM Reasoning
arXiv:2605.19358v1 Announce Type: new Abstract: Entropy-based deep reasoning has emerged as a promising direction for improving the reasoning capabilities of Large Language Models (LLMs), but existing