Eliciting Medical Reasoning with Knowledge-enhanced Data Synthesis: A Semi-Supervised Reinforcement Learning Approach
arXiv:2604.11547v1 Announce Type: cross Abstract: While large language models hold promise for complex medical applications, their development is hindered by the scarcity of high-quality reasoning dat