Towards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language Models
arXiv:2607.13860v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) have demonstrated remarkable success in 2D medical image understanding, their extension to 3D volumetric