GLINT: Sparsely Gated Vision-Language Alignment for Fine-Grained Radiology Representations
DGX agentarXiv:2606.03180v1 Announce Type: cross Abstract: Vision-language models (VLMs) for radiology have emerged as a scalable paradigm by leveraging image-report pairs naturally produced in clinical workfl