Look Before You Zoom: Adaptive Routing for the Resolution-Context Trade-off in Visual RAG
arXiv:2606.21968v1 Announce Type: new Abstract: Vision-Language Models (VLMs) struggle as query-relevant objects become smaller. To address this, recent training-free approaches dynamically retrieve a