A Cross-Modal Prompt Injection Attack against Large Vision-Language Models with Image-Only Perturbation
arXiv:2605.16090v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) have emerged as a powerful paradigm for multimodal intelligence, but their growing deployment also expands the at