VLMShield: Efficient and Robust Defense of Vision-Language Models against Malicious Prompts
DGX agentarXiv:2604.06502v1 Announce Type: new Abstract: Vision-Language Models (VLMs) face significant safety vulnerabilities from malicious prompt attacks due to weakened alignment during visual integration.