Unifying Adversarially Robust Model Experts in Vision-Language Models
DGX agentarXiv:2607.27897v1 Announce Type: new Abstract: Vision-language models (VLMs), such as CLIP, are vulnerable to adversarial attacks, posing a serious problem for real-life applications and deployment.