Improving Adversarial Robustness of Zero-Shot CLIP with Confidence-Aware Weighting
arXiv:2510.02913v2 Announce Type: replace Abstract: Vision-language models such as CLIP demonstrate impressive zero-shot generalization but remain highly vulnerable to adversarial attacks. Prior adver