Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Wanqi, Bai, Shuanghao, Mandic, Danilo P., Zhao, Qibin, Chen, Badong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Soft Prompt Generation for Domain Generalization
von: Bai, Shuanghao, et al.
Veröffentlicht: (2024)
von: Bai, Shuanghao, et al.
Veröffentlicht: (2024)
Dual-Path Stable Soft Prompt Generation for Domain Generalization
von: Zhang, Yuedi, et al.
Veröffentlicht: (2025)
von: Zhang, Yuedi, et al.
Veröffentlicht: (2025)
PromptTA: Prompt-driven Text Adapter for Source-free Domain Generalization
von: Zhang, Haoran, et al.
Veröffentlicht: (2024)
von: Zhang, Haoran, et al.
Veröffentlicht: (2024)
Prompt-based Distribution Alignment for Unsupervised Domain Adaptation
von: Bai, Shuanghao, et al.
Veröffentlicht: (2023)
von: Bai, Shuanghao, et al.
Veröffentlicht: (2023)
Adversarial Robustness for Visual Grounding of Multimodal Large Language Models
von: Gao, Kuofeng, et al.
Veröffentlicht: (2024)
von: Gao, Kuofeng, et al.
Veröffentlicht: (2024)
Revisiting Multimodal Positional Encoding in Vision-Language Models
von: Huang, Jie, et al.
Veröffentlicht: (2025)
von: Huang, Jie, et al.
Veröffentlicht: (2025)
Revisiting Shadow Detection from a Vision-Language Perspective
von: Wang, Yonghui, et al.
Veröffentlicht: (2026)
von: Wang, Yonghui, et al.
Veröffentlicht: (2026)
MMT-ARD: Multimodal Multi-Teacher Adversarial Distillation for Robust Vision-Language Models
von: Li, Yuqi, et al.
Veröffentlicht: (2025)
von: Li, Yuqi, et al.
Veröffentlicht: (2025)
A Unified Perspective on Adversarial Membership Manipulation in Vision Models
von: Gao, Ruize, et al.
Veröffentlicht: (2026)
von: Gao, Ruize, et al.
Veröffentlicht: (2026)
Survey of Adversarial Robustness in Multimodal Large Language Models
von: Jiang, Chengze, et al.
Veröffentlicht: (2025)
von: Jiang, Chengze, et al.
Veröffentlicht: (2025)
Probing the Robustness of Vision-Language Pretrained Models: A Multimodal Adversarial Attack Approach
von: Guan, Jiwei, et al.
Veröffentlicht: (2024)
von: Guan, Jiwei, et al.
Veröffentlicht: (2024)
Closed-Loop Bidirectional Prompting for Adversarial Robustness of Vision Language Models
von: Liu, Xiao, et al.
Veröffentlicht: (2026)
von: Liu, Xiao, et al.
Veröffentlicht: (2026)
On the Adversarial Robustness of 3D Large Vision-Language Models
von: Liu, Chao, et al.
Veröffentlicht: (2026)
von: Liu, Chao, et al.
Veröffentlicht: (2026)
Revisiting Prompt Pretraining of Vision-Language Models
von: Chen, Zhenyuan, et al.
Veröffentlicht: (2024)
von: Chen, Zhenyuan, et al.
Veröffentlicht: (2024)
TAPT: Test-Time Adversarial Prompt Tuning for Robust Inference in Vision-Language Models
von: Wang, Xin, et al.
Veröffentlicht: (2024)
von: Wang, Xin, et al.
Veröffentlicht: (2024)
Enhancing Adversarial Robustness of Vision-Language Models through Low-Rank Adaptation
von: Ji, Yuheng, et al.
Veröffentlicht: (2024)
von: Ji, Yuheng, et al.
Veröffentlicht: (2024)
Chain of Attack: On the Robustness of Vision-Language Models Against Transfer-Based Adversarial Attacks
von: Xie, Peng, et al.
Veröffentlicht: (2024)
von: Xie, Peng, et al.
Veröffentlicht: (2024)
Adversarial Guided Diffusion Models for Adversarial Purification
von: Lin, Guang, et al.
Veröffentlicht: (2024)
von: Lin, Guang, et al.
Veröffentlicht: (2024)
VLM4VLA: Revisiting Vision-Language-Models in Vision-Language-Action Models
von: Zhang, Jianke, et al.
Veröffentlicht: (2026)
von: Zhang, Jianke, et al.
Veröffentlicht: (2026)
Adversarial Robustness Analysis of Vision-Language Models in Medical Image Segmentation
von: Budathoki, Anjila, et al.
Veröffentlicht: (2025)
von: Budathoki, Anjila, et al.
Veröffentlicht: (2025)
AGC: Adaptive Geodesic Correction for Adversarial Robustness on Vision-Language Models
von: Li, Zhiwei, et al.
Veröffentlicht: (2026)
von: Li, Zhiwei, et al.
Veröffentlicht: (2026)
MAA: Meticulous Adversarial Attack against Vision-Language Pre-trained Models
von: Zhang, Peng-Fei, et al.
Veröffentlicht: (2025)
von: Zhang, Peng-Fei, et al.
Veröffentlicht: (2025)
Adversarial Training on Purification (AToP): Advancing Both Robustness and Generalization
von: Lin, Guang, et al.
Veröffentlicht: (2024)
von: Lin, Guang, et al.
Veröffentlicht: (2024)
StableVLA: Towards Robust Vision-Language-Action Models without Extra Data
von: Fu, Yiyang, et al.
Veröffentlicht: (2026)
von: Fu, Yiyang, et al.
Veröffentlicht: (2026)
Revisiting CroPA: A Reproducibility Study and Enhancements for Cross-Prompt Adversarial Transferability in Vision-Language Models
von: Mittal, Atharv, et al.
Veröffentlicht: (2025)
von: Mittal, Atharv, et al.
Veröffentlicht: (2025)
OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation
von: Cui, Can, et al.
Veröffentlicht: (2025)
von: Cui, Can, et al.
Veröffentlicht: (2025)
Jacobian Regularizer-based Neural Granger Causality
von: Zhou, Wanqi, et al.
Veröffentlicht: (2024)
von: Zhou, Wanqi, et al.
Veröffentlicht: (2024)
Self-Calibrated Consistency can Fight Back for Adversarial Robustness in Vision-Language Models
von: Liu, Jiaxiang, et al.
Veröffentlicht: (2025)
von: Liu, Jiaxiang, et al.
Veröffentlicht: (2025)
Evolution-based Region Adversarial Prompt Learning for Robustness Enhancement in Vision-Language Models
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
Robust SAM: On the Adversarial Robustness of Vision Foundation Models
von: Long, Jiahuan, et al.
Veröffentlicht: (2025)
von: Long, Jiahuan, et al.
Veröffentlicht: (2025)
Hyper Adversarial Tuning for Boosting Adversarial Robustness of Pretrained Large Vision Models
von: Lv, Kangtao, et al.
Veröffentlicht: (2024)
von: Lv, Kangtao, et al.
Veröffentlicht: (2024)
Revisiting Continual Semantic Segmentation with Pre-trained Vision Models
von: Zhang, Duzhen, et al.
Veröffentlicht: (2025)
von: Zhang, Duzhen, et al.
Veröffentlicht: (2025)
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking
von: Liu, Xinqi, et al.
Veröffentlicht: (2024)
von: Liu, Xinqi, et al.
Veröffentlicht: (2024)
VOPE: Revisiting Hallucination of Vision-Language Models in Voluntary Imagination Task
von: Long, Xingming, et al.
Veröffentlicht: (2025)
von: Long, Xingming, et al.
Veröffentlicht: (2025)
PDA: Text-Augmented Defense Framework for Robust Vision-Language Models against Adversarial Image Attacks
von: Xu, Jingning, et al.
Veröffentlicht: (2026)
von: Xu, Jingning, et al.
Veröffentlicht: (2026)
On the Adversarial Robustness of Large Vision-Language Models under Visual Token Compression
von: Zhang, Xinwei, et al.
Veröffentlicht: (2026)
von: Zhang, Xinwei, et al.
Veröffentlicht: (2026)
Automatic Robotic Development through Collaborative Framework by Large Language Models
von: Luan, Zhirong, et al.
Veröffentlicht: (2024)
von: Luan, Zhirong, et al.
Veröffentlicht: (2024)
VLATTACK: Multimodal Adversarial Attacks on Vision-Language Tasks via Pre-trained Models
von: Yin, Ziyi, et al.
Veröffentlicht: (2023)
von: Yin, Ziyi, et al.
Veröffentlicht: (2023)
CF-GO-Net: A Universal Distribution Learner via Characteristic Function Networks with Graph Optimizers
von: Yu, Zeyang, et al.
Veröffentlicht: (2024)
von: Yu, Zeyang, et al.
Veröffentlicht: (2024)
Double Visual Defense: Adversarial Pre-training and Instruction Tuning for Improving Vision-Language Model Robustness
von: Wang, Zeyu, et al.
Veröffentlicht: (2025)
von: Wang, Zeyu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Soft Prompt Generation for Domain Generalization
von: Bai, Shuanghao, et al.
Veröffentlicht: (2024) -
Dual-Path Stable Soft Prompt Generation for Domain Generalization
von: Zhang, Yuedi, et al.
Veröffentlicht: (2025) -
PromptTA: Prompt-driven Text Adapter for Source-free Domain Generalization
von: Zhang, Haoran, et al.
Veröffentlicht: (2024) -
Prompt-based Distribution Alignment for Unsupervised Domain Adaptation
von: Bai, Shuanghao, et al.
Veröffentlicht: (2023) -
Adversarial Robustness for Visual Grounding of Multimodal Large Language Models
von: Gao, Kuofeng, et al.
Veröffentlicht: (2024)