When and Where to Attack? Stage-wise Attention-Guided Adversarial Attack on Large Vision Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Kwak, Jaehyun, Cao, Nam, Cho, Boryeong, Lee, Segyu, Ahn, Sumyeong, Yun, Se-Young |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VACoDe: Visual Augmented Contrastive Decoding
by: Kim, Sihyeon, et al.
Published: (2024)
by: Kim, Sihyeon, et al.
Published: (2024)
UniSAFE: A Comprehensive Benchmark for Safety Evaluation of Unified Multimodal Models
by: Lee, Segyu, et al.
Published: (2026)
by: Lee, Segyu, et al.
Published: (2026)
Active Prompt Learning in Vision Language Models
by: Bang, Jihwan, et al.
Published: (2023)
by: Bang, Jihwan, et al.
Published: (2023)
FedDr+: Stabilizing Dot-regression with Global Feature Distillation for Federated Learning
by: Kim, Seongyoon, et al.
Published: (2024)
by: Kim, Seongyoon, et al.
Published: (2024)
Generative Visual Code Mobile World Models
by: Koh, Woosung, et al.
Published: (2026)
by: Koh, Woosung, et al.
Published: (2026)
Attention-Guided Patch-Wise Sparse Adversarial Attacks on Vision-Language-Action Models
by: Zhang, Naifu, et al.
Published: (2025)
by: Zhang, Naifu, et al.
Published: (2025)
QuRe: Query-Relevant Retrieval through Hard Negative Sampling in Composed Image Retrieval
by: Kwak, Jaehyun, et al.
Published: (2025)
by: Kwak, Jaehyun, et al.
Published: (2025)
Enhancing Targeted Adversarial Attacks on Large Vision-Language Models via Intermediate Projector
by: Cao, Yiming, et al.
Published: (2025)
by: Cao, Yiming, et al.
Published: (2025)
When Alignment Fails: Multimodal Adversarial Attacks on Vision-Language-Action Models
by: Yan, Yuping, et al.
Published: (2025)
by: Yan, Yuping, et al.
Published: (2025)
FOAM: Frequency and Operator Error-Based Adaptive Damping Method for Reducing Staleness-Oriented Error for Shampoo
by: Nam, Kyunghun, et al.
Published: (2026)
by: Nam, Kyunghun, et al.
Published: (2026)
SGHA-Attack: Semantic-Guided Hierarchical Alignment for Transferable Targeted Attacks on Vision-Language Models
by: Wang, Haobo, et al.
Published: (2026)
by: Wang, Haobo, et al.
Published: (2026)
Hijacking Vision-and-Language Navigation Agents with Adversarial Environmental Attacks
by: Yang, Zijiao, et al.
Published: (2024)
by: Yang, Zijiao, et al.
Published: (2024)
PA-Attack: Guiding Gray-Box Attacks on LVLM Vision Encoders with Prototypes and Attention
by: Mei, Hefei, et al.
Published: (2026)
by: Mei, Hefei, et al.
Published: (2026)
A Two-Stage Globally-Diverse Adversarial Attack for Vision-Language Pre-training Models
by: Chen, Wutao, et al.
Published: (2026)
by: Chen, Wutao, et al.
Published: (2026)
From Attack to Protection: Leveraging Watermarking Attack Network for Advanced Add-on Watermarking
by: Nam, Seung-Hun, et al.
Published: (2020)
by: Nam, Seung-Hun, et al.
Published: (2020)
Model-agnostic Adversarial Attack and Defense for Vision-Language-Action Models
by: Xu, Haochuan, et al.
Published: (2025)
by: Xu, Haochuan, et al.
Published: (2025)
When Surfaces Lie: Exploiting Wrinkle-Induced Attention Shift to Attack Vision-Language Models
by: Hu, Chengyin, et al.
Published: (2026)
by: Hu, Chengyin, et al.
Published: (2026)
Chain of Attack: On the Robustness of Vision-Language Models Against Transfer-Based Adversarial Attacks
by: Xie, Peng, et al.
Published: (2024)
by: Xie, Peng, et al.
Published: (2024)
GSE: Group-wise Sparse and Explainable Adversarial Attacks
by: Sadiku, Shpresim, et al.
Published: (2023)
by: Sadiku, Shpresim, et al.
Published: (2023)
AttackVLA: Benchmarking Adversarial and Backdoor Attacks on Vision-Language-Action Models
by: Li, Jiayu, et al.
Published: (2025)
by: Li, Jiayu, et al.
Published: (2025)
On Epistemic Uncertainty of Visual Tokens for Object Hallucinations in Large Vision-Language Models
by: Seo, Hoigi, et al.
Published: (2025)
by: Seo, Hoigi, et al.
Published: (2025)
Semantic-Aligned Adversarial Evolution Triangle for High-Transferability Vision-Language Attack
by: Jia, Xiaojun, et al.
Published: (2024)
by: Jia, Xiaojun, et al.
Published: (2024)
Break the Visual Perception: Adversarial Attacks Targeting Encoded Visual Tokens of Large Vision-Language Models
by: Wang, Yubo, et al.
Published: (2024)
by: Wang, Yubo, et al.
Published: (2024)
Visual Adversarial Attack on Vision-Language Models for Autonomous Driving
by: Zhang, Tianyuan, et al.
Published: (2024)
by: Zhang, Tianyuan, et al.
Published: (2024)
PG-Attack: A Precision-Guided Adversarial Attack Framework Against Vision Foundation Models for Autonomous Driving
by: Fu, Jiyuan, et al.
Published: (2024)
by: Fu, Jiyuan, et al.
Published: (2024)
Attention-Based Real-Time Defenses for Physical Adversarial Attacks in Vision Applications
by: Rossolini, Giulio, et al.
Published: (2023)
by: Rossolini, Giulio, et al.
Published: (2023)
Probing the Robustness of Vision-Language Pretrained Models: A Multimodal Adversarial Attack Approach
by: Guan, Jiwei, et al.
Published: (2024)
by: Guan, Jiwei, et al.
Published: (2024)
Transferable Adversarial Attacks on Black-Box Vision-Language Models
by: Hu, Kai, et al.
Published: (2025)
by: Hu, Kai, et al.
Published: (2025)
Erasing Thousands of Concepts: Towards Scalable and Practical Concept Erasure for Text-to-Image Diffusion Models
by: Seo, Hoigi, et al.
Published: (2026)
by: Seo, Hoigi, et al.
Published: (2026)
Adversarial-Guided Diffusion for Multimodal LLM Attacks
by: Xia, Chengwei, et al.
Published: (2025)
by: Xia, Chengwei, et al.
Published: (2025)
Towards Difficulty-Agnostic Efficient Transfer Learning for Vision-Language Models
by: Yang, Yongjin, et al.
Published: (2023)
by: Yang, Yongjin, et al.
Published: (2023)
MS-GAGA: Metric-Selective Guided Adversarial Generation Attack
by: Ho, Dion J. X., et al.
Published: (2025)
by: Ho, Dion J. X., et al.
Published: (2025)
Black-Box Adversarial Attack on Vision Language Models for Autonomous Driving
by: Wang, Lu, et al.
Published: (2025)
by: Wang, Lu, et al.
Published: (2025)
Text Adversarial Attacks with Dynamic Outputs
by: Wang, Wenqiang, et al.
Published: (2025)
by: Wang, Wenqiang, et al.
Published: (2025)
MAA: Meticulous Adversarial Attack against Vision-Language Pre-trained Models
by: Zhang, Peng-Fei, et al.
Published: (2025)
by: Zhang, Peng-Fei, et al.
Published: (2025)
When Background Matters: Breaking Medical Vision Language Models by Transferable Attack
by: Ghosh, Akash, et al.
Published: (2026)
by: Ghosh, Akash, et al.
Published: (2026)
Adversarial Prompt Injection Attack on Multimodal Large Language Models
by: Ding, Meiwen, et al.
Published: (2026)
by: Ding, Meiwen, et al.
Published: (2026)
A Generative Adversarial Approach to Adversarial Attacks Guided by Contrastive Language-Image Pre-trained Model
by: Soor, Sampriti, et al.
Published: (2025)
by: Soor, Sampriti, et al.
Published: (2025)
Attention-aggregated Attack for Boosting the Transferability of Facial Adversarial Examples
by: Li, Jian-Wei, et al.
Published: (2025)
by: Li, Jian-Wei, et al.
Published: (2025)
Downstream Transfer Attack: Adversarial Attacks on Downstream Models with Pre-trained Vision Transformers
by: Zheng, Weijie, et al.
Published: (2024)
by: Zheng, Weijie, et al.
Published: (2024)
Similar Items
-
VACoDe: Visual Augmented Contrastive Decoding
by: Kim, Sihyeon, et al.
Published: (2024) -
UniSAFE: A Comprehensive Benchmark for Safety Evaluation of Unified Multimodal Models
by: Lee, Segyu, et al.
Published: (2026) -
Active Prompt Learning in Vision Language Models
by: Bang, Jihwan, et al.
Published: (2023) -
FedDr+: Stabilizing Dot-regression with Global Feature Distillation for Federated Learning
by: Kim, Seongyoon, et al.
Published: (2024) -
Generative Visual Code Mobile World Models
by: Koh, Woosung, et al.
Published: (2026)