When Robots Obey the Patch: Universal Transferable Patch Attacks on Vision-Language-Action Models
Fuente:
arXiv
Saved in:
| Main Authors: | Lu, Hui, Yu, Yi, Yang, Yiming, Yi, Chenyu, Zhang, Qixin, Shen, Bingquan, Kot, Alex C., Jiang, Xudong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Universal Adversarial Attacks against Closed-Source MLLMs via Target-View Routed Meta Optimization
by: Lu, Hui, et al.
Published: (2026)
by: Lu, Hui, et al.
Published: (2026)
Semantic Deep Hiding for Robust Unlearnable Examples
by: Meng, Ruohan, et al.
Published: (2024)
by: Meng, Ruohan, et al.
Published: (2024)
MambaTAD: When State-Space Models Meet Long-Range Temporal Action Detection
by: Lu, Hui, et al.
Published: (2025)
by: Lu, Hui, et al.
Published: (2025)
VLA-Hijack: A Transferable Patch Attack against Vision-Language-Action Models via Visual Proprioception Hijacking
by: Fu, Jiyuan, et al.
Published: (2026)
by: Fu, Jiyuan, et al.
Published: (2026)
Brain-Inspired Stepwise Patch Merging for Vision Transformers
by: Yu, Yonghao, et al.
Published: (2024)
by: Yu, Yonghao, et al.
Published: (2024)
Attention-Guided Patch-Wise Sparse Adversarial Attacks on Vision-Language-Action Models
by: Zhang, Naifu, et al.
Published: (2025)
by: Zhang, Naifu, et al.
Published: (2025)
From Pretrain to Pain: Adversarial Vulnerability of Video Foundation Models Without Task Knowledge
by: Lu, Hui, et al.
Published: (2025)
by: Lu, Hui, et al.
Published: (2025)
Towards Physical World Backdoor Attacks against Skeleton Action Recognition
by: Zheng, Qichen, et al.
Published: (2024)
by: Zheng, Qichen, et al.
Published: (2024)
PhysPatch: A Physically Realizable and Transferable Adversarial Patch Attack for Multimodal Large Language Models-based Autonomous Driving Systems
by: Guo, Qi, et al.
Published: (2025)
by: Guo, Qi, et al.
Published: (2025)
PatchCue: Enhancing Vision-Language Model Reasoning with Patch-Based Visual Cues
by: Qi, Yukun, et al.
Published: (2026)
by: Qi, Yukun, et al.
Published: (2026)
Thermal Topology Collapse: Universal Physical Patch Attacks on Infrared Vision Systems
by: Hu, Chengyin, et al.
Published: (2026)
by: Hu, Chengyin, et al.
Published: (2026)
PAD: Patch-Agnostic Defense against Adversarial Patch Attacks
by: Jing, Lihua, et al.
Published: (2024)
by: Jing, Lihua, et al.
Published: (2024)
SAVER: Mitigating Hallucinations in Large Vision-Language Models via Style-Aware Visual Early Revision
by: Li, Zhaoxu, et al.
Published: (2025)
by: Li, Zhaoxu, et al.
Published: (2025)
BadPatch: Diffusion-Based Generation of Physical Adversarial Patches
by: Wang, Zhixiang, et al.
Published: (2024)
by: Wang, Zhixiang, et al.
Published: (2024)
Robust and Transferable Backdoor Attacks Against Deep Image Compression With Selective Frequency Prior
by: Yu, Yi, et al.
Published: (2024)
by: Yu, Yi, et al.
Published: (2024)
Tracing Copied Pixels and Regularizing Patch Affinity in Copy Detection
by: Lu, Yichen, et al.
Published: (2026)
by: Lu, Yichen, et al.
Published: (2026)
Patch is Enough: Naturalistic Adversarial Patch against Vision-Language Pre-training Models
by: Kong, Dehong, et al.
Published: (2024)
by: Kong, Dehong, et al.
Published: (2024)
Revealing Physical-World Semantic Vulnerabilities: Universal Adversarial Patches for Infrared Vision-Language Models
by: Hu, Chengyin, et al.
Published: (2026)
by: Hu, Chengyin, et al.
Published: (2026)
Defending against Patch-Based and Texture-Based Adversarial Attacks with Spectral Decomposition
by: Zhang, Wei, et al.
Published: (2026)
by: Zhang, Wei, et al.
Published: (2026)
OmniPatch: A Universal Adversarial Patch for ViT-CNN Cross-Architecture Transfer in Semantic Segmentation
by: Aggarwal, Aarush, et al.
Published: (2026)
by: Aggarwal, Aarush, et al.
Published: (2026)
Transferable Adversarial Attacks on SAM and Its Downstream Models
by: Xia, Song, et al.
Published: (2024)
by: Xia, Song, et al.
Published: (2024)
CDUPatch: Color-Driven Universal Adversarial Patch Attack for Dual-Modal Visible-Infrared Detectors
by: Long, Jiahuan, et al.
Published: (2025)
by: Long, Jiahuan, et al.
Published: (2025)
Adversarial Patch for 3D Local Feature Extractor
by: Pao, Yu Wen, et al.
Published: (2024)
by: Pao, Yu Wen, et al.
Published: (2024)
Patch Rebirth: Toward Fast and Transferable Model Inversion of Vision Transformers
by: Heo, Seongsoo, et al.
Published: (2025)
by: Heo, Seongsoo, et al.
Published: (2025)
Fight Fire with Fire: Combating Adversarial Patch Attacks using Pattern-randomized Defensive Patches
by: Feng, Jianan, et al.
Published: (2023)
by: Feng, Jianan, et al.
Published: (2023)
Concept-Based Masking: A Patch-Agnostic Defense Against Adversarial Patch Attacks
by: Mehrotra, Ayushi, et al.
Published: (2025)
by: Mehrotra, Ayushi, et al.
Published: (2025)
Open-set Anomaly Segmentation in Complex Scenarios
by: Xia, Song, et al.
Published: (2025)
by: Xia, Song, et al.
Published: (2025)
PASTA: A Patch-Agnostic Twofold-Stealthy Backdoor Attack on Vision Transformers
by: Liu, Dazhuang, et al.
Published: (2026)
by: Liu, Dazhuang, et al.
Published: (2026)
Time Is All It Takes: Spike-Retiming Attacks on Event-Driven Spiking Neural Networks
by: Yu, Yi, et al.
Published: (2026)
by: Yu, Yi, et al.
Published: (2026)
PFF-Net: Patch Feature Fitting for Point Cloud Normal Estimation
by: Li, Qing, et al.
Published: (2025)
by: Li, Qing, et al.
Published: (2025)
BB-Patch: BlackBox Adversarial Patch-Attack using Zeroth-Order Optimization
by: Kumar, Satyadwyoom, et al.
Published: (2024)
by: Kumar, Satyadwyoom, et al.
Published: (2024)
TIPSv2: Advancing Vision-Language Pretraining with Enhanced Patch-Text Alignment
by: Cao, Bingyi, et al.
Published: (2026)
by: Cao, Bingyi, et al.
Published: (2026)
Exploring Vision Transformers for 3D Human Motion-Language Models with Motion Patches
by: Yu, Qing, et al.
Published: (2024)
by: Yu, Qing, et al.
Published: (2024)
Seeing Realism from Simulation: Efficient Video Transfer for Vision-Language-Action Data Augmentation
by: Hui, Chenyu, et al.
Published: (2026)
by: Hui, Chenyu, et al.
Published: (2026)
Backdoor Attacks against No-Reference Image Quality Assessment Models via a Scalable Trigger
by: Yu, Yi, et al.
Published: (2024)
by: Yu, Yi, et al.
Published: (2024)
Information-Theoretic Constraints for Continual Vision-Language-Action Alignment
by: Zhao, Libang, et al.
Published: (2026)
by: Zhao, Libang, et al.
Published: (2026)
Efficient Vision-and-Language Pre-training with Text-Relevant Image Patch Selection
by: Ye, Wei, et al.
Published: (2024)
by: Ye, Wei, et al.
Published: (2024)
Rethinking Patch Dependence for Masked Autoencoders
by: Fu, Letian, et al.
Published: (2024)
by: Fu, Letian, et al.
Published: (2024)
VLA-Pro: Cross-Task Procedural Memory Transfer for Vision-Language-Action Models
by: Si, Shengyu, et al.
Published: (2026)
by: Si, Shengyu, et al.
Published: (2026)
When Background Matters: Breaking Medical Vision Language Models by Transferable Attack
by: Ghosh, Akash, et al.
Published: (2026)
by: Ghosh, Akash, et al.
Published: (2026)
Similar Items
-
Universal Adversarial Attacks against Closed-Source MLLMs via Target-View Routed Meta Optimization
by: Lu, Hui, et al.
Published: (2026) -
Semantic Deep Hiding for Robust Unlearnable Examples
by: Meng, Ruohan, et al.
Published: (2024) -
MambaTAD: When State-Space Models Meet Long-Range Temporal Action Detection
by: Lu, Hui, et al.
Published: (2025) -
VLA-Hijack: A Transferable Patch Attack against Vision-Language-Action Models via Visual Proprioception Hijacking
by: Fu, Jiyuan, et al.
Published: (2026) -
Brain-Inspired Stepwise Patch Merging for Vision Transformers
by: Yu, Yonghao, et al.
Published: (2024)