Adversarial Prompt Injection Attack on Multimodal Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Ding, Meiwen, Xia, Song, Kong, Chenqi, Jiang, Xudong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Feature-Space Smoothing: Certified Robustness of Deep Representations
di: Xia, Song, et al.
Pubblicazione: (2026)
di: Xia, Song, et al.
Pubblicazione: (2026)
Physical Prompt Injection Attacks on Large Vision-Language Models
di: Ling, Chen, et al.
Pubblicazione: (2026)
di: Ling, Chen, et al.
Pubblicazione: (2026)
Probing the Robustness of Vision-Language Pretrained Models: A Multimodal Adversarial Attack Approach
di: Guan, Jiwei, et al.
Pubblicazione: (2024)
di: Guan, Jiwei, et al.
Pubblicazione: (2024)
Adversarial Prompt Distillation for Vision-Language Models
di: Luo, Lin, et al.
Pubblicazione: (2024)
di: Luo, Lin, et al.
Pubblicazione: (2024)
Adversarial Prompt Tuning for Vision-Language Models
di: Zhang, Jiaming, et al.
Pubblicazione: (2023)
di: Zhang, Jiaming, et al.
Pubblicazione: (2023)
When Alignment Fails: Multimodal Adversarial Attacks on Vision-Language-Action Models
di: Yan, Yuping, et al.
Pubblicazione: (2025)
di: Yan, Yuping, et al.
Pubblicazione: (2025)
Multimodal Prompt Decoupling Attack on the Safety Filters in Text-to-Image Models
di: Peng, Xingkai, et al.
Pubblicazione: (2025)
di: Peng, Xingkai, et al.
Pubblicazione: (2025)
See&Trek: Training-Free Spatial Prompting for Multimodal Large Language Model
di: Li, Pengteng, et al.
Pubblicazione: (2025)
di: Li, Pengteng, et al.
Pubblicazione: (2025)
DarkLLM: Learning Language-Driven Adversarial Attacks with Large Language Models
di: Sun, Ye, et al.
Pubblicazione: (2026)
di: Sun, Ye, et al.
Pubblicazione: (2026)
Revisiting the Robust Generalization of Adversarial Prompt Tuning
di: Yang, Fan, et al.
Pubblicazione: (2024)
di: Yang, Fan, et al.
Pubblicazione: (2024)
Image-based Prompt Injection: Hijacking Multimodal LLMs through Visually Embedded Adversarial Instructions
di: Nagaraja, Neha, et al.
Pubblicazione: (2026)
di: Nagaraja, Neha, et al.
Pubblicazione: (2026)
Prompting Large Vision-Language Models for Compositional Reasoning
di: Ossowski, Timothy, et al.
Pubblicazione: (2024)
di: Ossowski, Timothy, et al.
Pubblicazione: (2024)
Prompt-Agnostic Adversarial Perturbation for Customized Diffusion Models
di: Wan, Cong, et al.
Pubblicazione: (2024)
di: Wan, Cong, et al.
Pubblicazione: (2024)
eXIAA: eXplainable Injections for Adversarial Attack
di: Pesce, Leonardo, et al.
Pubblicazione: (2025)
di: Pesce, Leonardo, et al.
Pubblicazione: (2025)
NAP-Tuning: Neural Augmented Prompt Tuning for Adversarially Robust Vision-Language Models
di: Zhang, Jiaming, et al.
Pubblicazione: (2025)
di: Zhang, Jiaming, et al.
Pubblicazione: (2025)
Temporal Grounding of Activities using Multimodal Large Language Models
di: Song, Young Chol
Pubblicazione: (2024)
di: Song, Young Chol
Pubblicazione: (2024)
SAKED: Mitigating Hallucination in Large Vision-Language Models via Stability-Aware Knowledge Enhanced Decoding
di: Li, Zhaoxu, et al.
Pubblicazione: (2026)
di: Li, Zhaoxu, et al.
Pubblicazione: (2026)
Downstream Transfer Attack: Adversarial Attacks on Downstream Models with Pre-trained Vision Transformers
di: Zheng, Weijie, et al.
Pubblicazione: (2024)
di: Zheng, Weijie, et al.
Pubblicazione: (2024)
Investigating Redundancy in Multimodal Large Language Models with Multiple Vision Encoders
di: Wang, Yizhou, et al.
Pubblicazione: (2025)
di: Wang, Yizhou, et al.
Pubblicazione: (2025)
XiHeFusion: Harnessing Large Language Models for Science Communication in Nuclear Fusion
di: Wang, Xiao, et al.
Pubblicazione: (2025)
di: Wang, Xiao, et al.
Pubblicazione: (2025)
AUVIC: Adversarial Unlearning of Visual Concepts for Multi-modal Large Language Models
di: Chen, Haokun, et al.
Pubblicazione: (2025)
di: Chen, Haokun, et al.
Pubblicazione: (2025)
Prompt-Aware Adapter: Towards Learning Adaptive Visual Tokens for Multimodal Large Language Models
di: Zhang, Yue, et al.
Pubblicazione: (2024)
di: Zhang, Yue, et al.
Pubblicazione: (2024)
AttackVLA: Benchmarking Adversarial and Backdoor Attacks on Vision-Language-Action Models
di: Li, Jiayu, et al.
Pubblicazione: (2025)
di: Li, Jiayu, et al.
Pubblicazione: (2025)
Black-Box Adversarial Attack on Vision Language Models for Autonomous Driving
di: Wang, Lu, et al.
Pubblicazione: (2025)
di: Wang, Lu, et al.
Pubblicazione: (2025)
When Robots Obey the Patch: Universal Transferable Patch Attacks on Vision-Language-Action Models
di: Lu, Hui, et al.
Pubblicazione: (2025)
di: Lu, Hui, et al.
Pubblicazione: (2025)
RefSR-Adv: Adversarial Attack on Reference-based Image Super-Resolution Models
di: Dai, Jiazhu, et al.
Pubblicazione: (2026)
di: Dai, Jiazhu, et al.
Pubblicazione: (2026)
Concept-based Adversarial Attack: a Probabilistic Perspective
di: Zhang, Andi, et al.
Pubblicazione: (2025)
di: Zhang, Andi, et al.
Pubblicazione: (2025)
Break the Visual Perception: Adversarial Attacks Targeting Encoded Visual Tokens of Large Vision-Language Models
di: Wang, Yubo, et al.
Pubblicazione: (2024)
di: Wang, Yubo, et al.
Pubblicazione: (2024)
WebInject: Prompt Injection Attack to Web Agents
di: Wang, Xilong, et al.
Pubblicazione: (2025)
di: Wang, Xilong, et al.
Pubblicazione: (2025)
Image-of-Thought Prompting for Visual Reasoning Refinement in Multimodal Large Language Models
di: Zhou, Qiji, et al.
Pubblicazione: (2024)
di: Zhou, Qiji, et al.
Pubblicazione: (2024)
Exploring the Transferability of Visual Prompting for Multimodal Large Language Models
di: Zhang, Yichi, et al.
Pubblicazione: (2024)
di: Zhang, Yichi, et al.
Pubblicazione: (2024)
QAPruner: Quantization-Aware Vision Token Pruning for Multimodal Large Language Models
di: Wang, Xinhao, et al.
Pubblicazione: (2026)
di: Wang, Xinhao, et al.
Pubblicazione: (2026)
MMR-AD: A Large-Scale Multimodal Dataset for Benchmarking General Anomaly Detection with Multimodal Large Language Models
di: Yao, Xincheng, et al.
Pubblicazione: (2026)
di: Yao, Xincheng, et al.
Pubblicazione: (2026)
Img-Diff: Contrastive Data Synthesis for Multimodal Large Language Models
di: Jiao, Qirui, et al.
Pubblicazione: (2024)
di: Jiao, Qirui, et al.
Pubblicazione: (2024)
A Survey on Benchmarks of Multimodal Large Language Models
di: Li, Jian, et al.
Pubblicazione: (2024)
di: Li, Jian, et al.
Pubblicazione: (2024)
PromptDx: Differentiable Prompt Tuning for Multimodal In-Context Alzheimer's Diagnosis
di: Zhong, Lujia, et al.
Pubblicazione: (2026)
di: Zhong, Lujia, et al.
Pubblicazione: (2026)
BLINK: Multimodal Large Language Models Can See but Not Perceive
di: Fu, Xingyu, et al.
Pubblicazione: (2024)
di: Fu, Xingyu, et al.
Pubblicazione: (2024)
2AFC Prompting of Large Multimodal Models for Image Quality Assessment
di: Zhu, Hanwei, et al.
Pubblicazione: (2024)
di: Zhu, Hanwei, et al.
Pubblicazione: (2024)
Transferable Adversarial Face Attack with Text Controlled Attribute
di: Li, Wenyun, et al.
Pubblicazione: (2024)
di: Li, Wenyun, et al.
Pubblicazione: (2024)
Efficient Multimodal Large Language Models: A Survey
di: Jin, Yizhang, et al.
Pubblicazione: (2024)
di: Jin, Yizhang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Feature-Space Smoothing: Certified Robustness of Deep Representations
di: Xia, Song, et al.
Pubblicazione: (2026) -
Physical Prompt Injection Attacks on Large Vision-Language Models
di: Ling, Chen, et al.
Pubblicazione: (2026) -
Probing the Robustness of Vision-Language Pretrained Models: A Multimodal Adversarial Attack Approach
di: Guan, Jiwei, et al.
Pubblicazione: (2024) -
Adversarial Prompt Distillation for Vision-Language Models
di: Luo, Lin, et al.
Pubblicazione: (2024) -
Adversarial Prompt Tuning for Vision-Language Models
di: Zhang, Jiaming, et al.
Pubblicazione: (2023)