Backdoor-Powered Prompt Injection Attacks Nullify Defense Methods
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Yulin, Li, Haoran, Sui, Yuan, Song, Yangqiu, Hooi, Bryan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Defense Against Prompt Injection Attack by Leveraging Attack Techniques
von: Chen, Yulin, et al.
Veröffentlicht: (2024)
von: Chen, Yulin, et al.
Veröffentlicht: (2024)
Can Indirect Prompt Injection Attacks Be Detected and Removed?
von: Chen, Yulin, et al.
Veröffentlicht: (2025)
von: Chen, Yulin, et al.
Veröffentlicht: (2025)
TopicAttack: An Indirect Prompt Injection Attack via Topic Transition
von: Chen, Yulin, et al.
Veröffentlicht: (2025)
von: Chen, Yulin, et al.
Veröffentlicht: (2025)
BaThe: Defense against the Jailbreak Attack in Multimodal Large Language Models by Treating Harmful Instruction as Backdoor Trigger
von: Chen, Yulin, et al.
Veröffentlicht: (2024)
von: Chen, Yulin, et al.
Veröffentlicht: (2024)
Robustness via Referencing: Defending against Prompt Injection Attacks by Referencing the Executed Instruction
von: Chen, Yulin, et al.
Veröffentlicht: (2025)
von: Chen, Yulin, et al.
Veröffentlicht: (2025)
WebAgentGuard: A Reasoning-Driven Guard Model for Detecting Prompt Injection Attacks in Web Agents
von: Chen, Yulin, et al.
Veröffentlicht: (2026)
von: Chen, Yulin, et al.
Veröffentlicht: (2026)
Privacy in Large Language Models: Attacks, Defenses and Future Directions
von: Li, Haoran, et al.
Veröffentlicht: (2023)
von: Li, Haoran, et al.
Veröffentlicht: (2023)
WARD: Adversarially Robust Defense of Web Agents Against Prompt Injections
von: Cao, Tri, et al.
Veröffentlicht: (2026)
von: Cao, Tri, et al.
Veröffentlicht: (2026)
VPI-Bench: Visual Prompt Injection Attacks for Computer-Use Agents
von: Cao, Tri, et al.
Veröffentlicht: (2025)
von: Cao, Tri, et al.
Veröffentlicht: (2025)
Simulate and Eliminate: Revoke Backdoors for Generative Large Language Models
von: Li, Haoran, et al.
Veröffentlicht: (2024)
von: Li, Haoran, et al.
Veröffentlicht: (2024)
A Critical Evaluation of Defenses against Prompt Injection Attacks
von: Jia, Yuqi, et al.
Veröffentlicht: (2025)
von: Jia, Yuqi, et al.
Veröffentlicht: (2025)
Distributed Backdoor Attacks on Federated Graph Learning and Certified Defenses
von: Yang, Yuxin, et al.
Veröffentlicht: (2024)
von: Yang, Yuxin, et al.
Veröffentlicht: (2024)
Multimodal Prompt Injection Attacks: Risks and Defenses for Modern LLMs
von: Yeo, Andrew, et al.
Veröffentlicht: (2025)
von: Yeo, Andrew, et al.
Veröffentlicht: (2025)
AegisAgent: An Autonomous Defense Agent Against Prompt Injection Attacks in LLM-HARs
von: Wang, Yihan, et al.
Veröffentlicht: (2025)
von: Wang, Yihan, et al.
Veröffentlicht: (2025)
E-SAGE: Explainability-based Defense Against Backdoor Attacks on Graph Neural Networks
von: Yuan, Dingqiang, et al.
Veröffentlicht: (2024)
von: Yuan, Dingqiang, et al.
Veröffentlicht: (2024)
System-Level Defense against Indirect Prompt Injection Attacks: An Information Flow Control Perspective
von: Wu, Fangzhou, et al.
Veröffentlicht: (2024)
von: Wu, Fangzhou, et al.
Veröffentlicht: (2024)
FATH: Authentication-based Test-time Defense against Indirect Prompt Injection Attacks
von: Wang, Jiongxiao, et al.
Veröffentlicht: (2024)
von: Wang, Jiongxiao, et al.
Veröffentlicht: (2024)
Adaptive Attacks Break Defenses Against Indirect Prompt Injection Attacks on LLM Agents
von: Zhan, Qiusi, et al.
Veröffentlicht: (2025)
von: Zhan, Qiusi, et al.
Veröffentlicht: (2025)
Injection, Attack and Erasure: Revocable Backdoor Attacks via Machine Unlearning
von: Song, Baogang, et al.
Veröffentlicht: (2025)
von: Song, Baogang, et al.
Veröffentlicht: (2025)
SUAD: Solid-Channel Ultrasound Injection Attack and Defense to Voice Assistants
von: Liu, Chao, et al.
Veröffentlicht: (2025)
von: Liu, Chao, et al.
Veröffentlicht: (2025)
Defending Against Prompt Injection With a Few DefensiveTokens
von: Chen, Sizhe, et al.
Veröffentlicht: (2025)
von: Chen, Sizhe, et al.
Veröffentlicht: (2025)
Attack as Defense: Run-time Backdoor Implantation for Image Content Protection
von: Zhang, Haichuan, et al.
Veröffentlicht: (2024)
von: Zhang, Haichuan, et al.
Veröffentlicht: (2024)
Backdoor Attacks and Defenses in Computer Vision Domain: A Survey
von: Abbasi, Bilal Hussain, et al.
Veröffentlicht: (2025)
von: Abbasi, Bilal Hussain, et al.
Veröffentlicht: (2025)
Prompt Injection Attack to Tool Selection in LLM Agents
von: Shi, Jiawen, et al.
Veröffentlicht: (2025)
von: Shi, Jiawen, et al.
Veröffentlicht: (2025)
Tit-for-Tat: Safeguarding Large Vision-Language Models Against Jailbreak Attacks via Adversarial Defense
von: Hao, Shuyang, et al.
Veröffentlicht: (2025)
von: Hao, Shuyang, et al.
Veröffentlicht: (2025)
MELON: Provable Defense Against Indirect Prompt Injection Attacks in AI Agents
von: Zhu, Kaijie, et al.
Veröffentlicht: (2025)
von: Zhu, Kaijie, et al.
Veröffentlicht: (2025)
A Multi-Agent LLM Defense Pipeline Against Prompt Injection Attacks
von: Hossain, S M Asif, et al.
Veröffentlicht: (2025)
von: Hossain, S M Asif, et al.
Veröffentlicht: (2025)
Backdoored Retrievers for Prompt Injection Attacks on Retrieval Augmented Generation of Large Language Models
von: Clop, Cody, et al.
Veröffentlicht: (2024)
von: Clop, Cody, et al.
Veröffentlicht: (2024)
Exploring Backdoor Attack and Defense for LLM-empowered Recommendations
von: Ning, Liangbo, et al.
Veröffentlicht: (2025)
von: Ning, Liangbo, et al.
Veröffentlicht: (2025)
Formalizing and Benchmarking Prompt Injection Attacks and Defenses
von: Liu, Yupei, et al.
Veröffentlicht: (2023)
von: Liu, Yupei, et al.
Veröffentlicht: (2023)
Defensive Prompt Patch: A Robust and Interpretable Defense of LLMs against Jailbreak Attacks
von: Xiong, Chen, et al.
Veröffentlicht: (2024)
von: Xiong, Chen, et al.
Veröffentlicht: (2024)
PromptShield: Deployable Detection for Prompt Injection Attacks
von: Jacob, Dennis, et al.
Veröffentlicht: (2025)
von: Jacob, Dennis, et al.
Veröffentlicht: (2025)
PromptArmor: Simple yet Effective Prompt Injection Defenses
von: Shi, Tianneng, et al.
Veröffentlicht: (2025)
von: Shi, Tianneng, et al.
Veröffentlicht: (2025)
Backdoor Threats in Variational Quantum Circuits: Taxonomy, Attacks, and Defenses
von: Jiang, Lei, et al.
Veröffentlicht: (2026)
von: Jiang, Lei, et al.
Veröffentlicht: (2026)
ICON: Indirect Prompt Injection Defense for Agents based on Inference-Time Correction
von: Wang, Che, et al.
Veröffentlicht: (2026)
von: Wang, Che, et al.
Veröffentlicht: (2026)
PINA: Prompt Injection Attack against Navigation Agents
von: Liu, Jiani, et al.
Veröffentlicht: (2026)
von: Liu, Jiani, et al.
Veröffentlicht: (2026)
Stealthy Dual-Trigger Backdoors: Attacking Prompt Tuning in LM-Empowered Graph Foundation Models
von: Xue, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Xue, Xiaoyu, et al.
Veröffentlicht: (2025)
The Defense Trilemma: Why Prompt Injection Defense Wrappers Fail?
von: Bhatt, Manish, et al.
Veröffentlicht: (2026)
von: Bhatt, Manish, et al.
Veröffentlicht: (2026)
The Attacker Moves Second: Stronger Adaptive Attacks Bypass Defenses Against Llm Jailbreaks and Prompt Injections
von: Nasr, Milad, et al.
Veröffentlicht: (2025)
von: Nasr, Milad, et al.
Veröffentlicht: (2025)
The Vulnerability of LLM Rankers to Prompt Injection Attacks
von: Yin, Yu, et al.
Veröffentlicht: (2026)
von: Yin, Yu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Defense Against Prompt Injection Attack by Leveraging Attack Techniques
von: Chen, Yulin, et al.
Veröffentlicht: (2024) -
Can Indirect Prompt Injection Attacks Be Detected and Removed?
von: Chen, Yulin, et al.
Veröffentlicht: (2025) -
TopicAttack: An Indirect Prompt Injection Attack via Topic Transition
von: Chen, Yulin, et al.
Veröffentlicht: (2025) -
BaThe: Defense against the Jailbreak Attack in Multimodal Large Language Models by Treating Harmful Instruction as Backdoor Trigger
von: Chen, Yulin, et al.
Veröffentlicht: (2024) -
Robustness via Referencing: Defending against Prompt Injection Attacks by Referencing the Executed Instruction
von: Chen, Yulin, et al.
Veröffentlicht: (2025)