Backdoored Retrievers for Prompt Injection Attacks on Retrieval Augmented Generation of Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Clop, Cody, Teglia, Yannick |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Phantom: General Backdoor Attacks on Retrieval Augmented Language Generation
by: Chaudhari, Harsh, et al.
Published: (2024)
by: Chaudhari, Harsh, et al.
Published: (2024)
Using Retriever Augmented Large Language Models for Attack Graph Generation
by: Prapty, Renascence Tarafder, et al.
Published: (2024)
by: Prapty, Renascence Tarafder, et al.
Published: (2024)
PoisonedRAG: Knowledge Corruption Attacks to Retrieval-Augmented Generation of Large Language Models
by: Zou, Wei, et al.
Published: (2024)
by: Zou, Wei, et al.
Published: (2024)
Backdoor Attacks on Deep Learning Face Detection
by: Roux, Quentin Le, et al.
Published: (2025)
by: Roux, Quentin Le, et al.
Published: (2025)
Backdooring Instruction-Tuned Large Language Models with Virtual Prompt Injection
by: Yan, Jun, et al.
Published: (2023)
by: Yan, Jun, et al.
Published: (2023)
Prompt Injection Attacks on Large Language Models in Oncology
by: Clusmann, Jan, et al.
Published: (2024)
by: Clusmann, Jan, et al.
Published: (2024)
SoK: On the Survivability of Backdoor Attacks on Unconstrained Face Recognition Systems
by: Roux, Quentin Le, et al.
Published: (2025)
by: Roux, Quentin Le, et al.
Published: (2025)
An Early Categorization of Prompt Injection Attacks on Large Language Models
by: Rossi, Sippo, et al.
Published: (2024)
by: Rossi, Sippo, et al.
Published: (2024)
Advancing TTP Analysis: Harnessing the Power of Large Language Models with Retrieval Augmented Generation
by: Fayyazi, Reza, et al.
Published: (2023)
by: Fayyazi, Reza, et al.
Published: (2023)
Poisoned-MRAG: Knowledge Poisoning Attacks to Multimodal Retrieval Augmented Generation
by: Liu, Yinuo, et al.
Published: (2025)
by: Liu, Yinuo, et al.
Published: (2025)
Securing Large Language Models (LLMs) from Prompt Injection Attacks
by: Suri, Omar Farooq Khan, et al.
Published: (2025)
by: Suri, Omar Farooq Khan, et al.
Published: (2025)
Traceback of Poisoning Attacks to Retrieval-Augmented Generation
by: Zhang, Baolei, et al.
Published: (2025)
by: Zhang, Baolei, et al.
Published: (2025)
BadChain: Backdoor Chain-of-Thought Prompting for Large Language Models
by: Xiang, Zhen, et al.
Published: (2024)
by: Xiang, Zhen, et al.
Published: (2024)
Practical Poisoning Attacks against Retrieval-Augmented Generation
by: Zhang, Baolei, et al.
Published: (2025)
by: Zhang, Baolei, et al.
Published: (2025)
Secure Retrieval-Augmented Generation against Poisoning Attacks
by: Cheng, Zirui, et al.
Published: (2025)
by: Cheng, Zirui, et al.
Published: (2025)
Benchmarking Poisoning Attacks against Retrieval-Augmented Generation
by: Zhang, Baolei, et al.
Published: (2025)
by: Zhang, Baolei, et al.
Published: (2025)
Krait: A Backdoor Attack Against Graph Prompt Tuning
by: Song, Ying, et al.
Published: (2024)
by: Song, Ying, et al.
Published: (2024)
Composite Backdoor Attacks Against Large Language Models
by: Huang, Hai, et al.
Published: (2023)
by: Huang, Hai, et al.
Published: (2023)
MARAGE: Transferable Multi-Model Adversarial Attack for Retrieval-Augmented Generation Data Extraction
by: Hu, Xiao, et al.
Published: (2025)
by: Hu, Xiao, et al.
Published: (2025)
Poisoning the Watchtower: Prompt Injection Attacks Against LLM-Augmented Security Operations Through Adversarial Log Content
by: Pandey, Rohan, et al.
Published: (2026)
by: Pandey, Rohan, et al.
Published: (2026)
Imperio: Language-Guided Backdoor Attacks for Arbitrary Model Control
by: Chow, Ka-Ho, et al.
Published: (2024)
by: Chow, Ka-Ho, et al.
Published: (2024)
Adaptive Attacks Break Defenses Against Indirect Prompt Injection Attacks on LLM Agents
by: Zhan, Qiusi, et al.
Published: (2025)
by: Zhan, Qiusi, et al.
Published: (2025)
PIShield: Detecting Prompt Injection Attacks via Intrinsic LLM Features
by: Zou, Wei, et al.
Published: (2025)
by: Zou, Wei, et al.
Published: (2025)
Backdoor Attacks on Discrete Graph Diffusion Models
by: Wang, Jiawen, et al.
Published: (2025)
by: Wang, Jiawen, et al.
Published: (2025)
SteganoBackdoor: Stealthy and Data-Efficient Backdoor Attacks on Language Models
by: Xue, Eric, et al.
Published: (2025)
by: Xue, Eric, et al.
Published: (2025)
DCMI: A Differential Calibration Membership Inference Attack Against Retrieval-Augmented Generation
by: Gao, Xinyu, et al.
Published: (2025)
by: Gao, Xinyu, et al.
Published: (2025)
Medusa: Cross-Modal Transferable Adversarial Attacks on Multimodal Medical Retrieval-Augmented Generation
by: Shang, Yingjia, et al.
Published: (2025)
by: Shang, Yingjia, et al.
Published: (2025)
BadImplant: Injection-based Multi-Targeted Graph Backdoor Attack
by: Khan, Md Nabi Newaz, et al.
Published: (2026)
by: Khan, Md Nabi Newaz, et al.
Published: (2026)
Neural Exec: Learning (and Learning from) Execution Triggers for Prompt Injection Attacks
by: Pasquini, Dario, et al.
Published: (2024)
by: Pasquini, Dario, et al.
Published: (2024)
A Multi-Agent LLM Defense Pipeline Against Prompt Injection Attacks
by: Hossain, S M Asif, et al.
Published: (2025)
by: Hossain, S M Asif, et al.
Published: (2025)
PIDP-Attack: Combining Prompt Injection with Database Poisoning Attacks on Retrieval-Augmented Generation Systems
by: Wang, Haozhen, et al.
Published: (2026)
by: Wang, Haozhen, et al.
Published: (2026)
DeepSweep: An Evaluation Framework for Mitigating DNN Backdoor Attacks using Data Augmentation
by: Qiu, Han, et al.
Published: (2020)
by: Qiu, Han, et al.
Published: (2020)
Watch Out for Your Guidance on Generation! Exploring Conditional Backdoor Attacks against Large Language Models
by: He, Jiaming, et al.
Published: (2024)
by: He, Jiaming, et al.
Published: (2024)
BadMerging: Backdoor Attacks Against Model Merging
by: Zhang, Jinghuai, et al.
Published: (2024)
by: Zhang, Jinghuai, et al.
Published: (2024)
UIBDiffusion: Universal Imperceptible Backdoor Attack for Diffusion Models
by: Han, Yuning, et al.
Published: (2024)
by: Han, Yuning, et al.
Published: (2024)
Persistent Backdoor Attacks in Continual Learning
by: Guo, Zhen, et al.
Published: (2024)
by: Guo, Zhen, et al.
Published: (2024)
Backdoor Attacks on Decentralised Post-Training
by: Ersoy, Oğuzhan, et al.
Published: (2026)
by: Ersoy, Oğuzhan, et al.
Published: (2026)
Awakening the Hydra: Stabilizing Multi-Concept Backdoor Injection in Text-to-Image Diffusion Models
by: Wang, Kai, et al.
Published: (2026)
by: Wang, Kai, et al.
Published: (2026)
Unlearning Backdoor Attacks through Gradient-Based Model Pruning
by: Dunnett, Kealan, et al.
Published: (2024)
by: Dunnett, Kealan, et al.
Published: (2024)
Diff-Cleanse: Identifying and Mitigating Backdoor Attacks in Diffusion Models
by: Hao, Jiang, et al.
Published: (2024)
by: Hao, Jiang, et al.
Published: (2024)
Similar Items
-
Phantom: General Backdoor Attacks on Retrieval Augmented Language Generation
by: Chaudhari, Harsh, et al.
Published: (2024) -
Using Retriever Augmented Large Language Models for Attack Graph Generation
by: Prapty, Renascence Tarafder, et al.
Published: (2024) -
PoisonedRAG: Knowledge Corruption Attacks to Retrieval-Augmented Generation of Large Language Models
by: Zou, Wei, et al.
Published: (2024) -
Backdoor Attacks on Deep Learning Face Detection
by: Roux, Quentin Le, et al.
Published: (2025) -
Backdooring Instruction-Tuned Large Language Models with Virtual Prompt Injection
by: Yan, Jun, et al.
Published: (2023)