ReliabilityRAG: Effective and Provably Robust Defense for RAG-based Web-Search
Fuente:
arXiv
Saved in:
| Main Authors: | Shen, Zeyu, Imana, Basileal, Wu, Tong, Xiang, Chong, Mittal, Prateek, Korolova, Aleksandra |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Certifiably Robust RAG against Retrieval Corruption
by: Xiang, Chong, et al.
Published: (2024)
by: Xiang, Chong, et al.
Published: (2024)
ConfusedPilot: Confused Deputy Risks in RAG-based LLMs
by: RoyChowdhury, Ayush, et al.
Published: (2024)
by: RoyChowdhury, Ayush, et al.
Published: (2024)
PatchCURE: Improving Certifiable Robustness, Model Utility, and Computation Efficiency of Adversarial Patch Defenses
by: Xiang, Chong, et al.
Published: (2023)
by: Xiang, Chong, et al.
Published: (2023)
Position: Towards Resilience Against Adversarial Examples
by: Dai, Sihui, et al.
Published: (2024)
by: Dai, Sihui, et al.
Published: (2024)
Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAG
by: Choudhary, Sarthak, et al.
Published: (2025)
by: Choudhary, Sarthak, et al.
Published: (2025)
Context manipulation attacks : Web agents are susceptible to corrupted memory
by: Patlan, Atharv Singh, et al.
Published: (2025)
by: Patlan, Atharv Singh, et al.
Published: (2025)
EmoRAG: Evaluating RAG Robustness to Symbolic Perturbations
by: Zhou, Xinyun, et al.
Published: (2025)
by: Zhou, Xinyun, et al.
Published: (2025)
Ward: Provable RAG Dataset Inference via LLM Watermarks
by: Jovanović, Nikola, et al.
Published: (2024)
by: Jovanović, Nikola, et al.
Published: (2024)
External Evaluation of Discrimination Mitigation Efforts in Meta's Ad Delivery
by: Imana, Basileal, et al.
Published: (2025)
by: Imana, Basileal, et al.
Published: (2025)
P$^2$RAG: Efficient Privacy-Preserving RAG Service Supporting Arbitrary Top-$k$ Retrieval
by: Ming, Yulong, et al.
Published: (2026)
by: Ming, Yulong, et al.
Published: (2026)
Semantic Chameleon: Corpus-Dependent Poisoning Attacks and Defenses in RAG Systems
by: Thornton, Scott
Published: (2026)
by: Thornton, Scott
Published: (2026)
PatchDEMUX: A Certifiably Robust Framework for Multi-label Classifiers Against Adversarial Patches
by: Jacob, Dennis, et al.
Published: (2025)
by: Jacob, Dennis, et al.
Published: (2025)
Auditing for Racial Discrimination in the Delivery of Education Ads
by: Imana, Basileal, et al.
Published: (2024)
by: Imana, Basileal, et al.
Published: (2024)
Auditing for Bias in Ad Delivery Using Inferred Demographic Attributes
by: Imana, Basileal, et al.
Published: (2024)
by: Imana, Basileal, et al.
Published: (2024)
ADMIT: Few-shot Knowledge Poisoning Attacks on RAG-based Fact Checking
by: Wu, Yutao, et al.
Published: (2025)
by: Wu, Yutao, et al.
Published: (2025)
WARD: Adversarially Robust Defense of Web Agents Against Prompt Injections
by: Cao, Tri, et al.
Published: (2026)
by: Cao, Tri, et al.
Published: (2026)
Subgraph Reconstruction Attacks on Graph RAG Deployments with Practical Defenses
by: Song, Minkyoo, et al.
Published: (2026)
by: Song, Minkyoo, et al.
Published: (2026)
RAG-targeted Adversarial Attack on LLM-based Threat Detection and Mitigation Framework
by: Ikbarieh, Seif, et al.
Published: (2025)
by: Ikbarieh, Seif, et al.
Published: (2025)
CyberRAG: An Agentic RAG cyber attack classification and reporting tool
by: Blefari, Francesco, et al.
Published: (2025)
by: Blefari, Francesco, et al.
Published: (2025)
RAG with Differential Privacy
by: Grislain, Nicolas
Published: (2024)
by: Grislain, Nicolas
Published: (2024)
Private-RAG: Answering Multiple Queries with LLMs while Keeping Your Data Private
by: Wu, Ruihan, et al.
Published: (2025)
by: Wu, Ruihan, et al.
Published: (2025)
Privacy-Aware RAG: Secure and Isolated Knowledge Retrieval
by: Zhou, Pengcheng, et al.
Published: (2025)
by: Zhou, Pengcheng, et al.
Published: (2025)
MELON: Provable Defense Against Indirect Prompt Injection Attacks in AI Agents
by: Zhu, Kaijie, et al.
Published: (2025)
by: Zhu, Kaijie, et al.
Published: (2025)
Real AI Agents with Fake Memories: Fatal Context Manipulation Attacks on Web3 Agents
by: Patlan, Atharv Singh, et al.
Published: (2025)
by: Patlan, Atharv Singh, et al.
Published: (2025)
Safeguarding Multimodal Knowledge Copyright in the RAG-as-a-Service Environment
by: Chen, Tianyu, et al.
Published: (2025)
by: Chen, Tianyu, et al.
Published: (2025)
ARCeR: an Agentic RAG for the Automated Definition of Cyber Ranges
by: Lupinacci, Matteo, et al.
Published: (2025)
by: Lupinacci, Matteo, et al.
Published: (2025)
RAGShield: Detecting Numerical Claim Manipulation in Government RAG Systems
by: Patil, KrishnaSaiReddy
Published: (2026)
by: Patil, KrishnaSaiReddy
Published: (2026)
The Hidden Threat in Plain Text: Attacking RAG Data Loaders
by: Castagnaro, Alberto, et al.
Published: (2025)
by: Castagnaro, Alberto, et al.
Published: (2025)
RAG Security and Privacy: Formalizing the Threat Model and Attack Surface
by: Arzanipour, Atousa, et al.
Published: (2025)
by: Arzanipour, Atousa, et al.
Published: (2025)
Bidirectional Intention Inference Enhances LLMs' Defense Against Multi-Turn Jailbreak Attacks
by: Tong, Haibo, et al.
Published: (2025)
by: Tong, Haibo, et al.
Published: (2025)
GraphRAG under Fire
by: Liang, Jiacheng, et al.
Published: (2025)
by: Liang, Jiacheng, et al.
Published: (2025)
Advanced Real-Time Fraud Detection Using RAG-Based LLMs
by: Singh, Gurjot, et al.
Published: (2025)
by: Singh, Gurjot, et al.
Published: (2025)
Unleashing Worms and Extracting Data: Escalating the Outcome of Attacks against RAG-based Inference in Scale and Severity Using Jailbreaking
by: Cohen, Stav, et al.
Published: (2024)
by: Cohen, Stav, et al.
Published: (2024)
AdaPhish: AI-Powered Adaptive Defense and Education Resource Against Deceptive Emails
by: Meguro, Rei, et al.
Published: (2025)
by: Meguro, Rei, et al.
Published: (2025)
Silent Leaks: Implicit Knowledge Extraction Attack on RAG Systems through Benign Queries
by: Wang, Yuhao, et al.
Published: (2025)
by: Wang, Yuhao, et al.
Published: (2025)
Cordon-MAS: Defending RAG against Knowledge Poisoning via Information-Flow Control
by: Yu, Zhe, et al.
Published: (2026)
by: Yu, Zhe, et al.
Published: (2026)
Leveraging RAG for Training-Free Alignment of LLMs
by: Halloran, John T.
Published: (2026)
by: Halloran, John T.
Published: (2026)
Spa-VLM: Stealthy Poisoning Attacks on RAG-based VLM
by: Yu, Lei, et al.
Published: (2025)
by: Yu, Lei, et al.
Published: (2025)
Uncovering, Explaining, and Mitigating the Superficial Safety of Backdoor Defense
by: Min, Rui, et al.
Published: (2024)
by: Min, Rui, et al.
Published: (2024)
Instructional Segment Embedding: Improving LLM Safety with Instruction Hierarchy
by: Wu, Tong, et al.
Published: (2024)
by: Wu, Tong, et al.
Published: (2024)
Similar Items
-
Certifiably Robust RAG against Retrieval Corruption
by: Xiang, Chong, et al.
Published: (2024) -
ConfusedPilot: Confused Deputy Risks in RAG-based LLMs
by: RoyChowdhury, Ayush, et al.
Published: (2024) -
PatchCURE: Improving Certifiable Robustness, Model Utility, and Computation Efficiency of Adversarial Patch Defenses
by: Xiang, Chong, et al.
Published: (2023) -
Position: Towards Resilience Against Adversarial Examples
by: Dai, Sihui, et al.
Published: (2024) -
Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAG
by: Choudhary, Sarthak, et al.
Published: (2025)