SD-RAG: A Prompt-Injection-Resilient Framework for Selective Disclosure in Retrieval-Augmented Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Masoud, Aiman Al, Arazzi, Marco, Nocera, Antonino |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LoRA as Oracle
von: Arazzi, Marco, et al.
Veröffentlicht: (2026)
von: Arazzi, Marco, et al.
Veröffentlicht: (2026)
A Deep Reinforcement Learning Approach for Security-Aware Service Acquisition in IoT
von: Arazzi, Marco, et al.
Veröffentlicht: (2024)
von: Arazzi, Marco, et al.
Veröffentlicht: (2024)
Privacy Preserving and Robust Aggregation for Cross-Silo Federated Learning in Non-IID Settings
von: Arazzi, Marco, et al.
Veröffentlicht: (2025)
von: Arazzi, Marco, et al.
Veröffentlicht: (2025)
Secure Federated Data Distillation
von: Arazzi, Marco, et al.
Veröffentlicht: (2025)
von: Arazzi, Marco, et al.
Veröffentlicht: (2025)
Subject Data Auditing via Source Inference Attack in Cross-Silo Federated Learning
von: Li, Jiaxin, et al.
Veröffentlicht: (2024)
von: Li, Jiaxin, et al.
Veröffentlicht: (2024)
Security in LLM-as-a-Judge: A Comprehensive SoK
von: Masoud, Aiman Al, et al.
Veröffentlicht: (2026)
von: Masoud, Aiman Al, et al.
Veröffentlicht: (2026)
SecureBreak -- A dataset towards safe and secure models
von: Arazzi, Marco, et al.
Veröffentlicht: (2026)
von: Arazzi, Marco, et al.
Veröffentlicht: (2026)
XBreaking: Understanding how LLMs security alignment can be broken
von: Arazzi, Marco, et al.
Veröffentlicht: (2025)
von: Arazzi, Marco, et al.
Veröffentlicht: (2025)
When Forgetting Triggers Backdoors: A Clean Unlearning Attack
von: Arazzi, Marco, et al.
Veröffentlicht: (2025)
von: Arazzi, Marco, et al.
Veröffentlicht: (2025)
Exploring Structural Complexity in Normative RAG with Graph-based approaches: A case study on the ETSI Standards
von: Masoud, Aiman Al, et al.
Veröffentlicht: (2026)
von: Masoud, Aiman Al, et al.
Veröffentlicht: (2026)
KDk: A Defense Mechanism Against Label Inference Attacks in Vertical Federated Learning
von: Arazzi, Marco, et al.
Veröffentlicht: (2024)
von: Arazzi, Marco, et al.
Veröffentlicht: (2024)
A Novel IoT Trust Model Leveraging Fully Distributed Behavioral Fingerprinting and Secure Delegation
von: Arazzi, Marco, et al.
Veröffentlicht: (2023)
von: Arazzi, Marco, et al.
Veröffentlicht: (2023)
PIDP-Attack: Combining Prompt Injection with Database Poisoning Attacks on Retrieval-Augmented Generation Systems
von: Wang, Haozhen, et al.
Veröffentlicht: (2026)
von: Wang, Haozhen, et al.
Veröffentlicht: (2026)
SoK: The Last Line of Defense: On Backdoor Defense Evaluation
von: Abad, Gorka, et al.
Veröffentlicht: (2025)
von: Abad, Gorka, et al.
Veröffentlicht: (2025)
Privacy-Preserving in Blockchain-based Federated Learning Systems
von: M., Sameera K., et al.
Veröffentlicht: (2024)
von: M., Sameera K., et al.
Veröffentlicht: (2024)
PR-Attack: Coordinated Prompt-RAG Attacks on Retrieval-Augmented Generation in Large Language Models via Bilevel Optimization
von: Jiao, Yang, et al.
Veröffentlicht: (2025)
von: Jiao, Yang, et al.
Veröffentlicht: (2025)
Prompt Injection as an Emerging Threat: Evaluating the Resilience of Large Language Models
von: Ganiuly, Daniyal, et al.
Veröffentlicht: (2025)
von: Ganiuly, Daniyal, et al.
Veröffentlicht: (2025)
Let's Focus: Focused Backdoor Attack against Federated Transfer Learning
von: Arazzi, Marco, et al.
Veröffentlicht: (2024)
von: Arazzi, Marco, et al.
Veröffentlicht: (2024)
ClawGuard: A Runtime Security Framework for Tool-Augmented LLM Agents Against Indirect Prompt Injection
von: Zhao, Wei, et al.
Veröffentlicht: (2026)
von: Zhao, Wei, et al.
Veröffentlicht: (2026)
PIR-RAG: A System for Private Information Retrieval in Retrieval-Augmented Generation
von: Wang, Baiqiang, et al.
Veröffentlicht: (2025)
von: Wang, Baiqiang, et al.
Veröffentlicht: (2025)
Overcoming the Retrieval Barrier: Indirect Prompt Injection in the Wild for LLM Systems
von: Chang, Hongyan, et al.
Veröffentlicht: (2026)
von: Chang, Hongyan, et al.
Veröffentlicht: (2026)
RAG-WM: An Efficient Black-Box Watermarking Approach for Retrieval-Augmented Generation of Large Language Models
von: Lv, Peizhuo, et al.
Veröffentlicht: (2025)
von: Lv, Peizhuo, et al.
Veröffentlicht: (2025)
SecureRAG-RTL: A Retrieval-Augmented, Multi-Agent, Zero-Shot LLM-Driven Framework for Hardware Vulnerability Detection
von: Hasan, Touseef, et al.
Veröffentlicht: (2026)
von: Hasan, Touseef, et al.
Veröffentlicht: (2026)
You Snooze, You Lose: Automatic Safety Alignment Restoration through Neural Weight Translation
von: Arazzi, Marco, et al.
Veröffentlicht: (2026)
von: Arazzi, Marco, et al.
Veröffentlicht: (2026)
Bypassing Prompt Injection Detectors through Evasive Injections
von: Rahman, Md Jahedur, et al.
Veröffentlicht: (2026)
von: Rahman, Md Jahedur, et al.
Veröffentlicht: (2026)
PromptLocate: Localizing Prompt Injection Attacks
von: Jia, Yuqi, et al.
Veröffentlicht: (2025)
von: Jia, Yuqi, et al.
Veröffentlicht: (2025)
The Good and The Bad: Exploring Privacy Issues in Retrieval-Augmented Generation (RAG)
von: Zeng, Shenglai, et al.
Veröffentlicht: (2024)
von: Zeng, Shenglai, et al.
Veröffentlicht: (2024)
Defeating Prompt Injections by Design
von: Debenedetti, Edoardo, et al.
Veröffentlicht: (2025)
von: Debenedetti, Edoardo, et al.
Veröffentlicht: (2025)
The Defense Trilemma: Why Prompt Injection Defense Wrappers Fail?
von: Bhatt, Manish, et al.
Veröffentlicht: (2026)
von: Bhatt, Manish, et al.
Veröffentlicht: (2026)
Toward Trustworthy Agentic AI: A Multimodal Framework for Preventing Prompt Injection Attacks
von: Syed, Toqeer Ali, et al.
Veröffentlicht: (2025)
von: Syed, Toqeer Ali, et al.
Veröffentlicht: (2025)
Do Multimodal RAG Systems Leak Data? A Comprehensive Evaluation of Membership Inference and Image Caption Retrieval Attacks
von: Al-Lawati, Ali, et al.
Veröffentlicht: (2026)
von: Al-Lawati, Ali, et al.
Veröffentlicht: (2026)
SafeRAG: Benchmarking Security in Retrieval-Augmented Generation of Large Language Model
von: Liang, Xun, et al.
Veröffentlicht: (2025)
von: Liang, Xun, et al.
Veröffentlicht: (2025)
PromptArmor: Simple yet Effective Prompt Injection Defenses
von: Shi, Tianneng, et al.
Veröffentlicht: (2025)
von: Shi, Tianneng, et al.
Veröffentlicht: (2025)
Provably Secure Retrieval-Augmented Generation
von: Zhou, Pengcheng, et al.
Veröffentlicht: (2025)
von: Zhou, Pengcheng, et al.
Veröffentlicht: (2025)
Epistemic Bias Injection: Biasing LLMs via Selective Context Retrieval
von: Wu, Hao, et al.
Veröffentlicht: (2025)
von: Wu, Hao, et al.
Veröffentlicht: (2025)
To Protect the LLM Agent Against the Prompt Injection Attack with Polymorphic Prompt
von: Wang, Zhilong, et al.
Veröffentlicht: (2025)
von: Wang, Zhilong, et al.
Veröffentlicht: (2025)
A Critical Evaluation of Defenses against Prompt Injection Attacks
von: Jia, Yuqi, et al.
Veröffentlicht: (2025)
von: Jia, Yuqi, et al.
Veröffentlicht: (2025)
CourtGuard: A Local, Multiagent Prompt Injection Classifier
von: Wu, Isaac, et al.
Veröffentlicht: (2025)
von: Wu, Isaac, et al.
Veröffentlicht: (2025)
Analysis of LLMs Against Prompt Injection and Jailbreak Attacks
von: Jaiswal, Piyush, et al.
Veröffentlicht: (2026)
von: Jaiswal, Piyush, et al.
Veröffentlicht: (2026)
Evaluation of Prompt Injection Defenses in Large Language Models
von: Deep, Priyal, et al.
Veröffentlicht: (2026)
von: Deep, Priyal, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
LoRA as Oracle
von: Arazzi, Marco, et al.
Veröffentlicht: (2026) -
A Deep Reinforcement Learning Approach for Security-Aware Service Acquisition in IoT
von: Arazzi, Marco, et al.
Veröffentlicht: (2024) -
Privacy Preserving and Robust Aggregation for Cross-Silo Federated Learning in Non-IID Settings
von: Arazzi, Marco, et al.
Veröffentlicht: (2025) -
Secure Federated Data Distillation
von: Arazzi, Marco, et al.
Veröffentlicht: (2025) -
Subject Data Auditing via Source Inference Attack in Cross-Silo Federated Learning
von: Li, Jiaxin, et al.
Veröffentlicht: (2024)