ConfusedPilot: Confused Deputy Risks in RAG-based LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | RoyChowdhury, Ayush, Luo, Mulong, Sahu, Prateek, Banerjee, Sarbartha, Tiwari, Mohit |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SoK: A Systems Perspective on Compound AI Threats and Countermeasures
von: Banerjee, Sarbartha, et al.
Veröffentlicht: (2024)
von: Banerjee, Sarbartha, et al.
Veröffentlicht: (2024)
Cascade: Composing Software-Hardware Attack Gadgets for Adversarial Threat Amplification in Compound AI Systems
von: Banerjee, Sarbartha, et al.
Veröffentlicht: (2026)
von: Banerjee, Sarbartha, et al.
Veröffentlicht: (2026)
Towards Reinforcement Learning for Exploration of Speculative Execution Vulnerabilities
von: Lai, Evan, et al.
Veröffentlicht: (2025)
von: Lai, Evan, et al.
Veröffentlicht: (2025)
Confusion is the Final Barrier: Rethinking Jailbreak Evaluation and Investigating the Real Misuse Threat of LLMs
von: Yan, Yu, et al.
Veröffentlicht: (2025)
von: Yan, Yu, et al.
Veröffentlicht: (2025)
Prompt Injection as Role Confusion
von: Ye, Charles, et al.
Veröffentlicht: (2026)
von: Ye, Charles, et al.
Veröffentlicht: (2026)
Speed Kills: Exploring Confused Deputy Attacks Through Edge AI Accelerators
von: Danduri, Datta Manikanta Sri Hari, et al.
Veröffentlicht: (2026)
von: Danduri, Datta Manikanta Sri Hari, et al.
Veröffentlicht: (2026)
Retrieval-Confused Generation is a Good Defender for Privacy Violation Attack of Large Language Models
von: Peng, Wanli, et al.
Veröffentlicht: (2025)
von: Peng, Wanli, et al.
Veröffentlicht: (2025)
ReliabilityRAG: Effective and Provably Robust Defense for RAG-based Web-Search
von: Shen, Zeyu, et al.
Veröffentlicht: (2025)
von: Shen, Zeyu, et al.
Veröffentlicht: (2025)
Obsidian: Cooperative State-Space Exploration for Performant Inference on Secure ML Accelerators
von: Banerjee, Sarbartha, et al.
Veröffentlicht: (2024)
von: Banerjee, Sarbartha, et al.
Veröffentlicht: (2024)
Give Them an Inch and They Will Take a Mile:Understanding and Measuring Caller Identity Confusion in MCP-Based AI Systems
von: Huang, Yuhang, et al.
Veröffentlicht: (2026)
von: Huang, Yuhang, et al.
Veröffentlicht: (2026)
False Friends in the Shell: Unveiling the Emoticon Semantic Confusion in Large Language Models
von: Jiang, Weipeng, et al.
Veröffentlicht: (2026)
von: Jiang, Weipeng, et al.
Veröffentlicht: (2026)
ConfusionPrompt: Practical Private Inference for Online Large Language Models
von: Mai, Peihua, et al.
Veröffentlicht: (2023)
von: Mai, Peihua, et al.
Veröffentlicht: (2023)
CloudLens: Modeling and Detecting Cloud Security Vulnerabilities
von: Kazdagli, Mikhail, et al.
Veröffentlicht: (2024)
von: Kazdagli, Mikhail, et al.
Veröffentlicht: (2024)
Diffuse or Confuse: A Diffusion Deepfake Speech Dataset
von: Firc, Anton, et al.
Veröffentlicht: (2024)
von: Firc, Anton, et al.
Veröffentlicht: (2024)
Epistemic Bias Injection: Biasing LLMs via Selective Context Retrieval
von: Wu, Hao, et al.
Veröffentlicht: (2025)
von: Wu, Hao, et al.
Veröffentlicht: (2025)
Advanced Real-Time Fraud Detection Using RAG-Based LLMs
von: Singh, Gurjot, et al.
Veröffentlicht: (2025)
von: Singh, Gurjot, et al.
Veröffentlicht: (2025)
On Evaluating the Durability of Safeguards for Open-Weight LLMs
von: Qi, Xiangyu, et al.
Veröffentlicht: (2024)
von: Qi, Xiangyu, et al.
Veröffentlicht: (2024)
Advancing Vulnerability Classification with BERT: A Multi-Objective Learning Model
von: Tiwari, Himanshu
Veröffentlicht: (2025)
von: Tiwari, Himanshu
Veröffentlicht: (2025)
Leveraging RAG for Training-Free Alignment of LLMs
von: Halloran, John T.
Veröffentlicht: (2026)
von: Halloran, John T.
Veröffentlicht: (2026)
Deep Learning Under Siege: Identifying Security Vulnerabilities and Risk Mitigation Strategies
von: Al-Karaki, Jamal, et al.
Veröffentlicht: (2024)
von: Al-Karaki, Jamal, et al.
Veröffentlicht: (2024)
RAG-targeted Adversarial Attack on LLM-based Threat Detection and Mitigation Framework
von: Ikbarieh, Seif, et al.
Veröffentlicht: (2025)
von: Ikbarieh, Seif, et al.
Veröffentlicht: (2025)
Deep Learning-based Dual Watermarking for Image Copyright Protection and Authentication
von: Padhi, Sudev Kumar, et al.
Veröffentlicht: (2025)
von: Padhi, Sudev Kumar, et al.
Veröffentlicht: (2025)
Safety Alignment Should Be Made More Than Just a Few Tokens Deep
von: Qi, Xiangyu, et al.
Veröffentlicht: (2024)
von: Qi, Xiangyu, et al.
Veröffentlicht: (2024)
Multimodal Prompt Injection Attacks: Risks and Defenses for Modern LLMs
von: Yeo, Andrew, et al.
Veröffentlicht: (2025)
von: Yeo, Andrew, et al.
Veröffentlicht: (2025)
Identifying and Addressing User-level Security Concerns in Smart Homes Using "Smaller" LLMs
von: Chowdhury, Hafijul Hoque, et al.
Veröffentlicht: (2025)
von: Chowdhury, Hafijul Hoque, et al.
Veröffentlicht: (2025)
Securing RAG: A Risk Assessment and Mitigation Framework
von: Ammann, Lukas, et al.
Veröffentlicht: (2025)
von: Ammann, Lukas, et al.
Veröffentlicht: (2025)
P$^2$RAG: Efficient Privacy-Preserving RAG Service Supporting Arbitrary Top-$k$ Retrieval
von: Ming, Yulong, et al.
Veröffentlicht: (2026)
von: Ming, Yulong, et al.
Veröffentlicht: (2026)
Understanding Secret Leakage Risks in Code LLMs: A Tokenization Perspective
von: Chen, Meifang, et al.
Veröffentlicht: (2026)
von: Chen, Meifang, et al.
Veröffentlicht: (2026)
CTRAPS: CTAP Client Impersonation and API Confusion on FIDO2
von: Casagrande, Marco, et al.
Veröffentlicht: (2024)
von: Casagrande, Marco, et al.
Veröffentlicht: (2024)
How Does Naming Affect LLMs on Code Analysis Tasks?
von: Wang, Zhilong, et al.
Veröffentlicht: (2023)
von: Wang, Zhilong, et al.
Veröffentlicht: (2023)
Context manipulation attacks : Web agents are susceptible to corrupted memory
von: Patlan, Atharv Singh, et al.
Veröffentlicht: (2025)
von: Patlan, Atharv Singh, et al.
Veröffentlicht: (2025)
AgenTRIM: Tool Risk Mitigation for Agentic AI
von: Betser, Roy, et al.
Veröffentlicht: (2026)
von: Betser, Roy, et al.
Veröffentlicht: (2026)
Analysing Safety Risks in LLMs Fine-Tuned with Pseudo-Malicious Cyber Security Data
von: ElZemity, Adel, et al.
Veröffentlicht: (2025)
von: ElZemity, Adel, et al.
Veröffentlicht: (2025)
Automated Framework to Evaluate and Harden LLM System Instructions against Encoding Attacks
von: Sahu, Anubhab, et al.
Veröffentlicht: (2026)
von: Sahu, Anubhab, et al.
Veröffentlicht: (2026)
Unleashing Worms and Extracting Data: Escalating the Outcome of Attacks against RAG-based Inference in Scale and Severity Using Jailbreaking
von: Cohen, Stav, et al.
Veröffentlicht: (2024)
von: Cohen, Stav, et al.
Veröffentlicht: (2024)
Large Language Models are Easily Confused: A Quantitative Metric, Security Implications and Typological Analysis
von: Chen, Yiyi, et al.
Veröffentlicht: (2024)
von: Chen, Yiyi, et al.
Veröffentlicht: (2024)
Privacy-Aware RAG: Secure and Isolated Knowledge Retrieval
von: Zhou, Pengcheng, et al.
Veröffentlicht: (2025)
von: Zhou, Pengcheng, et al.
Veröffentlicht: (2025)
The Sum Leaks More Than Its Parts: Compositional Privacy Risks and Mitigations in Multi-Agent Collaboration
von: Patil, Vaidehi, et al.
Veröffentlicht: (2025)
von: Patil, Vaidehi, et al.
Veröffentlicht: (2025)
One Leak Away: How Pretrained Model Exposure Amplifies Jailbreak Risks in Finetuned LLMs
von: Tan, Yixin, et al.
Veröffentlicht: (2025)
von: Tan, Yixin, et al.
Veröffentlicht: (2025)
Scaling Homomorphic Applications in Deployment
von: Marinelli, Ryan, et al.
Veröffentlicht: (2025)
von: Marinelli, Ryan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SoK: A Systems Perspective on Compound AI Threats and Countermeasures
von: Banerjee, Sarbartha, et al.
Veröffentlicht: (2024) -
Cascade: Composing Software-Hardware Attack Gadgets for Adversarial Threat Amplification in Compound AI Systems
von: Banerjee, Sarbartha, et al.
Veröffentlicht: (2026) -
Towards Reinforcement Learning for Exploration of Speculative Execution Vulnerabilities
von: Lai, Evan, et al.
Veröffentlicht: (2025) -
Confusion is the Final Barrier: Rethinking Jailbreak Evaluation and Investigating the Real Misuse Threat of LLMs
von: Yan, Yu, et al.
Veröffentlicht: (2025) -
Prompt Injection as Role Confusion
von: Ye, Charles, et al.
Veröffentlicht: (2026)