Epistemic Bias Injection: Biasing LLMs via Selective Context Retrieval
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Hao, Saxena, Prateek |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Balancing Privacy and Efficiency: Music Information Retrieval via Additive Homomorphic Encryption
by: Wang, William Zerong, et al.
Published: (2025)
by: Wang, William Zerong, et al.
Published: (2025)
The Phish, The Spam, and The Valid: Generating Feature-Rich Emails for Benchmarking LLMs
by: Toth, Rebeka, et al.
Published: (2025)
by: Toth, Rebeka, et al.
Published: (2025)
Please Don't Kill My Vibe: Empowering Agents with Data Flow Control
by: Summers, Charlie, et al.
Published: (2025)
by: Summers, Charlie, et al.
Published: (2025)
SAMEP: A Secure Protocol for Persistent Context Sharing Across AI Agents
by: Masoor, Hari
Published: (2025)
by: Masoor, Hari
Published: (2025)
The Impact of Event Data Partitioning on Privacy-aware Process Discovery
by: Lim, Jungeun, et al.
Published: (2025)
by: Lim, Jungeun, et al.
Published: (2025)
Practical and Ready-to-Use Methodology to Assess the re-identification Risk in Anonymized Datasets
by: Sondeck, Louis-Philippe, et al.
Published: (2025)
by: Sondeck, Louis-Philippe, et al.
Published: (2025)
An advanced data fabric architecture leveraging homomorphic encryption and federated learning
by: Rieyan, Sakib Anwar, et al.
Published: (2024)
by: Rieyan, Sakib Anwar, et al.
Published: (2024)
Tabular Data Synthesis with Differential Privacy: A Survey
by: Yang, Mengmeng, et al.
Published: (2024)
by: Yang, Mengmeng, et al.
Published: (2024)
Deontic Knowledge Graphs for Privacy Compliance in Multimodal Disaster Data Sharing
by: Echenim, Kelvin Uzoma, et al.
Published: (2026)
by: Echenim, Kelvin Uzoma, et al.
Published: (2026)
SD-RAG: A Prompt-Injection-Resilient Framework for Selective Disclosure in Retrieval-Augmented Generation
by: Masoud, Aiman Al, et al.
Published: (2026)
by: Masoud, Aiman Al, et al.
Published: (2026)
Attacking Byzantine Robust Aggregation in High Dimensions
by: Choudhary, Sarthak, et al.
Published: (2023)
by: Choudhary, Sarthak, et al.
Published: (2023)
Context manipulation attacks : Web agents are susceptible to corrupted memory
by: Patlan, Atharv Singh, et al.
Published: (2025)
by: Patlan, Atharv Singh, et al.
Published: (2025)
Decoding Latent Attack Surfaces in LLMs: Prompt Injection via HTML in Web Summarization
by: Verma, Ishaan, et al.
Published: (2025)
by: Verma, Ishaan, et al.
Published: (2025)
Are Your LLM-based Text-to-SQL Models Secure? Exploring SQL Injection via Backdoor Attacks
by: Lin, Meiyu, et al.
Published: (2025)
by: Lin, Meiyu, et al.
Published: (2025)
MemPot: Defending Against Memory Extraction Attack with Optimized Honeypots
by: Wang, Yuhao, et al.
Published: (2026)
by: Wang, Yuhao, et al.
Published: (2026)
Learning from Anonymized and Incomplete Tabular Data
by: Lange, Lucas, et al.
Published: (2026)
by: Lange, Lucas, et al.
Published: (2026)
Learning Federated Neural Graph Databases for Answering Complex Queries from Distributed Knowledge Graphs
by: Hu, Qi, et al.
Published: (2024)
by: Hu, Qi, et al.
Published: (2024)
Analysis of LLMs Against Prompt Injection and Jailbreak Attacks
by: Jaiswal, Piyush, et al.
Published: (2026)
by: Jaiswal, Piyush, et al.
Published: (2026)
Multimodal Prompt Injection Attacks: Risks and Defenses for Modern LLMs
by: Yeo, Andrew, et al.
Published: (2025)
by: Yeo, Andrew, et al.
Published: (2025)
Overcoming the Retrieval Barrier: Indirect Prompt Injection in the Wild for LLM Systems
by: Chang, Hongyan, et al.
Published: (2026)
by: Chang, Hongyan, et al.
Published: (2026)
ConfusedPilot: Confused Deputy Risks in RAG-based LLMs
by: RoyChowdhury, Ayush, et al.
Published: (2024)
by: RoyChowdhury, Ayush, et al.
Published: (2024)
PROMPTFUZZ: Harnessing Fuzzing Techniques for Robust Testing of Prompt Injection in LLMs
by: Yu, Jiahao, et al.
Published: (2024)
by: Yu, Jiahao, et al.
Published: (2024)
Invisible Threats from Model Context Protocol: Generating Stealthy Injection Payload via Tree-based Adaptive Search
by: Shen, Yulin, et al.
Published: (2026)
by: Shen, Yulin, et al.
Published: (2026)
PISanitizer: Preventing Prompt Injection to Long-Context LLMs via Prompt Sanitization
by: Geng, Runpeng, et al.
Published: (2025)
by: Geng, Runpeng, et al.
Published: (2025)
Unified Threat Detection and Mitigation Framework (UTDMF): Combating Prompt Injection, Deception, and Bias in Enterprise-Scale Transformers
by: KumarRavindran, Santhosh
Published: (2025)
by: KumarRavindran, Santhosh
Published: (2025)
Hijacking Large Audio-Language Models via Context-Agnostic and Imperceptible Auditory Prompt Injection
by: Chen, Meng, et al.
Published: (2026)
by: Chen, Meng, et al.
Published: (2026)
PIDP-Attack: Combining Prompt Injection with Database Poisoning Attacks on Retrieval-Augmented Generation Systems
by: Wang, Haozhen, et al.
Published: (2026)
by: Wang, Haozhen, et al.
Published: (2026)
bi-GRPO: Bidirectional Optimization for Jailbreak Backdoor Injection on LLMs
by: Ji, Wence, et al.
Published: (2025)
by: Ji, Wence, et al.
Published: (2025)
ShadowCode: Towards (Automatic) External Prompt Injection Attack against Code LLMs
by: Yang, Yuchen, et al.
Published: (2024)
by: Yang, Yuchen, et al.
Published: (2024)
AdapTools: Adaptive Tool-based Indirect Prompt Injection Attacks on Agentic LLMs
by: Wang, Che, et al.
Published: (2026)
by: Wang, Che, et al.
Published: (2026)
Know Thy Enemy: Securing LLMs Against Prompt Injection via Diverse Data Synthesis and Instruction-Level Chain-of-Thought Learning
by: Chang, Zhiyuan, et al.
Published: (2026)
by: Chang, Zhiyuan, et al.
Published: (2026)
Retrieval-Augmented LLMs for Security Incident Analysis
by: Cadet, Xavier, et al.
Published: (2026)
by: Cadet, Xavier, et al.
Published: (2026)
Too Easily Fooled? Prompt Injection Breaks LLMs on Frustratingly Simple Multiple-Choice Questions
by: Guo, Xuyang, et al.
Published: (2025)
by: Guo, Xuyang, et al.
Published: (2025)
On Evaluating the Durability of Safeguards for Open-Weight LLMs
by: Qi, Xiangyu, et al.
Published: (2024)
by: Qi, Xiangyu, et al.
Published: (2024)
Bypassing Prompt Injection Detectors through Evasive Injections
by: Rahman, Md Jahedur, et al.
Published: (2026)
by: Rahman, Md Jahedur, et al.
Published: (2026)
Enhancing Jailbreak Attacks on LLMs via Persona Prompts
by: Zhang, Zheng, et al.
Published: (2025)
by: Zhang, Zheng, et al.
Published: (2025)
CourtGuard: A Local, Multiagent Prompt Injection Classifier
by: Wu, Isaac, et al.
Published: (2025)
by: Wu, Isaac, et al.
Published: (2025)
Breaking the Protocol: Security Analysis of the Model Context Protocol Specification and Prompt Injection Vulnerabilities in Tool-Integrated LLM Agents
by: Maloyan, Narek, et al.
Published: (2026)
by: Maloyan, Narek, et al.
Published: (2026)
SecInfer: Preventing Prompt Injection via Inference-time Scaling
by: Liu, Yupei, et al.
Published: (2025)
by: Liu, Yupei, et al.
Published: (2025)
Injection, Attack and Erasure: Revocable Backdoor Attacks via Machine Unlearning
by: Song, Baogang, et al.
Published: (2025)
by: Song, Baogang, et al.
Published: (2025)
Similar Items
-
Balancing Privacy and Efficiency: Music Information Retrieval via Additive Homomorphic Encryption
by: Wang, William Zerong, et al.
Published: (2025) -
The Phish, The Spam, and The Valid: Generating Feature-Rich Emails for Benchmarking LLMs
by: Toth, Rebeka, et al.
Published: (2025) -
Please Don't Kill My Vibe: Empowering Agents with Data Flow Control
by: Summers, Charlie, et al.
Published: (2025) -
SAMEP: A Secure Protocol for Persistent Context Sharing Across AI Agents
by: Masoor, Hari
Published: (2025) -
The Impact of Event Data Partitioning on Privacy-aware Process Discovery
by: Lim, Jungeun, et al.
Published: (2025)