You Told Me to Do It: Measuring Instructional Text-induced Private Data Leakage in LLM Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Kao, Ching-Yu, Li, Xinfeng, Dai, Shenyu, Qiu, Tianze, Zhou, Pengcheng, Jiang, Eric Hanchen, Sperl, Philip |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
"Are You Sure?": An Empirical Study of Human Perception Vulnerability in LLM-Driven Agentic Systems
by: Li, Xinfeng, et al.
Published: (2026)
by: Li, Xinfeng, et al.
Published: (2026)
Security-by-Design for LLM-Based Code Generation: Leveraging Internal Representations for Concept-Driven Steering Mechanisms
by: Wendlinger, Maximilian, et al.
Published: (2026)
by: Wendlinger, Maximilian, et al.
Published: (2026)
Secret-Protected Evolution for Differentially Private Synthetic Text Generation
by: Wang, Tianze, et al.
Published: (2025)
by: Wang, Tianze, et al.
Published: (2025)
Observable Channels, Not Just Storage: Evaluating Privacy Leakage in LLM Agent Pipelines
by: Huang, Tao, et al.
Published: (2026)
by: Huang, Tao, et al.
Published: (2026)
RTBAS: Defending LLM Agents Against Prompt Injection and Privacy Leakage
by: Zhong, Peter Yong, et al.
Published: (2025)
by: Zhong, Peter Yong, et al.
Published: (2025)
Causality Laundering: Denial-Feedback Leakage in Tool-Calling LLM Agents
by: Chinaei, Mohammad Hossein
Published: (2026)
by: Chinaei, Mohammad Hossein
Published: (2026)
Select Me! When You Need a Tool: A Black-box Text Attack on Tool Selection
by: Chen, Liuji, et al.
Published: (2025)
by: Chen, Liuji, et al.
Published: (2025)
Private Data Leakage in Federated Human Activity Recognition for Wearable Healthcare Devices
by: Chen, Kongyang, et al.
Published: (2024)
by: Chen, Kongyang, et al.
Published: (2024)
Topology Matters: Measuring Memory Leakage in Multi-Agent LLMs
by: Liu, Jinbo, et al.
Published: (2025)
by: Liu, Jinbo, et al.
Published: (2025)
ContextLeak: Auditing Leakage in Private In-Context Learning Methods
by: Choi, Jacob, et al.
Published: (2025)
by: Choi, Jacob, et al.
Published: (2025)
LeakAgent: RL-based Red-teaming Agent for LLM Privacy Leakage
by: Nie, Yuzhou, et al.
Published: (2024)
by: Nie, Yuzhou, et al.
Published: (2024)
Credential Leakage in LLM Agent Skills: A Large-Scale Empirical Study
by: Chen, Zhihao, et al.
Published: (2026)
by: Chen, Zhihao, et al.
Published: (2026)
Breaking Euston: Recovering Private Inputs from Secure Inference by Exploiting Subspace Leakage
by: Zhao, Jiaqi, et al.
Published: (2026)
by: Zhao, Jiaqi, et al.
Published: (2026)
Do Androids Dream of Breaking the Game? Systematically Auditing AI Agent Benchmarks with BenchJack
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
Rethinking Side-Channel Analysis: Automated Discovery and Analysis of Side-Channel Leakage with LLM-Assisted Agents
by: Xu, Zhen, et al.
Published: (2026)
by: Xu, Zhen, et al.
Published: (2026)
IP Leakage Attacks Targeting LLM-Based Multi-Agent Systems
by: Wang, Liwen, et al.
Published: (2025)
by: Wang, Liwen, et al.
Published: (2025)
Privacy Leakage via Output Label Space and Differentially Private Continual Learning
by: Tobaben, Marlon, et al.
Published: (2024)
by: Tobaben, Marlon, et al.
Published: (2024)
Batch Me If You Can: Coverage-guided RPKI Fuzzing at Scale
by: Schulmann, Haya, et al.
Published: (2026)
by: Schulmann, Haya, et al.
Published: (2026)
Double-Adversarial Activation Anomaly Detection: Adversarial Autoencoders are Anomaly Generators
by: Schulze, J. -P., et al.
Published: (2021)
by: Schulze, J. -P., et al.
Published: (2021)
The Man Behind the Sound: Demystifying Audio Private Attribute Profiling via Multimodal Large Language Model Agents
by: Wang, Lixu, et al.
Published: (2025)
by: Wang, Lixu, et al.
Published: (2025)
Earn While You Reveal: Private Set Intersection that Rewards Participants
by: Abadi, Aydin
Published: (2023)
by: Abadi, Aydin
Published: (2023)
The Landscape of Prompt Injection Threats in LLM Agents: From Taxonomy to Analysis
by: Wang, Peiran, et al.
Published: (2026)
by: Wang, Peiran, et al.
Published: (2026)
Can You Keep a Secret? Involuntary Information Leakage in Language Model Writing
by: Holtzman, Ari, et al.
Published: (2026)
by: Holtzman, Ari, et al.
Published: (2026)
Thwart Me If You Can: An Empirical Analysis of Android Platform Armoring Against Stalkerware
by: Jadhav, Malvika, et al.
Published: (2025)
by: Jadhav, Malvika, et al.
Published: (2025)
DP-MGTD: Privacy-Preserving Machine-Generated Text Detection via Adaptive Differentially Private Entity Sanitization
by: Wang, Lionel Z., et al.
Published: (2026)
by: Wang, Lionel Z., et al.
Published: (2026)
Minimal Cascade Gradient Smoothing for Fast Transferable Preemptive Adversarial Defense
by: Wang, Hanrui, et al.
Published: (2024)
by: Wang, Hanrui, et al.
Published: (2024)
EdgeLeakage: Membership Information Leakage in Distributed Edge Intelligence Systems
by: Chen, Kongyang, et al.
Published: (2024)
by: Chen, Kongyang, et al.
Published: (2024)
CircuitGuard: Mitigating LLM Memorization in RTL Code Generation Against IP Leakage
by: Mashnoor, Nowfel, et al.
Published: (2025)
by: Mashnoor, Nowfel, et al.
Published: (2025)
Image-to-Text Logic Jailbreak: Your Imagination can Help You Do Anything
by: Zou, Xiaotian, et al.
Published: (2024)
by: Zou, Xiaotian, et al.
Published: (2024)
Network-Level Prompt and Trait Leakage in Local Research Agents
by: Jeong, Hyejun, et al.
Published: (2025)
by: Jeong, Hyejun, et al.
Published: (2025)
Real-Time Privacy Risk Measurement with Privacy Tokens for Gradient Leakage
by: Meng, Jiayang, et al.
Published: (2025)
by: Meng, Jiayang, et al.
Published: (2025)
Evaluating Privacy Leakage in Split Learning
by: Qiu, Xinchi, et al.
Published: (2023)
by: Qiu, Xinchi, et al.
Published: (2023)
A Vision for Access Control in LLM-based Agent Systems
by: Li, Xinfeng, et al.
Published: (2025)
by: Li, Xinfeng, et al.
Published: (2025)
Enhancing Leakage Attacks on Searchable Symmetric Encryption Using LLM-Based Synthetic Data Generation
by: Chiu, Joshua, et al.
Published: (2025)
by: Chiu, Joshua, et al.
Published: (2025)
You Can't Steal Nothing: Mitigating Prompt Leakages in LLMs via System Vectors
by: Cao, Bochuan, et al.
Published: (2025)
by: Cao, Bochuan, et al.
Published: (2025)
S-Leak: Leakage-Abuse Attack Against Efficient Conjunctive SSE via s-term Leakage
by: Su, Yue, et al.
Published: (2025)
by: Su, Yue, et al.
Published: (2025)
CLIOPATRA: Extracting Private Information from LLM Insights
by: Annamalai, Meenatchi Sundaram Muthu Selva, et al.
Published: (2026)
by: Annamalai, Meenatchi Sundaram Muthu Selva, et al.
Published: (2026)
Your Agent Is Mine: Measuring Malicious Intermediary Attacks on the LLM Supply Chain
by: Liu, Hanzhi, et al.
Published: (2026)
by: Liu, Hanzhi, et al.
Published: (2026)
Gotcha! I Know What You are Doing on the FPGA Cloud: Fingerprinting Co-Located Cloud FPGA Accelerators via Measuring Communication Links
by: Fang, Chongzhou, et al.
Published: (2023)
by: Fang, Chongzhou, et al.
Published: (2023)
I Know What You Bought Last Summer: Investigating User Data Leakage in E-Commerce Platforms
by: Vlachogiannakis, Ioannis, et al.
Published: (2025)
by: Vlachogiannakis, Ioannis, et al.
Published: (2025)
Similar Items
-
"Are You Sure?": An Empirical Study of Human Perception Vulnerability in LLM-Driven Agentic Systems
by: Li, Xinfeng, et al.
Published: (2026) -
Security-by-Design for LLM-Based Code Generation: Leveraging Internal Representations for Concept-Driven Steering Mechanisms
by: Wendlinger, Maximilian, et al.
Published: (2026) -
Secret-Protected Evolution for Differentially Private Synthetic Text Generation
by: Wang, Tianze, et al.
Published: (2025) -
Observable Channels, Not Just Storage: Evaluating Privacy Leakage in LLM Agent Pipelines
by: Huang, Tao, et al.
Published: (2026) -
RTBAS: Defending LLM Agents Against Prompt Injection and Privacy Leakage
by: Zhong, Peter Yong, et al.
Published: (2025)