Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory
Fuente:
arXiv
Saved in:
| Main Authors: | Mireshghallah, Niloofar, Kim, Hyunwoo, Zhou, Xuhui, Tsvetkov, Yulia, Sap, Maarten, Shokri, Reza, Choi, Yejin |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A False Sense of Privacy: Evaluating Textual Data Sanitization Beyond Surface-level Privacy Leakage
by: Xin, Rui, et al.
Published: (2025)
by: Xin, Rui, et al.
Published: (2025)
CIMemories: A Compositional Benchmark for Contextual Integrity of Persistent Memory in LLMs
by: Mireshghallah, Niloofar, et al.
Published: (2025)
by: Mireshghallah, Niloofar, et al.
Published: (2025)
PPMI: Privacy-Preserving LLM Interaction with Socratic Chain-of-Thought Reasoning and Homomorphically Encrypted Vector Databases
by: Bae, Yubeen, et al.
Published: (2025)
by: Bae, Yubeen, et al.
Published: (2025)
Position: Privacy Is Not Just Memorization!
by: Mireshghallah, Niloofar, et al.
Published: (2025)
by: Mireshghallah, Niloofar, et al.
Published: (2025)
Synthetic Data Can Mislead Evaluations: Membership Inference as Machine Text Detection
by: Naseh, Ali, et al.
Published: (2025)
by: Naseh, Ali, et al.
Published: (2025)
Privacy Checklist: Privacy Violation Detection Grounding on Contextual Integrity Theory
by: Li, Haoran, et al.
Published: (2024)
by: Li, Haoran, et al.
Published: (2024)
One (Thread) Can Keep a (PRNG) Secret, but not Two
by: Porat, Ehood, et al.
Published: (2026)
by: Porat, Ehood, et al.
Published: (2026)
Integrating Differential Privacy and Contextual Integrity
by: Benthall, Sebastian, et al.
Published: (2024)
by: Benthall, Sebastian, et al.
Published: (2024)
Can Large Language Models Really Recognize Your Name?
by: Pham, Dzung, et al.
Published: (2025)
by: Pham, Dzung, et al.
Published: (2025)
GoldCoin: Grounding Large Language Models in Privacy Laws via Contextual Integrity Theory
by: Fan, Wei, et al.
Published: (2024)
by: Fan, Wei, et al.
Published: (2024)
It Takes Two: Complementary Self-Distillation for Contextual Integrity in LLMs
by: Park, Sangwoo, et al.
Published: (2026)
by: Park, Sangwoo, et al.
Published: (2026)
Can AI Keep a Secret? Contextual Integrity Verification: A Provable Security Architecture for LLMs
by: Gupta, Aayush
Published: (2025)
by: Gupta, Aayush
Published: (2025)
Privacy Ripple Effects from Adding or Removing Personal Information in Language Model Training
by: Borkar, Jaydeep, et al.
Published: (2025)
by: Borkar, Jaydeep, et al.
Published: (2025)
Can You Keep a Secret? Involuntary Information Leakage in Language Model Writing
by: Holtzman, Ari, et al.
Published: (2026)
by: Holtzman, Ari, et al.
Published: (2026)
I Can Tell Your Secrets: Inferring Privacy Attributes from Mini-app Interaction History in Super-apps
by: Cai, Yifeng, et al.
Published: (2025)
by: Cai, Yifeng, et al.
Published: (2025)
Differentially Private Learning Needs Better Model Initialization and Self-Distillation
by: Ngong, Ivoline C., et al.
Published: (2024)
by: Ngong, Ivoline C., et al.
Published: (2024)
ReFuzz: Reusing Tests for Processor Fuzzing with Contextual Bandits
by: Chen, Chen, et al.
Published: (2025)
by: Chen, Chen, et al.
Published: (2025)
Privacy-Preserving Edge Federated Learning for Intelligent Mobile-Health Systems
by: Aminifar, Amin, et al.
Published: (2024)
by: Aminifar, Amin, et al.
Published: (2024)
Web Privacy based on Contextual Integrity: Measuring the Collapse of Online Contexts
by: Sivan-Sevilla, Ido, et al.
Published: (2024)
by: Sivan-Sevilla, Ido, et al.
Published: (2024)
Low-Cost High-Power Membership Inference Attacks
by: Zarifzadeh, Sajjad, et al.
Published: (2023)
by: Zarifzadeh, Sajjad, et al.
Published: (2023)
Holding Secrets Accountable: Auditing Privacy-Preserving Machine Learning
by: Lycklama, Hidde, et al.
Published: (2024)
by: Lycklama, Hidde, et al.
Published: (2024)
Privacy Without Losing Place: A Paradigm for Private Retrieval in Spatial RAGs
by: Edemacu, Kennedy, et al.
Published: (2026)
by: Edemacu, Kennedy, et al.
Published: (2026)
TrojanStego: Your Language Model Can Secretly Be A Steganographic Privacy Leaking Agent
by: Meier, Dominik, et al.
Published: (2025)
by: Meier, Dominik, et al.
Published: (2025)
The Canary's Echo: Auditing Privacy Risks of LLM-Generated Synthetic Text
by: Meeus, Matthieu, et al.
Published: (2025)
by: Meeus, Matthieu, et al.
Published: (2025)
Continual Pretraining on Encrypted Synthetic Data for Privacy-Preserving LLMs
by: Liu, Honghao, et al.
Published: (2026)
by: Liu, Honghao, et al.
Published: (2026)
Guarding Multiple Secrets: Enhanced Summary Statistic Privacy for Data Sharing
by: Wang, Shuaiqi, et al.
Published: (2024)
by: Wang, Shuaiqi, et al.
Published: (2024)
Privacy-Preserving LLMs Routing
by: Wu, Xidong, et al.
Published: (2026)
by: Wu, Xidong, et al.
Published: (2026)
Mind The Gap: Can Air-Gaps Keep Your Private Data Secure?
by: Guri, Mordechai
Published: (2024)
by: Guri, Mordechai
Published: (2024)
A Blockchain-Enhanced Framework for Privacy and Data Integrity in Crowdsourced Drone Services
by: Akram, Junaid, et al.
Published: (2024)
by: Akram, Junaid, et al.
Published: (2024)
$π$QLB: A Privacy-preserving with Integrity-assuring Query Language for Blockchain
by: Sohrabi, Nasrin, et al.
Published: (2022)
by: Sohrabi, Nasrin, et al.
Published: (2022)
CI-Work: Benchmarking Contextual Integrity in Enterprise LLM Agents
by: Fu, Wenjie, et al.
Published: (2026)
by: Fu, Wenjie, et al.
Published: (2026)
Covert Surveillance in Smart Devices: A SCOUR Framework Analysis of Youth Privacy Implications
by: Shouli, Austin, et al.
Published: (2025)
by: Shouli, Austin, et al.
Published: (2025)
Privacy Bias in Language Models: A Contextual Integrity-based Auditing Metric
by: Shvartzshnaider, Yan, et al.
Published: (2024)
by: Shvartzshnaider, Yan, et al.
Published: (2024)
When Machine Unlearning Meets Retrieval-Augmented Generation (RAG): Keep Secret or Forget Knowledge?
by: Wang, Shang, et al.
Published: (2024)
by: Wang, Shang, et al.
Published: (2024)
Do Vision-Language Models Respect Contextual Integrity in Location Disclosure?
by: Yang, Ruixin, et al.
Published: (2026)
by: Yang, Ruixin, et al.
Published: (2026)
Multiverse Privacy Theory for Contextual Risks in Complex User-AI Interactions
by: Gumusel, Ece
Published: (2025)
by: Gumusel, Ece
Published: (2025)
Protecting Vehicle Location Privacy with Contextually-Driven Synthetic Location Generation
by: Yadav, Sourabh, et al.
Published: (2024)
by: Yadav, Sourabh, et al.
Published: (2024)
Alpaca against Vicuna: Using LLMs to Uncover Memorization of LLMs
by: Kassem, Aly M., et al.
Published: (2024)
by: Kassem, Aly M., et al.
Published: (2024)
Blind-Touch: Homomorphic Encryption-Based Distributed Neural Network Inference for Privacy-Preserving Fingerprint Authentication
by: Choi, Hyunmin, et al.
Published: (2023)
by: Choi, Hyunmin, et al.
Published: (2023)
RedacBench: Can AI Erase Your Secrets?
by: Jeon, Hyunjun, et al.
Published: (2026)
by: Jeon, Hyunjun, et al.
Published: (2026)
Similar Items
-
A False Sense of Privacy: Evaluating Textual Data Sanitization Beyond Surface-level Privacy Leakage
by: Xin, Rui, et al.
Published: (2025) -
CIMemories: A Compositional Benchmark for Contextual Integrity of Persistent Memory in LLMs
by: Mireshghallah, Niloofar, et al.
Published: (2025) -
PPMI: Privacy-Preserving LLM Interaction with Socratic Chain-of-Thought Reasoning and Homomorphically Encrypted Vector Databases
by: Bae, Yubeen, et al.
Published: (2025) -
Position: Privacy Is Not Just Memorization!
by: Mireshghallah, Niloofar, et al.
Published: (2025) -
Synthetic Data Can Mislead Evaluations: Membership Inference as Machine Text Detection
by: Naseh, Ali, et al.
Published: (2025)