Do LLMs Really Memorize Personally Identifiable Information? Revisiting PII Leakage with a Cue-Controlled Memorization Framework
Fuente:
arXiv
Saved in:
| Main Authors: | Luo, Xiaoyu, Chen, Yiyi, Li, Qiongxiu, Bjerva, Johannes |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Shared Path: Unraveling Memorization in Multilingual LLMs through Language Similarities
by: Luo, Xiaoyu, et al.
Published: (2025)
by: Luo, Xiaoyu, et al.
Published: (2025)
Characterizing Memorization in Diffusion Language Models: Generalized Extraction and Sampling Effects
by: Luo, Xiaoyu, et al.
Published: (2026)
by: Luo, Xiaoyu, et al.
Published: (2026)
LAGO: Few-shot Crosslingual Embedding Inversion Attacks via Language Similarity-Aware Graph Optimization
by: Yu, Wenrui, et al.
Published: (2025)
by: Yu, Wenrui, et al.
Published: (2025)
Large Language Models are Easily Confused: A Quantitative Metric, Security Implications and Typological Analysis
by: Chen, Yiyi, et al.
Published: (2024)
by: Chen, Yiyi, et al.
Published: (2024)
Trustworthy Machine Learning via Memorization and the Granular Long-Tail: A Survey on Interactions, Tradeoffs, and Beyond
by: Li, Qiongxiu, et al.
Published: (2025)
by: Li, Qiongxiu, et al.
Published: (2025)
Semantic Leakage from Image Embeddings
by: Chen, Yiyi, et al.
Published: (2026)
by: Chen, Yiyi, et al.
Published: (2026)
Text Embedding Inversion Security for Multilingual Language Models
by: Chen, Yiyi, et al.
Published: (2024)
by: Chen, Yiyi, et al.
Published: (2024)
Memorization and Knowledge Injection in Gated LLMs
by: Pan, Xu, et al.
Published: (2025)
by: Pan, Xu, et al.
Published: (2025)
Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis
by: Djiré, Albérick Euraste, et al.
Published: (2025)
by: Djiré, Albérick Euraste, et al.
Published: (2025)
GLiNER2-PII: A Multilingual Model for Personally Identifiable Information Extraction
by: Zaratiana, Urchade, et al.
Published: (2026)
by: Zaratiana, Urchade, et al.
Published: (2026)
ALGEN: Few-shot Inversion Attacks on Textual Embeddings using Alignment and Generation
by: Chen, Yiyi, et al.
Published: (2025)
by: Chen, Yiyi, et al.
Published: (2025)
Memorization vs. Reasoning: Updating LLMs with New Knowledge
by: Li, Aochong Oliver, et al.
Published: (2025)
by: Li, Aochong Oliver, et al.
Published: (2025)
Do Audio LLMs Really LISTEN, or Just Transcribe? Measuring Lexical vs. Acoustic Emotion Cues Reliance
by: Chen, Jingyi, et al.
Published: (2025)
by: Chen, Jingyi, et al.
Published: (2025)
Data Compressibility Quantifies LLM Memorization
by: Huang, Yizhan, et al.
Published: (2025)
by: Huang, Yizhan, et al.
Published: (2025)
Memorization in Attention-only Transformers
by: Dana, Léo, et al.
Published: (2024)
by: Dana, Léo, et al.
Published: (2024)
LLMs and Memorization: On Quality and Specificity of Copyright Compliance
by: Mueller, Felix B, et al.
Published: (2024)
by: Mueller, Felix B, et al.
Published: (2024)
AlpsBench: An LLM Personalization Benchmark for Real-Dialogue Memorization and Preference Alignment
by: Xiao, Jianfei, et al.
Published: (2026)
by: Xiao, Jianfei, et al.
Published: (2026)
Context Memorization for Efficient Long Context Generation
by: Okoshi, Yasuyuki, et al.
Published: (2026)
by: Okoshi, Yasuyuki, et al.
Published: (2026)
Impact of Fine-Tuning Methods on Memorization in Large Language Models
by: Hou, Jie, et al.
Published: (2025)
by: Hou, Jie, et al.
Published: (2025)
Memorization in In-Context Learning
by: Golchin, Shahriar, et al.
Published: (2024)
by: Golchin, Shahriar, et al.
Published: (2024)
Protoknowledge Shapes Behaviour of LLMs in Downstream Tasks: Memorization and Generalization with Knowledge Graphs
by: Ranaldi, Federico, et al.
Published: (2025)
by: Ranaldi, Federico, et al.
Published: (2025)
Quantifying In-Context Reasoning Effects and Memorization Effects in LLMs
by: Lou, Siyu, et al.
Published: (2024)
by: Lou, Siyu, et al.
Published: (2024)
ROME: Memorization Insights from Text, Logits and Representation
by: Li, Bo, et al.
Published: (2024)
by: Li, Bo, et al.
Published: (2024)
Mitigating Unintended Memorization with LoRA in Federated Learning for LLMs
by: Bossy, Thierry, et al.
Published: (2025)
by: Bossy, Thierry, et al.
Published: (2025)
NLP Security and Ethics, in the Wild
by: Lent, Heather, et al.
Published: (2025)
by: Lent, Heather, et al.
Published: (2025)
Arithmetic with Language Models: from Memorization to Computation
by: Maltoni, Davide, et al.
Published: (2023)
by: Maltoni, Davide, et al.
Published: (2023)
Memorization in Fine-Tuned Large Language Models
by: Savine, Danil
Published: (2025)
by: Savine, Danil
Published: (2025)
Memorizing Documents with Guidance in Large Language Models
by: Park, Bumjin, et al.
Published: (2024)
by: Park, Bumjin, et al.
Published: (2024)
ATLAS: Learning to Optimally Memorize the Context at Test Time
by: Behrouz, Ali, et al.
Published: (2025)
by: Behrouz, Ali, et al.
Published: (2025)
Mitigating Memorization In Language Models
by: Sakarvadia, Mansi, et al.
Published: (2024)
by: Sakarvadia, Mansi, et al.
Published: (2024)
A Multi-Perspective Analysis of Memorization in Large Language Models
by: Chen, Bowen, et al.
Published: (2024)
by: Chen, Bowen, et al.
Published: (2024)
Memorizing is Not Enough: Deep Knowledge Injection Through Reasoning
by: Xu, Ruoxi, et al.
Published: (2025)
by: Xu, Ruoxi, et al.
Published: (2025)
MatheMagic: Generating Dynamic Mathematics Benchmarks Robust to Memorization
by: O'Brien, Dayyán, et al.
Published: (2025)
by: O'Brien, Dayyán, et al.
Published: (2025)
Position: Privacy Is Not Just Memorization!
by: Mireshghallah, Niloofar, et al.
Published: (2025)
by: Mireshghallah, Niloofar, et al.
Published: (2025)
Diagnosing Memorization in Chain-of-Thought Reasoning, One Token at a Time
by: Li, Huihan, et al.
Published: (2025)
by: Li, Huihan, et al.
Published: (2025)
Private Memorization Editing: Turning Memorization into a Defense to Strengthen Data Privacy in Large Language Models
by: Ruzzetti, Elena Sofia, et al.
Published: (2025)
by: Ruzzetti, Elena Sofia, et al.
Published: (2025)
Retentive or Forgetful? Diving into the Knowledge Memorizing Mechanism of Language Models
by: Cao, Boxi, et al.
Published: (2023)
by: Cao, Boxi, et al.
Published: (2023)
Memorization in Large Language Models in Medicine: Prevalence, Characteristics, and Implications
by: Li, Anran, et al.
Published: (2025)
by: Li, Anran, et al.
Published: (2025)
Undesirable Memorization in Large Language Models: A Survey
by: Satvaty, Ali, et al.
Published: (2024)
by: Satvaty, Ali, et al.
Published: (2024)
Unintended Memorization of Sensitive Information in Fine-Tuned Language Models
by: Szep, Marton, et al.
Published: (2026)
by: Szep, Marton, et al.
Published: (2026)
Similar Items
-
Shared Path: Unraveling Memorization in Multilingual LLMs through Language Similarities
by: Luo, Xiaoyu, et al.
Published: (2025) -
Characterizing Memorization in Diffusion Language Models: Generalized Extraction and Sampling Effects
by: Luo, Xiaoyu, et al.
Published: (2026) -
LAGO: Few-shot Crosslingual Embedding Inversion Attacks via Language Similarity-Aware Graph Optimization
by: Yu, Wenrui, et al.
Published: (2025) -
Large Language Models are Easily Confused: A Quantitative Metric, Security Implications and Typological Analysis
by: Chen, Yiyi, et al.
Published: (2024) -
Trustworthy Machine Learning via Memorization and the Granular Long-Tail: A Survey on Interactions, Tradeoffs, and Beyond
by: Li, Qiongxiu, et al.
Published: (2025)