The Pitfalls of Memorization: When Memorization Hurts Generalization
Fuente:
arXiv
Saved in:
| Main Authors: | Bayat, Reza, Pezeshki, Mohammad, Dohmatob, Elvis, Lopez-Paz, David, Vincent, Pascal |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Steering Large Language Model Activations in Sparse Spaces
by: Bayat, Reza, et al.
Published: (2025)
by: Bayat, Reza, et al.
Published: (2025)
Discovering environments with XRM
by: Pezeshki, Mohammad, et al.
Published: (2023)
by: Pezeshki, Mohammad, et al.
Published: (2023)
Memorization Sinks: Isolating Memorization during LLM Training
by: Ghosal, Gaurav R., et al.
Published: (2025)
by: Ghosal, Gaurav R., et al.
Published: (2025)
Why Less is More (Sometimes): A Theory of Data Curation
by: Dohmatob, Elvis, et al.
Published: (2025)
by: Dohmatob, Elvis, et al.
Published: (2025)
Model Collapse Demystified: The Case of Regression
by: Dohmatob, Elvis, et al.
Published: (2024)
by: Dohmatob, Elvis, et al.
Published: (2024)
Efficient Refusal Ablation in LLM through Optimal Transport
by: Nanfack, Geraldin, et al.
Published: (2026)
by: Nanfack, Geraldin, et al.
Published: (2026)
When Tables Leak: Attacking String Memorization in LLM-Based Tabular Data Generation
by: Ward, Joshua, et al.
Published: (2025)
by: Ward, Joshua, et al.
Published: (2025)
Critical Windows of Complexity Control: When Transformers Decide to Reason or Memorize
by: Ali, Sarwan
Published: (2026)
by: Ali, Sarwan
Published: (2026)
Improving the Scaling Laws of Synthetic Data with Deliberate Practice
by: Askari-Hemmat, Reyhane, et al.
Published: (2025)
by: Askari-Hemmat, Reyhane, et al.
Published: (2025)
Generalizability of Memorization Neural Networks
by: Yu, Lijia, et al.
Published: (2024)
by: Yu, Lijia, et al.
Published: (2024)
Decoding Generalization from Memorization in Deep Neural Networks
by: Ketha, Simran, et al.
Published: (2025)
by: Ketha, Simran, et al.
Published: (2025)
Memorization in In-Context Learning
by: Golchin, Shahriar, et al.
Published: (2024)
by: Golchin, Shahriar, et al.
Published: (2024)
Memorization in deep learning: A survey
by: Wei, Jiaheng, et al.
Published: (2024)
by: Wei, Jiaheng, et al.
Published: (2024)
Membership and Memorization in LLM Knowledge Distillation
by: Zhang, Ziqi, et al.
Published: (2025)
by: Zhang, Ziqi, et al.
Published: (2025)
On the Memorization of Consistency Distillation for Diffusion Models
by: Jiang, Bingqing, et al.
Published: (2026)
by: Jiang, Bingqing, et al.
Published: (2026)
Re-understanding Graph Unlearning through Memorization
by: Ding, Pengfei, et al.
Published: (2026)
by: Ding, Pengfei, et al.
Published: (2026)
Batch Normalization Amplifies Memorization and Privacy Risks
by: Doan, Ngoc Phu, et al.
Published: (2026)
by: Doan, Ngoc Phu, et al.
Published: (2026)
Mitigating Memorization In Language Models
by: Sakarvadia, Mansi, et al.
Published: (2024)
by: Sakarvadia, Mansi, et al.
Published: (2024)
Beyond Model Collapse: Scaling Up with Synthesized Data Requires Verification
by: Feng, Yunzhen, et al.
Published: (2024)
by: Feng, Yunzhen, et al.
Published: (2024)
Understanding and Mitigating Memorization in Generative Models via Sharpness of Probability Landscapes
by: Jeon, Dongjae, et al.
Published: (2024)
by: Jeon, Dongjae, et al.
Published: (2024)
Memorize Theorems, Not Instances: Probing SFT Generalization through Mathematical Reasoning
by: Peng, Ruiying, et al.
Published: (2026)
by: Peng, Ruiying, et al.
Published: (2026)
Analysis of the Memorization and Generalization Capabilities of AI Agents: Are Continual Learners Robust?
by: Kim, Minsu, et al.
Published: (2023)
by: Kim, Minsu, et al.
Published: (2023)
Memorization-Compression Cycles Improve Generalization
by: Yu, Fangyuan
Published: (2025)
by: Yu, Fangyuan
Published: (2025)
Compositional Risk Minimization
by: Mahajan, Divyat, et al.
Published: (2024)
by: Mahajan, Divyat, et al.
Published: (2024)
On Memorization in Diffusion Models
by: Gu, Xiangming, et al.
Published: (2023)
by: Gu, Xiangming, et al.
Published: (2023)
TNT: Improving Chunkwise Training for Test-Time Memorization
by: Li, Zeman, et al.
Published: (2025)
by: Li, Zeman, et al.
Published: (2025)
How Do Flow Matching Models Memorize and Generalize in Sample Data Subspaces?
by: Gao, Weiguo, et al.
Published: (2024)
by: Gao, Weiguo, et al.
Published: (2024)
RePCS: Diagnosing Data Memorization in LLM-Powered Retrieval-Augmented Generation
by: Anh, Le Vu, et al.
Published: (2025)
by: Anh, Le Vu, et al.
Published: (2025)
Titans: Learning to Memorize at Test Time
by: Behrouz, Ali, et al.
Published: (2024)
by: Behrouz, Ali, et al.
Published: (2024)
Detecting Memorization in Large Language Models
by: Slonski, Eduardo
Published: (2024)
by: Slonski, Eduardo
Published: (2024)
Memorization Dynamics of Fill-in-the-Middle Pretraining
by: von Arx, Tobias, et al.
Published: (2026)
by: von Arx, Tobias, et al.
Published: (2026)
Memorization: A Close Look at Books
by: Ma, Iris, et al.
Published: (2025)
by: Ma, Iris, et al.
Published: (2025)
Memorization Control in Diffusion Models from Denoising-centric Perspective
by: Vu, Thuy Phuong, et al.
Published: (2026)
by: Vu, Thuy Phuong, et al.
Published: (2026)
Beyond Frequency: The Role of Redundancy in Large Language Model Memorization
by: Zhang, Jie, et al.
Published: (2025)
by: Zhang, Jie, et al.
Published: (2025)
From Teacher to Student: Tracking Memorization Through Model Distillation
by: Singh, Simardeep
Published: (2025)
by: Singh, Simardeep
Published: (2025)
Position: Privacy Is Not Just Memorization!
by: Mireshghallah, Niloofar, et al.
Published: (2025)
by: Mireshghallah, Niloofar, et al.
Published: (2025)
Extracting Memorized Training Data via Decomposition
by: Su, Ellen, et al.
Published: (2024)
by: Su, Ellen, et al.
Published: (2024)
Unveiling Privacy, Memorization, and Input Curvature Links
by: Ravikumar, Deepak, et al.
Published: (2024)
by: Ravikumar, Deepak, et al.
Published: (2024)
Are Large Language Models Memorizing Bug Benchmarks?
by: Ramos, Daniel, et al.
Published: (2024)
by: Ramos, Daniel, et al.
Published: (2024)
Exploring Memorization in Fine-tuned Language Models
by: Zeng, Shenglai, et al.
Published: (2023)
by: Zeng, Shenglai, et al.
Published: (2023)
Similar Items
-
Steering Large Language Model Activations in Sparse Spaces
by: Bayat, Reza, et al.
Published: (2025) -
Discovering environments with XRM
by: Pezeshki, Mohammad, et al.
Published: (2023) -
Memorization Sinks: Isolating Memorization during LLM Training
by: Ghosal, Gaurav R., et al.
Published: (2025) -
Why Less is More (Sometimes): A Theory of Data Curation
by: Dohmatob, Elvis, et al.
Published: (2025) -
Model Collapse Demystified: The Case of Regression
by: Dohmatob, Elvis, et al.
Published: (2024)