Recite, Reconstruct, Recollect: Memorization in LMs as a Multifaceted Phenomenon
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Prashanth, USVSN Sai, Deng, Alvin, O'Brien, Kyle, S V, Jyothir, Khan, Mohammad Aflah, Borkar, Jaydeep, Choquette-Choo, Christopher A., Fuehne, Jacob Ray, Biderman, Stella, Ke, Tracy, Lee, Katherine, Saphra, Naomi |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Privacy Ripple Effects from Adding or Removing Personal Information in Language Model Training
par: Borkar, Jaydeep, et autres
Publié: (2025)
par: Borkar, Jaydeep, et autres
Publié: (2025)
Mind the Gap: Analyzing Lacunae with Transformer-Based Transcription
par: Borkar, Jaydeep, et autres
Publié: (2024)
par: Borkar, Jaydeep, et autres
Publié: (2024)
PolyPythias: Stability and Outliers across Fifty Language Model Pre-Training Runs
par: van der Wal, Oskar, et autres
Publié: (2025)
par: van der Wal, Oskar, et autres
Publié: (2025)
Mechanistic?
par: Saphra, Naomi, et autres
Publié: (2024)
par: Saphra, Naomi, et autres
Publié: (2024)
Memorization Dynamics in Knowledge Distillation for Language Models
par: Borkar, Jaydeep, et autres
Publié: (2026)
par: Borkar, Jaydeep, et autres
Publié: (2026)
Hidden Breakthroughs in Language Model Training
par: Kangaslahti, Sara, et autres
Publié: (2025)
par: Kangaslahti, Sara, et autres
Publié: (2025)
ChatGPT Doesn't Trust Chargers Fans: Guardrail Sensitivity in Context
par: Li, Victoria R., et autres
Publié: (2024)
par: Li, Victoria R., et autres
Publié: (2024)
Sometimes I am a Tree: Data Drives Unstable Hierarchical Generalization
par: Qin, Tian, et autres
Publié: (2024)
par: Qin, Tian, et autres
Publié: (2024)
MatheMagic: Generating Dynamic Mathematics Benchmarks Robust to Memorization
par: O'Brien, Dayyán, et autres
Publié: (2025)
par: O'Brien, Dayyán, et autres
Publié: (2025)
Recital Review
Publié: (2022)
Publié: (2022)
Deep Ignorance: Filtering Pretraining Data Builds Tamper-Resistant Safeguards into Open-Weight LLMs
par: O'Brien, Kyle, et autres
Publié: (2025)
par: O'Brien, Kyle, et autres
Publié: (2025)
Rethinking Memorization Measures and their Implications in Large Language Models
par: Ghosh, Bishwamittra, et autres
Publié: (2025)
par: Ghosh, Bishwamittra, et autres
Publié: (2025)
Bidirectional LMs are Better Knowledge Memorizers? A Benchmark for Real-world Knowledge Injection
par: Zhang, Yuwei, et autres
Publié: (2025)
par: Zhang, Yuwei, et autres
Publié: (2025)
A Taxonomy of Transcendence
par: Abreu, Natalie, et autres
Publié: (2025)
par: Abreu, Natalie, et autres
Publié: (2025)
First Tragedy, then Parse: History Repeats Itself in the New Era of Large Language Models
par: Saphra, Naomi, et autres
Publié: (2023)
par: Saphra, Naomi, et autres
Publié: (2023)
TRAM: Bridging Trust Regions and Sharpness Aware Minimization
par: Sherborne, Tom, et autres
Publié: (2023)
par: Sherborne, Tom, et autres
Publié: (2023)
Fast Forwarding Low-Rank Training
par: Rahamim, Adir, et autres
Publié: (2024)
par: Rahamim, Adir, et autres
Publié: (2024)
Linguistic Generalizations are not Rules: Impacts on Evaluation of LMs
par: Weissweiler, Leonie, et autres
Publié: (2025)
par: Weissweiler, Leonie, et autres
Publié: (2025)
Optimal Rates for $O(1)$-Smooth DP-SCO with a Single Epoch and Large Batches
par: Choquette-Choo, Christopher A., et autres
Publié: (2024)
par: Choquette-Choo, Christopher A., et autres
Publié: (2024)
Recollecting Resonances
Publié: (2020)
Publié: (2020)
Rote Learning Considered Useful: Generalizing over Memorized Data in LLMs
par: Wu, Qinyuan, et autres
Publié: (2025)
par: Wu, Qinyuan, et autres
Publié: (2025)
The Censorship Phenomenon in College and Research Libraries: An Investigation of the Canadian Prairie Provinces, 1980-1985.
par: Schrader, Alvin M., et autres
Publié: (1989)
par: Schrader, Alvin M., et autres
Publié: (1989)
Causal Drawbridges: Characterizing Gradient Blocking of Syntactic Islands in Transformer LMs
par: Boguraev, Sasha, et autres
Publié: (2026)
par: Boguraev, Sasha, et autres
Publié: (2026)
Latent State Models of Training Dynamics
par: Hu, Michael Y., et autres
Publié: (2023)
par: Hu, Michael Y., et autres
Publié: (2023)
Na ante-sala da discriminação: o preço dos atributos de sexo ecor no Brasil (19891999)
par: Ciro Biderman
Publié: (2004)
par: Ciro Biderman
Publié: (2004)
What Matters in Memorizing and Recalling Facts? Multifaceted Benchmarks for Knowledge Probing in Language Models
par: Zhao, Xin, et autres
Publié: (2024)
par: Zhao, Xin, et autres
Publié: (2024)
Privacy Amplification for Matrix Mechanisms
par: Choquette-Choo, Christopher A., et autres
Publié: (2023)
par: Choquette-Choo, Christopher A., et autres
Publié: (2023)
Recollections of a Nuclear War
par: Morrison, Philip
Publié: (1945)
par: Morrison, Philip
Publié: (1945)
Recollections of a Nuclear War
Publié: (1995)
Publié: (1995)
A suite of LMs comprehend puzzle statements as well as humans
par: Goldberg, Adele E, et autres
Publié: (2025)
par: Goldberg, Adele E, et autres
Publié: (2025)
LLM Circuit Analyses Are Consistent Across Training and Scale
par: Tigges, Curt, et autres
Publié: (2024)
par: Tigges, Curt, et autres
Publié: (2024)
Grokking Group Multiplication with Cosets
par: Stander, Dashiell, et autres
Publié: (2023)
par: Stander, Dashiell, et autres
Publié: (2023)
The Ghost in the Keys: A Disklavier Demo for Human-AI Musical Co-Creativity
par: Bradshaw, Louis, et autres
Publié: (2025)
par: Bradshaw, Louis, et autres
Publié: (2025)
Recite Your Ask Out Loud
Publié: (2025)
Publié: (2025)
Hubble: a Model Suite to Advance the Study of LLM Memorization
par: Wei, Johnny Tian-Zheng, et autres
Publié: (2025)
par: Wei, Johnny Tian-Zheng, et autres
Publié: (2025)
Benchmarks as Microscopes: A Call for Model Metrology
par: Saxon, Michael, et autres
Publié: (2024)
par: Saxon, Michael, et autres
Publié: (2024)
Random Scaling of Emergent Capabilities
par: Zhao, Rosie, et autres
Publié: (2025)
par: Zhao, Rosie, et autres
Publié: (2025)
Dynamic Masking Rate Schedules for MLM Pretraining
par: Ankner, Zachary, et autres
Publié: (2023)
par: Ankner, Zachary, et autres
Publié: (2023)
Auditing Private Prediction
par: Chadha, Karan, et autres
Publié: (2024)
par: Chadha, Karan, et autres
Publié: (2024)
Stochastic Approximation with Two Time Scales: The General Case
par: Borkar, Vivek S
Publié: (2024)
par: Borkar, Vivek S
Publié: (2024)
Documents similaires
-
Privacy Ripple Effects from Adding or Removing Personal Information in Language Model Training
par: Borkar, Jaydeep, et autres
Publié: (2025) -
Mind the Gap: Analyzing Lacunae with Transformer-Based Transcription
par: Borkar, Jaydeep, et autres
Publié: (2024) -
PolyPythias: Stability and Outliers across Fifty Language Model Pre-Training Runs
par: van der Wal, Oskar, et autres
Publié: (2025) -
Mechanistic?
par: Saphra, Naomi, et autres
Publié: (2024) -
Memorization Dynamics in Knowledge Distillation for Language Models
par: Borkar, Jaydeep, et autres
Publié: (2026)