Recite, Reconstruct, Recollect: Memorization in LMs as a Multifaceted Phenomenon
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Prashanth, USVSN Sai, Deng, Alvin, O'Brien, Kyle, S V, Jyothir, Khan, Mohammad Aflah, Borkar, Jaydeep, Choquette-Choo, Christopher A., Fuehne, Jacob Ray, Biderman, Stella, Ke, Tracy, Lee, Katherine, Saphra, Naomi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Privacy Ripple Effects from Adding or Removing Personal Information in Language Model Training
von: Borkar, Jaydeep, et al.
Veröffentlicht: (2025)
von: Borkar, Jaydeep, et al.
Veröffentlicht: (2025)
Mind the Gap: Analyzing Lacunae with Transformer-Based Transcription
von: Borkar, Jaydeep, et al.
Veröffentlicht: (2024)
von: Borkar, Jaydeep, et al.
Veröffentlicht: (2024)
PolyPythias: Stability and Outliers across Fifty Language Model Pre-Training Runs
von: van der Wal, Oskar, et al.
Veröffentlicht: (2025)
von: van der Wal, Oskar, et al.
Veröffentlicht: (2025)
Mechanistic?
von: Saphra, Naomi, et al.
Veröffentlicht: (2024)
von: Saphra, Naomi, et al.
Veröffentlicht: (2024)
Memorization Dynamics in Knowledge Distillation for Language Models
von: Borkar, Jaydeep, et al.
Veröffentlicht: (2026)
von: Borkar, Jaydeep, et al.
Veröffentlicht: (2026)
Hidden Breakthroughs in Language Model Training
von: Kangaslahti, Sara, et al.
Veröffentlicht: (2025)
von: Kangaslahti, Sara, et al.
Veröffentlicht: (2025)
ChatGPT Doesn't Trust Chargers Fans: Guardrail Sensitivity in Context
von: Li, Victoria R., et al.
Veröffentlicht: (2024)
von: Li, Victoria R., et al.
Veröffentlicht: (2024)
Sometimes I am a Tree: Data Drives Unstable Hierarchical Generalization
von: Qin, Tian, et al.
Veröffentlicht: (2024)
von: Qin, Tian, et al.
Veröffentlicht: (2024)
MatheMagic: Generating Dynamic Mathematics Benchmarks Robust to Memorization
von: O'Brien, Dayyán, et al.
Veröffentlicht: (2025)
von: O'Brien, Dayyán, et al.
Veröffentlicht: (2025)
Deep Ignorance: Filtering Pretraining Data Builds Tamper-Resistant Safeguards into Open-Weight LLMs
von: O'Brien, Kyle, et al.
Veröffentlicht: (2025)
von: O'Brien, Kyle, et al.
Veröffentlicht: (2025)
Recital Review
Veröffentlicht: (2022)
Veröffentlicht: (2022)
Rethinking Memorization Measures and their Implications in Large Language Models
von: Ghosh, Bishwamittra, et al.
Veröffentlicht: (2025)
von: Ghosh, Bishwamittra, et al.
Veröffentlicht: (2025)
Bidirectional LMs are Better Knowledge Memorizers? A Benchmark for Real-world Knowledge Injection
von: Zhang, Yuwei, et al.
Veröffentlicht: (2025)
von: Zhang, Yuwei, et al.
Veröffentlicht: (2025)
Linguistic Generalizations are not Rules: Impacts on Evaluation of LMs
von: Weissweiler, Leonie, et al.
Veröffentlicht: (2025)
von: Weissweiler, Leonie, et al.
Veröffentlicht: (2025)
A Taxonomy of Transcendence
von: Abreu, Natalie, et al.
Veröffentlicht: (2025)
von: Abreu, Natalie, et al.
Veröffentlicht: (2025)
First Tragedy, then Parse: History Repeats Itself in the New Era of Large Language Models
von: Saphra, Naomi, et al.
Veröffentlicht: (2023)
von: Saphra, Naomi, et al.
Veröffentlicht: (2023)
TRAM: Bridging Trust Regions and Sharpness Aware Minimization
von: Sherborne, Tom, et al.
Veröffentlicht: (2023)
von: Sherborne, Tom, et al.
Veröffentlicht: (2023)
Fast Forwarding Low-Rank Training
von: Rahamim, Adir, et al.
Veröffentlicht: (2024)
von: Rahamim, Adir, et al.
Veröffentlicht: (2024)
Optimal Rates for $O(1)$-Smooth DP-SCO with a Single Epoch and Large Batches
von: Choquette-Choo, Christopher A., et al.
Veröffentlicht: (2024)
von: Choquette-Choo, Christopher A., et al.
Veröffentlicht: (2024)
Recollecting Resonances
Veröffentlicht: (2020)
Veröffentlicht: (2020)
Rote Learning Considered Useful: Generalizing over Memorized Data in LLMs
von: Wu, Qinyuan, et al.
Veröffentlicht: (2025)
von: Wu, Qinyuan, et al.
Veröffentlicht: (2025)
The Censorship Phenomenon in College and Research Libraries: An Investigation of the Canadian Prairie Provinces, 1980-1985.
von: Schrader, Alvin M., et al.
Veröffentlicht: (1989)
von: Schrader, Alvin M., et al.
Veröffentlicht: (1989)
Causal Drawbridges: Characterizing Gradient Blocking of Syntactic Islands in Transformer LMs
von: Boguraev, Sasha, et al.
Veröffentlicht: (2026)
von: Boguraev, Sasha, et al.
Veröffentlicht: (2026)
Latent State Models of Training Dynamics
von: Hu, Michael Y., et al.
Veröffentlicht: (2023)
von: Hu, Michael Y., et al.
Veröffentlicht: (2023)
Na ante-sala da discriminação: o preço dos atributos de sexo ecor no Brasil (19891999)
von: Ciro Biderman
Veröffentlicht: (2004)
von: Ciro Biderman
Veröffentlicht: (2004)
What Matters in Memorizing and Recalling Facts? Multifaceted Benchmarks for Knowledge Probing in Language Models
von: Zhao, Xin, et al.
Veröffentlicht: (2024)
von: Zhao, Xin, et al.
Veröffentlicht: (2024)
Privacy Amplification for Matrix Mechanisms
von: Choquette-Choo, Christopher A., et al.
Veröffentlicht: (2023)
von: Choquette-Choo, Christopher A., et al.
Veröffentlicht: (2023)
Recollections of a Nuclear War
von: Morrison, Philip
Veröffentlicht: (1945)
von: Morrison, Philip
Veröffentlicht: (1945)
Recollections of a Nuclear War
Veröffentlicht: (1995)
Veröffentlicht: (1995)
A suite of LMs comprehend puzzle statements as well as humans
von: Goldberg, Adele E, et al.
Veröffentlicht: (2025)
von: Goldberg, Adele E, et al.
Veröffentlicht: (2025)
LLM Circuit Analyses Are Consistent Across Training and Scale
von: Tigges, Curt, et al.
Veröffentlicht: (2024)
von: Tigges, Curt, et al.
Veröffentlicht: (2024)
Grokking Group Multiplication with Cosets
von: Stander, Dashiell, et al.
Veröffentlicht: (2023)
von: Stander, Dashiell, et al.
Veröffentlicht: (2023)
The Ghost in the Keys: A Disklavier Demo for Human-AI Musical Co-Creativity
von: Bradshaw, Louis, et al.
Veröffentlicht: (2025)
von: Bradshaw, Louis, et al.
Veröffentlicht: (2025)
Recite Your Ask Out Loud
Veröffentlicht: (2025)
Veröffentlicht: (2025)
Hubble: a Model Suite to Advance the Study of LLM Memorization
von: Wei, Johnny Tian-Zheng, et al.
Veröffentlicht: (2025)
von: Wei, Johnny Tian-Zheng, et al.
Veröffentlicht: (2025)
Benchmarks as Microscopes: A Call for Model Metrology
von: Saxon, Michael, et al.
Veröffentlicht: (2024)
von: Saxon, Michael, et al.
Veröffentlicht: (2024)
Random Scaling of Emergent Capabilities
von: Zhao, Rosie, et al.
Veröffentlicht: (2025)
von: Zhao, Rosie, et al.
Veröffentlicht: (2025)
Dynamic Masking Rate Schedules for MLM Pretraining
von: Ankner, Zachary, et al.
Veröffentlicht: (2023)
von: Ankner, Zachary, et al.
Veröffentlicht: (2023)
Auditing Private Prediction
von: Chadha, Karan, et al.
Veröffentlicht: (2024)
von: Chadha, Karan, et al.
Veröffentlicht: (2024)
RELACIONAMENTO DE QUALIDADE NO COMÉRCIO ELETRÔNICO
von: Stella Naomi Moriguchi
Veröffentlicht: (2016)
von: Stella Naomi Moriguchi
Veröffentlicht: (2016)
Ähnliche Einträge
-
Privacy Ripple Effects from Adding or Removing Personal Information in Language Model Training
von: Borkar, Jaydeep, et al.
Veröffentlicht: (2025) -
Mind the Gap: Analyzing Lacunae with Transformer-Based Transcription
von: Borkar, Jaydeep, et al.
Veröffentlicht: (2024) -
PolyPythias: Stability and Outliers across Fifty Language Model Pre-Training Runs
von: van der Wal, Oskar, et al.
Veröffentlicht: (2025) -
Mechanistic?
von: Saphra, Naomi, et al.
Veröffentlicht: (2024) -
Memorization Dynamics in Knowledge Distillation for Language Models
von: Borkar, Jaydeep, et al.
Veröffentlicht: (2026)