Understanding Verbatim Memorization in LLMs Through Circuit Discovery
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lasy, Ilya, Knees, Peter, Woltran, Stefan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Demystifying Verbatim Memorization in Large Language Models
von: Huang, Jing, et al.
Veröffentlicht: (2024)
von: Huang, Jing, et al.
Veröffentlicht: (2024)
Revisiting Non-Verbatim Memorization in Large Language Models: The Role of Entity Surface Forms
von: Nishida, Yuto, et al.
Veröffentlicht: (2026)
von: Nishida, Yuto, et al.
Veröffentlicht: (2026)
Memorization or Reasoning? Exploring the Idiom Understanding of LLMs
von: Kim, Jisu, et al.
Veröffentlicht: (2025)
von: Kim, Jisu, et al.
Veröffentlicht: (2025)
Be like a Goldfish, Don't Memorize! Mitigating Memorization in Generative LLMs
von: Hans, Abhimanyu, et al.
Veröffentlicht: (2024)
von: Hans, Abhimanyu, et al.
Veröffentlicht: (2024)
Positional Fragility in LLMs: How Offset Effects Reshape Our Understanding of Memorization Risks
von: Xu, Yixuan, et al.
Veröffentlicht: (2025)
von: Xu, Yixuan, et al.
Veröffentlicht: (2025)
All Circuits Lead to Rome: Rethinking Functional Anisotropy in Circuit and Sheaf Discovery for LLMs
von: Chen, Xi, et al.
Veröffentlicht: (2026)
von: Chen, Xi, et al.
Veröffentlicht: (2026)
Long-Term Ad Memorability: Understanding & Generating Memorable Ads
von: SI, Harini, et al.
Veröffentlicht: (2023)
von: SI, Harini, et al.
Veröffentlicht: (2023)
Do LLMs "Feel"? Emotion Circuits Discovery and Control
von: Wang, Chenxi, et al.
Veröffentlicht: (2025)
von: Wang, Chenxi, et al.
Veröffentlicht: (2025)
Alpaca against Vicuna: Using LLMs to Uncover Memorization of LLMs
von: Kassem, Aly M., et al.
Veröffentlicht: (2024)
von: Kassem, Aly M., et al.
Veröffentlicht: (2024)
Mitigating Memorization in LLMs using Activation Steering
von: Suri, Manan, et al.
Veröffentlicht: (2025)
von: Suri, Manan, et al.
Veröffentlicht: (2025)
Spurious Rewards Paradox: Mechanistically Understanding How RLVR Activates Memorization Shortcuts in LLMs
von: Yan, Lecheng, et al.
Veröffentlicht: (2026)
von: Yan, Lecheng, et al.
Veröffentlicht: (2026)
ACL-Verbatim: hallucination-free question answering for research
von: Recski, Gábor, et al.
Veröffentlicht: (2026)
von: Recski, Gábor, et al.
Veröffentlicht: (2026)
Memorization in Language Models through the Lens of Intrinsic Dimension
von: Arnold, Stefan
Veröffentlicht: (2025)
von: Arnold, Stefan
Veröffentlicht: (2025)
Memorization and Knowledge Injection in Gated LLMs
von: Pan, Xu, et al.
Veröffentlicht: (2025)
von: Pan, Xu, et al.
Veröffentlicht: (2025)
Unveiling Over-Memorization in Finetuning LLMs for Reasoning Tasks
von: Ruan, Zhiwen, et al.
Veröffentlicht: (2025)
von: Ruan, Zhiwen, et al.
Veröffentlicht: (2025)
Guiding Generative Storytelling with Knowledge Graphs
von: Pan, Zhijun, et al.
Veröffentlicht: (2025)
von: Pan, Zhijun, et al.
Veröffentlicht: (2025)
Language Models May Verbatim Complete Text They Were Not Explicitly Trained On
von: Liu, Ken Ziyu, et al.
Veröffentlicht: (2025)
von: Liu, Ken Ziyu, et al.
Veröffentlicht: (2025)
Guess or Recall? Training CNNs to Classify and Localize Memorization in LLMs
von: Dentan, Jérémie, et al.
Veröffentlicht: (2025)
von: Dentan, Jérémie, et al.
Veröffentlicht: (2025)
MedMT-Bench: Can LLMs Memorize and Understand Long Multi-Turn Conversations in Medical Scenarios?
von: Yang, Lin, et al.
Veröffentlicht: (2026)
von: Yang, Lin, et al.
Veröffentlicht: (2026)
Mirage of Mastery: Memorization Tricks LLMs into Artificially Inflated Self-Knowledge
von: Kale, Sahil
Veröffentlicht: (2025)
von: Kale, Sahil
Veröffentlicht: (2025)
Rote Learning Considered Useful: Generalizing over Memorized Data in LLMs
von: Wu, Qinyuan, et al.
Veröffentlicht: (2025)
von: Wu, Qinyuan, et al.
Veröffentlicht: (2025)
Beyond Math: Stories as a Testbed for Memorization-Constrained Reasoning in LLMs
von: Jiang, Yuxuan, et al.
Veröffentlicht: (2024)
von: Jiang, Yuxuan, et al.
Veröffentlicht: (2024)
Memorizing is Not Enough: Deep Knowledge Injection Through Reasoning
von: Xu, Ruoxi, et al.
Veröffentlicht: (2025)
von: Xu, Ruoxi, et al.
Veröffentlicht: (2025)
Acoustically Precise Hesitation Tagging Is Essential for End-to-End Verbatim Transcription Systems
von: Lin, Jhen-Ke, et al.
Veröffentlicht: (2025)
von: Lin, Jhen-Ke, et al.
Veröffentlicht: (2025)
The Landscape of Memorization in LLMs: Mechanisms, Measurement, and Mitigation
von: Xiong, Alexander, et al.
Veröffentlicht: (2025)
von: Xiong, Alexander, et al.
Veröffentlicht: (2025)
Identifying Legal Holdings with LLMs: A Systematic Study of Performance, Scale, and Memorization
von: Arvin, Chuck
Veröffentlicht: (2025)
von: Arvin, Chuck
Veröffentlicht: (2025)
Do LLMs Really Memorize Personally Identifiable Information? Revisiting PII Leakage with a Cue-Controlled Memorization Framework
von: Luo, Xiaoyu, et al.
Veröffentlicht: (2026)
von: Luo, Xiaoyu, et al.
Veröffentlicht: (2026)
Alignment Whack-a-Mole : Finetuning Activates Verbatim Recall of Copyrighted Books in Large Language Models
von: Liu, Xinyue, et al.
Veröffentlicht: (2026)
von: Liu, Xinyue, et al.
Veröffentlicht: (2026)
Combining Voting and Abstract Argumentation to Understand Online Discussions
von: Bernreiter, Michael, et al.
Veröffentlicht: (2024)
von: Bernreiter, Michael, et al.
Veröffentlicht: (2024)
Through the Prism of Culture: Evaluating LLMs' Understanding of Indian Subcultures and Traditions
von: Chhikara, Garima, et al.
Veröffentlicht: (2025)
von: Chhikara, Garima, et al.
Veröffentlicht: (2025)
Memorization $\neq$ Understanding: Do Large Language Models Have the Ability of Scenario Cognition?
von: Ma, Boxiang, et al.
Veröffentlicht: (2025)
von: Ma, Boxiang, et al.
Veröffentlicht: (2025)
LLMs and Memorization: On Quality and Specificity of Copyright Compliance
von: Mueller, Felix B, et al.
Veröffentlicht: (2024)
von: Mueller, Felix B, et al.
Veröffentlicht: (2024)
UnSeenTimeQA: Time-Sensitive Question-Answering Beyond LLMs' Memorization
von: Uddin, Md Nayem, et al.
Veröffentlicht: (2024)
von: Uddin, Md Nayem, et al.
Veröffentlicht: (2024)
Benchmarking Chinese Commonsense Reasoning of LLMs: From Chinese-Specifics to Reasoning-Memorization Correlations
von: Sun, Jiaxing, et al.
Veröffentlicht: (2024)
von: Sun, Jiaxing, et al.
Veröffentlicht: (2024)
Memorization, Emergence, and Explaining Reversal Failures: A Controlled Study of Relational Semantics in LLMs
von: Zhu, Yihua, et al.
Veröffentlicht: (2026)
von: Zhu, Yihua, et al.
Veröffentlicht: (2026)
ParaPO: Aligning Language Models to Reduce Verbatim Reproduction of Pre-training Data
von: Chen, Tong, et al.
Veröffentlicht: (2025)
von: Chen, Tong, et al.
Veröffentlicht: (2025)
Memorization vs. Reasoning: Updating LLMs with New Knowledge
von: Li, Aochong Oliver, et al.
Veröffentlicht: (2025)
von: Li, Aochong Oliver, et al.
Veröffentlicht: (2025)
Shared Path: Unraveling Memorization in Multilingual LLMs through Language Similarities
von: Luo, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Luo, Xiaoyu, et al.
Veröffentlicht: (2025)
Do Localization Methods Actually Localize Memorized Data in LLMs? A Tale of Two Benchmarks
von: Chang, Ting-Yun, et al.
Veröffentlicht: (2023)
von: Chang, Ting-Yun, et al.
Veröffentlicht: (2023)
Beyond Memorization: Distinguishing between Reductive and Epistemic Reasoning in LLMs using Classic Logic Puzzles
von: Gabay, Adi, et al.
Veröffentlicht: (2026)
von: Gabay, Adi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Demystifying Verbatim Memorization in Large Language Models
von: Huang, Jing, et al.
Veröffentlicht: (2024) -
Revisiting Non-Verbatim Memorization in Large Language Models: The Role of Entity Surface Forms
von: Nishida, Yuto, et al.
Veröffentlicht: (2026) -
Memorization or Reasoning? Exploring the Idiom Understanding of LLMs
von: Kim, Jisu, et al.
Veröffentlicht: (2025) -
Be like a Goldfish, Don't Memorize! Mitigating Memorization in Generative LLMs
von: Hans, Abhimanyu, et al.
Veröffentlicht: (2024) -
Positional Fragility in LLMs: How Offset Effects Reshape Our Understanding of Memorization Risks
von: Xu, Yixuan, et al.
Veröffentlicht: (2025)