Generalisation First, Memorisation Second? Memorisation Localisation for Natural Language Classification Tasks
Fuente:
arXiv
Saved in:
| Main Authors: | Dankers, Verna, Titov, Ivan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Understanding Memorisation in LLMs: Dynamics, Influencing Factors, and Implications
by: Speicher, Till, et al.
Published: (2024)
by: Speicher, Till, et al.
Published: (2024)
Early Detection and Reduction of Memorisation for Domain Adaptation and Instruction Tuning
by: Slack, Dean L., et al.
Published: (2025)
by: Slack, Dean L., et al.
Published: (2025)
Memorization Inheritance in Sequence-Level Knowledge Distillation for Neural Machine Translation
by: Dankers, Verna, et al.
Published: (2025)
by: Dankers, Verna, et al.
Published: (2025)
Causal Estimation of Memorisation Profiles
by: Lesci, Pietro, et al.
Published: (2024)
by: Lesci, Pietro, et al.
Published: (2024)
Model Capacity Determines Grokking through Competing Memorisation and Generalisation Speeds
by: Song, Yiding, et al.
Published: (2026)
by: Song, Yiding, et al.
Published: (2026)
Traces of Memorisation in Large Language Models for Code
by: Al-Kaswan, Ali, et al.
Published: (2023)
by: Al-Kaswan, Ali, et al.
Published: (2023)
Memorisation, convergence and generalisation in generative models
by: Maillard, Antoine, et al.
Published: (2026)
by: Maillard, Antoine, et al.
Published: (2026)
Reducing Memorisation in Generative Models via Riemannian Bayesian Inference
by: Gegenfurtner, Johanna Marie, et al.
Published: (2026)
by: Gegenfurtner, Johanna Marie, et al.
Published: (2026)
Unlearning Traces the Influential Training Data of Language Models
by: Isonuma, Masaru, et al.
Published: (2024)
by: Isonuma, Masaru, et al.
Published: (2024)
Optimising Calls to Large Language Models with Uncertainty-Based Two-Tier Selection
by: Ramírez, Guillem, et al.
Published: (2024)
by: Ramírez, Guillem, et al.
Published: (2024)
Evaluating Subword Tokenization: Alien Subword Composition and OOV Generalization Challenge
by: Batsuren, Khuyagbaatar, et al.
Published: (2024)
by: Batsuren, Khuyagbaatar, et al.
Published: (2024)
Controlling What You Share: Assessing Language Model Adherence to Privacy Preferences
by: Ramírez, Guillem, et al.
Published: (2025)
by: Ramírez, Guillem, et al.
Published: (2025)
Memorisation and forgetting in a learning Hopfield neural network: bifurcation mechanisms, attractors and basins
by: Essex, Adam E., et al.
Published: (2025)
by: Essex, Adam E., et al.
Published: (2025)
M-Wanda: Improving One-Shot Pruning for Multilingual LLMs
by: Choenni, Rochelle, et al.
Published: (2025)
by: Choenni, Rochelle, et al.
Published: (2025)
ArabicNLU 2024: The First Arabic Natural Language Understanding Shared Task
by: Khalilia, Mohammed, et al.
Published: (2024)
by: Khalilia, Mohammed, et al.
Published: (2024)
Shared Doubt: Zero-shot Cross-Lingual Confidence Estimation for Language Models
by: Kyriakou, Athina, et al.
Published: (2026)
by: Kyriakou, Athina, et al.
Published: (2026)
Explanation Regularisation through the Lens of Attributions
by: Ferreira, Pedro, et al.
Published: (2024)
by: Ferreira, Pedro, et al.
Published: (2024)
Strengthening Structural Inductive Biases by Pre-training to Perform Syntactic Transformations
by: Lindemann, Matthias, et al.
Published: (2024)
by: Lindemann, Matthias, et al.
Published: (2024)
SIP: Injecting a Structural Inductive Bias into a Seq2Seq Model by Simulation
by: Lindemann, Matthias, et al.
Published: (2023)
by: Lindemann, Matthias, et al.
Published: (2023)
Truthful or Fabricated? Using Causal Attribution to Mitigate Reward Hacking in Explanations
by: Ferreira, Pedro, et al.
Published: (2025)
by: Ferreira, Pedro, et al.
Published: (2025)
Cache & Distil: Optimising API Calls to Large Language Models
by: Ramírez, Guillem, et al.
Published: (2023)
by: Ramírez, Guillem, et al.
Published: (2023)
Anthropomimetic Uncertainty: What Verbalized Uncertainty in Language Models is Missing
by: Ulmer, Dennis, et al.
Published: (2025)
by: Ulmer, Dennis, et al.
Published: (2025)
Detecting and Pruning Prominent but Detrimental Neurons in Large Language Models
by: Ali, Ameen, et al.
Published: (2025)
by: Ali, Ameen, et al.
Published: (2025)
Efficient and Private: Memorisation under differentially private parameter-efficient fine-tuning in language models
by: Ma, Olivia, et al.
Published: (2024)
by: Ma, Olivia, et al.
Published: (2024)
What's New in My Data? Novelty Exploration via Contrastive Generation
by: Isonuma, Masaru, et al.
Published: (2024)
by: Isonuma, Masaru, et al.
Published: (2024)
Finding Culture-Sensitive Neurons in Vision-Language Models
by: Zhao, Xiutian, et al.
Published: (2025)
by: Zhao, Xiutian, et al.
Published: (2025)
Mitigating Copy Bias in In-Context Learning through Neuron Pruning
by: Ali, Ameen, et al.
Published: (2024)
by: Ali, Ameen, et al.
Published: (2024)
Clarification as Supervision: Reinforcement Learning for Vision-Language Interfaces
by: Gkountouras, John, et al.
Published: (2025)
by: Gkountouras, John, et al.
Published: (2025)
Joint Localization and Activation Editing for Low-Resource Fine-Tuning
by: Lai, Wen, et al.
Published: (2025)
by: Lai, Wen, et al.
Published: (2025)
A Controllable Examination for Long-Context Language Models
by: Yang, Yijun, et al.
Published: (2025)
by: Yang, Yijun, et al.
Published: (2025)
FOLIO: Natural Language Reasoning with First-Order Logic
by: Han, Simeng, et al.
Published: (2022)
by: Han, Simeng, et al.
Published: (2022)
Emergence and Localisation of Semantic Role Circuits in LLMs
by: Aljaafari, Nura, et al.
Published: (2025)
by: Aljaafari, Nura, et al.
Published: (2025)
Do Generalisation Results Generalise?
by: Boglioni, Matteo, et al.
Published: (2025)
by: Boglioni, Matteo, et al.
Published: (2025)
Privacy-Preserving Models for Legal Natural Language Processing
by: Yin, Ying, et al.
Published: (2022)
by: Yin, Ying, et al.
Published: (2022)
A Systematic Evaluation of Large Language Models for Natural Language Generation Tasks
by: Ni, Xuanfan, et al.
Published: (2024)
by: Ni, Xuanfan, et al.
Published: (2024)
FOSSIL: Harnessing Feedback on Suboptimal Samples for Data-Efficient Generalisation with Imitation Learning for Embodied Vision-and-Language Tasks
by: McCallum, Sabrina, et al.
Published: (2025)
by: McCallum, Sabrina, et al.
Published: (2025)
Evaluation of Few-Shot Learning for Classification Tasks in the Polish Language
by: Hadeliya, Tsimur, et al.
Published: (2024)
by: Hadeliya, Tsimur, et al.
Published: (2024)
Caption First, VQA Second: Knowledge Density, Not Task Format, Drives Multimodal Scaling
by: Zou, Hongjian, et al.
Published: (2026)
by: Zou, Hongjian, et al.
Published: (2026)
Multi-Step Deductive Reasoning Over Natural Language: An Empirical Study on Out-of-Distribution Generalisation
by: Bao, Qiming, et al.
Published: (2022)
by: Bao, Qiming, et al.
Published: (2022)
Enhancing Long Document Long Form Summarisation with Self-Planning
by: Du, Xiaotang, et al.
Published: (2025)
by: Du, Xiaotang, et al.
Published: (2025)
Similar Items
-
Understanding Memorisation in LLMs: Dynamics, Influencing Factors, and Implications
by: Speicher, Till, et al.
Published: (2024) -
Early Detection and Reduction of Memorisation for Domain Adaptation and Instruction Tuning
by: Slack, Dean L., et al.
Published: (2025) -
Memorization Inheritance in Sequence-Level Knowledge Distillation for Neural Machine Translation
by: Dankers, Verna, et al.
Published: (2025) -
Causal Estimation of Memorisation Profiles
by: Lesci, Pietro, et al.
Published: (2024) -
Model Capacity Determines Grokking through Competing Memorisation and Generalisation Speeds
by: Song, Yiding, et al.
Published: (2026)