Parameter Alignment Mitigates Catastrophic Forgetting in Multilingual Expert Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Ahuja, Sanchit, Blevins, Terra |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Accommodation Goes Both Ways: Studying Linguistic Convergence Between Humans and Language Models
by: Blevins, Terra
Published: (2026)
by: Blevins, Terra
Published: (2026)
Breaking the Curse of Multilinguality with Cross-lingual Expert Language Models
by: Blevins, Terra, et al.
Published: (2024)
by: Blevins, Terra, et al.
Published: (2024)
Contamination Report for Multilingual Benchmarks
by: Ahuja, Sanchit, et al.
Published: (2024)
by: Ahuja, Sanchit, et al.
Published: (2024)
Comparing Hallucination Detection Metrics for Multilingual Generation
by: Kang, Haoqiang, et al.
Published: (2024)
by: Kang, Haoqiang, et al.
Published: (2024)
Conditions for Catastrophic Forgetting in Multilingual Translation
by: Liu, Danni, et al.
Published: (2025)
by: Liu, Danni, et al.
Published: (2025)
Model Tailor: Mitigating Catastrophic Forgetting in Multi-modal Large Language Models
by: Zhu, Didi, et al.
Published: (2024)
by: Zhu, Didi, et al.
Published: (2024)
Targeted Multilingual Adaptation for Low-resource Language Families
by: Downey, C. M., et al.
Published: (2024)
by: Downey, C. M., et al.
Published: (2024)
MYTE: Morphology-Driven Byte Encoding for Better and Fairer Multilingual Language Modeling
by: Limisiewicz, Tomasz, et al.
Published: (2024)
by: Limisiewicz, Tomasz, et al.
Published: (2024)
Mitigating Catastrophic Forgetting in Large Language Models with Self-Synthesized Rehearsal
by: Huang, Jianheng, et al.
Published: (2024)
by: Huang, Jianheng, et al.
Published: (2024)
Locate-then-Merge: Neuron-Level Parameter Fusion for Mitigating Catastrophic Forgetting in Multimodal LLMs
by: Yu, Zeping, et al.
Published: (2025)
by: Yu, Zeping, et al.
Published: (2025)
Mitigating Catastrophic Forgetting in Continual Learning through Model Growth
by: Süalp, Ege, et al.
Published: (2025)
by: Süalp, Ege, et al.
Published: (2025)
Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting
by: Ding, Fei, et al.
Published: (2025)
by: Ding, Fei, et al.
Published: (2025)
Sharpness-Aware Pretraining Mitigates Catastrophic Forgetting
by: Watts, Ishaan, et al.
Published: (2026)
by: Watts, Ishaan, et al.
Published: (2026)
Model-Dowser: Data-Free Importance Probing to Mitigate Catastrophic Forgetting in Multimodal Large Language Models
by: Hwang, Hyeontaek, et al.
Published: (2026)
by: Hwang, Hyeontaek, et al.
Published: (2026)
Scaling Laws for Multilingual Language Models
by: He, Yifei, et al.
Published: (2024)
by: He, Yifei, et al.
Published: (2024)
SelfAug: Mitigating Catastrophic Forgetting in Retrieval-Augmented Generation via Distribution Self-Alignment
by: Huang, Yuqing, et al.
Published: (2025)
by: Huang, Yuqing, et al.
Published: (2025)
Balancing Speciality and Versatility: A Coarse to Fine Framework for Mitigating Catastrophic Forgetting in Large Language Models
by: Zhang, Hengyuan, et al.
Published: (2024)
by: Zhang, Hengyuan, et al.
Published: (2024)
Do language models accommodate their users? A study of linguistic convergence
by: Blevins, Terra, et al.
Published: (2025)
by: Blevins, Terra, et al.
Published: (2025)
Revisiting Catastrophic Forgetting in Large Language Model Tuning
by: Li, Hongyu, et al.
Published: (2024)
by: Li, Hongyu, et al.
Published: (2024)
Relation Extraction or Pattern Matching? Unravelling the Generalisation Limits of Language Models for Biographical RE
by: Arzt, Varvara, et al.
Published: (2025)
by: Arzt, Varvara, et al.
Published: (2025)
Mitigating Catastrophic Forgetting in Target Language Adaptation of LLMs via Source-Shielded Updates
by: Yamaguchi, Atsuki, et al.
Published: (2025)
by: Yamaguchi, Atsuki, et al.
Published: (2025)
Demystifying Prompts in Language Models via Perplexity Estimation
by: Gonen, Hila, et al.
Published: (2022)
by: Gonen, Hila, et al.
Published: (2022)
Analyzing and Reducing Catastrophic Forgetting in Parameter Efficient Tuning
by: Ren, Weijieying, et al.
Published: (2024)
by: Ren, Weijieying, et al.
Published: (2024)
Understanding Catastrophic Forgetting in Language Models via Implicit Inference
by: Kotha, Suhas, et al.
Published: (2023)
by: Kotha, Suhas, et al.
Published: (2023)
Mitigating Catastrophic Forgetting in Mathematical Reasoning Finetuning through Mixed Training
by: Reynolds, John Graham
Published: (2025)
by: Reynolds, John Graham
Published: (2025)
Does Liking Yellow Imply Driving a School Bus? Semantic Leakage in Language Models
by: Gonen, Hila, et al.
Published: (2024)
by: Gonen, Hila, et al.
Published: (2024)
EfficientXLang: Towards Improving Token Efficiency Through Cross-Lingual Reasoning
by: Ahuja, Sanchit, et al.
Published: (2025)
by: Ahuja, Sanchit, et al.
Published: (2025)
Analyzing Mitigation Strategies for Catastrophic Forgetting in End-to-End Training of Spoken Language Models
by: Hsiao, Chi-Yuan, et al.
Published: (2025)
by: Hsiao, Chi-Yuan, et al.
Published: (2025)
A Comparative Empirical Study of Catastrophic Forgetting Mitigation in Sequential Task Adaptation for Continual Natural Language Processing Systems
by: Abrahamyan, Aram, et al.
Published: (2026)
by: Abrahamyan, Aram, et al.
Published: (2026)
Attractor Patch Networks: Reducing Catastrophic Forgetting with Routed Low-Rank Patch Experts
by: Shashank
Published: (2026)
by: Shashank
Published: (2026)
An Empirical Study of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning
by: Luo, Yun, et al.
Published: (2023)
by: Luo, Yun, et al.
Published: (2023)
Catastrophic Forgetting in LLMs: A Comparative Analysis Across Language Tasks
by: Haque, Naimul
Published: (2025)
by: Haque, Naimul
Published: (2025)
Can Small Language Models Use What They Retrieve? An Empirical Study of Retrieval Utilization Across Model Scale
by: Pandey, Sanchit
Published: (2026)
by: Pandey, Sanchit
Published: (2026)
Dynamic Orthogonal Continual Fine-tuning for Mitigating Catastrophic Forgettings
by: Zhang, Zhixin, et al.
Published: (2025)
by: Zhang, Zhixin, et al.
Published: (2025)
Mitigating Catastrophic Forgetting in Multi-domain Chinese Spelling Correction by Multi-stage Knowledge Transfer Framework
by: Xing, Peng, et al.
Published: (2024)
by: Xing, Peng, et al.
Published: (2024)
CURLoRA: Stable LLM Continual Fine-Tuning and Catastrophic Forgetting Mitigation
by: Fawi, Muhammad
Published: (2024)
by: Fawi, Muhammad
Published: (2024)
Calibration Is Not Enough: Evaluating Confidence Estimation Under Language Variations
by: Xia, Yuxi, et al.
Published: (2026)
by: Xia, Yuxi, et al.
Published: (2026)
Preventing Catastrophic Forgetting: Behavior-Aware Sampling for Safer Language Model Fine-Tuning
by: Pham, Anh, et al.
Published: (2025)
by: Pham, Anh, et al.
Published: (2025)
DOSA: A Dataset of Social Artifacts from Different Indian Geographical Subcultures
by: Seth, Agrima, et al.
Published: (2024)
by: Seth, Agrima, et al.
Published: (2024)
OPLoRA: Orthogonal Projection LoRA Prevents Catastrophic Forgetting during Parameter-Efficient Fine-Tuning
by: Xiong, Yifeng, et al.
Published: (2025)
by: Xiong, Yifeng, et al.
Published: (2025)
Similar Items
-
Accommodation Goes Both Ways: Studying Linguistic Convergence Between Humans and Language Models
by: Blevins, Terra
Published: (2026) -
Breaking the Curse of Multilinguality with Cross-lingual Expert Language Models
by: Blevins, Terra, et al.
Published: (2024) -
Contamination Report for Multilingual Benchmarks
by: Ahuja, Sanchit, et al.
Published: (2024) -
Comparing Hallucination Detection Metrics for Multilingual Generation
by: Kang, Haoqiang, et al.
Published: (2024) -
Conditions for Catastrophic Forgetting in Multilingual Translation
by: Liu, Danni, et al.
Published: (2025)