Collaborative Parameter Learning: Mitigating Forgetting via Parameter-Level Gradient Analysis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Mutian, Zhan, Zisen, Chen, Yutong, Li, Haolin, Wang, Kaiwen, Zheng, Kaili, Wang, Yuguang, Wang, Qi, Gao, Jiandong, Wu, Ji |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Decoupling Knowledge and Reasoning in LLMs: An Exploration Using Cognitive Dual-System Theory
von: Yang, Mutian, et al.
Veröffentlicht: (2025)
von: Yang, Mutian, et al.
Veröffentlicht: (2025)
Bayesian Parameter-Efficient Fine-Tuning for Overcoming Catastrophic Forgetting
von: Chen, Haolin, et al.
Veröffentlicht: (2024)
von: Chen, Haolin, et al.
Veröffentlicht: (2024)
Towards Revealing the Effectiveness of Small-Scale Fine-tuning in R1-style Reinforcement Learning
von: Chen, Yutong, et al.
Veröffentlicht: (2025)
von: Chen, Yutong, et al.
Veröffentlicht: (2025)
Locate-then-Merge: Neuron-Level Parameter Fusion for Mitigating Catastrophic Forgetting in Multimodal LLMs
von: Yu, Zeping, et al.
Veröffentlicht: (2025)
von: Yu, Zeping, et al.
Veröffentlicht: (2025)
CoScale-RL: Efficient Post-Training by Co-Scaling Data and Computation
von: Chen, Yutong, et al.
Veröffentlicht: (2026)
von: Chen, Yutong, et al.
Veröffentlicht: (2026)
Dynamic feature selection in medical predictive monitoring by reinforcement learning
von: Chen, Yutong, et al.
Veröffentlicht: (2024)
von: Chen, Yutong, et al.
Veröffentlicht: (2024)
InterMesh: Explicit Interaction-Aware End-to-End Multi-Person Human Mesh Recovery
von: Zheng, Kaili, et al.
Veröffentlicht: (2026)
von: Zheng, Kaili, et al.
Veröffentlicht: (2026)
Towards Metric-Aware Multi-Person Mesh Recovery by Jointly Optimizing Human Crowd in Camera Space
von: Wang, Kaiwen, et al.
Veröffentlicht: (2025)
von: Wang, Kaiwen, et al.
Veröffentlicht: (2025)
Parameter Alignment Mitigates Catastrophic Forgetting in Multilingual Expert Language Models
von: Ahuja, Sanchit, et al.
Veröffentlicht: (2026)
von: Ahuja, Sanchit, et al.
Veröffentlicht: (2026)
Analyzing and Reducing Catastrophic Forgetting in Parameter Efficient Tuning
von: Ren, Weijieying, et al.
Veröffentlicht: (2024)
von: Ren, Weijieying, et al.
Veröffentlicht: (2024)
Turning Back Without Forgetting: Selective Backward Refinement for Parameter-Efficient Continual Learning
von: Tiwari, Anushka, et al.
Veröffentlicht: (2026)
von: Tiwari, Anushka, et al.
Veröffentlicht: (2026)
BoxComm: Benchmarking Category-Aware Commentary Generation and Narration Rhythm in Boxing
von: Wang, Kaiwen, et al.
Veröffentlicht: (2026)
von: Wang, Kaiwen, et al.
Veröffentlicht: (2026)
DualCap: Enhancing Lightweight Image Captioning via Dual Retrieval with Similar Scenes Visual Prompts
von: Li, Binbin, et al.
Veröffentlicht: (2025)
von: Li, Binbin, et al.
Veröffentlicht: (2025)
Mitigating Visual Knowledge Forgetting in MLLM Instruction-tuning via Modality-decoupled Gradient Descent
von: Wu, Junda, et al.
Veröffentlicht: (2025)
von: Wu, Junda, et al.
Veröffentlicht: (2025)
Mitigating Forgetting in Continual Learning with Selective Gradient Projection
von: Singh, Anika, et al.
Veröffentlicht: (2026)
von: Singh, Anika, et al.
Veröffentlicht: (2026)
Alleviating Forgetfulness of Linear Attention by Hybrid Sparse Attention and Contextualized Learnable Token Eviction
von: He, Mutian, et al.
Veröffentlicht: (2025)
von: He, Mutian, et al.
Veröffentlicht: (2025)
Exploring the Impact of Parameter Update Magnitude on Forgetting and Generalization of Continual Learning
von: He, JinLi, et al.
Veröffentlicht: (2026)
von: He, JinLi, et al.
Veröffentlicht: (2026)
Machine Learning–Driven Screening of High‐Activity Antitumor Nanozymes Using an Ensemble Enzymatic Oracle System
von: Guanmeng Zhang, et al.
Veröffentlicht: (2025)
von: Guanmeng Zhang, et al.
Veröffentlicht: (2025)
Parameter Symmetry and Noise Equilibrium of Stochastic Gradient Descent
von: Ziyin, Liu, et al.
Veröffentlicht: (2024)
von: Ziyin, Liu, et al.
Veröffentlicht: (2024)
Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy
von: Xie, Yutong, et al.
Veröffentlicht: (2026)
von: Xie, Yutong, et al.
Veröffentlicht: (2026)
More Than Memory Savings: Zeroth-Order Optimization Mitigates Forgetting in Continual Learning
von: Yu, Wanhao, et al.
Veröffentlicht: (2025)
von: Yu, Wanhao, et al.
Veröffentlicht: (2025)
Universal and Parameter-free Gradient Sliding for Composite Optimization
von: Wu, Yan, et al.
Veröffentlicht: (2026)
von: Wu, Yan, et al.
Veröffentlicht: (2026)
Understanding Textual Capability Degradation in Speech LLMs via Parameter Importance Analysis
von: Wang, Chao, et al.
Veröffentlicht: (2025)
von: Wang, Chao, et al.
Veröffentlicht: (2025)
Towards Aligned Data Forgetting via Twin Machine Unlearning
von: Niu, Zhenxing, et al.
Veröffentlicht: (2025)
von: Niu, Zhenxing, et al.
Veröffentlicht: (2025)
Automated Federated Pipeline for Parameter-Efficient Fine-Tuning of Large Language Models
von: Fang, Zihan, et al.
Veröffentlicht: (2024)
von: Fang, Zihan, et al.
Veröffentlicht: (2024)
Deep Learning Powered Estimate of The Extrinsic Parameters on Unmanned Surface Vehicles
von: Shen, Yi, et al.
Veröffentlicht: (2024)
von: Shen, Yi, et al.
Veröffentlicht: (2024)
Impact Force Algorithm and Parameters of Rolling Stone Impact Pier in Mountain Area
von: Zi-Jian Wang, et al.
Veröffentlicht: (2024)
von: Zi-Jian Wang, et al.
Veröffentlicht: (2024)
On Catastrophic Forgetting in Low-Rank Decomposition-Based Parameter-Efficient Fine-Tuning
von: Ahmad, Muhammad, et al.
Veröffentlicht: (2026)
von: Ahmad, Muhammad, et al.
Veröffentlicht: (2026)
The Gradient Descent Adaptive Moment Parameter Estimation for Multi‐Frequency Sine Signal Systems
von: Kai Zhang, et al.
Veröffentlicht: (2025)
von: Kai Zhang, et al.
Veröffentlicht: (2025)
Mitigating Catastrophic Forgetting in Large Language Models with Forgetting-aware Pruning
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
Enhanced hardness and toughness in (V, Zr, Nb, Ta)C high‐entropy ceramics via core‐rim structure
von: Bo Zhao, et al.
Veröffentlicht: (2025)
von: Bo Zhao, et al.
Veröffentlicht: (2025)
Silent Sabotage During Fine-Tuning: Few-Shot Rationale Poisoning of Compact Medical LLMs
von: Xie, Jingyuan, et al.
Veröffentlicht: (2026)
von: Xie, Jingyuan, et al.
Veröffentlicht: (2026)
KnobTree: Intelligent Database Parameter Configuration via Explainable Reinforcement Learning
von: Chen, Jiahan, et al.
Veröffentlicht: (2024)
von: Chen, Jiahan, et al.
Veröffentlicht: (2024)
ColA: Collaborative Adaptation with Gradient Learning
von: Diao, Enmao, et al.
Veröffentlicht: (2024)
von: Diao, Enmao, et al.
Veröffentlicht: (2024)
Enhancing Large Language Model Performance with Gradient-Based Parameter Selection
von: Li, Haoling, et al.
Veröffentlicht: (2024)
von: Li, Haoling, et al.
Veröffentlicht: (2024)
Knowledge Gradient for Preference Learning
von: Wu, Kaiwen, et al.
Veröffentlicht: (2026)
von: Wu, Kaiwen, et al.
Veröffentlicht: (2026)
ParameterNet: Parameters Are All You Need
von: Han, Kai, et al.
Veröffentlicht: (2023)
von: Han, Kai, et al.
Veröffentlicht: (2023)
Parameter Importance-Driven Continual Learning for Foundation Models
von: Wang, Lingxiang, et al.
Veröffentlicht: (2025)
von: Wang, Lingxiang, et al.
Veröffentlicht: (2025)
SelfAug: Mitigating Catastrophic Forgetting in Retrieval-Augmented Generation via Distribution Self-Alignment
von: Huang, Yuqing, et al.
Veröffentlicht: (2025)
von: Huang, Yuqing, et al.
Veröffentlicht: (2025)
Save It All: Enabling Full Parameter Tuning for Federated Large Language Models via Cycle Block Gradient Descent
von: Wang, Lin, et al.
Veröffentlicht: (2024)
von: Wang, Lin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Decoupling Knowledge and Reasoning in LLMs: An Exploration Using Cognitive Dual-System Theory
von: Yang, Mutian, et al.
Veröffentlicht: (2025) -
Bayesian Parameter-Efficient Fine-Tuning for Overcoming Catastrophic Forgetting
von: Chen, Haolin, et al.
Veröffentlicht: (2024) -
Towards Revealing the Effectiveness of Small-Scale Fine-tuning in R1-style Reinforcement Learning
von: Chen, Yutong, et al.
Veröffentlicht: (2025) -
Locate-then-Merge: Neuron-Level Parameter Fusion for Mitigating Catastrophic Forgetting in Multimodal LLMs
von: Yu, Zeping, et al.
Veröffentlicht: (2025) -
CoScale-RL: Efficient Post-Training by Co-Scaling Data and Computation
von: Chen, Yutong, et al.
Veröffentlicht: (2026)