MSSR: Memory-Aware Adaptive Replay for Continual LLM Fine-Tuning
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Lu, Yiyang, He, Yu, Chen, Jianlong, Zha, Hongyuan |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Self-generated Replay Memories for Continual Neural Machine Translation
par: Resta, Michele, et autres
Publié: (2024)
par: Resta, Michele, et autres
Publié: (2024)
FOREVER: Forgetting Curve-Inspired Memory Replay for Language Model Continual Learning
par: Feng, Yujie, et autres
Publié: (2026)
par: Feng, Yujie, et autres
Publié: (2026)
AirLLM: Diffusion Policy-based Adaptive LoRA for Remote Fine-Tuning of LLM over the Air
par: Yang, Shiyi, et autres
Publié: (2025)
par: Yang, Shiyi, et autres
Publié: (2025)
SuRe: Surprise-Driven Prioritised Replay for Continual LLM Learning
par: Hazard, Hugo, et autres
Publié: (2025)
par: Hazard, Hugo, et autres
Publié: (2025)
Filter-then-Weight: Online Data Selection and Reweighting for LLM Fine-Tuning
par: Wang, Fangxin, et autres
Publié: (2026)
par: Wang, Fangxin, et autres
Publié: (2026)
CURLoRA: Stable LLM Continual Fine-Tuning and Catastrophic Forgetting Mitigation
par: Fawi, Muhammad
Publié: (2024)
par: Fawi, Muhammad
Publié: (2024)
GeRe: Towards Efficient Anti-Forgetting in Continual Learning of LLM via General Samples Replay
par: Zhang, Yunan, et autres
Publié: (2025)
par: Zhang, Yunan, et autres
Publié: (2025)
Improving Data Efficiency for LLM Reinforcement Fine-tuning Through Difficulty-targeted Online Data Selection and Rollout Replay
par: Sun, Yifan, et autres
Publié: (2025)
par: Sun, Yifan, et autres
Publié: (2025)
Understanding the Performance and Estimating the Cost of LLM Fine-Tuning
par: Xia, Yuchen, et autres
Publié: (2024)
par: Xia, Yuchen, et autres
Publié: (2024)
Improving Large Language Models with Concept-Aware Fine-Tuning
par: Chen, Michael K., et autres
Publié: (2025)
par: Chen, Michael K., et autres
Publié: (2025)
AdaZeta: Adaptive Zeroth-Order Tensor-Train Adaption for Memory-Efficient Large Language Models Fine-Tuning
par: Yang, Yifan, et autres
Publié: (2024)
par: Yang, Yifan, et autres
Publié: (2024)
Fine-Tune an SLM or Prompt an LLM? The Case of Generating Low-Code Workflows
par: Ayala, Orlando Marquez, et autres
Publié: (2025)
par: Ayala, Orlando Marquez, et autres
Publié: (2025)
Hallucination Detection in LLMs: Fast and Memory-Efficient Fine-Tuned Models
par: Arteaga, Gabriel Y., et autres
Publié: (2024)
par: Arteaga, Gabriel Y., et autres
Publié: (2024)
SibylSense: Adaptive Rubric Learning via Memory Tuning and Adversarial Probing
par: Xu, Yifei, et autres
Publié: (2026)
par: Xu, Yifei, et autres
Publié: (2026)
FutureSim: Replaying World Events to Evaluate Adaptive Agents
par: Goel, Shashwat, et autres
Publié: (2026)
par: Goel, Shashwat, et autres
Publié: (2026)
Selection of LLM Fine-Tuning Data based on Orthogonal Rules
par: Li, Xiaomin, et autres
Publié: (2024)
par: Li, Xiaomin, et autres
Publié: (2024)
LLMem: Estimating GPU Memory Usage for Fine-Tuning Pre-Trained LLMs
par: Kim, Taeho, et autres
Publié: (2024)
par: Kim, Taeho, et autres
Publié: (2024)
Internalizing LLM Reasoning via Discovery and Replay of Latent Actions
par: Shi, Zhenning, et autres
Publié: (2026)
par: Shi, Zhenning, et autres
Publié: (2026)
Selective Reflection-Tuning: Student-Selected Data Recycling for LLM Instruction-Tuning
par: Li, Ming, et autres
Publié: (2024)
par: Li, Ming, et autres
Publié: (2024)
Entropy-Adaptive Fine-Tuning: Resolving Confident Conflicts to Mitigate Forgetting
par: Diao, Muxi, et autres
Publié: (2026)
par: Diao, Muxi, et autres
Publié: (2026)
Position-Aware Parameter Efficient Fine-Tuning Approach for Reducing Positional Bias in LLMs
par: Zhang, Zheng, et autres
Publié: (2024)
par: Zhang, Zheng, et autres
Publié: (2024)
Reinforcement Fine-Tuning Naturally Mitigates Forgetting in Continual Post-Training
par: Lai, Song, et autres
Publié: (2025)
par: Lai, Song, et autres
Publié: (2025)
Stabilizing LLM Supervised Fine-Tuning via Explicit Distributional Control
par: Wang, Xinyu, et autres
Publié: (2026)
par: Wang, Xinyu, et autres
Publié: (2026)
Rethinking Parameter Sharing for LLM Fine-Tuning with Multiple LoRAs
par: Ban, Hao, et autres
Publié: (2025)
par: Ban, Hao, et autres
Publié: (2025)
Harmony in Divergence: Towards Fast, Accurate, and Memory-efficient Zeroth-order LLM Fine-tuning
par: Tan, Qitao, et autres
Publié: (2025)
par: Tan, Qitao, et autres
Publié: (2025)
RePO: Replay-Enhanced Policy Optimization
par: Li, Siheng, et autres
Publié: (2025)
par: Li, Siheng, et autres
Publié: (2025)
Inference-Aware Fine-Tuning for Best-of-N Sampling in Large Language Models
par: Chow, Yinlam, et autres
Publié: (2024)
par: Chow, Yinlam, et autres
Publié: (2024)
Revisiting Replay and Gradient Alignment for Continual Pre-Training of Large Language Models
par: Abbes, Istabrak, et autres
Publié: (2025)
par: Abbes, Istabrak, et autres
Publié: (2025)
Not All Adapters Matter: Selective Adapter Freezing for Memory-Efficient Fine-Tuning of Language Models
par: Son, Hyegang, et autres
Publié: (2024)
par: Son, Hyegang, et autres
Publié: (2024)
Aligning Backchannel and Dialogue Context Representations via Contrastive LLM Fine-Tuning
par: Qian, Livia, et autres
Publié: (2026)
par: Qian, Livia, et autres
Publié: (2026)
Mitigating Training Imbalance in LLM Fine-Tuning via Selective Parameter Merging
par: Ju, Yiming, et autres
Publié: (2024)
par: Ju, Yiming, et autres
Publié: (2024)
Secure LLM Fine-Tuning via Safety-Aware Probing
par: Wu, Chengcan, et autres
Publié: (2025)
par: Wu, Chengcan, et autres
Publié: (2025)
Embedding Enhancement via Fine-Tuned Language Models for Learner-Item Cognitive Modeling
par: Liu, Yuanhao, et autres
Publié: (2026)
par: Liu, Yuanhao, et autres
Publié: (2026)
Neural Parameter Search for Slimmer Fine-Tuned Models and Better Transfer
par: Du, Guodong, et autres
Publié: (2025)
par: Du, Guodong, et autres
Publié: (2025)
EMORL: Ensemble Multi-Objective Reinforcement Learning for Efficient and Flexible LLM Fine-Tuning
par: Kong, Lingxiao, et autres
Publié: (2025)
par: Kong, Lingxiao, et autres
Publié: (2025)
LLMs Meet Finance: Fine-Tuning Foundation Models for the Open FinLLM Leaderboard
par: Rao, Varun, et autres
Publié: (2025)
par: Rao, Varun, et autres
Publié: (2025)
Parameter-Efficient Routed Fine-Tuning: Mixture-of-Experts Demands Mixture of Adaptation Modules
par: Liu, Yilun, et autres
Publié: (2025)
par: Liu, Yilun, et autres
Publié: (2025)
Parameter-Efficient Fine-Tuning with Discrete Fourier Transform
par: Gao, Ziqi, et autres
Publié: (2024)
par: Gao, Ziqi, et autres
Publié: (2024)
Proximal Supervised Fine-Tuning
par: Zhu, Wenhong, et autres
Publié: (2025)
par: Zhu, Wenhong, et autres
Publié: (2025)
MatryoshkaLoRA: Learning Accurate Hierarchical Low-Rank Representations for LLM Fine-Tuning
par: Modoranu, Ionut-Vlad, et autres
Publié: (2026)
par: Modoranu, Ionut-Vlad, et autres
Publié: (2026)
Documents similaires
-
Self-generated Replay Memories for Continual Neural Machine Translation
par: Resta, Michele, et autres
Publié: (2024) -
FOREVER: Forgetting Curve-Inspired Memory Replay for Language Model Continual Learning
par: Feng, Yujie, et autres
Publié: (2026) -
AirLLM: Diffusion Policy-based Adaptive LoRA for Remote Fine-Tuning of LLM over the Air
par: Yang, Shiyi, et autres
Publié: (2025) -
SuRe: Surprise-Driven Prioritised Replay for Continual LLM Learning
par: Hazard, Hugo, et autres
Publié: (2025) -
Filter-then-Weight: Online Data Selection and Reweighting for LLM Fine-Tuning
par: Wang, Fangxin, et autres
Publié: (2026)