Train with Perturbation, Infer after Merging: A Two-Stage Framework for Continual Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qiu, Haomiao, Zhang, Miao, Qiao, Ziyue, Nie, Liqiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SplitLoRA: Balancing Stability and Plasticity in Continual Learning Through Gradient Space Splitting
von: Qiu, Haomiao, et al.
Veröffentlicht: (2025)
von: Qiu, Haomiao, et al.
Veröffentlicht: (2025)
Benchmarking Post-Training Quantization in LLMs: Comprehensive Taxonomy, Unified Evaluation, and Comparative Analysis
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2025)
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2025)
GraphDancer: Training LLMs to Explore and Reason over Graphs via Two-Stage Curriculum Post-Training
von: Bai, Yuyang, et al.
Veröffentlicht: (2026)
von: Bai, Yuyang, et al.
Veröffentlicht: (2026)
Limited Reference, Reliable Generation: A Two-Component Framework for Tabular Data Generation in Low-Data Regimes
von: Jiang, Mingxuan, et al.
Veröffentlicht: (2025)
von: Jiang, Mingxuan, et al.
Veröffentlicht: (2025)
FedDRL: A Trustworthy Federated Learning Model Fusion Method Based on Staged Reinforcement Learning
von: Chen, Leiming, et al.
Veröffentlicht: (2023)
von: Chen, Leiming, et al.
Veröffentlicht: (2023)
CAT Merging: A Training-Free Approach for Resolving Conflicts in Model Merging
von: Sun, Wenju, et al.
Veröffentlicht: (2025)
von: Sun, Wenju, et al.
Veröffentlicht: (2025)
Merge then Realign: Simple and Effective Modality-Incremental Continual Learning for Multimodal LLMs
von: Zhang, Dingkun, et al.
Veröffentlicht: (2025)
von: Zhang, Dingkun, et al.
Veröffentlicht: (2025)
Tunable MAGMAX: Preference-Aware Model Merging for Continual Learning
von: Hiroshima, Kei, et al.
Veröffentlicht: (2026)
von: Hiroshima, Kei, et al.
Veröffentlicht: (2026)
MergeMix: Optimizing Mid-Training Data Mixtures via Learnable Model Merging
von: Wang, Jiapeng, et al.
Veröffentlicht: (2026)
von: Wang, Jiapeng, et al.
Veröffentlicht: (2026)
Modular Delta Merging with Orthogonal Constraints: A Scalable Framework for Continual and Reversible Model Composition
von: Khan, Haris, et al.
Veröffentlicht: (2025)
von: Khan, Haris, et al.
Veröffentlicht: (2025)
TAET: Two-Stage Adversarial Equalization Training on Long-Tailed Distributions
von: YuHang, Wang, et al.
Veröffentlicht: (2025)
von: YuHang, Wang, et al.
Veröffentlicht: (2025)
Training-free Heterogeneous Model Merging
von: Xu, Zhengqi, et al.
Veröffentlicht: (2024)
von: Xu, Zhengqi, et al.
Veröffentlicht: (2024)
The Two-Stage Decision-Sampling Hypothesis: Understanding the Emergence of Self-Reflection in RL-Trained LLMs
von: Zhao, Zibo, et al.
Veröffentlicht: (2026)
von: Zhao, Zibo, et al.
Veröffentlicht: (2026)
BlendRL: A Framework for Merging Symbolic and Neural Policy Learning
von: Shindo, Hikaru, et al.
Veröffentlicht: (2024)
von: Shindo, Hikaru, et al.
Veröffentlicht: (2024)
MAny: Merge Anything for Multimodal Continual Instruction Tuning
von: Gao, Zijian, et al.
Veröffentlicht: (2026)
von: Gao, Zijian, et al.
Veröffentlicht: (2026)
One Train for Two Tasks: An Encrypted Traffic Classification Framework Using Supervised Contrastive Learning
von: Zhang, Haozhen, et al.
Veröffentlicht: (2024)
von: Zhang, Haozhen, et al.
Veröffentlicht: (2024)
Non-Neighbors Also Matter to Kriging: A New Contrastive-Prototypical Learning
von: Li, Zhishuai, et al.
Veröffentlicht: (2024)
von: Li, Zhishuai, et al.
Veröffentlicht: (2024)
Rethinking Graph Contrastive Learning through Relative Similarity Preservation
von: Ning, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Ning, Zhiyuan, et al.
Veröffentlicht: (2025)
DASH: Fast Differentiable Architecture Search for Hybrid Attention in Minutes on a Single GPU
von: Chen, Weizhe, et al.
Veröffentlicht: (2026)
von: Chen, Weizhe, et al.
Veröffentlicht: (2026)
Fine, I'll Merge It Myself: A Multi-Fidelity Framework for Automated Model Merging
von: Su, Guinan, et al.
Veröffentlicht: (2025)
von: Su, Guinan, et al.
Veröffentlicht: (2025)
Two-Stage Learned Decomposition for Scalable Routing on Multigraphs
von: Rydin, Filip, et al.
Veröffentlicht: (2026)
von: Rydin, Filip, et al.
Veröffentlicht: (2026)
Learning Scenario Reduction for Two-Stage Robust Optimization with Discrete Uncertainty
von: Lin, Tianjue, et al.
Veröffentlicht: (2026)
von: Lin, Tianjue, et al.
Veröffentlicht: (2026)
FCOS: A Two-Stage Recoverable Model Pruning Framework for Automatic Modulation Recognition
von: Lu, Yao, et al.
Veröffentlicht: (2025)
von: Lu, Yao, et al.
Veröffentlicht: (2025)
Revisiting Plasticity in Visual Reinforcement Learning: Data, Modules and Training Stages
von: Ma, Guozheng, et al.
Veröffentlicht: (2023)
von: Ma, Guozheng, et al.
Veröffentlicht: (2023)
Training-free LLM Merging for Multi-task Learning
von: Fu, Zichuan, et al.
Veröffentlicht: (2025)
von: Fu, Zichuan, et al.
Veröffentlicht: (2025)
OptProver: Bridging Olympiad and Optimization through Continual Training in Formal Theorem Proving
von: Li, Chenyi, et al.
Veröffentlicht: (2026)
von: Li, Chenyi, et al.
Veröffentlicht: (2026)
Hessian Aware Low-Rank Perturbation for Order-Robust Continual Learning
von: Li, Jiaqi, et al.
Veröffentlicht: (2023)
von: Li, Jiaqi, et al.
Veröffentlicht: (2023)
Toward a Holistic Approach to Continual Model Merging
von: Phan, Hoang, et al.
Veröffentlicht: (2025)
von: Phan, Hoang, et al.
Veröffentlicht: (2025)
Unlocking the Potential of Continual Model Merging: An ODE Perspective
von: Lin, Lihong, et al.
Veröffentlicht: (2026)
von: Lin, Lihong, et al.
Veröffentlicht: (2026)
A Hybrid Framework for Spatial Interpolation: Merging Data-driven with Domain Knowledge
von: Zhang, Cong, et al.
Veröffentlicht: (2024)
von: Zhang, Cong, et al.
Veröffentlicht: (2024)
PSO-Merging: Merging Models Based on Particle Swarm Optimization
von: Zhang, Kehao, et al.
Veröffentlicht: (2025)
von: Zhang, Kehao, et al.
Veröffentlicht: (2025)
Optimal Brain Iterative Merging: Mitigating Interference in LLM Merging
von: Wang, Zhixiang, et al.
Veröffentlicht: (2025)
von: Wang, Zhixiang, et al.
Veröffentlicht: (2025)
SimMerge: Learning to Select Merge Operators from Similarity Signals
von: Bolton, Oliver, et al.
Veröffentlicht: (2026)
von: Bolton, Oliver, et al.
Veröffentlicht: (2026)
REFORMER: A ChatGPT-Driven Data Synthesis Framework Elevating Text-to-SQL Models
von: Liu, Shenyang, et al.
Veröffentlicht: (2025)
von: Liu, Shenyang, et al.
Veröffentlicht: (2025)
Large EEG-U-Transformer for Time-Step Level Detection Without Pre-Training
von: Wu, Kerui, et al.
Veröffentlicht: (2025)
von: Wu, Kerui, et al.
Veröffentlicht: (2025)
Superpose Task-specific Features for Model Merging
von: Qiu, Haiquan, et al.
Veröffentlicht: (2025)
von: Qiu, Haiquan, et al.
Veröffentlicht: (2025)
Cyborg Data: Merging Human with AI Generated Training Data
von: North, Kai, et al.
Veröffentlicht: (2025)
von: North, Kai, et al.
Veröffentlicht: (2025)
Bayesian Federated Learning for Continual Training
von: Milasheuski, Usevalad, et al.
Veröffentlicht: (2025)
von: Milasheuski, Usevalad, et al.
Veröffentlicht: (2025)
MIN-Merging: Merge the Important Neurons for Model Merging
von: Liang, Yunfei
Veröffentlicht: (2025)
von: Liang, Yunfei
Veröffentlicht: (2025)
DynamicLight: Two-Stage Dynamic Traffic Signal Timing
von: Zhang, Liang, et al.
Veröffentlicht: (2022)
von: Zhang, Liang, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
SplitLoRA: Balancing Stability and Plasticity in Continual Learning Through Gradient Space Splitting
von: Qiu, Haomiao, et al.
Veröffentlicht: (2025) -
Benchmarking Post-Training Quantization in LLMs: Comprehensive Taxonomy, Unified Evaluation, and Comparative Analysis
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2025) -
GraphDancer: Training LLMs to Explore and Reason over Graphs via Two-Stage Curriculum Post-Training
von: Bai, Yuyang, et al.
Veröffentlicht: (2026) -
Limited Reference, Reliable Generation: A Two-Component Framework for Tabular Data Generation in Low-Data Regimes
von: Jiang, Mingxuan, et al.
Veröffentlicht: (2025) -
FedDRL: A Trustworthy Federated Learning Model Fusion Method Based on Staged Reinforcement Learning
von: Chen, Leiming, et al.
Veröffentlicht: (2023)