On Understanding of the Dynamics of Model Capacity in Continual Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chakraborty, Supriyo, Raghavan, Krishnan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Training Dynamics Underlying Language Model Scaling Laws: Loss Deceleration and Zero-Sum Learning
von: Mircea, Andrei, et al.
Veröffentlicht: (2025)
von: Mircea, Andrei, et al.
Veröffentlicht: (2025)
The Impossibility of Inverse Permutation Learning in Transformer Models
von: Alur, Rohan, et al.
Veröffentlicht: (2025)
von: Alur, Rohan, et al.
Veröffentlicht: (2025)
Optimizing Reasoning Efficiency through Prompt Difficulty Prediction
von: Zhao, Bo, et al.
Veröffentlicht: (2025)
von: Zhao, Bo, et al.
Veröffentlicht: (2025)
Dense Backpropagation Improves Training for Sparse Mixture-of-Experts
von: Panda, Ashwinee, et al.
Veröffentlicht: (2025)
von: Panda, Ashwinee, et al.
Veröffentlicht: (2025)
SPEAR-MM: Selective Parameter Evaluation and Restoration via Model Merging for Efficient Financial LLM Adaptation
von: Kapusuzoglu, Berkcan, et al.
Veröffentlicht: (2025)
von: Kapusuzoglu, Berkcan, et al.
Veröffentlicht: (2025)
Self-Controlled Dynamic Expansion Model for Continual Learning
von: Wu, Runqing, et al.
Veröffentlicht: (2025)
von: Wu, Runqing, et al.
Veröffentlicht: (2025)
The Effect of Architecture During Continual Learning
von: Hahn, Allyson, et al.
Veröffentlicht: (2026)
von: Hahn, Allyson, et al.
Veröffentlicht: (2026)
Capacity-Constrained Continual Learning
von: Wen, Zheng, et al.
Veröffentlicht: (2025)
von: Wen, Zheng, et al.
Veröffentlicht: (2025)
LLM Surgery: Efficient Knowledge Unlearning and Editing in Large Language Models
von: Veldanda, Akshaj Kumar, et al.
Veröffentlicht: (2024)
von: Veldanda, Akshaj Kumar, et al.
Veröffentlicht: (2024)
Understanding the Dynamics of Demonstration Conflict in In-Context Learning
von: Jiao, Difan, et al.
Veröffentlicht: (2026)
von: Jiao, Difan, et al.
Veröffentlicht: (2026)
Understanding the Learning Dynamics of Alignment with Human Feedback
von: Im, Shawn, et al.
Veröffentlicht: (2024)
von: Im, Shawn, et al.
Veröffentlicht: (2024)
Learning Regularizers: Learning Optimizers that can Regularize
von: Sahoo, Suraj Kumar, et al.
Veröffentlicht: (2025)
von: Sahoo, Suraj Kumar, et al.
Veröffentlicht: (2025)
Dynamic Graph Structure Estimation for Learning Multivariate Point Process using Spiking Neural Networks
von: Chakraborty, Biswadeep, et al.
Veröffentlicht: (2025)
von: Chakraborty, Biswadeep, et al.
Veröffentlicht: (2025)
Condensation-Concatenation Framework for Dynamic Graph Continual Learning
von: Yan, Tingxu, et al.
Veröffentlicht: (2025)
von: Yan, Tingxu, et al.
Veröffentlicht: (2025)
CLeAN: Continual Learning Adaptive Normalization in Dynamic Environments
von: Marasco, Isabella, et al.
Veröffentlicht: (2026)
von: Marasco, Isabella, et al.
Veröffentlicht: (2026)
An Effective Dynamic Gradient Calibration Method for Continual Learning
von: Lin, Weichen, et al.
Veröffentlicht: (2024)
von: Lin, Weichen, et al.
Veröffentlicht: (2024)
Calibration of Continual Learning Models
von: Li, Lanpei, et al.
Veröffentlicht: (2024)
von: Li, Lanpei, et al.
Veröffentlicht: (2024)
Divide And Conquer: Learning Chaotic Dynamical Systems With Multistep Penalty Neural Ordinary Differential Equations
von: Chakraborty, Dibyajyoti, et al.
Veröffentlicht: (2024)
von: Chakraborty, Dibyajyoti, et al.
Veröffentlicht: (2024)
SPARC: Subspace-Aware Prompt Adaptation for Robust Continual Learning in LLMs
von: Jayasuriya, Dinithi, et al.
Veröffentlicht: (2025)
von: Jayasuriya, Dinithi, et al.
Veröffentlicht: (2025)
DESIRE: Dynamic Knowledge Consolidation for Rehearsal-Free Continual Learning
von: Guo, Haiyang, et al.
Veröffentlicht: (2024)
von: Guo, Haiyang, et al.
Veröffentlicht: (2024)
Continual Learning for Adaptable Car-Following in Dynamic Traffic Environments
von: Chen, Xianda, et al.
Veröffentlicht: (2024)
von: Chen, Xianda, et al.
Veröffentlicht: (2024)
Information as Structural Alignment: A Dynamical Theory of Continual Learning
von: Negulescu, Radu
Veröffentlicht: (2026)
von: Negulescu, Radu
Veröffentlicht: (2026)
Understanding Simplicity Bias towards Compositional Mappings via Learning Dynamics
von: Ren, Yi, et al.
Veröffentlicht: (2024)
von: Ren, Yi, et al.
Veröffentlicht: (2024)
Learning to Route for Dynamic Adapter Composition in Continual Learning with Language Models
von: Araujo, Vladimir, et al.
Veröffentlicht: (2024)
von: Araujo, Vladimir, et al.
Veröffentlicht: (2024)
Recasting Continual Learning as Sequence Modeling
von: Lee, Soochan, et al.
Veröffentlicht: (2023)
von: Lee, Soochan, et al.
Veröffentlicht: (2023)
Energy-Based Models for Continual Learning
von: Li, Shuang, et al.
Veröffentlicht: (2020)
von: Li, Shuang, et al.
Veröffentlicht: (2020)
Measurement Scheduling for ICU Patients with Offline Reinforcement Learning
von: Ji, Zongliang, et al.
Veröffentlicht: (2024)
von: Ji, Zongliang, et al.
Veröffentlicht: (2024)
Noise-Tolerant Coreset-Based Class Incremental Continual Learning
von: Mucllari, Edison, et al.
Veröffentlicht: (2025)
von: Mucllari, Edison, et al.
Veröffentlicht: (2025)
Goal Discovery with Causal Capacity for Efficient Reinforcement Learning
von: Yu, Yan, et al.
Veröffentlicht: (2025)
von: Yu, Yan, et al.
Veröffentlicht: (2025)
Learning Dynamics in Continual Pre-Training for Large Language Models
von: Wang, Xingjin, et al.
Veröffentlicht: (2025)
von: Wang, Xingjin, et al.
Veröffentlicht: (2025)
Fast and Featureless Node Representation Learning with Partial Pairwise Supervision
von: Chakraborty, Sujan, et al.
Veröffentlicht: (2026)
von: Chakraborty, Sujan, et al.
Veröffentlicht: (2026)
Dynamic Continual Learning: Harnessing Parameter Uncertainty for Improved Network Adaptation
von: Angelini, Christopher, et al.
Veröffentlicht: (2025)
von: Angelini, Christopher, et al.
Veröffentlicht: (2025)
Leveraging Parameter Space Symmetries for Reasoning Skill Transfer in LLMs
von: Horoi, Stefan, et al.
Veröffentlicht: (2025)
von: Horoi, Stefan, et al.
Veröffentlicht: (2025)
Is One Layer Enough? Understanding Inference Dynamics in Tabular Foundation Models
von: Balef, Amir Rezaei, et al.
Veröffentlicht: (2026)
von: Balef, Amir Rezaei, et al.
Veröffentlicht: (2026)
Embeddings to Diagnosis: Latent Fragility under Agentic Perturbations in Clinical LLMs
von: Vijayaraj, Raj Krishnan
Veröffentlicht: (2025)
von: Vijayaraj, Raj Krishnan
Veröffentlicht: (2025)
A Dynamical Systems-Inspired Pruning Strategy for Addressing Oversmoothing in Graph Neural Networks
von: Chakraborty, Biswadeep, et al.
Veröffentlicht: (2024)
von: Chakraborty, Biswadeep, et al.
Veröffentlicht: (2024)
Pareto Continual Learning: Preference-Conditioned Learning and Adaption for Dynamic Stability-Plasticity Trade-off
von: Lai, Song, et al.
Veröffentlicht: (2025)
von: Lai, Song, et al.
Veröffentlicht: (2025)
On Token's Dilemma: Dynamic MoE with Drift-Aware Token Assignment for Continual Learning of Large Vision Language Models
von: Zhao, Chongyang, et al.
Veröffentlicht: (2026)
von: Zhao, Chongyang, et al.
Veröffentlicht: (2026)
STX-Search: Explanation Search for Continuous Dynamic Spatio-Temporal Models
von: Anwar, Saif, et al.
Veröffentlicht: (2025)
von: Anwar, Saif, et al.
Veröffentlicht: (2025)
ED2: Environment Dynamics Decomposition World Models for Continuous Control
von: Hao, Jianye, et al.
Veröffentlicht: (2021)
von: Hao, Jianye, et al.
Veröffentlicht: (2021)
Ähnliche Einträge
-
Training Dynamics Underlying Language Model Scaling Laws: Loss Deceleration and Zero-Sum Learning
von: Mircea, Andrei, et al.
Veröffentlicht: (2025) -
The Impossibility of Inverse Permutation Learning in Transformer Models
von: Alur, Rohan, et al.
Veröffentlicht: (2025) -
Optimizing Reasoning Efficiency through Prompt Difficulty Prediction
von: Zhao, Bo, et al.
Veröffentlicht: (2025) -
Dense Backpropagation Improves Training for Sparse Mixture-of-Experts
von: Panda, Ashwinee, et al.
Veröffentlicht: (2025) -
SPEAR-MM: Selective Parameter Evaluation and Restoration via Model Merging for Efficient Financial LLM Adaptation
von: Kapusuzoglu, Berkcan, et al.
Veröffentlicht: (2025)