An Effective Dynamic Gradient Calibration Method for Continual Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lin, Weichen, Chen, Jiaxiang, Huang, Ruomin, Ding, Hu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multimodal Representation Learning using Adaptive Graph Construction
von: Huang, Weichen
Veröffentlicht: (2024)
von: Huang, Weichen
Veröffentlicht: (2024)
Calibration of Continual Learning Models
von: Li, Lanpei, et al.
Veröffentlicht: (2024)
von: Li, Lanpei, et al.
Veröffentlicht: (2024)
Selective Learning: Towards Robust Calibration with Dynamic Regularization
von: Han, Zongbo, et al.
Veröffentlicht: (2024)
von: Han, Zongbo, et al.
Veröffentlicht: (2024)
Logit Dynamics in Softmax Policy Gradient Methods
von: Li, Yingru
Veröffentlicht: (2025)
von: Li, Yingru
Veröffentlicht: (2025)
ProRL: Effective Reinforcement Learning for Proactive Recommendation via Rectified Policy Gradient Estimation
von: Hou, Hongru, et al.
Veröffentlicht: (2026)
von: Hou, Hongru, et al.
Veröffentlicht: (2026)
Just-In-Time Reinforcement Learning: Continual Learning in LLM Agents Without Gradient Updates
von: Li, Yibo, et al.
Veröffentlicht: (2026)
von: Li, Yibo, et al.
Veröffentlicht: (2026)
Learning General Policies with Policy Gradient Methods
von: Ståhlberg, Simon, et al.
Veröffentlicht: (2025)
von: Ståhlberg, Simon, et al.
Veröffentlicht: (2025)
An Advantage-based Optimization Method for Reinforcement Learning in Large Action Space
von: Lin, Hai, et al.
Veröffentlicht: (2024)
von: Lin, Hai, et al.
Veröffentlicht: (2024)
Dirichlet-Based Prediction Calibration for Learning with Noisy Labels
von: Zong, Chen-Chen, et al.
Veröffentlicht: (2024)
von: Zong, Chen-Chen, et al.
Veröffentlicht: (2024)
Elastic Multi-Gradient Descent for Parallel Continual Learning
von: Lyu, Fan, et al.
Veröffentlicht: (2024)
von: Lyu, Fan, et al.
Veröffentlicht: (2024)
Self-Controlled Dynamic Expansion Model for Continual Learning
von: Wu, Runqing, et al.
Veröffentlicht: (2025)
von: Wu, Runqing, et al.
Veröffentlicht: (2025)
A Selective Learning Method for Temporal Graph Continual Learning
von: Liu, Hanmo, et al.
Veröffentlicht: (2025)
von: Liu, Hanmo, et al.
Veröffentlicht: (2025)
Policy Gradient Methods for Non-Markovian Reinforcement Learning
von: Kar, Avik, et al.
Veröffentlicht: (2026)
von: Kar, Avik, et al.
Veröffentlicht: (2026)
Fisher-Orthogonal Projected Natural Gradient Descent for Continual Learning
von: Garg, Ishir, et al.
Veröffentlicht: (2026)
von: Garg, Ishir, et al.
Veröffentlicht: (2026)
Merge then Realign: Simple and Effective Modality-Incremental Continual Learning for Multimodal LLMs
von: Zhang, Dingkun, et al.
Veröffentlicht: (2025)
von: Zhang, Dingkun, et al.
Veröffentlicht: (2025)
In-Context Learning can Perform Continual Learning Like Humans
von: Kang, Liuwang, et al.
Veröffentlicht: (2025)
von: Kang, Liuwang, et al.
Veröffentlicht: (2025)
On the Convergence of Continual Learning with Adaptive Methods
von: Han, Seungyub, et al.
Veröffentlicht: (2024)
von: Han, Seungyub, et al.
Veröffentlicht: (2024)
Enhancing Deep Learning with Optimized Gradient Descent: Bridging Numerical Methods and Neural Network Training
von: Ma, Yuhan, et al.
Veröffentlicht: (2024)
von: Ma, Yuhan, et al.
Veröffentlicht: (2024)
AttriReBoost: A Gradient-Free Propagation Optimization Method for Cold Start Mitigation in Attribute Missing Graphs
von: Li, Mengran, et al.
Veröffentlicht: (2025)
von: Li, Mengran, et al.
Veröffentlicht: (2025)
Federated Natural Policy Gradient and Actor Critic Methods for Multi-task Reinforcement Learning
von: Yang, Tong, et al.
Veröffentlicht: (2023)
von: Yang, Tong, et al.
Veröffentlicht: (2023)
Calibration and Transformation-Free Weight-Only LLMs Quantization via Dynamic Grouping
von: Zheng, Xinzhe, et al.
Veröffentlicht: (2025)
von: Zheng, Xinzhe, et al.
Veröffentlicht: (2025)
Continual Diffuser (CoD): Mastering Continual Offline Reinforcement Learning with Experience Rehearsal
von: Hu, Jifeng, et al.
Veröffentlicht: (2024)
von: Hu, Jifeng, et al.
Veröffentlicht: (2024)
MANGO: Meta-Adaptive Network Gradient Optimization for Online Continual Learning
von: Awasthi, Ankita, et al.
Veröffentlicht: (2026)
von: Awasthi, Ankita, et al.
Veröffentlicht: (2026)
Continual Learning for Adaptable Car-Following in Dynamic Traffic Environments
von: Chen, Xianda, et al.
Veröffentlicht: (2024)
von: Chen, Xianda, et al.
Veröffentlicht: (2024)
Balance Reward and Safety Optimization for Safe Reinforcement Learning: A Perspective of Gradient Manipulation
von: Gu, Shangding, et al.
Veröffentlicht: (2024)
von: Gu, Shangding, et al.
Veröffentlicht: (2024)
Attention Sinks Induce Gradient Sinks: Massive Activations as Gradient Regulators in Transformers
von: Chen, Yihong, et al.
Veröffentlicht: (2026)
von: Chen, Yihong, et al.
Veröffentlicht: (2026)
Integrated Gradient Correlation: a Dataset-wise Attribution Method
von: Lelièvre, Pierre, et al.
Veröffentlicht: (2024)
von: Lelièvre, Pierre, et al.
Veröffentlicht: (2024)
On the Global Optimality of Policy Gradient Methods in General Utility Reinforcement Learning
von: Barakat, Anas, et al.
Veröffentlicht: (2024)
von: Barakat, Anas, et al.
Veröffentlicht: (2024)
PG-Rainbow: Using Distributional Reinforcement Learning in Policy Gradient Methods
von: Jeon, WooJae, et al.
Veröffentlicht: (2024)
von: Jeon, WooJae, et al.
Veröffentlicht: (2024)
An Integrated Fusion Framework for Ensemble Learning Leveraging Gradient Boosting and Fuzzy Rule-Based Models
von: Li, Jinbo, et al.
Veröffentlicht: (2025)
von: Li, Jinbo, et al.
Veröffentlicht: (2025)
GradientStabilizer:Fix the Norm, Not the Gradient
von: Huang, Tianjin, et al.
Veröffentlicht: (2025)
von: Huang, Tianjin, et al.
Veröffentlicht: (2025)
ColA: Collaborative Adaptation with Gradient Learning
von: Diao, Enmao, et al.
Veröffentlicht: (2024)
von: Diao, Enmao, et al.
Veröffentlicht: (2024)
Reinforcement Learning via Value Gradient Flow
von: Xu, Haoran, et al.
Veröffentlicht: (2026)
von: Xu, Haoran, et al.
Veröffentlicht: (2026)
Pareto Continual Learning: Preference-Conditioned Learning and Adaption for Dynamic Stability-Plasticity Trade-off
von: Lai, Song, et al.
Veröffentlicht: (2025)
von: Lai, Song, et al.
Veröffentlicht: (2025)
Gradient Boosting Reinforcement Learning
von: Fuhrer, Benjamin, et al.
Veröffentlicht: (2024)
von: Fuhrer, Benjamin, et al.
Veröffentlicht: (2024)
Principled Curriculum Learning using Parameter Continuation Methods
von: Pathak, Harsh Nilesh, et al.
Veröffentlicht: (2025)
von: Pathak, Harsh Nilesh, et al.
Veröffentlicht: (2025)
Survey on Recent Progress of AI for Chemistry: Methods, Applications, and Opportunities
von: Ding, Hu, et al.
Veröffentlicht: (2025)
von: Ding, Hu, et al.
Veröffentlicht: (2025)
Revisiting Softmax Masking: Stop Gradient for Enhancing Stability in Replay-based Continual Learning
von: Kim, Hoyong, et al.
Veröffentlicht: (2023)
von: Kim, Hoyong, et al.
Veröffentlicht: (2023)
Calibrated Dataset Condensation for Faster Hyperparameter Search
von: Ding, Mucong, et al.
Veröffentlicht: (2024)
von: Ding, Mucong, et al.
Veröffentlicht: (2024)
On Understanding of the Dynamics of Model Capacity in Continual Learning
von: Chakraborty, Supriyo, et al.
Veröffentlicht: (2025)
von: Chakraborty, Supriyo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Multimodal Representation Learning using Adaptive Graph Construction
von: Huang, Weichen
Veröffentlicht: (2024) -
Calibration of Continual Learning Models
von: Li, Lanpei, et al.
Veröffentlicht: (2024) -
Selective Learning: Towards Robust Calibration with Dynamic Regularization
von: Han, Zongbo, et al.
Veröffentlicht: (2024) -
Logit Dynamics in Softmax Policy Gradient Methods
von: Li, Yingru
Veröffentlicht: (2025) -
ProRL: Effective Reinforcement Learning for Proactive Recommendation via Rectified Policy Gradient Estimation
von: Hou, Hongru, et al.
Veröffentlicht: (2026)