Hidden Failure Modes of Gradient Modification under Adam in Continual Learning, and Adaptive Decoupled Moment Routing as a Repair
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, Yuelin, Yu, Zhenbo, Cheng, Zhengxue, Liu, Wei, Song, Li |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Maximizing Rollout Informativeness under a Fixed Budget: A Submodular View of Tree Search for Tool-Use Agentic Reinforcement Learning
by: Hu, Yuelin, et al.
Published: (2026)
by: Hu, Yuelin, et al.
Published: (2026)
AuditRepairBench: A Paired-Execution Trace Corpus for Evaluator-Channel Ranking Instability in Agent Repair
by: Hu, Yuelin, et al.
Published: (2026)
by: Hu, Yuelin, et al.
Published: (2026)
Predictable Gradient Manifolds in Deep Learning: Temporal Path-Length and Intrinsic Rank as a Complexity Regime
by: Calvo, Anherutowa
Published: (2026)
by: Calvo, Anherutowa
Published: (2026)
Intervention-Assisted Policy Gradient Methods for Online Stochastic Queuing Network Optimization: Technical Report
by: Wigmore, Jerrod, et al.
Published: (2024)
by: Wigmore, Jerrod, et al.
Published: (2024)
Recurrent Memory-Augmented Transformers with Chunked Attention for Long-Context Language Modeling
by: Kashyap, Ankit
Published: (2025)
by: Kashyap, Ankit
Published: (2025)
StepScorer: Accelerating Reinforcement Learning with Step-wise Scoring and Psychological Regret Modeling
by: Xu, Zhe
Published: (2026)
by: Xu, Zhe
Published: (2026)
4OPS: Structural Difficulty Modeling in Integer Arithmetic Puzzles
by: Zeytuncu, Yunus E.
Published: (2026)
by: Zeytuncu, Yunus E.
Published: (2026)
Unified Unbiased Variance Estimation for Maximum Mean Discrepancy: Robust Finite-Sample Performance with Imbalanced Data and Exact Acceleration under Null and Alternative Hypotheses
by: Zhong, Shijie, et al.
Published: (2026)
by: Zhong, Shijie, et al.
Published: (2026)
Learning Nonlinear Regime Transitions via Semi-Parametric State-Space Models
by: Hiremath, Prakul Sunil
Published: (2026)
by: Hiremath, Prakul Sunil
Published: (2026)
Causality-Driven Neural Network Repair: Challenges and Opportunities
by: Vares, Fatemeh, et al.
Published: (2025)
by: Vares, Fatemeh, et al.
Published: (2025)
Grammatically-Guided Sparse Attention for Efficient and Interpretable Transformers
by: Pratyush, Spandan
Published: (2026)
by: Pratyush, Spandan
Published: (2026)
Time-to-Injury Forecasting in Elite Female Football: A DeepHit Survival Approach
by: Catterall, Victoria, et al.
Published: (2026)
by: Catterall, Victoria, et al.
Published: (2026)
Emotion-Inspired Learning Signals (EILS): A Homeostatic Framework for Adaptive Autonomous Agents
by: Tiwari, Dhruv
Published: (2025)
by: Tiwari, Dhruv
Published: (2025)
A Prescriptive Framework for Determining Optimal Days for Short-Term Traffic Counts
by: Mukwaya, Arthur, et al.
Published: (2025)
by: Mukwaya, Arthur, et al.
Published: (2025)
CBR -- Boosting Adaptive Classification By Retrieval of Encrypted Network Traffic with Out-of-distribution
by: Lukach, Amir, et al.
Published: (2024)
by: Lukach, Amir, et al.
Published: (2024)
On the approximation capability of GNNs in node classification/regression tasks
by: D'Inverno, Giuseppe Alessio, et al.
Published: (2021)
by: D'Inverno, Giuseppe Alessio, et al.
Published: (2021)
TRIM: Achieving Extreme Sparsity with Targeted Row-wise Iterative Metric-driven Pruning
by: Beck, Florentin, et al.
Published: (2025)
by: Beck, Florentin, et al.
Published: (2025)
The Architecture of Errors: From Universal Impossibility to Patch-Local LLM Reliability
by: Arbuzov, Mikhail L., et al.
Published: (2026)
by: Arbuzov, Mikhail L., et al.
Published: (2026)
Scattered Forest Search: Smarter Code Space Exploration with LLMs
by: Light, Jonathan, et al.
Published: (2024)
by: Light, Jonathan, et al.
Published: (2024)
A Multidisciplinary Approach to Telegram Data Analysis
by: Varbanov, Velizar, et al.
Published: (2024)
by: Varbanov, Velizar, et al.
Published: (2024)
Alignment Verifiability in Large Language Models: Normative Indistinguishability under Behavioral Evaluation
by: Santos-Grueiro, Igor
Published: (2026)
by: Santos-Grueiro, Igor
Published: (2026)
Reinforcement Learning Controlled Adaptive PSO for Task Offloading in IIoT Edge Computing
by: Perera, Minod, et al.
Published: (2025)
by: Perera, Minod, et al.
Published: (2025)
Learning Real-Life Approval Elections
by: Faliszewski, Piotr, et al.
Published: (2026)
by: Faliszewski, Piotr, et al.
Published: (2026)
UniPROT: Uniform Prototype Selection via Partial Optimal Transport with Submodular Guarantees
by: Chanda, Prateek, et al.
Published: (2026)
by: Chanda, Prateek, et al.
Published: (2026)
Sparse Knowledge Distillation: A Mathematical Framework for Probability-Domain Temperature Scaling and Multi-Stage Compression
by: Flouro, Aaron R., et al.
Published: (2026)
by: Flouro, Aaron R., et al.
Published: (2026)
Recursive Meta-Distillation: An Axiomatic Framework for Iterative Knowledge Refinement
by: Flouro, Aaron R., et al.
Published: (2026)
by: Flouro, Aaron R., et al.
Published: (2026)
Primal-Dual Sample Complexity Bounds for Constrained Markov Decision Processes with Multiple Constraints
by: Buckley, Max, et al.
Published: (2025)
by: Buckley, Max, et al.
Published: (2025)
When Do Early-Exit Networks Generalize? A PAC-Bayesian Theory of Adaptive Depth
by: Guo, Dongxin, et al.
Published: (2026)
by: Guo, Dongxin, et al.
Published: (2026)
CoupleEvo: Evolving Heuristics for Coupled Optimization Problems Using Large Language Models
by: Bömer, Thomas, et al.
Published: (2026)
by: Bömer, Thomas, et al.
Published: (2026)
Explaining Bayesian Optimization by Shapley Values Facilitates Human-AI Collaboration
by: Rodemann, Julian, et al.
Published: (2024)
by: Rodemann, Julian, et al.
Published: (2024)
Exploring LLM-based Verilog Code Generation with Data-Efficient Fine-Tuning and Testbench Automation
by: Chen, Mu-Chi, et al.
Published: (2026)
by: Chen, Mu-Chi, et al.
Published: (2026)
Optimizing Job Allocation using Reinforcement Learning with Graph Neural Networks
by: Quaedvlieg, Lars C. P. M.
Published: (2025)
by: Quaedvlieg, Lars C. P. M.
Published: (2025)
DELTA: Variational Disentangled Learning for Privacy-Preserving Data Reprogramming
by: Malarkkan, Arun Vignesh, et al.
Published: (2025)
by: Malarkkan, Arun Vignesh, et al.
Published: (2025)
Enhanced and Efficient Reasoning in Large Learning Models
by: Valiant, Leslie G.
Published: (2026)
by: Valiant, Leslie G.
Published: (2026)
Vision-Guided Iterative Refinement for Frontend Code Generation
by: Sansford, Hannah, et al.
Published: (2026)
by: Sansford, Hannah, et al.
Published: (2026)
The Hidden Attention of Mamba Models
by: Ali, Ameen, et al.
Published: (2024)
by: Ali, Ameen, et al.
Published: (2024)
Graph Coloring for Multi-Task Learning
by: Patapati, Santosh
Published: (2025)
by: Patapati, Santosh
Published: (2025)
LLM Agents for Generating Microservice-based Applications: how complex is your specification?
by: Yellin, Daniel M.
Published: (2025)
by: Yellin, Daniel M.
Published: (2025)
Diffusion-MPC in Discrete Domains: Feasibility Constraints, Horizon Effects, and Critic Alignment: Case study with Tetris
by: Wang, Haochuan Kevin
Published: (2026)
by: Wang, Haochuan Kevin
Published: (2026)
Chain of Unit-Physics: A Primitive-Centric Approach to Scientific Code Synthesis
by: Sharma, Vansh, et al.
Published: (2025)
by: Sharma, Vansh, et al.
Published: (2025)
Similar Items
-
Maximizing Rollout Informativeness under a Fixed Budget: A Submodular View of Tree Search for Tool-Use Agentic Reinforcement Learning
by: Hu, Yuelin, et al.
Published: (2026) -
AuditRepairBench: A Paired-Execution Trace Corpus for Evaluator-Channel Ranking Instability in Agent Repair
by: Hu, Yuelin, et al.
Published: (2026) -
Predictable Gradient Manifolds in Deep Learning: Temporal Path-Length and Intrinsic Rank as a Complexity Regime
by: Calvo, Anherutowa
Published: (2026) -
Intervention-Assisted Policy Gradient Methods for Online Stochastic Queuing Network Optimization: Technical Report
by: Wigmore, Jerrod, et al.
Published: (2024) -
Recurrent Memory-Augmented Transformers with Chunked Attention for Long-Context Language Modeling
by: Kashyap, Ankit
Published: (2025)