FOAM: Frequency and Operator Error-Based Adaptive Damping Method for Reducing Staleness-Oriented Error for Shampoo
Fuente:
arXiv
Saved in:
| Main Authors: | Nam, Kyunghun, Ahn, Sumyeong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Purifying Shampoo: Investigating Shampoo's Heuristics by Decomposing its Preconditioner
by: Eschenhagen, Runa, et al.
Published: (2025)
by: Eschenhagen, Runa, et al.
Published: (2025)
Latent Veracity Inference for Identifying Errors in Stepwise Reasoning
by: Kim, Minsu, et al.
Published: (2025)
by: Kim, Minsu, et al.
Published: (2025)
SOAP: Improving and Stabilizing Shampoo using Adam
by: Vyas, Nikhil, et al.
Published: (2024)
by: Vyas, Nikhil, et al.
Published: (2024)
Enhancing Stability in Training Conditional Generative Adversarial Networks via Selective Data Matching
by: Kong, Kyeongbo, et al.
Published: (2021)
by: Kong, Kyeongbo, et al.
Published: (2021)
4-bit Shampoo for Memory-Efficient Network Training
by: Wang, Sike, et al.
Published: (2024)
by: Wang, Sike, et al.
Published: (2024)
Clarifying Shampoo: Adapting Spectral Descent to Stochasticity and the Parameter Trajectory
by: Eschenhagen, Runa, et al.
Published: (2026)
by: Eschenhagen, Runa, et al.
Published: (2026)
Pro-KLShampoo: Projected KL-Shampoo with Whitening Recovered by Orthogonalization
by: Sun, Ruotong, et al.
Published: (2026)
by: Sun, Ruotong, et al.
Published: (2026)
Symmetric Q-learning: Reducing Skewness of Bellman Error in Online Reinforcement Learning
by: Omura, Motoki, et al.
Published: (2024)
by: Omura, Motoki, et al.
Published: (2024)
Error Diversity Matters: An Error-Resistant Ensemble Method for Unsupervised Dependency Parsing
by: Shayegh, Behzad, et al.
Published: (2024)
by: Shayegh, Behzad, et al.
Published: (2024)
TextBFGS: A Case-Based Reasoning Approach to Code Optimization via Error-Operator Retrieval
by: Zhang, Zizheng, et al.
Published: (2026)
by: Zhang, Zizheng, et al.
Published: (2026)
Towards Reducing Diagnostic Errors with Interpretable Risk Prediction
by: McInerney, Denis Jered, et al.
Published: (2024)
by: McInerney, Denis Jered, et al.
Published: (2024)
Degree of Staleness-Aware Data Updating in Federated Learning
by: Liu, Tao, et al.
Published: (2025)
by: Liu, Tao, et al.
Published: (2025)
FedPSA: Modeling Behavioral Staleness in Asynchronous Federated Learning
by: Lu, Chaoyi, et al.
Published: (2026)
by: Lu, Chaoyi, et al.
Published: (2026)
FedStale: leveraging stale client updates in federated learning
by: Rodio, Angelo, et al.
Published: (2024)
by: Rodio, Angelo, et al.
Published: (2024)
Bellman Error Centering
by: Chen, Xingguo, et al.
Published: (2025)
by: Chen, Xingguo, et al.
Published: (2025)
FOAM: Blocked State Folding for Memory-Efficient LLM Training
by: Wen, Ziqing, et al.
Published: (2025)
by: Wen, Ziqing, et al.
Published: (2025)
Error Adjustment Based on Spatiotemporal Correlation Fusion for Traffic Forecasting
by: Liu, Fuqiang, et al.
Published: (2025)
by: Liu, Fuqiang, et al.
Published: (2025)
Revisiting Gradient Staleness: Evaluating Distance Metrics for Asynchronous Federated Learning Aggregation
by: Wilhelm, Patrick, et al.
Published: (2026)
by: Wilhelm, Patrick, et al.
Published: (2026)
Learning from Yesterday's Error: An Efficient Online Learning Method for Traffic Demand Prediction
by: Huang, Xiannan, et al.
Published: (2026)
by: Huang, Xiannan, et al.
Published: (2026)
SEF: A Method for Computing Prediction Intervals by Shifting the Error Function in Neural Networks
by: Aretos, E. V., et al.
Published: (2024)
by: Aretos, E. V., et al.
Published: (2024)
Bounding-Box Inference for Error-Aware Model-Based Reinforcement Learning
by: Talvitie, Erin J., et al.
Published: (2024)
by: Talvitie, Erin J., et al.
Published: (2024)
Error Analysis of Shapley Value-Based Model Explanations: An Informative Perspective
by: Zhao, Ningsheng, et al.
Published: (2024)
by: Zhao, Ningsheng, et al.
Published: (2024)
Learning the Error Patterns of Language Models
by: Kim, Jinwoo, et al.
Published: (2026)
by: Kim, Jinwoo, et al.
Published: (2026)
Score Based Error Correcting Code Decoder
by: Helvits, Alon, et al.
Published: (2026)
by: Helvits, Alon, et al.
Published: (2026)
RILQ: Rank-Insensitive LoRA-based Quantization Error Compensation for Boosting 2-bit Large Language Model Accuracy
by: Lee, Geonho, et al.
Published: (2024)
by: Lee, Geonho, et al.
Published: (2024)
Self-Error-Instruct: Generalizing from Errors for LLMs Mathematical Reasoning
by: Yu, Erxin, et al.
Published: (2025)
by: Yu, Erxin, et al.
Published: (2025)
On the Error-Correcting Effects of Stochasticity in Discrete Diffusion
by: Yuan, William, et al.
Published: (2026)
by: Yuan, William, et al.
Published: (2026)
Error-Driven Prompt Optimization for Arithmetic Reasoning
by: Pándy, Árpád, et al.
Published: (2025)
by: Pándy, Árpád, et al.
Published: (2025)
Z-Error Loss for Training Neural Networks
by: Godin, Guillaume
Published: (2025)
by: Godin, Guillaume
Published: (2025)
On Minimizing Adversarial Counterfactual Error in Adversarial RL
by: Belaire, Roman, et al.
Published: (2024)
by: Belaire, Roman, et al.
Published: (2024)
Building Trust in PINNs: Error Estimation through Finite Difference Methods
by: Krasowski, Aleksander, et al.
Published: (2026)
by: Krasowski, Aleksander, et al.
Published: (2026)
Quantifying Calibration Error in Neural Networks Through Evidence-Based Theory
by: Ouattara, Koffi Ismael, et al.
Published: (2024)
by: Ouattara, Koffi Ismael, et al.
Published: (2024)
Prosperity before Collapse: How Far Can Off-Policy RL Reach with Stale Data on LLMs?
by: Zheng, Haizhong, et al.
Published: (2025)
by: Zheng, Haizhong, et al.
Published: (2025)
Analyzing Error Sources in Global Feature Effect Estimation
by: Heiß, Timo, et al.
Published: (2026)
by: Heiß, Timo, et al.
Published: (2026)
Reward Learning through Ranking Mean Squared Error
by: Kharyal, Chaitanya, et al.
Published: (2026)
by: Kharyal, Chaitanya, et al.
Published: (2026)
Dissecting Quantization Error: A Concentration-Alignment Perspective
by: Federici, Marco, et al.
Published: (2026)
by: Federici, Marco, et al.
Published: (2026)
Improving Label Error Detection and Elimination with Uncertainty Quantification
by: Jakubik, Johannes, et al.
Published: (2024)
by: Jakubik, Johannes, et al.
Published: (2024)
On the Collapse Errors Induced by the Deterministic Sampler for Diffusion Models
by: Zhang, Yi, et al.
Published: (2025)
by: Zhang, Yi, et al.
Published: (2025)
Beyond the Norms: Detecting Prediction Errors in Regression Models
by: Altieri, Andres, et al.
Published: (2024)
by: Altieri, Andres, et al.
Published: (2024)
Bounding the Worst-class Error: A Boosting Approach
by: Saito, Yuya, et al.
Published: (2023)
by: Saito, Yuya, et al.
Published: (2023)
Similar Items
-
Purifying Shampoo: Investigating Shampoo's Heuristics by Decomposing its Preconditioner
by: Eschenhagen, Runa, et al.
Published: (2025) -
Latent Veracity Inference for Identifying Errors in Stepwise Reasoning
by: Kim, Minsu, et al.
Published: (2025) -
SOAP: Improving and Stabilizing Shampoo using Adam
by: Vyas, Nikhil, et al.
Published: (2024) -
Enhancing Stability in Training Conditional Generative Adversarial Networks via Selective Data Matching
by: Kong, Kyeongbo, et al.
Published: (2021) -
4-bit Shampoo for Memory-Efficient Network Training
by: Wang, Sike, et al.
Published: (2024)