Incremental Gauss-Newton Descent for Machine Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Korbit, Mikalai, Zanon, Mario |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fast Gauss-Newton for Multiclass Cross-Entropy
by: Korbit, Mikalai, et al.
Published: (2026)
by: Korbit, Mikalai, et al.
Published: (2026)
Exact Gauss-Newton Optimization for Training Deep Neural Networks
by: Korbit, Mikalai, et al.
Published: (2024)
by: Korbit, Mikalai, et al.
Published: (2024)
Second-Order, First-Class: A Composable Stack for Curvature-Aware Training
by: Korbit, Mikalai, et al.
Published: (2026)
by: Korbit, Mikalai, et al.
Published: (2026)
Gauss-Newton Natural Gradient Descent for Shape Learning
by: King, James, et al.
Published: (2026)
by: King, James, et al.
Published: (2026)
Incremental Gauss--Newton Methods with Superlinear Convergence Rates
by: Zhou, Zhiling, et al.
Published: (2024)
by: Zhou, Zhiling, et al.
Published: (2024)
Error whitening: Why Gauss-Newton outperforms Newton
by: McKay, Maricela Best, et al.
Published: (2026)
by: McKay, Maricela Best, et al.
Published: (2026)
Gauss-Newton Temporal Difference Learning with Nonlinear Function Approximation
by: Ke, Zhifa, et al.
Published: (2023)
by: Ke, Zhifa, et al.
Published: (2023)
Regularized Gauss-Newton for Optimizing Overparameterized Neural Networks
by: Adeoye, Adeyemi D., et al.
Published: (2024)
by: Adeoye, Adeyemi D., et al.
Published: (2024)
Sharpened Lazy Incremental Quasi-Newton Method
by: Lahoti, Aakash, et al.
Published: (2023)
by: Lahoti, Aakash, et al.
Published: (2023)
An Adaptive Stochastic Gradient Method with Non-negative Gauss-Newton Stepsizes
by: Orvieto, Antonio, et al.
Published: (2024)
by: Orvieto, Antonio, et al.
Published: (2024)
Incremental Quasi-Newton Methods with Faster Superlinear Convergence Rates
by: Liu, Zhuanghua, et al.
Published: (2024)
by: Liu, Zhuanghua, et al.
Published: (2024)
Non-Asymptotic Optimization and Generalization Bounds for Stochastic Gauss-Newton in Overparameterized Models
by: Cayci, Semih
Published: (2025)
by: Cayci, Semih
Published: (2025)
Preconditioned Norms: A Unified Framework for Steepest Descent, Quasi-Newton and Adaptive Methods
by: Veprikov, Andrey, et al.
Published: (2025)
by: Veprikov, Andrey, et al.
Published: (2025)
Incremental Gradient Descent with Small Epoch Counts is Surprisingly Slow on Ill-Conditioned Problems
by: Kim, Yujun, et al.
Published: (2025)
by: Kim, Yujun, et al.
Published: (2025)
Quadratic Gradient: A Unified Framework Bridging Gradient Descent and Newton-Type Methods by Synthesizing Hessians and Gradients
by: Chiang, John
Published: (2022)
by: Chiang, John
Published: (2022)
A Riemannian Optimization Perspective of the Gauss-Newton Method for Feedforward Neural Networks
by: Cayci, Semih
Published: (2024)
by: Cayci, Semih
Published: (2024)
A Gauss-Newton Approach for Min-Max Optimization in Generative Adversarial Networks
by: Mishra, Neel, et al.
Published: (2024)
by: Mishra, Neel, et al.
Published: (2024)
RGNMR: A Gauss-Newton method for robust matrix completion with theoretical guarantees
by: Laufer, Eilon Vaknin, et al.
Published: (2025)
by: Laufer, Eilon Vaknin, et al.
Published: (2025)
Reconstructing Physics-Informed Machine Learning for Traffic Flow Modeling: a Multi-Gradient Descent and Pareto Learning Approach
by: Lei, Yuan-Zheng, et al.
Published: (2025)
by: Lei, Yuan-Zheng, et al.
Published: (2025)
Stochastic Adaptive Gradient Descent Without Descent
by: Aujol, Jean-François, et al.
Published: (2025)
by: Aujol, Jean-François, et al.
Published: (2025)
Learning Provably Improves the Convergence of Gradient Descent
by: Song, Qingyu, et al.
Published: (2025)
by: Song, Qingyu, et al.
Published: (2025)
Enhancing Fractional Gradient Descent with Learned Optimizers
by: Sobotka, Jan, et al.
Published: (2025)
by: Sobotka, Jan, et al.
Published: (2025)
Incremental Learning of Sparse Attention Patterns in Transformers
by: Yüksel, Oğuz Kaan, et al.
Published: (2026)
by: Yüksel, Oğuz Kaan, et al.
Published: (2026)
A Mirror Descent Perspective of Smoothed Sign Descent
by: Wang, Shuyang, et al.
Published: (2024)
by: Wang, Shuyang, et al.
Published: (2024)
On the Convergence of Gradient Descent on Learning Transformers with Residual Connections
by: Qin, Zhen, et al.
Published: (2025)
by: Qin, Zhen, et al.
Published: (2025)
Provable and Practical Online Learning Rate Adaptation with Hypergradient Descent
by: Chu, Ya-Chi, et al.
Published: (2025)
by: Chu, Ya-Chi, et al.
Published: (2025)
RESIST: Resilient Decentralized Learning Using Consensus Gradient Descent
by: Fang, Cheng, et al.
Published: (2025)
by: Fang, Cheng, et al.
Published: (2025)
The Method of Infinite Descent
by: Batley, Reza T., et al.
Published: (2025)
by: Batley, Reza T., et al.
Published: (2025)
Corner Gradient Descent
by: Yarotsky, Dmitry
Published: (2025)
by: Yarotsky, Dmitry
Published: (2025)
Random Function Descent
by: Benning, Felix, et al.
Published: (2023)
by: Benning, Felix, et al.
Published: (2023)
Last Iterate Convergence of Incremental Methods and Applications in Continual Learning
by: Cai, Xufeng, et al.
Published: (2024)
by: Cai, Xufeng, et al.
Published: (2024)
Value Mirror Descent for Reinforcement Learning
by: Jia, Zhichao, et al.
Published: (2026)
by: Jia, Zhichao, et al.
Published: (2026)
When Descent Is Too Stable: Event-Triggered Hamiltonian Learning to Optimize
by: Wang, Yi, et al.
Published: (2026)
by: Wang, Yi, et al.
Published: (2026)
Convergence of Muon with Newton-Schulz
by: Kim, Gyu Yeol, et al.
Published: (2026)
by: Kim, Gyu Yeol, et al.
Published: (2026)
A Local Polyak-Lojasiewicz and Descent Lemma of Gradient Descent For Overparametrized Linear Models
by: Xu, Ziqing, et al.
Published: (2025)
by: Xu, Ziqing, et al.
Published: (2025)
Faster Acceleration for Steepest Descent
by: Bai, Cedar Site, et al.
Published: (2024)
by: Bai, Cedar Site, et al.
Published: (2024)
Adaptive Conditional Gradient Descent
by: Khademi, Abbas, et al.
Published: (2025)
by: Khademi, Abbas, et al.
Published: (2025)
Parameter-free Mirror Descent
by: Jacobsen, Andrew, et al.
Published: (2022)
by: Jacobsen, Andrew, et al.
Published: (2022)
Mirror Descent on Riemannian Manifolds
by: Jiang, Jiaxin, et al.
Published: (2026)
by: Jiang, Jiaxin, et al.
Published: (2026)
$k$-SVD with Gradient Descent
by: Jedra, Yassir, et al.
Published: (2025)
by: Jedra, Yassir, et al.
Published: (2025)
Similar Items
-
Fast Gauss-Newton for Multiclass Cross-Entropy
by: Korbit, Mikalai, et al.
Published: (2026) -
Exact Gauss-Newton Optimization for Training Deep Neural Networks
by: Korbit, Mikalai, et al.
Published: (2024) -
Second-Order, First-Class: A Composable Stack for Curvature-Aware Training
by: Korbit, Mikalai, et al.
Published: (2026) -
Gauss-Newton Natural Gradient Descent for Shape Learning
by: King, James, et al.
Published: (2026) -
Incremental Gauss--Newton Methods with Superlinear Convergence Rates
by: Zhou, Zhiling, et al.
Published: (2024)