Exact Gauss-Newton Optimization for Training Deep Neural Networks
Fuente:
arXiv
Saved in:
| Main Authors: | Korbit, Mikalai, Adeoye, Adeyemi D., Bemporad, Alberto, Zanon, Mario |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fast Gauss-Newton for Multiclass Cross-Entropy
by: Korbit, Mikalai, et al.
Published: (2026)
by: Korbit, Mikalai, et al.
Published: (2026)
Incremental Gauss-Newton Descent for Machine Learning
by: Korbit, Mikalai, et al.
Published: (2024)
by: Korbit, Mikalai, et al.
Published: (2024)
Regularized Gauss-Newton for Optimizing Overparameterized Neural Networks
by: Adeoye, Adeyemi D., et al.
Published: (2024)
by: Adeoye, Adeyemi D., et al.
Published: (2024)
Second-Order, First-Class: A Composable Stack for Curvature-Aware Training
by: Korbit, Mikalai, et al.
Published: (2026)
by: Korbit, Mikalai, et al.
Published: (2026)
Self-concordant smoothing in proximal quasi-Newton algorithms for large-scale convex composite optimization
by: Adeoye, Adeyemi D., et al.
Published: (2023)
by: Adeoye, Adeyemi D., et al.
Published: (2023)
A proximal augmented Lagrangian method for nonconvex optimization with equality and inequality constraints
by: Adeoye, Adeyemi D., et al.
Published: (2025)
by: Adeoye, Adeyemi D., et al.
Published: (2025)
Exact, Tractable Gauss-Newton Optimization in Deep Reversible Architectures Reveal Poor Generalization
by: Buffelli, Davide, et al.
Published: (2024)
by: Buffelli, Davide, et al.
Published: (2024)
Theoretical characterisation of the Gauss-Newton conditioning in Neural Networks
by: Zhao, Jim, et al.
Published: (2024)
by: Zhao, Jim, et al.
Published: (2024)
Approximating Full Conformal Prediction for Neural Network Regression with Gauss-Newton Influence
by: Tailor, Dharmesh, et al.
Published: (2025)
by: Tailor, Dharmesh, et al.
Published: (2025)
Global and Preference-based Optimization with Mixed Variables using Piecewise Affine Surrogates
by: Zhu, Mengjia, et al.
Published: (2023)
by: Zhu, Mengjia, et al.
Published: (2023)
A Riemannian Optimization Perspective of the Gauss-Newton Method for Feedforward Neural Networks
by: Cayci, Semih
Published: (2024)
by: Cayci, Semih
Published: (2024)
An L-BFGS-B approach for linear and nonlinear system identification under $\ell_1$ and group-Lasso regularization
by: Bemporad, Alberto
Published: (2024)
by: Bemporad, Alberto
Published: (2024)
Gauss-Newton Unlearning for the LLM Era
by: McKinney, Lev, et al.
Published: (2026)
by: McKinney, Lev, et al.
Published: (2026)
A Gauss-Newton Approach for Min-Max Optimization in Generative Adversarial Networks
by: Mishra, Neel, et al.
Published: (2024)
by: Mishra, Neel, et al.
Published: (2024)
Error whitening: Why Gauss-Newton outperforms Newton
by: McKay, Maricela Best, et al.
Published: (2026)
by: McKay, Maricela Best, et al.
Published: (2026)
Parametric Nonconvex Optimization via Convex Surrogates
by: Wang, Renzi, et al.
Published: (2026)
by: Wang, Renzi, et al.
Published: (2026)
FINDER: Stochastic Mirroring of Noisy Quasi-Newton Search and Deep Network Training
by: Suman, Uttam, et al.
Published: (2024)
by: Suman, Uttam, et al.
Published: (2024)
Non-Asymptotic Optimization and Generalization Bounds for Stochastic Gauss-Newton in Overparameterized Models
by: Cayci, Semih
Published: (2025)
by: Cayci, Semih
Published: (2025)
The Potential of Second-Order Optimization for LLMs: A Study with Full Gauss-Newton
by: Abreu, Natalie, et al.
Published: (2025)
by: Abreu, Natalie, et al.
Published: (2025)
Efficient identification of linear, parameter-varying, and nonlinear systems with noise models
by: Bemporad, Alberto, et al.
Published: (2025)
by: Bemporad, Alberto, et al.
Published: (2025)
Deep Neural Network Training as Random Effects: An Optimization-Inference Duality
by: Yao, Minhao, et al.
Published: (2026)
by: Yao, Minhao, et al.
Published: (2026)
A Structure-Guided Gauss-Newton Method for Shallow ReLU Neural Network
by: Cai, Zhiqiang, et al.
Published: (2024)
by: Cai, Zhiqiang, et al.
Published: (2024)
An active learning method for solving competitive multi-agent decision-making and control problems
by: Fabiani, Filippo, et al.
Published: (2022)
by: Fabiani, Filippo, et al.
Published: (2022)
Incremental Gauss--Newton Methods with Superlinear Convergence Rates
by: Zhou, Zhiling, et al.
Published: (2024)
by: Zhou, Zhiling, et al.
Published: (2024)
Gauss-Newton Natural Gradient Descent for Shape Learning
by: King, James, et al.
Published: (2026)
by: King, James, et al.
Published: (2026)
Learning Low-Dimensional Embeddings for Black-Box Optimization
by: Busetto, Riccardo, et al.
Published: (2025)
by: Busetto, Riccardo, et al.
Published: (2025)
Q-Newton: Hybrid Quantum-Classical Scheduling for Accelerating Neural Network Training with Newton's Gradient Descent
by: Li, Pingzhi, et al.
Published: (2024)
by: Li, Pingzhi, et al.
Published: (2024)
Stein Variational Newton Neural Network Ensembles
by: Flöge, Klemens, et al.
Published: (2024)
by: Flöge, Klemens, et al.
Published: (2024)
ReBoot: Encrypted Training of Deep Neural Networks with CKKS Bootstrapping
by: Pirillo, Alberto, et al.
Published: (2025)
by: Pirillo, Alberto, et al.
Published: (2025)
On the Hardness of Training Deep Neural Networks Discretely
by: Doron-Arad, Ilan
Published: (2024)
by: Doron-Arad, Ilan
Published: (2024)
Gauss-Newton Temporal Difference Learning with Nonlinear Function Approximation
by: Ke, Zhifa, et al.
Published: (2023)
by: Ke, Zhifa, et al.
Published: (2023)
A Majorization-Minimization Gauss-Newton Method for 1-Bit Matrix Completion
by: Liu, Xiaoqian, et al.
Published: (2023)
by: Liu, Xiaoqian, et al.
Published: (2023)
On Newton's Method to Unlearn Neural Networks
by: Bui, Nhung, et al.
Published: (2024)
by: Bui, Nhung, et al.
Published: (2024)
Learning Morphisms with Gauss-Newton Approximation for Growing Networks
by: Lawton, Neal, et al.
Published: (2024)
by: Lawton, Neal, et al.
Published: (2024)
Training Neural Networks by Optimizing Neuron Positions
by: Erb, Laura, et al.
Published: (2025)
by: Erb, Laura, et al.
Published: (2025)
Learning Lyapunov terminal costs from data for complexity reduction in nonlinear model predictive control
by: Shokhjakhon Abdufattokhov, et al.
Published: (2024)
by: Shokhjakhon Abdufattokhov, et al.
Published: (2024)
Learning disturbance models for offset-free reference tracking
by: Krupa, Pablo, et al.
Published: (2023)
by: Krupa, Pablo, et al.
Published: (2023)
On Two-Player Scalar Discrete-Time Linear Quadratic Games
by: Cavalagli, Chiara, et al.
Published: (2026)
by: Cavalagli, Chiara, et al.
Published: (2026)
Adam or Gauss-Newton? A Comparative Study In Terms of Basis Alignment and SGD Noise
by: Liu, Bingbin, et al.
Published: (2025)
by: Liu, Bingbin, et al.
Published: (2025)
An Adaptive Stochastic Gradient Method with Non-negative Gauss-Newton Stepsizes
by: Orvieto, Antonio, et al.
Published: (2024)
by: Orvieto, Antonio, et al.
Published: (2024)
Similar Items
-
Fast Gauss-Newton for Multiclass Cross-Entropy
by: Korbit, Mikalai, et al.
Published: (2026) -
Incremental Gauss-Newton Descent for Machine Learning
by: Korbit, Mikalai, et al.
Published: (2024) -
Regularized Gauss-Newton for Optimizing Overparameterized Neural Networks
by: Adeoye, Adeyemi D., et al.
Published: (2024) -
Second-Order, First-Class: A Composable Stack for Curvature-Aware Training
by: Korbit, Mikalai, et al.
Published: (2026) -
Self-concordant smoothing in proximal quasi-Newton algorithms for large-scale convex composite optimization
by: Adeoye, Adeyemi D., et al.
Published: (2023)