Frugality in second-order optimization: floating-point approximations for Newton's method
Fuente:
arXiv
Saved in:
| Main Authors: | Carrino, Giuseppe, Piccolomini, Elena Loli, Riccietti, Elisa, Mary, Theo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Introduction to optimization methods for training SciML models
by: Kopaničáková, Alena, et al.
Published: (2026)
by: Kopaničáková, Alena, et al.
Published: (2026)
Fuzzy hyperparameters update in a second order optimization
by: Bensadok, Abdelaziz, et al.
Published: (2024)
by: Bensadok, Abdelaziz, et al.
Published: (2024)
A second-order method landing on the Stiefel manifold via Newton$\unicode{x2013}$Schulz iteration
by: Xiong, Xinhui, et al.
Published: (2026)
by: Xiong, Xinhui, et al.
Published: (2026)
A second-order-like optimizer with adaptive gradient scaling for deep learning
by: Bolte, Jérôme, et al.
Published: (2024)
by: Bolte, Jérôme, et al.
Published: (2024)
The Newton-Muon Optimizer
by: Du, Zhehang, et al.
Published: (2026)
by: Du, Zhehang, et al.
Published: (2026)
Space-Variant Total Variation boosted by learning techniques in few-view tomographic imaging
by: Morotti, Elena, et al.
Published: (2024)
by: Morotti, Elena, et al.
Published: (2024)
How Well Can Transformers Emulate In-context Newton's Method?
by: Giannou, Angeliki, et al.
Published: (2024)
by: Giannou, Angeliki, et al.
Published: (2024)
A Riemannian Optimization Perspective of the Gauss-Newton Method for Feedforward Neural Networks
by: Cayci, Semih
Published: (2024)
by: Cayci, Semih
Published: (2024)
Tutorial on amortized optimization
by: Amos, Brandon
Published: (2022)
by: Amos, Brandon
Published: (2022)
High-order expansion of Neural Ordinary Differential Equations flows
by: Izzo, Dario, et al.
Published: (2025)
by: Izzo, Dario, et al.
Published: (2025)
Linear attention is (maybe) all you need (to understand transformer optimization)
by: Ahn, Kwangjun, et al.
Published: (2023)
by: Ahn, Kwangjun, et al.
Published: (2023)
Deep learning enhanced mixed integer optimization: Learning to reduce model dimensionality
by: Triantafyllou, Niki, et al.
Published: (2024)
by: Triantafyllou, Niki, et al.
Published: (2024)
A physics-informed Bayesian optimization method for rapid development of electrical machines
by: Asef, Pedram, et al.
Published: (2025)
by: Asef, Pedram, et al.
Published: (2025)
A space-decoupling framework for optimization on bounded-rank matrices with orthogonally invariant constraints
by: Yang, Yan, et al.
Published: (2025)
by: Yang, Yan, et al.
Published: (2025)
Stochastic interior-point methods for smooth conic optimization with applications
by: He, Chuan, et al.
Published: (2024)
by: He, Chuan, et al.
Published: (2024)
High-dimensional mixed-categorical Gaussian processes with application to multidisciplinary design optimization for a green aircraft
by: Saves, Paul, et al.
Published: (2023)
by: Saves, Paul, et al.
Published: (2023)
Newton-CG methods for nonconvex unconstrained optimization with Hölder continuous Hessian
by: He, Chuan, et al.
Published: (2023)
by: He, Chuan, et al.
Published: (2023)
A multiobjective continuation method to compute the regularization path of deep neural networks
by: Amakor, Augustina C., et al.
Published: (2023)
by: Amakor, Augustina C., et al.
Published: (2023)
Stabilizing reinforcement learning control: A modular framework for optimizing over all stable behavior
by: Lawrence, Nathan P., et al.
Published: (2023)
by: Lawrence, Nathan P., et al.
Published: (2023)
Convergence and sample complexity of natural policy gradient primal-dual methods for constrained MDPs
by: Ding, Dongsheng, et al.
Published: (2022)
by: Ding, Dongsheng, et al.
Published: (2022)
A multilevel stochastic regularized first-order method with application to finite sum minimization
by: Marini, Filippo, et al.
Published: (2024)
by: Marini, Filippo, et al.
Published: (2024)
CLeAN: Continual Learning Adaptive Normalization in Dynamic Environments
by: Marasco, Isabella, et al.
Published: (2026)
by: Marasco, Isabella, et al.
Published: (2026)
On the consistent reasoning paradox of intelligence and optimal trust in AI: The power of 'I don't know'
by: Bastounis, Alexander, et al.
Published: (2024)
by: Bastounis, Alexander, et al.
Published: (2024)
Learning optimal objective values for MILP
by: Scavuzzo, Lara, et al.
Published: (2024)
by: Scavuzzo, Lara, et al.
Published: (2024)
Adaptive Weighted Total Variation boosted by learning techniques in few-view tomographic imaging
by: Morotti, Elena, et al.
Published: (2025)
by: Morotti, Elena, et al.
Published: (2025)
On the optimization dynamics of RLVR: Gradient gap and step size thresholds
by: Suk, Joe, et al.
Published: (2025)
by: Suk, Joe, et al.
Published: (2025)
Byzantine Machine Learning: MultiKrum and an optimal notion of robustness
by: Bareilles, Gilles, et al.
Published: (2026)
by: Bareilles, Gilles, et al.
Published: (2026)
$\ell_1$-norm rank-one symmetric matrix factorization has no spurious second-order stationary points
by: Guan, Jiewen, et al.
Published: (2024)
by: Guan, Jiewen, et al.
Published: (2024)
A distributed semismooth Newton based augmented Lagrangian method for distributed optimization
by: Ma, Qihao, et al.
Published: (2026)
by: Ma, Qihao, et al.
Published: (2026)
An adaptively inexact first-order method for bilevel optimization with application to hyperparameter learning
by: Salehi, Mohammad Sadegh, et al.
Published: (2023)
by: Salehi, Mohammad Sadegh, et al.
Published: (2023)
An inexact Bregman proximal point method and its acceleration version for unbalanced optimal transport
by: Chen, Xiang, et al.
Published: (2024)
by: Chen, Xiang, et al.
Published: (2024)
SMiLE: Provably Enforcing Global Relational Properties in Neural Networks
by: Francobaldi, Matteo, et al.
Published: (2025)
by: Francobaldi, Matteo, et al.
Published: (2025)
Zeroth-Order Optimization Finds Flat Minima
by: Zhang, Liang, et al.
Published: (2025)
by: Zhang, Liang, et al.
Published: (2025)
Q3R: Quadratic Reweighted Rank Regularizer for Effective Low-Rank Training
by: Ghosh, Ipsita, et al.
Published: (2025)
by: Ghosh, Ipsita, et al.
Published: (2025)
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models
by: Zhao, Pengxiang, et al.
Published: (2025)
by: Zhao, Pengxiang, et al.
Published: (2025)
On Some Tunable Multi-fidelity Bayesian Optimization Frameworks
by: Manoj, Arjun, et al.
Published: (2025)
by: Manoj, Arjun, et al.
Published: (2025)
AI2STOW: End-to-End Deep Reinforcement Learning to Construct Master Stowage Plans under Demand Uncertainty
by: Van Twiller, Jaike, et al.
Published: (2025)
by: Van Twiller, Jaike, et al.
Published: (2025)
Quantization through Piecewise-Affine Regularization: Optimization and Statistical Guarantees
by: Ma, Jianhao, et al.
Published: (2025)
by: Ma, Jianhao, et al.
Published: (2025)
Closing the Loop: Coordinating Inventory and Recommendation via Deep Reinforcement Learning on Multiple Timescales
by: Jiang, Jinyang, et al.
Published: (2025)
by: Jiang, Jinyang, et al.
Published: (2025)
On the Optimal Construction of Unbiased Gradient Estimators for Zeroth-Order Optimization
by: Ma, Shaocong, et al.
Published: (2025)
by: Ma, Shaocong, et al.
Published: (2025)
Similar Items
-
Introduction to optimization methods for training SciML models
by: Kopaničáková, Alena, et al.
Published: (2026) -
Fuzzy hyperparameters update in a second order optimization
by: Bensadok, Abdelaziz, et al.
Published: (2024) -
A second-order method landing on the Stiefel manifold via Newton$\unicode{x2013}$Schulz iteration
by: Xiong, Xinhui, et al.
Published: (2026) -
A second-order-like optimizer with adaptive gradient scaling for deep learning
by: Bolte, Jérôme, et al.
Published: (2024) -
The Newton-Muon Optimizer
by: Du, Zhehang, et al.
Published: (2026)