Quantitative Convergences of Lie Group Momentum Optimizers
Fuente:
arXiv
Saved in:
| Main Authors: | Kong, Lingkai, Tao, Molei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Convergence of Kinetic Langevin Monte Carlo on Lie groups
by: Kong, Lingkai, et al.
Published: (2024)
by: Kong, Lingkai, et al.
Published: (2024)
Policy Gradient with Second Order Momentum
by: Sun, Tianyu
Published: (2025)
by: Sun, Tianyu
Published: (2025)
A Family of Controllable Momentum Coefficients for Forward-Backward Accelerated Algorithms
by: Fu, Mingwei, et al.
Published: (2025)
by: Fu, Mingwei, et al.
Published: (2025)
Last-Iterate Convergence of Randomized Kaczmarz and SGD with Greedy Step Size
by: Dereziński, Michał, et al.
Published: (2026)
by: Dereziński, Michał, et al.
Published: (2026)
Modified Loss of Momentum Gradient Descent: Fine-Grained Analysis
by: Cattaneo, Matias D., et al.
Published: (2025)
by: Cattaneo, Matias D., et al.
Published: (2025)
Anderson Acceleration in Nonsmooth Problems: Local Convergence via Active Manifold Identification
by: Li, Kexin, et al.
Published: (2024)
by: Li, Kexin, et al.
Published: (2024)
Convergence of Adam for Non-convex Objectives: Relaxed Hyperparameters and Non-ergodic Case
by: He, Meixuan, et al.
Published: (2023)
by: He, Meixuan, et al.
Published: (2023)
Convergence of two-timescale gradient descent ascent dynamics: finite-dimensional and mean-field perspectives
by: An, Jing, et al.
Published: (2025)
by: An, Jing, et al.
Published: (2025)
On the Convergence of the Gradient Descent Method with Stochastic Fixed-point Rounding Errors under the Polyak-Lojasiewicz Inequality
by: Xia, Lu, et al.
Published: (2023)
by: Xia, Lu, et al.
Published: (2023)
Nonlinear Dimensionality Reduction Techniques for Bayesian Optimization
by: Long, Luo, et al.
Published: (2025)
by: Long, Luo, et al.
Published: (2025)
Adaptive Proximal Gradient Method for Convex Optimization
by: Malitsky, Yura, et al.
Published: (2023)
by: Malitsky, Yura, et al.
Published: (2023)
Scalable Acceleration for Classification-Based Derivative-Free Optimization
by: Han, Tianyi, et al.
Published: (2023)
by: Han, Tianyi, et al.
Published: (2023)
Super Gradient Descent: Global Optimization requires Global Gradient
by: Achour, Seifeddine
Published: (2024)
by: Achour, Seifeddine
Published: (2024)
Primal-Dual Methods for Nonsmooth Nonconvex Optimization with Orthogonality Constraints
by: Zhu, Linglingzhi, et al.
Published: (2026)
by: Zhu, Linglingzhi, et al.
Published: (2026)
Enhanced Adaptive Gradient Algorithms for Nonconvex-PL Minimax Optimization
by: Huang, Feihu, et al.
Published: (2023)
by: Huang, Feihu, et al.
Published: (2023)
A Gauss-Newton Approach for Min-Max Optimization in Generative Adversarial Networks
by: Mishra, Neel, et al.
Published: (2024)
by: Mishra, Neel, et al.
Published: (2024)
Fast Unconstrained Optimization via Hessian Averaging and Adaptive Gradient Sampling Methods
by: O'Leary-Roseberry, Thomas, et al.
Published: (2024)
by: O'Leary-Roseberry, Thomas, et al.
Published: (2024)
Shape Derivative-Informed Neural Operators with Application to Risk-Averse Shape Optimization
by: Gong, Xindi, et al.
Published: (2026)
by: Gong, Xindi, et al.
Published: (2026)
Adaptive Lipschitz-Free Conditional Gradient Methods for Stochastic Composite Nonconvex Optimization
by: Yuan, Ganzhao
Published: (2026)
by: Yuan, Ganzhao
Published: (2026)
UAdam: Unified Adam-Type Algorithmic Framework for Non-Convex Stochastic Optimization
by: Jiang, Yiming, et al.
Published: (2023)
by: Jiang, Yiming, et al.
Published: (2023)
OptEMA: Adaptive Exponential Moving Average for Stochastic Optimization with Zero-Noise Optimality
by: Yuan, Ganzhao
Published: (2026)
by: Yuan, Ganzhao
Published: (2026)
A Block Coordinate Descent Method for Nonsmooth Composite Optimization under Orthogonality Constraints
by: Yuan, Ganzhao
Published: (2023)
by: Yuan, Ganzhao
Published: (2023)
End-to-End Mesh Optimization of a Hybrid Deep Learning Black-Box PDE Solver
by: Ma, Shaocong, et al.
Published: (2024)
by: Ma, Shaocong, et al.
Published: (2024)
Gradient is All You Need? How Consensus-Based Optimization can be Interpreted as a Stochastic Relaxation of Gradient Descent
by: Riedl, Konstantin, et al.
Published: (2023)
by: Riedl, Konstantin, et al.
Published: (2023)
Bayesian Optimization on Networks
by: Li, Wenwen, et al.
Published: (2025)
by: Li, Wenwen, et al.
Published: (2025)
Convergence analysis of Lie and Strang splitting for operator-valued differential Riccati equations
by: Hansen, Eskil, et al.
Published: (2025)
by: Hansen, Eskil, et al.
Published: (2025)
On the Width Scaling of Neural Optimizers Under Matrix Operator Norms I: Row/Column Normalization and Hyperparameter Transfer
by: Xu, Ruihan, et al.
Published: (2026)
by: Xu, Ruihan, et al.
Published: (2026)
Trust-Region Sequential Quadratic Programming for Stochastic Optimization with Random Models
by: Fang, Yuchen, et al.
Published: (2024)
by: Fang, Yuchen, et al.
Published: (2024)
Convergence Analysis of Fractional Gradient Descent
by: Aggarwal, Ashwani
Published: (2023)
by: Aggarwal, Ashwani
Published: (2023)
A Type II Hamiltonian Variational Principle and Adjoint Systems for Lie Groups
by: Tran, Brian K., et al.
Published: (2023)
by: Tran, Brian K., et al.
Published: (2023)
Randomized Kaczmarz Methods with Beyond-Krylov Convergence
by: Dereziński, Michał, et al.
Published: (2025)
by: Dereziński, Michał, et al.
Published: (2025)
Learning rate adaptive stochastic gradient descent optimization methods: numerical simulations for deep learning methods for partial differential equations and convergence analyses
by: Dereich, Steffen, et al.
Published: (2024)
by: Dereich, Steffen, et al.
Published: (2024)
On the numerical reliability of nonsmooth autodiff: a MaxPool case study
by: Boustany, Ryan
Published: (2024)
by: Boustany, Ryan
Published: (2024)
Subhomogeneous Deep Equilibrium Models
by: Sittoni, Pietro, et al.
Published: (2024)
by: Sittoni, Pietro, et al.
Published: (2024)
Flattened one-bit stochastic gradient descent: compressed distributed optimization with controlled variance
by: Stollenwerk, Alexander, et al.
Published: (2024)
by: Stollenwerk, Alexander, et al.
Published: (2024)
Efficient Trajectory Inference in Wasserstein Space Using Consecutive Averaging
by: Banerjee, Amartya, et al.
Published: (2024)
by: Banerjee, Amartya, et al.
Published: (2024)
KANtrol: A Physics-Informed Kolmogorov-Arnold Network Framework for Solving Multi-Dimensional and Fractional Optimal Control Problems
by: Aghaei, Alireza Afzal
Published: (2024)
by: Aghaei, Alireza Afzal
Published: (2024)
Real-time optimal control of high-dimensional parametrized systems by deep learning-based reduced order models
by: Tomasetto, Matteo, et al.
Published: (2024)
by: Tomasetto, Matteo, et al.
Published: (2024)
Cubic regularized subspace Newton for non-convex optimization
by: Zhao, Jim, et al.
Published: (2024)
by: Zhao, Jim, et al.
Published: (2024)
Towards Quantifying the Preconditioning Effect of Adam
by: Das, Rudrajit, et al.
Published: (2024)
by: Das, Rudrajit, et al.
Published: (2024)
Similar Items
-
Convergence of Kinetic Langevin Monte Carlo on Lie groups
by: Kong, Lingkai, et al.
Published: (2024) -
Policy Gradient with Second Order Momentum
by: Sun, Tianyu
Published: (2025) -
A Family of Controllable Momentum Coefficients for Forward-Backward Accelerated Algorithms
by: Fu, Mingwei, et al.
Published: (2025) -
Last-Iterate Convergence of Randomized Kaczmarz and SGD with Greedy Step Size
by: Dereziński, Michał, et al.
Published: (2026) -
Modified Loss of Momentum Gradient Descent: Fine-Grained Analysis
by: Cattaneo, Matias D., et al.
Published: (2025)