Local linear convergence of gradient methods for overparameterized Gaussian mixtures
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Jingxing, Charisopoulos, Vasileios, Fazel, Maryam |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Toward Global Convergence of Gradient EM for Over-Parameterized Gaussian Mixture Models
by: Xu, Weihang, et al.
Published: (2024)
by: Xu, Weihang, et al.
Published: (2024)
Solving Inverse Problems with Deep Linear Neural Networks: Global Convergence Guarantees for Gradient Descent with Weight Decay
by: Laus, Hannah, et al.
Published: (2025)
by: Laus, Hannah, et al.
Published: (2025)
Problem-dependent convergence bounds for randomized linear gradient compression
by: Flynn, Thomas, et al.
Published: (2024)
by: Flynn, Thomas, et al.
Published: (2024)
Almost sure convergence rates of stochastic gradient methods under gradient domination
by: Weissmann, Simon, et al.
Published: (2024)
by: Weissmann, Simon, et al.
Published: (2024)
Linear regression with overparameterized linear neural networks: Tight upper and lower bounds for implicit $\ell^1$-regularization
by: Matt, Hannes, et al.
Published: (2025)
by: Matt, Hannes, et al.
Published: (2025)
Finite Sample Identification of Partially Observed Bilinear Dynamical Systems
by: Sattar, Yahya, et al.
Published: (2025)
by: Sattar, Yahya, et al.
Published: (2025)
A short proof of near-linear convergence of adaptive gradient descent under fourth-order growth and convexity
by: Davis, Damek, et al.
Published: (2026)
by: Davis, Damek, et al.
Published: (2026)
Explore-then-Commit for Nonstationary Linear Bandits with Latent Dynamics
by: Choi, Sunmook, et al.
Published: (2025)
by: Choi, Sunmook, et al.
Published: (2025)
Sub-optimality of the Separation Principle for Quadratic Control from Bilinear Observations
by: Sattar, Yahya, et al.
Published: (2025)
by: Sattar, Yahya, et al.
Published: (2025)
Global Convergence of Four-Layer Matrix Factorization under Random Initialization
by: Luo, Minrui, et al.
Published: (2025)
by: Luo, Minrui, et al.
Published: (2025)
Preconditioned subgradient method for composite optimization: overparameterization and fast convergence
by: Díaz, Mateo, et al.
Published: (2025)
by: Díaz, Mateo, et al.
Published: (2025)
On the global convergence of gradient descent for wide shallow models with bounded nonlinearities
by: Petit, Romain, et al.
Published: (2026)
by: Petit, Romain, et al.
Published: (2026)
Iteratively reweighted kernel machines efficiently learn sparse functions
by: Zhu, Libin, et al.
Published: (2025)
by: Zhu, Libin, et al.
Published: (2025)
High-dimensional Limit of SGD for Diagonal Linear Networks
by: Malaxechebarría, Begoña García, et al.
Published: (2026)
by: Malaxechebarría, Begoña García, et al.
Published: (2026)
Learning to optimize with guarantees: a complete characterization of linearly convergent algorithms
by: Martin, Andrea, et al.
Published: (2025)
by: Martin, Andrea, et al.
Published: (2025)
Online reinforcement learning via sparse Gaussian mixture model Q-functions
by: Vu, Minh, et al.
Published: (2025)
by: Vu, Minh, et al.
Published: (2025)
A stochastic gradient method for trilevel optimization
by: Giovannelli, Tommaso, et al.
Published: (2025)
by: Giovannelli, Tommaso, et al.
Published: (2025)
Contractivity and linear convergence in bilinear saddle-point problems: An operator-theoretic approach
by: Dirren, Colin, et al.
Published: (2024)
by: Dirren, Colin, et al.
Published: (2024)
Auto-conditioned primal-dual hybrid gradient method and alternating direction method of multipliers
by: Lan, Guanghui, et al.
Published: (2024)
by: Lan, Guanghui, et al.
Published: (2024)
Learning rate adaptive stochastic gradient descent optimization methods: numerical simulations for deep learning methods for partial differential equations and convergence analyses
by: Dereich, Steffen, et al.
Published: (2024)
by: Dereich, Steffen, et al.
Published: (2024)
Mean-square and linear convergence of a stochastic proximal point algorithm in metric spaces of nonpositive curvature
by: Pischke, Nicholas
Published: (2025)
by: Pischke, Nicholas
Published: (2025)
Local convergence of simultaneous min-max algorithms to differential equilibrium on Riemannian manifold
by: Zhang, Sixin
Published: (2024)
by: Zhang, Sixin
Published: (2024)
Global convergence of gradient descent for phase retrieval
by: Fougereux, Théodore, et al.
Published: (2024)
by: Fougereux, Théodore, et al.
Published: (2024)
Non-convergence to global minimizers for Adam and stochastic gradient descent optimization and constructions of local minimizers in the training of artificial neural networks
by: Jentzen, Arnulf, et al.
Published: (2024)
by: Jentzen, Arnulf, et al.
Published: (2024)
Unregularized limit of stochastic gradient method for Wasserstein distributionally robust optimization
by: Le, Tam
Published: (2025)
by: Le, Tam
Published: (2025)
Nonlinear tomographic reconstruction via nonsmooth optimization
by: Charisopoulos, Vasileios, et al.
Published: (2024)
by: Charisopoulos, Vasileios, et al.
Published: (2024)
Geometry and convergence of natural policy gradient methods
by: Müller, Johannes, et al.
Published: (2022)
by: Müller, Johannes, et al.
Published: (2022)
Adaptive multi-gradient methods for quasiconvex vector optimization and applications to multi-task learning
by: Minh, Nguyen Anh, et al.
Published: (2024)
by: Minh, Nguyen Anh, et al.
Published: (2024)
Convergence rates of stochastic gradient method with independent sequences of step-size and momentum weight
by: Hwang, Wen-Liang
Published: (2024)
by: Hwang, Wen-Liang
Published: (2024)
Projected gradient methods for nonconvex and stochastic smooth optimization: new complexities and auto-conditioned stepsizes
by: Lan, Guanghui, et al.
Published: (2024)
by: Lan, Guanghui, et al.
Published: (2024)
Online Cluster-Based Parameter Control for Metaheuristic
by: Tatsis, Vasileios A., et al.
Published: (2025)
by: Tatsis, Vasileios A., et al.
Published: (2025)
On the convergence rate of noisy Bayesian Optimization with Expected Improvement
by: Wang, Jingyi, et al.
Published: (2025)
by: Wang, Jingyi, et al.
Published: (2025)
Accelerated zero-order SGD under high-order smoothness and overparameterized regime
by: Bychkov, Georgii, et al.
Published: (2024)
by: Bychkov, Georgii, et al.
Published: (2024)
The generator gradient estimator is an adjoint state method for stochastic differential equations
by: Badolle, Quentin, et al.
Published: (2024)
by: Badolle, Quentin, et al.
Published: (2024)
On improving generalization in a class of learning problems with the method of small parameters for weakly-controlled optimal gradient systems
by: Befekadu, Getachew K.
Published: (2024)
by: Befekadu, Getachew K.
Published: (2024)
A least-square method for non-asymptotic identification in linear switching control
by: Sun, Haoyuan, et al.
Published: (2024)
by: Sun, Haoyuan, et al.
Published: (2024)
Full error analysis of policy gradient learning algorithms for exploratory linear quadratic mean-field control problem in continuous time with common noise
by: Frikha, Noufel, et al.
Published: (2024)
by: Frikha, Noufel, et al.
Published: (2024)
Generalized EXTRA stochastic gradient Langevin dynamics
by: Gurbuzbalaban, Mert, et al.
Published: (2024)
by: Gurbuzbalaban, Mert, et al.
Published: (2024)
No-Rank Tensor Decomposition Using Metric Learning
by: Bagherian, Maryam
Published: (2025)
by: Bagherian, Maryam
Published: (2025)
Toward generalizable learning of all (linear) first-order methods via memory augmented Transformers
by: Dutta, Sanchayan, et al.
Published: (2024)
by: Dutta, Sanchayan, et al.
Published: (2024)
Similar Items
-
Toward Global Convergence of Gradient EM for Over-Parameterized Gaussian Mixture Models
by: Xu, Weihang, et al.
Published: (2024) -
Solving Inverse Problems with Deep Linear Neural Networks: Global Convergence Guarantees for Gradient Descent with Weight Decay
by: Laus, Hannah, et al.
Published: (2025) -
Problem-dependent convergence bounds for randomized linear gradient compression
by: Flynn, Thomas, et al.
Published: (2024) -
Almost sure convergence rates of stochastic gradient methods under gradient domination
by: Weissmann, Simon, et al.
Published: (2024) -
Linear regression with overparameterized linear neural networks: Tight upper and lower bounds for implicit $\ell^1$-regularization
by: Matt, Hannes, et al.
Published: (2025)