Error dynamics of mini-batch gradient descent with random reshuffling for least squares regression
Fuente:
arXiv
Saved in:
| Main Authors: | Lok, Jackie, Sonthalia, Rishi, Rebrova, Elizaveta |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On Regularization via Early Stopping for Least Squares Regression
by: Sonthalia, Rishi, et al.
Published: (2024)
by: Sonthalia, Rishi, et al.
Published: (2024)
Subspace-constrained randomized coordinate descent for linear systems with good low-rank matrix approximations
by: Lok, Jackie, et al.
Published: (2025)
by: Lok, Jackie, et al.
Published: (2025)
Towards Universal Convergence of Backward Error in Linear System Solvers
by: Dereziński, Michał, et al.
Published: (2026)
by: Dereziński, Michał, et al.
Published: (2026)
Random-reshuffled SARAH does not need a full gradient computations
by: Beznosikov, Aleksandr, et al.
Published: (2021)
by: Beznosikov, Aleksandr, et al.
Published: (2021)
Compressed and distributed least-squares regression: convergence rates with applications to Federated Learning
by: Philippenko, Constantin, et al.
Published: (2023)
by: Philippenko, Constantin, et al.
Published: (2023)
A stochastic gradient descent algorithm with random search directions
by: Gbaguidi, Eméric
Published: (2025)
by: Gbaguidi, Eméric
Published: (2025)
Robust stochastic first order methods in heavy-tailed noise via medoid mini-batch gradient sampling
by: Vukovic, Manojlo, et al.
Published: (2026)
by: Vukovic, Manojlo, et al.
Published: (2026)
On the stability of gradient descent with second order dynamics for time-varying cost functions
by: Gibson, Travis E., et al.
Published: (2024)
by: Gibson, Travis E., et al.
Published: (2024)
New logarithmic step size for stochastic gradient descent
by: Shamaee, M. Soheil, et al.
Published: (2024)
by: Shamaee, M. Soheil, et al.
Published: (2024)
Geometry-Preserving Neural Architectures on Manifolds with Boundary
by: Elamvazhuthi, Karthik, et al.
Published: (2026)
by: Elamvazhuthi, Karthik, et al.
Published: (2026)
Randomized Kaczmarz Methods with Beyond-Krylov Convergence
by: Dereziński, Michał, et al.
Published: (2025)
by: Dereziński, Michał, et al.
Published: (2025)
A least-square method for non-asymptotic identification in linear switching control
by: Sun, Haoyuan, et al.
Published: (2024)
by: Sun, Haoyuan, et al.
Published: (2024)
Error bounds for particle gradient descent, and extensions of the log-Sobolev and Talagrand inequalities
by: Caprio, Rocco, et al.
Published: (2024)
by: Caprio, Rocco, et al.
Published: (2024)
On the global convergence of gradient descent for wide shallow models with bounded nonlinearities
by: Petit, Romain, et al.
Published: (2026)
by: Petit, Romain, et al.
Published: (2026)
The duality structure gradient descent algorithm: analysis and applications to neural networks
by: Flynn, Thomas
Published: (2017)
by: Flynn, Thomas
Published: (2017)
Concentration inequalities for semidefinite least squares based on data
by: Fabiani, Filippo, et al.
Published: (2025)
by: Fabiani, Filippo, et al.
Published: (2025)
Long-time dynamics and universality of nonconvex gradient descent
by: Han, Qiyang
Published: (2025)
by: Han, Qiyang
Published: (2025)
Convergence of continuous-time stochastic gradient descent with applications to deep neural networks
by: Lugosi, Gabor, et al.
Published: (2024)
by: Lugosi, Gabor, et al.
Published: (2024)
Fine-grained Analysis and Faster Algorithms for Iteratively Solving Linear Systems
by: Dereziński, Michał, et al.
Published: (2024)
by: Dereziński, Michał, et al.
Published: (2024)
Learning nonnegative matrix factorizations from compressed data
by: Chaudhry, Abraar, et al.
Published: (2024)
by: Chaudhry, Abraar, et al.
Published: (2024)
Convergence of two-timescale gradient descent ascent dynamics: finite-dimensional and mean-field perspectives
by: An, Jing, et al.
Published: (2025)
by: An, Jing, et al.
Published: (2025)
Convergence of stochastic gradient descent under a local Lojasiewicz condition for deep neural networks
by: An, Jing, et al.
Published: (2023)
by: An, Jing, et al.
Published: (2023)
Risk reversal for least squares estimators under nested convex constraints
by: Al-Ghattas, Omar
Published: (2026)
by: Al-Ghattas, Omar
Published: (2026)
Global convergence of gradient descent for phase retrieval
by: Fougereux, Théodore, et al.
Published: (2024)
by: Fougereux, Théodore, et al.
Published: (2024)
A short proof of near-linear convergence of adaptive gradient descent under fourth-order growth and convexity
by: Davis, Damek, et al.
Published: (2026)
by: Davis, Damek, et al.
Published: (2026)
Bayesian dynamic scheduling of multipurpose batch processes under incomplete look-ahead information
by: Zheng, Taicheng, et al.
Published: (2025)
by: Zheng, Taicheng, et al.
Published: (2025)
When majority rules, minority loses: bias amplification of gradient descent
by: Bachoc, François, et al.
Published: (2025)
by: Bachoc, François, et al.
Published: (2025)
Precise gradient descent training dynamics for finite-width multi-layer neural networks
by: Han, Qiyang, et al.
Published: (2025)
by: Han, Qiyang, et al.
Published: (2025)
Non-convergence to global minimizers for Adam and stochastic gradient descent optimization and constructions of local minimizers in the training of artificial neural networks
by: Jentzen, Arnulf, et al.
Published: (2024)
by: Jentzen, Arnulf, et al.
Published: (2024)
Convergence of projected stochastic natural gradient variational inference for various step size and sample or batch size schedules
by: Guilmeau, Thomas, et al.
Published: (2026)
by: Guilmeau, Thomas, et al.
Published: (2026)
Manifold constrained steepest descent
by: Yang, Kaiwei, et al.
Published: (2026)
by: Yang, Kaiwei, et al.
Published: (2026)
Proximal random reshuffling under local Lipschitz continuity
by: Josz, Cedric, et al.
Published: (2024)
by: Josz, Cedric, et al.
Published: (2024)
Flattened one-bit stochastic gradient descent: compressed distributed optimization with controlled variance
by: Stollenwerk, Alexander, et al.
Published: (2024)
by: Stollenwerk, Alexander, et al.
Published: (2024)
Convergence of flow-based generative models via proximal gradient descent in Wasserstein space
by: Cheng, Xiuyuan, et al.
Published: (2023)
by: Cheng, Xiuyuan, et al.
Published: (2023)
Problem-dependent convergence bounds for randomized linear gradient compression
by: Flynn, Thomas, et al.
Published: (2024)
by: Flynn, Thomas, et al.
Published: (2024)
Generalized EXTRA stochastic gradient Langevin dynamics
by: Gurbuzbalaban, Mert, et al.
Published: (2024)
by: Gurbuzbalaban, Mert, et al.
Published: (2024)
High-dimensional scaling limits and fluctuations of online least-squares SGD with smooth covariance
by: Balasubramanian, Krishnakumar, et al.
Published: (2023)
by: Balasubramanian, Krishnakumar, et al.
Published: (2023)
Learning mirror maps in policy mirror descent
by: Alfano, Carlo, et al.
Published: (2024)
by: Alfano, Carlo, et al.
Published: (2024)
Riemannian coordinate descent algorithms on matrix manifolds
by: Han, Andi, et al.
Published: (2024)
by: Han, Andi, et al.
Published: (2024)
Stochastic gradient with least-squares control variates
by: Nobile, Fabio, et al.
Published: (2025)
by: Nobile, Fabio, et al.
Published: (2025)
Similar Items
-
On Regularization via Early Stopping for Least Squares Regression
by: Sonthalia, Rishi, et al.
Published: (2024) -
Subspace-constrained randomized coordinate descent for linear systems with good low-rank matrix approximations
by: Lok, Jackie, et al.
Published: (2025) -
Towards Universal Convergence of Backward Error in Linear System Solvers
by: Dereziński, Michał, et al.
Published: (2026) -
Random-reshuffled SARAH does not need a full gradient computations
by: Beznosikov, Aleksandr, et al.
Published: (2021) -
Compressed and distributed least-squares regression: convergence rates with applications to Federated Learning
by: Philippenko, Constantin, et al.
Published: (2023)