The Optimality of (Accelerated) SGD for High-Dimensional Quadratic Optimization
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Haihan, Liu, Yuanshi, Chen, Qianwen, Fang, Cong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Accelerating Single-Pass SGD for Generalized Linear Prediction
por: Chen, Qian, et al.
Publicado: (2026)
por: Chen, Qian, et al.
Publicado: (2026)
Optimal Projection-Free Adaptive SGD for Matrix Optimization
por: Kovalev, Dmitry
Publicado: (2026)
por: Kovalev, Dmitry
Publicado: (2026)
Understanding Outer Optimizers in Local SGD: Learning Rates, Momentum, and Acceleration
por: Khaled, Ahmed, et al.
Publicado: (2025)
por: Khaled, Ahmed, et al.
Publicado: (2025)
Accelerated Optimization Landscape of Linear-Quadratic Regulator
por: Feng, Lechen, et al.
Publicado: (2023)
por: Feng, Lechen, et al.
Publicado: (2023)
From PowerSGD to PowerSGD+: Low-Rank Gradient Compression for Distributed Optimization with Convergence Guarantees
por: Xie, Shengping, et al.
Publicado: (2025)
por: Xie, Shengping, et al.
Publicado: (2025)
SGD with Adaptive Preconditioning: Unified Analysis and Momentum Acceleration
por: Kovalev, Dmitry
Publicado: (2025)
por: Kovalev, Dmitry
Publicado: (2025)
Shadowheart SGD: Distributed Asynchronous SGD with Optimal Time Complexity Under Arbitrary Computation and Communication Heterogeneity
por: Tyurin, Alexander, et al.
Publicado: (2024)
por: Tyurin, Alexander, et al.
Publicado: (2024)
SGD with Partial Hessian for Deep Neural Networks Optimization
por: Sun, Ying, et al.
Publicado: (2024)
por: Sun, Ying, et al.
Publicado: (2024)
SLowcal-SGD: Slow Query Points Improve Local-SGD for Stochastic Convex Optimization
por: Dahan, Tehila, et al.
Publicado: (2023)
por: Dahan, Tehila, et al.
Publicado: (2023)
Two-Timescale Optimization Framework for Sparse-Feedback Linear-Quadratic Optimal Control
por: Feng, Lechen, et al.
Publicado: (2024)
por: Feng, Lechen, et al.
Publicado: (2024)
Sign-SGD via Parameter-Free Optimization
por: Medyakov, Daniil, et al.
Publicado: (2025)
por: Medyakov, Daniil, et al.
Publicado: (2025)
Near-Optimal Tensor PCA via Normalized Stochastic Gradient Ascent with Overparameterization
por: Ding, Shihong, et al.
Publicado: (2025)
por: Ding, Shihong, et al.
Publicado: (2025)
Optimal Algorithms in Linear Regression under Covariate Shift: On the Importance of Precondition
por: Liu, Yuanshi, et al.
Publicado: (2025)
por: Liu, Yuanshi, et al.
Publicado: (2025)
Revisiting Gradient Normalization and Clipping for Nonconvex SGD under Heavy-Tailed Noise: Necessity, Sufficiency, and Acceleration
por: Sun, Tao, et al.
Publicado: (2024)
por: Sun, Tao, et al.
Publicado: (2024)
On the Provable Suboptimality of Momentum SGD in Nonstationary Stochastic Optimization
por: Sahu, Sharan, et al.
Publicado: (2026)
por: Sahu, Sharan, et al.
Publicado: (2026)
Bias-Optimal Bounds for SGD: A Computer-Aided Lyapunov Analysis
por: Cortild, Daniel, et al.
Publicado: (2025)
por: Cortild, Daniel, et al.
Publicado: (2025)
Adaptive SGD with Line-Search and Polyak Stepsizes: Nonconvex Convergence and Accelerated Rates
por: Wu, Haotian
Publicado: (2025)
por: Wu, Haotian
Publicado: (2025)
ROOT-SGD: Sharp Nonasymptotics and Near-Optimal Asymptotics in a Single Algorithm
por: Li, Chris Junchi, et al.
Publicado: (2020)
por: Li, Chris Junchi, et al.
Publicado: (2020)
Optimal Growth Schedules for Batch Size and Learning Rate in SGD that Reduce SFO Complexity
por: Umeda, Hikaru, et al.
Publicado: (2025)
por: Umeda, Hikaru, et al.
Publicado: (2025)
SketchySGD: Reliable Stochastic Optimization via Randomized Curvature Estimates
por: Frangella, Zachary, et al.
Publicado: (2022)
por: Frangella, Zachary, et al.
Publicado: (2022)
Non-Euclidean SGD for Structured Optimization: Unified Analysis and Improved Rates
por: Kovalev, Dmitry, et al.
Publicado: (2025)
por: Kovalev, Dmitry, et al.
Publicado: (2025)
Proactive DP: A Multple Target Optimization Framework for DP-SGD
por: van Dijk, Marten, et al.
Publicado: (2021)
por: van Dijk, Marten, et al.
Publicado: (2021)
Nonconvex Optimization Framework for Group-Sparse Feedback Linear-Quadratic Optimal Control: Penalty Approach
por: Feng, Lechen, et al.
Publicado: (2025)
por: Feng, Lechen, et al.
Publicado: (2025)
A Unified Lyapunov-IQC Framework for Uniform Stability of Smooth Quadratic First-Order Accelerated Optimizers
por: Li, Don, et al.
Publicado: (2026)
por: Li, Don, et al.
Publicado: (2026)
A Minibatch-SGD-Based Learning Meta-Policy for Inventory Systems with Myopic Optimal Policy
por: Lyu, Jiameng, et al.
Publicado: (2024)
por: Lyu, Jiameng, et al.
Publicado: (2024)
Enhancing Stochastic Optimization for Statistical Efficiency Using ROOT-SGD with Diminishing Stepsize
por: Li, Chris Junchi
Publicado: (2024)
por: Li, Chris Junchi
Publicado: (2024)
Tight Long-Term Tail Decay of (Clipped) SGD in Non-Convex Optimization
por: Armacki, Aleksandar, et al.
Publicado: (2026)
por: Armacki, Aleksandar, et al.
Publicado: (2026)
Leveraging Coordinate Momentum in SignSGD and Muon: Memory-Optimized Zero-Order
por: Petrov, Egor, et al.
Publicado: (2025)
por: Petrov, Egor, et al.
Publicado: (2025)
Nonconvex Optimization Framework for Group-Sparse Feedback Linear-Quadratic Optimal Control: Non-Penalty Approach
por: Feng, Lechen, et al.
Publicado: (2025)
por: Feng, Lechen, et al.
Publicado: (2025)
Making SGD Parameter-Free
por: Carmon, Yair, et al.
Publicado: (2022)
por: Carmon, Yair, et al.
Publicado: (2022)
On the Trajectories of SGD Without Replacement
por: Beneventano, Pierfrancesco
Publicado: (2023)
por: Beneventano, Pierfrancesco
Publicado: (2023)
Differentially Private Clipped-SGD: High-Probability Convergence with Arbitrary Clipping Level
por: Khah, Saleh Vatan, et al.
Publicado: (2025)
por: Khah, Saleh Vatan, et al.
Publicado: (2025)
$μ^2$-SGD: Stable Stochastic Optimization via a Double Momentum Mechanism
por: Dahan, Tehila, et al.
Publicado: (2023)
por: Dahan, Tehila, et al.
Publicado: (2023)
Dual-Delayed Asynchronous SGD for Arbitrarily Heterogeneous Data
por: Wang, Xiaolu, et al.
Publicado: (2024)
por: Wang, Xiaolu, et al.
Publicado: (2024)
High-Probability Bounds for SGD under the Polyak-Lojasiewicz Condition with Markovian Noise
por: Kar, Avik, et al.
Publicado: (2026)
por: Kar, Avik, et al.
Publicado: (2026)
Demystifying SGD with Doubly Stochastic Gradients
por: Kim, Kyurae, et al.
Publicado: (2024)
por: Kim, Kyurae, et al.
Publicado: (2024)
Dimension-adapted Momentum Outscales SGD
por: Ferbach, Damien, et al.
Publicado: (2025)
por: Ferbach, Damien, et al.
Publicado: (2025)
Heavy-Tail Phenomenon in Decentralized SGD
por: Gurbuzbalaban, Mert, et al.
Publicado: (2022)
por: Gurbuzbalaban, Mert, et al.
Publicado: (2022)
VAMO: Efficient Zeroth-Order Variance Reduction for SGD with Faster Convergence
por: Chen, Jiahe, et al.
Publicado: (2025)
por: Chen, Jiahe, et al.
Publicado: (2025)
Does SGD Seek Flatness or Sharpness? An Exactly Solvable Model
por: Xu, Yizhou, et al.
Publicado: (2026)
por: Xu, Yizhou, et al.
Publicado: (2026)
Ejemplares similares
-
Accelerating Single-Pass SGD for Generalized Linear Prediction
por: Chen, Qian, et al.
Publicado: (2026) -
Optimal Projection-Free Adaptive SGD for Matrix Optimization
por: Kovalev, Dmitry
Publicado: (2026) -
Understanding Outer Optimizers in Local SGD: Learning Rates, Momentum, and Acceleration
por: Khaled, Ahmed, et al.
Publicado: (2025) -
Accelerated Optimization Landscape of Linear-Quadratic Regulator
por: Feng, Lechen, et al.
Publicado: (2023) -
From PowerSGD to PowerSGD+: Low-Rank Gradient Compression for Distributed Optimization with Convergence Guarantees
por: Xie, Shengping, et al.
Publicado: (2025)