Dual Space Preconditioning for Gradient Descent in the Overparameterized Regime
Fuente:
arXiv
Saved in:
| Main Authors: | Ghane, Reza, Akhtiamov, Danil, Hassibi, Babak |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Precise Performance of Linear Denoisers in the Proportional Regime
by: Ghane, Reza, et al.
Published: (2026)
by: Ghane, Reza, et al.
Published: (2026)
Universality in Transfer Learning for Linear Models
by: Ghane, Reza, et al.
Published: (2024)
by: Ghane, Reza, et al.
Published: (2024)
One-Bit Quantization and Sparsification for Multiclass Linear Classification with Strong Regularization
by: Ghane, Reza, et al.
Published: (2024)
by: Ghane, Reza, et al.
Published: (2024)
One-Bit Quantization for Random Features Models
by: Akhtiamov, Danil, et al.
Published: (2025)
by: Akhtiamov, Danil, et al.
Published: (2025)
Implicit Bias and Convergence of Matrix Stochastic Mirror Descent
by: Akhtiamov, Danil, et al.
Published: (2026)
by: Akhtiamov, Danil, et al.
Published: (2026)
Gaussian Universality for Diffusion Models
by: Ghane, Reza, et al.
Published: (2025)
by: Ghane, Reza, et al.
Published: (2025)
A Novel Gaussian Min-Max Theorem and its Applications
by: Akhtiamov, Danil, et al.
Published: (2024)
by: Akhtiamov, Danil, et al.
Published: (2024)
A Precise Performance Analysis of the Randomized Singular Value Decomposition
by: Akhtiamov, Danil, et al.
Published: (2025)
by: Akhtiamov, Danil, et al.
Published: (2025)
Robust Mean Estimation With Auxiliary Samples
by: Han, Barron, et al.
Published: (2025)
by: Han, Barron, et al.
Published: (2025)
Preconditioned Gradient Descent for Overparameterized Nonconvex Burer--Monteiro Factorization with Global Optimality Certification
by: Zhang, Gavin, et al.
Published: (2022)
by: Zhang, Gavin, et al.
Published: (2022)
Mirror and Preconditioned Gradient Descent in Wasserstein Space
by: Bonet, Clément, et al.
Published: (2024)
by: Bonet, Clément, et al.
Published: (2024)
Benefits of Early Stopping in Gradient Descent for Overparameterized Logistic Regression
by: Wu, Jingfeng, et al.
Published: (2025)
by: Wu, Jingfeng, et al.
Published: (2025)
Estimation of Toeplitz Covariance Matrices using Overparameterized Gradient Descent
by: Busbib, Daniel, et al.
Published: (2025)
by: Busbib, Daniel, et al.
Published: (2025)
Beyond Quadratic Costs: A Bregman Divergence Approach to H$_\infty$ Control
by: Hajar, Joudi, et al.
Published: (2025)
by: Hajar, Joudi, et al.
Published: (2025)
Beyond Quadratic Costs in LQR: Bregman Divergence Control
by: Hassibi, Babak, et al.
Published: (2025)
by: Hassibi, Babak, et al.
Published: (2025)
Characterizing Dynamical Stability of Stochastic Gradient Descent in Overparameterized Learning
by: Chemnitz, Dennis, et al.
Published: (2024)
by: Chemnitz, Dennis, et al.
Published: (2024)
Optimal Implicit Bias in Linear Regression
by: Varma, Kanumuri Nithin, et al.
Published: (2025)
by: Varma, Kanumuri Nithin, et al.
Published: (2025)
Bias of Stochastic Gradient Descent or the Architecture: Disentangling the Effects of Overparameterization of Neural Networks
by: Peleg, Amit, et al.
Published: (2024)
by: Peleg, Amit, et al.
Published: (2024)
Effectiveness of Distributed Gradient Descent with Local Steps for Overparameterized Models
by: Zhu, Heng, et al.
Published: (2024)
by: Zhu, Heng, et al.
Published: (2024)
On the Convergence Behavior of Preconditioned Gradient Descent Toward the Rich Learning Regime
by: Jiang, Shuai, et al.
Published: (2026)
by: Jiang, Shuai, et al.
Published: (2026)
Preconditioning for Accelerated Gradient Descent Optimization and Regularization
by: Ye, Qiang
Published: (2024)
by: Ye, Qiang
Published: (2024)
Privacy for Free in the Overparameterized Regime
by: Bombari, Simone, et al.
Published: (2024)
by: Bombari, Simone, et al.
Published: (2024)
Distributionally Robust K-Means Clustering
by: Malik, Vikrant, et al.
Published: (2026)
by: Malik, Vikrant, et al.
Published: (2026)
Adaptive Step Sizes for Preconditioned Stochastic Gradient Descent
by: Köhne, Frederik, et al.
Published: (2023)
by: Köhne, Frederik, et al.
Published: (2023)
The Power of Preconditioning in Overparameterized Low-Rank Matrix Sensing
by: Xu, Xingyu, et al.
Published: (2023)
by: Xu, Xingyu, et al.
Published: (2023)
Preconditioned Gradient Descent for Over-Parameterized Nonconvex Matrix Factorization
by: Zhang, Gavin, et al.
Published: (2025)
by: Zhang, Gavin, et al.
Published: (2025)
NysAct: A Scalable Preconditioned Gradient Descent using Nystrom Approximation
by: Seung, Hyunseok, et al.
Published: (2025)
by: Seung, Hyunseok, et al.
Published: (2025)
Learn to Change the World: Multi-level Reinforcement Learning with Model-Changing Actions
by: Lu, Ziqing, et al.
Published: (2025)
by: Lu, Ziqing, et al.
Published: (2025)
Variational Stochastic Gradient Descent for Deep Neural Networks
by: Chen, Haotian, et al.
Published: (2024)
by: Chen, Haotian, et al.
Published: (2024)
Guarantees of a Preconditioned Subgradient Algorithm for Overparameterized Asymmetric Low-rank Matrix Recovery
by: Giampouras, Paris, et al.
Published: (2024)
by: Giampouras, Paris, et al.
Published: (2024)
Stochastic Gradient Descent in the Saddle-to-Saddle Regime of Deep Linear Networks
by: Corlouer, Guillaume, et al.
Published: (2026)
by: Corlouer, Guillaume, et al.
Published: (2026)
Efficient Low-Tubal-Rank Tensor Estimation via Alternating Preconditioned Gradient Descent
by: Liu, Zhiyu, et al.
Published: (2025)
by: Liu, Zhiyu, et al.
Published: (2025)
Sharp Generalization for Nonparametric Regression in Interpolation Space by Over-Parameterized Neural Networks Trained with Preconditioned Gradient Descent and Early Stopping
by: Yang, Yingzhen, et al.
Published: (2024)
by: Yang, Yingzhen, et al.
Published: (2024)
Efficient Over-parameterized Matrix Sensing from Noisy Measurements via Alternating Preconditioned Gradient Descent
by: Liu, Zhiyu, et al.
Published: (2025)
by: Liu, Zhiyu, et al.
Published: (2025)
How Does Label Noise Gradient Descent Improve Generalization in the Low SNR Regime?
by: Huang, Wei, et al.
Published: (2025)
by: Huang, Wei, et al.
Published: (2025)
Double Descent and Overparameterization in Particle Physics Data
by: Vigl, Matthias, et al.
Published: (2025)
by: Vigl, Matthias, et al.
Published: (2025)
Sharp Global Guarantees for Nonconvex Low-rank Recovery in the Noisy Overparameterized Regime
by: Zhang, Richard Y.
Published: (2021)
by: Zhang, Richard Y.
Published: (2021)
Automated Feature Labeling with Token-Space Gradient Descent
by: Schulz, Julian, et al.
Published: (2025)
by: Schulz, Julian, et al.
Published: (2025)
Armijo Line-search Can Make (Stochastic) Gradient Descent Provably Faster
by: Vaswani, Sharan, et al.
Published: (2025)
by: Vaswani, Sharan, et al.
Published: (2025)
Fast and Accurate Estimation of Low-Rank Matrices from Noisy Measurements via Preconditioned Non-Convex Gradient Descent
by: Zhang, Gavin, et al.
Published: (2023)
by: Zhang, Gavin, et al.
Published: (2023)
Similar Items
-
Precise Performance of Linear Denoisers in the Proportional Regime
by: Ghane, Reza, et al.
Published: (2026) -
Universality in Transfer Learning for Linear Models
by: Ghane, Reza, et al.
Published: (2024) -
One-Bit Quantization and Sparsification for Multiclass Linear Classification with Strong Regularization
by: Ghane, Reza, et al.
Published: (2024) -
One-Bit Quantization for Random Features Models
by: Akhtiamov, Danil, et al.
Published: (2025) -
Implicit Bias and Convergence of Matrix Stochastic Mirror Descent
by: Akhtiamov, Danil, et al.
Published: (2026)