Implicit Bias and Convergence of Matrix Stochastic Mirror Descent
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Akhtiamov, Danil, Ghane, Reza, Pooladzandi, Omead, Hassibi, Babak |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Precise Performance of Linear Denoisers in the Proportional Regime
von: Ghane, Reza, et al.
Veröffentlicht: (2026)
von: Ghane, Reza, et al.
Veröffentlicht: (2026)
Dual Space Preconditioning for Gradient Descent in the Overparameterized Regime
von: Ghane, Reza, et al.
Veröffentlicht: (2026)
von: Ghane, Reza, et al.
Veröffentlicht: (2026)
Universality in Transfer Learning for Linear Models
von: Ghane, Reza, et al.
Veröffentlicht: (2024)
von: Ghane, Reza, et al.
Veröffentlicht: (2024)
One-Bit Quantization and Sparsification for Multiclass Linear Classification with Strong Regularization
von: Ghane, Reza, et al.
Veröffentlicht: (2024)
von: Ghane, Reza, et al.
Veröffentlicht: (2024)
One-Bit Quantization for Random Features Models
von: Akhtiamov, Danil, et al.
Veröffentlicht: (2025)
von: Akhtiamov, Danil, et al.
Veröffentlicht: (2025)
Gaussian Universality for Diffusion Models
von: Ghane, Reza, et al.
Veröffentlicht: (2025)
von: Ghane, Reza, et al.
Veröffentlicht: (2025)
Beyond Quadratic Costs in LQR: Bregman Divergence Control
von: Hassibi, Babak, et al.
Veröffentlicht: (2025)
von: Hassibi, Babak, et al.
Veröffentlicht: (2025)
Convergence and Implicit Bias of Gradient Descent on Continual Linear Classification
von: Jung, Hyunji, et al.
Veröffentlicht: (2025)
von: Jung, Hyunji, et al.
Veröffentlicht: (2025)
A Novel Gaussian Min-Max Theorem and its Applications
von: Akhtiamov, Danil, et al.
Veröffentlicht: (2024)
von: Akhtiamov, Danil, et al.
Veröffentlicht: (2024)
A Precise Performance Analysis of the Randomized Singular Value Decomposition
von: Akhtiamov, Danil, et al.
Veröffentlicht: (2025)
von: Akhtiamov, Danil, et al.
Veröffentlicht: (2025)
On the Convergence of Policy in Unregularized Policy Mirror Descent
von: Lin, Dachao, et al.
Veröffentlicht: (2022)
von: Lin, Dachao, et al.
Veröffentlicht: (2022)
On the Convergence of Policy Mirror Descent with Temporal Difference Evaluation
von: Liu, Jiacai, et al.
Veröffentlicht: (2025)
von: Liu, Jiacai, et al.
Veröffentlicht: (2025)
Implicit Bias of Mirror Flow on Separable Data
von: Pesme, Scott, et al.
Veröffentlicht: (2024)
von: Pesme, Scott, et al.
Veröffentlicht: (2024)
Convergence of Policy Mirror Descent Beyond Compatible Function Approximation
von: Sherman, Uri, et al.
Veröffentlicht: (2025)
von: Sherman, Uri, et al.
Veröffentlicht: (2025)
Convergence of Alternating Gradient Descent for Matrix Factorization
von: Ward, Rachel, et al.
Veröffentlicht: (2023)
von: Ward, Rachel, et al.
Veröffentlicht: (2023)
Implicit Bias of Gradient Descent for Non-Homogeneous Deep Networks
von: Cai, Yuhang, et al.
Veröffentlicht: (2025)
von: Cai, Yuhang, et al.
Veröffentlicht: (2025)
Implicit Bias of Spectral Descent and Muon on Multiclass Separable Data
von: Fan, Chen, et al.
Veröffentlicht: (2025)
von: Fan, Chen, et al.
Veröffentlicht: (2025)
Zeroth-Order Stochastic Mirror Descent Algorithms for Minimax Excess Risk Optimization
von: Gu, Zhihao, et al.
Veröffentlicht: (2024)
von: Gu, Zhihao, et al.
Veröffentlicht: (2024)
Convergence Analysis of Stochastic Gradient Descent with MCMC Estimators
von: Li, Tianyou, et al.
Veröffentlicht: (2023)
von: Li, Tianyou, et al.
Veröffentlicht: (2023)
Convergence of Gradient Descent with Small Initialization for Unregularized Matrix Completion
von: Ma, Jianhao, et al.
Veröffentlicht: (2024)
von: Ma, Jianhao, et al.
Veröffentlicht: (2024)
Implicit Bias and Fast Convergence Rates for Self-attention
von: Vasudeva, Bhavya, et al.
Veröffentlicht: (2024)
von: Vasudeva, Bhavya, et al.
Veröffentlicht: (2024)
Mirror Descent on Riemannian Manifolds
von: Jiang, Jiaxin, et al.
Veröffentlicht: (2026)
von: Jiang, Jiaxin, et al.
Veröffentlicht: (2026)
Parameter-free Mirror Descent
von: Jacobsen, Andrew, et al.
Veröffentlicht: (2022)
von: Jacobsen, Andrew, et al.
Veröffentlicht: (2022)
A Mirror Descent Perspective of Smoothed Sign Descent
von: Wang, Shuyang, et al.
Veröffentlicht: (2024)
von: Wang, Shuyang, et al.
Veröffentlicht: (2024)
Convergence of Steepest Descent and Adam under Non-Uniform Smoothness
von: Vaswani, Sharan, et al.
Veröffentlicht: (2026)
von: Vaswani, Sharan, et al.
Veröffentlicht: (2026)
Exponential Convergence of (Stochastic) Gradient Descent for Separable Logistic Regression
von: Kale, Sacchit, et al.
Veröffentlicht: (2026)
von: Kale, Sacchit, et al.
Veröffentlicht: (2026)
On the Convergence of Stochastic Gradient Descent with Perturbed Forward-Backward Passes
von: Kong, Boao, et al.
Veröffentlicht: (2026)
von: Kong, Boao, et al.
Veröffentlicht: (2026)
Faster Convergence of Stochastic Accelerated Gradient Descent under Interpolation
von: Mishkin, Aaron, et al.
Veröffentlicht: (2024)
von: Mishkin, Aaron, et al.
Veröffentlicht: (2024)
Nonsmooth Implicit Differentiation: Deterministic and Stochastic Convergence Rates
von: Grazzi, Riccardo, et al.
Veröffentlicht: (2024)
von: Grazzi, Riccardo, et al.
Veröffentlicht: (2024)
Robust Mean Estimation With Auxiliary Samples
von: Han, Barron, et al.
Veröffentlicht: (2025)
von: Han, Barron, et al.
Veröffentlicht: (2025)
Convergence of Implicit Gradient Descent for Training Two-Layer Physics-Informed Neural Networks
von: Xu, Xianliang, et al.
Veröffentlicht: (2024)
von: Xu, Xianliang, et al.
Veröffentlicht: (2024)
Faster Convergence of Riemannian Stochastic Gradient Descent with Increasing Batch Size
von: Oowada, Kanata, et al.
Veröffentlicht: (2025)
von: Oowada, Kanata, et al.
Veröffentlicht: (2025)
Coupling-based Convergence Diagnostic and Stepsize Scheme for Stochastic Gradient Descent
von: Li, Xiang, et al.
Veröffentlicht: (2024)
von: Li, Xiang, et al.
Veröffentlicht: (2024)
Armijo Line-search Can Make (Stochastic) Gradient Descent Provably Faster
von: Vaswani, Sharan, et al.
Veröffentlicht: (2025)
von: Vaswani, Sharan, et al.
Veröffentlicht: (2025)
Mirror Descent on Reproducing Kernel Banach Spaces
von: Kumar, Akash, et al.
Veröffentlicht: (2024)
von: Kumar, Akash, et al.
Veröffentlicht: (2024)
Mirror and Preconditioned Gradient Descent in Wasserstein Space
von: Bonet, Clément, et al.
Veröffentlicht: (2024)
von: Bonet, Clément, et al.
Veröffentlicht: (2024)
Using Stochastic Gradient Descent to Smooth Nonconvex Functions: Analysis of Implicit Graduated Optimization
von: Sato, Naoki, et al.
Veröffentlicht: (2023)
von: Sato, Naoki, et al.
Veröffentlicht: (2023)
Towards Noise-adaptive, Problem-adaptive (Accelerated) Stochastic Gradient Descent
von: Vaswani, Sharan, et al.
Veröffentlicht: (2021)
von: Vaswani, Sharan, et al.
Veröffentlicht: (2021)
A Novel Framework for Policy Mirror Descent with General Parameterization and Linear Convergence
von: Alfano, Carlo, et al.
Veröffentlicht: (2023)
von: Alfano, Carlo, et al.
Veröffentlicht: (2023)
Implicit Bias in Matrix Factorization and its Explicit Realization in a New Architecture
von: Hou, Yikun, et al.
Veröffentlicht: (2025)
von: Hou, Yikun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Precise Performance of Linear Denoisers in the Proportional Regime
von: Ghane, Reza, et al.
Veröffentlicht: (2026) -
Dual Space Preconditioning for Gradient Descent in the Overparameterized Regime
von: Ghane, Reza, et al.
Veröffentlicht: (2026) -
Universality in Transfer Learning for Linear Models
von: Ghane, Reza, et al.
Veröffentlicht: (2024) -
One-Bit Quantization and Sparsification for Multiclass Linear Classification with Strong Regularization
von: Ghane, Reza, et al.
Veröffentlicht: (2024) -
One-Bit Quantization for Random Features Models
von: Akhtiamov, Danil, et al.
Veröffentlicht: (2025)