Generalization Guarantees of Gradient Descent for Multi-Layer Neural Networks
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Puyu, Lei, Yunwen, Wang, Di, Ying, Yiming, Zhou, Ding-Xuan |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Optimal Rates for Generalization of Gradient Descent for Deep ReLU Classification
by: Li, Yuanfan, et al.
Published: (2025)
by: Li, Yuanfan, et al.
Published: (2025)
Optimization, Generalization and Differential Privacy Bounds for Gradient Descent on Kolmogorov-Arnold Networks
by: Wang, Puyu, et al.
Published: (2026)
by: Wang, Puyu, et al.
Published: (2026)
Stochastic Gradient Descent with Momentum is Algorithmically Stable
by: Lei, Yunwen, et al.
Published: (2026)
by: Lei, Yunwen, et al.
Published: (2026)
Generalization analysis with deep ReLU networks for metric and similarity learning
by: Zhou, Junyu, et al.
Published: (2024)
by: Zhou, Junyu, et al.
Published: (2024)
Stability-based Generalization Analysis of Randomized Coordinate Descent for Pairwise Learning
by: Wu, Liang, et al.
Published: (2025)
by: Wu, Liang, et al.
Published: (2025)
Generalization Bounds for Rank-sparse Neural Networks
by: Ledent, Antoine, et al.
Published: (2025)
by: Ledent, Antoine, et al.
Published: (2025)
On Discriminative Probabilistic Modeling for Self-Supervised Representation Learning
by: Wang, Bokun, et al.
Published: (2024)
by: Wang, Bokun, et al.
Published: (2024)
Beyond Cross-Validation: Adaptive Parameter Selection for Kernel-Based Gradient Descents
by: Liu, Xiaotong, et al.
Published: (2026)
by: Liu, Xiaotong, et al.
Published: (2026)
Towards Initialization-dependent and Non-vacuous Generalization Bounds for Overparameterized Shallow Neural Networks
by: Lei, Yunwen, et al.
Published: (2026)
by: Lei, Yunwen, et al.
Published: (2026)
Recovery Guarantees of Unsupervised Neural Networks for Inverse Problems trained with Gradient Descent
by: Buskulic, Nathan, et al.
Published: (2024)
by: Buskulic, Nathan, et al.
Published: (2024)
Distributed Gradient Descent for Functional Learning
by: Yu, Zhan, et al.
Published: (2023)
by: Yu, Zhan, et al.
Published: (2023)
Generalization Bounds of Stochastic Gradient Descent in Homogeneous Neural Networks
by: Ma, Wenquan, et al.
Published: (2026)
by: Ma, Wenquan, et al.
Published: (2026)
Stochastic Gradient Descent for Two-layer Neural Networks
by: Cao, Dinghao, et al.
Published: (2024)
by: Cao, Dinghao, et al.
Published: (2024)
Convergence of Implicit Gradient Descent for Training Two-Layer Physics-Informed Neural Networks
by: Xu, Xianliang, et al.
Published: (2024)
by: Xu, Xianliang, et al.
Published: (2024)
Fine-grained Analysis of Non-parametric Estimation for Pairwise Learning
by: Zhou, Junyu, et al.
Published: (2023)
by: Zhou, Junyu, et al.
Published: (2023)
How Does Gradient Descent Learn Features -- A Local Analysis for Regularized Two-Layer Neural Networks
by: Zhou, Mo, et al.
Published: (2024)
by: Zhou, Mo, et al.
Published: (2024)
Generalization Guarantees on Data-Driven Tuning of Gradient Descent with Langevin Updates
by: Goyal, Saumya, et al.
Published: (2026)
by: Goyal, Saumya, et al.
Published: (2026)
Generalization and Optimization of SGD with Lookahead
by: Li, Kangcheng, et al.
Published: (2025)
by: Li, Kangcheng, et al.
Published: (2025)
Towards Understanding Generalization in DP-GD: A Case Study in Training Two-Layer CNNs
by: Shi, Zhongjie, et al.
Published: (2025)
by: Shi, Zhongjie, et al.
Published: (2025)
Statistical Guarantees for High-Dimensional Stochastic Gradient Descent
by: Li, Jiaqi, et al.
Published: (2025)
by: Li, Jiaqi, et al.
Published: (2025)
Learning Theory of the SVRG: Generalization and Convergence Analysis
by: Lei, Yunwen, et al.
Published: (2026)
by: Lei, Yunwen, et al.
Published: (2026)
Solving Inverse Problems with Deep Linear Neural Networks: Global Convergence Guarantees for Gradient Descent with Weight Decay
by: Laus, Hannah, et al.
Published: (2025)
by: Laus, Hannah, et al.
Published: (2025)
Gradient Flow Convergence Guarantee for General Neural Network Architectures
by: Jakhmola, Yash
Published: (2025)
by: Jakhmola, Yash
Published: (2025)
Variational Stochastic Gradient Descent for Deep Neural Networks
by: Chen, Haotian, et al.
Published: (2024)
by: Chen, Haotian, et al.
Published: (2024)
Classification with Deep Neural Networks and Logistic Loss
by: Zhang, Zihan, et al.
Published: (2023)
by: Zhang, Zihan, et al.
Published: (2023)
Do Neural Networks Need Gradient Descent to Generalize? A Theoretical Study
by: Alexander, Yotam, et al.
Published: (2025)
by: Alexander, Yotam, et al.
Published: (2025)
Gradient Manipulation in Distributed Stochastic Gradient Descent with Strategic Agents: Truthful Incentives with Convergence Guarantees
by: Chen, Ziqin, et al.
Published: (2026)
by: Chen, Ziqin, et al.
Published: (2026)
Optimal Convergence Rates of Deep Neural Network Classifiers
by: Zhang, Zihan, et al.
Published: (2025)
by: Zhang, Zihan, et al.
Published: (2025)
A Theoretical Analysis of Noise Geometry in Stochastic Gradient Descent
by: Wang, Mingze, et al.
Published: (2023)
by: Wang, Mingze, et al.
Published: (2023)
Hybrid Coordinate Descent for Efficient Neural Network Learning Using Line Search and Gradient Descent
by: Hsiao, Yen-Che, et al.
Published: (2024)
by: Hsiao, Yen-Che, et al.
Published: (2024)
Convergence Analysis of Natural Gradient Descent for Over-parameterized Physics-Informed Neural Networks
by: Xu, Xianliang, et al.
Published: (2024)
by: Xu, Xianliang, et al.
Published: (2024)
Randomized Pairwise Learning with Adaptive Sampling: A PAC-Bayes Analysis
by: Zhou, Sijia, et al.
Published: (2025)
by: Zhou, Sijia, et al.
Published: (2025)
On the Theory of Continual Learning with Gradient Descent for Neural Networks
by: Taheri, Hossein, et al.
Published: (2025)
by: Taheri, Hossein, et al.
Published: (2025)
The Double Descent Behavior in Two Layer Neural Network for Binary Classification
by: Abeykoon, Chathurika S, et al.
Published: (2025)
by: Abeykoon, Chathurika S, et al.
Published: (2025)
Gradient Descent Finds Over-Parameterized Neural Networks with Sharp Generalization for Nonparametric Regression
by: Yang, Yingzhen, et al.
Published: (2024)
by: Yang, Yingzhen, et al.
Published: (2024)
On Multi-Stage Loss Dynamics in Neural Networks: Mechanisms of Plateau and Descent Stages
by: Chen, Zheng-An, et al.
Published: (2024)
by: Chen, Zheng-An, et al.
Published: (2024)
Layer-diverse Negative Sampling for Graph Neural Networks
by: Duan, Wei, et al.
Published: (2024)
by: Duan, Wei, et al.
Published: (2024)
Minibatch and Local SGD: Algorithmic Stability and Linear Speedup in Generalization
by: Lei, Yunwen, et al.
Published: (2023)
by: Lei, Yunwen, et al.
Published: (2023)
Gradient Descent Algorithm Survey
by: Fucheng, Deng, et al.
Published: (2025)
by: Fucheng, Deng, et al.
Published: (2025)
Provable Guarantees for Nonlinear Feature Learning in Three-Layer Neural Networks
by: Nichani, Eshaan, et al.
Published: (2023)
by: Nichani, Eshaan, et al.
Published: (2023)
Similar Items
-
Optimal Rates for Generalization of Gradient Descent for Deep ReLU Classification
by: Li, Yuanfan, et al.
Published: (2025) -
Optimization, Generalization and Differential Privacy Bounds for Gradient Descent on Kolmogorov-Arnold Networks
by: Wang, Puyu, et al.
Published: (2026) -
Stochastic Gradient Descent with Momentum is Algorithmically Stable
by: Lei, Yunwen, et al.
Published: (2026) -
Generalization analysis with deep ReLU networks for metric and similarity learning
by: Zhou, Junyu, et al.
Published: (2024) -
Stability-based Generalization Analysis of Randomized Coordinate Descent for Pairwise Learning
by: Wu, Liang, et al.
Published: (2025)