Towards Initialization-dependent and Non-vacuous Generalization Bounds for Overparameterized Shallow Neural Networks
Fuente:
arXiv
Saved in:
| Main Authors: | Lei, Yunwen, Xie, Yufeng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Prompts Generalize with Low Data: Non-vacuous Generalization Bounds for Optimizing Prompts with More Informative Priors
by: Madras, David, et al.
Published: (2025)
by: Madras, David, et al.
Published: (2025)
Provable Generalization in Overparameterized Neural Nets
by: Dhingra, Aviral
Published: (2025)
by: Dhingra, Aviral
Published: (2025)
Learning Theory of the SVRG: Generalization and Convergence Analysis
by: Lei, Yunwen, et al.
Published: (2026)
by: Lei, Yunwen, et al.
Published: (2026)
Minibatch and Local SGD: Algorithmic Stability and Linear Speedup in Generalization
by: Lei, Yunwen, et al.
Published: (2023)
by: Lei, Yunwen, et al.
Published: (2023)
Bootstrap SGD: Algorithmic Stability and Robustness
by: Christmann, Andreas, et al.
Published: (2024)
by: Christmann, Andreas, et al.
Published: (2024)
The Role of Symmetry in Optimizing Overparameterized Networks
by: Sareen, Kusha, et al.
Published: (2026)
by: Sareen, Kusha, et al.
Published: (2026)
The Spectral Bias of Shallow Neural Network Learning is Shaped by the Choice of Non-linearity
by: Sahs, Justin, et al.
Published: (2025)
by: Sahs, Justin, et al.
Published: (2025)
Generalization Bounds for Rank-sparse Neural Networks
by: Ledent, Antoine, et al.
Published: (2025)
by: Ledent, Antoine, et al.
Published: (2025)
Stochastic Gradient Descent with Momentum is Algorithmically Stable
by: Lei, Yunwen, et al.
Published: (2026)
by: Lei, Yunwen, et al.
Published: (2026)
Machine Unlearning under Overparameterization
by: Block, Jacob L., et al.
Published: (2025)
by: Block, Jacob L., et al.
Published: (2025)
From Shallow Bayesian Neural Networks to Gaussian Processes: General Convergence, Identifiability and Scalable Inference
by: de Araújo, Gracielle Antunes, et al.
Published: (2026)
by: de Araújo, Gracielle Antunes, et al.
Published: (2026)
On the Convergence of Overparameterized Problems: Inherent Properties of the Compositional Structure of Neural Networks
by: de Oliveira, Arthur Castello Branco, et al.
Published: (2025)
by: de Oliveira, Arthur Castello Branco, et al.
Published: (2025)
Entropic Confinement and Mode Connectivity in Overparameterized Neural Networks
by: Di Carlo, Luca, et al.
Published: (2025)
by: Di Carlo, Luca, et al.
Published: (2025)
Principal Components for Neural Network Initialization
by: Phan, Nhan, et al.
Published: (2025)
by: Phan, Nhan, et al.
Published: (2025)
Neural Network Verification with Branch-and-Bound for General Nonlinearities
by: Shi, Zhouxing, et al.
Published: (2024)
by: Shi, Zhouxing, et al.
Published: (2024)
Quotient Geometry, Effective Curvature, and Implicit Bias in Simple Shallow Neural Networks
by: Dong, Hang-Cheng, et al.
Published: (2026)
by: Dong, Hang-Cheng, et al.
Published: (2026)
Analysis of Overparameterization in Continual Learning under a Linear Model
by: Goldfarb, Daniel, et al.
Published: (2025)
by: Goldfarb, Daniel, et al.
Published: (2025)
Large Language Models Are Overparameterized Text Encoders
by: K, Thennal D, et al.
Published: (2024)
by: K, Thennal D, et al.
Published: (2024)
IDInit: A Universal and Stable Initialization Method for Neural Network Training
by: Pan, Yu, et al.
Published: (2025)
by: Pan, Yu, et al.
Published: (2025)
Robust Weight Initialization for Tanh Neural Networks with Fixed Point Analysis
by: Lee, Hyunwoo, et al.
Published: (2024)
by: Lee, Hyunwoo, et al.
Published: (2024)
Soft Contamination Means Benchmarks Test Shallow Generalization
by: Spiesberger, Ari, et al.
Published: (2026)
by: Spiesberger, Ari, et al.
Published: (2026)
On the Probabilistic Learnability of Compact Neural Network Preimage Bounds
by: Marzari, Luca, et al.
Published: (2025)
by: Marzari, Luca, et al.
Published: (2025)
Teasing Apart Architecture and Initial Weights as Sources of Inductive Bias in Neural Networks
by: Bencomo, Gianluca, et al.
Published: (2025)
by: Bencomo, Gianluca, et al.
Published: (2025)
Learning Non-Vacuous Generalization Bounds from Optimization
by: Tan, Chengli, et al.
Published: (2022)
by: Tan, Chengli, et al.
Published: (2022)
Towards Generalization of Graph Neural Networks for AC Optimal Power Flow
by: Arowolo, Olayiwola, et al.
Published: (2025)
by: Arowolo, Olayiwola, et al.
Published: (2025)
Variance-Dependent Regret Bounds for Non-stationary Linear Bandits
by: Wang, Zhiyong, et al.
Published: (2024)
by: Wang, Zhiyong, et al.
Published: (2024)
Non-convolutional Graph Neural Networks
by: Wang, Yuanqing, et al.
Published: (2024)
by: Wang, Yuanqing, et al.
Published: (2024)
A Generalization Bound for Nearly-Linear Networks
by: Golikov, Eugene
Published: (2024)
by: Golikov, Eugene
Published: (2024)
Compressible Dynamics in Deep Overparameterized Low-Rank Learning & Adaptation
by: Yaras, Can, et al.
Published: (2024)
by: Yaras, Can, et al.
Published: (2024)
LION-DG: Layer-Informed Initialization with Deep Gradient Protocols for Accelerated Neural Network Training
by: Kim, Hyunjun
Published: (2026)
by: Kim, Hyunjun
Published: (2026)
Solving Probabilistic Verification Problems of Neural Networks using Branch and Bound
by: Boetius, David, et al.
Published: (2024)
by: Boetius, David, et al.
Published: (2024)
Provably Bounding Neural Network Preimages
by: Kotha, Suhas, et al.
Published: (2023)
by: Kotha, Suhas, et al.
Published: (2023)
Effective Sample Size and Generalization Bounds for Temporal Networks
by: Gahtan, Barak, et al.
Published: (2025)
by: Gahtan, Barak, et al.
Published: (2025)
Deep Minds and Shallow Probes
by: Lee, Su Hyeong, et al.
Published: (2026)
by: Lee, Su Hyeong, et al.
Published: (2026)
Non-Euclidean Spatial Graph Neural Network
by: Zhang, Zheng, et al.
Published: (2023)
by: Zhang, Zheng, et al.
Published: (2023)
SpanGNN: Towards Memory-Efficient Graph Neural Networks via Spanning Subgraph Training
by: Gu, Xizhi, et al.
Published: (2024)
by: Gu, Xizhi, et al.
Published: (2024)
Scaling Laws and Spectra of Shallow Neural Networks in the Feature Learning Regime
by: Defilippis, Leonardo, et al.
Published: (2025)
by: Defilippis, Leonardo, et al.
Published: (2025)
Machine-Learning-Enhanced Non-Invasive Testing for MASLD Fibrosis: Shallow-Deep Neural Networks Versus FIB-4, Tabular Foundation Models, and Large Language Models
by: Angelakis, Athanasios, et al.
Published: (2026)
by: Angelakis, Athanasios, et al.
Published: (2026)
Network Interdiction Goes Neural
by: Zhang, Lei, et al.
Published: (2024)
by: Zhang, Lei, et al.
Published: (2024)
Knowledge Distillation in Wide Neural Networks: Risk Bound, Data Efficiency and Imperfect Teacher
by: Ji, Guangda, et al.
Published: (2020)
by: Ji, Guangda, et al.
Published: (2020)
Similar Items
-
Prompts Generalize with Low Data: Non-vacuous Generalization Bounds for Optimizing Prompts with More Informative Priors
by: Madras, David, et al.
Published: (2025) -
Provable Generalization in Overparameterized Neural Nets
by: Dhingra, Aviral
Published: (2025) -
Learning Theory of the SVRG: Generalization and Convergence Analysis
by: Lei, Yunwen, et al.
Published: (2026) -
Minibatch and Local SGD: Algorithmic Stability and Linear Speedup in Generalization
by: Lei, Yunwen, et al.
Published: (2023) -
Bootstrap SGD: Algorithmic Stability and Robustness
by: Christmann, Andreas, et al.
Published: (2024)