How many samples are needed to train a deep neural network?
Fuente:
arXiv
Saved in:
| Main Authors: | Golestaneh, Pegah, Taheri, Mahsa, Lederer, Johannes |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Regularization can make diffusion models more efficient
by: Taheri, Mahsa, et al.
Published: (2025)
by: Taheri, Mahsa, et al.
Published: (2025)
Statistical Guarantees for Approximate Stationary Points of Shallow Neural Networks
by: Taheri, Mahsa, et al.
Published: (2022)
by: Taheri, Mahsa, et al.
Published: (2022)
Non-asymptotic error bounds for probability flow ODEs under weak log-concavity
by: Kremling, Gitte, et al.
Published: (2025)
by: Kremling, Gitte, et al.
Published: (2025)
Adaptive tail index estimation: minimal assumptions and non-asymptotic guarantees
by: Lederer, Johannes, et al.
Published: (2025)
by: Lederer, Johannes, et al.
Published: (2025)
On the VC dimension of deep group convolutional neural networks
by: Sepliarskaia, Anna, et al.
Published: (2024)
by: Sepliarskaia, Anna, et al.
Published: (2024)
Affine Invariance in Continuous-Domain Convolutional Neural Networks
by: Mohaddes, Ali, et al.
Published: (2023)
by: Mohaddes, Ali, et al.
Published: (2023)
The loss landscape of deep linear neural networks: a second-order analysis
by: Achour, El Mehdi, et al.
Published: (2021)
by: Achour, El Mehdi, et al.
Published: (2021)
Posterior and variational inference for deep neural networks with heavy-tailed weights
by: Castillo, Ismaël, et al.
Published: (2024)
by: Castillo, Ismaël, et al.
Published: (2024)
How many asymmetric communities are there in multi-layer directed networks?
by: Qing, Huan
Published: (2026)
by: Qing, Huan
Published: (2026)
PAC-Bayesian risk bounds for fully connected deep neural network with Gaussian priors
by: Mai, The Tien
Published: (2025)
by: Mai, The Tien
Published: (2025)
Diffusion-based Denoising Beats Vanilla Score Matching in Parameter Estimation: A Theoretical Explanation
by: Schwienhorst, Benedikt Lütke, et al.
Published: (2026)
by: Schwienhorst, Benedikt Lütke, et al.
Published: (2026)
Nonlinear spiked covariance matrices and signal propagation in deep neural networks
by: Wang, Zhichao, et al.
Published: (2024)
by: Wang, Zhichao, et al.
Published: (2024)
Causal inference through multi-stage learning and doubly robust deep neural networks
by: Zhang, Yuqian, et al.
Published: (2024)
by: Zhang, Yuqian, et al.
Published: (2024)
Gaussian mixture layers for neural networks
by: Chewi, Sinho, et al.
Published: (2025)
by: Chewi, Sinho, et al.
Published: (2025)
On the rates of convergence for learning with convolutional neural networks
by: Yang, Yunfei, et al.
Published: (2024)
by: Yang, Yunfei, et al.
Published: (2024)
How many measurements are enough? Bayesian recovery in inverse problems with general distributions
by: Adcock, Ben, et al.
Published: (2025)
by: Adcock, Ben, et al.
Published: (2025)
Mind the spikes: Benign overfitting of kernels and neural networks in fixed dimension
by: Haas, Moritz, et al.
Published: (2023)
by: Haas, Moritz, et al.
Published: (2023)
The surrogate Gibbs-posterior of a corrected stochastic MALA: Towards uncertainty quantification for neural networks
by: Bieringer, Sebastian, et al.
Published: (2023)
by: Bieringer, Sebastian, et al.
Published: (2023)
Nonparametric regression using over-parameterized shallow ReLU neural networks
by: Yang, Yunfei, et al.
Published: (2023)
by: Yang, Yunfei, et al.
Published: (2023)
Precise gradient descent training dynamics for finite-width multi-layer neural networks
by: Han, Qiyang, et al.
Published: (2025)
by: Han, Qiyang, et al.
Published: (2025)
Sliding down the stairs: how correlated latent variables accelerate learning with neural networks
by: Bardone, Lorenzo, et al.
Published: (2024)
by: Bardone, Lorenzo, et al.
Published: (2024)
Do we really need the Rademacher complexities?
by: Bartl, Daniel, et al.
Published: (2025)
by: Bartl, Daniel, et al.
Published: (2025)
Information-theoretic reduction of deep neural networks to linear models in the overparametrized proportional regime
by: Camilli, Francesco, et al.
Published: (2025)
by: Camilli, Francesco, et al.
Published: (2025)
Gaussian random field approximation via Stein's method with applications to wide random neural networks
by: Balasubramanian, Krishnakumar, et al.
Published: (2023)
by: Balasubramanian, Krishnakumar, et al.
Published: (2023)
Statistically guided deep learning
by: Kohler, Michael, et al.
Published: (2025)
by: Kohler, Michael, et al.
Published: (2025)
Nonparametric logistic regression with deep learning
by: Yara, Atsutomo, et al.
Published: (2024)
by: Yara, Atsutomo, et al.
Published: (2024)
A general framework for deep learning
by: Kengne, William, et al.
Published: (2025)
by: Kengne, William, et al.
Published: (2025)
Higher-order accurate two-sample network inference and network hashing
by: Shao, Meijia, et al.
Published: (2022)
by: Shao, Meijia, et al.
Published: (2022)
How many labelers do you have? A closer look at gold-standard labels
by: Cheng, Chen, et al.
Published: (2022)
by: Cheng, Chen, et al.
Published: (2022)
Robust deep learning from weakly dependent data
by: Kengne, William, et al.
Published: (2024)
by: Kengne, William, et al.
Published: (2024)
Misclassification bounds for PAC-Bayesian sparse deep learning
by: Mai, The Tien
Published: (2024)
by: Mai, The Tien
Published: (2024)
Gradient descent for deep equilibrium single-index models
by: Dandapanthula, Sanjit, et al.
Published: (2025)
by: Dandapanthula, Sanjit, et al.
Published: (2025)
Adaptive deep learning for nonlinear time series models
by: Kurisu, Daisuke, et al.
Published: (2022)
by: Kurisu, Daisuke, et al.
Published: (2022)
Towards a mathematical theory for consistency training in diffusion models
by: Li, Gen, et al.
Published: (2024)
by: Li, Gen, et al.
Published: (2024)
Streaming data recovery via Bayesian tensor train decomposition
by: Huang, Yunyu, et al.
Published: (2023)
by: Huang, Yunyu, et al.
Published: (2023)
Marginal and training-conditional guarantees in one-shot federated conformal prediction
by: Humbert, Pierre, et al.
Published: (2024)
by: Humbert, Pierre, et al.
Published: (2024)
A Statistical Theory of Contrastive Pre-training and Multimodal Generative AI
by: Oko, Kazusato, et al.
Published: (2025)
by: Oko, Kazusato, et al.
Published: (2025)
The twin peaks of learning neural networks
by: Demyanenko, Elizaveta, et al.
Published: (2024)
by: Demyanenko, Elizaveta, et al.
Published: (2024)
Optimal training-conditional regret for online conformal prediction
by: Liang, Jiadong, et al.
Published: (2026)
by: Liang, Jiadong, et al.
Published: (2026)
The curse of overparametrization in adversarial training: Precise analysis of robust generalization for random features regression
by: Hassani, Hamed, et al.
Published: (2022)
by: Hassani, Hamed, et al.
Published: (2022)
Similar Items
-
Regularization can make diffusion models more efficient
by: Taheri, Mahsa, et al.
Published: (2025) -
Statistical Guarantees for Approximate Stationary Points of Shallow Neural Networks
by: Taheri, Mahsa, et al.
Published: (2022) -
Non-asymptotic error bounds for probability flow ODEs under weak log-concavity
by: Kremling, Gitte, et al.
Published: (2025) -
Adaptive tail index estimation: minimal assumptions and non-asymptotic guarantees
by: Lederer, Johannes, et al.
Published: (2025) -
On the VC dimension of deep group convolutional neural networks
by: Sepliarskaia, Anna, et al.
Published: (2024)