Asymptotic convexity of wide and shallow neural networks
Fuente:
arXiv
Saved in:
| Main Authors: | Borkar, Vivek, Pandit, Parthe |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adynamical systems view of training generativemodels and the memorization phenomenon
by: Athreya, Siva, et al.
Published: (2026)
by: Athreya, Siva, et al.
Published: (2026)
A theoretical basis for model collapse in recursive training
by: Borkar, Vivek Shripad
Published: (2025)
by: Borkar, Vivek Shripad
Published: (2025)
Generalization error bounds for two-layer neural networks with Lipschitz loss function
by: Nguwi, Jiang Yu, et al.
Published: (2026)
by: Nguwi, Jiang Yu, et al.
Published: (2026)
Genus expansion for non-linear random matrix ensembles with applications to neural networks
by: Cirone, Nicola Muca, et al.
Published: (2024)
by: Cirone, Nicola Muca, et al.
Published: (2024)
In almost all shallow analytic neural network optimization landscapes, efficient minimizers have strongly convex neighborhoods
by: Benning, Felix, et al.
Published: (2025)
by: Benning, Felix, et al.
Published: (2025)
On the Rashomon ratio of infinite hypothesis sets
by: Coupkova, Evzenie, et al.
Published: (2024)
by: Coupkova, Evzenie, et al.
Published: (2024)
Universal Approximation Theorem and error bounds for quantum neural networks and quantum reservoirs
by: Gonon, Lukas, et al.
Published: (2023)
by: Gonon, Lukas, et al.
Published: (2023)
A Lipschitz spaces view of infinitely wide shallow neural networks
by: Bartolucci, Francesca, et al.
Published: (2024)
by: Bartolucci, Francesca, et al.
Published: (2024)
Spectral complexity of deep neural networks
by: Di Lillo, Simmaco, et al.
Published: (2024)
by: Di Lillo, Simmaco, et al.
Published: (2024)
Phase Transitions in the Fluctuations of Functionals of Random Neural Networks
by: Di Lillo, Simmaco, et al.
Published: (2026)
by: Di Lillo, Simmaco, et al.
Published: (2026)
Large Deviations of Gaussian Neural Networks with ReLU activation
by: Vogel, Quirin
Published: (2024)
by: Vogel, Quirin
Published: (2024)
The Feature Speed Formula: a flexible approach to scale hyper-parameters of deep neural networks
by: Chizat, Lénaïc, et al.
Published: (2023)
by: Chizat, Lénaïc, et al.
Published: (2023)
Concentration of measure for non-linear random matrices with applications to neural networks and non-commutative polynomials
by: Adamczak, Radosław
Published: (2025)
by: Adamczak, Radosław
Published: (2025)
Upper and lower bounds for the Lipschitz constant of random neural networks
by: Geuchen, Paul, et al.
Published: (2023)
by: Geuchen, Paul, et al.
Published: (2023)
A Generalization Bound for a Family of Implicit Networks
by: Fung, Samy Wu, et al.
Published: (2024)
by: Fung, Samy Wu, et al.
Published: (2024)
Regime-Aware Conditional Neural Processes with Multi-Criteria Decision Support for Operational Electricity Price Forecasting
by: Das, Abhinav, et al.
Published: (2025)
by: Das, Abhinav, et al.
Published: (2025)
Mathematical Introduction to Deep Learning: Methods, Implementations, and Theory
by: Jentzen, Arnulf, et al.
Published: (2023)
by: Jentzen, Arnulf, et al.
Published: (2023)
Near-optimal estimates for the $\ell^p$-Lipschitz constants of deep random ReLU neural networks
by: Dirksen, Sjoerd, et al.
Published: (2025)
by: Dirksen, Sjoerd, et al.
Published: (2025)
Convergence of gradient descent for deep neural networks
by: Chatterjee, Sourav
Published: (2022)
by: Chatterjee, Sourav
Published: (2022)
A ZeNN architecture to avoid the Gaussian trap
by: Carvalho, Luís, et al.
Published: (2025)
by: Carvalho, Luís, et al.
Published: (2025)
Convergence Analysis of Real-time Recurrent Learning (RTRL) for a class of Recurrent Neural Networks
by: Lam, Samuel Chun-Hei, et al.
Published: (2025)
by: Lam, Samuel Chun-Hei, et al.
Published: (2025)
The Positivity of the Neural Tangent Kernel
by: Carvalho, Luís, et al.
Published: (2024)
by: Carvalho, Luís, et al.
Published: (2024)
Kernel Limit for a Class of Recurrent Neural Networks Trained on Ergodic Data Sequences
by: Lam, Samuel Chun-Hei, et al.
Published: (2023)
by: Lam, Samuel Chun-Hei, et al.
Published: (2023)
On the existence of minimizers in shallow residual ReLU neural network optimization landscapes
by: Dereich, Steffen, et al.
Published: (2023)
by: Dereich, Steffen, et al.
Published: (2023)
Hybrid deep additive neural networks
by: Kim, Gyu Min, et al.
Published: (2024)
by: Kim, Gyu Min, et al.
Published: (2024)
Entropic bounds for conditionally Gaussian vectors and applications to neural networks
by: Celli, Lucia, et al.
Published: (2025)
by: Celli, Lucia, et al.
Published: (2025)
Deep neural networks with dependent weights: Gaussian Process mixture limit, heavy tails, sparsity and compressibility
by: Lee, Hoil, et al.
Published: (2022)
by: Lee, Hoil, et al.
Published: (2022)
Efficient kernel surrogates for neural network-based regression
by: Qadeer, Saad, et al.
Published: (2023)
by: Qadeer, Saad, et al.
Published: (2023)
A new architecture of high-order deep neural networks that learn martingales
by: Ninomiya, Syoiti, et al.
Published: (2025)
by: Ninomiya, Syoiti, et al.
Published: (2025)
Large and moderate deviations for Gaussian neural networks
by: Macci, Claudio, et al.
Published: (2024)
by: Macci, Claudio, et al.
Published: (2024)
Universality of Kernel Random Matrices and Kernel Regression in the Quadratic Regime
by: Pandit, Parthe, et al.
Published: (2024)
by: Pandit, Parthe, et al.
Published: (2024)
Large deviations of one-hidden-layer neural networks
by: Hirsch, Christian, et al.
Published: (2024)
by: Hirsch, Christian, et al.
Published: (2024)
Comparison of generalised additive models and neural networks in applications: A systematic review
by: Doohan, Jessica, et al.
Published: (2025)
by: Doohan, Jessica, et al.
Published: (2025)
Overlap-aware meta-learning attention to enhance hypergraph neural networks for node classification
by: Yang, Murong, et al.
Published: (2025)
by: Yang, Murong, et al.
Published: (2025)
Explicit neural network classifiers for non-separable data
by: Ewald, Patrícia Muñoz
Published: (2025)
by: Ewald, Patrícia Muñoz
Published: (2025)
Global law of conjugate kernel random matrices with heavy-tailed weights
by: Guionnet, Alice, et al.
Published: (2025)
by: Guionnet, Alice, et al.
Published: (2025)
Posterior Bayesian Neural Networks with Dependent Weights
by: Apollonio, Nicola, et al.
Published: (2025)
by: Apollonio, Nicola, et al.
Published: (2025)
Separation capacity of linear reservoirs with random connectivity matrix
by: Boutaib, Youness
Published: (2024)
by: Boutaib, Youness
Published: (2024)
Bayesian sparsification for deep neural networks with Bayesian model reduction
by: Marković, Dimitrije, et al.
Published: (2023)
by: Marković, Dimitrije, et al.
Published: (2023)
Dimensionality reduction and width of deep neural networks based on topological degree theory
by: Yang, Xiao-Song
Published: (2025)
by: Yang, Xiao-Song
Published: (2025)
Similar Items
-
Adynamical systems view of training generativemodels and the memorization phenomenon
by: Athreya, Siva, et al.
Published: (2026) -
A theoretical basis for model collapse in recursive training
by: Borkar, Vivek Shripad
Published: (2025) -
Generalization error bounds for two-layer neural networks with Lipschitz loss function
by: Nguwi, Jiang Yu, et al.
Published: (2026) -
Genus expansion for non-linear random matrix ensembles with applications to neural networks
by: Cirone, Nicola Muca, et al.
Published: (2024) -
In almost all shallow analytic neural network optimization landscapes, efficient minimizers have strongly convex neighborhoods
by: Benning, Felix, et al.
Published: (2025)