Exact capacity of the \emph{wide} hidden layer treelike neural networks with generic activations
Fuente:
arXiv
Saved in:
| Main Author: | Stojnic, Mihailo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fixed width treelike neural networks capacity analysis -- generic activations
by: Stojnic, Mihailo
Published: (2024)
by: Stojnic, Mihailo
Published: (2024)
Ultrametric OGP - parametric RDT \emph{symmetric} binary perceptron connection
by: Stojnic, Mihailo
Published: (2026)
by: Stojnic, Mihailo
Published: (2026)
Capacity of the Hebbian-Hopfield network associative memory
by: Stojnic, Mihailo
Published: (2024)
by: Stojnic, Mihailo
Published: (2024)
Binary perceptron computational gap -- a parametric fl RDT view
by: Stojnic, Mihailo
Published: (2025)
by: Stojnic, Mihailo
Published: (2025)
Parametric RDT approach to computational gap of symmetric binary perceptron
by: Stojnic, Mihailo
Published: (2026)
by: Stojnic, Mihailo
Published: (2026)
Deep ReLU networks -- injectivity capacity upper bounds
by: Stojnic, Mihailo
Published: (2024)
by: Stojnic, Mihailo
Published: (2024)
Injectivity capacity of ReLU gates
by: Stojnic, Mihailo
Published: (2024)
by: Stojnic, Mihailo
Published: (2024)
Rare dense solutions clusters in asymmetric binary perceptrons -- local entropy via fully lifted RDT
by: Stojnic, Mihailo
Published: (2025)
by: Stojnic, Mihailo
Published: (2025)
Ground state energies of multipartite $p$-spin models -- partially lifted RDT view
by: Stojnic, Mihailo
Published: (2025)
by: Stojnic, Mihailo
Published: (2025)
CLuP practically achieves $\sim 1.77$ positive and $\sim 0.33$ negative Hopfield model ground state free energy
by: Stojnic, Mihailo
Published: (2025)
by: Stojnic, Mihailo
Published: (2025)
A CLuP algorithm to practically achieve $\sim 0.76$ SK--model ground state free energy
by: Stojnic, Mihailo
Published: (2025)
by: Stojnic, Mihailo
Published: (2025)
Exact full-RSB SAT/UNSAT transition in infinitely wide two-layer neural networks
by: Annesi, Brandon L., et al.
Published: (2024)
by: Annesi, Brandon L., et al.
Published: (2024)
Bayes-optimal learning of an extensive-width neural network from quadratically many samples
by: Maillard, Antoine, et al.
Published: (2024)
by: Maillard, Antoine, et al.
Published: (2024)
Generalization performance of narrow one-hidden layer networks in the teacher-student setting
by: Ortiz, Rodrigo Pérez, et al.
Published: (2025)
by: Ortiz, Rodrigo Pérez, et al.
Published: (2025)
Fundamental Limits of Matrix Sensing: Exact Asymptotics, Universality, and Applications
by: Xu, Yizhou, et al.
Published: (2025)
by: Xu, Yizhou, et al.
Published: (2025)
Solution space and storage capacity of fully connected two-layer neural networks with generic activation functions
by: Nishiyama, Sota, et al.
Published: (2024)
by: Nishiyama, Sota, et al.
Published: (2024)
The twin peaks of learning neural networks
by: Demyanenko, Elizaveta, et al.
Published: (2024)
by: Demyanenko, Elizaveta, et al.
Published: (2024)
High-dimensional manifold of solutions in neural networks: insights from statistical physics
by: Malatesta, Enrico M.
Published: (2023)
by: Malatesta, Enrico M.
Published: (2023)
Fully lifted \emph{blirp} interpolation -- a large deviation view
by: Stojnic, Mihailo
Published: (2025)
by: Stojnic, Mihailo
Published: (2025)
A large deviation view of \emph{stationarized} fully lifted blirp interpolation
by: Stojnic, Mihailo
Published: (2025)
by: Stojnic, Mihailo
Published: (2025)
Properties of the geometry of solutions and capacity of multi-layer neural networks with Rectified Linear Units activations
by: Baldassi, Carlo, et al.
Published: (2019)
by: Baldassi, Carlo, et al.
Published: (2019)
Analysis of Diffusion Models for Manifold Data
by: George, Anand Jerry, et al.
Published: (2025)
by: George, Anand Jerry, et al.
Published: (2025)
Storage capacity of perceptron with variable selection
by: Xu, Yingying, et al.
Published: (2025)
by: Xu, Yingying, et al.
Published: (2025)
The maximum-average subtensor problem: equilibrium and out-of-equilibrium properties
by: Erba, Vittorio, et al.
Published: (2025)
by: Erba, Vittorio, et al.
Published: (2025)
A universal compression theory for lottery ticket hypothesis and neural scaling laws
by: Wang, Hong-Yi, et al.
Published: (2025)
by: Wang, Hong-Yi, et al.
Published: (2025)
Injectivity of ReLU networks: perspectives from statistical physics
by: Maillard, Antoine, et al.
Published: (2023)
by: Maillard, Antoine, et al.
Published: (2023)
Statistical mechanics of extensive-width Bayesian neural networks near interpolation
by: Barbier, Jean, et al.
Published: (2025)
by: Barbier, Jean, et al.
Published: (2025)
Optimal generalisation and learning transition in extensive-width shallow neural networks near interpolation
by: Barbier, Jean, et al.
Published: (2025)
by: Barbier, Jean, et al.
Published: (2025)
A Fourier perspective on the learning dynamics of neural networks: from sample complexities to mechanistic insights
by: Ricci, Fabiola, et al.
Published: (2026)
by: Ricci, Fabiola, et al.
Published: (2026)
Sharp feature-learning transitions and Bayes-optimal neural scaling laws in extensive-width networks
by: Nguyen, Minh-Toan, et al.
Published: (2026)
by: Nguyen, Minh-Toan, et al.
Published: (2026)
Spectral Phase Transition and Optimal PCA in Block-Structured Spiked models
by: Mergny, Pierre, et al.
Published: (2024)
by: Mergny, Pierre, et al.
Published: (2024)
Precise asymptotic analysis of Sobolev training for random feature models
by: Fisher, Katharine E, et al.
Published: (2025)
by: Fisher, Katharine E, et al.
Published: (2025)
Gaussian Universality of Perceptrons with Random Labels
by: Gerace, Federica, et al.
Published: (2022)
by: Gerace, Federica, et al.
Published: (2022)
The Random Subsequence Model and Uniform Codes for the Deletion Channel
by: Jeong, Ryan, et al.
Published: (2026)
by: Jeong, Ryan, et al.
Published: (2026)
The replica-symmetric free energy for Ising spin glasses with orthogonally invariant couplings
by: Fan, Zhou, et al.
Published: (2021)
by: Fan, Zhou, et al.
Published: (2021)
Information-theoretic reduction of deep neural networks to linear models in the overparametrized proportional regime
by: Camilli, Francesco, et al.
Published: (2025)
by: Camilli, Francesco, et al.
Published: (2025)
A generalized neural tangent kernel for surrogate gradient learning
by: Eilers, Luke, et al.
Published: (2024)
by: Eilers, Luke, et al.
Published: (2024)
Rigorous Asymptotics for First-Order Algorithms Through the Dynamical Cavity Method
by: Dandi, Yatin, et al.
Published: (2026)
by: Dandi, Yatin, et al.
Published: (2026)
Stochastic Interpolants: A Unifying Framework for Flows and Diffusions
by: Albergo, Michael S., et al.
Published: (2023)
by: Albergo, Michael S., et al.
Published: (2023)
When resampling/reweighting improves feature learning in imbalanced classification?: A toy-model study
by: Obuchi, Tomoyuki, et al.
Published: (2024)
by: Obuchi, Tomoyuki, et al.
Published: (2024)
Similar Items
-
Fixed width treelike neural networks capacity analysis -- generic activations
by: Stojnic, Mihailo
Published: (2024) -
Ultrametric OGP - parametric RDT \emph{symmetric} binary perceptron connection
by: Stojnic, Mihailo
Published: (2026) -
Capacity of the Hebbian-Hopfield network associative memory
by: Stojnic, Mihailo
Published: (2024) -
Binary perceptron computational gap -- a parametric fl RDT view
by: Stojnic, Mihailo
Published: (2025) -
Parametric RDT approach to computational gap of symmetric binary perceptron
by: Stojnic, Mihailo
Published: (2026)