Bayes-optimal learning of an extensive-width neural network from quadratically many samples
Fuente:
arXiv
Saved in:
| Main Authors: | Maillard, Antoine, Troiani, Emanuele, Martin, Simon, Krzakala, Florent, Zdeborová, Lenka |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fundamental Limits of Matrix Sensing: Exact Asymptotics, Universality, and Applications
by: Xu, Yizhou, et al.
Published: (2025)
by: Xu, Yizhou, et al.
Published: (2025)
The Nuclear Route: Sharp Asymptotics of ERM in Overparameterized Quadratic Networks
by: Erba, Vittorio, et al.
Published: (2025)
by: Erba, Vittorio, et al.
Published: (2025)
Bayes optimal learning of attention-indexed models
by: Boncoraglio, Fabrizio, et al.
Published: (2025)
by: Boncoraglio, Fabrizio, et al.
Published: (2025)
Fundamental limits of learning in sequence multi-index models and deep attention networks: High-dimensional asymptotics and sharp thresholds
by: Troiani, Emanuele, et al.
Published: (2025)
by: Troiani, Emanuele, et al.
Published: (2025)
On the Atypical Solutions of the Symmetric Binary Perceptron
by: Barbier, Damien, et al.
Published: (2023)
by: Barbier, Damien, et al.
Published: (2023)
Low-rank Matrix Estimation with Inhomogeneous Noise
by: Guionnet, Alice, et al.
Published: (2022)
by: Guionnet, Alice, et al.
Published: (2022)
Single-Head Attention in High Dimensions: A Theory of Generalization, Weights Spectra, and Scaling Laws
by: Boncoraglio, Fabrizio, et al.
Published: (2025)
by: Boncoraglio, Fabrizio, et al.
Published: (2025)
Statistical mechanics of the maximum-average submatrix problem
by: Erba, Vittorio, et al.
Published: (2023)
by: Erba, Vittorio, et al.
Published: (2023)
The committee machine: Computational to statistical gaps in learning a two-layers neural network
by: Aubin, Benjamin, et al.
Published: (2018)
by: Aubin, Benjamin, et al.
Published: (2018)
Bilinear Sequence Regression: A Model for Learning from Long Sequences of High-dimensional Tokens
by: Erba, Vittorio, et al.
Published: (2024)
by: Erba, Vittorio, et al.
Published: (2024)
The phase diagram of compressed sensing with $\ell_0$-norm regularization
by: Barbier, Damien, et al.
Published: (2024)
by: Barbier, Damien, et al.
Published: (2024)
Gaussian Universality of Perceptrons with Random Labels
by: Gerace, Federica, et al.
Published: (2022)
by: Gerace, Federica, et al.
Published: (2022)
Sampling with flows, diffusion and autoregressive neural networks: A spin-glass perspective
by: Ghio, Davide, et al.
Published: (2023)
by: Ghio, Davide, et al.
Published: (2023)
Learning with Restricted Boltzmann Machines: Asymptotics of AMP and GD in High Dimensions
by: Xu, Yizhou, et al.
Published: (2025)
by: Xu, Yizhou, et al.
Published: (2025)
Quenches in the Sherrington-Kirkpatrick model
by: Erba, Vittorio, et al.
Published: (2024)
by: Erba, Vittorio, et al.
Published: (2024)
Fundamental computational limits of weak learnability in high-dimensional multi-index models
by: Troiani, Emanuele, et al.
Published: (2024)
by: Troiani, Emanuele, et al.
Published: (2024)
The maximum-average subtensor problem: equilibrium and out-of-equilibrium properties
by: Erba, Vittorio, et al.
Published: (2025)
by: Erba, Vittorio, et al.
Published: (2025)
Computational Thresholds in Multi-Modal Learning via the Spiked Matrix-Tensor Model
by: Tabanelli, Hugo, et al.
Published: (2025)
by: Tabanelli, Hugo, et al.
Published: (2025)
The planted XY model: thermodynamics and inference
by: Chen, Siyu, et al.
Published: (2022)
by: Chen, Siyu, et al.
Published: (2022)
Asymptotics of feature learning in two-layer networks after one gradient-step
by: Cui, Hugo, et al.
Published: (2024)
by: Cui, Hugo, et al.
Published: (2024)
Scaling Laws and Spectra of Shallow Neural Networks in the Feature Learning Regime
by: Defilippis, Leonardo, et al.
Published: (2025)
by: Defilippis, Leonardo, et al.
Published: (2025)
Asymptotics of SGD in Sequence-Single Index Models and Single-Layer Attention Networks
by: Arnaboldi, Luca, et al.
Published: (2025)
by: Arnaboldi, Luca, et al.
Published: (2025)
Rigorous Asymptotics for First-Order Algorithms Through the Dynamical Cavity Method
by: Dandi, Yatin, et al.
Published: (2026)
by: Dandi, Yatin, et al.
Published: (2026)
Sequential Dynamics in Ising Spin Glasses
by: Dandi, Yatin, et al.
Published: (2025)
by: Dandi, Yatin, et al.
Published: (2025)
Average-case matrix discrepancy: satisfiability bounds
by: Maillard, Antoine
Published: (2024)
by: Maillard, Antoine
Published: (2024)
Analysis of Bootstrap and Subsampling in High-dimensional Regularized Regression
by: Clarté, Lucas, et al.
Published: (2024)
by: Clarté, Lucas, et al.
Published: (2024)
Injectivity of ReLU networks: perspectives from statistical physics
by: Maillard, Antoine, et al.
Published: (2023)
by: Maillard, Antoine, et al.
Published: (2023)
Spectral Phase Transition and Optimal PCA in Block-Structured Spiked models
by: Mergny, Pierre, et al.
Published: (2024)
by: Mergny, Pierre, et al.
Published: (2024)
Fixed width treelike neural networks capacity analysis -- generic activations
by: Stojnic, Mihailo
Published: (2024)
by: Stojnic, Mihailo
Published: (2024)
Sharp feature-learning transitions and Bayes-optimal neural scaling laws in extensive-width networks
by: Nguyen, Minh-Toan, et al.
Published: (2026)
by: Nguyen, Minh-Toan, et al.
Published: (2026)
Optimal thresholds and algorithms for a model of multi-modal learning in high dimensions
by: Keup, Christian, et al.
Published: (2024)
by: Keup, Christian, et al.
Published: (2024)
A solvable high-dimensional model where nonlinear autoencoders learn structure invisible to PCA while test loss misaligns with generalization
by: Mendes, Vicente Conde, et al.
Published: (2026)
by: Mendes, Vicente Conde, et al.
Published: (2026)
Exact capacity of the \emph{wide} hidden layer treelike neural networks with generic activations
by: Stojnic, Mihailo
Published: (2024)
by: Stojnic, Mihailo
Published: (2024)
Statistical physics analysis of graph neural networks: Approaching optimality in the contextual stochastic block model
by: Duranthon, O., et al.
Published: (2025)
by: Duranthon, O., et al.
Published: (2025)
Compressed sensing with l0-norm: statistical physics analysis and algorithms for signal recovery
by: Barbier, D., et al.
Published: (2023)
by: Barbier, D., et al.
Published: (2023)
Dynamical Cavity Method for Hypergraphs and its Application to Quenches in the k-XOR-SAT Problem
by: Maier, Aude, et al.
Published: (2024)
by: Maier, Aude, et al.
Published: (2024)
Dynamical Phase Transitions in Graph Cellular Automata
by: Behrens, Freya, et al.
Published: (2023)
by: Behrens, Freya, et al.
Published: (2023)
Exact threshold for approximate ellipsoid fitting of random points
by: Bandeira, Afonso S., et al.
Published: (2023)
by: Bandeira, Afonso S., et al.
Published: (2023)
High-dimensional Asymptotics of Denoising Autoencoders
by: Cui, Hugo, et al.
Published: (2023)
by: Cui, Hugo, et al.
Published: (2023)
Building Conformal Prediction Intervals with Approximate Message Passing
by: Clarté, Lucas, et al.
Published: (2024)
by: Clarté, Lucas, et al.
Published: (2024)
Similar Items
-
Fundamental Limits of Matrix Sensing: Exact Asymptotics, Universality, and Applications
by: Xu, Yizhou, et al.
Published: (2025) -
The Nuclear Route: Sharp Asymptotics of ERM in Overparameterized Quadratic Networks
by: Erba, Vittorio, et al.
Published: (2025) -
Bayes optimal learning of attention-indexed models
by: Boncoraglio, Fabrizio, et al.
Published: (2025) -
Fundamental limits of learning in sequence multi-index models and deep attention networks: High-dimensional asymptotics and sharp thresholds
by: Troiani, Emanuele, et al.
Published: (2025) -
On the Atypical Solutions of the Symmetric Binary Perceptron
by: Barbier, Damien, et al.
Published: (2023)