Applying statistical learning theory to deep learning
Fuente:
arXiv
Saved in:
| Main Authors: | Gerbelot, Cédric, Karagulyan, Avetik, Karp, Stefani, Ravichandran, Kavya, Stern, Menachem, Srebro, Nathan |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Asymptotic theory of in-context learning by linear attention
by: Lu, Yue M., et al.
Published: (2024)
by: Lu, Yue M., et al.
Published: (2024)
Spring-block theory of feature learning in deep neural networks
by: Shi, Cheng, et al.
Published: (2024)
by: Shi, Cheng, et al.
Published: (2024)
A unified theory of feature learning in RNNs and DNNs
by: Bauer, Jan P., et al.
Published: (2026)
by: Bauer, Jan P., et al.
Published: (2026)
A solvable model of learning generative diffusion: theory and insights
by: Cui, Hugo, et al.
Published: (2025)
by: Cui, Hugo, et al.
Published: (2025)
Implicit bias produces neural scaling laws in learning curves, from perceptrons to deep networks
by: D'Amico, Francesco, et al.
Published: (2025)
by: D'Amico, Francesco, et al.
Published: (2025)
Fundamental limits of learning in sequence multi-index models and deep attention networks: High-dimensional asymptotics and sharp thresholds
by: Troiani, Emanuele, et al.
Published: (2025)
by: Troiani, Emanuele, et al.
Published: (2025)
A solvable high-dimensional model where nonlinear autoencoders learn structure invisible to PCA while test loss misaligns with generalization
by: Mendes, Vicente Conde, et al.
Published: (2026)
by: Mendes, Vicente Conde, et al.
Published: (2026)
High-dimensional learning of narrow neural networks
by: Cui, Hugo
Published: (2024)
by: Cui, Hugo
Published: (2024)
Dynamics of Meta-learning Representation in the Teacher-student Scenario
by: Wang, Hui, et al.
Published: (2024)
by: Wang, Hui, et al.
Published: (2024)
Machine learning for sustainable geoenergy: uncertainty, physics and decision-ready inference
by: Menke, Hannah P., et al.
Published: (2026)
by: Menke, Hannah P., et al.
Published: (2026)
Deep networks learn to parse uniform-depth context-free languages from local statistics
by: Parley, Jack T., et al.
Published: (2026)
by: Parley, Jack T., et al.
Published: (2026)
Asymptotics of feature learning in two-layer networks after one gradient-step
by: Cui, Hugo, et al.
Published: (2024)
by: Cui, Hugo, et al.
Published: (2024)
Adaptive kernel predictors from feature-learning infinite limits of neural networks
by: Lauditi, Clarissa, et al.
Published: (2025)
by: Lauditi, Clarissa, et al.
Published: (2025)
Optimal thresholds and algorithms for a model of multi-modal learning in high dimensions
by: Keup, Christian, et al.
Published: (2024)
by: Keup, Christian, et al.
Published: (2024)
Transient learning dynamics drive escape from sharp valleys in Stochastic Gradient Descent
by: Yang, Ning, et al.
Published: (2026)
by: Yang, Ning, et al.
Published: (2026)
Microscopic imprints of learned solutions in adaptive resistor networks
by: Guzman, Marcel, et al.
Published: (2024)
by: Guzman, Marcel, et al.
Published: (2024)
Online unsupervised Hebbian learning in deep photonic neuromorphic networks
by: Li, Xi, et al.
Published: (2026)
by: Li, Xi, et al.
Published: (2026)
A statistical physics framework for optimal learning
by: Mignacco, Francesca, et al.
Published: (2025)
by: Mignacco, Francesca, et al.
Published: (2025)
Emergent weight morphologies in deep neural networks
by: de Jong, Pascal, et al.
Published: (2025)
by: de Jong, Pascal, et al.
Published: (2025)
Generative diffusion for perceptron problems: statistical physics analysis and efficient algorithms
by: Demyanenko, Elizaveta, et al.
Published: (2025)
by: Demyanenko, Elizaveta, et al.
Published: (2025)
Physical networks become what they learn
by: Stern, Menachem, et al.
Published: (2024)
by: Stern, Menachem, et al.
Published: (2024)
Finite-time Lyapunov exponents of deep neural networks
by: Storm, L., et al.
Published: (2023)
by: Storm, L., et al.
Published: (2023)
Computing frustration and near-monotonicity in deep neural networks
by: Wendin, Joel, et al.
Published: (2025)
by: Wendin, Joel, et al.
Published: (2025)
DCEM: A deep complementary energy method for solid mechanics
by: Wang, Yizheng, et al.
Published: (2023)
by: Wang, Yizheng, et al.
Published: (2023)
Statistical physics of deep learning: Optimal learning of a multi-layer perceptron near interpolation
by: Barbier, Jean, et al.
Published: (2025)
by: Barbier, Jean, et al.
Published: (2025)
Contrastive learning in tunable dynamical systems
by: Stern, Menachem, et al.
Published: (2026)
by: Stern, Menachem, et al.
Published: (2026)
Critical feature learning in deep neural networks
by: Fischer, Kirsten, et al.
Published: (2024)
by: Fischer, Kirsten, et al.
Published: (2024)
Deep neural networks from the perspective of ergodic theory
by: Zhang, Fan
Published: (2023)
by: Zhang, Fan
Published: (2023)
Field theory for optimal signal propagation in ResNets
by: Fischer, Kirsten, et al.
Published: (2023)
by: Fischer, Kirsten, et al.
Published: (2023)
Towards a theory of how the structure of language is acquired by deep neural networks
by: Cagnetta, Francesco, et al.
Published: (2024)
by: Cagnetta, Francesco, et al.
Published: (2024)
Learning curves theory for hierarchically compositional data with power-law distributed features
by: Cagnetta, Francesco, et al.
Published: (2025)
by: Cagnetta, Francesco, et al.
Published: (2025)
Universal characteristics of deep neural network loss surfaces from random matrix theory
by: Baskerville, Nicholas P, et al.
Published: (2022)
by: Baskerville, Nicholas P, et al.
Published: (2022)
Distinct mechanisms underlying in-context learning in transformers
by: Gibson, Cole, et al.
Published: (2026)
by: Gibson, Cole, et al.
Published: (2026)
Bayes optimal learning of attention-indexed models
by: Boncoraglio, Fabrizio, et al.
Published: (2025)
by: Boncoraglio, Fabrizio, et al.
Published: (2025)
Liquid and solid layers in a thermal deep learning machine
by: Huang, Gang, et al.
Published: (2025)
by: Huang, Gang, et al.
Published: (2025)
Injectivity of ReLU networks: perspectives from statistical physics
by: Maillard, Antoine, et al.
Published: (2023)
by: Maillard, Antoine, et al.
Published: (2023)
The committee machine: Computational to statistical gaps in learning a two-layers neural network
by: Aubin, Benjamin, et al.
Published: (2018)
by: Aubin, Benjamin, et al.
Published: (2018)
Dataset-learning duality and emergent criticality
by: Kukleva, Ekaterina, et al.
Published: (2024)
by: Kukleva, Ekaterina, et al.
Published: (2024)
Coding schemes in neural networks learning classification tasks
by: van Meegen, Alexander, et al.
Published: (2024)
by: van Meegen, Alexander, et al.
Published: (2024)
Cell reprogramming design by transfer learning of functional transcriptional networks
by: Wytock, Thomas P., et al.
Published: (2024)
by: Wytock, Thomas P., et al.
Published: (2024)
Similar Items
-
Asymptotic theory of in-context learning by linear attention
by: Lu, Yue M., et al.
Published: (2024) -
Spring-block theory of feature learning in deep neural networks
by: Shi, Cheng, et al.
Published: (2024) -
A unified theory of feature learning in RNNs and DNNs
by: Bauer, Jan P., et al.
Published: (2026) -
A solvable model of learning generative diffusion: theory and insights
by: Cui, Hugo, et al.
Published: (2025) -
Implicit bias produces neural scaling laws in learning curves, from perceptrons to deep networks
by: D'Amico, Francesco, et al.
Published: (2025)