A Theory of Generalization in Deep Learning
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Litman, Elon, Guo, Gabe |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
You Need Better Attention Priors
par: Litman, Elon, et autres
Publié: (2026)
par: Litman, Elon, et autres
Publié: (2026)
The Origin of Edge of Stability
par: Litman, Elon
Publié: (2026)
par: Litman, Elon
Publié: (2026)
Scaled-Dot-Product Attention as One-Sided Entropic Optimal Transport
par: Litman, Elon
Publié: (2025)
par: Litman, Elon
Publié: (2025)
Equilibrium Propagation Without Limits
par: Litman, Elon
Publié: (2025)
par: Litman, Elon
Publié: (2025)
ABC: Any-Subset Autoregression via Non-Markovian Diffusion Bridges in Continuous Time and Space
par: Guo, Gabe, et autres
Publié: (2026)
par: Guo, Gabe, et autres
Publié: (2026)
Reviving Any-Subset Autoregressive Models with Principled Parallel Sampling and Speculative Decoding
par: Guo, Gabe, et autres
Publié: (2025)
par: Guo, Gabe, et autres
Publié: (2025)
Learning When to Stop: Adaptive Latent Reasoning via Reinforcement Learning
par: Ning, Alex, et autres
Publié: (2025)
par: Ning, Alex, et autres
Publié: (2025)
A Generalized Information Bottleneck Theory of Deep Learning
par: Westphal, Charles, et autres
Publié: (2025)
par: Westphal, Charles, et autres
Publié: (2025)
There Will Be a Scientific Theory of Deep Learning
par: Simon, Jamie, et autres
Publié: (2026)
par: Simon, Jamie, et autres
Publié: (2026)
The Propagation Field: A Geometric Substrate Theory of Deep Learning
par: Gu, Xingrui
Publié: (2026)
par: Gu, Xingrui
Publié: (2026)
Generalized Regularized Evidential Deep Learning Models: Theory and Comprehensive Evaluation
par: Pandey, Deep Shankar, et autres
Publié: (2025)
par: Pandey, Deep Shankar, et autres
Publié: (2025)
A Survey on Statistical Theory of Deep Learning: Approximation, Training Dynamics, and Generative Models
par: Suh, Namjoon, et autres
Publié: (2024)
par: Suh, Namjoon, et autres
Publié: (2024)
Artificial Neural Network and Deep Learning: Fundamentals and Theory
par: Hammad, M. M.
Publié: (2024)
par: Hammad, M. M.
Publié: (2024)
Controllable Data Generation by Deep Learning: A Review
par: Wang, Shiyu, et autres
Publié: (2022)
par: Wang, Shiyu, et autres
Publié: (2022)
Conjugate Learning Theory: Uncovering the Mechanisms of Trainability and Generalization in Deep Neural Networks
par: Qi, Binchuan
Publié: (2026)
par: Qi, Binchuan
Publié: (2026)
SETOL: A Semi-Empirical Theory of (Deep) Learning
par: Martin, Charles H, et autres
Publié: (2025)
par: Martin, Charles H, et autres
Publié: (2025)
A Practical Theory of Generalization in Selectivity Learning
par: Wu, Peizhi, et autres
Publié: (2024)
par: Wu, Peizhi, et autres
Publié: (2024)
Deep ARTMAP: Generalized Hierarchical Learning with Adaptive Resonance Theory
par: Melton, Niklas M., et autres
Publié: (2025)
par: Melton, Niklas M., et autres
Publié: (2025)
Randomness Helps Rigor: A Probabilistic Learning Rate Scheduler Bridging Theory and Deep Learning Practice
par: Devapriya, Dahlia, et autres
Publié: (2024)
par: Devapriya, Dahlia, et autres
Publié: (2024)
On the Ability of Deep Networks to Learn Symmetries from Data: A Neural Kernel Theory
par: Perin, Andrea, et autres
Publié: (2024)
par: Perin, Andrea, et autres
Publié: (2024)
When Deep Learning Meets Polyhedral Theory: A Survey
par: Huchette, Joey, et autres
Publié: (2023)
par: Huchette, Joey, et autres
Publié: (2023)
Random Matrix Theory for Deep Learning: Beyond Eigenvalues of Linear Models
par: Liao, Zhenyu, et autres
Publié: (2025)
par: Liao, Zhenyu, et autres
Publié: (2025)
In-Context Learning Is Provably Bayesian Inference: A Generalization Theory for Meta-Learning
par: Wakayama, Tomoya, et autres
Publié: (2025)
par: Wakayama, Tomoya, et autres
Publié: (2025)
A Deep Learning Framework for Multi-Operator Learning: Architectures and Approximation Theory
par: Weihs, Adrien, et autres
Publié: (2025)
par: Weihs, Adrien, et autres
Publié: (2025)
A Generative Deep Learning Approach for Crash Severity Modeling with Imbalanced Data
par: Chen, Junlan, et autres
Publié: (2024)
par: Chen, Junlan, et autres
Publié: (2024)
Gradient-Based Multi-Objective Deep Learning: Algorithms, Theories, Applications, and Beyond
par: Chen, Weiyu, et autres
Publié: (2025)
par: Chen, Weiyu, et autres
Publié: (2025)
A General Theory for Compositional Generalization
par: Fu, Jingwen, et autres
Publié: (2024)
par: Fu, Jingwen, et autres
Publié: (2024)
Pragmatic Policy Development via Interpretable Behavior Cloning
par: Matsson, Anton, et autres
Publié: (2025)
par: Matsson, Anton, et autres
Publié: (2025)
Position: A Theory of Deep Learning Must Include Compositional Sparsity
par: Danhofer, David A., et autres
Publié: (2025)
par: Danhofer, David A., et autres
Publié: (2025)
Advancing Molecular Machine Learning Representations with Stereoelectronics-Infused Molecular Graphs
par: Boiko, Daniil A., et autres
Publié: (2024)
par: Boiko, Daniil A., et autres
Publié: (2024)
Generative Flow Networks: Theory and Applications to Structure Learning
par: Deleu, Tristan
Publié: (2025)
par: Deleu, Tristan
Publié: (2025)
Probabilistic Learning and Generation in Deep Sequence Models
par: Chen, Wenlong
Publié: (2026)
par: Chen, Wenlong
Publié: (2026)
Feature Selection as Deep Sequential Generative Learning
par: Ying, Wangyang, et autres
Publié: (2024)
par: Ying, Wangyang, et autres
Publié: (2024)
Generalization Analysis for Deep Contrastive Representation Learning
par: Hieu, Nong Minh, et autres
Publié: (2024)
par: Hieu, Nong Minh, et autres
Publié: (2024)
Cross-Dataset Generalization in Deep Learning
par: Zhang, Xuyu, et autres
Publié: (2024)
par: Zhang, Xuyu, et autres
Publié: (2024)
Enhancing Accuracy in Deep Learning Using Random Matrix Theory
par: Berlyand, Leonid, et autres
Publié: (2023)
par: Berlyand, Leonid, et autres
Publié: (2023)
Copresheaf Topological Neural Networks: A Generalized Deep Learning Framework
par: Hajij, Mustafa, et autres
Publié: (2025)
par: Hajij, Mustafa, et autres
Publié: (2025)
Recurrent Expansion: A Pathway Toward the Next Generation of Deep Learning
par: Berghout, Tarek
Publié: (2025)
par: Berghout, Tarek
Publié: (2025)
Deep Learning on Graphs for Mobile Network Topology Generation
par: Meli, Felix Nannesson, et autres
Publié: (2025)
par: Meli, Felix Nannesson, et autres
Publié: (2025)
On Rademacher Complexity-based Generalization Bounds for Deep Learning
par: Truong, Lan V.
Publié: (2022)
par: Truong, Lan V.
Publié: (2022)
Documents similaires
-
You Need Better Attention Priors
par: Litman, Elon, et autres
Publié: (2026) -
The Origin of Edge of Stability
par: Litman, Elon
Publié: (2026) -
Scaled-Dot-Product Attention as One-Sided Entropic Optimal Transport
par: Litman, Elon
Publié: (2025) -
Equilibrium Propagation Without Limits
par: Litman, Elon
Publié: (2025) -
ABC: Any-Subset Autoregression via Non-Markovian Diffusion Bridges in Continuous Time and Space
par: Guo, Gabe, et autres
Publié: (2026)