Salvato in:
| Autori principali: | Rawal, Divit, DeWeese, Michael R. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2605.01288 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Saddle Hierarchy in Dense Associative Memory
di: Thériault, Robin, et al.
Pubblicazione: (2025)
di: Thériault, Robin, et al.
Pubblicazione: (2025)
Spectral Dynamics in Deep Networks: Feature Learning, Outlier Escape, and Learning Rate Transfer
di: Lauditi, Clarissa, et al.
Pubblicazione: (2026)
di: Lauditi, Clarissa, et al.
Pubblicazione: (2026)
Deep Learning as Neural Low-Degree Filtering: A Spectral Theory of Hierarchical Feature Learning
di: Dandi, Yatin, et al.
Pubblicazione: (2026)
di: Dandi, Yatin, et al.
Pubblicazione: (2026)
Dynamical Mean-Field Theory of Self-Attention Neural Networks
di: Poc-López, Ángel, et al.
Pubblicazione: (2024)
di: Poc-López, Ángel, et al.
Pubblicazione: (2024)
Precise Dynamics of Diagonal Linear Networks: A Unifying Analysis by Dynamical Mean-Field Theory
di: Nishiyama, Sota, et al.
Pubblicazione: (2025)
di: Nishiyama, Sota, et al.
Pubblicazione: (2025)
From Kernels to Features: A Multi-Scale Adaptive Theory of Feature Learning
di: Rubin, Noa, et al.
Pubblicazione: (2025)
di: Rubin, Noa, et al.
Pubblicazione: (2025)
The Training Process of Many Deep Networks Explores the Same Low-Dimensional Manifold
di: Mao, Jialin, et al.
Pubblicazione: (2023)
di: Mao, Jialin, et al.
Pubblicazione: (2023)
Kernel Renormalization in Bayesian Deep Neural Networks: the Equivalent Wishart Ansatz in the Proportional Regime
di: Baglioni, Paolo, et al.
Pubblicazione: (2026)
di: Baglioni, Paolo, et al.
Pubblicazione: (2026)
How Deep Networks Learn Sparse and Hierarchical Data: the Sparse Random Hierarchy Model
di: Tomasini, Umberto, et al.
Pubblicazione: (2024)
di: Tomasini, Umberto, et al.
Pubblicazione: (2024)
Statistical Physics of Deep Neural Networks: Generalization Capability, Beyond the Infinite Width, and Feature Learning
di: Ariosto, Sebastiano
Pubblicazione: (2025)
di: Ariosto, Sebastiano
Pubblicazione: (2025)
Deep Linear Network Training Dynamics from Random Initialization: Data, Width, Depth, and Hyperparameter Transfer
di: Bordelon, Blake, et al.
Pubblicazione: (2025)
di: Bordelon, Blake, et al.
Pubblicazione: (2025)
A Fast Algorithm to Simulate Nonlinear Resistive Networks
di: Scellier, Benjamin
Pubblicazione: (2024)
di: Scellier, Benjamin
Pubblicazione: (2024)
Training Dynamics of Nonlinear Contrastive Learning Model in the High Dimensional Limit
di: Meng, Lineghuan, et al.
Pubblicazione: (2024)
di: Meng, Lineghuan, et al.
Pubblicazione: (2024)
Escape dynamics and implicit bias of one-pass SGD in overparameterized quadratic networks
di: Bocchi, Dario, et al.
Pubblicazione: (2026)
di: Bocchi, Dario, et al.
Pubblicazione: (2026)
Graph Neural Networks Do Not Always Oversmooth
di: Epping, Bastian, et al.
Pubblicazione: (2024)
di: Epping, Bastian, et al.
Pubblicazione: (2024)
Applications of Statistical Field Theory in Deep Learning
di: Ringel, Zohar, et al.
Pubblicazione: (2025)
di: Ringel, Zohar, et al.
Pubblicazione: (2025)
Theory of Speciation Transitions in Diffusion Models with General Class Structure
di: Achilli, Beatrice, et al.
Pubblicazione: (2026)
di: Achilli, Beatrice, et al.
Pubblicazione: (2026)
Parameter Symmetry Potentially Unifies Deep Learning Theory
di: Ziyin, Liu, et al.
Pubblicazione: (2025)
di: Ziyin, Liu, et al.
Pubblicazione: (2025)
Theory of Scaling Laws for In-Context Regression: Depth, Width, Context and Time
di: Bordelon, Blake, et al.
Pubblicazione: (2025)
di: Bordelon, Blake, et al.
Pubblicazione: (2025)
Deep neural networks from the perspective of ergodic theory
di: Zhang, Fan
Pubblicazione: (2023)
di: Zhang, Fan
Pubblicazione: (2023)
The Physics of Data and Tasks: Theories of Locality and Compositionality in Deep Learning
di: Favero, Alessandro
Pubblicazione: (2025)
di: Favero, Alessandro
Pubblicazione: (2025)
Theory of Optimal Learning Rate Schedules and Scaling Laws for a Random Feature Model
di: Bordelon, Blake, et al.
Pubblicazione: (2026)
di: Bordelon, Blake, et al.
Pubblicazione: (2026)
High-Dimensional Limit of Stochastic Gradient Flow via Dynamical Mean-Field Theory
di: Nishiyama, Sota, et al.
Pubblicazione: (2026)
di: Nishiyama, Sota, et al.
Pubblicazione: (2026)
Approximation Theory for Neural Networks: Old and New
di: Mukherjee, Soumendu Sundar, et al.
Pubblicazione: (2026)
di: Mukherjee, Soumendu Sundar, et al.
Pubblicazione: (2026)
A Random-Matrix Criterion for Initializing Gated Recurrent Neural Networks
di: Fioratti, Tommaso, et al.
Pubblicazione: (2026)
di: Fioratti, Tommaso, et al.
Pubblicazione: (2026)
A Federated Many-to-One Hopfield model for associative Neural Networks
di: Alessandrelli, Andrea, et al.
Pubblicazione: (2026)
di: Alessandrelli, Andrea, et al.
Pubblicazione: (2026)
Formation of Representations in Neural Networks
di: Ziyin, Liu, et al.
Pubblicazione: (2024)
di: Ziyin, Liu, et al.
Pubblicazione: (2024)
Convergence Acceleration of Markov Chain Monte Carlo-based Gradient Descent by Deep Unfolding
di: Hagiwara, Ryo, et al.
Pubblicazione: (2024)
di: Hagiwara, Ryo, et al.
Pubblicazione: (2024)
Transfer Learning in Infinite Width Feature Learning Networks
di: Lauditi, Clarissa, et al.
Pubblicazione: (2025)
di: Lauditi, Clarissa, et al.
Pubblicazione: (2025)
The Rules-and-Facts Model for Simultaneous Generalization and Memorization in Neural Networks
di: Farné, Gabriele, et al.
Pubblicazione: (2026)
di: Farné, Gabriele, et al.
Pubblicazione: (2026)
Demolition and Reinforcement of Memories in Spin-Glass-like Neural Networks
di: Ventura, Enrico
Pubblicazione: (2024)
di: Ventura, Enrico
Pubblicazione: (2024)
Growing Neural Networks: Dynamic Evolution through Gradient Descent
di: Radhakrishnan, Anil, et al.
Pubblicazione: (2025)
di: Radhakrishnan, Anil, et al.
Pubblicazione: (2025)
Dynamical Decoupling of Generalization and Overfitting in Large Two-Layer Networks
di: Montanari, Andrea, et al.
Pubblicazione: (2025)
di: Montanari, Andrea, et al.
Pubblicazione: (2025)
Benchmarking Graph Neural Networks in Solving Hard Constraint Satisfaction Problems
di: Skenderi, Geri, et al.
Pubblicazione: (2026)
di: Skenderi, Geri, et al.
Pubblicazione: (2026)
Controlled Langevin Dynamics for Sampling of Feedforward Neural Networks Trained with Minibatches
di: Zambon, Alessandro, et al.
Pubblicazione: (2026)
di: Zambon, Alessandro, et al.
Pubblicazione: (2026)
Statistical Mechanics Calculations Using Variational Autoregressive Networks and Quantum Annealing
di: Tamura, Yuta, et al.
Pubblicazione: (2024)
di: Tamura, Yuta, et al.
Pubblicazione: (2024)
Grokking as a First Order Phase Transition in Two Layer Networks
di: Rubin, Noa, et al.
Pubblicazione: (2023)
di: Rubin, Noa, et al.
Pubblicazione: (2023)
Initial Guessing Bias: How Untrained Networks Favor Some Classes
di: Francazi, Emanuele, et al.
Pubblicazione: (2023)
di: Francazi, Emanuele, et al.
Pubblicazione: (2023)
STEM Diffraction Pattern Analysis with Deep Learning Networks
di: Wissel, Sebastian, et al.
Pubblicazione: (2025)
di: Wissel, Sebastian, et al.
Pubblicazione: (2025)
Dynamical Learning in Deep Asymmetric Recurrent Neural Networks
di: Badalotti, Davide, et al.
Pubblicazione: (2025)
di: Badalotti, Davide, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Saddle Hierarchy in Dense Associative Memory
di: Thériault, Robin, et al.
Pubblicazione: (2025) -
Spectral Dynamics in Deep Networks: Feature Learning, Outlier Escape, and Learning Rate Transfer
di: Lauditi, Clarissa, et al.
Pubblicazione: (2026) -
Deep Learning as Neural Low-Degree Filtering: A Spectral Theory of Hierarchical Feature Learning
di: Dandi, Yatin, et al.
Pubblicazione: (2026) -
Dynamical Mean-Field Theory of Self-Attention Neural Networks
di: Poc-López, Ángel, et al.
Pubblicazione: (2024) -
Precise Dynamics of Diagonal Linear Networks: A Unifying Analysis by Dynamical Mean-Field Theory
di: Nishiyama, Sota, et al.
Pubblicazione: (2025)