Field theory for optimal signal propagation in ResNets
Fuente:
arXiv
Saved in:
| Main Authors: | Fischer, Kirsten, Dahmen, David, Helias, Moritz |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A unified theory of feature learning in RNNs and DNNs
by: Bauer, Jan P., et al.
Published: (2026)
by: Bauer, Jan P., et al.
Published: (2026)
From Kernels to Features: A Multi-Scale Adaptive Theory of Feature Learning
by: Rubin, Noa, et al.
Published: (2025)
by: Rubin, Noa, et al.
Published: (2025)
Critical feature learning in deep neural networks
by: Fischer, Kirsten, et al.
Published: (2024)
by: Fischer, Kirsten, et al.
Published: (2024)
Reduction of interaction order in hard combinatorial optimization via conditionally independent degrees of freedom
by: Ciobanu, Alexandru, et al.
Published: (2025)
by: Ciobanu, Alexandru, et al.
Published: (2025)
Effect of Synaptic Heterogeneity on Neuronal Coordination
by: Layer, Moritz, et al.
Published: (2023)
by: Layer, Moritz, et al.
Published: (2023)
Graph Neural Networks Do Not Always Oversmooth
by: Epping, Bastian, et al.
Published: (2024)
by: Epping, Bastian, et al.
Published: (2024)
Applications of Statistical Field Theory in Deep Learning
by: Ringel, Zohar, et al.
Published: (2025)
by: Ringel, Zohar, et al.
Published: (2025)
Linking Network and Neuron-level Correlations by Renormalized Field Theory
by: Dick, Michael, et al.
Published: (2023)
by: Dick, Michael, et al.
Published: (2023)
Beyond-mean-field fluctuations for the solution of constraint satisfaction problems
by: Foos, Niklas, et al.
Published: (2025)
by: Foos, Niklas, et al.
Published: (2025)
Dynamics of neural scaling laws in random feature regression with powerlaw-distributed kernel eigenvalues
by: Kramp, Jakob, et al.
Published: (2026)
by: Kramp, Jakob, et al.
Published: (2026)
Renormalization group for deep neural networks: Universality of learning and scaling laws
by: Coppola, Gorka Peraza, et al.
Published: (2025)
by: Coppola, Gorka Peraza, et al.
Published: (2025)
Lecture notes: From Gaussian processes to feature learning
by: Helias, Moritz, et al.
Published: (2026)
by: Helias, Moritz, et al.
Published: (2026)
Two failure modes of deep transformers and how to avoid them: a unified theory of signal propagation at initialisation
by: Giorlandino, Alessio, et al.
Published: (2025)
by: Giorlandino, Alessio, et al.
Published: (2025)
Perfect reconstruction of sparse signals using nonconvexity control and one-step RSB message passing
by: Gu, Xiaosi, et al.
Published: (2025)
by: Gu, Xiaosi, et al.
Published: (2025)
Finite-size scaling of hetero-associative retrieval in continuous-signal-driven Ising spin systems
by: Ladiana, Andrea
Published: (2026)
by: Ladiana, Andrea
Published: (2026)
Restoring balance: principled under/oversampling of data for optimal classification
by: Loffredo, Emanuele, et al.
Published: (2024)
by: Loffredo, Emanuele, et al.
Published: (2024)
Applying statistical learning theory to deep learning
by: Gerbelot, Cédric, et al.
Published: (2023)
by: Gerbelot, Cédric, et al.
Published: (2023)
Asymptotic theory of in-context learning by linear attention
by: Lu, Yue M., et al.
Published: (2024)
by: Lu, Yue M., et al.
Published: (2024)
Deep neural networks from the perspective of ergodic theory
by: Zhang, Fan
Published: (2023)
by: Zhang, Fan
Published: (2023)
A solvable model of learning generative diffusion: theory and insights
by: Cui, Hugo, et al.
Published: (2025)
by: Cui, Hugo, et al.
Published: (2025)
Statistical physics analysis of graph neural networks: Approaching optimality in the contextual stochastic block model
by: Duranthon, O., et al.
Published: (2025)
by: Duranthon, O., et al.
Published: (2025)
Dynamical Mean-Field Theory of Self-Attention Neural Networks
by: Poc-López, Ángel, et al.
Published: (2024)
by: Poc-López, Ángel, et al.
Published: (2024)
Learning curves theory for hierarchically compositional data with power-law distributed features
by: Cagnetta, Francesco, et al.
Published: (2025)
by: Cagnetta, Francesco, et al.
Published: (2025)
$L_0$ Regularization of Field-Aware Factorization Machine through Ising Model
by: Okamoto, Yasuharu
Published: (2024)
by: Okamoto, Yasuharu
Published: (2024)
High-Dimensional Limit of Stochastic Gradient Flow via Dynamical Mean-Field Theory
by: Nishiyama, Sota, et al.
Published: (2026)
by: Nishiyama, Sota, et al.
Published: (2026)
Precise Dynamics of Diagonal Linear Networks: A Unifying Analysis by Dynamical Mean-Field Theory
by: Nishiyama, Sota, et al.
Published: (2025)
by: Nishiyama, Sota, et al.
Published: (2025)
Soft Quantization: Model Compression Via Weight Coupling
by: Bernstein, Daniel T., et al.
Published: (2026)
by: Bernstein, Daniel T., et al.
Published: (2026)
Random walk models for the propagation of signalling molecules in one-dimensional spatial networks and their continuum limit
by: Mehrpooya, Adel, et al.
Published: (2023)
by: Mehrpooya, Adel, et al.
Published: (2023)
Sampling at intermediate temperatures is optimal for training large language models in protein structure prediction
by: Ghiringhelli, L., et al.
Published: (2026)
by: Ghiringhelli, L., et al.
Published: (2026)
Benchmarking Graph Neural Networks in Solving Hard Constraint Satisfaction Problems
by: Skenderi, Geri, et al.
Published: (2026)
by: Skenderi, Geri, et al.
Published: (2026)
Analytic theory of dropout regularization
by: Mori, Francesco, et al.
Published: (2025)
by: Mori, Francesco, et al.
Published: (2025)
Bayes optimal learning of attention-indexed models
by: Boncoraglio, Fabrizio, et al.
Published: (2025)
by: Boncoraglio, Fabrizio, et al.
Published: (2025)
The RL Perceptron: Generalisation Dynamics of Policy Learning in High Dimensions
by: Patel, Nishil, et al.
Published: (2023)
by: Patel, Nishil, et al.
Published: (2023)
The Quantization Model of Neural Scaling
by: Michaud, Eric J., et al.
Published: (2023)
by: Michaud, Eric J., et al.
Published: (2023)
Introduction to Latent Variable Energy-Based Models: A Path Towards Autonomous Machine Intelligence
by: Dawid, Anna, et al.
Published: (2023)
by: Dawid, Anna, et al.
Published: (2023)
How does training shape the Riemannian geometry of neural network representations?
by: Zavatone-Veth, Jacob A., et al.
Published: (2023)
by: Zavatone-Veth, Jacob A., et al.
Published: (2023)
Grokking as a First Order Phase Transition in Two Layer Networks
by: Rubin, Noa, et al.
Published: (2023)
by: Rubin, Noa, et al.
Published: (2023)
A universal approximation theorem for nonlinear resistive networks
by: Scellier, Benjamin, et al.
Published: (2023)
by: Scellier, Benjamin, et al.
Published: (2023)
High-dimensional Asymptotics of Denoising Autoencoders
by: Cui, Hugo, et al.
Published: (2023)
by: Cui, Hugo, et al.
Published: (2023)
On the different regimes of Stochastic Gradient Descent
by: Sclocchi, Antonio, et al.
Published: (2023)
by: Sclocchi, Antonio, et al.
Published: (2023)
Similar Items
-
A unified theory of feature learning in RNNs and DNNs
by: Bauer, Jan P., et al.
Published: (2026) -
From Kernels to Features: A Multi-Scale Adaptive Theory of Feature Learning
by: Rubin, Noa, et al.
Published: (2025) -
Critical feature learning in deep neural networks
by: Fischer, Kirsten, et al.
Published: (2024) -
Reduction of interaction order in hard combinatorial optimization via conditionally independent degrees of freedom
by: Ciobanu, Alexandru, et al.
Published: (2025) -
Effect of Synaptic Heterogeneity on Neuronal Coordination
by: Layer, Moritz, et al.
Published: (2023)