Saved in:
| Main Authors: | Ruben, Benjamin S., Tong, William L., Chaudhry, Hamza Tahir, Pehlevan, Cengiz |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2412.05418 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Infinite Limits of Multi-head Transformer Dynamics
by: Bordelon, Blake, et al.
Published: (2024)
by: Bordelon, Blake, et al.
Published: (2024)
Learning Curves for Noisy Heterogeneous Feature-Subsampled Ridge Ensembles
by: Ruben, Benjamin S., et al.
Published: (2023)
by: Ruben, Benjamin S., et al.
Published: (2023)
How Feature Learning Can Improve Neural Scaling Laws
by: Bordelon, Blake, et al.
Published: (2024)
by: Bordelon, Blake, et al.
Published: (2024)
A Dynamical Model of Neural Scaling Laws
by: Bordelon, Blake, et al.
Published: (2024)
by: Bordelon, Blake, et al.
Published: (2024)
Theory of Scaling Laws for In-Context Regression: Depth, Width, Context and Time
by: Bordelon, Blake, et al.
Published: (2025)
by: Bordelon, Blake, et al.
Published: (2025)
Disordered Dynamics in High Dimensions: Connections to Random Matrices and Machine Learning
by: Bordelon, Blake, et al.
Published: (2026)
by: Bordelon, Blake, et al.
Published: (2026)
Deep Linear Network Training Dynamics from Random Initialization: Data, Width, Depth, and Hyperparameter Transfer
by: Bordelon, Blake, et al.
Published: (2025)
by: Bordelon, Blake, et al.
Published: (2025)
Transfer Learning in Infinite Width Feature Learning Networks
by: Lauditi, Clarissa, et al.
Published: (2025)
by: Lauditi, Clarissa, et al.
Published: (2025)
A note on the dynamics of extended-context disordered kinetic spin models
by: Zavatone-Veth, Jacob A., et al.
Published: (2025)
by: Zavatone-Veth, Jacob A., et al.
Published: (2025)
Scaling and renormalization in high-dimensional regression
by: Atanasov, Alexander, et al.
Published: (2024)
by: Atanasov, Alexander, et al.
Published: (2024)
Spectral Dynamics in Deep Networks: Feature Learning, Outlier Escape, and Learning Rate Transfer
by: Lauditi, Clarissa, et al.
Published: (2026)
by: Lauditi, Clarissa, et al.
Published: (2026)
Nadaraya-Watson kernel smoothing as a random energy model
by: Zavatone-Veth, Jacob A., et al.
Published: (2024)
by: Zavatone-Veth, Jacob A., et al.
Published: (2024)
Adaptive kernel predictors from feature-learning infinite limits of neural networks
by: Lauditi, Clarissa, et al.
Published: (2025)
by: Lauditi, Clarissa, et al.
Published: (2025)
A solvable model of learning generative diffusion: theory and insights
by: Cui, Hugo, et al.
Published: (2025)
by: Cui, Hugo, et al.
Published: (2025)
Risk and cross validation in ridge regression with correlated samples
by: Atanasov, Alexander, et al.
Published: (2024)
by: Atanasov, Alexander, et al.
Published: (2024)
Grokking as the Transition from Lazy to Rich Training Dynamics
by: Kumar, Tanishq, et al.
Published: (2023)
by: Kumar, Tanishq, et al.
Published: (2023)
Theory of Optimal Learning Rate Schedules and Scaling Laws for a Random Feature Model
by: Bordelon, Blake, et al.
Published: (2026)
by: Bordelon, Blake, et al.
Published: (2026)
How does training shape the Riemannian geometry of neural network representations?
by: Zavatone-Veth, Jacob A., et al.
Published: (2023)
by: Zavatone-Veth, Jacob A., et al.
Published: (2023)
Two-Point Deterministic Equivalence for Stochastic Gradient Dynamics in Linear Models
by: Atanasov, Alexander, et al.
Published: (2025)
by: Atanasov, Alexander, et al.
Published: (2025)
Asymptotic theory of in-context learning by linear attention
by: Lu, Yue M., et al.
Published: (2024)
by: Lu, Yue M., et al.
Published: (2024)
Dynamically Learning to Integrate in Recurrent Neural Networks
by: Bordelon, Blake, et al.
Published: (2025)
by: Bordelon, Blake, et al.
Published: (2025)
Dropout Universality: Scaling Laws and Optimal Scheduling at the Edge-of-Chaos
by: Sarmiento, Lucas Fernandez
Published: (2026)
by: Sarmiento, Lucas Fernandez
Published: (2026)
Unbound States and Mixed Bound--Unbound Phases in Near-Infinitely Deep Potentials
by: Cheng, Shujie, et al.
Published: (2026)
by: Cheng, Shujie, et al.
Published: (2026)
Ergodicity Breaking and High-Dimensional Chaos in Random Recurrent Networks
by: Martorell, Carles, et al.
Published: (2025)
by: Martorell, Carles, et al.
Published: (2025)
Probing Band-center Anomaly with the Kernel Polynomial Method
by: Khan, N. A., et al.
Published: (2019)
by: Khan, N. A., et al.
Published: (2019)
Scaling of Static and Dynamical Properties of Random Anisotropy Magnets
by: Garanin, Dmitry A., et al.
Published: (2024)
by: Garanin, Dmitry A., et al.
Published: (2024)
Asymmetric Scaling Laws from Sparse Features
by: Sous, John, et al.
Published: (2026)
by: Sous, John, et al.
Published: (2026)
From Kernels to Features: A Multi-Scale Adaptive Theory of Feature Learning
by: Rubin, Noa, et al.
Published: (2025)
by: Rubin, Noa, et al.
Published: (2025)
Explaining Neural Scaling Laws
by: Bahri, Yasaman, et al.
Published: (2021)
by: Bahri, Yasaman, et al.
Published: (2021)
From Chaos to Synchrony in Recurrent Excitatory-Inhibitory Networks with Target-Specific Inhibition
by: Martorell, Carles, et al.
Published: (2026)
by: Martorell, Carles, et al.
Published: (2026)
Neural Scaling Laws Rooted in the Data Distribution
by: Brill, Ari
Published: (2024)
by: Brill, Ari
Published: (2024)
Scaling Laws and Spectra of Shallow Neural Networks in the Feature Learning Regime
by: Defilippis, Leonardo, et al.
Published: (2025)
by: Defilippis, Leonardo, et al.
Published: (2025)
Compensation between the parameters of the Jonschers's Universal Relaxation Law in disordered materials
by: Papathanassiou, Anthony N., et al.
Published: (2025)
by: Papathanassiou, Anthony N., et al.
Published: (2025)
Random Features Hopfield Networks generalize retrieval to previously unseen examples
by: Kalaj, Silvio, et al.
Published: (2024)
by: Kalaj, Silvio, et al.
Published: (2024)
Glass-like Caging with Random Planes
by: Bonnet, Gilles, et al.
Published: (2023)
by: Bonnet, Gilles, et al.
Published: (2023)
Random transverse and longitudinal field Ising chains
by: Petö, Tamás, et al.
Published: (2025)
by: Petö, Tamás, et al.
Published: (2025)
The Spectral Boundary of Block Structured Random Matrices
by: Patil, Nirbhay, et al.
Published: (2023)
by: Patil, Nirbhay, et al.
Published: (2023)
Random Apollonian networks with tailored clustering coefficient
by: Souza, Eduardo M. K., et al.
Published: (2024)
by: Souza, Eduardo M. K., et al.
Published: (2024)
Scaling Laws and Representation Learning in Simple Hierarchical Languages: Transformers vs. Convolutional Architectures
by: Cagnetta, Francesco, et al.
Published: (2025)
by: Cagnetta, Francesco, et al.
Published: (2025)
Spectral and Entanglement Properties of the Random Exchange Heisenberg Chain
by: Gao, Yilun, et al.
Published: (2024)
by: Gao, Yilun, et al.
Published: (2024)
Similar Items
-
Infinite Limits of Multi-head Transformer Dynamics
by: Bordelon, Blake, et al.
Published: (2024) -
Learning Curves for Noisy Heterogeneous Feature-Subsampled Ridge Ensembles
by: Ruben, Benjamin S., et al.
Published: (2023) -
How Feature Learning Can Improve Neural Scaling Laws
by: Bordelon, Blake, et al.
Published: (2024) -
A Dynamical Model of Neural Scaling Laws
by: Bordelon, Blake, et al.
Published: (2024) -
Theory of Scaling Laws for In-Context Regression: Depth, Width, Context and Time
by: Bordelon, Blake, et al.
Published: (2025)