Transfer Learning in Infinite Width Feature Learning Networks
Fuente:
arXiv
Guardado en:
| Autores principales: | Lauditi, Clarissa, Bordelon, Blake, Pehlevan, Cengiz |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Spectral Dynamics in Deep Networks: Feature Learning, Outlier Escape, and Learning Rate Transfer
por: Lauditi, Clarissa, et al.
Publicado: (2026)
por: Lauditi, Clarissa, et al.
Publicado: (2026)
Adaptive kernel predictors from feature-learning infinite limits of neural networks
por: Lauditi, Clarissa, et al.
Publicado: (2025)
por: Lauditi, Clarissa, et al.
Publicado: (2025)
Deep Linear Network Training Dynamics from Random Initialization: Data, Width, Depth, and Hyperparameter Transfer
por: Bordelon, Blake, et al.
Publicado: (2025)
por: Bordelon, Blake, et al.
Publicado: (2025)
Disordered Dynamics in High Dimensions: Connections to Random Matrices and Machine Learning
por: Bordelon, Blake, et al.
Publicado: (2026)
por: Bordelon, Blake, et al.
Publicado: (2026)
How Feature Learning Can Improve Neural Scaling Laws
por: Bordelon, Blake, et al.
Publicado: (2024)
por: Bordelon, Blake, et al.
Publicado: (2024)
Theory of Scaling Laws for In-Context Regression: Depth, Width, Context and Time
por: Bordelon, Blake, et al.
Publicado: (2025)
por: Bordelon, Blake, et al.
Publicado: (2025)
Infinite Limits of Multi-head Transformer Dynamics
por: Bordelon, Blake, et al.
Publicado: (2024)
por: Bordelon, Blake, et al.
Publicado: (2024)
A Dynamical Model of Neural Scaling Laws
por: Bordelon, Blake, et al.
Publicado: (2024)
por: Bordelon, Blake, et al.
Publicado: (2024)
Dynamically Learning to Integrate in Recurrent Neural Networks
por: Bordelon, Blake, et al.
Publicado: (2025)
por: Bordelon, Blake, et al.
Publicado: (2025)
Grokking as the Transition from Lazy to Rich Training Dynamics
por: Kumar, Tanishq, et al.
Publicado: (2023)
por: Kumar, Tanishq, et al.
Publicado: (2023)
Two-Point Deterministic Equivalence for Stochastic Gradient Dynamics in Linear Models
por: Atanasov, Alexander, et al.
Publicado: (2025)
por: Atanasov, Alexander, et al.
Publicado: (2025)
Theory of Optimal Learning Rate Schedules and Scaling Laws for a Random Feature Model
por: Bordelon, Blake, et al.
Publicado: (2026)
por: Bordelon, Blake, et al.
Publicado: (2026)
Learning Curves for Noisy Heterogeneous Feature-Subsampled Ridge Ensembles
por: Ruben, Benjamin S., et al.
Publicado: (2023)
por: Ruben, Benjamin S., et al.
Publicado: (2023)
Statistical Physics of Deep Neural Networks: Generalization Capability, Beyond the Infinite Width, and Feature Learning
por: Ariosto, Sebastiano
Publicado: (2025)
por: Ariosto, Sebastiano
Publicado: (2025)
Random Features Hopfield Networks generalize retrieval to previously unseen examples
por: Kalaj, Silvio, et al.
Publicado: (2024)
por: Kalaj, Silvio, et al.
Publicado: (2024)
Nadaraya-Watson kernel smoothing as a random energy model
por: Zavatone-Veth, Jacob A., et al.
Publicado: (2024)
por: Zavatone-Veth, Jacob A., et al.
Publicado: (2024)
No Free Lunch From Random Feature Ensembles: Scaling Laws and Near-Optimality Conditions
por: Ruben, Benjamin S., et al.
Publicado: (2024)
por: Ruben, Benjamin S., et al.
Publicado: (2024)
A solvable model of learning generative diffusion: theory and insights
por: Cui, Hugo, et al.
Publicado: (2025)
por: Cui, Hugo, et al.
Publicado: (2025)
Risk and cross validation in ridge regression with correlated samples
por: Atanasov, Alexander, et al.
Publicado: (2024)
por: Atanasov, Alexander, et al.
Publicado: (2024)
Scaling and renormalization in high-dimensional regression
por: Atanasov, Alexander, et al.
Publicado: (2024)
por: Atanasov, Alexander, et al.
Publicado: (2024)
How does training shape the Riemannian geometry of neural network representations?
por: Zavatone-Veth, Jacob A., et al.
Publicado: (2023)
por: Zavatone-Veth, Jacob A., et al.
Publicado: (2023)
Asymptotic theory of in-context learning by linear attention
por: Lu, Yue M., et al.
Publicado: (2024)
por: Lu, Yue M., et al.
Publicado: (2024)
A note on the dynamics of extended-context disordered kinetic spin models
por: Zavatone-Veth, Jacob A., et al.
Publicado: (2025)
por: Zavatone-Veth, Jacob A., et al.
Publicado: (2025)
Generalization performance of narrow one-hidden layer networks in the teacher-student setting
por: Ortiz, Rodrigo Pérez, et al.
Publicado: (2025)
por: Ortiz, Rodrigo Pérez, et al.
Publicado: (2025)
From Kernels to Features: A Multi-Scale Adaptive Theory of Feature Learning
por: Rubin, Noa, et al.
Publicado: (2025)
por: Rubin, Noa, et al.
Publicado: (2025)
Deep Learning as Neural Low-Degree Filtering: A Spectral Theory of Hierarchical Feature Learning
por: Dandi, Yatin, et al.
Publicado: (2026)
por: Dandi, Yatin, et al.
Publicado: (2026)
Algorithmic Task Capture, Computational Complexity, and Inductive Bias of Infinite Transformers
por: Davidovich, Orit, et al.
Publicado: (2026)
por: Davidovich, Orit, et al.
Publicado: (2026)
High-Dimensional Analysis of Gradient Flow for Extensive-Width Quadratic Neural Networks
por: Martin, Simon, et al.
Publicado: (2026)
por: Martin, Simon, et al.
Publicado: (2026)
How Deep Networks Learn Sparse and Hierarchical Data: the Sparse Random Hierarchy Model
por: Tomasini, Umberto, et al.
Publicado: (2024)
por: Tomasini, Umberto, et al.
Publicado: (2024)
Scaling Laws and Spectra of Shallow Neural Networks in the Feature Learning Regime
por: Defilippis, Leonardo, et al.
Publicado: (2025)
por: Defilippis, Leonardo, et al.
Publicado: (2025)
Prototype Analysis in Hopfield Networks with Hebbian Learning
por: McAlister, Hayden, et al.
Publicado: (2024)
por: McAlister, Hayden, et al.
Publicado: (2024)
Physical Reinforcement Learning
por: Dillavou, Sam, et al.
Publicado: (2025)
por: Dillavou, Sam, et al.
Publicado: (2025)
Exact Learning Dynamics of In-Context Learning in Linear Transformers and Its Application to Non-Linear Transformers
por: Mainali, Nischal, et al.
Publicado: (2025)
por: Mainali, Nischal, et al.
Publicado: (2025)
Impact of dendritic non-linearities on the computational capabilities of neurons
por: Lauditi, Clarissa, et al.
Publicado: (2024)
por: Lauditi, Clarissa, et al.
Publicado: (2024)
Learning and extrapolating scale-invariant processes
por: Alvez-Canepa, Anaclara, et al.
Publicado: (2026)
por: Alvez-Canepa, Anaclara, et al.
Publicado: (2026)
Hebbian Learning from First Principles
por: Albanese, Linda, et al.
Publicado: (2024)
por: Albanese, Linda, et al.
Publicado: (2024)
Asymptotics of Learning with Deep Structured (Random) Features
por: Schröder, Dominik, et al.
Publicado: (2024)
por: Schröder, Dominik, et al.
Publicado: (2024)
Nonlocal Monte Carlo via Reinforcement Learning
por: Dobrynin, Dmitrii, et al.
Publicado: (2025)
por: Dobrynin, Dmitrii, et al.
Publicado: (2025)
The effect of priors on Learning with Restricted Boltzmann Machines
por: Manzan, Gianluca, et al.
Publicado: (2024)
por: Manzan, Gianluca, et al.
Publicado: (2024)
Learning Linear Regression with Low-Rank Tasks in-Context
por: Takanami, Kaito, et al.
Publicado: (2025)
por: Takanami, Kaito, et al.
Publicado: (2025)
Ejemplares similares
-
Spectral Dynamics in Deep Networks: Feature Learning, Outlier Escape, and Learning Rate Transfer
por: Lauditi, Clarissa, et al.
Publicado: (2026) -
Adaptive kernel predictors from feature-learning infinite limits of neural networks
por: Lauditi, Clarissa, et al.
Publicado: (2025) -
Deep Linear Network Training Dynamics from Random Initialization: Data, Width, Depth, and Hyperparameter Transfer
por: Bordelon, Blake, et al.
Publicado: (2025) -
Disordered Dynamics in High Dimensions: Connections to Random Matrices and Machine Learning
por: Bordelon, Blake, et al.
Publicado: (2026) -
How Feature Learning Can Improve Neural Scaling Laws
por: Bordelon, Blake, et al.
Publicado: (2024)