Adaptive kernel predictors from feature-learning infinite limits of neural networks
Fuente:
arXiv
Salvato in:
| Autori principali: | Lauditi, Clarissa, Bordelon, Blake, Pehlevan, Cengiz |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Transfer Learning in Infinite Width Feature Learning Networks
di: Lauditi, Clarissa, et al.
Pubblicazione: (2025)
di: Lauditi, Clarissa, et al.
Pubblicazione: (2025)
Spectral Dynamics in Deep Networks: Feature Learning, Outlier Escape, and Learning Rate Transfer
di: Lauditi, Clarissa, et al.
Pubblicazione: (2026)
di: Lauditi, Clarissa, et al.
Pubblicazione: (2026)
Deep Linear Network Training Dynamics from Random Initialization: Data, Width, Depth, and Hyperparameter Transfer
di: Bordelon, Blake, et al.
Pubblicazione: (2025)
di: Bordelon, Blake, et al.
Pubblicazione: (2025)
Disordered Dynamics in High Dimensions: Connections to Random Matrices and Machine Learning
di: Bordelon, Blake, et al.
Pubblicazione: (2026)
di: Bordelon, Blake, et al.
Pubblicazione: (2026)
How Feature Learning Can Improve Neural Scaling Laws
di: Bordelon, Blake, et al.
Pubblicazione: (2024)
di: Bordelon, Blake, et al.
Pubblicazione: (2024)
A Dynamical Model of Neural Scaling Laws
di: Bordelon, Blake, et al.
Pubblicazione: (2024)
di: Bordelon, Blake, et al.
Pubblicazione: (2024)
Theory of Scaling Laws for In-Context Regression: Depth, Width, Context and Time
di: Bordelon, Blake, et al.
Pubblicazione: (2025)
di: Bordelon, Blake, et al.
Pubblicazione: (2025)
Infinite Limits of Multi-head Transformer Dynamics
di: Bordelon, Blake, et al.
Pubblicazione: (2024)
di: Bordelon, Blake, et al.
Pubblicazione: (2024)
Grokking as the Transition from Lazy to Rich Training Dynamics
di: Kumar, Tanishq, et al.
Pubblicazione: (2023)
di: Kumar, Tanishq, et al.
Pubblicazione: (2023)
Two-Point Deterministic Equivalence for Stochastic Gradient Dynamics in Linear Models
di: Atanasov, Alexander, et al.
Pubblicazione: (2025)
di: Atanasov, Alexander, et al.
Pubblicazione: (2025)
Dynamically Learning to Integrate in Recurrent Neural Networks
di: Bordelon, Blake, et al.
Pubblicazione: (2025)
di: Bordelon, Blake, et al.
Pubblicazione: (2025)
Nadaraya-Watson kernel smoothing as a random energy model
di: Zavatone-Veth, Jacob A., et al.
Pubblicazione: (2024)
di: Zavatone-Veth, Jacob A., et al.
Pubblicazione: (2024)
How does training shape the Riemannian geometry of neural network representations?
di: Zavatone-Veth, Jacob A., et al.
Pubblicazione: (2023)
di: Zavatone-Veth, Jacob A., et al.
Pubblicazione: (2023)
A solvable model of learning generative diffusion: theory and insights
di: Cui, Hugo, et al.
Pubblicazione: (2025)
di: Cui, Hugo, et al.
Pubblicazione: (2025)
Theory of Optimal Learning Rate Schedules and Scaling Laws for a Random Feature Model
di: Bordelon, Blake, et al.
Pubblicazione: (2026)
di: Bordelon, Blake, et al.
Pubblicazione: (2026)
Learning Curves for Noisy Heterogeneous Feature-Subsampled Ridge Ensembles
di: Ruben, Benjamin S., et al.
Pubblicazione: (2023)
di: Ruben, Benjamin S., et al.
Pubblicazione: (2023)
Risk and cross validation in ridge regression with correlated samples
di: Atanasov, Alexander, et al.
Pubblicazione: (2024)
di: Atanasov, Alexander, et al.
Pubblicazione: (2024)
Scaling and renormalization in high-dimensional regression
di: Atanasov, Alexander, et al.
Pubblicazione: (2024)
di: Atanasov, Alexander, et al.
Pubblicazione: (2024)
Asymptotic theory of in-context learning by linear attention
di: Lu, Yue M., et al.
Pubblicazione: (2024)
di: Lu, Yue M., et al.
Pubblicazione: (2024)
No Free Lunch From Random Feature Ensembles: Scaling Laws and Near-Optimality Conditions
di: Ruben, Benjamin S., et al.
Pubblicazione: (2024)
di: Ruben, Benjamin S., et al.
Pubblicazione: (2024)
Random Features Hopfield Networks generalize retrieval to previously unseen examples
di: Kalaj, Silvio, et al.
Pubblicazione: (2024)
di: Kalaj, Silvio, et al.
Pubblicazione: (2024)
Spring-block theory of feature learning in deep neural networks
di: Shi, Cheng, et al.
Pubblicazione: (2024)
di: Shi, Cheng, et al.
Pubblicazione: (2024)
High-dimensional learning of narrow neural networks
di: Cui, Hugo
Pubblicazione: (2024)
di: Cui, Hugo
Pubblicazione: (2024)
A note on the dynamics of extended-context disordered kinetic spin models
di: Zavatone-Veth, Jacob A., et al.
Pubblicazione: (2025)
di: Zavatone-Veth, Jacob A., et al.
Pubblicazione: (2025)
Generalization performance of narrow one-hidden layer networks in the teacher-student setting
di: Ortiz, Rodrigo Pérez, et al.
Pubblicazione: (2025)
di: Ortiz, Rodrigo Pérez, et al.
Pubblicazione: (2025)
Asymptotics of feature learning in two-layer networks after one gradient-step
di: Cui, Hugo, et al.
Pubblicazione: (2024)
di: Cui, Hugo, et al.
Pubblicazione: (2024)
Implicit bias produces neural scaling laws in learning curves, from perceptrons to deep networks
di: D'Amico, Francesco, et al.
Pubblicazione: (2025)
di: D'Amico, Francesco, et al.
Pubblicazione: (2025)
Exact full-RSB SAT/UNSAT transition in infinitely wide two-layer neural networks
di: Annesi, Brandon L., et al.
Pubblicazione: (2024)
di: Annesi, Brandon L., et al.
Pubblicazione: (2024)
Critical feature learning in deep neural networks
di: Fischer, Kirsten, et al.
Pubblicazione: (2024)
di: Fischer, Kirsten, et al.
Pubblicazione: (2024)
Deep neural networks from the perspective of ergodic theory
di: Zhang, Fan
Pubblicazione: (2023)
di: Zhang, Fan
Pubblicazione: (2023)
A unified theory of feature learning in RNNs and DNNs
di: Bauer, Jan P., et al.
Pubblicazione: (2026)
di: Bauer, Jan P., et al.
Pubblicazione: (2026)
Emergent weight morphologies in deep neural networks
di: de Jong, Pascal, et al.
Pubblicazione: (2025)
di: de Jong, Pascal, et al.
Pubblicazione: (2025)
On the role of non-linear latent features in bipartite generative neural networks
di: Bonnaire, Tony, et al.
Pubblicazione: (2025)
di: Bonnaire, Tony, et al.
Pubblicazione: (2025)
A generalized neural tangent kernel for surrogate gradient learning
di: Eilers, Luke, et al.
Pubblicazione: (2024)
di: Eilers, Luke, et al.
Pubblicazione: (2024)
Supervised and Unsupervised protocols for hetero-associative neural networks
di: Alessandrelli, Andrea, et al.
Pubblicazione: (2025)
di: Alessandrelli, Andrea, et al.
Pubblicazione: (2025)
Computing frustration and near-monotonicity in deep neural networks
di: Wendin, Joel, et al.
Pubblicazione: (2025)
di: Wendin, Joel, et al.
Pubblicazione: (2025)
Finite-time Lyapunov exponents of deep neural networks
di: Storm, L., et al.
Pubblicazione: (2023)
di: Storm, L., et al.
Pubblicazione: (2023)
Fundamental limits of learning in sequence multi-index models and deep attention networks: High-dimensional asymptotics and sharp thresholds
di: Troiani, Emanuele, et al.
Pubblicazione: (2025)
di: Troiani, Emanuele, et al.
Pubblicazione: (2025)
Training neural networks with structured noise improves classification and generalization
di: Benedetti, Marco, et al.
Pubblicazione: (2023)
di: Benedetti, Marco, et al.
Pubblicazione: (2023)
Coding schemes in neural networks learning classification tasks
di: van Meegen, Alexander, et al.
Pubblicazione: (2024)
di: van Meegen, Alexander, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Transfer Learning in Infinite Width Feature Learning Networks
di: Lauditi, Clarissa, et al.
Pubblicazione: (2025) -
Spectral Dynamics in Deep Networks: Feature Learning, Outlier Escape, and Learning Rate Transfer
di: Lauditi, Clarissa, et al.
Pubblicazione: (2026) -
Deep Linear Network Training Dynamics from Random Initialization: Data, Width, Depth, and Hyperparameter Transfer
di: Bordelon, Blake, et al.
Pubblicazione: (2025) -
Disordered Dynamics in High Dimensions: Connections to Random Matrices and Machine Learning
di: Bordelon, Blake, et al.
Pubblicazione: (2026) -
How Feature Learning Can Improve Neural Scaling Laws
di: Bordelon, Blake, et al.
Pubblicazione: (2024)