How does training shape the Riemannian geometry of neural network representations?
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zavatone-Veth, Jacob A., Yang, Sheng, Rubinfien, Julian A., Pehlevan, Cengiz |
|---|---|
| Format: | Preprint |
| Publié: |
2023
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Nadaraya-Watson kernel smoothing as a random energy model
par: Zavatone-Veth, Jacob A., et autres
Publié: (2024)
par: Zavatone-Veth, Jacob A., et autres
Publié: (2024)
A note on the dynamics of extended-context disordered kinetic spin models
par: Zavatone-Veth, Jacob A., et autres
Publié: (2025)
par: Zavatone-Veth, Jacob A., et autres
Publié: (2025)
Risk and cross validation in ridge regression with correlated samples
par: Atanasov, Alexander, et autres
Publié: (2024)
par: Atanasov, Alexander, et autres
Publié: (2024)
Scaling and renormalization in high-dimensional regression
par: Atanasov, Alexander, et autres
Publié: (2024)
par: Atanasov, Alexander, et autres
Publié: (2024)
Dynamically Learning to Integrate in Recurrent Neural Networks
par: Bordelon, Blake, et autres
Publié: (2025)
par: Bordelon, Blake, et autres
Publié: (2025)
Two-Point Deterministic Equivalence for Stochastic Gradient Dynamics in Linear Models
par: Atanasov, Alexander, et autres
Publié: (2025)
par: Atanasov, Alexander, et autres
Publié: (2025)
Asymptotic theory of in-context learning by linear attention
par: Lu, Yue M., et autres
Publié: (2024)
par: Lu, Yue M., et autres
Publié: (2024)
Adaptive kernel predictors from feature-learning infinite limits of neural networks
par: Lauditi, Clarissa, et autres
Publié: (2025)
par: Lauditi, Clarissa, et autres
Publié: (2025)
How Feature Learning Can Improve Neural Scaling Laws
par: Bordelon, Blake, et autres
Publié: (2024)
par: Bordelon, Blake, et autres
Publié: (2024)
Disordered Dynamics in High Dimensions: Connections to Random Matrices and Machine Learning
par: Bordelon, Blake, et autres
Publié: (2026)
par: Bordelon, Blake, et autres
Publié: (2026)
Deep Linear Network Training Dynamics from Random Initialization: Data, Width, Depth, and Hyperparameter Transfer
par: Bordelon, Blake, et autres
Publié: (2025)
par: Bordelon, Blake, et autres
Publié: (2025)
Transfer Learning in Infinite Width Feature Learning Networks
par: Lauditi, Clarissa, et autres
Publié: (2025)
par: Lauditi, Clarissa, et autres
Publié: (2025)
A Dynamical Model of Neural Scaling Laws
par: Bordelon, Blake, et autres
Publié: (2024)
par: Bordelon, Blake, et autres
Publié: (2024)
Learning Curves for Noisy Heterogeneous Feature-Subsampled Ridge Ensembles
par: Ruben, Benjamin S., et autres
Publié: (2023)
par: Ruben, Benjamin S., et autres
Publié: (2023)
A solvable model of learning generative diffusion: theory and insights
par: Cui, Hugo, et autres
Publié: (2025)
par: Cui, Hugo, et autres
Publié: (2025)
Infinite Limits of Multi-head Transformer Dynamics
par: Bordelon, Blake, et autres
Publié: (2024)
par: Bordelon, Blake, et autres
Publié: (2024)
Theory of Scaling Laws for In-Context Regression: Depth, Width, Context and Time
par: Bordelon, Blake, et autres
Publié: (2025)
par: Bordelon, Blake, et autres
Publié: (2025)
Grokking as the Transition from Lazy to Rich Training Dynamics
par: Kumar, Tanishq, et autres
Publié: (2023)
par: Kumar, Tanishq, et autres
Publié: (2023)
Spectral Dynamics in Deep Networks: Feature Learning, Outlier Escape, and Learning Rate Transfer
par: Lauditi, Clarissa, et autres
Publié: (2026)
par: Lauditi, Clarissa, et autres
Publié: (2026)
No Free Lunch From Random Feature Ensembles: Scaling Laws and Near-Optimality Conditions
par: Ruben, Benjamin S., et autres
Publié: (2024)
par: Ruben, Benjamin S., et autres
Publié: (2024)
Applications of information geometry to spiking neural network behavior
par: Crosser, Jacob T., et autres
Publié: (2023)
par: Crosser, Jacob T., et autres
Publié: (2023)
Symmetry in language statistics shapes the geometry of model representations
par: Karkada, Dhruva, et autres
Publié: (2026)
par: Karkada, Dhruva, et autres
Publié: (2026)
Properties of the geometry of solutions and capacity of multi-layer neural networks with Rectified Linear Units activations
par: Baldassi, Carlo, et autres
Publié: (2019)
par: Baldassi, Carlo, et autres
Publié: (2019)
Emergent weight morphologies in deep neural networks
par: de Jong, Pascal, et autres
Publié: (2025)
par: de Jong, Pascal, et autres
Publié: (2025)
High-dimensional learning of narrow neural networks
par: Cui, Hugo
Publié: (2024)
par: Cui, Hugo
Publié: (2024)
Deep neural networks from the perspective of ergodic theory
par: Zhang, Fan
Publié: (2023)
par: Zhang, Fan
Publié: (2023)
Finite-time Lyapunov exponents of deep neural networks
par: Storm, L., et autres
Publié: (2023)
par: Storm, L., et autres
Publié: (2023)
Supervised and Unsupervised protocols for hetero-associative neural networks
par: Alessandrelli, Andrea, et autres
Publié: (2025)
par: Alessandrelli, Andrea, et autres
Publié: (2025)
Computing frustration and near-monotonicity in deep neural networks
par: Wendin, Joel, et autres
Publié: (2025)
par: Wendin, Joel, et autres
Publié: (2025)
Quantum sequel of neural network training
par: Zhang, Hao, et autres
Publié: (2025)
par: Zhang, Hao, et autres
Publié: (2025)
Training neural networks with structured noise improves classification and generalization
par: Benedetti, Marco, et autres
Publié: (2023)
par: Benedetti, Marco, et autres
Publié: (2023)
Boundary between noise and information applied to filtering neural network weight matrices
par: Staats, Max, et autres
Publié: (2022)
par: Staats, Max, et autres
Publié: (2022)
Improving deep neural network performance through sampling
par: Ghantasala, Lakshmi A., et autres
Publié: (2025)
par: Ghantasala, Lakshmi A., et autres
Publié: (2025)
Neuronal correlations shape the scaling behavior of memory capacity and nonlinear computational capability of reservoir recurrent neural networks
par: Takasu, Shotaro, et autres
Publié: (2025)
par: Takasu, Shotaro, et autres
Publié: (2025)
Generalized hetero-associative neural networks
par: Agliari, Elena, et autres
Publié: (2024)
par: Agliari, Elena, et autres
Publié: (2024)
Statistical physics analysis of graph neural networks: Approaching optimality in the contextual stochastic block model
par: Duranthon, O., et autres
Publié: (2025)
par: Duranthon, O., et autres
Publié: (2025)
Implicit bias produces neural scaling laws in learning curves, from perceptrons to deep networks
par: D'Amico, Francesco, et autres
Publié: (2025)
par: D'Amico, Francesco, et autres
Publié: (2025)
Solution space and storage capacity of fully connected two-layer neural networks with generic activation functions
par: Nishiyama, Sota, et autres
Publié: (2024)
par: Nishiyama, Sota, et autres
Publié: (2024)
Pruning-induced phases in fully-connected neural networks: the eumentia, the dementia, and the amentia
par: Pan, Haining, et autres
Publié: (2026)
par: Pan, Haining, et autres
Publié: (2026)
Critical feature learning in deep neural networks
par: Fischer, Kirsten, et autres
Publié: (2024)
par: Fischer, Kirsten, et autres
Publié: (2024)
Documents similaires
-
Nadaraya-Watson kernel smoothing as a random energy model
par: Zavatone-Veth, Jacob A., et autres
Publié: (2024) -
A note on the dynamics of extended-context disordered kinetic spin models
par: Zavatone-Veth, Jacob A., et autres
Publié: (2025) -
Risk and cross validation in ridge regression with correlated samples
par: Atanasov, Alexander, et autres
Publié: (2024) -
Scaling and renormalization in high-dimensional regression
par: Atanasov, Alexander, et autres
Publié: (2024) -
Dynamically Learning to Integrate in Recurrent Neural Networks
par: Bordelon, Blake, et autres
Publié: (2025)