Is Grokking a Computational Glass Relaxation?
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Xiaotian, Shang, Yue, Yang, Entao, Zhang, Ge |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Grokking as the Transition from Lazy to Rich Training Dynamics
by: Kumar, Tanishq, et al.
Published: (2023)
by: Kumar, Tanishq, et al.
Published: (2023)
Grokking as a First Order Phase Transition in Two Layer Networks
by: Rubin, Noa, et al.
Published: (2023)
by: Rubin, Noa, et al.
Published: (2023)
Grokking at the Edge of Linear Separability
by: Beck, Alon, et al.
Published: (2024)
by: Beck, Alon, et al.
Published: (2024)
Grokking vs. Learning: Same Features, Different Encodings
by: Manning-Coe, Dmitry, et al.
Published: (2025)
by: Manning-Coe, Dmitry, et al.
Published: (2025)
Grokking in Linear Estimators -- A Solvable Model that Groks without Understanding
by: Levi, Noam, et al.
Published: (2023)
by: Levi, Noam, et al.
Published: (2023)
Demolition and Reinforcement of Memories in Spin-Glass-like Neural Networks
by: Ventura, Enrico
Published: (2024)
by: Ventura, Enrico
Published: (2024)
Grokking in the Ising Model
by: Hutchison, Karolina, et al.
Published: (2025)
by: Hutchison, Karolina, et al.
Published: (2025)
Predicting and Interpreting Energy Barriers of Metallic Glasses with Graph Neural Networks
by: Li, Haoyu, et al.
Published: (2023)
by: Li, Haoyu, et al.
Published: (2023)
Grokking as Dimensional Phase Transition in Neural Networks
by: Wang, Ping
Published: (2026)
by: Wang, Ping
Published: (2026)
Dimensional Criticality at Grokking Across MLPs and Transformers
by: Wang, Ping
Published: (2026)
by: Wang, Ping
Published: (2026)
Computing frustration and near-monotonicity in deep neural networks
by: Wendin, Joel, et al.
Published: (2025)
by: Wendin, Joel, et al.
Published: (2025)
Transient learning dynamics drive escape from sharp valleys in Stochastic Gradient Descent
by: Yang, Ning, et al.
Published: (2026)
by: Yang, Ning, et al.
Published: (2026)
Deep neural networks from the perspective of ergodic theory
by: Zhang, Fan
Published: (2023)
by: Zhang, Fan
Published: (2023)
Algorithmic Task Capture, Computational Complexity, and Inductive Bias of Infinite Transformers
by: Davidovich, Orit, et al.
Published: (2026)
by: Davidovich, Orit, et al.
Published: (2026)
Grokking Modular Polynomials
by: Doshi, Darshil, et al.
Published: (2024)
by: Doshi, Darshil, et al.
Published: (2024)
Computational Thresholds in Multi-Modal Learning via the Spiked Matrix-Tensor Model
by: Tabanelli, Hugo, et al.
Published: (2025)
by: Tabanelli, Hugo, et al.
Published: (2025)
A solvable model of learning generative diffusion: theory and insights
by: Cui, Hugo, et al.
Published: (2025)
by: Cui, Hugo, et al.
Published: (2025)
A Spin Glass Characterization of Neural Networks
by: Li, Jun
Published: (2025)
by: Li, Jun
Published: (2025)
Asymptotic theory of in-context learning by linear attention
by: Lu, Yue M., et al.
Published: (2024)
by: Lu, Yue M., et al.
Published: (2024)
Asymptotics of feature learning in two-layer networks after one gradient-step
by: Cui, Hugo, et al.
Published: (2024)
by: Cui, Hugo, et al.
Published: (2024)
A Generative Diffusion Model for Amorphous Materials
by: Yang, Kai, et al.
Published: (2025)
by: Yang, Kai, et al.
Published: (2025)
How does training shape the Riemannian geometry of neural network representations?
by: Zavatone-Veth, Jacob A., et al.
Published: (2023)
by: Zavatone-Veth, Jacob A., et al.
Published: (2023)
The Training Process of Many Deep Networks Explores the Same Low-Dimensional Manifold
by: Mao, Jialin, et al.
Published: (2023)
by: Mao, Jialin, et al.
Published: (2023)
Pattern Expansion of Spin Glasses
by: Shen, Mutian, et al.
Published: (2026)
by: Shen, Mutian, et al.
Published: (2026)
Nadaraya-Watson kernel smoothing as a random energy model
by: Zavatone-Veth, Jacob A., et al.
Published: (2024)
by: Zavatone-Veth, Jacob A., et al.
Published: (2024)
Asymptotic generalization error of a single-layer graph convolutional network
by: Duranthon, O., et al.
Published: (2024)
by: Duranthon, O., et al.
Published: (2024)
Optimal thresholds and algorithms for a model of multi-modal learning in high dimensions
by: Keup, Christian, et al.
Published: (2024)
by: Keup, Christian, et al.
Published: (2024)
Theory of Optimal Learning Rate Schedules and Scaling Laws for a Random Feature Model
by: Bordelon, Blake, et al.
Published: (2026)
by: Bordelon, Blake, et al.
Published: (2026)
Regularization, early-stopping and dreaming: a Hopfield-like setup to address generalization and overfitting
by: Agliari, Elena, et al.
Published: (2023)
by: Agliari, Elena, et al.
Published: (2023)
Neural Networks as Spin Models: From Glass to Hidden Order Through Training
by: Barney, Richard, et al.
Published: (2024)
by: Barney, Richard, et al.
Published: (2024)
Graph Learning Metallic Glass Discovery from Wikipedia
by: Ouyang, K. -C., et al.
Published: (2025)
by: Ouyang, K. -C., et al.
Published: (2025)
Learning with Restricted Boltzmann Machines: Asymptotics of AMP and GD in High Dimensions
by: Xu, Yizhou, et al.
Published: (2025)
by: Xu, Yizhou, et al.
Published: (2025)
Statistical physics analysis of graph neural networks: Approaching optimality in the contextual stochastic block model
by: Duranthon, O., et al.
Published: (2025)
by: Duranthon, O., et al.
Published: (2025)
Siamese Neural Network for Label-Efficient Critical Phenomena Prediction in 3D Percolation Models
by: Wang, Shanshan, et al.
Published: (2025)
by: Wang, Shanshan, et al.
Published: (2025)
Perfect reconstruction of sparse signals using nonconvexity control and one-step RSB message passing
by: Gu, Xiaosi, et al.
Published: (2025)
by: Gu, Xiaosi, et al.
Published: (2025)
Saddle Hierarchy in Dense Associative Memory
by: Thériault, Robin, et al.
Published: (2025)
by: Thériault, Robin, et al.
Published: (2025)
Statistical Physics of Deep Neural Networks: Generalization Capability, Beyond the Infinite Width, and Feature Learning
by: Ariosto, Sebastiano
Published: (2025)
by: Ariosto, Sebastiano
Published: (2025)
Nonlocal Monte Carlo via Reinforcement Learning
by: Dobrynin, Dmitrii, et al.
Published: (2025)
by: Dobrynin, Dmitrii, et al.
Published: (2025)
Graph Neural Network Approach to Predicting Magnetization in Quasi-One-Dimensional Ising Systems
by: Slavin, V., et al.
Published: (2025)
by: Slavin, V., et al.
Published: (2025)
Exploring the Energy Landscape of RBMs: Reciprocal Space Insights into Bosons, Hierarchical Learning and Symmetry Breaking
by: Toledo-Marin, J. Quetzalcóatl, et al.
Published: (2025)
by: Toledo-Marin, J. Quetzalcóatl, et al.
Published: (2025)
Similar Items
-
Grokking as the Transition from Lazy to Rich Training Dynamics
by: Kumar, Tanishq, et al.
Published: (2023) -
Grokking as a First Order Phase Transition in Two Layer Networks
by: Rubin, Noa, et al.
Published: (2023) -
Grokking at the Edge of Linear Separability
by: Beck, Alon, et al.
Published: (2024) -
Grokking vs. Learning: Same Features, Different Encodings
by: Manning-Coe, Dmitry, et al.
Published: (2025) -
Grokking in Linear Estimators -- A Solvable Model that Groks without Understanding
by: Levi, Noam, et al.
Published: (2023)