Entropic Confinement and Mode Connectivity in Overparameterized Neural Networks
Fuente:
arXiv
Saved in:
| Main Authors: | Di Carlo, Luca, Goddard, Chase, Schwab, David J. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
When can in-context learning generalize out of task distribution?
by: Goddard, Chase, et al.
Published: (2025)
by: Goddard, Chase, et al.
Published: (2025)
Spectral Architecture Search for Neural Network Models
by: Peri, Gianluca, et al.
Published: (2025)
by: Peri, Gianluca, et al.
Published: (2025)
Statistical Mechanics and Artificial Neural Networks: Principles, Models, and Applications
by: Böttcher, Lucas, et al.
Published: (2024)
by: Böttcher, Lucas, et al.
Published: (2024)
Estimating Global Input Relevance and Enforcing Sparse Representations with a Scalable Spectral Neural Network Approach
by: Chicchi, Lorenzo, et al.
Published: (2024)
by: Chicchi, Lorenzo, et al.
Published: (2024)
Generalization vs. Specialization under Concept Shift
by: Nguyen, Alex, et al.
Published: (2024)
by: Nguyen, Alex, et al.
Published: (2024)
A Generative Neural Annealer for Black-Box Combinatorial Optimization
by: Zhang, Yuan-Hang, et al.
Published: (2025)
by: Zhang, Yuan-Hang, et al.
Published: (2025)
Stable Attractors for Neural networks classification via Ordinary Differential Equations (SA-nODE)
by: Marino, Raffaele, et al.
Published: (2023)
by: Marino, Raffaele, et al.
Published: (2023)
Deterministic versus stochastic dynamical classifiers: opposing random adversarial attacks with noise
by: Chicchi, Lorenzo, et al.
Published: (2024)
by: Chicchi, Lorenzo, et al.
Published: (2024)
Deep Neural Nets as Hamiltonians
by: Winer, Mike, et al.
Published: (2025)
by: Winer, Mike, et al.
Published: (2025)
Escape dynamics and implicit bias of one-pass SGD in overparameterized quadratic networks
by: Bocchi, Dario, et al.
Published: (2026)
by: Bocchi, Dario, et al.
Published: (2026)
Phase transitions in the mini-batch size for sparse and dense two-layer neural networks
by: Marino, Raffaele, et al.
Published: (2023)
by: Marino, Raffaele, et al.
Published: (2023)
A Minimal Model of Representation Collapse: Frustration, Stop-Gradient, and Dynamics
by: Yao, Louie Hong, et al.
Published: (2026)
by: Yao, Louie Hong, et al.
Published: (2026)
Inferring Higher-Order Couplings with Neural Networks
by: Decelle, Aurélien, et al.
Published: (2025)
by: Decelle, Aurélien, et al.
Published: (2025)
Gaussian Universality in Neural Network Dynamics with Generalized Structured Input Distributions
by: Bae, Jaeyong, et al.
Published: (2024)
by: Bae, Jaeyong, et al.
Published: (2024)
An Analytical Characterization of Sloppiness in Neural Networks: Insights from Linear Models
by: Mao, Jialin, et al.
Published: (2025)
by: Mao, Jialin, et al.
Published: (2025)
Universal Scaling Laws of Absorbing Phase Transitions in Artificial Deep Neural Networks
by: Tamai, Keiichi, et al.
Published: (2023)
by: Tamai, Keiichi, et al.
Published: (2023)
Data coarse graining can improve model performance
by: Nguyen, Alex, et al.
Published: (2025)
by: Nguyen, Alex, et al.
Published: (2025)
Performance of machine-learning-assisted Monte Carlo in sampling from simple statistical physics models
by: Del Bono, Luca Maria, et al.
Published: (2025)
by: Del Bono, Luca Maria, et al.
Published: (2025)
Demonstrating Real Advantage of Machine-Learning-Enhanced Monte Carlo for Combinatorial Optimization
by: Del Bono, Luca Maria, et al.
Published: (2025)
by: Del Bono, Luca Maria, et al.
Published: (2025)
The critical slowing down in diffusion models
by: Del Bono, Luca Maria, et al.
Published: (2026)
by: Del Bono, Luca Maria, et al.
Published: (2026)
Biased Generalization in Diffusion Models
by: Garnier-Brun, Jerome, et al.
Published: (2026)
by: Garnier-Brun, Jerome, et al.
Published: (2026)
What makes it possible to learn probability distributions in the natural world?
by: Bialek, William, et al.
Published: (2020)
by: Bialek, William, et al.
Published: (2020)
Stochastic Resetting Accelerates Policy Convergence in Reinforcement Learning
by: Zhou, Jello, et al.
Published: (2026)
by: Zhou, Jello, et al.
Published: (2026)
Pseudo-likelihood produces associative memories able to generalize, even for asymmetric couplings
by: D'Amico, Francesco, et al.
Published: (2025)
by: D'Amico, Francesco, et al.
Published: (2025)
Message Passing Variational Autoregressive Network for Solving Intractable Ising Models
by: Ma, Qunlong, et al.
Published: (2024)
by: Ma, Qunlong, et al.
Published: (2024)
Neural Thermodynamics: Entropic Forces in Deep and Universal Representation Learning
by: Ziyin, Liu, et al.
Published: (2025)
by: Ziyin, Liu, et al.
Published: (2025)
Intuition emerges in Maximum Caliber models at criticality
by: Arola-Fernández, Lluís
Published: (2025)
by: Arola-Fernández, Lluís
Published: (2025)
Emergent Slow Thinking in LLMs as Inverse Tree Freezing
by: Hu, Sihan, et al.
Published: (2025)
by: Hu, Sihan, et al.
Published: (2025)
Preferential attachment and power-law degree distributions in heterogeneous multilayer hypergraphs
by: Di Lauro, Francesco, et al.
Published: (2025)
by: Di Lauro, Francesco, et al.
Published: (2025)
Inference in Spreading Processes with Neural-Network Priors
by: Ghio, Davide, et al.
Published: (2025)
by: Ghio, Davide, et al.
Published: (2025)
Dissecting a Small Artificial Neural Network
by: Yang, Xiguang, et al.
Published: (2025)
by: Yang, Xiguang, et al.
Published: (2025)
How do Probabilistic Graphical Models and Graph Neural Networks Look at Network Data?
by: Lapenna, Michela, et al.
Published: (2025)
by: Lapenna, Michela, et al.
Published: (2025)
Disorder-Induced Anomalous Mobility Enhancement in Confined Geometries
by: Shafir, Dan, et al.
Published: (2024)
by: Shafir, Dan, et al.
Published: (2024)
Overparametrization bends the landscape: BBP transitions at initialization in simple Neural Networks
by: Annesi, Brandon Livio, et al.
Published: (2025)
by: Annesi, Brandon Livio, et al.
Published: (2025)
Analytic theory of dropout regularization
by: Mori, Francesco, et al.
Published: (2025)
by: Mori, Francesco, et al.
Published: (2025)
The Copycat Perceptron: Smashing Barriers Through Collective Learning
by: Catania, Giovanni, et al.
Published: (2023)
by: Catania, Giovanni, et al.
Published: (2023)
Fundamental operating regimes, hyper-parameter fine-tuning and glassiness: towards an interpretable replica-theory for trained restricted Boltzmann machines
by: Fachechi, Alberto, et al.
Published: (2024)
by: Fachechi, Alberto, et al.
Published: (2024)
The autoregressive neural network architecture of the Boltzmann distribution of pairwise interacting spins systems
by: Biazzo, Indaco
Published: (2023)
by: Biazzo, Indaco
Published: (2023)
Distinct mechanisms underlying in-context learning in transformers
by: Gibson, Cole, et al.
Published: (2026)
by: Gibson, Cole, et al.
Published: (2026)
Interpreting the Synchronization Gap: The Hidden Mechanism Inside Diffusion Transformers
by: Albrychiewicz, Emil, et al.
Published: (2026)
by: Albrychiewicz, Emil, et al.
Published: (2026)
Similar Items
-
When can in-context learning generalize out of task distribution?
by: Goddard, Chase, et al.
Published: (2025) -
Spectral Architecture Search for Neural Network Models
by: Peri, Gianluca, et al.
Published: (2025) -
Statistical Mechanics and Artificial Neural Networks: Principles, Models, and Applications
by: Böttcher, Lucas, et al.
Published: (2024) -
Estimating Global Input Relevance and Enforcing Sparse Representations with a Scalable Spectral Neural Network Approach
by: Chicchi, Lorenzo, et al.
Published: (2024) -
Generalization vs. Specialization under Concept Shift
by: Nguyen, Alex, et al.
Published: (2024)