Decoupling Dynamical Richness from Representation Learning: Towards Practical Measurement
Fuente:
arXiv
Salvato in:
| Autori principali: | Nam, Yoonsoo, Fonseca, Nayara, Lee, Seok Hyeong, Mingard, Chris, Goring, Niclas, Harzli, Ouns El, Erturk, Abdurrahman Hadi, Hayou, Soufiane, Louis, Ard A. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A simple mean field model of feature learning
di: Göring, Niclas, et al.
Pubblicazione: (2025)
di: Göring, Niclas, et al.
Pubblicazione: (2025)
Feature learning is decoupled from generalization in high capacity neural networks
di: Göring, Niclas Alexander, et al.
Pubblicazione: (2025)
di: Göring, Niclas Alexander, et al.
Pubblicazione: (2025)
An exactly solvable model for emergence and scaling laws in the multitask sparse parity problem
di: Nam, Yoonsoo, et al.
Pubblicazione: (2024)
di: Nam, Yoonsoo, et al.
Pubblicazione: (2024)
Sufficient Conditions for Stability of Minimum-Norm Interpolating Deep ReLU Networks
di: Harzli, Ouns El, et al.
Pubblicazione: (2026)
di: Harzli, Ouns El, et al.
Pubblicazione: (2026)
Supervised sparse auto-encoders for interpretable and compositional representations
di: Harzli, Ouns El, et al.
Pubblicazione: (2026)
di: Harzli, Ouns El, et al.
Pubblicazione: (2026)
Successive vertex orderings of connected graphs
di: Agrawal, Prarthana, et al.
Pubblicazione: (2026)
di: Agrawal, Prarthana, et al.
Pubblicazione: (2026)
Characterising the Inductive Biases of Neural Networks on Boolean Data
di: Mingard, Chris, et al.
Pubblicazione: (2025)
di: Mingard, Chris, et al.
Pubblicazione: (2025)
Exploiting the equivalence between quantum neural networks and perceptrons
di: Mingard, Chris, et al.
Pubblicazione: (2024)
di: Mingard, Chris, et al.
Pubblicazione: (2024)
Position: Solve Layerwise Linear Models First to Understand Neural Dynamical Phenomena (Neural Collapse, Emergence, Lazy/Rich Regime, and Grokking)
di: Nam, Yoonsoo, et al.
Pubblicazione: (2025)
di: Nam, Yoonsoo, et al.
Pubblicazione: (2025)
A Proof of Learning Rate Transfer under $μ$P
di: Hayou, Soufiane
Pubblicazione: (2025)
di: Hayou, Soufiane
Pubblicazione: (2025)
Optimal Embedding Learning Rate in LLMs: The Effect of Vocabulary Size
di: Hayou, Soufiane, et al.
Pubblicazione: (2025)
di: Hayou, Soufiane, et al.
Pubblicazione: (2025)
Deep neural networks have an inbuilt Occam's razor
di: Mingard, Chris, et al.
Pubblicazione: (2023)
di: Mingard, Chris, et al.
Pubblicazione: (2023)
Learning Rate Scaling across LoRA Ranks and Transfer to Full Finetuning
di: Chen, Nan, et al.
Pubblicazione: (2026)
di: Chen, Nan, et al.
Pubblicazione: (2026)
The Impact of Initialization on LoRA Finetuning Dynamics
di: Hayou, Soufiane, et al.
Pubblicazione: (2024)
di: Hayou, Soufiane, et al.
Pubblicazione: (2024)
The Myth of Expert Specialization in MoEs: Why Routing Reflects Geometry, Not Necessarily Domain Expertise
di: Wang, Xi, et al.
Pubblicazione: (2026)
di: Wang, Xi, et al.
Pubblicazione: (2026)
PLoP: Precise LoRA Placement for Efficient Finetuning of Large Models
di: Hayou, Soufiane, et al.
Pubblicazione: (2025)
di: Hayou, Soufiane, et al.
Pubblicazione: (2025)
LoRA+: Efficient Low Rank Adaptation of Large Models
di: Hayou, Soufiane, et al.
Pubblicazione: (2024)
di: Hayou, Soufiane, et al.
Pubblicazione: (2024)
From Neural Networks to Logical Theories: The Correspondence between Fibring Modal Logics and Fibring Neural Networks
di: Harzli, Ouns El, et al.
Pubblicazione: (2025)
di: Harzli, Ouns El, et al.
Pubblicazione: (2025)
Leave-one-out Distinguishability in Machine Learning
di: Ye, Jiayuan, et al.
Pubblicazione: (2023)
di: Ye, Jiayuan, et al.
Pubblicazione: (2023)
On the Stability of the Jacobian Matrix in Deep Neural Networks
di: Dadoun, Benjamin, et al.
Pubblicazione: (2025)
di: Dadoun, Benjamin, et al.
Pubblicazione: (2025)
How Bad is Training on Synthetic Data? A Statistical Analysis of Language Model Collapse
di: Seddik, Mohamed El Amine, et al.
Pubblicazione: (2024)
di: Seddik, Mohamed El Amine, et al.
Pubblicazione: (2024)
Maximizing the Potential of Synthetic Data: Insights from Random Matrix Theory
di: Firdoussi, Aymane El, et al.
Pubblicazione: (2024)
di: Firdoussi, Aymane El, et al.
Pubblicazione: (2024)
CLIPSE -- a minimalistic CLIP-based image search engine for research
di: Göring, Steve
Pubblicazione: (2025)
di: Göring, Steve
Pubblicazione: (2025)
$μ$pscaling small models: Principled warm starts and hyperparameter transfer
di: Ma, Yuxin, et al.
Pubblicazione: (2026)
di: Ma, Yuxin, et al.
Pubblicazione: (2026)
Out-of-Domain Generalization in Dynamical Systems Reconstruction
di: Göring, Niclas, et al.
Pubblicazione: (2024)
di: Göring, Niclas, et al.
Pubblicazione: (2024)
Towards solving industrial integer linear programs with Decoded Quantum Interferometry
di: Sabater, Francesc, et al.
Pubblicazione: (2025)
di: Sabater, Francesc, et al.
Pubblicazione: (2025)
Closed-form $\ell_r$ norm scaling with data for overparameterized linear regression and diagonal linear networks under $\ell_p$ bias
di: Zhang, Shuofeng, et al.
Pubblicazione: (2025)
di: Zhang, Shuofeng, et al.
Pubblicazione: (2025)
Position: Many generalization measures for deep learning are fragile
di: Zhang, Shuofeng, et al.
Pubblicazione: (2025)
di: Zhang, Shuofeng, et al.
Pubblicazione: (2025)
Properties of uniformly $3$-connected graphs
di: Göring, Frank, et al.
Pubblicazione: (2022)
di: Göring, Frank, et al.
Pubblicazione: (2022)
The Library and the Writing Centre Build a Workshop: Exploring the Impact of an Asynchronous Online Academic Integrity Course
di: Ard, Stephanie Evers, et al.
Pubblicazione: (2019)
di: Ard, Stephanie Evers, et al.
Pubblicazione: (2019)
Economic policy uncertainty and international corporate leasing
di: Goutham Abotula, et al.
Pubblicazione: (2026)
di: Goutham Abotula, et al.
Pubblicazione: (2026)
Decoupling Representation and Learning in Genetic Programming: the LaSER Approach
di: Le, Nam H., et al.
Pubblicazione: (2025)
di: Le, Nam H., et al.
Pubblicazione: (2025)
Evaluation Under Imperfect Benchmarks and Ratings: A Case Study in Text Simplification
di: Liu, Joseph, et al.
Pubblicazione: (2025)
di: Liu, Joseph, et al.
Pubblicazione: (2025)
Burden the Hand: Regulatory Intensity and Payout Policy
di: Yahia Abdelbar, et al.
Pubblicazione: (2026)
di: Yahia Abdelbar, et al.
Pubblicazione: (2026)
Cotype zeta functions enumerating subalgebras of $R$-algebras
di: Lee, Seok Hyeong, et al.
Pubblicazione: (2025)
di: Lee, Seok Hyeong, et al.
Pubblicazione: (2025)
Lifting problem for universal quadratic forms over totally real cubic number fields
di: Kim, Daejun, et al.
Pubblicazione: (2023)
di: Kim, Daejun, et al.
Pubblicazione: (2023)
Engineering of Bacterial‐Hybrid System via Interfacial Nanotechnology and Biological Compatibilization Strategy
di: Seok Hyeong Bu, et al.
Pubblicazione: (2025)
di: Seok Hyeong Bu, et al.
Pubblicazione: (2025)
Discussions on Driven Cavity Flow
di: Erturk, E.
Pubblicazione: (2004)
di: Erturk, E.
Pubblicazione: (2004)
Reflections on Currency Crises
di: Korkut Erturk
Pubblicazione: (2004)
di: Korkut Erturk
Pubblicazione: (2004)
Sport und Studienerfolg
di: Göring, Arne, et al.
Pubblicazione: (2020)
di: Göring, Arne, et al.
Pubblicazione: (2020)
Documenti analoghi
-
A simple mean field model of feature learning
di: Göring, Niclas, et al.
Pubblicazione: (2025) -
Feature learning is decoupled from generalization in high capacity neural networks
di: Göring, Niclas Alexander, et al.
Pubblicazione: (2025) -
An exactly solvable model for emergence and scaling laws in the multitask sparse parity problem
di: Nam, Yoonsoo, et al.
Pubblicazione: (2024) -
Sufficient Conditions for Stability of Minimum-Norm Interpolating Deep ReLU Networks
di: Harzli, Ouns El, et al.
Pubblicazione: (2026) -
Supervised sparse auto-encoders for interpretable and compositional representations
di: Harzli, Ouns El, et al.
Pubblicazione: (2026)