Compressibility Measures Complexity: Minimum Description Length Meets Singular Learning Theory
Fuente:
arXiv
Saved in:
| Main Authors: | Urdshals, Einar, Lau, Edmund, Hoogland, Jesse, van Wingerden, Stan, Murfet, Daniel |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Differentiation and Specialization of Attention Heads via the Refined Local Learning Coefficient
by: Wang, George, et al.
Published: (2024)
by: Wang, George, et al.
Published: (2024)
Interpreting Reinforcement Learning Agents with Susceptibilities
by: Elliott, Chris, et al.
Published: (2026)
by: Elliott, Chris, et al.
Published: (2026)
The Local Learning Coefficient: A Singularity-Aware Complexity Measure
by: Lau, Edmund, et al.
Published: (2023)
by: Lau, Edmund, et al.
Published: (2023)
Stagewise Reinforcement Learning and the Geometry of the Regret Landscape
by: Elliott, Chris, et al.
Published: (2026)
by: Elliott, Chris, et al.
Published: (2026)
Structure Development in List-Sorting Transformers
by: Urdshals, Einar, et al.
Published: (2025)
by: Urdshals, Einar, et al.
Published: (2025)
Structural Inference: Interpreting Small Language Models with Susceptibilities
by: Baker, Garrett, et al.
Published: (2025)
by: Baker, Garrett, et al.
Published: (2025)
Towards Spectroscopy: Susceptibility Clusters in Language Models
by: Gordon, Andrew, et al.
Published: (2026)
by: Gordon, Andrew, et al.
Published: (2026)
Dynamics of Transient Structure in In-Context Linear Regression Transformers
by: Carroll, Liam, et al.
Published: (2025)
by: Carroll, Liam, et al.
Published: (2025)
Programs as Singularities
by: Murfet, Daniel, et al.
Published: (2025)
by: Murfet, Daniel, et al.
Published: (2025)
Linear Response Estimators for Singular Statistical Models
by: Elliott, Chris, et al.
Published: (2026)
by: Elliott, Chris, et al.
Published: (2026)
Loss Landscape Degeneracy and Stagewise Development in Transformers
by: Hoogland, Jesse, et al.
Published: (2024)
by: Hoogland, Jesse, et al.
Published: (2024)
From Global to Local: A Scalable Benchmark for Local Posterior Sampling
by: Hitchcock, Rohan, et al.
Published: (2025)
by: Hitchcock, Rohan, et al.
Published: (2025)
The Loss Kernel: A Geometric Probe for Deep Learning Interpretability
by: Adam, Maxwell, et al.
Published: (2025)
by: Adam, Maxwell, et al.
Published: (2025)
Modes of Sequence Models and Learning Coefficients
by: Chen, Zhongtian, et al.
Published: (2025)
by: Chen, Zhongtian, et al.
Published: (2025)
A Theory of Machine Understanding via the Minimum Description Length Principle
by: Zhang, Canlin, et al.
Published: (2025)
by: Zhang, Canlin, et al.
Published: (2025)
Minimum Description Length and Generalization Guarantees for Representation Learning
by: Sefidgaran, Milad, et al.
Published: (2024)
by: Sefidgaran, Milad, et al.
Published: (2024)
Beyond Tokens in Language Models: Interpreting Activations through Text Genre Chunks
by: Benito-Rodriguez, Éloïse, et al.
Published: (2025)
by: Benito-Rodriguez, Éloïse, et al.
Published: (2025)
You Are What You Eat -- AI Alignment Requires Understanding How Data Shapes Structure and Generalisation
by: Lehalleur, Simon Pepin, et al.
Published: (2025)
by: Lehalleur, Simon Pepin, et al.
Published: (2025)
Patterning: The Dual of Interpretability
by: Wang, George, et al.
Published: (2026)
by: Wang, George, et al.
Published: (2026)
Susceptibilities and Patterning: A Primer on Linear Response in Bayesian Learning
by: Elliott, Chris, et al.
Published: (2026)
by: Elliott, Chris, et al.
Published: (2026)
A Minimum Description Length Approach to Regularization in Neural Networks
by: Abudy, Matan, et al.
Published: (2025)
by: Abudy, Matan, et al.
Published: (2025)
Dark Matter-induced electron excitations in silicon and germanium with Deep Learning
by: Catena, Riccardo, et al.
Published: (2024)
by: Catena, Riccardo, et al.
Published: (2024)
Influence Dynamics and Stagewise Data Attribution
by: Lee, Jin Hwa, et al.
Published: (2025)
by: Lee, Jin Hwa, et al.
Published: (2025)
MDL-Pool: Adaptive Multilevel Graph Pooling Based on Minimum Description Length
by: von Pichowski, Jan, et al.
Published: (2024)
by: von Pichowski, Jan, et al.
Published: (2024)
Complexity Reduction in Machine Learning-Based Wireless Positioning: Minimum Description Features
by: Oh, Myeung Suk, et al.
Published: (2024)
by: Oh, Myeung Suk, et al.
Published: (2024)
Minimum Description Feature Selection for Complexity Reduction in Machine Learning-based Wireless Positioning
by: Oh, Myeung Suk, et al.
Published: (2024)
by: Oh, Myeung Suk, et al.
Published: (2024)
Estimating the Local Learning Coefficient at Scale
by: Furman, Zach, et al.
Published: (2024)
by: Furman, Zach, et al.
Published: (2024)
Minimum Description Length based Granular-Ball Tree Regularization for Spectral Clustering
by: Xian, Zeqiang, et al.
Published: (2026)
by: Xian, Zeqiang, et al.
Published: (2026)
Embryology of a Language Model
by: Wang, George, et al.
Published: (2025)
by: Wang, George, et al.
Published: (2025)
Bayesian Influence Functions for Hessian-Free Data Attribution
by: Kreer, Philipp Alexander, et al.
Published: (2025)
by: Kreer, Philipp Alexander, et al.
Published: (2025)
Excess Description Length of Learning Generalizable Predictors
by: Donoway, Elizabeth, et al.
Published: (2026)
by: Donoway, Elizabeth, et al.
Published: (2026)
Bridging Kolmogorov Complexity and Deep Learning: Asymptotically Optimal Description Length Objectives for Transformers
by: Shaw, Peter, et al.
Published: (2025)
by: Shaw, Peter, et al.
Published: (2025)
Provably Minimum-Length Conformal Prediction Sets for Ordinal Classification
by: Zhang, Zijian, et al.
Published: (2025)
by: Zhang, Zijian, et al.
Published: (2025)
A Boundary-Aware Non-parametric Granular-Ball Classifier Based on Minimum Description Length
by: Xian, Zeqiang, et al.
Published: (2026)
by: Xian, Zeqiang, et al.
Published: (2026)
Granger Causal Inference in Multivariate Hawkes Processes by Minimum Message Length
by: Hlavackova-Schindler, Katerina, et al.
Published: (2023)
by: Hlavackova-Schindler, Katerina, et al.
Published: (2023)
Fed-CVLC: Compressing Federated Learning Communications with Variable-Length Codes
by: Su, Xiaoxin, et al.
Published: (2024)
by: Su, Xiaoxin, et al.
Published: (2024)
A Concise Mathematical Description of Active Inference in Discrete Time
by: van Oostrum, Jesse, et al.
Published: (2024)
by: van Oostrum, Jesse, et al.
Published: (2024)
Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization
by: Amaefuna, Theophilus, et al.
Published: (2026)
by: Amaefuna, Theophilus, et al.
Published: (2026)
Multitask learning with semiempirical orbital charges enables sample-efficient MLIPs
by: Neporozhnii, Ihor, et al.
Published: (2026)
by: Neporozhnii, Ihor, et al.
Published: (2026)
A Basin-Selection Perspective on Grokking via Singular Learning Theory
by: Cullen, Ben, et al.
Published: (2026)
by: Cullen, Ben, et al.
Published: (2026)
Similar Items
-
Differentiation and Specialization of Attention Heads via the Refined Local Learning Coefficient
by: Wang, George, et al.
Published: (2024) -
Interpreting Reinforcement Learning Agents with Susceptibilities
by: Elliott, Chris, et al.
Published: (2026) -
The Local Learning Coefficient: A Singularity-Aware Complexity Measure
by: Lau, Edmund, et al.
Published: (2023) -
Stagewise Reinforcement Learning and the Geometry of the Regret Landscape
by: Elliott, Chris, et al.
Published: (2026) -
Structure Development in List-Sorting Transformers
by: Urdshals, Einar, et al.
Published: (2025)