A Compression Perspective on Simplicity Bias
Fuente:
arXiv
Saved in:
| Main Authors: | Marty, Tom, Elmoznino, Eric, Gagnon, Leo, Kasetty, Tejas, Nishikawa-Toomey, Mizu, Mittal, Sarthak, Lajoie, Guillaume, Sridhar, Dhanya |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Next-Token Prediction Should be Ambiguity-Sensitive: A Meta-Learning Perspective
by: Gagnon, Leo, et al.
Published: (2025)
by: Gagnon, Leo, et al.
Published: (2025)
In-context learning and Occam's razor
by: Elmoznino, Eric, et al.
Published: (2024)
by: Elmoznino, Eric, et al.
Published: (2024)
Does learning the right latent variables necessarily improve in-context learning?
by: Mittal, Sarthak, et al.
Published: (2024)
by: Mittal, Sarthak, et al.
Published: (2024)
Beyond Distribution Sharpening: The Importance of Task Rewards
by: Mittal, Sarthak, et al.
Published: (2026)
by: Mittal, Sarthak, et al.
Published: (2026)
Evaluating Interventional Reasoning Capabilities of Large Language Models
by: Kasetty, Tejas, et al.
Published: (2024)
by: Kasetty, Tejas, et al.
Published: (2024)
A Complexity-Based Theory of Compositionality
by: Elmoznino, Eric, et al.
Published: (2024)
by: Elmoznino, Eric, et al.
Published: (2024)
Iterative Amortized Inference: Unifying In-Context Learning and Learned Optimizers
by: Mittal, Sarthak, et al.
Published: (2025)
by: Mittal, Sarthak, et al.
Published: (2025)
In-Context Parametric Inference: Point or Distribution Estimators?
by: Mittal, Sarthak, et al.
Published: (2025)
by: Mittal, Sarthak, et al.
Published: (2025)
Discrete, compositional, and symbolic representations through attractor dynamics
by: Nam, Andrew, et al.
Published: (2023)
by: Nam, Andrew, et al.
Published: (2023)
Amortized In-Context Bayesian Posterior Estimation
by: Mittal, Sarthak, et al.
Published: (2025)
by: Mittal, Sarthak, et al.
Published: (2025)
Understanding Simplicity Bias towards Compositional Mappings via Learning Dynamics
by: Ren, Yi, et al.
Published: (2024)
by: Ren, Yi, et al.
Published: (2024)
Bayesian learning of Causal Structure and Mechanisms with GFlowNets and Variational Bayes
by: Nishikawa-Toomey, Mizu, et al.
Published: (2022)
by: Nishikawa-Toomey, Mizu, et al.
Published: (2022)
Sparse Shift Autoencoders for Identifying Concepts from Large Language Model Activations
by: Joshi, Shruti, et al.
Published: (2025)
by: Joshi, Shruti, et al.
Published: (2025)
SimBa: Simplicity Bias for Scaling Up Parameters in Deep Reinforcement Learning
by: Lee, Hojoon, et al.
Published: (2024)
by: Lee, Hojoon, et al.
Published: (2024)
Changing the Training Data Distribution to Reduce Simplicity Bias Improves In-distribution Generalization
by: Nguyen, Dang, et al.
Published: (2024)
by: Nguyen, Dang, et al.
Published: (2024)
Masked Autoencoders that Feel the Heart: Unveiling Simplicity Bias for ECG Analyses
by: Xu, He-Yang, et al.
Published: (2025)
by: Xu, He-Yang, et al.
Published: (2025)
Do Quantum Neural Networks have Simplicity Bias?
by: Pointing, Jessica
Published: (2024)
by: Pointing, Jessica
Published: (2024)
Accelerating Training with Neuron Interaction and Nowcasting Networks
by: Knyazev, Boris, et al.
Published: (2024)
by: Knyazev, Boris, et al.
Published: (2024)
BoxRL-NNV: Boxed Refinement of Latin Hypercube Samples for Neural Network Verification
by: Das, Sarthak
Published: (2025)
by: Das, Sarthak
Published: (2025)
BONSAI: Bayesian Optimization with Natural Simplicity and Interpretability
by: Daulton, Samuel, et al.
Published: (2026)
by: Daulton, Samuel, et al.
Published: (2026)
From Isolation to Entanglement: When Do Interpretability Methods Identify and Disentangle Known Concepts?
by: Mueller, Aaron, et al.
Published: (2025)
by: Mueller, Aaron, et al.
Published: (2025)
JEDI: Jointly Embedded Inference of Neural Dynamics
by: Jamkhandi, Anirudh, et al.
Published: (2026)
by: Jamkhandi, Anirudh, et al.
Published: (2026)
Universal Reinforcement Learning in Coalgebras: Asynchronous Stochastic Computation via Conduction
by: Mahadevan, Sridhar
Published: (2025)
by: Mahadevan, Sridhar
Published: (2025)
Consciousness as a Functor
by: Mahadevan, Sridhar
Published: (2025)
by: Mahadevan, Sridhar
Published: (2025)
GAIA: Categorical Foundations of Generative AI
by: Mahadevan, Sridhar
Published: (2024)
by: Mahadevan, Sridhar
Published: (2024)
Universal Imitation Games
by: Mahadevan, Sridhar
Published: (2024)
by: Mahadevan, Sridhar
Published: (2024)
From Sparsity to Simplicity: Enabling Simpler Sequential Replacements via Sparse Attention Distillation
by: Ren, Yuxin, et al.
Published: (2026)
by: Ren, Yuxin, et al.
Published: (2026)
T2IBias: Uncovering Societal Bias Encoded in the Latent Space of Text-to-Image Generative Models
by: Sufian, Abu, et al.
Published: (2025)
by: Sufian, Abu, et al.
Published: (2025)
A Probabilistic Perspective on Unlearning and Alignment for Large Language Models
by: Scholten, Yan, et al.
Published: (2024)
by: Scholten, Yan, et al.
Published: (2024)
Emergent Bias and Fairness in Multi-Agent Decision Systems
by: Madigan, Maeve, et al.
Published: (2025)
by: Madigan, Maeve, et al.
Published: (2025)
Learning a Generic Value-Selection Heuristic Inside a Constraint Programming Solver
by: Marty, Tom, et al.
Published: (2023)
by: Marty, Tom, et al.
Published: (2023)
Artificial Intelligence Ecosystem for Automating Self-Directed Teaching
by: Gotavade, Tejas Satish
Published: (2024)
by: Gotavade, Tejas Satish
Published: (2024)
Mind the Gap: A Causal Perspective on Bias Amplification in Prediction & Decision-Making
by: Plecko, Drago, et al.
Published: (2024)
by: Plecko, Drago, et al.
Published: (2024)
Simplicity Prevails: Rethinking Negative Preference Optimization for LLM Unlearning
by: Fan, Chongyu, et al.
Published: (2024)
by: Fan, Chongyu, et al.
Published: (2024)
Amortizing intractable inference in large language models
by: Hu, Edward J., et al.
Published: (2023)
by: Hu, Edward J., et al.
Published: (2023)
Mixture Density Networks for Classification with an Application to Product Bundling
by: Gugulothu, Narendhar, et al.
Published: (2024)
by: Gugulothu, Narendhar, et al.
Published: (2024)
Score-of-Mixture Training: Training One-Step Generative Models Made Simple via Score Estimation of Mixture Distributions
by: Jayashankar, Tejas, et al.
Published: (2025)
by: Jayashankar, Tejas, et al.
Published: (2025)
WorkArena: How Capable Are Web Agents at Solving Common Knowledge Work Tasks?
by: Drouin, Alexandre, et al.
Published: (2024)
by: Drouin, Alexandre, et al.
Published: (2024)
How Ensemble Learning Balances Accuracy and Overfitting: A Bias-Variance Perspective on Tabular Data
by: Mohammad, Zubair Ahmed
Published: (2025)
by: Mohammad, Zubair Ahmed
Published: (2025)
Differentiable Tree Search Network
by: Mittal, Dixant, et al.
Published: (2024)
by: Mittal, Dixant, et al.
Published: (2024)
Similar Items
-
Next-Token Prediction Should be Ambiguity-Sensitive: A Meta-Learning Perspective
by: Gagnon, Leo, et al.
Published: (2025) -
In-context learning and Occam's razor
by: Elmoznino, Eric, et al.
Published: (2024) -
Does learning the right latent variables necessarily improve in-context learning?
by: Mittal, Sarthak, et al.
Published: (2024) -
Beyond Distribution Sharpening: The Importance of Task Rewards
by: Mittal, Sarthak, et al.
Published: (2026) -
Evaluating Interventional Reasoning Capabilities of Large Language Models
by: Kasetty, Tejas, et al.
Published: (2024)