Cross-Entropy Is All You Need To Invert the Data Generating Process
Fuente:
arXiv
Saved in:
| Main Authors: | Reizinger, Patrik, Bizeul, Alice, Juhos, Attila, Vogt, Julia E., Balestriero, Randall, Brendel, Wieland, Klindt, David |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Position: An Empirically Grounded Identifiability Theory Will Accelerate Self-Supervised Learning Research
by: Reizinger, Patrik, et al.
Published: (2025)
by: Reizinger, Patrik, et al.
Published: (2025)
InfoNCE: Identifying the Gap Between Theory and Practice
by: Rusak, Evgenia, et al.
Published: (2024)
by: Rusak, Evgenia, et al.
Published: (2024)
Estimating Treatment Effects with Independent Component Analysis
by: Reizinger, Patrik, et al.
Published: (2025)
by: Reizinger, Patrik, et al.
Published: (2025)
Causality is Key for Interpretability Claims to Generalise
by: Joshi, Shruti, et al.
Published: (2026)
by: Joshi, Shruti, et al.
Published: (2026)
Identifiable Exchangeable Mechanisms for Causal Structure and Representation Learning
by: Reizinger, Patrik, et al.
Published: (2024)
by: Reizinger, Patrik, et al.
Published: (2024)
Occam's Razor for Self Supervised Learning: What is Sufficient to Learn Good Representations?
by: Ibrahim, Mark, et al.
Published: (2024)
by: Ibrahim, Mark, et al.
Published: (2024)
Skill Learning via Policy Diversity Yields Identifiable Representations for Reinforcement Learning
by: Reizinger, Patrik, et al.
Published: (2025)
by: Reizinger, Patrik, et al.
Published: (2025)
An Interventional Perspective on Identifiability in Gaussian LTI Systems with Independent Component Analysis
by: Rajendran, Goutham, et al.
Published: (2023)
by: Rajendran, Goutham, et al.
Published: (2023)
Rule Extrapolation in Language Models: A Study of Compositional Generalization on OOD Prompts
by: Mészáros, Anna, et al.
Published: (2024)
by: Mészáros, Anna, et al.
Published: (2024)
Provable Compositional Generalization for Object-Centric Learning
by: Wiedemer, Thaddäus, et al.
Published: (2023)
by: Wiedemer, Thaddäus, et al.
Published: (2023)
Position: Understanding LLMs Requires More Than Statistical Generalization
by: Reizinger, Patrik, et al.
Published: (2024)
by: Reizinger, Patrik, et al.
Published: (2024)
In Search of Forgotten Domain Generalization
by: Mayilvahanan, Prasanna, et al.
Published: (2024)
by: Mayilvahanan, Prasanna, et al.
Published: (2024)
Who Guards the Guardians? The Challenges of Evaluating Identifiability of Learned Representations
by: Joshi, Shruti, et al.
Published: (2026)
by: Joshi, Shruti, et al.
Published: (2026)
No Location Left Behind: Measuring and Improving the Fairness of Implicit Representations for Earth Data
by: Cai, Daniel, et al.
Published: (2025)
by: Cai, Daniel, et al.
Published: (2025)
DIET-CP: Lightweight and Data Efficient Self Supervised Continued Pretraining
by: Rodas, Bryan, et al.
Published: (2025)
by: Rodas, Bryan, et al.
Published: (2025)
From superposition to sparse codes: interpretable representations in neural networks
by: Klindt, David, et al.
Published: (2025)
by: Klindt, David, et al.
Published: (2025)
Task Priors: Enhancing Model Evaluation by Considering the Entire Space of Downstream Tasks
by: Patel, Niket, et al.
Published: (2025)
by: Patel, Niket, et al.
Published: (2025)
SAFE: A Novel Approach to AI Weather Evaluation through Stratified Assessments of Forecasts over Earth
by: Masi, Nick, et al.
Published: (2025)
by: Masi, Nick, et al.
Published: (2025)
ALLoRA: Adaptive Learning Rate Mitigates LoRA Fatal Flaws
by: Huang, Hai, et al.
Published: (2024)
by: Huang, Hai, et al.
Published: (2024)
Data Whitening Improves Sparse Autoencoder Learning
by: Saraswatula, Ashwin, et al.
Published: (2025)
by: Saraswatula, Ashwin, et al.
Published: (2025)
From Pixels to Components: Eigenvector Masking for Visual Representation Learning
by: Bizeul, Alice, et al.
Published: (2025)
by: Bizeul, Alice, et al.
Published: (2025)
Characterizing Large Language Model Geometry Helps Solve Toxicity Detection and Generation
by: Balestriero, Randall, et al.
Published: (2023)
by: Balestriero, Randall, et al.
Published: (2023)
A Probabilistic Model Behind Self-Supervised Learning
by: Bizeul, Alice, et al.
Published: (2024)
by: Bizeul, Alice, et al.
Published: (2024)
Fast and Exact Enumeration of Deep Networks Partitions Regions
by: Balestriero, Randall, et al.
Published: (2024)
by: Balestriero, Randall, et al.
Published: (2024)
Interpretable Diffusion Models with B-cos Networks
by: Bernold, Nicola, et al.
Published: (2025)
by: Bernold, Nicola, et al.
Published: (2025)
Attention is All You Need Until You Need Retention
by: Yaslioglu, M. Murat
Published: (2025)
by: Yaslioglu, M. Murat
Published: (2025)
Transduction is All You Need for Structured Data Workflows
by: Gliozzo, Alfio, et al.
Published: (2025)
by: Gliozzo, Alfio, et al.
Published: (2025)
LeJEPA: Provable and Scalable Self-Supervised Learning Without the Heuristics
by: Balestriero, Randall, et al.
Published: (2025)
by: Balestriero, Randall, et al.
Published: (2025)
Learning by Reconstruction Produces Uninformative Features For Perception
by: Balestriero, Randall, et al.
Published: (2024)
by: Balestriero, Randall, et al.
Published: (2024)
LLMs on the Line: Data Determines Loss-to-Loss Scaling Laws
by: Mayilvahanan, Prasanna, et al.
Published: (2025)
by: Mayilvahanan, Prasanna, et al.
Published: (2025)
Contrast Is All You Need
by: Kilic, Burak, et al.
Published: (2023)
by: Kilic, Burak, et al.
Published: (2023)
Context is All You Need
by: Delanois, Jean Erik, et al.
Published: (2026)
by: Delanois, Jean Erik, et al.
Published: (2026)
FineVision: Open Data Is All You Need
by: Wiedmann, Luis, et al.
Published: (2025)
by: Wiedmann, Luis, et al.
Published: (2025)
Common Sense Is All You Need
by: Latapie, Hugo
Published: (2025)
by: Latapie, Hugo
Published: (2025)
The Fair Language Model Paradox
by: Pinto, Andrea, et al.
Published: (2024)
by: Pinto, Andrea, et al.
Published: (2024)
[MASK] is All You Need
by: Hu, Vincent Tao, et al.
Published: (2024)
by: Hu, Vincent Tao, et al.
Published: (2024)
Information Gain Is Not All You Need
by: Ericson, Ludvig, et al.
Published: (2025)
by: Ericson, Ludvig, et al.
Published: (2025)
Cooperation Is All You Need
by: Adeel, Ahsan, et al.
Published: (2023)
by: Adeel, Ahsan, et al.
Published: (2023)
Does CLIP's Generalization Performance Mainly Stem from High Train-Test Similarity?
by: Mayilvahanan, Prasanna, et al.
Published: (2023)
by: Mayilvahanan, Prasanna, et al.
Published: (2023)
Exploitation Is All You Need... for Exploration
by: Rentschler, Micah, et al.
Published: (2025)
by: Rentschler, Micah, et al.
Published: (2025)
Similar Items
-
Position: An Empirically Grounded Identifiability Theory Will Accelerate Self-Supervised Learning Research
by: Reizinger, Patrik, et al.
Published: (2025) -
InfoNCE: Identifying the Gap Between Theory and Practice
by: Rusak, Evgenia, et al.
Published: (2024) -
Estimating Treatment Effects with Independent Component Analysis
by: Reizinger, Patrik, et al.
Published: (2025) -
Causality is Key for Interpretability Claims to Generalise
by: Joshi, Shruti, et al.
Published: (2026) -
Identifiable Exchangeable Mechanisms for Causal Structure and Representation Learning
by: Reizinger, Patrik, et al.
Published: (2024)