Toward Identifiable Sparse Autoencoders
Fuente:
arXiv
Saved in:
| Main Authors: | Nelson, Walter, Karaletsos, Theofanis, Locatello, Francesco |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Statistical and structural identifiability in representation learning
by: Nelson, Walter, et al.
Published: (2026)
by: Nelson, Walter, et al.
Published: (2026)
Modelling Cellular Perturbations with the Sparse Additive Mechanism Shift Variational Autoencoder
by: Bereket, Michael, et al.
Published: (2023)
by: Bereket, Michael, et al.
Published: (2023)
MorphGen: Controllable and Morphologically Plausible Generative Cell-Imaging
by: Demirel, Berker, et al.
Published: (2025)
by: Demirel, Berker, et al.
Published: (2025)
Adjusting Pretrained Backbones for Performativity
by: Demirel, Berker, et al.
Published: (2024)
by: Demirel, Berker, et al.
Published: (2024)
Compositional Deep Probabilistic Models of DNA Encoded Libraries
by: Chen, Benson, et al.
Published: (2023)
by: Chen, Benson, et al.
Published: (2023)
Learning Explicit Single-Cell Dynamics Using ODE Representations
by: von Bassewitz, Jan-Philipp, et al.
Published: (2025)
by: von Bassewitz, Jan-Philipp, et al.
Published: (2025)
Channel Vision Transformers: An Image Is Worth 1 x 16 x 16 Words
by: Bao, Yujia, et al.
Published: (2023)
by: Bao, Yujia, et al.
Published: (2023)
Causal Learning with the Invariance Principle
by: Montagna, Francesco, et al.
Published: (2026)
by: Montagna, Francesco, et al.
Published: (2026)
Controlling Transient Amplification Improves Long-horizon Rollouts
by: Pervez, Adeel, et al.
Published: (2026)
by: Pervez, Adeel, et al.
Published: (2026)
Calibrated Test-Time Guidance for Bayesian Inference
by: Geyfman, Daniel, et al.
Published: (2026)
by: Geyfman, Daniel, et al.
Published: (2026)
Parallel Token Prediction for Language Models
by: Draxler, Felix, et al.
Published: (2025)
by: Draxler, Felix, et al.
Published: (2025)
The Rate-Distortion-Polysemanticity Tradeoff in SAEs
by: Mencattini, Tommaso, et al.
Published: (2026)
by: Mencattini, Tommaso, et al.
Published: (2026)
Variational Control for Guidance in Diffusion Models
by: Pandey, Kushagra, et al.
Published: (2025)
by: Pandey, Kushagra, et al.
Published: (2025)
Identifying General Mechanism Shifts in Linear Causal Representations
by: Chen, Tianyu, et al.
Published: (2024)
by: Chen, Tianyu, et al.
Published: (2024)
Do Sparse Autoencoders Identify Reasoning Features in Language Models?
by: Ma, George, et al.
Published: (2026)
by: Ma, George, et al.
Published: (2026)
Addressing Instrument-Outcome Confounding in Mendelian Randomization through Representation Learning
by: Huang, Shimeng, et al.
Published: (2026)
by: Huang, Shimeng, et al.
Published: (2026)
Learning Discrete Diffusion of Graphs via Free-Energy Gradient Flows
by: Rancati, Dario, et al.
Published: (2026)
by: Rancati, Dario, et al.
Published: (2026)
Marrying Causal Representation Learning with Dynamical Systems for Science
by: Yao, Dingling, et al.
Published: (2024)
by: Yao, Dingling, et al.
Published: (2024)
Mechanistic PDE Networks for Discovery of Governing Equations
by: Pervez, Adeel, et al.
Published: (2025)
by: Pervez, Adeel, et al.
Published: (2025)
Towards Principled Evaluations of Sparse Autoencoders for Interpretability and Control
by: Makelov, Aleksandar, et al.
Published: (2024)
by: Makelov, Aleksandar, et al.
Published: (2024)
Exploratory Causal Inference in SAEnce
by: Mencattini, Tommaso, et al.
Published: (2025)
by: Mencattini, Tommaso, et al.
Published: (2025)
Demystifying amortized causal discovery with transformers
by: Montagna, Francesco, et al.
Published: (2024)
by: Montagna, Francesco, et al.
Published: (2024)
How Well Do LLMs Understand Drug Mechanisms? A Knowledge + Reasoning Evaluation Dataset
by: Mohan, Sunil, et al.
Published: (2025)
by: Mohan, Sunil, et al.
Published: (2025)
Identifiable Object-Centric Representation Learning via Probabilistic Slot Attention
by: Kori, Avinash, et al.
Published: (2024)
by: Kori, Avinash, et al.
Published: (2024)
Ensembling Sparse Autoencoders
by: Gadgil, Soham, et al.
Published: (2025)
by: Gadgil, Soham, et al.
Published: (2025)
Towards a holistic understanding of Selection Bias for Causal Effect Identification
by: Qiu, Yiwen, et al.
Published: (2026)
by: Qiu, Yiwen, et al.
Published: (2026)
Preference Instability in Reward Models: Detection and Mitigation via Sparse Autoencoders
by: Liu, Shunchang, et al.
Published: (2026)
by: Liu, Shunchang, et al.
Published: (2026)
High-dimensional Analysis of Synthetic Data Selection
by: Rezaei, Parham, et al.
Published: (2025)
by: Rezaei, Parham, et al.
Published: (2025)
Navigating the Latent Space Dynamics of Neural Models
by: Fumero, Marco, et al.
Published: (2025)
by: Fumero, Marco, et al.
Published: (2025)
Towards Understanding the Robustness of Sparse Autoencoders
by: Saiyed, Ahson, et al.
Published: (2026)
by: Saiyed, Ahson, et al.
Published: (2026)
Beyond Input Activations: Identifying Influential Latents by Gradient Sparse Autoencoders
by: Shu, Dong, et al.
Published: (2025)
by: Shu, Dong, et al.
Published: (2025)
Scalable Single-Cell Gene Expression Generation with Latent Diffusion Models
by: Palla, Giovanni, et al.
Published: (2025)
by: Palla, Giovanni, et al.
Published: (2025)
Towards Interpretable Protein Structure Prediction with Sparse Autoencoders
by: Parsan, Nithin, et al.
Published: (2025)
by: Parsan, Nithin, et al.
Published: (2025)
Out-of-Distribution Detection with Relative Angles
by: Demirel, Berker, et al.
Published: (2024)
by: Demirel, Berker, et al.
Published: (2024)
Re-envisioning Euclid Galaxy Morphology: Identifying and Interpreting Features with Sparse Autoencoders
by: Wu, John F., et al.
Published: (2025)
by: Wu, John F., et al.
Published: (2025)
Analysis of Variational Sparse Autoencoders
by: Baker, Zachary, et al.
Published: (2025)
by: Baker, Zachary, et al.
Published: (2025)
Mechanistic Neural Networks for Scientific Machine Learning
by: Pervez, Adeel, et al.
Published: (2024)
by: Pervez, Adeel, et al.
Published: (2024)
Sparse Autoencoders, Again?
by: Lu, Yin, et al.
Published: (2025)
by: Lu, Yin, et al.
Published: (2025)
Sparse Shift Autoencoders for Identifying Concepts from Large Language Model Activations
by: Joshi, Shruti, et al.
Published: (2025)
by: Joshi, Shruti, et al.
Published: (2025)
Decomposing The Dark Matter of Sparse Autoencoders
by: Engels, Joshua, et al.
Published: (2024)
by: Engels, Joshua, et al.
Published: (2024)
Similar Items
-
Statistical and structural identifiability in representation learning
by: Nelson, Walter, et al.
Published: (2026) -
Modelling Cellular Perturbations with the Sparse Additive Mechanism Shift Variational Autoencoder
by: Bereket, Michael, et al.
Published: (2023) -
MorphGen: Controllable and Morphologically Plausible Generative Cell-Imaging
by: Demirel, Berker, et al.
Published: (2025) -
Adjusting Pretrained Backbones for Performativity
by: Demirel, Berker, et al.
Published: (2024) -
Compositional Deep Probabilistic Models of DNA Encoded Libraries
by: Chen, Benson, et al.
Published: (2023)