Supervised sparse auto-encoders for interpretable and compositional representations
Fuente:
arXiv
Saved in:
| Main Authors: | Harzli, Ouns El, Wallner, Hugo, Nam, Yoonsoo, Tao, Haixuan Xavier |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sufficient Conditions for Stability of Minimum-Norm Interpolating Deep ReLU Networks
by: Harzli, Ouns El, et al.
Published: (2026)
by: Harzli, Ouns El, et al.
Published: (2026)
From Neural Networks to Logical Theories: The Correspondence between Fibring Modal Logics and Fibring Neural Networks
by: Harzli, Ouns El, et al.
Published: (2025)
by: Harzli, Ouns El, et al.
Published: (2025)
L-MAE: Longitudinal masked auto-encoder with time and severity-aware encoding for diabetic retinopathy progression prediction
by: Zeghlache, Rachid, et al.
Published: (2024)
by: Zeghlache, Rachid, et al.
Published: (2024)
Unsupervised decoding of encoded reasoning using language model interpretability
by: Fang, Ching, et al.
Published: (2025)
by: Fang, Ching, et al.
Published: (2025)
Weight-sparse transformers have interpretable circuits
by: Gao, Leo, et al.
Published: (2025)
by: Gao, Leo, et al.
Published: (2025)
Exploiting the equivalence between quantum neural networks and perceptrons
by: Mingard, Chris, et al.
Published: (2024)
by: Mingard, Chris, et al.
Published: (2024)
Decoupling Dynamical Richness from Representation Learning: Towards Practical Measurement
by: Nam, Yoonsoo, et al.
Published: (2024)
by: Nam, Yoonsoo, et al.
Published: (2024)
Discrete, compositional, and symbolic representations through attractor dynamics
by: Nam, Andrew, et al.
Published: (2023)
by: Nam, Andrew, et al.
Published: (2023)
Can sparse autoencoders be used to decompose and interpret steering vectors?
by: Mayne, Harry, et al.
Published: (2024)
by: Mayne, Harry, et al.
Published: (2024)
A representational framework for learning and encoding structurally enriched trajectories in complex agent environments
by: Catarau-Cotutiu, Corina, et al.
Published: (2025)
by: Catarau-Cotutiu, Corina, et al.
Published: (2025)
Advancing Algorithmic Approaches to Probabilistic Argumentation under the Constellation Approach
by: Popescu, Andrei, et al.
Published: (2024)
by: Popescu, Andrei, et al.
Published: (2024)
Exploring bat song syllable representations in self-supervised audio encoders
by: Kloots, Marianne de Heer, et al.
Published: (2024)
by: Kloots, Marianne de Heer, et al.
Published: (2024)
Can LLMs interpret figurative language as humans do?: surface-level vs representational similarity
by: Bollepally, Samhita, et al.
Published: (2026)
by: Bollepally, Samhita, et al.
Published: (2026)
Granular-ball computing: an efficient, robust, and interpretable adaptive multi-granularity representation and computation method
by: Xia, Shuyin, et al.
Published: (2023)
by: Xia, Shuyin, et al.
Published: (2023)
Implicit neural representations as a coordinate-based framework for continuous environmental field reconstruction from sparse ecological observations
by: Pregowska, Agnieszka, et al.
Published: (2026)
by: Pregowska, Agnieszka, et al.
Published: (2026)
An interpretable unsupervised representation learning for high precision measurement in particle physics
by: Lv, Xing-Jian, et al.
Published: (2025)
by: Lv, Xing-Jian, et al.
Published: (2025)
Heterogeneous network and graph attention auto-encoder for LncRNA-disease association prediction
by: Liu, Jin-Xing, et al.
Published: (2024)
by: Liu, Jin-Xing, et al.
Published: (2024)
A new approach for encoding code and assisting code understanding
by: Fan, Mengdan, et al.
Published: (2024)
by: Fan, Mengdan, et al.
Published: (2024)
Scaling and evaluating sparse autoencoders
by: Gao, Leo, et al.
Published: (2024)
by: Gao, Leo, et al.
Published: (2024)
The role of positional encodings in the ARC benchmark
by: Costa, Guilherme H. Bandeira, et al.
Published: (2025)
by: Costa, Guilherme H. Bandeira, et al.
Published: (2025)
Refining Transcripts With TV Subtitles by Prompt-Based Weakly Supervised Training of ASR
by: Zhao, Xinnian, et al.
Published: (2025)
by: Zhao, Xinnian, et al.
Published: (2025)
Instantiations and Computational Aspects of Non-Flat Assumption-based Argumentation
by: Lehtonen, Tuomo, et al.
Published: (2024)
by: Lehtonen, Tuomo, et al.
Published: (2024)
Physically-Grounded Goal Imagination: Physics-Informed Variational Autoencoder for Self-Supervised Reinforcement Learning
by: Nguyen, Lan Thi Ha, et al.
Published: (2025)
by: Nguyen, Lan Thi Ha, et al.
Published: (2025)
Explanations that reveal all through the definition of encoding
by: Puli, Aahlad, et al.
Published: (2024)
by: Puli, Aahlad, et al.
Published: (2024)
Beyond Generation: Unlocking Universal Editing via Self-Supervised Fine-Tuning
by: Chen, Harold Haodong, et al.
Published: (2024)
by: Chen, Harold Haodong, et al.
Published: (2024)
Building spatial world models from sparse transitional episodic memories
by: He, Zizhan, et al.
Published: (2025)
by: He, Zizhan, et al.
Published: (2025)
Timer-S1: A Billion-Scale Time Series Foundation Model with Serial Scaling
by: Liu, Yong, et al.
Published: (2026)
by: Liu, Yong, et al.
Published: (2026)
Quantum feature encoding optimization
by: Fioravanti, Tommaso, et al.
Published: (2025)
by: Fioravanti, Tommaso, et al.
Published: (2025)
Positional encoding is not the same as context: A study on positional encoding for sequential recommendation
by: Lopez-Avila, Alejo, et al.
Published: (2024)
by: Lopez-Avila, Alejo, et al.
Published: (2024)
DiffDub: Person-generic Visual Dubbing Using Inpainting Renderer with Diffusion Auto-encoder
by: Liu, Tao, et al.
Published: (2023)
by: Liu, Tao, et al.
Published: (2023)
LaTiM: Longitudinal representation learning in continuous-time models to predict disease progression
by: Zeghlache, Rachid, et al.
Published: (2024)
by: Zeghlache, Rachid, et al.
Published: (2024)
S-SONDO: Self-Supervised Knowledge Distillation for General Audio Foundation Models
by: Adlouni, Mohammed Ali El, et al.
Published: (2026)
by: Adlouni, Mohammed Ali El, et al.
Published: (2026)
An Explainable Gaussian Process Auto-encoder for Tabular Data
by: Zhang, Wei, et al.
Published: (2025)
by: Zhang, Wei, et al.
Published: (2025)
Scaling laws for language encoding models in fMRI
by: Antonello, Richard, et al.
Published: (2023)
by: Antonello, Richard, et al.
Published: (2023)
Auto-encoding Molecules: Graph-Matching Capabilities Matter
by: Cunow, Magnus, et al.
Published: (2025)
by: Cunow, Magnus, et al.
Published: (2025)
Benchmarking data encoding methods in Quantum Machine Learning
by: Zang, Orlane, et al.
Published: (2025)
by: Zang, Orlane, et al.
Published: (2025)
Do traveling waves make good positional encodings?
by: van de Geijn, Chase, et al.
Published: (2025)
by: van de Geijn, Chase, et al.
Published: (2025)
Self-supervised network distillation: an effective approach to exploration in sparse reward environments
by: Pecháč, Matej, et al.
Published: (2023)
by: Pecháč, Matej, et al.
Published: (2023)
The quest for the GRAph Level autoEncoder (GRALE)
by: Krzakala, Paul, et al.
Published: (2025)
by: Krzakala, Paul, et al.
Published: (2025)
PIP: Positional-encoding Image Prior
by: Shabtay, Nimrod, et al.
Published: (2022)
by: Shabtay, Nimrod, et al.
Published: (2022)
Similar Items
-
Sufficient Conditions for Stability of Minimum-Norm Interpolating Deep ReLU Networks
by: Harzli, Ouns El, et al.
Published: (2026) -
From Neural Networks to Logical Theories: The Correspondence between Fibring Modal Logics and Fibring Neural Networks
by: Harzli, Ouns El, et al.
Published: (2025) -
L-MAE: Longitudinal masked auto-encoder with time and severity-aware encoding for diabetic retinopathy progression prediction
by: Zeghlache, Rachid, et al.
Published: (2024) -
Unsupervised decoding of encoded reasoning using language model interpretability
by: Fang, Ching, et al.
Published: (2025) -
Weight-sparse transformers have interpretable circuits
by: Gao, Leo, et al.
Published: (2025)