Compositionality Unlocks Deep Interpretable Models
Fuente:
arXiv
Saved in:
| Main Authors: | Dooms, Thomas, Gauderis, Ward, Wiggins, Geraint A., Oramas, Jose |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bilinear autoencoders find interpretable manifolds
by: Dooms, Thomas, et al.
Published: (2026)
by: Dooms, Thomas, et al.
Published: (2026)
From Mechanistic to Compositional Interpretability
by: Gauderis, Ward, et al.
Published: (2026)
by: Gauderis, Ward, et al.
Published: (2026)
Finding Manifolds With Bilinear Autoencoders
by: Dooms, Thomas, et al.
Published: (2025)
by: Dooms, Thomas, et al.
Published: (2025)
Bilinear MLPs enable weight-based mechanistic interpretability
by: Pearce, Michael T., et al.
Published: (2024)
by: Pearce, Michael T., et al.
Published: (2024)
Quantum Methods for Managing Ambiguity in Natural Language Processing
by: Eisinger, Jurek, et al.
Published: (2025)
by: Eisinger, Jurek, et al.
Published: (2025)
Deep Model Interpretation with Limited Data : A Coreset-based Approach
by: Behzadi-Khormouji, Hamed, et al.
Published: (2024)
by: Behzadi-Khormouji, Hamed, et al.
Published: (2024)
Tokenized SAEs: Disentangling SAE Reconstructions
by: Dooms, Thomas, et al.
Published: (2025)
by: Dooms, Thomas, et al.
Published: (2025)
BioOSS: A Bio-Inspired Oscillatory State System with Spatio-Temporal Dynamics
by: Yuan, Zhongju, et al.
Published: (2025)
by: Yuan, Zhongju, et al.
Published: (2025)
Towards a Formal Creativity Theory: Preliminary results in Novelty and Transformativeness
by: Santo, Luís Espírito, et al.
Published: (2024)
by: Santo, Luís Espírito, et al.
Published: (2024)
A novel Reservoir Architecture for Periodic Time Series Prediction
by: Yuan, Zhongju, et al.
Published: (2024)
by: Yuan, Zhongju, et al.
Published: (2024)
Fractals made Practical: Denoising Diffusion as Partitioned Iterated Function Systems
by: Dooms, Ann
Published: (2026)
by: Dooms, Ann
Published: (2026)
Weight-based Decomposition: A Case for Bilinear MLPs
by: Pearce, Michael T., et al.
Published: (2024)
by: Pearce, Michael T., et al.
Published: (2024)
When Are Two Networks the Same? Tensor Similarity for Mechanistic Interpretability
by: Gonzalez, ML Nissen, et al.
Published: (2026)
by: Gonzalez, ML Nissen, et al.
Published: (2026)
Smooth InfoMax -- Towards Easier Post-Hoc Interpretability
by: Denoodt, Fabian, et al.
Published: (2024)
by: Denoodt, Fabian, et al.
Published: (2024)
Efficient Post-Hoc Uncertainty Calibration via Variance-Based Smoothing
by: Denoodt, Fabian, et al.
Published: (2025)
by: Denoodt, Fabian, et al.
Published: (2025)
Towards the Characterization of Representations Learned via Capsule-based Network Architectures
by: Tawalbeh, Saja, et al.
Published: (2023)
by: Tawalbeh, Saja, et al.
Published: (2023)
Recursive Self-Aggregation Unlocks Deep Thinking in Large Language Models
by: Venkatraman, Siddarth, et al.
Published: (2025)
by: Venkatraman, Siddarth, et al.
Published: (2025)
KnowIt: Deep Time Series Modeling and Interpretation
by: Theunissen, M. W., et al.
Published: (2025)
by: Theunissen, M. W., et al.
Published: (2025)
LHT: Statistically-Driven Oblique Decision Trees for Interpretable Classification
by: Li, Hongyi, et al.
Published: (2025)
by: Li, Hongyi, et al.
Published: (2025)
Topological Interpretability for Deep-Learning
by: Spannaus, Adam, et al.
Published: (2023)
by: Spannaus, Adam, et al.
Published: (2023)
Learning Hyperplane Tree: A Piecewise Linear and Fully Interpretable Decision-making Framework
by: Li, Hongyi, et al.
Published: (2025)
by: Li, Hongyi, et al.
Published: (2025)
Rank-1 LoRAs Encode Interpretable Reasoning Signals
by: Ward, Jake, et al.
Published: (2025)
by: Ward, Jake, et al.
Published: (2025)
Sparse Deep Additive Model with Interactions: Enhancing Interpretability and Predictability
by: Hung, Yi-Ting, et al.
Published: (2025)
by: Hung, Yi-Ting, et al.
Published: (2025)
Interpretability in Deep Time Series Models Demands Semantic Alignment
by: De Felice, Giovanni, et al.
Published: (2026)
by: De Felice, Giovanni, et al.
Published: (2026)
Unlocking Historical Clinical Trial Data with ALIGN: A Compositional Large Language Model System for Medical Coding
by: Seedat, Nabeel, et al.
Published: (2024)
by: Seedat, Nabeel, et al.
Published: (2024)
Unlocking the Black Box of Latent Reasoning: An Interpretability-Guided Approach to Intervention
by: Chang, Shuochen, et al.
Published: (2026)
by: Chang, Shuochen, et al.
Published: (2026)
A Unified Theory of Compositionality, Modularity, and Interpretability in Markov Decision Processes
by: Ringstrom, Thomas J., et al.
Published: (2025)
by: Ringstrom, Thomas J., et al.
Published: (2025)
Transcoder Adapters for Reasoning-Model Diffing
by: Hu, Nathan, et al.
Published: (2026)
by: Hu, Nathan, et al.
Published: (2026)
Hybrid Physics and Deep Learning Model for Interpretable Vehicle State Prediction
by: Baier, Alexandra, et al.
Published: (2021)
by: Baier, Alexandra, et al.
Published: (2021)
Compositional Function Networks: A High-Performance Alternative to Deep Neural Networks with Built-in Interpretability
by: Li, Fang
Published: (2025)
by: Li, Fang
Published: (2025)
Interpretable Deep Clustering for Tabular Data
by: Svirsky, Jonathan, et al.
Published: (2023)
by: Svirsky, Jonathan, et al.
Published: (2023)
Unlocking Interpretability for RF Sensing: A Complex-Valued White-Box Transformer
by: Zhang, Xie, et al.
Published: (2025)
by: Zhang, Xie, et al.
Published: (2025)
Unlocking Biomedical Insights: Hierarchical Attention Networks for High-Dimensional Data Interpretation
by: Nair, Rekha R, et al.
Published: (2025)
by: Nair, Rekha R, et al.
Published: (2025)
An Interpretable Deep Learning Model for General Insurance Pricing
by: Laub, Patrick J., et al.
Published: (2025)
by: Laub, Patrick J., et al.
Published: (2025)
Mixtures of Experts Unlock Parameter Scaling for Deep RL
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
Higher Resolution, Better Generalization: Unlocking Visual Scaling in Deep Reinforcement Learning
by: Trumpp, Raphael, et al.
Published: (2026)
by: Trumpp, Raphael, et al.
Published: (2026)
Interpreting Neurons in Deep Vision Networks with Language Models
by: Bai, Nicholas, et al.
Published: (2024)
by: Bai, Nicholas, et al.
Published: (2024)
Interpreting Deep Neural Networks with the Package innsight
by: Koenen, Niklas, et al.
Published: (2023)
by: Koenen, Niklas, et al.
Published: (2023)
Interpretable Maximum Margin Deep Anomaly Detection
by: Yang, Zhiji, et al.
Published: (2026)
by: Yang, Zhiji, et al.
Published: (2026)
Unlocking Compositional Generalization in Continual Few-Shot Learning
by: Nguyen-Lam, Phu-Quy, et al.
Published: (2026)
by: Nguyen-Lam, Phu-Quy, et al.
Published: (2026)
Similar Items
-
Bilinear autoencoders find interpretable manifolds
by: Dooms, Thomas, et al.
Published: (2026) -
From Mechanistic to Compositional Interpretability
by: Gauderis, Ward, et al.
Published: (2026) -
Finding Manifolds With Bilinear Autoencoders
by: Dooms, Thomas, et al.
Published: (2025) -
Bilinear MLPs enable weight-based mechanistic interpretability
by: Pearce, Michael T., et al.
Published: (2024) -
Quantum Methods for Managing Ambiguity in Natural Language Processing
by: Eisinger, Jurek, et al.
Published: (2025)