Saved in:
| Main Authors: | Winnicki, John, Gnanasekaran, Abeynaya, Darve, Eric |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.23829 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient Variational Quantum Linear Solver for Structured Sparse Matrices
by: Gnanasekaran, Abeynaya, et al.
Published: (2024)
by: Gnanasekaran, Abeynaya, et al.
Published: (2024)
Efficient Quantum Access Model for Sparse Structured Matrices using Linear Combination of Things
by: Gnanasekaran, Abeynaya, et al.
Published: (2025)
by: Gnanasekaran, Abeynaya, et al.
Published: (2025)
Discovering Algorithms with Computational Language Processing
by: Bourdais, Theo, et al.
Published: (2025)
by: Bourdais, Theo, et al.
Published: (2025)
On the Emergence of Ergodic Dynamics in Unique Games
by: Sahai, Tuhin, et al.
Published: (2024)
by: Sahai, Tuhin, et al.
Published: (2024)
Variational Quantum Framework for Partial Differential Equation Constrained Optimization
by: Surana, Amit, et al.
Published: (2024)
by: Surana, Amit, et al.
Published: (2024)
Measuring Sparse Autoencoder Feature Sensitivity
by: Tian, Claire, et al.
Published: (2025)
by: Tian, Claire, et al.
Published: (2025)
The Geometry of Concepts: Sparse Autoencoder Feature Structure
by: Li, Yuxiao, et al.
Published: (2024)
by: Li, Yuxiao, et al.
Published: (2024)
From Token Lists to Graph Motifs: Weisfeiler-Lehman Analysis of Sparse Autoencoder Features
by: Fernandez-Boullon, Ruben, et al.
Published: (2026)
by: Fernandez-Boullon, Ruben, et al.
Published: (2026)
Sparse Autoencoder Features for Classifications and Transferability
by: Gallifant, Jack, et al.
Published: (2025)
by: Gallifant, Jack, et al.
Published: (2025)
Variational Quantum Framework for Nonlinear PDE Constrained Optimization Using Carleman Linearization
by: Gnanasekaran, Abeynaya, et al.
Published: (2024)
by: Gnanasekaran, Abeynaya, et al.
Published: (2024)
Adaptive Sparse Allocation with Mutual Choice & Feature Choice Sparse Autoencoders
by: Ayonrinde, Kola
Published: (2024)
by: Ayonrinde, Kola
Published: (2024)
Learning Multi-Level Features with Matryoshka Sparse Autoencoders
by: Bussmann, Bart, et al.
Published: (2025)
by: Bussmann, Bart, et al.
Published: (2025)
Feature Starvation as Geometric Instability in Sparse Autoencoders
by: Chaudhry, Faris, et al.
Published: (2026)
by: Chaudhry, Faris, et al.
Published: (2026)
Causal Interpretation of Sparse Autoencoder Features in Vision
by: Han, Sangyu, et al.
Published: (2025)
by: Han, Sangyu, et al.
Published: (2025)
Feature Hedging: Correlated Features Break Narrow Sparse Autoencoders
by: Chanin, David, et al.
Published: (2025)
by: Chanin, David, et al.
Published: (2025)
A is for Absorption: Studying Feature Splitting and Absorption in Sparse Autoencoders
by: Chanin, David, et al.
Published: (2024)
by: Chanin, David, et al.
Published: (2024)
Graph-Regularized Sparse Autoencoders for LLM Safety Steering
by: Yeon, Jehyeok, et al.
Published: (2025)
by: Yeon, Jehyeok, et al.
Published: (2025)
Improving Steering Vectors by Targeting Sparse Autoencoder Features
by: Chalnev, Sviatoslav, et al.
Published: (2024)
by: Chalnev, Sviatoslav, et al.
Published: (2024)
Sparse but Wrong: Incorrect L0 Leads to Incorrect Features in Sparse Autoencoders
by: Chanin, David, et al.
Published: (2025)
by: Chanin, David, et al.
Published: (2025)
Semantic Optimal Transport for Sparse Autoencoder Feature Matching and Circuit Compression
by: Cao, Tue M., et al.
Published: (2026)
by: Cao, Tue M., et al.
Published: (2026)
AbsTopK: Rethinking Sparse Autoencoders For Bidirectional Features
by: Zhu, Xudong, et al.
Published: (2025)
by: Zhu, Xudong, et al.
Published: (2025)
From Atoms to Trees: Building a Structured Feature Forest with Hierarchical Sparse Autoencoders
by: Luo, Yifan, et al.
Published: (2026)
by: Luo, Yifan, et al.
Published: (2026)
Empowering GraphRAG with Knowledge Filtering and Integration
by: Guo, Kai, et al.
Published: (2025)
by: Guo, Kai, et al.
Published: (2025)
Sparse Autoencoders, Again?
by: Lu, Yin, et al.
Published: (2025)
by: Lu, Yin, et al.
Published: (2025)
Rethinking Sparse Autoencoders: Select-and-Project for Fairness and Control from Encoder Features Alone
by: Bărbălau, Antonio, et al.
Published: (2025)
by: Bărbălau, Antonio, et al.
Published: (2025)
Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models
by: Pach, Mateusz, et al.
Published: (2025)
by: Pach, Mateusz, et al.
Published: (2025)
Are Sparse Autoencoder Benchmarks Reliable?
by: Chanin, David
Published: (2026)
by: Chanin, David
Published: (2026)
Constrain Alignment with Sparse Autoencoders
by: Yin, Qingyu, et al.
Published: (2024)
by: Yin, Qingyu, et al.
Published: (2024)
LLMs are Overconfident: Evaluating Confidence Interval Calibration with FermiEval
by: Epstein, Elliot L., et al.
Published: (2025)
by: Epstein, Elliot L., et al.
Published: (2025)
Towards Interpretable Protein Structure Prediction with Sparse Autoencoders
by: Parsan, Nithin, et al.
Published: (2025)
by: Parsan, Nithin, et al.
Published: (2025)
Allocate Marginal Reviews to Borderline Papers Using LLM Comparative Ranking
by: Epstein, Elliot L., et al.
Published: (2026)
by: Epstein, Elliot L., et al.
Published: (2026)
Mechanistic Knobs in LLMs: Retrieving and Steering High-Order Semantic Features via Sparse Autoencoders
by: Zhang, Ruikang, et al.
Published: (2026)
by: Zhang, Ruikang, et al.
Published: (2026)
Improving Sparse Autoencoder with Dynamic Attention
by: Wang, Dongsheng, et al.
Published: (2026)
by: Wang, Dongsheng, et al.
Published: (2026)
BatchTopK Sparse Autoencoders
by: Bussmann, Bart, et al.
Published: (2024)
by: Bussmann, Bart, et al.
Published: (2024)
Sparse Feature Circuits: Discovering and Editing Interpretable Causal Graphs in Language Models
by: Marks, Samuel, et al.
Published: (2024)
by: Marks, Samuel, et al.
Published: (2024)
Are Sparse Autoencoders Useful? A Case Study in Sparse Probing
by: Kantamneni, Subhash, et al.
Published: (2025)
by: Kantamneni, Subhash, et al.
Published: (2025)
SparseRM: A Lightweight Preference Modeling with Sparse Autoencoder
by: Liu, Dengcan, et al.
Published: (2025)
by: Liu, Dengcan, et al.
Published: (2025)
Taming Polysemanticity in LLMs: Provable Feature Recovery via Sparse Autoencoders
by: Chen, Siyu, et al.
Published: (2025)
by: Chen, Siyu, et al.
Published: (2025)
Sparse Autoencoders for Hypothesis Generation
by: Movva, Rajiv, et al.
Published: (2025)
by: Movva, Rajiv, et al.
Published: (2025)
Quantifying Feature Space Universality Across Large Language Models via Sparse Autoencoders
by: Lan, Michael, et al.
Published: (2024)
by: Lan, Michael, et al.
Published: (2024)
Similar Items
-
Efficient Variational Quantum Linear Solver for Structured Sparse Matrices
by: Gnanasekaran, Abeynaya, et al.
Published: (2024) -
Efficient Quantum Access Model for Sparse Structured Matrices using Linear Combination of Things
by: Gnanasekaran, Abeynaya, et al.
Published: (2025) -
Discovering Algorithms with Computational Language Processing
by: Bourdais, Theo, et al.
Published: (2025) -
On the Emergence of Ergodic Dynamics in Unique Games
by: Sahai, Tuhin, et al.
Published: (2024) -
Variational Quantum Framework for Partial Differential Equation Constrained Optimization
by: Surana, Amit, et al.
Published: (2024)