Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual Steering
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chatzoudis, Gerasimos, Li, Zhuowei, Moran, Gemma E., Wang, Hao, Metaxas, Dimitris N. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Can Cross-Layer Transcoders Replace Vision Transformer Activations? An Interpretable Perspective on Vision
von: Chatzoudis, Gerasimos, et al.
Veröffentlicht: (2026)
von: Chatzoudis, Gerasimos, et al.
Veröffentlicht: (2026)
LUCID-SAE: Learning Unified Vision-Language Sparse Codes for Interpretable Concept Discovery
von: Gu, Difei, et al.
Veröffentlicht: (2026)
von: Gu, Difei, et al.
Veröffentlicht: (2026)
The Hidden Life of Tokens: Reducing Hallucination of Large Vision-Language Models via Visual Information Steering
von: Li, Zhuowei, et al.
Veröffentlicht: (2025)
von: Li, Zhuowei, et al.
Veröffentlicht: (2025)
Test-Time Spectrum-Aware Latent Steering for Zero-Shot Generalization in Vision-Language Models
von: Dafnis, Konstantinos M., et al.
Veröffentlicht: (2025)
von: Dafnis, Konstantinos M., et al.
Veröffentlicht: (2025)
Training Like a Medical Resident: Context-Prior Learning Toward Universal Medical Image Segmentation
von: Gao, Yunhe, et al.
Veröffentlicht: (2023)
von: Gao, Yunhe, et al.
Veröffentlicht: (2023)
Steering Rectified Flow Models in the Vector Field for Controlled Image Generation
von: Patel, Maitreya, et al.
Veröffentlicht: (2024)
von: Patel, Maitreya, et al.
Veröffentlicht: (2024)
Show and Segment: Universal Medical Image Segmentation via In-Context Learning
von: Gao, Yunhe, et al.
Veröffentlicht: (2025)
von: Gao, Yunhe, et al.
Veröffentlicht: (2025)
MPDiT: Multi-Patch Global-to-Local Transformer Architecture For Efficient Flow Matching and Diffusion Model
von: Dao, Quan, et al.
Veröffentlicht: (2026)
von: Dao, Quan, et al.
Veröffentlicht: (2026)
Improving Visual Reasoning with Iterative Evidence Refinement
von: Shi, Zeru, et al.
Veröffentlicht: (2026)
von: Shi, Zeru, et al.
Veröffentlicht: (2026)
Neural Deformable Models for 3D Bi-Ventricular Heart Shape Reconstruction and Modeling from 2D Sparse Cardiac Magnetic Resonance Imaging
von: Ye, Meng, et al.
Veröffentlicht: (2023)
von: Ye, Meng, et al.
Veröffentlicht: (2023)
Score-Guided Diffusion for 3D Human Recovery
von: Stathopoulos, Anastasis, et al.
Veröffentlicht: (2024)
von: Stathopoulos, Anastasis, et al.
Veröffentlicht: (2024)
Aligning Human Knowledge with Visual Concepts Towards Explainable Medical Image Classification
von: Gao, Yunhe, et al.
Veröffentlicht: (2024)
von: Gao, Yunhe, et al.
Veröffentlicht: (2024)
Interpretable and Testable Vision Features via Sparse Autoencoders
von: Stevens, Samuel, et al.
Veröffentlicht: (2025)
von: Stevens, Samuel, et al.
Veröffentlicht: (2025)
Anatomy-VLM: A Fine-grained Vision-Language Model for Medical Interpretation
von: Gu, Difei, et al.
Veröffentlicht: (2025)
von: Gu, Difei, et al.
Veröffentlicht: (2025)
Instantaneous Perception of Moving Objects in 3D
von: Liu, Di, et al.
Veröffentlicht: (2024)
von: Liu, Di, et al.
Veröffentlicht: (2024)
Self-Corrected Flow Distillation for Consistent One-Step and Few-Step Text-to-Image Generation
von: Dao, Quan, et al.
Veröffentlicht: (2024)
von: Dao, Quan, et al.
Veröffentlicht: (2024)
Interpretability Transfer from Language to Vision via Sparse Autoencoders
von: Kravets, Alexey, et al.
Veröffentlicht: (2026)
von: Kravets, Alexey, et al.
Veröffentlicht: (2026)
Causal Interpretation of Sparse Autoencoder Features in Vision
von: Han, Sangyu, et al.
Veröffentlicht: (2025)
von: Han, Sangyu, et al.
Veröffentlicht: (2025)
Interpretable and Steerable Concept Bottleneck Sparse Autoencoders
von: Kulkarni, Akshay, et al.
Veröffentlicht: (2025)
von: Kulkarni, Akshay, et al.
Veröffentlicht: (2025)
Beyond Semantics: Disentangling Information Scope in Sparse Autoencoders for CLIP
von: Ro, Yusung, et al.
Veröffentlicht: (2026)
von: Ro, Yusung, et al.
Veröffentlicht: (2026)
Mammo-SAE: Interpreting Breast Cancer Concept Learning with Sparse Autoencoders
von: Nakka, Krishna Kanth
Veröffentlicht: (2025)
von: Nakka, Krishna Kanth
Veröffentlicht: (2025)
Sparse Autoencoders enable Robust and Interpretable Fine-tuning of CLIP models
von: Morelli, Fabian, et al.
Veröffentlicht: (2026)
von: Morelli, Fabian, et al.
Veröffentlicht: (2026)
Continuous Spatio-Temporal Memory Networks for 4D Cardiac Cine MRI Segmentation
von: Ye, Meng, et al.
Veröffentlicht: (2024)
von: Ye, Meng, et al.
Veröffentlicht: (2024)
Interpreting CLIP with Hierarchical Sparse Autoencoders
von: Zaigrajew, Vladimir, et al.
Veröffentlicht: (2025)
von: Zaigrajew, Vladimir, et al.
Veröffentlicht: (2025)
Sparse Autoencoders for Interpretable Medical Image Representation Learning
von: Wesp, Philipp, et al.
Veröffentlicht: (2026)
von: Wesp, Philipp, et al.
Veröffentlicht: (2026)
Towards Open-Ended Visual Scientific Discovery with Sparse Autoencoders
von: Stevens, Samuel, et al.
Veröffentlicht: (2025)
von: Stevens, Samuel, et al.
Veröffentlicht: (2025)
StreamFlow: Theory, Algorithm, and Implementation for High-Efficiency Rectified Flow Generation
von: Fang, Sen, et al.
Veröffentlicht: (2025)
von: Fang, Sen, et al.
Veröffentlicht: (2025)
Multiscale Vector-Quantized Variational Autoencoder for Endoscopic Image Synthesis
von: Diamantis, Dimitrios E., et al.
Veröffentlicht: (2025)
von: Diamantis, Dimitrios E., et al.
Veröffentlicht: (2025)
Steering LVLMs via Sparse Autoencoder for Hallucination Mitigation
von: Hua, Zhenglin, et al.
Veröffentlicht: (2025)
von: Hua, Zhenglin, et al.
Veröffentlicht: (2025)
RAC: Rectified Flow Auto Coder
von: Fang, Sen, et al.
Veröffentlicht: (2026)
von: Fang, Sen, et al.
Veröffentlicht: (2026)
Universal Sparse Autoencoders: Interpretable Cross-Model Concept Alignment
von: Thasarathan, Harrish, et al.
Veröffentlicht: (2025)
von: Thasarathan, Harrish, et al.
Veröffentlicht: (2025)
Disentangling Visual Priors: Unsupervised Learning of Scene Interpretations with Compositional Autoencoder
von: Krawiec, Krzysztof, et al.
Veröffentlicht: (2024)
von: Krawiec, Krzysztof, et al.
Veröffentlicht: (2024)
LoR-VP: Low-Rank Visual Prompting for Efficient Vision Model Adaptation
von: Jin, Can, et al.
Veröffentlicht: (2025)
von: Jin, Can, et al.
Veröffentlicht: (2025)
Resolving Inconsistent Semantics in Multi-Dataset Image Segmentation
von: Zhangli, Qilong, et al.
Veröffentlicht: (2024)
von: Zhangli, Qilong, et al.
Veröffentlicht: (2024)
How to Trace Latent Generative Model Generated Images without Artificial Watermark?
von: Wang, Zhenting, et al.
Veröffentlicht: (2024)
von: Wang, Zhenting, et al.
Veröffentlicht: (2024)
LouvreSAE: Sparse Autoencoders for Interpretable and Controllable Style Transfer
von: Panda, Raina, et al.
Veröffentlicht: (2025)
von: Panda, Raina, et al.
Veröffentlicht: (2025)
Residualized Temporal Sparse Autoencoders for Interpreting Diffusion Models
von: Yeung, Calvin, et al.
Veröffentlicht: (2026)
von: Yeung, Calvin, et al.
Veröffentlicht: (2026)
Beyond Pixels: Semi-Supervised Semantic Segmentation with a Multi-scale Patch-based Multi-Label Classifier
von: Howlader, Prantik, et al.
Veröffentlicht: (2024)
von: Howlader, Prantik, et al.
Veröffentlicht: (2024)
Efficient Unsupervised Visual Representation Learning with Explicit Cluster Balancing
von: Metaxas, Ioannis Maniadis, et al.
Veröffentlicht: (2024)
von: Metaxas, Ioannis Maniadis, et al.
Veröffentlicht: (2024)
Beyond Quantity: Distribution-Aware Labeling for Visual Grounding
von: Zhang, Yichi, et al.
Veröffentlicht: (2025)
von: Zhang, Yichi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Can Cross-Layer Transcoders Replace Vision Transformer Activations? An Interpretable Perspective on Vision
von: Chatzoudis, Gerasimos, et al.
Veröffentlicht: (2026) -
LUCID-SAE: Learning Unified Vision-Language Sparse Codes for Interpretable Concept Discovery
von: Gu, Difei, et al.
Veröffentlicht: (2026) -
The Hidden Life of Tokens: Reducing Hallucination of Large Vision-Language Models via Visual Information Steering
von: Li, Zhuowei, et al.
Veröffentlicht: (2025) -
Test-Time Spectrum-Aware Latent Steering for Zero-Shot Generalization in Vision-Language Models
von: Dafnis, Konstantinos M., et al.
Veröffentlicht: (2025) -
Training Like a Medical Resident: Context-Prior Learning Toward Universal Medical Image Segmentation
von: Gao, Yunhe, et al.
Veröffentlicht: (2023)