Interpretable and Testable Vision Features via Sparse Autoencoders
Fuente:
arXiv
Salvato in:
| Autori principali: | Stevens, Samuel, Chao, Wei-Lun, Berger-Wolf, Tanya, Su, Yu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards Open-Ended Visual Scientific Discovery with Sparse Autoencoders
di: Stevens, Samuel, et al.
Pubblicazione: (2025)
di: Stevens, Samuel, et al.
Pubblicazione: (2025)
Causal Interpretation of Sparse Autoencoder Features in Vision
di: Han, Sangyu, et al.
Pubblicazione: (2025)
di: Han, Sangyu, et al.
Pubblicazione: (2025)
Interpretability Transfer from Language to Vision via Sparse Autoencoders
di: Kravets, Alexey, et al.
Pubblicazione: (2026)
di: Kravets, Alexey, et al.
Pubblicazione: (2026)
A Simple Interpretable Transformer for Fine-Grained Image Classification and Analysis
di: Paul, Dipanjyoti, et al.
Pubblicazione: (2023)
di: Paul, Dipanjyoti, et al.
Pubblicazione: (2023)
Prompt-CAM: Making Vision Transformers Interpretable for Fine-Grained Analysis
di: Chowdhury, Arpita, et al.
Pubblicazione: (2025)
di: Chowdhury, Arpita, et al.
Pubblicazione: (2025)
BioCLIP: A Vision Foundation Model for the Tree of Life
di: Stevens, Samuel, et al.
Pubblicazione: (2023)
di: Stevens, Samuel, et al.
Pubblicazione: (2023)
Finer-CAM: Spotting the Difference Reveals Finer Details for Visual Explanation
di: Zhang, Ziheng, et al.
Pubblicazione: (2025)
di: Zhang, Ziheng, et al.
Pubblicazione: (2025)
Leveraging Latent Visual Reasoning in Silence
di: Zhu, Dongyao, et al.
Pubblicazione: (2026)
di: Zhu, Dongyao, et al.
Pubblicazione: (2026)
BioCAP: Exploiting Synthetic Captions Beyond Labels in Biological Foundation Models
di: Zhang, Ziheng, et al.
Pubblicazione: (2025)
di: Zhang, Ziheng, et al.
Pubblicazione: (2025)
Interpretable and Steerable Concept Bottleneck Sparse Autoencoders
di: Kulkarni, Akshay, et al.
Pubblicazione: (2025)
di: Kulkarni, Akshay, et al.
Pubblicazione: (2025)
An X-Ray Is Worth 15 Features: Sparse Autoencoders for Interpretable Radiology Report Generation
di: Abdulaal, Ahmed, et al.
Pubblicazione: (2024)
di: Abdulaal, Ahmed, et al.
Pubblicazione: (2024)
Mammo-SAE: Interpreting Breast Cancer Concept Learning with Sparse Autoencoders
di: Nakka, Krishna Kanth
Pubblicazione: (2025)
di: Nakka, Krishna Kanth
Pubblicazione: (2025)
Sparse Autoencoders enable Robust and Interpretable Fine-tuning of CLIP models
di: Morelli, Fabian, et al.
Pubblicazione: (2026)
di: Morelli, Fabian, et al.
Pubblicazione: (2026)
Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models
di: Pach, Mateusz, et al.
Pubblicazione: (2025)
di: Pach, Mateusz, et al.
Pubblicazione: (2025)
Reviving the Context: Camera Trap Species Classification as Link Prediction on Multimodal Knowledge Graphs
di: Pahuja, Vardaan, et al.
Pubblicazione: (2023)
di: Pahuja, Vardaan, et al.
Pubblicazione: (2023)
Sparse Autoencoders for Interpretable Medical Image Representation Learning
di: Wesp, Philipp, et al.
Pubblicazione: (2026)
di: Wesp, Philipp, et al.
Pubblicazione: (2026)
Leveraging Sparse LiDAR for RAFT-Stereo: A Depth Pre-Fill Perspective
di: Yoo, Jinsu, et al.
Pubblicazione: (2025)
di: Yoo, Jinsu, et al.
Pubblicazione: (2025)
Interpreting CLIP with Hierarchical Sparse Autoencoders
di: Zaigrajew, Vladimir, et al.
Pubblicazione: (2025)
di: Zaigrajew, Vladimir, et al.
Pubblicazione: (2025)
Universal Sparse Autoencoders: Interpretable Cross-Model Concept Alignment
di: Thasarathan, Harrish, et al.
Pubblicazione: (2025)
di: Thasarathan, Harrish, et al.
Pubblicazione: (2025)
Mind the (Data) Gap: Evaluating Vision Systems in Small Data Applications
di: Stevens, Samuel, et al.
Pubblicazione: (2025)
di: Stevens, Samuel, et al.
Pubblicazione: (2025)
Probing the Representational Power of Sparse Autoencoders in Vision Models
di: Olson, Matthew Lyle, et al.
Pubblicazione: (2025)
di: Olson, Matthew Lyle, et al.
Pubblicazione: (2025)
Analyzing Hierarchical Structure in Vision Models with Sparse Autoencoders
di: Olson, Matthew Lyle, et al.
Pubblicazione: (2025)
di: Olson, Matthew Lyle, et al.
Pubblicazione: (2025)
BioBench: A Blueprint to Move Beyond ImageNet for Scientific ML Benchmarks
di: Stevens, Samuel
Pubblicazione: (2025)
di: Stevens, Samuel
Pubblicazione: (2025)
Task-Oriented Human Grasp Synthesis via Context- and Task-Aware Diffusers
di: Liu, An-Lun, et al.
Pubblicazione: (2025)
di: Liu, An-Lun, et al.
Pubblicazione: (2025)
BioCLIP 2: Emergent Properties from Scaling Hierarchical Contrastive Learning
di: Gu, Jianyang, et al.
Pubblicazione: (2025)
di: Gu, Jianyang, et al.
Pubblicazione: (2025)
LouvreSAE: Sparse Autoencoders for Interpretable and Controllable Style Transfer
di: Panda, Raina, et al.
Pubblicazione: (2025)
di: Panda, Raina, et al.
Pubblicazione: (2025)
Autism Spectrum Disorder Classification with Interpretability in Children based on Structural MRI Features Extracted using Contrastive Variational Autoencoder
di: Ma, Ruimin, et al.
Pubblicazione: (2023)
di: Ma, Ruimin, et al.
Pubblicazione: (2023)
Residualized Temporal Sparse Autoencoders for Interpreting Diffusion Models
di: Yeung, Calvin, et al.
Pubblicazione: (2026)
di: Yeung, Calvin, et al.
Pubblicazione: (2026)
SPARC: Concept-Aligned Sparse Autoencoders for Cross-Model and Cross-Modal Interpretability
di: Nasiri-Sarvi, Ali, et al.
Pubblicazione: (2025)
di: Nasiri-Sarvi, Ali, et al.
Pubblicazione: (2025)
Sparse but not Simpler: A Multi-Level Interpretability Analysis of Vision Transformers
di: Zhang, Siyu
Pubblicazione: (2026)
di: Zhang, Siyu
Pubblicazione: (2026)
SAUCE: Selective Concept Unlearning in Vision-Language Models with Sparse Autoencoders
di: Li, Qing, et al.
Pubblicazione: (2025)
di: Li, Qing, et al.
Pubblicazione: (2025)
Dataset Distillation via Vision-Language Category Prototype
di: Zou, Yawen, et al.
Pubblicazione: (2025)
di: Zou, Yawen, et al.
Pubblicazione: (2025)
Model-Agnostic Gender Bias Control for Text-to-Image Generation via Sparse Autoencoder
di: Wu, Chao, et al.
Pubblicazione: (2025)
di: Wu, Chao, et al.
Pubblicazione: (2025)
Interpreting Low-level Vision Models with Causal Effect Maps
di: Hu, Jinfan, et al.
Pubblicazione: (2024)
di: Hu, Jinfan, et al.
Pubblicazione: (2024)
SPG: Sparse-Projected Guides with Sparse Autoencoders for Zero-Shot Anomaly Detection
di: Nanaumi, Tomoyasu, et al.
Pubblicazione: (2026)
di: Nanaumi, Tomoyasu, et al.
Pubblicazione: (2026)
Fine-Tuning is Fine, if Calibrated
di: Mai, Zheda, et al.
Pubblicazione: (2024)
di: Mai, Zheda, et al.
Pubblicazione: (2024)
TIDE : Temporal-Aware Sparse Autoencoders for Interpretable Diffusion Transformers in Image Generation
di: Huang, Victor Shea-Jay, et al.
Pubblicazione: (2025)
di: Huang, Victor Shea-Jay, et al.
Pubblicazione: (2025)
SparseFormer: Detecting Objects in HRW Shots via Sparse Vision Transformer
di: Li, Wenxi, et al.
Pubblicazione: (2025)
di: Li, Wenxi, et al.
Pubblicazione: (2025)
Deep Feature Consistent Variational Autoencoder
di: Hou, Xianxu, et al.
Pubblicazione: (2016)
di: Hou, Xianxu, et al.
Pubblicazione: (2016)
Aligning Information Capacity Between Vision and Language via Dense-to-Sparse Feature Distillation for Image-Text Matching
di: Liu, Yang, et al.
Pubblicazione: (2025)
di: Liu, Yang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Towards Open-Ended Visual Scientific Discovery with Sparse Autoencoders
di: Stevens, Samuel, et al.
Pubblicazione: (2025) -
Causal Interpretation of Sparse Autoencoder Features in Vision
di: Han, Sangyu, et al.
Pubblicazione: (2025) -
Interpretability Transfer from Language to Vision via Sparse Autoencoders
di: Kravets, Alexey, et al.
Pubblicazione: (2026) -
A Simple Interpretable Transformer for Fine-Grained Image Classification and Analysis
di: Paul, Dipanjyoti, et al.
Pubblicazione: (2023) -
Prompt-CAM: Making Vision Transformers Interpretable for Fine-Grained Analysis
di: Chowdhury, Arpita, et al.
Pubblicazione: (2025)