SALVE: Sparse Autoencoder-Latent Vector Editing for Mechanistic Control of Neural Networks
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Flovik, Vegard |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Achieving Data Efficient Neural Networks with Hybrid Concept-based Models
von: Opsahl, Tobias A., et al.
Veröffentlicht: (2024)
von: Opsahl, Tobias A., et al.
Veröffentlicht: (2024)
Steering Sparse Autoencoder Latents to Control Dynamic Head Pruning in Vision Transformers (Student Abstract)
von: Lee, Yousung, et al.
Veröffentlicht: (2026)
von: Lee, Yousung, et al.
Veröffentlicht: (2026)
Interpreting CLIP with Hierarchical Sparse Autoencoders
von: Zaigrajew, Vladimir, et al.
Veröffentlicht: (2025)
von: Zaigrajew, Vladimir, et al.
Veröffentlicht: (2025)
Residualized Temporal Sparse Autoencoders for Interpreting Diffusion Models
von: Yeung, Calvin, et al.
Veröffentlicht: (2026)
von: Yeung, Calvin, et al.
Veröffentlicht: (2026)
i-MAE: Are Latent Representations in Masked Autoencoders Linearly Separable?
von: Zhang, Kevin, et al.
Veröffentlicht: (2022)
von: Zhang, Kevin, et al.
Veröffentlicht: (2022)
Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models
von: Pach, Mateusz, et al.
Veröffentlicht: (2025)
von: Pach, Mateusz, et al.
Veröffentlicht: (2025)
One-Step is Enough: Sparse Autoencoders for Text-to-Image Diffusion Models
von: Surkov, Viacheslav, et al.
Veröffentlicht: (2024)
von: Surkov, Viacheslav, et al.
Veröffentlicht: (2024)
Kolmogorov-Arnold Network Autoencoders
von: Moradi, Mohammadamin, et al.
Veröffentlicht: (2024)
von: Moradi, Mohammadamin, et al.
Veröffentlicht: (2024)
Contrastive Denoising Score for Text-guided Latent Diffusion Image Editing
von: Nam, Hyelin, et al.
Veröffentlicht: (2023)
von: Nam, Hyelin, et al.
Veröffentlicht: (2023)
Steering LVLMs via Sparse Autoencoder for Hallucination Mitigation
von: Hua, Zhenglin, et al.
Veröffentlicht: (2025)
von: Hua, Zhenglin, et al.
Veröffentlicht: (2025)
FastCAV: Efficient Computation of Concept Activation Vectors for Explaining Deep Neural Networks
von: Schmalwasser, Laines, et al.
Veröffentlicht: (2025)
von: Schmalwasser, Laines, et al.
Veröffentlicht: (2025)
Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual Steering
von: Chatzoudis, Gerasimos, et al.
Veröffentlicht: (2025)
von: Chatzoudis, Gerasimos, et al.
Veröffentlicht: (2025)
Unsupervised Panoptic Interpretation of Latent Spaces in GANs Using Space-Filling Vector Quantization
von: Vali, Mohammad Hassan, et al.
Veröffentlicht: (2024)
von: Vali, Mohammad Hassan, et al.
Veröffentlicht: (2024)
k* Distribution: Evaluating the Latent Space of Deep Neural Networks using Local Neighborhood Analysis
von: Kotyan, Shashank, et al.
Veröffentlicht: (2023)
von: Kotyan, Shashank, et al.
Veröffentlicht: (2023)
Sparse Autoencoders as Plug-and-Play Firewalls for Adversarial Attack Detection in VLMs
von: Wang, Hao, et al.
Veröffentlicht: (2026)
von: Wang, Hao, et al.
Veröffentlicht: (2026)
GeoSAE: Geometric Prior-Guided Layer-Wise Sparse Autoencoder Annotation of Brain MRI Foundation Models
von: Nerrise, Favour, et al.
Veröffentlicht: (2026)
von: Nerrise, Favour, et al.
Veröffentlicht: (2026)
B-SMALL: A Bayesian Neural Network approach to Sparse Model-Agnostic Meta-Learning
von: Madan, Anish, et al.
Veröffentlicht: (2021)
von: Madan, Anish, et al.
Veröffentlicht: (2021)
A Gray-box Attack against Latent Diffusion Model-based Image Editing by Posterior Collapse
von: Guo, Zhongliang, et al.
Veröffentlicht: (2024)
von: Guo, Zhongliang, et al.
Veröffentlicht: (2024)
Efficient Model Editing with Task-Localized Sparse Fine-tuning
von: Iurada, Leonardo, et al.
Veröffentlicht: (2025)
von: Iurada, Leonardo, et al.
Veröffentlicht: (2025)
CSC-Unet: A Novel Convolutional Sparse Coding Strategy Based Neural Network for Semantic Segmentation
von: Tang, Haitong, et al.
Veröffentlicht: (2021)
von: Tang, Haitong, et al.
Veröffentlicht: (2021)
ED-NeRF: Efficient Text-Guided Editing of 3D Scene with Latent Space NeRF
von: Park, Jangho, et al.
Veröffentlicht: (2023)
von: Park, Jangho, et al.
Veröffentlicht: (2023)
SALAD: Skeleton-aware Latent Diffusion for Text-driven Motion Generation and Editing
von: Hong, Seokhyeon, et al.
Veröffentlicht: (2025)
von: Hong, Seokhyeon, et al.
Veröffentlicht: (2025)
From Radiologist Report to Image Label: Assessing Latent Dirichlet Allocation in Training Neural Networks for Orthopedic Radiograph Classification
von: Olczak, Jakub, et al.
Veröffentlicht: (2024)
von: Olczak, Jakub, et al.
Veröffentlicht: (2024)
Taming Latent Diffusion Model for Neural Radiance Field Inpainting
von: Lin, Chieh Hubert, et al.
Veröffentlicht: (2024)
von: Lin, Chieh Hubert, et al.
Veröffentlicht: (2024)
Improving the Diffusability of Autoencoders
von: Skorokhodov, Ivan, et al.
Veröffentlicht: (2025)
von: Skorokhodov, Ivan, et al.
Veröffentlicht: (2025)
Navigating Neural Space: Revisiting Concept Activation Vectors to Overcome Directional Divergence
von: Pahde, Frederik, et al.
Veröffentlicht: (2022)
von: Pahde, Frederik, et al.
Veröffentlicht: (2022)
Diffusion Autoencoders are Scalable Image Tokenizers
von: Chen, Yinbo, et al.
Veröffentlicht: (2025)
von: Chen, Yinbo, et al.
Veröffentlicht: (2025)
MeanFlow Transformers with Representation Autoencoders
von: Hu, Zheyuan, et al.
Veröffentlicht: (2025)
von: Hu, Zheyuan, et al.
Veröffentlicht: (2025)
Controllable Lung Nodule Synthesis via Histogram-Regularized Latent Diffusion Models
von: Kannan, Arunkumar, et al.
Veröffentlicht: (2026)
von: Kannan, Arunkumar, et al.
Veröffentlicht: (2026)
Beyond Adapter Retrieval: Latent Geometry-Preserving Composition via Sparse Task Projection
von: Jin, Pengfei, et al.
Veröffentlicht: (2024)
von: Jin, Pengfei, et al.
Veröffentlicht: (2024)
Masked Autoencoders Are Effective Tokenizers for Diffusion Models
von: Chen, Hao, et al.
Veröffentlicht: (2025)
von: Chen, Hao, et al.
Veröffentlicht: (2025)
Beyond the Known: Adversarial Autoencoders in Novelty Detection
von: Asad, Muhammad, et al.
Veröffentlicht: (2024)
von: Asad, Muhammad, et al.
Veröffentlicht: (2024)
CL-MAE: Curriculum-Learned Masked Autoencoders
von: Madan, Neelu, et al.
Veröffentlicht: (2023)
von: Madan, Neelu, et al.
Veröffentlicht: (2023)
Towards a Mechanistic Explanation of Diffusion Model Generalization
von: Niedoba, Matthew, et al.
Veröffentlicht: (2024)
von: Niedoba, Matthew, et al.
Veröffentlicht: (2024)
Signal Collapse in One-Shot Pruning: When Sparse Models Fail to Distinguish Neural Representations
von: Saikumar, Dhananjay, et al.
Veröffentlicht: (2025)
von: Saikumar, Dhananjay, et al.
Veröffentlicht: (2025)
Explaining Bayesian Neural Networks
von: Bykov, Kirill, et al.
Veröffentlicht: (2021)
von: Bykov, Kirill, et al.
Veröffentlicht: (2021)
Winfree Oscillatory Neural Network
von: Dai, Jiawen, et al.
Veröffentlicht: (2026)
von: Dai, Jiawen, et al.
Veröffentlicht: (2026)
On Diversity in Discriminative Neural Networks
von: Oubaha, Brahim, et al.
Veröffentlicht: (2024)
von: Oubaha, Brahim, et al.
Veröffentlicht: (2024)
Hyperbolic Busemann Neural Networks
von: Chen, Ziheng, et al.
Veröffentlicht: (2026)
von: Chen, Ziheng, et al.
Veröffentlicht: (2026)
Latent Zoning Network: A Unified Principle for Generative Modeling, Representation Learning, and Classification
von: Lin, Zinan, et al.
Veröffentlicht: (2025)
von: Lin, Zinan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Achieving Data Efficient Neural Networks with Hybrid Concept-based Models
von: Opsahl, Tobias A., et al.
Veröffentlicht: (2024) -
Steering Sparse Autoencoder Latents to Control Dynamic Head Pruning in Vision Transformers (Student Abstract)
von: Lee, Yousung, et al.
Veröffentlicht: (2026) -
Interpreting CLIP with Hierarchical Sparse Autoencoders
von: Zaigrajew, Vladimir, et al.
Veröffentlicht: (2025) -
Residualized Temporal Sparse Autoencoders for Interpreting Diffusion Models
von: Yeung, Calvin, et al.
Veröffentlicht: (2026) -
i-MAE: Are Latent Representations in Masked Autoencoders Linearly Separable?
von: Zhang, Kevin, et al.
Veröffentlicht: (2022)