Mechanistic Interpretability of Diffusion Models: Circuit-Level Analysis and Causal Validation
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Roy, Dip |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Bayesian Autoencoder for Medical Anomaly Detection: Uncertainty-Aware Approach for Brain 2 MRI Analysis
von: Roy, Dip
Veröffentlicht: (2025)
von: Roy, Dip
Veröffentlicht: (2025)
Causal Tracing of Object Representations in Large Vision Language Models: Mechanistic Interpretability and Hallucination Mitigation
von: Li, Qiming, et al.
Veröffentlicht: (2025)
von: Li, Qiming, et al.
Veröffentlicht: (2025)
Counting Circuits: Mechanistic Interpretability of Visual Reasoning in Large Vision-Language Models
von: Che, Liwei, et al.
Veröffentlicht: (2026)
von: Che, Liwei, et al.
Veröffentlicht: (2026)
Dissecting and Mitigating Diffusion Bias via Mechanistic Interpretability
von: Shi, Yingdong, et al.
Veröffentlicht: (2025)
von: Shi, Yingdong, et al.
Veröffentlicht: (2025)
Scale Alone Does not Improve Mechanistic Interpretability in Vision Models
von: Zimmermann, Roland S., et al.
Veröffentlicht: (2023)
von: Zimmermann, Roland S., et al.
Veröffentlicht: (2023)
Certified Circuits: Stability Guarantees for Mechanistic Circuits
von: Anani, Alaa, et al.
Veröffentlicht: (2026)
von: Anani, Alaa, et al.
Veröffentlicht: (2026)
MCAM: Multimodal Causal Analysis Model for Ego-Vehicle-Level Driving Video Understanding
von: Cheng, Tongtong, et al.
Veröffentlicht: (2025)
von: Cheng, Tongtong, et al.
Veröffentlicht: (2025)
Causal Diffusion Transformers for Generative Modeling
von: Deng, Chaorui, et al.
Veröffentlicht: (2024)
von: Deng, Chaorui, et al.
Veröffentlicht: (2024)
Where Reliability Lives in Vision-Language Models: A Mechanistic Study of Attention, Hidden States, and Causal Circuits
von: Mann, Logan, et al.
Veröffentlicht: (2026)
von: Mann, Logan, et al.
Veröffentlicht: (2026)
From Local to Global to Mechanistic: An iERF-Centered Unified Framework for Interpreting Vision Models
von: Kim, Yearim, et al.
Veröffentlicht: (2026)
von: Kim, Yearim, et al.
Veröffentlicht: (2026)
Attention Sinks in Diffusion Transformers: A Causal Analysis
von: Wu, Fangzheng, et al.
Veröffentlicht: (2026)
von: Wu, Fangzheng, et al.
Veröffentlicht: (2026)
Interpreting Low-level Vision Models with Causal Effect Maps
von: Hu, Jinfan, et al.
Veröffentlicht: (2024)
von: Hu, Jinfan, et al.
Veröffentlicht: (2024)
Sparse but not Simpler: A Multi-Level Interpretability Analysis of Vision Transformers
von: Zhang, Siyu
Veröffentlicht: (2026)
von: Zhang, Siyu
Veröffentlicht: (2026)
CausalDiff: Causality-Inspired Disentanglement via Diffusion Model for Adversarial Defense
von: Zhang, Mingkun, et al.
Veröffentlicht: (2024)
von: Zhang, Mingkun, et al.
Veröffentlicht: (2024)
Brain Imaging-to-Graph Generation using Adversarial Hierarchical Diffusion Models for MCI Causality Analysis
von: Zuo, Qiankun, et al.
Veröffentlicht: (2023)
von: Zuo, Qiankun, et al.
Veröffentlicht: (2023)
Integrating Anatomical Priors into a Causal Diffusion Model
von: Li, Binxu, et al.
Veröffentlicht: (2025)
von: Li, Binxu, et al.
Veröffentlicht: (2025)
Causal Motion Diffusion Models for Autoregressive Motion Generation
von: Yu, Qing, et al.
Veröffentlicht: (2026)
von: Yu, Qing, et al.
Veröffentlicht: (2026)
Radiologist-Guided Causal Concept Bottleneck Models for Chest X-Ray Interpretation
von: Rafferty, Amy, et al.
Veröffentlicht: (2026)
von: Rafferty, Amy, et al.
Veröffentlicht: (2026)
Do VLMs Have Bad Eyes? Diagnosing Compositional Failures via Mechanistic Interpretability
von: Aravindan, Ashwath Vaithinathan, et al.
Veröffentlicht: (2025)
von: Aravindan, Ashwath Vaithinathan, et al.
Veröffentlicht: (2025)
CURE: Concept Unlearning via Orthogonal Representation Editing in Diffusion Models
von: Biswas, Shristi Das, et al.
Veröffentlicht: (2025)
von: Biswas, Shristi Das, et al.
Veröffentlicht: (2025)
HEART: Hyperspherical Embedding Alignment via Kent-Representation Traversal in Diffusion Models
von: Roy, Arani, et al.
Veröffentlicht: (2026)
von: Roy, Arani, et al.
Veröffentlicht: (2026)
FineCausal: A Causal-Based Framework for Interpretable Fine-Grained Action Quality Assessment
von: Han, Ruisheng, et al.
Veröffentlicht: (2025)
von: Han, Ruisheng, et al.
Veröffentlicht: (2025)
Video Diffusion Models are Training-free Motion Interpreter and Controller
von: Xiao, Zeqi, et al.
Veröffentlicht: (2024)
von: Xiao, Zeqi, et al.
Veröffentlicht: (2024)
Event-Level Detection of Surgical Instrument Handovers in Videos with Interpretable Vision Models
von: Katsarou, Katerina, et al.
Veröffentlicht: (2026)
von: Katsarou, Katerina, et al.
Veröffentlicht: (2026)
C$^2$MIL: Synchronizing Semantic and Topological Causalities in Multiple Instance Learning for Robust and Interpretable Survival Analysis
von: Cen, Min, et al.
Veröffentlicht: (2025)
von: Cen, Min, et al.
Veröffentlicht: (2025)
Interpretable Failure Detection with Human-Level Concepts
von: Nguyen, Kien X., et al.
Veröffentlicht: (2025)
von: Nguyen, Kien X., et al.
Veröffentlicht: (2025)
Causal Deciphering and Inpainting in Spatio-Temporal Dynamics via Diffusion Model
von: Duan, Yifan, et al.
Veröffentlicht: (2024)
von: Duan, Yifan, et al.
Veröffentlicht: (2024)
RFDM: Residual Flow Diffusion Model for Efficient Causal Video Editing
von: Salehi, Mohammadreza, et al.
Veröffentlicht: (2026)
von: Salehi, Mohammadreza, et al.
Veröffentlicht: (2026)
EDITOR: Effective and Interpretable Prompt Inversion for Text-to-Image Diffusion Models
von: Li, Mingzhe, et al.
Veröffentlicht: (2025)
von: Li, Mingzhe, et al.
Veröffentlicht: (2025)
Med-SegLens: Latent-Level Model Diffing for Interpretable Medical Image Segmentation
von: Ahmed, Salma J., et al.
Veröffentlicht: (2026)
von: Ahmed, Salma J., et al.
Veröffentlicht: (2026)
LLM-Powered Flood Depth Estimation from Social Media Imagery: A Vision-Language Model Framework with Mechanistic Interpretability for Transportation Resilience
von: Fuad, Nafis, et al.
Veröffentlicht: (2026)
von: Fuad, Nafis, et al.
Veröffentlicht: (2026)
DiffusionPID: Interpreting Diffusion via Partial Information Decomposition
von: Zawar, Rushikesh, et al.
Veröffentlicht: (2024)
von: Zawar, Rushikesh, et al.
Veröffentlicht: (2024)
Causal Interpretability for Adversarial Robustness: A Hybrid Generative Classification Approach
von: Zhao, Chunheng, et al.
Veröffentlicht: (2024)
von: Zhao, Chunheng, et al.
Veröffentlicht: (2024)
Padding Tone: A Mechanistic Analysis of Padding Tokens in T2I Models
von: Toker, Michael, et al.
Veröffentlicht: (2025)
von: Toker, Michael, et al.
Veröffentlicht: (2025)
A Causal Diffusion Model for Video Reconstruction from Ultra-Low-Bitrate Representations
von: Eteke, Cem, et al.
Veröffentlicht: (2026)
von: Eteke, Cem, et al.
Veröffentlicht: (2026)
An Ordinal Diffusion Model for Generating Medical Images with Different Severity Levels
von: Takezaki, Shumpei, et al.
Veröffentlicht: (2024)
von: Takezaki, Shumpei, et al.
Veröffentlicht: (2024)
Towards Seamless Interaction: Causal Turn-Level Modeling of Interactive 3D Conversational Head Dynamics
von: Chen, Junjie, et al.
Veröffentlicht: (2025)
von: Chen, Junjie, et al.
Veröffentlicht: (2025)
On Mechanistic Knowledge Localization in Text-to-Image Generative Models
von: Basu, Samyadeep, et al.
Veröffentlicht: (2024)
von: Basu, Samyadeep, et al.
Veröffentlicht: (2024)
The Dual Power of Interpretable Token Embeddings: Jailbreaking Attacks and Defenses for Diffusion Model Unlearning
von: Chen, Siyi, et al.
Veröffentlicht: (2025)
von: Chen, Siyi, et al.
Veröffentlicht: (2025)
ProMark: Proactive Diffusion Watermarking for Causal Attribution
von: Asnani, Vishal, et al.
Veröffentlicht: (2024)
von: Asnani, Vishal, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Bayesian Autoencoder for Medical Anomaly Detection: Uncertainty-Aware Approach for Brain 2 MRI Analysis
von: Roy, Dip
Veröffentlicht: (2025) -
Causal Tracing of Object Representations in Large Vision Language Models: Mechanistic Interpretability and Hallucination Mitigation
von: Li, Qiming, et al.
Veröffentlicht: (2025) -
Counting Circuits: Mechanistic Interpretability of Visual Reasoning in Large Vision-Language Models
von: Che, Liwei, et al.
Veröffentlicht: (2026) -
Dissecting and Mitigating Diffusion Bias via Mechanistic Interpretability
von: Shi, Yingdong, et al.
Veröffentlicht: (2025) -
Scale Alone Does not Improve Mechanistic Interpretability in Vision Models
von: Zimmermann, Roland S., et al.
Veröffentlicht: (2023)