Causal Attribution via Activation Patching
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Izadi, Amirmohammad, Banayeeanzade, Mohammadali, Mirrokni, Alireza, Hasani, Hosein, Bagherian, Mobin, Mehri, Faridoun, Baghshah, Mahdieh Soleymani |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Understanding Counting Mechanisms in Large Language and Vision-Language Models
von: Hasani, Hosein, et al.
Veröffentlicht: (2025)
von: Hasani, Hosein, et al.
Veröffentlicht: (2025)
Uncovering Grounding IDs: How External Cues Shape Multimodal Binding
von: Hasani, Hosein, et al.
Veröffentlicht: (2025)
von: Hasani, Hosein, et al.
Veröffentlicht: (2025)
LibraGrad: Balancing Gradient Flow for Universally Better Vision Transformer Attributions
von: Mehri, Faridoun, et al.
Veröffentlicht: (2024)
von: Mehri, Faridoun, et al.
Veröffentlicht: (2024)
Mechanistic Interpretability of Large-Scale Counting in LLMs through a System-2 Strategy
von: Hasani, Hosein, et al.
Veröffentlicht: (2026)
von: Hasani, Hosein, et al.
Veröffentlicht: (2026)
Visual Structures Helps Visual Reasoning: Addressing the Binding Problem in VLMs
von: Izadi, Amirmohammad, et al.
Veröffentlicht: (2025)
von: Izadi, Amirmohammad, et al.
Veröffentlicht: (2025)
Attention Overlap Is Responsible for The Entity Missing Problem in Text-to-image Diffusion Models!
von: Marioriyad, Arash, et al.
Veröffentlicht: (2024)
von: Marioriyad, Arash, et al.
Veröffentlicht: (2024)
Analyzing CLIP's Performance Limitations in Multi-Object Scenarios: A Controlled High-Resolution Study
von: Abbasi, Reza, et al.
Veröffentlicht: (2025)
von: Abbasi, Reza, et al.
Veröffentlicht: (2025)
CLIP Under the Microscope: A Fine-Grained Analysis of Multi-Object Representation
von: Abbasi, Reza, et al.
Veröffentlicht: (2025)
von: Abbasi, Reza, et al.
Veröffentlicht: (2025)
Improving 3D Few-Shot Segmentation with Inference-Time Pseudo-Labeling
von: Mozafari, Mohammad, et al.
Veröffentlicht: (2024)
von: Mozafari, Mohammad, et al.
Veröffentlicht: (2024)
HaloProbe: Bayesian Detection and Mitigation of Object Hallucinations in Vision-Language Models
von: Zohrabi, Reihaneh, et al.
Veröffentlicht: (2026)
von: Zohrabi, Reihaneh, et al.
Veröffentlicht: (2026)
Spurious-Aware Prototype Refinement for Reliable Out-of-Distribution Detection
von: Zohrabi, Reihaneh, et al.
Veröffentlicht: (2025)
von: Zohrabi, Reihaneh, et al.
Veröffentlicht: (2025)
ComAlign: Compositional Alignment in Vision-Language Models
von: Abdollah, Ali, et al.
Veröffentlicht: (2024)
von: Abdollah, Ali, et al.
Veröffentlicht: (2024)
T2I-FineEval: Fine-Grained Compositional Metric for Text-to-Image Evaluation
von: Hosseini, Seyed Mohammad Hadi, et al.
Veröffentlicht: (2025)
von: Hosseini, Seyed Mohammad Hadi, et al.
Veröffentlicht: (2025)
Language Plays a Pivotal Role in the Object-Attribute Compositional Generalization of CLIP
von: Abbasi, Reza, et al.
Veröffentlicht: (2024)
von: Abbasi, Reza, et al.
Veröffentlicht: (2024)
Deciphering the Role of Representation Disentanglement: Investigating Compositional Generalization in CLIP Models
von: Abbasi, Reza, et al.
Veröffentlicht: (2024)
von: Abbasi, Reza, et al.
Veröffentlicht: (2024)
Dilated Balanced Cross Entropy Loss for Medical Image Segmentation
von: Hosseini, Seyed Mohsen, et al.
Veröffentlicht: (2024)
von: Hosseini, Seyed Mohsen, et al.
Veröffentlicht: (2024)
Classification of Breast Cancer Histopathology Images using a Modified Supervised Contrastive Learning Method
von: Sani, Matina Mahdizadeh, et al.
Veröffentlicht: (2024)
von: Sani, Matina Mahdizadeh, et al.
Veröffentlicht: (2024)
Why Settle for Mid: A Probabilistic Viewpoint to Spatial Relationship Alignment in Text-to-image Models
von: Rezaei, Parham, et al.
Veröffentlicht: (2025)
von: Rezaei, Parham, et al.
Veröffentlicht: (2025)
Diffusion Beats Autoregressive: An Evaluation of Compositional Generation in Text-to-Image Models
von: Marioriyad, Arash, et al.
Veröffentlicht: (2024)
von: Marioriyad, Arash, et al.
Veröffentlicht: (2024)
Eye-Q: A Multilingual Benchmark for Visual Word Puzzle Solving and Image-to-Phrase Reasoning
von: Najar, Ali, et al.
Veröffentlicht: (2026)
von: Najar, Ali, et al.
Veröffentlicht: (2026)
Trained Models Tell Us How to Make Them Robust to Spurious Correlation without Group Annotation
von: Ghaznavi, Mahdi, et al.
Veröffentlicht: (2024)
von: Ghaznavi, Mahdi, et al.
Veröffentlicht: (2024)
Fine-Grained Alignment and Noise Refinement for Compositional Text-to-Image Generation
von: Izadi, Amir Mohammad, et al.
Veröffentlicht: (2025)
von: Izadi, Amir Mohammad, et al.
Veröffentlicht: (2025)
Infinity and Beyond: Compositional Alignment in VAR and Diffusion T2I Models
von: Shahabadi, Hossein, et al.
Veröffentlicht: (2025)
von: Shahabadi, Hossein, et al.
Veröffentlicht: (2025)
Hidden Meanings in Plain Sight: RebusBench for Evaluating Cognitive Visual Reasoning
von: Kasaei, Seyed Amir, et al.
Veröffentlicht: (2026)
von: Kasaei, Seyed Amir, et al.
Veröffentlicht: (2026)
Decompose-and-Compose: A Compositional Approach to Mitigating Spurious Correlation
von: Noohdani, Fahimeh Hosseini, et al.
Veröffentlicht: (2024)
von: Noohdani, Fahimeh Hosseini, et al.
Veröffentlicht: (2024)
No Concept Left Behind: Test-Time Optimization for Compositional Text-to-Image Generation
von: Sameti, Mohammad Hossein, et al.
Veröffentlicht: (2025)
von: Sameti, Mohammad Hossein, et al.
Veröffentlicht: (2025)
Evaluating the Evaluators: Metrics for Compositional Text-to-Image Generation
von: Kasaei, Seyed Amir, et al.
Veröffentlicht: (2025)
von: Kasaei, Seyed Amir, et al.
Veröffentlicht: (2025)
CARINOX: Inference-time Scaling with Category-Aware Reward-based Initial Noise Optimization and Exploration
von: Kasaei, Seyed Amir, et al.
Veröffentlicht: (2025)
von: Kasaei, Seyed Amir, et al.
Veröffentlicht: (2025)
RODEO: Robust Outlier Detection via Exposing Adaptive Out-of-Distribution Samples
von: Mirzaei, Hossein, et al.
Veröffentlicht: (2025)
von: Mirzaei, Hossein, et al.
Veröffentlicht: (2025)
HyCoVAD: A Hybrid SSL-LLM Model for Complex Video Anomaly Detection
von: Hemmatyar, Mohammad Mahdi, et al.
Veröffentlicht: (2025)
von: Hemmatyar, Mohammad Mahdi, et al.
Veröffentlicht: (2025)
CAREL: Instruction-guided reinforcement learning with cross-modal auxiliary objectives
von: Saghafian, Armin, et al.
Veröffentlicht: (2024)
von: Saghafian, Armin, et al.
Veröffentlicht: (2024)
Face-LLaVA: Facial Expression and Attribute Understanding through Instruction Tuning
von: Chaubey, Ashutosh, et al.
Veröffentlicht: (2025)
von: Chaubey, Ashutosh, et al.
Veröffentlicht: (2025)
GABInsight: Exploring Gender-Activity Binding Bias in Vision-Language Models
von: Abdollahi, Ali, et al.
Veröffentlicht: (2024)
von: Abdollahi, Ali, et al.
Veröffentlicht: (2024)
A Contrastive Teacher-Student Framework for Novelty Detection under Style Shifts
von: Mirzaei, Hossein, et al.
Veröffentlicht: (2025)
von: Mirzaei, Hossein, et al.
Veröffentlicht: (2025)
ProMark: Proactive Diffusion Watermarking for Causal Attribution
von: Asnani, Vishal, et al.
Veröffentlicht: (2024)
von: Asnani, Vishal, et al.
Veröffentlicht: (2024)
Patch-Based Spatial Authorship Attribution in Human-Robot Collaborative Paintings
von: Chen, Eric, et al.
Veröffentlicht: (2026)
von: Chen, Eric, et al.
Veröffentlicht: (2026)
Causal Physics Steering in Video World Models via Concept Activation Vectors
von: Alam, Nahid
Veröffentlicht: (2026)
von: Alam, Nahid
Veröffentlicht: (2026)
From Activation to Causality: Discovery of Causal Visual Representations in the Human Brain
von: Golbari, Yuval, et al.
Veröffentlicht: (2026)
von: Golbari, Yuval, et al.
Veröffentlicht: (2026)
All Patches Matter, More Patches Better: Enhance AI-Generated Image Detection via Panoptic Patch Learning
von: Yang, Zheng, et al.
Veröffentlicht: (2025)
von: Yang, Zheng, et al.
Veröffentlicht: (2025)
LaFAM: Unsupervised Feature Attribution with Label-free Activation Maps
von: Karjauv, Aray, et al.
Veröffentlicht: (2024)
von: Karjauv, Aray, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Understanding Counting Mechanisms in Large Language and Vision-Language Models
von: Hasani, Hosein, et al.
Veröffentlicht: (2025) -
Uncovering Grounding IDs: How External Cues Shape Multimodal Binding
von: Hasani, Hosein, et al.
Veröffentlicht: (2025) -
LibraGrad: Balancing Gradient Flow for Universally Better Vision Transformer Attributions
von: Mehri, Faridoun, et al.
Veröffentlicht: (2024) -
Mechanistic Interpretability of Large-Scale Counting in LLMs through a System-2 Strategy
von: Hasani, Hosein, et al.
Veröffentlicht: (2026) -
Visual Structures Helps Visual Reasoning: Addressing the Binding Problem in VLMs
von: Izadi, Amirmohammad, et al.
Veröffentlicht: (2025)