On the Faithfulness of Vision Transformer Explanations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Junyi, Kang, Weitai, Tang, Hao, Hong, Yuan, Yan, Yan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Token Transformation Matters: Towards Faithful Post-hoc Explanation for Vision Transformer
von: Wu, Junyi, et al.
Veröffentlicht: (2024)
von: Wu, Junyi, et al.
Veröffentlicht: (2024)
Visual Grounding with Attention-Driven Constraint Balancing
von: Kang, Weitai, et al.
Veröffentlicht: (2024)
von: Kang, Weitai, et al.
Veröffentlicht: (2024)
SegVG: Transferring Object Bounding Box to Segmentation for Visual Grounding
von: Kang, Weitai, et al.
Veröffentlicht: (2024)
von: Kang, Weitai, et al.
Veröffentlicht: (2024)
ACTRESS: Active Retraining for Semi-supervised Visual Grounding
von: Kang, Weitai, et al.
Veröffentlicht: (2024)
von: Kang, Weitai, et al.
Veröffentlicht: (2024)
Intent3D: 3D Object Detection in RGB-D Scans Based on Human Intention
von: Kang, Weitai, et al.
Veröffentlicht: (2024)
von: Kang, Weitai, et al.
Veröffentlicht: (2024)
VGent: Visual Grounding via Modular Design for Disentangling Reasoning and Prediction
von: Kang, Weitai, et al.
Veröffentlicht: (2025)
von: Kang, Weitai, et al.
Veröffentlicht: (2025)
3DResT: A Strong Baseline for Semi-Supervised 3D Referring Expression Segmentation
von: Chen, Wenxin, et al.
Veröffentlicht: (2025)
von: Chen, Wenxin, et al.
Veröffentlicht: (2025)
CogniMap3D: Cognitive 3D Mapping and Rapid Retrieval
von: Wang, Feiran, et al.
Veröffentlicht: (2026)
von: Wang, Feiran, et al.
Veröffentlicht: (2026)
Explanation-Driven Counterfactual Testing for Faithfulness in Vision-Language Model Explanations
von: Ding, Sihao, et al.
Veröffentlicht: (2025)
von: Ding, Sihao, et al.
Veröffentlicht: (2025)
Robin3D: Improving 3D Large Language Model via Robust Instruction Tuning
von: Kang, Weitai, et al.
Veröffentlicht: (2024)
von: Kang, Weitai, et al.
Veröffentlicht: (2024)
PTQ4DiT: Post-training Quantization for Diffusion Transformers
von: Wu, Junyi, et al.
Veröffentlicht: (2024)
von: Wu, Junyi, et al.
Veröffentlicht: (2024)
Faithful Counterfactual Visual Explanations (FCVE)
von: Khan, Bismillah, et al.
Veröffentlicht: (2025)
von: Khan, Bismillah, et al.
Veröffentlicht: (2025)
GIFT: A Framework Towards Global Interpretable Faithful Textual Explanations of Vision Classifiers
von: Zablocki, Éloi, et al.
Veröffentlicht: (2024)
von: Zablocki, Éloi, et al.
Veröffentlicht: (2024)
ExpVG: Investigating the Design Space of Visual Grounding in Multimodal Large Language Model
von: Kang, Weitai, et al.
Veröffentlicht: (2025)
von: Kang, Weitai, et al.
Veröffentlicht: (2025)
SAM-Sode: Towards Faithful Explanations for Tiny Bacteria Detection
von: Tan, Wanying, et al.
Veröffentlicht: (2026)
von: Tan, Wanying, et al.
Veröffentlicht: (2026)
Accelerating Vision Transformers on Brain Processing Unit
von: Tang, Jinchi, et al.
Veröffentlicht: (2026)
von: Tang, Jinchi, et al.
Veröffentlicht: (2026)
QuEST: Low-bit Diffusion Model Quantization via Efficient Selective Finetuning
von: Wang, Haoxuan, et al.
Veröffentlicht: (2024)
von: Wang, Haoxuan, et al.
Veröffentlicht: (2024)
[Re] Improving Interpretation Faithfulness for Vision Transformers
von: Kurek, Izabela, et al.
Veröffentlicht: (2025)
von: Kurek, Izabela, et al.
Veröffentlicht: (2025)
Dataset Quantization with Active Learning based Adaptive Sampling
von: Zhao, Zhenghao, et al.
Veröffentlicht: (2024)
von: Zhao, Zhenghao, et al.
Veröffentlicht: (2024)
Inline Critic Steers Image Editing
von: Kang, Weitai, et al.
Veröffentlicht: (2026)
von: Kang, Weitai, et al.
Veröffentlicht: (2026)
There is More to Attention: Statistical Filtering Enhances Explanations in Vision Transformers
von: Ayyar, Meghna P, et al.
Veröffentlicht: (2025)
von: Ayyar, Meghna P, et al.
Veröffentlicht: (2025)
Improving Interpretation Faithfulness for Vision Transformers
von: Hu, Lijie, et al.
Veröffentlicht: (2023)
von: Hu, Lijie, et al.
Veröffentlicht: (2023)
Faithful Extreme Image Rescaling with Learnable Reversible Transformation and Semantic Priors
von: Wei, Hao, et al.
Veröffentlicht: (2026)
von: Wei, Hao, et al.
Veröffentlicht: (2026)
Zero-Shot Faithful Textual Explanations via Directional-Derivative Influence on Predictions
von: Yamauchi, Toshinori, et al.
Veröffentlicht: (2026)
von: Yamauchi, Toshinori, et al.
Veröffentlicht: (2026)
Empowering CAM-Based Methods with Capability to Generate Fine-Grained and High-Faithfulness Explanations
von: Qiu, Changqing, et al.
Veröffentlicht: (2023)
von: Qiu, Changqing, et al.
Veröffentlicht: (2023)
Inpainting the Gaps: A Novel Framework for Evaluating Explanation Methods in Vision Transformers
von: Badisa, Lokesh, et al.
Veröffentlicht: (2024)
von: Badisa, Lokesh, et al.
Veröffentlicht: (2024)
From Particles to Fields: Reframing Photon Mapping with Continuous Gaussian Photon Fields
von: Tao, Jiachen, et al.
Veröffentlicht: (2025)
von: Tao, Jiachen, et al.
Veröffentlicht: (2025)
Token Transforming: A Unified and Training-Free Token Compression Framework for Vision Transformer Acceleration
von: Zeng, Fanhu, et al.
Veröffentlicht: (2025)
von: Zeng, Fanhu, et al.
Veröffentlicht: (2025)
Fusion of regional and sparse attention in Vision Transformers
von: Ibtehaz, Nabil, et al.
Veröffentlicht: (2024)
von: Ibtehaz, Nabil, et al.
Veröffentlicht: (2024)
A Novel Vision Transformer for Camera-LiDAR Fusion based Traffic Object Segmentation
von: Tahves, Toomas, et al.
Veröffentlicht: (2025)
von: Tahves, Toomas, et al.
Veröffentlicht: (2025)
Centroid-centered Modeling for Efficient Vision Transformer Pre-training
von: Yan, Xin, et al.
Veröffentlicht: (2023)
von: Yan, Xin, et al.
Veröffentlicht: (2023)
Context-Aware Decoding for Faithful Vision-Language Generation
von: Fazli, Mehrdad, et al.
Veröffentlicht: (2026)
von: Fazli, Mehrdad, et al.
Veröffentlicht: (2026)
Improving Network Interpretability via Explanation Consistency Evaluation
von: Wu, Hefeng, et al.
Veröffentlicht: (2024)
von: Wu, Hefeng, et al.
Veröffentlicht: (2024)
SVLTA: Benchmarking Vision-Language Temporal Alignment via Synthetic Video Situation
von: Du, Hao, et al.
Veröffentlicht: (2025)
von: Du, Hao, et al.
Veröffentlicht: (2025)
Vision Transformers with Self-Distilled Registers
von: Chen, Yinjie, et al.
Veröffentlicht: (2025)
von: Chen, Yinjie, et al.
Veröffentlicht: (2025)
ACC-ViT : Atrous Convolution's Comeback in Vision Transformers
von: Ibtehaz, Nabil, et al.
Veröffentlicht: (2024)
von: Ibtehaz, Nabil, et al.
Veröffentlicht: (2024)
CNNs, Transformers, Hybrid, and Vision Language Models for Skin Cancer Detection
von: Dey, Durjoy, et al.
Veröffentlicht: (2026)
von: Dey, Durjoy, et al.
Veröffentlicht: (2026)
WeakTr: Exploring Plain Vision Transformer for Weakly-supervised Semantic Segmentation
von: Zhu, Lianghui, et al.
Veröffentlicht: (2023)
von: Zhu, Lianghui, et al.
Veröffentlicht: (2023)
Efficient Multimodal Dataset Distillation via Generative Models
von: Zhao, Zhenghao, et al.
Veröffentlicht: (2025)
von: Zhao, Zhenghao, et al.
Veröffentlicht: (2025)
TraceFlow: Dynamic 3D Reconstruction of Specular Scenes Driven by Ray Tracing
von: Tao, Jiachen, et al.
Veröffentlicht: (2025)
von: Tao, Jiachen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Token Transformation Matters: Towards Faithful Post-hoc Explanation for Vision Transformer
von: Wu, Junyi, et al.
Veröffentlicht: (2024) -
Visual Grounding with Attention-Driven Constraint Balancing
von: Kang, Weitai, et al.
Veröffentlicht: (2024) -
SegVG: Transferring Object Bounding Box to Segmentation for Visual Grounding
von: Kang, Weitai, et al.
Veröffentlicht: (2024) -
ACTRESS: Active Retraining for Semi-supervised Visual Grounding
von: Kang, Weitai, et al.
Veröffentlicht: (2024) -
Intent3D: 3D Object Detection in RGB-D Scans Based on Human Intention
von: Kang, Weitai, et al.
Veröffentlicht: (2024)