Multi-Head Explainer: A General Framework to Improve Explainability in CNNs and Transformers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Bohang, Liò, Pietro |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TraNCE: Transformative Non-linear Concept Explainer for CNNs
von: Akpudo, Ugochukwu Ejike, et al.
Veröffentlicht: (2025)
von: Akpudo, Ugochukwu Ejike, et al.
Veröffentlicht: (2025)
Enhancing Surgical Documentation through Multimodal Visual-Temporal Transformers and Generative AI
von: Georgenthum, Hugo, et al.
Veröffentlicht: (2025)
von: Georgenthum, Hugo, et al.
Veröffentlicht: (2025)
Less is More: The Influence of Pruning on the Explainability of CNNs
von: Merkle, Florian, et al.
Veröffentlicht: (2023)
von: Merkle, Florian, et al.
Veröffentlicht: (2023)
Automated Image Captioning with CNNs and Transformers
von: Cahyono, Joshua Adrian, et al.
Veröffentlicht: (2024)
von: Cahyono, Joshua Adrian, et al.
Veröffentlicht: (2024)
Evaluating the Stability of Semantic Concept Representations in CNNs for Robust Explainability
von: Mikriukov, Georgii, et al.
Veröffentlicht: (2023)
von: Mikriukov, Georgii, et al.
Veröffentlicht: (2023)
Explainable embeddings with Distance Explainer
von: Meijer, Christiaan, et al.
Veröffentlicht: (2025)
von: Meijer, Christiaan, et al.
Veröffentlicht: (2025)
Towards Explainable AI: Multi-Modal Transformer for Video-based Image Description Generation
von: Agarwal, Lakshita, et al.
Veröffentlicht: (2025)
von: Agarwal, Lakshita, et al.
Veröffentlicht: (2025)
Towards Optimal Trade-offs in Knowledge Distillation for CNNs and Vision Transformers at the Edge
von: Violos, John, et al.
Veröffentlicht: (2024)
von: Violos, John, et al.
Veröffentlicht: (2024)
When CNNs Outperform Transformers and Mambas: Revisiting Deep Architectures for Dental Caries Segmentation
von: Ghimire, Aashish, et al.
Veröffentlicht: (2025)
von: Ghimire, Aashish, et al.
Veröffentlicht: (2025)
Learning Online Scale Transformation for Talking Head Video Generation
von: Hong, Fa-Ting, et al.
Veröffentlicht: (2024)
von: Hong, Fa-Ting, et al.
Veröffentlicht: (2024)
Efficient and Concise Explanations for Object Detection with Gaussian-Class Activation Mapping Explainer
von: Nguyen, Quoc Khanh, et al.
Veröffentlicht: (2024)
von: Nguyen, Quoc Khanh, et al.
Veröffentlicht: (2024)
Entropy-Lens: Uncovering Decision Strategies in LLMs
von: Ali, Riccardo, et al.
Veröffentlicht: (2025)
von: Ali, Riccardo, et al.
Veröffentlicht: (2025)
A Framework for Evaluating Zero-Shot Image Generation in Concept-based Explainability
von: Astolfi, Giacomo, et al.
Veröffentlicht: (2026)
von: Astolfi, Giacomo, et al.
Veröffentlicht: (2026)
FiVL: A Framework for Improved Vision-Language Alignment through the Lens of Training, Evaluation and Explainability
von: Aflalo, Estelle, et al.
Veröffentlicht: (2024)
von: Aflalo, Estelle, et al.
Veröffentlicht: (2024)
A Fully Transformer Based Multimodal Framework for Explainable Cancer Image Segmentation Using Radiology Reports
von: Adahada, Enobong, et al.
Veröffentlicht: (2025)
von: Adahada, Enobong, et al.
Veröffentlicht: (2025)
UniHead: Unifying Multi-Perception for Detection Heads
von: Zhou, Hantao, et al.
Veröffentlicht: (2023)
von: Zhou, Hantao, et al.
Veröffentlicht: (2023)
Faithful Attention Explainer: Verbalizing Decisions Based on Discriminative Features
von: Rong, Yao, et al.
Veröffentlicht: (2024)
von: Rong, Yao, et al.
Veröffentlicht: (2024)
General vs Domain-Specific CNNs: Understanding Pretraining Effects on Brain MRI Tumor Classification
von: Abedini, Helia, et al.
Veröffentlicht: (2025)
von: Abedini, Helia, et al.
Veröffentlicht: (2025)
Beyond CNNs: Efficient Fine-Tuning of Multi-Modal LLMs for Object Detection on Low-Data Regimes
von: Elamon, Nirmal, et al.
Veröffentlicht: (2025)
von: Elamon, Nirmal, et al.
Veröffentlicht: (2025)
Explainable Face Recognition via Improved Localization
von: Shadman, Rashik, et al.
Veröffentlicht: (2025)
von: Shadman, Rashik, et al.
Veröffentlicht: (2025)
Pruning By Explaining Revisited: Optimizing Attribution Methods to Prune CNNs and Transformers
von: Hatefi, Sayed Mohammad Vakilzadeh, et al.
Veröffentlicht: (2024)
von: Hatefi, Sayed Mohammad Vakilzadeh, et al.
Veröffentlicht: (2024)
RAIDX: A Retrieval-Augmented Generation and GRPO Reinforcement Learning Framework for Explainable Deepfake Detection
von: Li, Tianxiao, et al.
Veröffentlicht: (2025)
von: Li, Tianxiao, et al.
Veröffentlicht: (2025)
DiTFastAttnV2: Head-wise Attention Compression for Multi-Modality Diffusion Transformers
von: Zhang, Hanling, et al.
Veröffentlicht: (2025)
von: Zhang, Hanling, et al.
Veröffentlicht: (2025)
PointExplainer: Towards Transparent Parkinson's Disease Diagnosis
von: Wang, Xuechao, et al.
Veröffentlicht: (2025)
von: Wang, Xuechao, et al.
Veröffentlicht: (2025)
TVE: Learning Meta-attribution for Transferable Vision Explainer
von: Wang, Guanchu, et al.
Veröffentlicht: (2023)
von: Wang, Guanchu, et al.
Veröffentlicht: (2023)
RNNs, CNNs and Transformers in Human Action Recognition: A Survey and a Hybrid Model
von: Alomar, Khaled, et al.
Veröffentlicht: (2024)
von: Alomar, Khaled, et al.
Veröffentlicht: (2024)
Enhancing CNNs robustness to occlusions with bioinspired filters for border completion
von: Coutinho, Catarina P., et al.
Veröffentlicht: (2025)
von: Coutinho, Catarina P., et al.
Veröffentlicht: (2025)
Integrative CAM: Adaptive Layer Fusion for Comprehensive Interpretation of CNNs
von: Singh, Aniket K., et al.
Veröffentlicht: (2024)
von: Singh, Aniket K., et al.
Veröffentlicht: (2024)
Enhancing Osteoporosis Detection: An Explainable Multi-Modal Learning Framework with Feature Fusion and Variable Clustering
von: Chagahi, Mehdi Hosseini, et al.
Veröffentlicht: (2024)
von: Chagahi, Mehdi Hosseini, et al.
Veröffentlicht: (2024)
Head Forcing: Long Autoregressive Video Generation via Head Heterogeneity
von: Tian, Jiahao, et al.
Veröffentlicht: (2026)
von: Tian, Jiahao, et al.
Veröffentlicht: (2026)
Inspecting Explainability of Transformer Models with Additional Statistical Information
von: Nguyen, Hoang C., et al.
Veröffentlicht: (2023)
von: Nguyen, Hoang C., et al.
Veröffentlicht: (2023)
Do Transformers Understand Ancient Roman Coin Motifs Better than CNNs?
von: Reid, David, et al.
Veröffentlicht: (2026)
von: Reid, David, et al.
Veröffentlicht: (2026)
DAWN: Dynamic Frame Avatar with Non-autoregressive Diffusion Framework for Talking Head Video Generation
von: Cheng, Hanbo, et al.
Veröffentlicht: (2024)
von: Cheng, Hanbo, et al.
Veröffentlicht: (2024)
Explainable, Multi-modal Wound Infection Classification from Images Augmented with Generated Captions
von: Busaranuvong, Palawat, et al.
Veröffentlicht: (2025)
von: Busaranuvong, Palawat, et al.
Veröffentlicht: (2025)
Evaluating the Impact of Compression Techniques on the Robustness of CNNs under Natural Corruptions
von: Da Silva, Itallo Patrick Castro Alves, et al.
Veröffentlicht: (2025)
von: Da Silva, Itallo Patrick Castro Alves, et al.
Veröffentlicht: (2025)
Reliable or Deceptive? Investigating Gated Features for Smooth Visual Explanations in CNNs
von: Mitra, Soham, et al.
Veröffentlicht: (2024)
von: Mitra, Soham, et al.
Veröffentlicht: (2024)
Vehicle Classification under Extreme Imbalance: A Comparative Study of Ensemble Learning and CNNs
von: Syarubany, Abu Hanif Muhammad
Veröffentlicht: (2025)
von: Syarubany, Abu Hanif Muhammad
Veröffentlicht: (2025)
Promoting Shape Bias in CNNs: Frequency-Based and Contrastive Regularization for Corruption Robustness
von: Ranabhat, Robin Narsingh, et al.
Veröffentlicht: (2025)
von: Ranabhat, Robin Narsingh, et al.
Veröffentlicht: (2025)
CTA-Net: A CNN-Transformer Aggregation Network for Improving Multi-Scale Feature Extraction
von: Meng, Chunlei, et al.
Veröffentlicht: (2024)
von: Meng, Chunlei, et al.
Veröffentlicht: (2024)
MSVIT: Improving Spiking Vision Transformer Using Multi-scale Attention Fusion
von: Hua, Wei, et al.
Veröffentlicht: (2025)
von: Hua, Wei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
TraNCE: Transformative Non-linear Concept Explainer for CNNs
von: Akpudo, Ugochukwu Ejike, et al.
Veröffentlicht: (2025) -
Enhancing Surgical Documentation through Multimodal Visual-Temporal Transformers and Generative AI
von: Georgenthum, Hugo, et al.
Veröffentlicht: (2025) -
Less is More: The Influence of Pruning on the Explainability of CNNs
von: Merkle, Florian, et al.
Veröffentlicht: (2023) -
Automated Image Captioning with CNNs and Transformers
von: Cahyono, Joshua Adrian, et al.
Veröffentlicht: (2024) -
Evaluating the Stability of Semantic Concept Representations in CNNs for Robust Explainability
von: Mikriukov, Georgii, et al.
Veröffentlicht: (2023)