Comparing the Decision-Making Mechanisms by Transformers and CNNs via Explanation Methods
Fuente:
arXiv
Saved in:
| Main Authors: | Jiang, Mingqi, Khorram, Saeed, Fuxin, Li |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Taming the Tail in Class-Conditional GANs: Knowledge Sharing via Unconditional Training at Lower Resolutions
by: Khorram, Saeed, et al.
Published: (2024)
by: Khorram, Saeed, et al.
Published: (2024)
GS4: Generalizable Sparse Splatting Semantic SLAM
by: Jiang, Mingqi, et al.
Published: (2025)
by: Jiang, Mingqi, et al.
Published: (2025)
A Comparative Study of Vision Transformers and CNNs for Few-Shot Rigid Transformation and Fundamental Matrix Estimation
by: Kaya, Alon, et al.
Published: (2025)
by: Kaya, Alon, et al.
Published: (2025)
OA-CNNs: Omni-Adaptive Sparse CNNs for 3D Semantic Segmentation
by: Peng, Bohao, et al.
Published: (2024)
by: Peng, Bohao, et al.
Published: (2024)
Object Dynamics Modeling with Hierarchical Point Cloud-based Representations
by: Kim, Chanho, et al.
Published: (2024)
by: Kim, Chanho, et al.
Published: (2024)
Point-based Instance Completion with Scene Constraints
by: Khademi, Wesley, et al.
Published: (2025)
by: Khademi, Wesley, et al.
Published: (2025)
PointRecon: Online Point-based 3D Reconstruction via Ray-based 2D-3D Matching
by: Ziwen, Chen, et al.
Published: (2024)
by: Ziwen, Chen, et al.
Published: (2024)
Vision Transformers for Kidney Stone Image Classification: A Comparative Study with CNNs
by: Reyes-Amezcua, Ivan, et al.
Published: (2025)
by: Reyes-Amezcua, Ivan, et al.
Published: (2025)
Reliable or Deceptive? Investigating Gated Features for Smooth Visual Explanations in CNNs
by: Mitra, Soham, et al.
Published: (2024)
by: Mitra, Soham, et al.
Published: (2024)
B-cos Alignment for Inherently Interpretable CNNs and Vision Transformers
by: Böhle, Moritz, et al.
Published: (2023)
by: Böhle, Moritz, et al.
Published: (2023)
Deep Network Pruning: A Comparative Study on CNNs in Face Recognition
by: Alonso-Fernandez, Fernando, et al.
Published: (2024)
by: Alonso-Fernandez, Fernando, et al.
Published: (2024)
Automated Image Captioning with CNNs and Transformers
by: Cahyono, Joshua Adrian, et al.
Published: (2024)
by: Cahyono, Joshua Adrian, et al.
Published: (2024)
Re-Prompting SAM 3 via Object Retrieval: 3rd of the 5th PVUW MOSE Track
by: Gao, Mingqi, et al.
Published: (2026)
by: Gao, Mingqi, et al.
Published: (2026)
CNNs, Transformers, Hybrid, and Vision Language Models for Skin Cancer Detection
by: Dey, Durjoy, et al.
Published: (2026)
by: Dey, Durjoy, et al.
Published: (2026)
Zero-Shot Textual Explanations via Translating Decision-Critical Features
by: Yamauchi, Toshinori, et al.
Published: (2025)
by: Yamauchi, Toshinori, et al.
Published: (2025)
Are Explanations Helpful? A Comparative Analysis of Explainability Methods in Skin Lesion Classifiers
by: Paccotacya-Yanque, Rosa Y. G., et al.
Published: (2024)
by: Paccotacya-Yanque, Rosa Y. G., et al.
Published: (2024)
PicoEyes: Unified Gaze Estimation Framework for Mixed Reality with a Large-Scale Multi-View Dataset
by: Duan, Fuxin, et al.
Published: (2026)
by: Duan, Fuxin, et al.
Published: (2026)
Bridging the Gap: Fusing CNNs and Transformers to Decode the Elegance of Handwritten Arabic Script
by: Boufenar, Chaouki, et al.
Published: (2025)
by: Boufenar, Chaouki, et al.
Published: (2025)
Combining Transformers and CNNs for Efficient Object Detection in High-Resolution Satellite Imagery
by: Drapier, Nicolas, et al.
Published: (2025)
by: Drapier, Nicolas, et al.
Published: (2025)
Inpainting the Gaps: A Novel Framework for Evaluating Explanation Methods in Vision Transformers
by: Badisa, Lokesh, et al.
Published: (2024)
by: Badisa, Lokesh, et al.
Published: (2024)
Pruning By Explaining Revisited: Optimizing Attribution Methods to Prune CNNs and Transformers
by: Hatefi, Sayed Mohammad Vakilzadeh, et al.
Published: (2024)
by: Hatefi, Sayed Mohammad Vakilzadeh, et al.
Published: (2024)
Enhancing Robustness in Post-Processing Watermarking: An Ensemble Attack Network Using CNNs and Transformers
by: Huang, Tzuhsuan, et al.
Published: (2025)
by: Huang, Tzuhsuan, et al.
Published: (2025)
On the Faithfulness of Vision Transformer Explanations
by: Wu, Junyi, et al.
Published: (2024)
by: Wu, Junyi, et al.
Published: (2024)
Learning a Particle Dynamics Model with Real-world Videos
by: Kim, Chanho, et al.
Published: (2026)
by: Kim, Chanho, et al.
Published: (2026)
Point Linguist Model: Segment Any Object via Bridged Large 3D-Language Model
by: Huang, Zhuoxu, et al.
Published: (2025)
by: Huang, Zhuoxu, et al.
Published: (2025)
Jointly Training and Pruning CNNs via Learnable Agent Guidance and Alignment
by: Ganjdanesh, Alireza, et al.
Published: (2024)
by: Ganjdanesh, Alireza, et al.
Published: (2024)
CAPO: Reinforcing Consistent Reasoning in Medical Decision-Making
by: Jiang, Songtao, et al.
Published: (2025)
by: Jiang, Songtao, et al.
Published: (2025)
Lightweight Channel Attention for Efficient CNNs
by: Kanaparthi, Prem Babu, et al.
Published: (2026)
by: Kanaparthi, Prem Babu, et al.
Published: (2026)
A Category-Fragment Segmentation Framework for Pelvic Fracture Segmentation in X-ray Images
by: Liu, Daiqi, et al.
Published: (2025)
by: Liu, Daiqi, et al.
Published: (2025)
Long-Term 3D Point Tracking By Cost Volume Fusion
by: Nguyen, Hung, et al.
Published: (2024)
by: Nguyen, Hung, et al.
Published: (2024)
Exploring Synergistic Ensemble Learning: Uniting CNNs, MLP-Mixers, and Vision Transformers to Enhance Image Classification
by: Bashar, Mk, et al.
Published: (2025)
by: Bashar, Mk, et al.
Published: (2025)
Mind-of-Director: Multi-modal Agent-Driven Film Previsualization via Collaborative Decision-Making
by: Nan, Shufeng, et al.
Published: (2026)
by: Nan, Shufeng, et al.
Published: (2026)
Joint Point Cloud Upsampling and Cleaning with Octree-based CNNs
by: Li, Jihe, et al.
Published: (2024)
by: Li, Jihe, et al.
Published: (2024)
Improving Network Interpretability via Explanation Consistency Evaluation
by: Wu, Hefeng, et al.
Published: (2024)
by: Wu, Hefeng, et al.
Published: (2024)
Bioinspired CNNs for border completion in occluded images
by: Coutinho, Catarina P., et al.
Published: (2026)
by: Coutinho, Catarina P., et al.
Published: (2026)
CNNs for Style Transfer of Digital to Film Photography
by: Mackenzie, Pierre, et al.
Published: (2024)
by: Mackenzie, Pierre, et al.
Published: (2024)
MVPainter: Accurate and Detailed 3D Texture Generation via Multi-View Diffusion with Geometric Control
by: Shao, Mingqi, et al.
Published: (2025)
by: Shao, Mingqi, et al.
Published: (2025)
Understanding CNNs from excitations
by: Ying, Zijian, et al.
Published: (2022)
by: Ying, Zijian, et al.
Published: (2022)
Praxis-VLM: Vision-Grounded Decision Making via Text-Driven Reinforcement Learning
by: Hu, Zhe, et al.
Published: (2025)
by: Hu, Zhe, et al.
Published: (2025)
Transforming Precision: A Comparative Analysis of Vision Transformers, CNNs, and Traditional ML for Knee Osteoarthritis Severity Diagnosis
by: Apon, Tasnim Sakib, et al.
Published: (2024)
by: Apon, Tasnim Sakib, et al.
Published: (2024)
Similar Items
-
Taming the Tail in Class-Conditional GANs: Knowledge Sharing via Unconditional Training at Lower Resolutions
by: Khorram, Saeed, et al.
Published: (2024) -
GS4: Generalizable Sparse Splatting Semantic SLAM
by: Jiang, Mingqi, et al.
Published: (2025) -
A Comparative Study of Vision Transformers and CNNs for Few-Shot Rigid Transformation and Fundamental Matrix Estimation
by: Kaya, Alon, et al.
Published: (2025) -
OA-CNNs: Omni-Adaptive Sparse CNNs for 3D Semantic Segmentation
by: Peng, Bohao, et al.
Published: (2024) -
Object Dynamics Modeling with Hierarchical Point Cloud-based Representations
by: Kim, Chanho, et al.
Published: (2024)