Dilated Convolution with Learnable Spacings makes visual models more aligned with humans: a Grad-CAM study
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chamas, Rabih, Khalfaoui-Hassani, Ismail, Masquelier, Timothee |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning Delays in Spiking Neural Networks using Dilated Convolutions with Learnable Spacings
von: Hammouamri, Ilyass, et al.
Veröffentlicht: (2023)
von: Hammouamri, Ilyass, et al.
Veröffentlicht: (2023)
Dilated Convolution with Learnable Spacings
von: Khalfaoui-Hassani, Ismail
Veröffentlicht: (2024)
von: Khalfaoui-Hassani, Ismail
Veröffentlicht: (2024)
Polynomial, trigonometric, and tropical activations
von: Khalfaoui-Hassani, Ismail, et al.
Veröffentlicht: (2025)
von: Khalfaoui-Hassani, Ismail, et al.
Veröffentlicht: (2025)
Self-supervised video pretraining yields robust and more human-aligned visual representations
von: Parthasarathy, Nikhil, et al.
Veröffentlicht: (2022)
von: Parthasarathy, Nikhil, et al.
Veröffentlicht: (2022)
An Attention Infused Deep Learning System with Grad-CAM Visualization for Early Screening of Glaucoma
von: Swaminathan, Ramanathan
Veröffentlicht: (2025)
von: Swaminathan, Ramanathan
Veröffentlicht: (2025)
Weakly Supervised Pneumonia Localization from Chest X-Rays Using Deep Neural Network and Grad-CAM Explanations
von: Shahi, Kiran, et al.
Veröffentlicht: (2025)
von: Shahi, Kiran, et al.
Veröffentlicht: (2025)
Toward Reliable and Explainable Nail Disease Classification: Leveraging Adversarial Training and Grad-CAM Visualization
von: Hossain, Farzia, et al.
Veröffentlicht: (2026)
von: Hossain, Farzia, et al.
Veröffentlicht: (2026)
RapidNet: Multi-Level Dilated Convolution Based Mobile Backbone
von: Munir, Mustafa, et al.
Veröffentlicht: (2024)
von: Munir, Mustafa, et al.
Veröffentlicht: (2024)
A Clinically Interpretable Deep CNN Framework for Early Chronic Kidney Disease Prediction Using Grad-CAM-Based Explainable AI
von: Ayub, Anas Bin, et al.
Veröffentlicht: (2025)
von: Ayub, Anas Bin, et al.
Veröffentlicht: (2025)
Enhancing Satellite Object Localization with Dilated Convolutions and Attention-aided Spatial Pooling
von: Mostafa, Seraj Al Mahmud, et al.
Veröffentlicht: (2025)
von: Mostafa, Seraj Al Mahmud, et al.
Veröffentlicht: (2025)
Vision Transformer attention alignment with human visual perception in aesthetic object evaluation
von: Carrasco, Miguel, et al.
Veröffentlicht: (2025)
von: Carrasco, Miguel, et al.
Veröffentlicht: (2025)
Enhanced SegNet with Integrated Grad-CAM for Interpretable Retinal Layer Segmentation in OCT Images
von: Saky, S M Asiful Islam, et al.
Veröffentlicht: (2025)
von: Saky, S M Asiful Islam, et al.
Veröffentlicht: (2025)
Explainable Image Similarity: Integrating Siamese Networks and Grad-CAM
von: Livieris, Ioannis E., et al.
Veröffentlicht: (2023)
von: Livieris, Ioannis E., et al.
Veröffentlicht: (2023)
Quantifying the human visual exposome with vision language models
von: Rominger, Christian, et al.
Veröffentlicht: (2026)
von: Rominger, Christian, et al.
Veröffentlicht: (2026)
Reimagining Linear Probing: Kolmogorov-Arnold Networks in Transfer Learning
von: Shen, Sheng, et al.
Veröffentlicht: (2024)
von: Shen, Sheng, et al.
Veröffentlicht: (2024)
DilateQuant: Accurate and Efficient Diffusion Quantization via Weight Dilation
von: Liu, Xuewen, et al.
Veröffentlicht: (2024)
von: Liu, Xuewen, et al.
Veröffentlicht: (2024)
A Cascaded Dilated Convolution Approach for Mpox Lesion Classification
von: Deshmukh, Ayush
Veröffentlicht: (2024)
von: Deshmukh, Ayush
Veröffentlicht: (2024)
MARS: Paying more attention to visual attributes for text-based person search
von: Ergasti, Alex, et al.
Veröffentlicht: (2024)
von: Ergasti, Alex, et al.
Veröffentlicht: (2024)
DeltaSpace: A Semantic-aligned Feature Space for Flexible Text-guided Image Editing
von: Lyu, Yueming, et al.
Veröffentlicht: (2023)
von: Lyu, Yueming, et al.
Veröffentlicht: (2023)
How to Evaluate and Refine your CAM
von: Domeniconi, Luca, et al.
Veröffentlicht: (2026)
von: Domeniconi, Luca, et al.
Veröffentlicht: (2026)
Towards aligned body representations in vision models
von: Gizdov, Andrey, et al.
Veröffentlicht: (2025)
von: Gizdov, Andrey, et al.
Veröffentlicht: (2025)
Minimal Sufficient Views: A DNN model making predictions with more evidence has higher accuracy
von: Kawano, Keisuke, et al.
Veröffentlicht: (2024)
von: Kawano, Keisuke, et al.
Veröffentlicht: (2024)
Data-Free Group-Wise Fully Quantized Winograd Convolution via Learnable Scales
von: Pan, Shuokai, et al.
Veröffentlicht: (2024)
von: Pan, Shuokai, et al.
Veröffentlicht: (2024)
Explaining generative diffusion models via visual analysis for interpretable decision-making process
von: Park, Ji-Hoon, et al.
Veröffentlicht: (2024)
von: Park, Ji-Hoon, et al.
Veröffentlicht: (2024)
Generalizing GradCAM for Embedding Networks
von: Bachhawat, Mudit
Veröffentlicht: (2024)
von: Bachhawat, Mudit
Veröffentlicht: (2024)
Improving generalization by mimicking the human visual diet
von: Madan, Spandan, et al.
Veröffentlicht: (2022)
von: Madan, Spandan, et al.
Veröffentlicht: (2022)
Med-CAM: Minimal Evidence for Explaining Medical Decision Making
von: Suhail, Pirzada, et al.
Veröffentlicht: (2026)
von: Suhail, Pirzada, et al.
Veröffentlicht: (2026)
Integrative CAM: Adaptive Layer Fusion for Comprehensive Interpretation of CNNs
von: Singh, Aniket K., et al.
Veröffentlicht: (2024)
von: Singh, Aniket K., et al.
Veröffentlicht: (2024)
CAM-VFD: Cross-Attention Multimodal Video Forgery Detection
von: Elkhodary, Hoda Osama, et al.
Veröffentlicht: (2026)
von: Elkhodary, Hoda Osama, et al.
Veröffentlicht: (2026)
Can visual language models resolve textual ambiguity with visual cues? Let visual puns tell you!
von: Chung, Jiwan, et al.
Veröffentlicht: (2024)
von: Chung, Jiwan, et al.
Veröffentlicht: (2024)
Finer-CAM: Spotting the Difference Reveals Finer Details for Visual Explanation
von: Zhang, Ziheng, et al.
Veröffentlicht: (2025)
von: Zhang, Ziheng, et al.
Veröffentlicht: (2025)
Prompt-CAM: Making Vision Transformers Interpretable for Fine-Grained Analysis
von: Chowdhury, Arpita, et al.
Veröffentlicht: (2025)
von: Chowdhury, Arpita, et al.
Veröffentlicht: (2025)
Feature CAM: Interpretable AI in Image Classification
von: Clement, Frincy, et al.
Veröffentlicht: (2024)
von: Clement, Frincy, et al.
Veröffentlicht: (2024)
Implicit Deformable Medical Image Registration with Learnable Kernels
von: Fogarollo, Stefano, et al.
Veröffentlicht: (2025)
von: Fogarollo, Stefano, et al.
Veröffentlicht: (2025)
Deep learning models are vulnerable, but adversarial examples are even more vulnerable
von: Li, Jun, et al.
Veröffentlicht: (2025)
von: Li, Jun, et al.
Veröffentlicht: (2025)
PDM-SSD: Single-Stage Three-Dimensional Object Detector With Point Dilation
von: Liang, Ao, et al.
Veröffentlicht: (2025)
von: Liang, Ao, et al.
Veröffentlicht: (2025)
CAM-Seg: A Continuous-valued Embedding Approach for Semantic Image Generation
von: Ahmed, Masud, et al.
Veröffentlicht: (2025)
von: Ahmed, Masud, et al.
Veröffentlicht: (2025)
EmoCAM: Toward Understanding What Drives CNN-based Emotion Recognition
von: Doulfoukar, Youssef, et al.
Veröffentlicht: (2024)
von: Doulfoukar, Youssef, et al.
Veröffentlicht: (2024)
Attention Guided CAM: Visual Explanations of Vision Transformer Guided by Self-Attention
von: Leem, Saebom, et al.
Veröffentlicht: (2024)
von: Leem, Saebom, et al.
Veröffentlicht: (2024)
VLA-Mark: A cross modal watermark for large vision-language alignment model
von: Liu, Shuliang, et al.
Veröffentlicht: (2025)
von: Liu, Shuliang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Learning Delays in Spiking Neural Networks using Dilated Convolutions with Learnable Spacings
von: Hammouamri, Ilyass, et al.
Veröffentlicht: (2023) -
Dilated Convolution with Learnable Spacings
von: Khalfaoui-Hassani, Ismail
Veröffentlicht: (2024) -
Polynomial, trigonometric, and tropical activations
von: Khalfaoui-Hassani, Ismail, et al.
Veröffentlicht: (2025) -
Self-supervised video pretraining yields robust and more human-aligned visual representations
von: Parthasarathy, Nikhil, et al.
Veröffentlicht: (2022) -
An Attention Infused Deep Learning System with Grad-CAM Visualization for Early Screening of Glaucoma
von: Swaminathan, Ramanathan
Veröffentlicht: (2025)