KPCA-CAM: Visual Explainability of Deep Computer Vision Models using Kernel PCA

Fuente: arXiv
Guardado en:
Detalles Bibliográficos
Autores principales: Karmani, Sachin, Sivakaran, Thanushon, Prasad, Gaurav, Ali, Mehmet, Yang, Wenbo, Tang, Sheyang
Formato: Preprint
Publicado: 2024
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866916417281982464
author Karmani, Sachin
Sivakaran, Thanushon
Prasad, Gaurav
Ali, Mehmet
Yang, Wenbo
Tang, Sheyang
author_facet Karmani, Sachin
Sivakaran, Thanushon
Prasad, Gaurav
Ali, Mehmet
Yang, Wenbo
Tang, Sheyang
contents Deep learning models often function as black boxes, providing no straightforward reasoning for their predictions. This is particularly true for computer vision models, which process tensors of pixel values to generate outcomes in tasks such as image classification and object detection. To elucidate the reasoning of these models, class activation maps (CAMs) are used to highlight salient regions that influence a model's output. This research introduces KPCA-CAM, a technique designed to enhance the interpretability of Convolutional Neural Networks (CNNs) through improved class activation maps. KPCA-CAM leverages Principal Component Analysis (PCA) with the kernel trick to capture nonlinear relationships within CNN activations more effectively. By mapping data into higher-dimensional spaces with kernel functions and extracting principal components from this transformed hyperplane, KPCA-CAM provides more accurate representations of the underlying data manifold. This enables a deeper understanding of the features influencing CNN decisions. Empirical evaluations on the ILSVRC dataset across different CNN models demonstrate that KPCA-CAM produces more precise activation maps, providing clearer insights into the model's reasoning compared to existing CAM algorithms. This research advances CAM techniques, equipping researchers and practitioners with a powerful tool to gain deeper insights into CNN decision-making processes and overall behaviors.
format Preprint
id arxiv_https___arxiv_org_abs_2410_00267
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle KPCA-CAM: Visual Explainability of Deep Computer Vision Models using Kernel PCA
Karmani, Sachin
Sivakaran, Thanushon
Prasad, Gaurav
Ali, Mehmet
Yang, Wenbo
Tang, Sheyang
Computer Vision and Pattern Recognition
Artificial Intelligence
Deep learning models often function as black boxes, providing no straightforward reasoning for their predictions. This is particularly true for computer vision models, which process tensors of pixel values to generate outcomes in tasks such as image classification and object detection. To elucidate the reasoning of these models, class activation maps (CAMs) are used to highlight salient regions that influence a model's output. This research introduces KPCA-CAM, a technique designed to enhance the interpretability of Convolutional Neural Networks (CNNs) through improved class activation maps. KPCA-CAM leverages Principal Component Analysis (PCA) with the kernel trick to capture nonlinear relationships within CNN activations more effectively. By mapping data into higher-dimensional spaces with kernel functions and extracting principal components from this transformed hyperplane, KPCA-CAM provides more accurate representations of the underlying data manifold. This enables a deeper understanding of the features influencing CNN decisions. Empirical evaluations on the ILSVRC dataset across different CNN models demonstrate that KPCA-CAM produces more precise activation maps, providing clearer insights into the model's reasoning compared to existing CAM algorithms. This research advances CAM techniques, equipping researchers and practitioners with a powerful tool to gain deeper insights into CNN decision-making processes and overall behaviors.
title KPCA-CAM: Visual Explainability of Deep Computer Vision Models using Kernel PCA
topic Computer Vision and Pattern Recognition
Artificial Intelligence
url https://arxiv.org/abs/2410.00267