DeepFRI Demystified: Interpretability vs. Accuracy in AI Protein Function Prediction

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Krishna, Ananya, Simon, Valentina, Kohli, Arjan
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866915645243785216
author Krishna, Ananya
Simon, Valentina
Kohli, Arjan
author_facet Krishna, Ananya
Simon, Valentina
Kohli, Arjan
contents Machine learning technologies for protein function prediction are black box models. Despite their potential to identify key drug targets with high accuracy and accelerate therapy development, the adoption of these methods depends on verifying their findings. This study evaluates DeepFRI, a leading Graph Convolutional Network (GCN) based tool, using advanced explainability techniques (GradCAM, Excitation Backpropagation, and PGExplainer) and adversarial robustness tests. Our findings reveal that the model's predictions often prioritize conserved motifs over truly deterministic residues, complicating the identification of functional sites. Quantitative analyses show that explainability methods differ significantly in granularity, with GradCAM providing broad relevance and PGExplainer pinpointing specific active sites. These results highlight tradeoffs between accuracy and interpretability, suggesting areas for improvement in DeepFRI's architecture to enhance its trustworthiness in drug discovery and regulatory settings.
format Preprint
id arxiv_https___arxiv_org_abs_2512_00642
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle DeepFRI Demystified: Interpretability vs. Accuracy in AI Protein Function Prediction
Krishna, Ananya
Simon, Valentina
Kohli, Arjan
Biomolecules
Machine learning technologies for protein function prediction are black box models. Despite their potential to identify key drug targets with high accuracy and accelerate therapy development, the adoption of these methods depends on verifying their findings. This study evaluates DeepFRI, a leading Graph Convolutional Network (GCN) based tool, using advanced explainability techniques (GradCAM, Excitation Backpropagation, and PGExplainer) and adversarial robustness tests. Our findings reveal that the model's predictions often prioritize conserved motifs over truly deterministic residues, complicating the identification of functional sites. Quantitative analyses show that explainability methods differ significantly in granularity, with GradCAM providing broad relevance and PGExplainer pinpointing specific active sites. These results highlight tradeoffs between accuracy and interpretability, suggesting areas for improvement in DeepFRI's architecture to enhance its trustworthiness in drug discovery and regulatory settings.
title DeepFRI Demystified: Interpretability vs. Accuracy in AI Protein Function Prediction
topic Biomolecules
url https://arxiv.org/abs/2512.00642