A principled approach for comparing Variable Importance

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Reyero-Lobo, Angel, Neuvial, Pierre, Thirion, Bertrand
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866908551325155328
author Reyero-Lobo, Angel
Neuvial, Pierre
Thirion, Bertrand
author_facet Reyero-Lobo, Angel
Neuvial, Pierre
Thirion, Bertrand
contents Variable importance measures (VIMs) aim to quantify the contribution of each input covariate to the predictability of a given output. With the growing interest in explainable AI, numerous VIMs have been proposed, many of which are heuristic in nature. This is often justified by the inherent subjectivity of the notion of importance. This raises important questions regarding usage: What makes a good VIM? How can we compare different VIMs? In this paper, we address these questions by: (1) proposing an axiomatic framework that bridges the gap between variable importance and variable selection. This framework formalizes the intuitive principle that features providing no additional information should not be assigned importance. It helps avoid false positives due to spurious correlations, which can arise with popular methods such as Shapley values; and (2) introducing a general pipeline for constructing VIMs, which clarifies the objective of various VIMs and thus facilitates meaningful comparisons. This approach is natural in statistics, but the literature has diverged from it. Finally, we provide an extensive set of examples to guide practitioners in selecting and estimating appropriate indices aligned with their specific goals and data.
format Preprint
id arxiv_https___arxiv_org_abs_2507_17306
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle A principled approach for comparing Variable Importance
Reyero-Lobo, Angel
Neuvial, Pierre
Thirion, Bertrand
Methodology
Statistics Theory
Variable importance measures (VIMs) aim to quantify the contribution of each input covariate to the predictability of a given output. With the growing interest in explainable AI, numerous VIMs have been proposed, many of which are heuristic in nature. This is often justified by the inherent subjectivity of the notion of importance. This raises important questions regarding usage: What makes a good VIM? How can we compare different VIMs? In this paper, we address these questions by: (1) proposing an axiomatic framework that bridges the gap between variable importance and variable selection. This framework formalizes the intuitive principle that features providing no additional information should not be assigned importance. It helps avoid false positives due to spurious correlations, which can arise with popular methods such as Shapley values; and (2) introducing a general pipeline for constructing VIMs, which clarifies the objective of various VIMs and thus facilitates meaningful comparisons. This approach is natural in statistics, but the literature has diverged from it. Finally, we provide an extensive set of examples to guide practitioners in selecting and estimating appropriate indices aligned with their specific goals and data.
title A principled approach for comparing Variable Importance
topic Methodology
Statistics Theory
url https://arxiv.org/abs/2507.17306