CONFEX: Uncertainty-Aware Counterfactual Explanations with Conformal Guarantees

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Bilkhoo, Aman, Hosseini, Mehran, Kazemi, Milad, Paoletti, Nicola
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866914108816752640
author Bilkhoo, Aman
Hosseini, Mehran
Kazemi, Milad
Paoletti, Nicola
author_facet Bilkhoo, Aman
Hosseini, Mehran
Kazemi, Milad
Paoletti, Nicola
contents Counterfactual explanations (CFXs) provide human-understandable justifications for model predictions, enabling actionable recourse and enhancing interpretability. To be reliable, CFXs must avoid regions of high predictive uncertainty, where explanations may be misleading or inapplicable. However, existing methods often neglect uncertainty or lack principled mechanisms for incorporating it with formal guarantees. We propose CONFEX, a novel method for generating uncertainty-aware counterfactual explanations using Conformal Prediction (CP) and Mixed-Integer Linear Programming (MILP). CONFEX explanations are designed to provide local coverage guarantees, addressing the issue that CFX generation violates exchangeability. To do so, we develop a novel localised CP procedure that enjoys an efficient MILP encoding by leveraging an offline tree-based partitioning of the input space. This way, CONFEX generates CFXs with rigorous guarantees on both predictive uncertainty and optimality. We evaluate CONFEX against state-of-the-art methods across diverse benchmarks and metrics, demonstrating that our uncertainty-aware approach yields robust and plausible explanations.
format Preprint
id arxiv_https___arxiv_org_abs_2510_19754
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle CONFEX: Uncertainty-Aware Counterfactual Explanations with Conformal Guarantees
Bilkhoo, Aman
Hosseini, Mehran
Kazemi, Milad
Paoletti, Nicola
Machine Learning
Counterfactual explanations (CFXs) provide human-understandable justifications for model predictions, enabling actionable recourse and enhancing interpretability. To be reliable, CFXs must avoid regions of high predictive uncertainty, where explanations may be misleading or inapplicable. However, existing methods often neglect uncertainty or lack principled mechanisms for incorporating it with formal guarantees. We propose CONFEX, a novel method for generating uncertainty-aware counterfactual explanations using Conformal Prediction (CP) and Mixed-Integer Linear Programming (MILP). CONFEX explanations are designed to provide local coverage guarantees, addressing the issue that CFX generation violates exchangeability. To do so, we develop a novel localised CP procedure that enjoys an efficient MILP encoding by leveraging an offline tree-based partitioning of the input space. This way, CONFEX generates CFXs with rigorous guarantees on both predictive uncertainty and optimality. We evaluate CONFEX against state-of-the-art methods across diverse benchmarks and metrics, demonstrating that our uncertainty-aware approach yields robust and plausible explanations.
title CONFEX: Uncertainty-Aware Counterfactual Explanations with Conformal Guarantees
topic Machine Learning
url https://arxiv.org/abs/2510.19754