Decision from Suboptimal Classifiers: Excess Risk Pre- and Post-Calibration

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Perez-Lebel, Alexandre, Varoquaux, Gael, Koyejo, Sanmi, Doutreligne, Matthieu, Morvan, Marine Le
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866913753926205440
author Perez-Lebel, Alexandre
Varoquaux, Gael
Koyejo, Sanmi
Doutreligne, Matthieu
Morvan, Marine Le
author_facet Perez-Lebel, Alexandre
Varoquaux, Gael
Koyejo, Sanmi
Doutreligne, Matthieu
Morvan, Marine Le
contents Probabilistic classifiers are central for making informed decisions under uncertainty. Based on the maximum expected utility principle, optimal decision rules can be derived using the posterior class probabilities and misclassification costs. Yet, in practice only learned approximations of the oracle posterior probabilities are available. In this work, we quantify the excess risk (a.k.a. regret) incurred using approximate posterior probabilities in batch binary decision-making. We provide analytical expressions for miscalibration-induced regret ($R^{\mathrm{CL}}$), as well as tight and informative upper and lower bounds on the regret of calibrated classifiers ($R^{\mathrm{GL}}$). These expressions allow us to identify regimes where recalibration alone addresses most of the regret, and regimes where the regret is dominated by the grouping loss, which calls for post-training beyond recalibration. Crucially, both $R^{\mathrm{CL}}$ and $R^{\mathrm{GL}}$ can be estimated in practice using a calibration curve and a recent grouping loss estimator. On NLP experiments, we show that these quantities identify when the expected gain of more advanced post-training is worth the operational cost. Finally, we highlight the potential of multicalibration approaches as efficient alternatives to costlier fine-tuning approaches.
format Preprint
id arxiv_https___arxiv_org_abs_2503_18025
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Decision from Suboptimal Classifiers: Excess Risk Pre- and Post-Calibration
Perez-Lebel, Alexandre
Varoquaux, Gael
Koyejo, Sanmi
Doutreligne, Matthieu
Morvan, Marine Le
Machine Learning
Artificial Intelligence
Probabilistic classifiers are central for making informed decisions under uncertainty. Based on the maximum expected utility principle, optimal decision rules can be derived using the posterior class probabilities and misclassification costs. Yet, in practice only learned approximations of the oracle posterior probabilities are available. In this work, we quantify the excess risk (a.k.a. regret) incurred using approximate posterior probabilities in batch binary decision-making. We provide analytical expressions for miscalibration-induced regret ($R^{\mathrm{CL}}$), as well as tight and informative upper and lower bounds on the regret of calibrated classifiers ($R^{\mathrm{GL}}$). These expressions allow us to identify regimes where recalibration alone addresses most of the regret, and regimes where the regret is dominated by the grouping loss, which calls for post-training beyond recalibration. Crucially, both $R^{\mathrm{CL}}$ and $R^{\mathrm{GL}}$ can be estimated in practice using a calibration curve and a recent grouping loss estimator. On NLP experiments, we show that these quantities identify when the expected gain of more advanced post-training is worth the operational cost. Finally, we highlight the potential of multicalibration approaches as efficient alternatives to costlier fine-tuning approaches.
title Decision from Suboptimal Classifiers: Excess Risk Pre- and Post-Calibration
topic Machine Learning
Artificial Intelligence
url https://arxiv.org/abs/2503.18025