Flexible and Efficient Drift Detection without Labels

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Tan, Nelvin, Shih, Yu-Ching, Yang, Dong, Salunkhe, Amol
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866915520783056896
author Tan, Nelvin
Shih, Yu-Ching
Yang, Dong
Salunkhe, Amol
author_facet Tan, Nelvin
Shih, Yu-Ching
Yang, Dong
Salunkhe, Amol
contents Machine learning models are being increasingly used to automate decisions in almost every domain, and ensuring the performance of these models is crucial for ensuring high quality machine learning enabled services. Ensuring concept drift is detected early is thus of the highest importance. A lot of research on concept drift has focused on the supervised case that assumes the true labels of supervised tasks are available immediately after making predictions. Controlling for false positives while monitoring the performance of predictive models used to make inference from extremely large datasets periodically, where the true labels are not instantly available, becomes extremely challenging. We propose a flexible and efficient concept drift detection algorithm that uses classical statistical process control in a label-less setting to accurately detect concept drifts. We show empirically that under computational constraints, our approach has better statistical power than previous known methods. Furthermore, we introduce a new semi-supervised drift detection framework to model the scenario of detecting drift (without labels) given prior detections, and show how our drift detection algorithm can be incorporated effectively into this framework. We demonstrate promising performance via numerical simulations.
format Preprint
id arxiv_https___arxiv_org_abs_2506_08734
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Flexible and Efficient Drift Detection without Labels
Tan, Nelvin
Shih, Yu-Ching
Yang, Dong
Salunkhe, Amol
Machine Learning
Applications
Machine learning models are being increasingly used to automate decisions in almost every domain, and ensuring the performance of these models is crucial for ensuring high quality machine learning enabled services. Ensuring concept drift is detected early is thus of the highest importance. A lot of research on concept drift has focused on the supervised case that assumes the true labels of supervised tasks are available immediately after making predictions. Controlling for false positives while monitoring the performance of predictive models used to make inference from extremely large datasets periodically, where the true labels are not instantly available, becomes extremely challenging. We propose a flexible and efficient concept drift detection algorithm that uses classical statistical process control in a label-less setting to accurately detect concept drifts. We show empirically that under computational constraints, our approach has better statistical power than previous known methods. Furthermore, we introduce a new semi-supervised drift detection framework to model the scenario of detecting drift (without labels) given prior detections, and show how our drift detection algorithm can be incorporated effectively into this framework. We demonstrate promising performance via numerical simulations.
title Flexible and Efficient Drift Detection without Labels
topic Machine Learning
Applications
url https://arxiv.org/abs/2506.08734