Learning from Label Proportions and Covariate-shifted Instances

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Singh, Sagalpreet, Sharma, Navodita, Havaldar, Shreyas, Saket, Rishi, Raghuveer, Aravindan
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866913941554200576
author Singh, Sagalpreet
Sharma, Navodita
Havaldar, Shreyas
Saket, Rishi
Raghuveer, Aravindan
author_facet Singh, Sagalpreet
Sharma, Navodita
Havaldar, Shreyas
Saket, Rishi
Raghuveer, Aravindan
contents In many applications, especially due to lack of supervision or privacy concerns, the training data is grouped into bags of instances (feature-vectors) and for each bag we have only an aggregate label derived from the instance-labels in the bag. In learning from label proportions (LLP) the aggregate label is the average of the instance-labels in a bag, and a significant body of work has focused on training models in the LLP setting to predict instance-labels. In practice however, the training data may have fully supervised albeit covariate-shifted source data, along with the usual target data with bag-labels, and we wish to train a good instance-level predictor on the target domain. We call this the covariate-shifted hybrid LLP problem. Fully supervised covariate shifted data often has useful training signals and the goal is to leverage them for better predictive performance in the hybrid LLP setting. To achieve this, we develop methods for hybrid LLP which naturally incorporate the target bag-labels along with the source instance-labels, in the domain adaptation framework. Apart from proving theoretical guarantees bounding the target generalization error, we also conduct experiments on several publicly available datasets showing that our methods outperform LLP and domain adaptation baselines as well techniques from previous related work.
format Preprint
id arxiv_https___arxiv_org_abs_2411_12334
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Learning from Label Proportions and Covariate-shifted Instances
Singh, Sagalpreet
Sharma, Navodita
Havaldar, Shreyas
Saket, Rishi
Raghuveer, Aravindan
Machine Learning
In many applications, especially due to lack of supervision or privacy concerns, the training data is grouped into bags of instances (feature-vectors) and for each bag we have only an aggregate label derived from the instance-labels in the bag. In learning from label proportions (LLP) the aggregate label is the average of the instance-labels in a bag, and a significant body of work has focused on training models in the LLP setting to predict instance-labels. In practice however, the training data may have fully supervised albeit covariate-shifted source data, along with the usual target data with bag-labels, and we wish to train a good instance-level predictor on the target domain. We call this the covariate-shifted hybrid LLP problem. Fully supervised covariate shifted data often has useful training signals and the goal is to leverage them for better predictive performance in the hybrid LLP setting. To achieve this, we develop methods for hybrid LLP which naturally incorporate the target bag-labels along with the source instance-labels, in the domain adaptation framework. Apart from proving theoretical guarantees bounding the target generalization error, we also conduct experiments on several publicly available datasets showing that our methods outperform LLP and domain adaptation baselines as well techniques from previous related work.
title Learning from Label Proportions and Covariate-shifted Instances
topic Machine Learning
url https://arxiv.org/abs/2411.12334