Addressing Bias in VLMs for Glaucoma Detection Without Protected Attribute Supervision

Fuente: arXiv
Guardado en:
Detalles Bibliográficos
Autores principales: Akash, Ahsan Habib, Murray, Greg, Amireskandari, Annahita, Palko, Joel, Laxson, Carol, Bhattarai, Binod, Gyawali, Prashnna
Formato: Preprint
Publicado: 2025
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866913986556985344
author Akash, Ahsan Habib
Murray, Greg
Amireskandari, Annahita
Palko, Joel
Laxson, Carol
Bhattarai, Binod
Gyawali, Prashnna
author_facet Akash, Ahsan Habib
Murray, Greg
Amireskandari, Annahita
Palko, Joel
Laxson, Carol
Bhattarai, Binod
Gyawali, Prashnna
contents Vision-Language Models (VLMs) have achieved remarkable success on multimodal tasks such as image-text retrieval and zero-shot classification, yet they can exhibit demographic biases even when explicit protected attributes are absent during training. In this work, we focus on automated glaucoma screening from retinal fundus images, a critical application given that glaucoma is a leading cause of irreversible blindness and disproportionately affects underserved populations. Building on a reweighting-based contrastive learning framework, we introduce an attribute-agnostic debiasing method that (i) infers proxy subgroups via unsupervised clustering of image-image embeddings, (ii) computes gradient-similarity weights between the CLIP-style multimodal loss and a SimCLR-style image-pair contrastive loss, and (iii) applies these weights in a joint, top-$k$ weighted objective to upweight underperforming clusters. This label-free approach adaptively targets the hardest examples, thereby reducing subgroup disparities. We evaluate our method on the Harvard FairVLMed glaucoma subset, reporting Equalized Odds Distance (EOD), Equalized Subgroup AUC (ES AUC), and Groupwise AUC to demonstrate equitable performance across inferred demographic subgroups.
format Preprint
id arxiv_https___arxiv_org_abs_2508_09087
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Addressing Bias in VLMs for Glaucoma Detection Without Protected Attribute Supervision
Akash, Ahsan Habib
Murray, Greg
Amireskandari, Annahita
Palko, Joel
Laxson, Carol
Bhattarai, Binod
Gyawali, Prashnna
Computer Vision and Pattern Recognition
Vision-Language Models (VLMs) have achieved remarkable success on multimodal tasks such as image-text retrieval and zero-shot classification, yet they can exhibit demographic biases even when explicit protected attributes are absent during training. In this work, we focus on automated glaucoma screening from retinal fundus images, a critical application given that glaucoma is a leading cause of irreversible blindness and disproportionately affects underserved populations. Building on a reweighting-based contrastive learning framework, we introduce an attribute-agnostic debiasing method that (i) infers proxy subgroups via unsupervised clustering of image-image embeddings, (ii) computes gradient-similarity weights between the CLIP-style multimodal loss and a SimCLR-style image-pair contrastive loss, and (iii) applies these weights in a joint, top-$k$ weighted objective to upweight underperforming clusters. This label-free approach adaptively targets the hardest examples, thereby reducing subgroup disparities. We evaluate our method on the Harvard FairVLMed glaucoma subset, reporting Equalized Odds Distance (EOD), Equalized Subgroup AUC (ES AUC), and Groupwise AUC to demonstrate equitable performance across inferred demographic subgroups.
title Addressing Bias in VLMs for Glaucoma Detection Without Protected Attribute Supervision
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2508.09087