Semi-Supervised Multi-Task Learning for Interpretable Quality As- sessment of Fundus Images

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Telesco, Lucas Gabriel, Nejamkin, Danila, Mata, Estefanía, Filizzola, Francisco, Wignall, Kevin, Troilo, Lucía Franco, Cenoz, María de los Angeles, Thompson, Melissa, Leguía, Mercedes, Larrabide, Ignacio, Orlando, José Ignacio
Natura: Preprint
Pubblicazione: 2025
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866914161243455488
author Telesco, Lucas Gabriel
Nejamkin, Danila
Mata, Estefanía
Filizzola, Francisco
Wignall, Kevin
Troilo, Lucía Franco
Cenoz, María de los Angeles
Thompson, Melissa
Leguía, Mercedes
Larrabide, Ignacio
Orlando, José Ignacio
author_facet Telesco, Lucas Gabriel
Nejamkin, Danila
Mata, Estefanía
Filizzola, Francisco
Wignall, Kevin
Troilo, Lucía Franco
Cenoz, María de los Angeles
Thompson, Melissa
Leguía, Mercedes
Larrabide, Ignacio
Orlando, José Ignacio
contents Retinal image quality assessment (RIQA) supports computer-aided diagnosis of eye diseases. However, most tools classify only overall image quality, without indicating acquisition defects to guide recapture. This gap is mainly due to the high cost of detailed annotations. In this paper, we aim to mitigate this limitation by introducing a hybrid semi-supervised learning approach that combines manual labels for overall quality with pseudo-labels of quality details within a multi-task framework. Our objective is to obtain more interpretable RIQA models without requiring extensive manual labeling. Pseudo-labels are generated by a Teacher model trained on a small dataset and then used to fine-tune a pre-trained model in a multi-task setting. Using a ResNet-18 backbone, we show that these weak annotations improve quality assessment over single-task baselines (F1: 0.875 vs. 0.863 on EyeQ, and 0.778 vs. 0.763 on DeepDRiD), matching or surpassing existing methods. The multi-task model achieved performance statistically comparable to the Teacher for most detail prediction tasks (p > 0.05). In a newly annotated EyeQ subset released with this paper, our model performed similarly to experts, suggesting that pseudo-label noise aligns with expert variability. Our main finding is that the proposed semi-supervised approach not only improves overall quality assessment but also provides interpretable feedback on capture conditions (illumination, clarity, contrast). This enhances interpretability at no extra manual labeling cost and offers clinically actionable outputs to guide image recapture.
format Preprint
id arxiv_https___arxiv_org_abs_2511_13353
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Semi-Supervised Multi-Task Learning for Interpretable Quality As- sessment of Fundus Images
Telesco, Lucas Gabriel
Nejamkin, Danila
Mata, Estefanía
Filizzola, Francisco
Wignall, Kevin
Troilo, Lucía Franco
Cenoz, María de los Angeles
Thompson, Melissa
Leguía, Mercedes
Larrabide, Ignacio
Orlando, José Ignacio
Computer Vision and Pattern Recognition
Artificial Intelligence
Retinal image quality assessment (RIQA) supports computer-aided diagnosis of eye diseases. However, most tools classify only overall image quality, without indicating acquisition defects to guide recapture. This gap is mainly due to the high cost of detailed annotations. In this paper, we aim to mitigate this limitation by introducing a hybrid semi-supervised learning approach that combines manual labels for overall quality with pseudo-labels of quality details within a multi-task framework. Our objective is to obtain more interpretable RIQA models without requiring extensive manual labeling. Pseudo-labels are generated by a Teacher model trained on a small dataset and then used to fine-tune a pre-trained model in a multi-task setting. Using a ResNet-18 backbone, we show that these weak annotations improve quality assessment over single-task baselines (F1: 0.875 vs. 0.863 on EyeQ, and 0.778 vs. 0.763 on DeepDRiD), matching or surpassing existing methods. The multi-task model achieved performance statistically comparable to the Teacher for most detail prediction tasks (p > 0.05). In a newly annotated EyeQ subset released with this paper, our model performed similarly to experts, suggesting that pseudo-label noise aligns with expert variability. Our main finding is that the proposed semi-supervised approach not only improves overall quality assessment but also provides interpretable feedback on capture conditions (illumination, clarity, contrast). This enhances interpretability at no extra manual labeling cost and offers clinically actionable outputs to guide image recapture.
title Semi-Supervised Multi-Task Learning for Interpretable Quality As- sessment of Fundus Images
topic Computer Vision and Pattern Recognition
Artificial Intelligence
url https://arxiv.org/abs/2511.13353