Leveraging Human-Machine Interactions for Computer Vision Dataset Quality Enhancement

Fuente: arXiv
Guardado en:
Detalles Bibliográficos
Autores principales: Anzaku, Esla Timothy, Hong, Hyesoo, Park, Jin-Woo, Yang, Wonjun, Kim, Kangmin, Won, JongBum, Herath, Deshika Vinoshani Kumari, Van Messem, Arnout, De Neve, Wesley
Formato: Preprint
Publicado: 2024
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866911768360517632
author Anzaku, Esla Timothy
Hong, Hyesoo
Park, Jin-Woo
Yang, Wonjun
Kim, Kangmin
Won, JongBum
Herath, Deshika Vinoshani Kumari
Van Messem, Arnout
De Neve, Wesley
author_facet Anzaku, Esla Timothy
Hong, Hyesoo
Park, Jin-Woo
Yang, Wonjun
Kim, Kangmin
Won, JongBum
Herath, Deshika Vinoshani Kumari
Van Messem, Arnout
De Neve, Wesley
contents Large-scale datasets for single-label multi-class classification, such as \emph{ImageNet-1k}, have been instrumental in advancing deep learning and computer vision. However, a critical and often understudied aspect is the comprehensive quality assessment of these datasets, especially regarding potential multi-label annotation errors. In this paper, we introduce a lightweight, user-friendly, and scalable framework that synergizes human and machine intelligence for efficient dataset validation and quality enhancement. We term this novel framework \emph{Multilabelfy}. Central to Multilabelfy is an adaptable web-based platform that systematically guides annotators through the re-evaluation process, effectively leveraging human-machine interactions to enhance dataset quality. By using Multilabelfy on the ImageNetV2 dataset, we found that approximately $47.88\%$ of the images contained at least two labels, underscoring the need for more rigorous assessments of such influential datasets. Furthermore, our analysis showed a negative correlation between the number of potential labels per image and model top-1 accuracy, illuminating a crucial factor in model evaluation and selection. Our open-source framework, Multilabelfy, offers a convenient, lightweight solution for dataset enhancement, emphasizing multi-label proportions. This study tackles major challenges in dataset integrity and provides key insights into model performance evaluation. Moreover, it underscores the advantages of integrating human expertise with machine capabilities to produce more robust models and trustworthy data development. The source code for Multilabelfy will be available at https://github.com/esla/Multilabelfy. \keywords{Computer Vision \and Dataset Quality Enhancement \and Dataset Validation \and Human-Computer Interaction \and Multi-label Annotation.}
format Preprint
id arxiv_https___arxiv_org_abs_2401_17736
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Leveraging Human-Machine Interactions for Computer Vision Dataset Quality Enhancement
Anzaku, Esla Timothy
Hong, Hyesoo
Park, Jin-Woo
Yang, Wonjun
Kim, Kangmin
Won, JongBum
Herath, Deshika Vinoshani Kumari
Van Messem, Arnout
De Neve, Wesley
Computer Vision and Pattern Recognition
Large-scale datasets for single-label multi-class classification, such as \emph{ImageNet-1k}, have been instrumental in advancing deep learning and computer vision. However, a critical and often understudied aspect is the comprehensive quality assessment of these datasets, especially regarding potential multi-label annotation errors. In this paper, we introduce a lightweight, user-friendly, and scalable framework that synergizes human and machine intelligence for efficient dataset validation and quality enhancement. We term this novel framework \emph{Multilabelfy}. Central to Multilabelfy is an adaptable web-based platform that systematically guides annotators through the re-evaluation process, effectively leveraging human-machine interactions to enhance dataset quality. By using Multilabelfy on the ImageNetV2 dataset, we found that approximately $47.88\%$ of the images contained at least two labels, underscoring the need for more rigorous assessments of such influential datasets. Furthermore, our analysis showed a negative correlation between the number of potential labels per image and model top-1 accuracy, illuminating a crucial factor in model evaluation and selection. Our open-source framework, Multilabelfy, offers a convenient, lightweight solution for dataset enhancement, emphasizing multi-label proportions. This study tackles major challenges in dataset integrity and provides key insights into model performance evaluation. Moreover, it underscores the advantages of integrating human expertise with machine capabilities to produce more robust models and trustworthy data development. The source code for Multilabelfy will be available at https://github.com/esla/Multilabelfy. \keywords{Computer Vision \and Dataset Quality Enhancement \and Dataset Validation \and Human-Computer Interaction \and Multi-label Annotation.}
title Leveraging Human-Machine Interactions for Computer Vision Dataset Quality Enhancement
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2401.17736