Quality Issues in Machine Learning Software Systems

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Côté, Pierre-Olivier, Nikanjam, Amin, Bouchoucha, Rached, Basta, Ilan, Abidi, Mouna, Khomh, Foutse
Format: Preprint
Published: 2023
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866929447387529216
author Côté, Pierre-Olivier
Nikanjam, Amin
Bouchoucha, Rached
Basta, Ilan
Abidi, Mouna
Khomh, Foutse
author_facet Côté, Pierre-Olivier
Nikanjam, Amin
Bouchoucha, Rached
Basta, Ilan
Abidi, Mouna
Khomh, Foutse
contents Context: An increasing demand is observed in various domains to employ Machine Learning (ML) for solving complex problems. ML models are implemented as software components and deployed in Machine Learning Software Systems (MLSSs). Problem: There is a strong need for ensuring the serving quality of MLSSs. False or poor decisions of such systems can lead to malfunction of other systems, significant financial losses, or even threats to human life. The quality assurance of MLSSs is considered a challenging task and currently is a hot research topic. Objective: This paper aims to investigate the characteristics of real quality issues in MLSSs from the viewpoint of practitioners. This empirical study aims to identify a catalog of quality issues in MLSSs. Method: We conduct a set of interviews with practitioners/experts, to gather insights about their experience and practices when dealing with quality issues. We validate the identified quality issues via a survey with ML practitioners. Results: Based on the content of 37 interviews, we identified 18 recurring quality issues and 21 strategies to mitigate them. For each identified issue, we describe the causes and consequences according to the practitioners' experience. Conclusion: We believe the catalog of issues developed in this study will allow the community to develop efficient quality assurance tools for ML models and MLSSs. A replication package of our study is available on our public GitHub repository
format Preprint
id arxiv_https___arxiv_org_abs_2306_15007
institution arXiv
publishDate 2023
record_format arxiv
spellingShingle Quality Issues in Machine Learning Software Systems
Côté, Pierre-Olivier
Nikanjam, Amin
Bouchoucha, Rached
Basta, Ilan
Abidi, Mouna
Khomh, Foutse
Software Engineering
Machine Learning
68T05
Context: An increasing demand is observed in various domains to employ Machine Learning (ML) for solving complex problems. ML models are implemented as software components and deployed in Machine Learning Software Systems (MLSSs). Problem: There is a strong need for ensuring the serving quality of MLSSs. False or poor decisions of such systems can lead to malfunction of other systems, significant financial losses, or even threats to human life. The quality assurance of MLSSs is considered a challenging task and currently is a hot research topic. Objective: This paper aims to investigate the characteristics of real quality issues in MLSSs from the viewpoint of practitioners. This empirical study aims to identify a catalog of quality issues in MLSSs. Method: We conduct a set of interviews with practitioners/experts, to gather insights about their experience and practices when dealing with quality issues. We validate the identified quality issues via a survey with ML practitioners. Results: Based on the content of 37 interviews, we identified 18 recurring quality issues and 21 strategies to mitigate them. For each identified issue, we describe the causes and consequences according to the practitioners' experience. Conclusion: We believe the catalog of issues developed in this study will allow the community to develop efficient quality assurance tools for ML models and MLSSs. A replication package of our study is available on our public GitHub repository
title Quality Issues in Machine Learning Software Systems
topic Software Engineering
Machine Learning
68T05
url https://arxiv.org/abs/2306.15007