Documenting Ethical Considerations in Open Source AI Models

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Gao, Haoyu, Zahedi, Mansooreh, Treude, Christoph, Rosenstock, Sarita, Cheong, Marc
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866916310212935680
author Gao, Haoyu
Zahedi, Mansooreh
Treude, Christoph
Rosenstock, Sarita
Cheong, Marc
author_facet Gao, Haoyu
Zahedi, Mansooreh
Treude, Christoph
Rosenstock, Sarita
Cheong, Marc
contents Background: The development of AI-enabled software heavily depends on AI model documentation, such as model cards, due to different domain expertise between software engineers and model developers. From an ethical standpoint, AI model documentation conveys critical information on ethical considerations along with mitigation strategies for downstream developers to ensure the delivery of ethically compliant software. However, knowledge on such documentation practice remains scarce. Aims: The objective of our study is to investigate how developers document ethical aspects of open source AI models in practice, aiming at providing recommendations for future documentation endeavours. Method: We selected three sources of documentation on GitHub and Hugging Face, and developed a keyword set to identify ethics-related documents systematically. After filtering an initial set of 2,347 documents, we identified 265 relevant ones and performed thematic analysis to derive the themes of ethical considerations. Results: Six themes emerge, with the three largest ones being model behavioural risks, model use cases, and model risk mitigation. Conclusions: Our findings reveal that open source AI model documentation focuses on articulating ethical problem statements and use case restrictions. We further provide suggestions to various stakeholders for improving documentation practice regarding ethical considerations.
format Preprint
id arxiv_https___arxiv_org_abs_2406_18071
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Documenting Ethical Considerations in Open Source AI Models
Gao, Haoyu
Zahedi, Mansooreh
Treude, Christoph
Rosenstock, Sarita
Cheong, Marc
Software Engineering
Background: The development of AI-enabled software heavily depends on AI model documentation, such as model cards, due to different domain expertise between software engineers and model developers. From an ethical standpoint, AI model documentation conveys critical information on ethical considerations along with mitigation strategies for downstream developers to ensure the delivery of ethically compliant software. However, knowledge on such documentation practice remains scarce. Aims: The objective of our study is to investigate how developers document ethical aspects of open source AI models in practice, aiming at providing recommendations for future documentation endeavours. Method: We selected three sources of documentation on GitHub and Hugging Face, and developed a keyword set to identify ethics-related documents systematically. After filtering an initial set of 2,347 documents, we identified 265 relevant ones and performed thematic analysis to derive the themes of ethical considerations. Results: Six themes emerge, with the three largest ones being model behavioural risks, model use cases, and model risk mitigation. Conclusions: Our findings reveal that open source AI model documentation focuses on articulating ethical problem statements and use case restrictions. We further provide suggestions to various stakeholders for improving documentation practice regarding ethical considerations.
title Documenting Ethical Considerations in Open Source AI Models
topic Software Engineering
url https://arxiv.org/abs/2406.18071