Evaluation of AI Ethics Tools in Language Models: A Developers' Perspective Case Study

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Silva, Jhessica, Moreira, Diego A. B., Santos, Gabriel O. dos, Ferreira, Alef, Maia, Helena, Avila, Sandra, Pedrini, Helio
Format: Preprint
Veröffentlicht: 2025
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866916051604733952
author Silva, Jhessica
Moreira, Diego A. B.
Santos, Gabriel O. dos
Ferreira, Alef
Maia, Helena
Avila, Sandra
Pedrini, Helio
author_facet Silva, Jhessica
Moreira, Diego A. B.
Santos, Gabriel O. dos
Ferreira, Alef
Maia, Helena
Avila, Sandra
Pedrini, Helio
contents In Artificial Intelligence (AI), language models have gained significant importance due to the widespread adoption of systems capable of simulating realistic conversations with humans through text generation. Because of their impact on society, developing and deploying these language models must be done responsibly, with attention to their negative impacts and possible harms. In this scenario, the number of AI Ethics Tools (AIETs) publications has recently increased. These AIETs are designed to help developers, companies, governments, and other stakeholders establish trust, transparency, and responsibility with their technologies by bringing accepted values to guide AI's design, development, and use stages. However, many AIETs lack good documentation, examples of use, and proof of their effectiveness in practice. This paper presents a methodology for evaluating AIETs in language models. Our approach involved an extensive literature survey on 213 AIETs, and after applying inclusion and exclusion criteria, we selected four AIETs: Model Cards, ALTAI, FactSheets, and Harms Modeling. For evaluation, we applied AIETs to language models developed for the Portuguese language, conducting 35 hours of interviews with their developers. The evaluation considered the developers' perspective on the AIETs' use and quality in helping to identify ethical considerations about their model. The results suggest that the applied AIETs serve as a guide for formulating general ethical considerations about language models. However, we note that they do not address unique aspects of these models, such as idiomatic expressions. Additionally, these AIETs did not help to identify potential negative impacts of models for the Portuguese language.
format Preprint
id arxiv_https___arxiv_org_abs_2512_15791
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Evaluation of AI Ethics Tools in Language Models: A Developers' Perspective Case Study
Silva, Jhessica
Moreira, Diego A. B.
Santos, Gabriel O. dos
Ferreira, Alef
Maia, Helena
Avila, Sandra
Pedrini, Helio
Computers and Society
Artificial Intelligence
Computation and Language
In Artificial Intelligence (AI), language models have gained significant importance due to the widespread adoption of systems capable of simulating realistic conversations with humans through text generation. Because of their impact on society, developing and deploying these language models must be done responsibly, with attention to their negative impacts and possible harms. In this scenario, the number of AI Ethics Tools (AIETs) publications has recently increased. These AIETs are designed to help developers, companies, governments, and other stakeholders establish trust, transparency, and responsibility with their technologies by bringing accepted values to guide AI's design, development, and use stages. However, many AIETs lack good documentation, examples of use, and proof of their effectiveness in practice. This paper presents a methodology for evaluating AIETs in language models. Our approach involved an extensive literature survey on 213 AIETs, and after applying inclusion and exclusion criteria, we selected four AIETs: Model Cards, ALTAI, FactSheets, and Harms Modeling. For evaluation, we applied AIETs to language models developed for the Portuguese language, conducting 35 hours of interviews with their developers. The evaluation considered the developers' perspective on the AIETs' use and quality in helping to identify ethical considerations about their model. The results suggest that the applied AIETs serve as a guide for formulating general ethical considerations about language models. However, we note that they do not address unique aspects of these models, such as idiomatic expressions. Additionally, these AIETs did not help to identify potential negative impacts of models for the Portuguese language.
title Evaluation of AI Ethics Tools in Language Models: A Developers' Perspective Case Study
topic Computers and Society
Artificial Intelligence
Computation and Language
url https://arxiv.org/abs/2512.15791