Applications and Implications of Large Language Models in Qualitative Analysis: A New Frontier for Empirical Software Engineering

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Leça, Matheus de Morais, Valença, Lucas, Santos, Reydne, Santos, Ronnie de Souza
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866909531063189504
author Leça, Matheus de Morais
Valença, Lucas
Santos, Reydne
Santos, Ronnie de Souza
author_facet Leça, Matheus de Morais
Valença, Lucas
Santos, Reydne
Santos, Ronnie de Souza
contents The use of large language models (LLMs) for qualitative analysis is gaining attention in various fields, including software engineering, where qualitative methods are essential for understanding human and social factors. This study aimed to investigate how LLMs are currently used in qualitative analysis and their potential applications in software engineering research, focusing on the benefits, limitations, and practices associated with their use. A systematic mapping study was conducted, analyzing 21 relevant studies to explore reported uses of LLMs for qualitative analysis. The findings indicate that LLMs are primarily used for tasks such as coding, thematic analysis, and data categorization, offering benefits like increased efficiency and support for new researchers. However, limitations such as output variability, challenges in capturing nuanced perspectives, and ethical concerns related to privacy and transparency were also identified. The study emphasizes the need for structured strategies and guidelines to optimize LLM use in qualitative research within software engineering, enhancing their effectiveness while addressing ethical considerations. While LLMs show promise in supporting qualitative analysis, human expertise remains crucial for interpreting data, and ongoing exploration of best practices will be vital for their successful integration into empirical software engineering research.
format Preprint
id arxiv_https___arxiv_org_abs_2412_06564
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Applications and Implications of Large Language Models in Qualitative Analysis: A New Frontier for Empirical Software Engineering
Leça, Matheus de Morais
Valença, Lucas
Santos, Reydne
Santos, Ronnie de Souza
Software Engineering
The use of large language models (LLMs) for qualitative analysis is gaining attention in various fields, including software engineering, where qualitative methods are essential for understanding human and social factors. This study aimed to investigate how LLMs are currently used in qualitative analysis and their potential applications in software engineering research, focusing on the benefits, limitations, and practices associated with their use. A systematic mapping study was conducted, analyzing 21 relevant studies to explore reported uses of LLMs for qualitative analysis. The findings indicate that LLMs are primarily used for tasks such as coding, thematic analysis, and data categorization, offering benefits like increased efficiency and support for new researchers. However, limitations such as output variability, challenges in capturing nuanced perspectives, and ethical concerns related to privacy and transparency were also identified. The study emphasizes the need for structured strategies and guidelines to optimize LLM use in qualitative research within software engineering, enhancing their effectiveness while addressing ethical considerations. While LLMs show promise in supporting qualitative analysis, human expertise remains crucial for interpreting data, and ongoing exploration of best practices will be vital for their successful integration into empirical software engineering research.
title Applications and Implications of Large Language Models in Qualitative Analysis: A New Frontier for Empirical Software Engineering
topic Software Engineering
url https://arxiv.org/abs/2412.06564