Evaluating LLMs Effectiveness in Detecting and Correcting Test Smells: An Empirical Study
Fuente:
arXiv
Saved in:
| Main Authors: | Santana Jr, E. G., Junior, Jander Pereira Santos, Almeida, Erlon P., Ahmed, Iftekhar, Neto, Paulo Anselmo da Mota Silveira, de Almeida, Eduardo Santana |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Test Smell: A Parasitic Energy Consumer in Software Testing
by: Misu, Md Rakib Hossain, et al.
Published: (2023)
by: Misu, Md Rakib Hossain, et al.
Published: (2023)
Which Prompting Technique Should I Use? An Empirical Investigation of Prompting Techniques for Software Engineering Tasks
by: Santana Jr, E. G., et al.
Published: (2025)
by: Santana Jr, E. G., et al.
Published: (2025)
What Makes a Great Software Quality Assurance Engineer?
by: Farias, Roselane Silva, et al.
Published: (2024)
by: Farias, Roselane Silva, et al.
Published: (2024)
Please do not go: understanding turnover of software engineers from different perspectives
by: Carvalho, Michelle Larissa Luciano, et al.
Published: (2024)
by: Carvalho, Michelle Larissa Luciano, et al.
Published: (2024)
Beyond Strict Rules: Assessing the Effectiveness of Large Language Models for Code Smell Detection
by: Souza, Saymon, et al.
Published: (2026)
by: Souza, Saymon, et al.
Published: (2026)
Prompting in the Wild: An Empirical Study of Prompt Evolution in Software Repositories
by: Tafreshipour, Mahan, et al.
Published: (2024)
by: Tafreshipour, Mahan, et al.
Published: (2024)
Developers' Perceptions on the Impact of ChatGPT in Software Development: A Survey
by: Vaillant, Thiago S., et al.
Published: (2024)
by: Vaillant, Thiago S., et al.
Published: (2024)
An Empirical Evaluation of Code Smell Detection in Angular Applications
by: Nunes, Maykon, et al.
Published: (2026)
by: Nunes, Maykon, et al.
Published: (2026)
Agentic LMs: Hunting Down Test Smells
by: Melo, Rian, et al.
Published: (2025)
by: Melo, Rian, et al.
Published: (2025)
Evaluating Large Language Models in Detecting Test Smells
by: Lucas, Keila, et al.
Published: (2024)
by: Lucas, Keila, et al.
Published: (2024)
An Empirical Study on Automatically Detecting AI-Generated Source Code: How Far Are We?
by: Suh, Hyunjae, et al.
Published: (2024)
by: Suh, Hyunjae, et al.
Published: (2024)
Inside Out: Uncovering How Comment Internalization Steers LLMs for Better or Worse
by: Imani, Aaron, et al.
Published: (2025)
by: Imani, Aaron, et al.
Published: (2025)
An Event-Driven Tool for Context-Aware Code Smell Detection Using SmellDSL
by: Viegas, Matheus dos Santos, et al.
Published: (2026)
by: Viegas, Matheus dos Santos, et al.
Published: (2026)
Evaluating Multi‐Label Machine Learning Models for Smart Home Environments
by: Diego Corrêa da Silva, et al.
Published: (2025)
by: Diego Corrêa da Silva, et al.
Published: (2025)
Does Documentation Matter? An Empirical Study of Practitioners' Perspective on Open-Source Software Adoption
by: Imani, Aaron, et al.
Published: (2024)
by: Imani, Aaron, et al.
Published: (2024)
Is Multi-Agent Debate (MAD) the Silver Bullet? An Empirical Analysis of MAD in Code Summarization and Translation
by: Chun, Jina, et al.
Published: (2025)
by: Chun, Jina, et al.
Published: (2025)
ToffA-DSPL: an approach of trade-off analysis for designing dynamic software product lines
by: Carvalho, Michelle Larissa Luciano, et al.
Published: (2024)
by: Carvalho, Michelle Larissa Luciano, et al.
Published: (2024)
Empirical Study of the Docker Smells Impact on the Image Size
by: Durieux, Thomas
Published: (2023)
by: Durieux, Thomas
Published: (2023)
Empirical Characterization of Logging Smells in Machine Learning Code
by: Foalem, Patrick Loic, et al.
Published: (2026)
by: Foalem, Patrick Loic, et al.
Published: (2026)
Empirical Characterization of Logging Smells in Machine Learning Code
by: Foalem, Patrick Loic, et al.
Published: (2026)
by: Foalem, Patrick Loic, et al.
Published: (2026)
Investigating the Performance of Small Language Models in Detecting Test Smells in Manual Test Cases
by: Lucas, Keila, et al.
Published: (2025)
by: Lucas, Keila, et al.
Published: (2025)
From Bias To Improved Prompts: A Case Study of Bias Mitigation of Clone Detection Models
by: Chen, QiHong, et al.
Published: (2025)
by: Chen, QiHong, et al.
Published: (2025)
Cache-Related Smells in GitLab CI/CD: Comprehensive Catalog, Automated Detection, and Empirical Evidence
by: Urdih, Francesco, et al.
Published: (2026)
by: Urdih, Francesco, et al.
Published: (2026)
Beyond Self-learned Attention: Mitigating Attention Bias in Transformer-based Models Using Attention Guidance
by: Gesi, Jiri, et al.
Published: (2024)
by: Gesi, Jiri, et al.
Published: (2024)
xNose: A Test Smell Detector for C#
by: Paul, Partha P., et al.
Published: (2024)
by: Paul, Partha P., et al.
Published: (2024)
A Comprehensive Evaluation of Parameter-Efficient Fine-Tuning on Code Smell Detection
by: Zhang, Beiqi, et al.
Published: (2024)
by: Zhang, Beiqi, et al.
Published: (2024)
Beyond Output Correctness: Benchmarking and Evaluating Large Language Model Reasoning in Coding Tasks
by: Li, Yuangang, et al.
Published: (2026)
by: Li, Yuangang, et al.
Published: (2026)
On the Effectiveness of LLMs for Manual Test Verifications
by: Peixoto, Myron David Lucena Campos, et al.
Published: (2024)
by: Peixoto, Myron David Lucena Campos, et al.
Published: (2024)
Context Conquers Parameters: Outperforming Proprietary LLM in Commit Message Generation
by: Imani, Aaron, et al.
Published: (2024)
by: Imani, Aaron, et al.
Published: (2024)
Software Testing with Large Language Models: An Interview Study with Practitioners
by: Santana, Maria Deolinda, et al.
Published: (2025)
by: Santana, Maria Deolinda, et al.
Published: (2025)
Test Co‐Evolution in Software Projects: A Large‐Scale Empirical Study
by: Charles Miranda, et al.
Published: (2025)
by: Charles Miranda, et al.
Published: (2025)
An Empirical Study of Interaction Smells in Multi-Turn Human-LLM Collaborative Code Generation
by: Zhang, Binquan, et al.
Published: (2026)
by: Zhang, Binquan, et al.
Published: (2026)
Specification and Detection of LLM Code Smells
by: Mahmoudi, Brahim, et al.
Published: (2025)
by: Mahmoudi, Brahim, et al.
Published: (2025)
Towards Automated Detection of Inline Code Comment Smells
by: Oztas, Ipek, et al.
Published: (2025)
by: Oztas, Ipek, et al.
Published: (2025)
The Role of the Retrospective Meetings in Detecting, Refactoring and Monitoring Community Smells
by: Dantas, Carlos, et al.
Published: (2025)
by: Dantas, Carlos, et al.
Published: (2025)
Exploring the Effectiveness of LLMs in Automated Logging Generation: An Empirical Study
by: Li, Yichen, et al.
Published: (2023)
by: Li, Yichen, et al.
Published: (2023)
Characterizing Requirements Smells
by: Gentili, Emanuele, et al.
Published: (2024)
by: Gentili, Emanuele, et al.
Published: (2024)
ML Code Smells: From Specification to Detection
by: Mahmoudi, Brahim, et al.
Published: (2025)
by: Mahmoudi, Brahim, et al.
Published: (2025)
Smells Depend on the Context: An Interview Study of Issue Tracking Problems and Smells in Practice
by: Montgomery, Lloyd, et al.
Published: (2026)
by: Montgomery, Lloyd, et al.
Published: (2026)
Understanding API Usage and Testing: An Empirical Study of C Libraries
by: Zaki, Ahmed, et al.
Published: (2025)
by: Zaki, Ahmed, et al.
Published: (2025)
Similar Items
-
Test Smell: A Parasitic Energy Consumer in Software Testing
by: Misu, Md Rakib Hossain, et al.
Published: (2023) -
Which Prompting Technique Should I Use? An Empirical Investigation of Prompting Techniques for Software Engineering Tasks
by: Santana Jr, E. G., et al.
Published: (2025) -
What Makes a Great Software Quality Assurance Engineer?
by: Farias, Roselane Silva, et al.
Published: (2024) -
Please do not go: understanding turnover of software engineers from different perspectives
by: Carvalho, Michelle Larissa Luciano, et al.
Published: (2024) -
Beyond Strict Rules: Assessing the Effectiveness of Large Language Models for Code Smell Detection
by: Souza, Saymon, et al.
Published: (2026)