Evaluating Large Language Models in Detecting Test Smells
Fuente:
arXiv
Guardado en:
| Autores principales: | Lucas, Keila, Gheyi, Rohit, Soares, Elvys, Ribeiro, Márcio, Machado, Ivan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Investigating the Performance of Small Language Models in Detecting Test Smells in Manual Test Cases
por: Lucas, Keila, et al.
Publicado: (2025)
por: Lucas, Keila, et al.
Publicado: (2025)
Agentic LMs: Hunting Down Test Smells
por: Melo, Rian, et al.
Publicado: (2025)
por: Melo, Rian, et al.
Publicado: (2025)
A Catalog of Transformations to Remove Smells From Natural Language Tests
por: Aranda, Manoel, et al.
Publicado: (2024)
por: Aranda, Manoel, et al.
Publicado: (2024)
Evaluating the Effectiveness of Small Language Models in Detecting Refactoring Bugs
por: Gheyi, Rohit, et al.
Publicado: (2025)
por: Gheyi, Rohit, et al.
Publicado: (2025)
Evaluating the Capability of LLMs in Identifying Compilation Errors in Configurable Systems
por: Albuquerque, Lucas, et al.
Publicado: (2024)
por: Albuquerque, Lucas, et al.
Publicado: (2024)
Code Generation with Small Language Models: A Codeforces-Based Study
por: Souza, Débora, et al.
Publicado: (2025)
por: Souza, Débora, et al.
Publicado: (2025)
Bugs in the Shadows: Static Detection of Faulty Python Refactorings
por: Oliveira, Jonhnanthan, et al.
Publicado: (2025)
por: Oliveira, Jonhnanthan, et al.
Publicado: (2025)
Foundation Models as Oracles for Refactoring Correctness Detection
por: Gheyi, Rohit, et al.
Publicado: (2026)
por: Gheyi, Rohit, et al.
Publicado: (2026)
Variability-Aware Detection and Repair of Compilation Errors Using Foundation Models in Configurable Systems
por: Gheyi, Rohit, et al.
Publicado: (2026)
por: Gheyi, Rohit, et al.
Publicado: (2026)
RefModel: Detecting Refactorings using Foundation Models
por: Simões, Pedro, et al.
Publicado: (2025)
por: Simões, Pedro, et al.
Publicado: (2025)
An Empirical Evaluation of Code Smell Detection in Angular Applications
por: Nunes, Maykon, et al.
Publicado: (2026)
por: Nunes, Maykon, et al.
Publicado: (2026)
Assessing Python Style Guides: An Eye-Tracking Study with Novice Developers
por: Roberto, Pablo, et al.
Publicado: (2024)
por: Roberto, Pablo, et al.
Publicado: (2024)
Adoption of Large Language Models in Scrum Management: Insights from Brazilian Practitioners
por: Perkusich, Mirko, et al.
Publicado: (2026)
por: Perkusich, Mirko, et al.
Publicado: (2026)
Evaluating LLMs Effectiveness in Detecting and Correcting Test Smells: An Empirical Study
por: Santana Jr, E. G., et al.
Publicado: (2025)
por: Santana Jr, E. G., et al.
Publicado: (2025)
Beyond Strict Rules: Assessing the Effectiveness of Large Language Models for Code Smell Detection
por: Souza, Saymon, et al.
Publicado: (2026)
por: Souza, Saymon, et al.
Publicado: (2026)
Quality Assessment of Python Tests Generated by Large Language Models
por: Alves, Victor, et al.
Publicado: (2025)
por: Alves, Victor, et al.
Publicado: (2025)
Test Smell: A Parasitic Energy Consumer in Software Testing
por: Misu, Md Rakib Hossain, et al.
Publicado: (2023)
por: Misu, Md Rakib Hossain, et al.
Publicado: (2023)
Beyond Code, We Are People: A Systematic Mapping of 25 Years of Literature on Soft Skills in Agile Development Teams
por: Lima, Israely, et al.
Publicado: (2026)
por: Lima, Israely, et al.
Publicado: (2026)
An Event-Driven Tool for Context-Aware Code Smell Detection Using SmellDSL
por: Viegas, Matheus dos Santos, et al.
Publicado: (2026)
por: Viegas, Matheus dos Santos, et al.
Publicado: (2026)
Refactoring for Novices in Java: An Eye Tracking Study on the Extract vs. Inline Methods
por: da Costa, José Aldo Silva, et al.
Publicado: (2026)
por: da Costa, José Aldo Silva, et al.
Publicado: (2026)
xNose: A Test Smell Detector for C#
por: Paul, Partha P., et al.
Publicado: (2024)
por: Paul, Partha P., et al.
Publicado: (2024)
A Comprehensive Evaluation of Parameter-Efficient Fine-Tuning on Code Smell Detection
por: Zhang, Beiqi, et al.
Publicado: (2024)
por: Zhang, Beiqi, et al.
Publicado: (2024)
Automated Detection of Inter-Language Design Smells in Multi-Language Deep Learning Frameworks
por: Li, Zengyang, et al.
Publicado: (2024)
por: Li, Zengyang, et al.
Publicado: (2024)
On the Evaluation of Large Language Models in Unit Test Generation
por: Yang, Lin, et al.
Publicado: (2024)
por: Yang, Lin, et al.
Publicado: (2024)
Towards Automated Detection of Inline Code Comment Smells
por: Oztas, Ipek, et al.
Publicado: (2025)
por: Oztas, Ipek, et al.
Publicado: (2025)
Comparative Evaluation of Large Language Models for Test-Skeleton Generation
por: Boorlagadda, Subhang, et al.
Publicado: (2025)
por: Boorlagadda, Subhang, et al.
Publicado: (2025)
The Role of the Retrospective Meetings in Detecting, Refactoring and Monitoring Community Smells
por: Dantas, Carlos, et al.
Publicado: (2025)
por: Dantas, Carlos, et al.
Publicado: (2025)
Characterizing Requirements Smells
por: Gentili, Emanuele, et al.
Publicado: (2024)
por: Gentili, Emanuele, et al.
Publicado: (2024)
Smells Depend on the Context: An Interview Study of Issue Tracking Problems and Smells in Practice
por: Montgomery, Lloyd, et al.
Publicado: (2026)
por: Montgomery, Lloyd, et al.
Publicado: (2026)
How Propense Are Large Language Models at Producing Code Smells? A Benchmarking Study
por: Velasco, Alejandro, et al.
Publicado: (2024)
por: Velasco, Alejandro, et al.
Publicado: (2024)
Specification and Detection of LLM Code Smells
por: Mahmoudi, Brahim, et al.
Publicado: (2025)
por: Mahmoudi, Brahim, et al.
Publicado: (2025)
From Docs to Descriptions: Smell-Aware Evaluation of MCP Server Descriptions
por: Wang, Peiran, et al.
Publicado: (2026)
por: Wang, Peiran, et al.
Publicado: (2026)
Code Smells in Clojure: Initial Findings from a Grey Literature Review
por: Araújo, Walber, et al.
Publicado: (2026)
por: Araújo, Walber, et al.
Publicado: (2026)
TestBench: Evaluating Class-Level Test Case Generation Capability of Large Language Models
por: Zhang, Quanjun, et al.
Publicado: (2024)
por: Zhang, Quanjun, et al.
Publicado: (2024)
STELLAR: A Search-Based Testing Framework for Large Language Model Applications
por: Sorokin, Lev, et al.
Publicado: (2026)
por: Sorokin, Lev, et al.
Publicado: (2026)
Detecting and Fixing API Misuses of Data Science Libraries Using Large Language Models
por: Galappaththi, Akalanka, et al.
Publicado: (2025)
por: Galappaththi, Akalanka, et al.
Publicado: (2025)
Prompt Learning for Multi-Label Code Smell Detection: A Promising Approach
por: Liu, Haiyang, et al.
Publicado: (2024)
por: Liu, Haiyang, et al.
Publicado: (2024)
PyExamine A Comprehensive, UnOpinionated Smell Detection Tool for Python
por: Shivashankar, Karthik, et al.
Publicado: (2025)
por: Shivashankar, Karthik, et al.
Publicado: (2025)
Drowzee: Metamorphic Testing for Fact-Conflicting Hallucination Detection in Large Language Models
por: Li, Ningke, et al.
Publicado: (2024)
por: Li, Ningke, et al.
Publicado: (2024)
ML Code Smells: From Specification to Detection
por: Mahmoudi, Brahim, et al.
Publicado: (2025)
por: Mahmoudi, Brahim, et al.
Publicado: (2025)
Ejemplares similares
-
Investigating the Performance of Small Language Models in Detecting Test Smells in Manual Test Cases
por: Lucas, Keila, et al.
Publicado: (2025) -
Agentic LMs: Hunting Down Test Smells
por: Melo, Rian, et al.
Publicado: (2025) -
A Catalog of Transformations to Remove Smells From Natural Language Tests
por: Aranda, Manoel, et al.
Publicado: (2024) -
Evaluating the Effectiveness of Small Language Models in Detecting Refactoring Bugs
por: Gheyi, Rohit, et al.
Publicado: (2025) -
Evaluating the Capability of LLMs in Identifying Compilation Errors in Configurable Systems
por: Albuquerque, Lucas, et al.
Publicado: (2024)