Investigating the Performance of Small Language Models in Detecting Test Smells in Manual Test Cases
Fuente:
arXiv
Saved in:
| Main Authors: | Lucas, Keila, Gheyi, Rohit, Ribeiro, Márcio, Palomba, Fabio, Martins, Luana, Soares, Elvys |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating Large Language Models in Detecting Test Smells
by: Lucas, Keila, et al.
Published: (2024)
by: Lucas, Keila, et al.
Published: (2024)
Agentic LMs: Hunting Down Test Smells
by: Melo, Rian, et al.
Published: (2025)
by: Melo, Rian, et al.
Published: (2025)
A Catalog of Transformations to Remove Smells From Natural Language Tests
by: Aranda, Manoel, et al.
Published: (2024)
by: Aranda, Manoel, et al.
Published: (2024)
Evaluating the Effectiveness of Small Language Models in Detecting Refactoring Bugs
by: Gheyi, Rohit, et al.
Published: (2025)
by: Gheyi, Rohit, et al.
Published: (2025)
Code Generation with Small Language Models: A Codeforces-Based Study
by: Souza, Débora, et al.
Published: (2025)
by: Souza, Débora, et al.
Published: (2025)
Evaluating the Capability of LLMs in Identifying Compilation Errors in Configurable Systems
by: Albuquerque, Lucas, et al.
Published: (2024)
by: Albuquerque, Lucas, et al.
Published: (2024)
Bugs in the Shadows: Static Detection of Faulty Python Refactorings
by: Oliveira, Jonhnanthan, et al.
Published: (2025)
by: Oliveira, Jonhnanthan, et al.
Published: (2025)
Foundation Models as Oracles for Refactoring Correctness Detection
by: Gheyi, Rohit, et al.
Published: (2026)
by: Gheyi, Rohit, et al.
Published: (2026)
Variability-Aware Detection and Repair of Compilation Errors Using Foundation Models in Configurable Systems
by: Gheyi, Rohit, et al.
Published: (2026)
by: Gheyi, Rohit, et al.
Published: (2026)
RefModel: Detecting Refactorings using Foundation Models
by: Simões, Pedro, et al.
Published: (2025)
by: Simões, Pedro, et al.
Published: (2025)
Assessing Python Style Guides: An Eye-Tracking Study with Novice Developers
by: Roberto, Pablo, et al.
Published: (2024)
by: Roberto, Pablo, et al.
Published: (2024)
How Do Communities of ML-Enabled Systems Smell? A Cross-Sectional Study on the Prevalence of Community Smells
by: Annunziata, Giusy, et al.
Published: (2025)
by: Annunziata, Giusy, et al.
Published: (2025)
Test Smell: A Parasitic Energy Consumer in Software Testing
by: Misu, Md Rakib Hossain, et al.
Published: (2023)
by: Misu, Md Rakib Hossain, et al.
Published: (2023)
Evaluating LLMs Effectiveness in Detecting and Correcting Test Smells: An Empirical Study
by: Santana Jr, E. G., et al.
Published: (2025)
by: Santana Jr, E. G., et al.
Published: (2025)
Socio-Technical Well-Being of Quantum Software Communities: An Overview on Community Smells
by: Lambiase, Stefano, et al.
Published: (2026)
by: Lambiase, Stefano, et al.
Published: (2026)
When Code Smells Meet ML: On the Lifecycle of ML-specific Code Smells in ML-enabled Systems
by: Recupito, Gilberto, et al.
Published: (2024)
by: Recupito, Gilberto, et al.
Published: (2024)
On the Effectiveness of LLMs for Manual Test Verifications
by: Peixoto, Myron David Lucena Campos, et al.
Published: (2024)
by: Peixoto, Myron David Lucena Campos, et al.
Published: (2024)
xNose: A Test Smell Detector for C#
by: Paul, Partha P., et al.
Published: (2024)
by: Paul, Partha P., et al.
Published: (2024)
Investigating Technical Debt Types, Issues, and Solutions in Serverless Computing
by: Perera, Hasini Sumalee, et al.
Published: (2026)
by: Perera, Hasini Sumalee, et al.
Published: (2026)
Adoption of Large Language Models in Scrum Management: Insights from Brazilian Practitioners
by: Perkusich, Mirko, et al.
Published: (2026)
by: Perkusich, Mirko, et al.
Published: (2026)
Investigating the Role of Cultural Values in Adopting Large Language Models for Software Engineering
by: Lambiase, Stefano, et al.
Published: (2024)
by: Lambiase, Stefano, et al.
Published: (2024)
Automated Test Case Repair Using Language Models
by: Yaraghi, Ahmadreza Saboor, et al.
Published: (2024)
by: Yaraghi, Ahmadreza Saboor, et al.
Published: (2024)
Beyond Strict Rules: Assessing the Effectiveness of Large Language Models for Code Smell Detection
by: Souza, Saymon, et al.
Published: (2026)
by: Souza, Saymon, et al.
Published: (2026)
ADPerf: Investigating and Testing Performance in Autonomous Driving Systems
by: Pham, Tri Minh-Triet, et al.
Published: (2025)
by: Pham, Tri Minh-Triet, et al.
Published: (2025)
An Event-Driven Tool for Context-Aware Code Smell Detection Using SmellDSL
by: Viegas, Matheus dos Santos, et al.
Published: (2026)
by: Viegas, Matheus dos Santos, et al.
Published: (2026)
TestBench: Evaluating Class-Level Test Case Generation Capability of Large Language Models
by: Zhang, Quanjun, et al.
Published: (2024)
by: Zhang, Quanjun, et al.
Published: (2024)
From Requirements to Test Cases: An NLP-Based Approach for High-Performance ECU Test Case Automation
by: Medeshetty, Nikitha, et al.
Published: (2025)
by: Medeshetty, Nikitha, et al.
Published: (2025)
Large Language Models as Test Case Generators: Performance Evaluation and Enhancement
by: Li, Kefan, et al.
Published: (2024)
by: Li, Kefan, et al.
Published: (2024)
LlamaRestTest: Effective REST API Testing with Small Language Models
by: Kim, Myeongsoo, et al.
Published: (2025)
by: Kim, Myeongsoo, et al.
Published: (2025)
Refactoring for Novices in Java: An Eye Tracking Study on the Extract vs. Inline Methods
by: da Costa, José Aldo Silva, et al.
Published: (2026)
by: da Costa, José Aldo Silva, et al.
Published: (2026)
No More Manual Tests? Evaluating and Improving ChatGPT for Unit Test Generation
by: Yuan, Zhiqiang, et al.
Published: (2023)
by: Yuan, Zhiqiang, et al.
Published: (2023)
TESTEVAL: Benchmarking Large Language Models for Test Case Generation
by: Wang, Wenhan, et al.
Published: (2024)
by: Wang, Wenhan, et al.
Published: (2024)
Automated Detection of Inter-Language Design Smells in Multi-Language Deep Learning Frameworks
by: Li, Zengyang, et al.
Published: (2024)
by: Li, Zengyang, et al.
Published: (2024)
Classification, Challenges, and Automated Approaches to Handle Non-Functional Requirements in ML-Enabled Systems: A Systematic Literature Review
by: De Martino, Vincenzo, et al.
Published: (2023)
by: De Martino, Vincenzo, et al.
Published: (2023)
FedCSD: A Federated Learning Based Approach for Code-Smell Detection
by: Alawadi, Sadi, et al.
Published: (2023)
by: Alawadi, Sadi, et al.
Published: (2023)
Investigating The Smells of LLM Generated Code
by: Paul, Debalina Ghosh, et al.
Published: (2025)
by: Paul, Debalina Ghosh, et al.
Published: (2025)
On the Impact of Requirements Smells in Prompts: The Case of Automated Traceability
by: Vogelsang, Andreas, et al.
Published: (2025)
by: Vogelsang, Andreas, et al.
Published: (2025)
Acceptance Test Generation with Large Language Models: An Industrial Case Study
by: Ferreira, Margarida, et al.
Published: (2025)
by: Ferreira, Margarida, et al.
Published: (2025)
Clean Code, Better Models: Enhancing LLM Performance with Smell-Cleaned Dataset
by: Xue, Zhipeng, et al.
Published: (2025)
by: Xue, Zhipeng, et al.
Published: (2025)
Towards Automated Detection of Inline Code Comment Smells
by: Oztas, Ipek, et al.
Published: (2025)
by: Oztas, Ipek, et al.
Published: (2025)
Similar Items
-
Evaluating Large Language Models in Detecting Test Smells
by: Lucas, Keila, et al.
Published: (2024) -
Agentic LMs: Hunting Down Test Smells
by: Melo, Rian, et al.
Published: (2025) -
A Catalog of Transformations to Remove Smells From Natural Language Tests
by: Aranda, Manoel, et al.
Published: (2024) -
Evaluating the Effectiveness of Small Language Models in Detecting Refactoring Bugs
by: Gheyi, Rohit, et al.
Published: (2025) -
Code Generation with Small Language Models: A Codeforces-Based Study
by: Souza, Débora, et al.
Published: (2025)