On the Effectiveness of LLMs for Manual Test Verifications
Fuente:
arXiv
Guardado en:
| Autores principales: | Peixoto, Myron David Lucena Campos, Baia, Davy de Medeiros, Nascimento, Nathalia, Alencar, Paulo, Fonseca, Baldoino, Ribeiro, Márcio |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Vulnerability Detection with Interprocedural Context in Multiple Languages: Assessing Effectiveness and Cost of Modern LLMs
por: Lira, Kevin, et al.
Publicado: (2026)
por: Lira, Kevin, et al.
Publicado: (2026)
Designing Empirical Studies on LLM-Based Code Generation: Towards a Reference Framework
por: Nascimento, Nathalia, et al.
Publicado: (2025)
por: Nascimento, Nathalia, et al.
Publicado: (2025)
Foundation Models as Oracles for Refactoring Correctness Detection
por: Gheyi, Rohit, et al.
Publicado: (2026)
por: Gheyi, Rohit, et al.
Publicado: (2026)
Automated Non-Functional Requirements Generation in Software Engineering with Large Language Models: A Comparative Study
por: Almonte, Jomar Thomas, et al.
Publicado: (2025)
por: Almonte, Jomar Thomas, et al.
Publicado: (2025)
Towards Automated Formal Verification of Backend Systems with LLMs
por: Xu, Kangping, et al.
Publicado: (2025)
por: Xu, Kangping, et al.
Publicado: (2025)
LLMs for Test Input Generation for Semantic Caches
por: Rasool, Zafaryab, et al.
Publicado: (2024)
por: Rasool, Zafaryab, et al.
Publicado: (2024)
Chain of Targeted Verification Questions to Improve the Reliability of Code Generated by LLMs
por: Ngassom, Sylvain Kouemo, et al.
Publicado: (2024)
por: Ngassom, Sylvain Kouemo, et al.
Publicado: (2024)
Accessible Smart Contracts Verification: Synthesizing Formal Models with Tamed LLMs
por: Corazza, Jan, et al.
Publicado: (2025)
por: Corazza, Jan, et al.
Publicado: (2025)
Verification-Guided Context Optimization for Tool Calling via Hierarchical LLMs-as-Editors
por: Li, Henger, et al.
Publicado: (2025)
por: Li, Henger, et al.
Publicado: (2025)
RAG-Verus: Repository-Level Program Verification with LLMs using Retrieval Augmented Generation
por: Zhong, Sicheng, et al.
Publicado: (2025)
por: Zhong, Sicheng, et al.
Publicado: (2025)
LlamaRestTest: Effective REST API Testing with Small Language Models
por: Kim, Myeongsoo, et al.
Publicado: (2025)
por: Kim, Myeongsoo, et al.
Publicado: (2025)
LLMs for Automated Unit Test Generation and Assessment in Java: The AgoneTest Framework
por: Lops, Andrea, et al.
Publicado: (2025)
por: Lops, Andrea, et al.
Publicado: (2025)
UnitTenX: Generating Tests for Legacy Packages with AI Agents Powered by Formal Verification
por: Charalambous, Yiannis, et al.
Publicado: (2025)
por: Charalambous, Yiannis, et al.
Publicado: (2025)
WIP: Assessing the Effectiveness of ChatGPT in Preparatory Testing Activities
por: Haldar, Susmita, et al.
Publicado: (2025)
por: Haldar, Susmita, et al.
Publicado: (2025)
AutoRestTest: A Tool for Automated REST API Testing Using LLMs and MARL
por: Stennett, Tyler, et al.
Publicado: (2025)
por: Stennett, Tyler, et al.
Publicado: (2025)
Variability-Aware Machine Learning Model Selection: Feature Modeling, Instantiation, and Experimental Case Study
por: Tavares, Cristina, et al.
Publicado: (2024)
por: Tavares, Cristina, et al.
Publicado: (2024)
Evaluating the Effectiveness of LLMs in Fixing Maintainability Issues in Real-World Projects
por: Nunes, Henrique, et al.
Publicado: (2025)
por: Nunes, Henrique, et al.
Publicado: (2025)
TAM-Eval: Evaluating LLMs for Automated Unit Test Maintenance
por: Bruches, Elena, et al.
Publicado: (2026)
por: Bruches, Elena, et al.
Publicado: (2026)
ReCatcher: Towards LLMs Regression Testing for Code Generation
por: Abbassi, Altaf Allah, et al.
Publicado: (2025)
por: Abbassi, Altaf Allah, et al.
Publicado: (2025)
The Potential of LLMs in Automating Software Testing: From Generation to Reporting
por: Sherifi, Betim, et al.
Publicado: (2024)
por: Sherifi, Betim, et al.
Publicado: (2024)
Can LLMs Enable Verification in Mainstream Programming?
por: Shefer, Aleksandr, et al.
Publicado: (2025)
por: Shefer, Aleksandr, et al.
Publicado: (2025)
Mutation-based Consistency Testing for Evaluating the Code Understanding Capability of LLMs
por: Li, Ziyu, et al.
Publicado: (2024)
por: Li, Ziyu, et al.
Publicado: (2024)
TREAT: A Code LLMs Trustworthiness / Reliability Evaluation and Testing Framework
por: Gao, Shuzheng, et al.
Publicado: (2025)
por: Gao, Shuzheng, et al.
Publicado: (2025)
Unify and Triumph: Polyglot, Diverse, and Self-Consistent Generation of Unit Tests with LLMs
por: Khelladi, Djamel Eddine, et al.
Publicado: (2025)
por: Khelladi, Djamel Eddine, et al.
Publicado: (2025)
Harnessing the Power of LLMs: Automating Unit Test Generation for High-Performance Computing
por: Karanjai, Rabimba, et al.
Publicado: (2024)
por: Karanjai, Rabimba, et al.
Publicado: (2024)
LogiCase: Effective Test Case Generation from Logical Description in Competitive Programming
por: Sung, Sicheol, et al.
Publicado: (2025)
por: Sung, Sicheol, et al.
Publicado: (2025)
Toward Effective AI Governance: A Review of Principles
por: Ribeiro, Danilo, et al.
Publicado: (2025)
por: Ribeiro, Danilo, et al.
Publicado: (2025)
Verification and Validation of Autonomous Systems
por: Shetiya, Sneha Sudhir, et al.
Publicado: (2024)
por: Shetiya, Sneha Sudhir, et al.
Publicado: (2024)
PyGraft: Configurable Generation of Synthetic Schemas and Knowledge Graphs at Your Fingertips
por: Hubert, Nicolas, et al.
Publicado: (2023)
por: Hubert, Nicolas, et al.
Publicado: (2023)
Helping LLMs Improve Code Generation Using Feedback from Testing and Static Analysis
por: Dolcetti, Greta, et al.
Publicado: (2024)
por: Dolcetti, Greta, et al.
Publicado: (2024)
LLMs in the Heart of Differential Testing: A Case Study on a Medical Rule Engine
por: Isaku, Erblin, et al.
Publicado: (2024)
por: Isaku, Erblin, et al.
Publicado: (2024)
Automating a Complete Software Test Process Using LLMs: An Automotive Case Study
por: Wang, Shuai, et al.
Publicado: (2025)
por: Wang, Shuai, et al.
Publicado: (2025)
LLMs are All You Need? Improving Fuzz Testing for MOJO with Large Language Models
por: Huang, Linghan, et al.
Publicado: (2025)
por: Huang, Linghan, et al.
Publicado: (2025)
LLM4DS: Evaluating Large Language Models for Data Science Code Generation
por: Nascimento, Nathalia, et al.
Publicado: (2024)
por: Nascimento, Nathalia, et al.
Publicado: (2024)
Programming with Data: Test-Driven Data Engineering for Self-Improving LLMs from Raw Corpora
por: Pan, Chenkai, et al.
Publicado: (2026)
por: Pan, Chenkai, et al.
Publicado: (2026)
VeriStruct: AI-assisted Automated Verification of Data-Structure Modules in Verus
por: Sun, Chuyue, et al.
Publicado: (2025)
por: Sun, Chuyue, et al.
Publicado: (2025)
Explicating Tacit Regulatory Knowledge from LLMs to Auto-Formalize Requirements for Compliance Test Case Generation
por: Xue, Zhiyi, et al.
Publicado: (2026)
por: Xue, Zhiyi, et al.
Publicado: (2026)
AgentGuard: Runtime Verification of AI Agents
por: Koohestani, Roham
Publicado: (2025)
por: Koohestani, Roham
Publicado: (2025)
Can LLMs Reason Like Automated Theorem Provers for Rust Verification? VCoT-Bench: Evaluating via Verification Chain of Thought
por: Xie, Zichen, et al.
Publicado: (2026)
por: Xie, Zichen, et al.
Publicado: (2026)
RESTestBench: A Benchmark for Evaluating the Effectiveness of LLM-Generated REST API Test Cases from NL Requirements
por: Kogler, Leon, et al.
Publicado: (2026)
por: Kogler, Leon, et al.
Publicado: (2026)
Ejemplares similares
-
Vulnerability Detection with Interprocedural Context in Multiple Languages: Assessing Effectiveness and Cost of Modern LLMs
por: Lira, Kevin, et al.
Publicado: (2026) -
Designing Empirical Studies on LLM-Based Code Generation: Towards a Reference Framework
por: Nascimento, Nathalia, et al.
Publicado: (2025) -
Foundation Models as Oracles for Refactoring Correctness Detection
por: Gheyi, Rohit, et al.
Publicado: (2026) -
Automated Non-Functional Requirements Generation in Software Engineering with Large Language Models: A Comparative Study
por: Almonte, Jomar Thomas, et al.
Publicado: (2025) -
Towards Automated Formal Verification of Backend Systems with LLMs
por: Xu, Kangping, et al.
Publicado: (2025)