Sakura: An Approach for Generating Complex Tests from Natural Language Test Descriptions
Fuente:
arXiv
Saved in:
| Main Authors: | Stennett, Tyler, Pan, Rangeet, McGinn, Bridget, Orso, Alessandro, Sinha, Saurabh |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hamster: A Large-Scale Study and Characterization of Developer-Written Tests
by: Pan, Rangeet, et al.
Published: (2025)
by: Pan, Rangeet, et al.
Published: (2025)
A Multi-Agent Approach for REST API Testing with Semantic Graphs and LLM-Driven Inputs
by: Kim, Myeongsoo, et al.
Published: (2024)
by: Kim, Myeongsoo, et al.
Published: (2024)
AutoRestTest: A Tool for Automated REST API Testing Using LLMs and MARL
by: Stennett, Tyler, et al.
Published: (2025)
by: Stennett, Tyler, et al.
Published: (2025)
Leveraging Large Language Models to Improve REST API Testing
by: Kim, Myeongsoo, et al.
Published: (2023)
by: Kim, Myeongsoo, et al.
Published: (2023)
SAINT: Service-level Integration Test Generation with Program Analysis and LLM-based Agents
by: Pan, Rangeet, et al.
Published: (2025)
by: Pan, Rangeet, et al.
Published: (2025)
LlamaRestTest: Effective REST API Testing with Small Language Models
by: Kim, Myeongsoo, et al.
Published: (2025)
by: Kim, Myeongsoo, et al.
Published: (2025)
ASTER: Natural and Multi-language Unit Test Generation with LLMs
by: Pan, Rangeet, et al.
Published: (2024)
by: Pan, Rangeet, et al.
Published: (2024)
Otter: Generating Tests from Issues to Validate SWE Patches
by: Ahmed, Toufique, et al.
Published: (2025)
by: Ahmed, Toufique, et al.
Published: (2025)
TDD-Bench Verified: Can LLMs Generate Tests for Issues Before They Get Resolved?
by: Ahmed, Toufique, et al.
Published: (2024)
by: Ahmed, Toufique, et al.
Published: (2024)
Codellm-Devkit: A Framework for Contextualizing Code LLMs with Program Analysis Insights
by: Krishna, Rahul, et al.
Published: (2024)
by: Krishna, Rahul, et al.
Published: (2024)
Advancing Automated In-Isolation Validation in Repository-Level Code Translation
by: Ke, Kaiyao, et al.
Published: (2025)
by: Ke, Kaiyao, et al.
Published: (2025)
AlphaTrans: A Neuro-Symbolic Compositional Approach for Repository-Level Code Translation and Validation
by: Ibrahimzada, Ali Reza, et al.
Published: (2024)
by: Ibrahimzada, Ali Reza, et al.
Published: (2024)
ScarfBench: A Benchmark for Cross-Framework Application Migration in Enterprise Java
by: Pavuluri, Advait, et al.
Published: (2026)
by: Pavuluri, Advait, et al.
Published: (2026)
Lost in Translation: A Study of Bugs Introduced by Large Language Models while Translating Code
by: Pan, Rangeet, et al.
Published: (2023)
by: Pan, Rangeet, et al.
Published: (2023)
Requirements Coverage-Guided Minimization for Natural Language Test Cases
by: Pan, Rongqi, et al.
Published: (2025)
by: Pan, Rongqi, et al.
Published: (2025)
Towards Test Generation from Task Description for Mobile Testing with Multi-modal Reasoning
by: Huynh, Hieu, et al.
Published: (2025)
by: Huynh, Hieu, et al.
Published: (2025)
Generating REST API Tests With Descriptive Names
by: Garrett, Philip, et al.
Published: (2025)
by: Garrett, Philip, et al.
Published: (2025)
Improving Program Debloating with 1-DU Chain Minimality
by: Kim, Myeongsoo, et al.
Published: (2024)
by: Kim, Myeongsoo, et al.
Published: (2024)
VISOR: A Vision-Language Model-based Test Oracle for Testing Robots
by: Saurabh, Prasun, et al.
Published: (2026)
by: Saurabh, Prasun, et al.
Published: (2026)
Redefining Crowdsourced Test Report Prioritization: An Innovative Approach with Large Language Model
by: Ling, Yuchen, et al.
Published: (2024)
by: Ling, Yuchen, et al.
Published: (2024)
Toward Generation of Test Cases from Task Descriptions via History-aware Planning
by: Cao, Duy, et al.
Published: (2025)
by: Cao, Duy, et al.
Published: (2025)
Rethinking Basis Path Testing: Mixed Integer Programming Approach for Test Path Set Generation
by: Wei, Chao, et al.
Published: (2026)
by: Wei, Chao, et al.
Published: (2026)
From Natural Language to Executable Properties for Property-based Testing of Mobile Apps
by: Xiong, Yiheng, et al.
Published: (2026)
by: Xiong, Yiheng, et al.
Published: (2026)
TestWeaver: Execution-aware, Feedback-driven Regression Testing Generation with Large Language Models
by: Le, Cuong Chi, et al.
Published: (2025)
by: Le, Cuong Chi, et al.
Published: (2025)
TestBench: Evaluating Class-Level Test Case Generation Capability of Large Language Models
by: Zhang, Quanjun, et al.
Published: (2024)
by: Zhang, Quanjun, et al.
Published: (2024)
Teralizer: Semantics-Based Test Generalization from Conventional Unit Tests to Property-Based Tests
by: Glock, Johann, et al.
Published: (2025)
by: Glock, Johann, et al.
Published: (2025)
Test smells in LLM-Generated Unit Tests
by: Ouédraogo, Wendkûuni C., et al.
Published: (2024)
by: Ouédraogo, Wendkûuni C., et al.
Published: (2024)
Test Wars: A Comparative Study of SBST, Symbolic Execution, and LLM-Based Approaches to Unit Test Generation
by: Abdullin, Azat, et al.
Published: (2025)
by: Abdullin, Azat, et al.
Published: (2025)
WebTestPilot: Agentic End-to-End Web Testing against Natural Language Specification by Inferring Oracles with Symbolized GUI Elements
by: Teoh, Xiwen, et al.
Published: (2026)
by: Teoh, Xiwen, et al.
Published: (2026)
On the Evaluation of Large Language Models in Unit Test Generation
by: Yang, Lin, et al.
Published: (2024)
by: Yang, Lin, et al.
Published: (2024)
Metamorphic Testing of Large Language Models for Natural Language Processing
by: Cho, Steven, et al.
Published: (2025)
by: Cho, Steven, et al.
Published: (2025)
Generalizing Test Cases for Comprehensive Test Scenario Coverage
by: Qi, Binhang, et al.
Published: (2026)
by: Qi, Binhang, et al.
Published: (2026)
Test Plan Generation for Live Testing of Cloud Services
by: Jebbar, Oussama, et al.
Published: (2025)
by: Jebbar, Oussama, et al.
Published: (2025)
DiffSpec: Differential Testing with LLMs using Natural Language Specifications and Code Artifacts
by: Rao, Nikitha, et al.
Published: (2024)
by: Rao, Nikitha, et al.
Published: (2024)
Usage, Effects and Requirements for AI Coding Assistants in the Enterprise: An Empirical Study
by: Vukovic, Maja, et al.
Published: (2026)
by: Vukovic, Maja, et al.
Published: (2026)
DISTINCT: A Description-Guided Branch-Consistency Analysis Framework for Non-Regressive Test Case Generation
by: Xue, Pengyu, et al.
Published: (2025)
by: Xue, Pengyu, et al.
Published: (2025)
Issue2Test: Generating Reproducing Test Cases from Issue Reports
by: Nashid, Noor, et al.
Published: (2025)
by: Nashid, Noor, et al.
Published: (2025)
A Generic Approach to Fix Test Flakiness in Real-World Projects
by: Chen, Yang, et al.
Published: (2024)
by: Chen, Yang, et al.
Published: (2024)
CooTest: An Automated Testing Approach for V2X Communication Systems
by: Guo, An, et al.
Published: (2024)
by: Guo, An, et al.
Published: (2024)
TestGenEval: A Real World Unit Test Generation and Test Completion Benchmark
by: Jain, Kush, et al.
Published: (2024)
by: Jain, Kush, et al.
Published: (2024)
Similar Items
-
Hamster: A Large-Scale Study and Characterization of Developer-Written Tests
by: Pan, Rangeet, et al.
Published: (2025) -
A Multi-Agent Approach for REST API Testing with Semantic Graphs and LLM-Driven Inputs
by: Kim, Myeongsoo, et al.
Published: (2024) -
AutoRestTest: A Tool for Automated REST API Testing Using LLMs and MARL
by: Stennett, Tyler, et al.
Published: (2025) -
Leveraging Large Language Models to Improve REST API Testing
by: Kim, Myeongsoo, et al.
Published: (2023) -
SAINT: Service-level Integration Test Generation with Program Analysis and LLM-based Agents
by: Pan, Rangeet, et al.
Published: (2025)