Toward Generation of Test Cases from Task Descriptions via History-aware Planning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cao, Duy, Nguyen, Phu, Le, Vy, Nguyen, Tien N., Nguyen, Vu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
KAT: Dependency-aware Automated API Testing with Large Language Models
von: Le, Tri, et al.
Veröffentlicht: (2024)
von: Le, Tri, et al.
Veröffentlicht: (2024)
Towards Test Generation from Task Description for Mobile Testing with Multi-modal Reasoning
von: Huynh, Hieu, et al.
Veröffentlicht: (2025)
von: Huynh, Hieu, et al.
Veröffentlicht: (2025)
Segment-Based Test Case Prioritization: A Multi-objective Approach
von: Huynh, Hieu, et al.
Veröffentlicht: (2024)
von: Huynh, Hieu, et al.
Veröffentlicht: (2024)
TestWeaver: Execution-aware, Feedback-driven Regression Testing Generation with Large Language Models
von: Le, Cuong Chi, et al.
Veröffentlicht: (2025)
von: Le, Cuong Chi, et al.
Veröffentlicht: (2025)
RBCTest: Leveraging LLMs to Mine and Verify Oracles of API Response Bodies for RESTful API Testing
von: Huynh, Hieu, et al.
Veröffentlicht: (2025)
von: Huynh, Hieu, et al.
Veröffentlicht: (2025)
MEMRES: A Memory-Augmented Resolver with Confidence Cascade for Agentic Python Dependency Resolution
von: Minh, Dao Sy Duy, et al.
Veröffentlicht: (2026)
von: Minh, Dao Sy Duy, et al.
Veröffentlicht: (2026)
Automated Description Generation for Software Patches
von: Vu, Thanh Trong, et al.
Veröffentlicht: (2024)
von: Vu, Thanh Trong, et al.
Veröffentlicht: (2024)
The Effect of Code Obfuscation on Human Program Comprehension
von: Nguyen, Anh H. N., et al.
Veröffentlicht: (2026)
von: Nguyen, Anh H. N., et al.
Veröffentlicht: (2026)
Enhancing Program Repair with Specification Guidance and Intermediate Behavioral Signals
von: Le-Anh, Minh, et al.
Veröffentlicht: (2026)
von: Le-Anh, Minh, et al.
Veröffentlicht: (2026)
Early-Stage Prediction of Review Effort in AI-Generated Pull Requests
von: Minh, Dao Sy Duy, et al.
Veröffentlicht: (2026)
von: Minh, Dao Sy Duy, et al.
Veröffentlicht: (2026)
Reinforcement Learning-Based REST API Testing with Multi-Coverage
von: Nguyen, Tien-Quang, et al.
Veröffentlicht: (2024)
von: Nguyen, Tien-Quang, et al.
Veröffentlicht: (2024)
HyperAgent: Generalist Software Engineering Agents to Solve Coding Tasks at Scale
von: Phan, Huy Nhat, et al.
Veröffentlicht: (2024)
von: Phan, Huy Nhat, et al.
Veröffentlicht: (2024)
Model-Agnostic Correctness Assessment for LLM-Generated Code via Dynamic Internal Representation Selection
von: Vu, Thanh Trong, et al.
Veröffentlicht: (2025)
von: Vu, Thanh Trong, et al.
Veröffentlicht: (2025)
When Retriever Meets Generator: A Joint Model for Code Comment Generation
von: Le, Tien P. T., et al.
Veröffentlicht: (2025)
von: Le, Tien P. T., et al.
Veröffentlicht: (2025)
Semantic Evolution over Populations for LLM-Guided Automated Program Repair
von: Le, Cuong Chi, et al.
Veröffentlicht: (2026)
von: Le, Cuong Chi, et al.
Veröffentlicht: (2026)
CollabCoder: Plan-Code Co-Evolution via Collaborative Decision-Making for Efficient Code Generation
von: Doan, Duy Tung, et al.
Veröffentlicht: (2026)
von: Doan, Duy Tung, et al.
Veröffentlicht: (2026)
Automated Web Application Testing: End-to-End Test Case Generation with Large Language Models and Screen Transition Graphs
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
Towards Reliable Evaluation of Neural Program Repair with Natural Robustness Testing
von: Le-Cong, Thanh, et al.
Veröffentlicht: (2024)
von: Le-Cong, Thanh, et al.
Veröffentlicht: (2024)
SpecMind: Cognitively Inspired, Interactive Multi-Turn Framework for Postcondition Inference
von: Le, Cuong Chi, et al.
Veröffentlicht: (2026)
von: Le, Cuong Chi, et al.
Veröffentlicht: (2026)
Evaluating Classical Software Process Models as Coordination Mechanisms for LLM-Based Software Generation
von: Ha, Duc Minh, et al.
Veröffentlicht: (2025)
von: Ha, Duc Minh, et al.
Veröffentlicht: (2025)
VisualCoder: Guiding Large Language Models in Code Execution with Fine-grained Multimodal Chain-of-Thought Reasoning
von: Le, Cuong Chi, et al.
Veröffentlicht: (2024)
von: Le, Cuong Chi, et al.
Veröffentlicht: (2024)
CodeLSI: Leveraging Foundation Models for Automated Code Generation with Low-Rank Optimization and Domain-Specific Instruction Tuning
von: Le, Huy, et al.
Veröffentlicht: (2025)
von: Le, Huy, et al.
Veröffentlicht: (2025)
Repeton: Structured Bug Repair with ReAct-Guided Patch-and-Test Cycles
von: Vinh, Nguyen Phu, et al.
Veröffentlicht: (2025)
von: Vinh, Nguyen Phu, et al.
Veröffentlicht: (2025)
Fuzzwise: Intelligent Initial Corpus Generation for Fuzzing
von: Dhulipala, Hridya, et al.
Veröffentlicht: (2025)
von: Dhulipala, Hridya, et al.
Veröffentlicht: (2025)
CABENCH: Benchmarking Composable AI for Solving Complex Tasks through Composing Ready-to-Use Models
von: Pham, Tung-Thuy, et al.
Veröffentlicht: (2025)
von: Pham, Tung-Thuy, et al.
Veröffentlicht: (2025)
Data-Driven Evidence-Based Syntactic Sugar Design
von: OBrien, David, et al.
Veröffentlicht: (2024)
von: OBrien, David, et al.
Veröffentlicht: (2024)
Are the Majority of Public Computational Notebooks Pathologically Non-Executable?
von: Nguyen, Tien, et al.
Veröffentlicht: (2025)
von: Nguyen, Tien, et al.
Veröffentlicht: (2025)
Assessing Large Language Models for Stabilizing Numerical Expressions in Scientific Software
von: Nguyen, Tien, et al.
Veröffentlicht: (2026)
von: Nguyen, Tien, et al.
Veröffentlicht: (2026)
Cerberus: Multi-Agent Reasoning and Coverage-Guided Exploration for Static Detection of Runtime Errors
von: Dhulipala, Hridya, et al.
Veröffentlicht: (2025)
von: Dhulipala, Hridya, et al.
Veröffentlicht: (2025)
Encoding Version History Context for Better Code Representation
von: Nguyen, Huy, et al.
Veröffentlicht: (2024)
von: Nguyen, Huy, et al.
Veröffentlicht: (2024)
Verifying DNN-based Semantic Communication Against Generative Adversarial Noise
von: Le, Thanh, et al.
Veröffentlicht: (2026)
von: Le, Thanh, et al.
Veröffentlicht: (2026)
SQLong: Enhanced NL2SQL for Longer Contexts with LLMs
von: Nguyen, Dai Quoc, et al.
Veröffentlicht: (2025)
von: Nguyen, Dai Quoc, et al.
Veröffentlicht: (2025)
Rectifier: Code Translation with Corrector via LLMs
von: Yin, Xin, et al.
Veröffentlicht: (2024)
von: Yin, Xin, et al.
Veröffentlicht: (2024)
Program Structure-aware Language Models: Targeted Software Testing beyond Textual Semantics
von: Tran, Khang, et al.
Veröffentlicht: (2026)
von: Tran, Khang, et al.
Veröffentlicht: (2026)
Correctness Assessment of Code Generated by Large Language Models Using Internal Representations
von: Bui, Tuan-Dung, et al.
Veröffentlicht: (2025)
von: Bui, Tuan-Dung, et al.
Veröffentlicht: (2025)
Generating Critical Scenarios for Testing Automated Driving Systems
von: Nguyen, Trung-Hieu, et al.
Veröffentlicht: (2024)
von: Nguyen, Trung-Hieu, et al.
Veröffentlicht: (2024)
CodeMMLU: A Multi-Task Benchmark for Assessing Code Understanding & Reasoning Capabilities of CodeLLMs
von: Manh, Dung Nguyen, et al.
Veröffentlicht: (2024)
von: Manh, Dung Nguyen, et al.
Veröffentlicht: (2024)
Toward Explaining Large Language Models in Software Engineering Tasks
von: Vitale, Antonio, et al.
Veröffentlicht: (2025)
von: Vitale, Antonio, et al.
Veröffentlicht: (2025)
An Empirical Study on Capability of Large Language Models in Understanding Code Semantics
von: Nguyen, Thu-Trang, et al.
Veröffentlicht: (2024)
von: Nguyen, Thu-Trang, et al.
Veröffentlicht: (2024)
Finetuning LLMs for Automatic Form Interaction on Web-Browser in Selenium Testing Framework
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
KAT: Dependency-aware Automated API Testing with Large Language Models
von: Le, Tri, et al.
Veröffentlicht: (2024) -
Towards Test Generation from Task Description for Mobile Testing with Multi-modal Reasoning
von: Huynh, Hieu, et al.
Veröffentlicht: (2025) -
Segment-Based Test Case Prioritization: A Multi-objective Approach
von: Huynh, Hieu, et al.
Veröffentlicht: (2024) -
TestWeaver: Execution-aware, Feedback-driven Regression Testing Generation with Large Language Models
von: Le, Cuong Chi, et al.
Veröffentlicht: (2025) -
RBCTest: Leveraging LLMs to Mine and Verify Oracles of API Response Bodies for RESTful API Testing
von: Huynh, Hieu, et al.
Veröffentlicht: (2025)