TestForge: Feedback-Driven, Agentic Test Suite Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Jain, Kush, Goues, Claire Le |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving Examples in Web API Specifications using Iterated-Calls In-Context Learning
by: Jain, Kush, et al.
Published: (2025)
by: Jain, Kush, et al.
Published: (2025)
TestGenEval: A Real World Unit Test Generation and Test Completion Benchmark
by: Jain, Kush, et al.
Published: (2024)
by: Jain, Kush, et al.
Published: (2024)
HerAgent: Rethinking the Automated Environment Deployment via Hierarchical Test Pyramid
by: Li, Xiang, et al.
Published: (2026)
by: Li, Xiang, et al.
Published: (2026)
Are Large Language Models Memorizing Bug Benchmarks?
by: Ramos, Daniel, et al.
Published: (2024)
by: Ramos, Daniel, et al.
Published: (2024)
DiffSpec: Differential Testing with LLMs using Natural Language Specifications and Code Artifacts
by: Rao, Nikitha, et al.
Published: (2024)
by: Rao, Nikitha, et al.
Published: (2024)
TestDecision: Sequential Test Suite Generation via Greedy Optimization and Reinforcement Learning
by: Wang, Guoqing, et al.
Published: (2026)
by: Wang, Guoqing, et al.
Published: (2026)
Automatic, Expressive, and Scalable Fuzzing with Stitching
by: Green, Harrison, et al.
Published: (2026)
by: Green, Harrison, et al.
Published: (2026)
PreciseBugCollector: Extensible, Executable and Precise Bug-fix Collection
by: Ye, He, et al.
Published: (2023)
by: Ye, He, et al.
Published: (2023)
Reusable Test Suites for Reinforcement Learning
by: Betten, Jørn Eirik, et al.
Published: (2025)
by: Betten, Jørn Eirik, et al.
Published: (2025)
On the Effectiveness of Modular Testing in EvoSuite
by: Dinella, Elizabeth
Published: (2026)
by: Dinella, Elizabeth
Published: (2026)
Empirical Derivations from an Evolving Test Suite
by: Ruohonen, Jukka, et al.
Published: (2025)
by: Ruohonen, Jukka, et al.
Published: (2025)
Idioms: Neural Decompilation With Joint Code and Type Definition Prediction
by: Dramko, Luke, et al.
Published: (2025)
by: Dramko, Luke, et al.
Published: (2025)
TestWeaver: Execution-aware, Feedback-driven Regression Testing Generation with Large Language Models
by: Le, Cuong Chi, et al.
Published: (2025)
by: Le, Cuong Chi, et al.
Published: (2025)
Automatically Removing Unnecessary Stubbings from Test Suites
by: Li, Mengzhen, et al.
Published: (2024)
by: Li, Mengzhen, et al.
Published: (2024)
PenForge: On-the-Fly Expert Agent Construction for Automated Penetration Testing
by: Huang, Huihui, et al.
Published: (2026)
by: Huang, Huihui, et al.
Published: (2026)
What is a "bug"? On subjectivity, epistemic power, and implications for software research
by: Widder, David Gray, et al.
Published: (2024)
by: Widder, David Gray, et al.
Published: (2024)
Ever-Improving Test Suite by Leveraging Large Language Models
by: Qiu, Ketai
Published: (2025)
by: Qiu, Ketai
Published: (2025)
Scalable Similarity-Aware Test Suite Minimization with Reinforcement Learning
by: Gu, Sijia, et al.
Published: (2024)
by: Gu, Sijia, et al.
Published: (2024)
Efficient Incremental Code Coverage Analysis for Regression Test Suites
by: Wang, Jiale Amber, et al.
Published: (2024)
by: Wang, Jiale Amber, et al.
Published: (2024)
Regression Test Suite for Payment Switch using jPOS
by: Sardesai, Atharv, et al.
Published: (2022)
by: Sardesai, Atharv, et al.
Published: (2022)
EvoGPT: Leveraging LLM-Driven Seed Diversity to Improve Search-Based Test Suite Generation
by: Broide, Lior, et al.
Published: (2025)
by: Broide, Lior, et al.
Published: (2025)
BugForge: Constructing and Utilizing DBMS Bug Repository to Enhance DBMS Testing
by: Li, Dawei, et al.
Published: (2026)
by: Li, Dawei, et al.
Published: (2026)
E-Test: E'er-Improving Test Suites
by: Qiu, Ketai, et al.
Published: (2025)
by: Qiu, Ketai, et al.
Published: (2025)
A System for Automated Unit Test Generation Using Large Language Models and Assessment of Generated Test Suites
by: Lops, Andrea, et al.
Published: (2024)
by: Lops, Andrea, et al.
Published: (2024)
FeedbackLLM: Metadata driven Multi-Agentic Language Agnostic Test Case Generator with Evolving prompt and Coverage Feedback
by: Jasti, Kushal, et al.
Published: (2026)
by: Jasti, Kushal, et al.
Published: (2026)
A Test Suite for Efficient Robustness Evaluation of Face Recognition Systems
by: Zhang, Ruihan, et al.
Published: (2025)
by: Zhang, Ruihan, et al.
Published: (2025)
Temporal Modeling of Change History for Black-Box Test Suite Minimization
by: Asif, Kamruzzaman, et al.
Published: (2026)
by: Asif, Kamruzzaman, et al.
Published: (2026)
Fast, Fine-Grained Equivalence Checking for Neural Decompilers
by: Dramko, Luke, et al.
Published: (2025)
by: Dramko, Luke, et al.
Published: (2025)
Are Benchmark Tests Strong Enough? Mutation-Guided Diagnosis and Augmentation of Regression Suites
by: Li, Chenglin, et al.
Published: (2026)
by: Li, Chenglin, et al.
Published: (2026)
Improving Spectrum-Based Localization of Multiple Faults by Iterative Test Suite Reduction
by: Callaghan, Dylan, et al.
Published: (2023)
by: Callaghan, Dylan, et al.
Published: (2023)
Revisiting Unnaturalness for Automated Program Repair in the Era of Large Language Models
by: Yang, Aidan Z. H., et al.
Published: (2024)
by: Yang, Aidan Z. H., et al.
Published: (2024)
Feature-Driven End-To-End Test Generation
by: Alian, Parsa, et al.
Published: (2024)
by: Alian, Parsa, et al.
Published: (2024)
Heterogeneous Prompting and Execution Feedback for SWE Issue Test Generation and Selection
by: Ahmed, Toufique, et al.
Published: (2025)
by: Ahmed, Toufique, et al.
Published: (2025)
ReFuzzer: Feedback-Driven Approach to Enhance Validity of LLM-Generated Test Programs
by: Shree, Iti, et al.
Published: (2025)
by: Shree, Iti, et al.
Published: (2025)
Agentic LMs: Hunting Down Test Smells
by: Melo, Rian, et al.
Published: (2025)
by: Melo, Rian, et al.
Published: (2025)
Testing Agentic Workflows with Structural Coverage Criteria
by: Kahani, Nafiseh, et al.
Published: (2026)
by: Kahani, Nafiseh, et al.
Published: (2026)
Automated Test Suite Enhancement Using Large Language Models with Few-shot Prompting
by: Chudic, Alex, et al.
Published: (2026)
by: Chudic, Alex, et al.
Published: (2026)
Adversarial Reasoning for Repair Based on Inferred Program Intent
by: Ye, He, et al.
Published: (2025)
by: Ye, He, et al.
Published: (2025)
QuanForge: A Mutation Testing Framework for Quantum Neural Networks
by: Shao, Minqi, et al.
Published: (2026)
by: Shao, Minqi, et al.
Published: (2026)
TDFlow: Agentic Workflows for Test Driven Development
by: Han, Kevin, et al.
Published: (2025)
by: Han, Kevin, et al.
Published: (2025)
Similar Items
-
Improving Examples in Web API Specifications using Iterated-Calls In-Context Learning
by: Jain, Kush, et al.
Published: (2025) -
TestGenEval: A Real World Unit Test Generation and Test Completion Benchmark
by: Jain, Kush, et al.
Published: (2024) -
HerAgent: Rethinking the Automated Environment Deployment via Hierarchical Test Pyramid
by: Li, Xiang, et al.
Published: (2026) -
Are Large Language Models Memorizing Bug Benchmarks?
by: Ramos, Daniel, et al.
Published: (2024) -
DiffSpec: Differential Testing with LLMs using Natural Language Specifications and Code Artifacts
by: Rao, Nikitha, et al.
Published: (2024)