Navigating the growing field of research on AI for software testing -- the taxonomy for AI-augmented software testing and an ontology-driven literature survey
Fuente:
arXiv
Saved in:
| Main Author: | Schieferdecker, Ina K. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Augmenting software engineering with AI and developing it further towards AI-assisted model-driven software engineering
by: Schieferdecker, Ina K.
Published: (2024)
by: Schieferdecker, Ina K.
Published: (2024)
Fuzzing the brain: Automated stress testing for the safety of ML-driven neurostimulation
by: Downing, Mara, et al.
Published: (2025)
by: Downing, Mara, et al.
Published: (2025)
In industrial embedded software, are some compilation errors easier to localize and fix than others?
by: Fu, Han, et al.
Published: (2024)
by: Fu, Han, et al.
Published: (2024)
Enhancing Differential Testing With LLMs For Testing Deep Learning Libraries
by: Li, Meiziniu, et al.
Published: (2024)
by: Li, Meiziniu, et al.
Published: (2024)
Demystifying the Silence of Correctness Bugs in PyTorch Compiler
by: Li, Meiziniu, et al.
Published: (2026)
by: Li, Meiziniu, et al.
Published: (2026)
COMET: Coverage-guided Model Generation For Deep Learning Library Testing
by: Li, Meiziniu, et al.
Published: (2022)
by: Li, Meiziniu, et al.
Published: (2022)
What is a "bug"? On subjectivity, epistemic power, and implications for software research
by: Widder, David Gray, et al.
Published: (2024)
by: Widder, David Gray, et al.
Published: (2024)
LSPRAG: LSP-Guided RAG for Language-Agnostic Real-Time Unit Test Generation
by: Go, Gwihwan, et al.
Published: (2025)
by: Go, Gwihwan, et al.
Published: (2025)
AgentAssay: Token-Efficient Regression Testing for Non-Deterministic AI Agent Workflows
by: Bhardwaj, Varun Pratap
Published: (2026)
by: Bhardwaj, Varun Pratap
Published: (2026)
AIRA: AI-Induced Risk Audit: A Structured Inspection Framework for AI-Generated Code
by: Parris, William M.
Published: (2026)
by: Parris, William M.
Published: (2026)
Auto-repair without test cases: How LLMs fix compilation errors in large industrial embedded code
by: Fu, Han, et al.
Published: (2025)
by: Fu, Han, et al.
Published: (2025)
Two is Better Than One: Digital Siblings to Improve Autonomous Driving Testing
by: Biagiola, Matteo, et al.
Published: (2023)
by: Biagiola, Matteo, et al.
Published: (2023)
The Productivity-Reliability Paradox: Specification-Driven Governance for AI-Augmented Software Development
by: Farrag, Sabry E.
Published: (2026)
by: Farrag, Sabry E.
Published: (2026)
Automated structural testing of LLM-based agents: methods, framework, and case studies
by: Kohl, Jens, et al.
Published: (2026)
by: Kohl, Jens, et al.
Published: (2026)
Abstain and Validate: A Dual-LLM Policy for Reducing Noise in Agentic Program Repair
by: Cambronero, José, et al.
Published: (2025)
by: Cambronero, José, et al.
Published: (2025)
Testing of Deep Reinforcement Learning Agents with Surrogate Models
by: Biagiola, Matteo, et al.
Published: (2023)
by: Biagiola, Matteo, et al.
Published: (2023)
The Future of AI-Driven Software Engineering
by: Terragni, Valerio, et al.
Published: (2024)
by: Terragni, Valerio, et al.
Published: (2024)
On the Mistaken Assumption of Interchangeable Deep Reinforcement Learning Implementations
by: Hundal, Rajdeep Singh, et al.
Published: (2025)
by: Hundal, Rajdeep Singh, et al.
Published: (2025)
Harnessing the Power of Large Language Models for Software Testing Education: A Focus on ISTQB Syllabus
by: Ngo, Tuan-Phong, et al.
Published: (2025)
by: Ngo, Tuan-Phong, et al.
Published: (2025)
Assessing Data Augmentation-Induced Bias in Training and Testing of Machine Learning Models
by: More, Riddhi, et al.
Published: (2025)
by: More, Riddhi, et al.
Published: (2025)
An Analysis of LLM Fine-Tuning and Few-Shot Learning for Flaky Test Detection and Classification
by: More, Riddhi, et al.
Published: (2025)
by: More, Riddhi, et al.
Published: (2025)
A Systematic Approach for Assessing Large Language Models' Test Case Generation Capability
by: Chang, Hung-Fu, et al.
Published: (2025)
by: Chang, Hung-Fu, et al.
Published: (2025)
From Untestable to Testable: Metamorphic Testing in the Age of LLMs
by: Terragni, Valerio
Published: (2026)
by: Terragni, Valerio
Published: (2026)
Towards Explainable Test Case Prioritisation with Learning-to-Rank Models
by: Ramírez, Aurora, et al.
Published: (2024)
by: Ramírez, Aurora, et al.
Published: (2024)
Test-Driven AI Agent Definition (TDAD): Compiling Tool-Using Agents from Behavioral Specifications
by: Rehan, Tzafrir
Published: (2026)
by: Rehan, Tzafrir
Published: (2026)
Automated Vulnerability Detection Using Deep Learning Technique
by: Yang, Guan-Yan, et al.
Published: (2024)
by: Yang, Guan-Yan, et al.
Published: (2024)
SpecOps: A Fully Automated AI Agent Testing Framework in Real-World GUI Environments
by: Ahmed, Syed Yusuf, et al.
Published: (2026)
by: Ahmed, Syed Yusuf, et al.
Published: (2026)
Supporting software engineering tasks with agentic AI: Demonstration on document retrieval and test scenario generation
by: Kica, Marian, et al.
Published: (2026)
by: Kica, Marian, et al.
Published: (2026)
Generative transformations and patterns in LLM-native approaches for software verification and falsification
by: Braberman, Víctor A., et al.
Published: (2024)
by: Braberman, Víctor A., et al.
Published: (2024)
SemLoc: Structured Grounding of Free-Form LLM Reasoning for Fault Localization
by: Yang, Zhaorui, et al.
Published: (2026)
by: Yang, Zhaorui, et al.
Published: (2026)
Benchmarking Generative AI Models for Deep Learning Test Input Generation
by: Maryam, et al.
Published: (2024)
by: Maryam, et al.
Published: (2024)
Towards a Probabilistic Framework for Analyzing and Improving LLM-Enabled Software
by: Baldonado, Juan Manuel, et al.
Published: (2025)
by: Baldonado, Juan Manuel, et al.
Published: (2025)
Monitoring Agentic Systems Before They're Reliable
by: Boston, Marisa Ferrara, et al.
Published: (2026)
by: Boston, Marisa Ferrara, et al.
Published: (2026)
CodeTracer: Towards Traceable Agent States
by: Li, Han, et al.
Published: (2026)
by: Li, Han, et al.
Published: (2026)
The Explabox: Model-Agnostic Machine Learning Transparency & Analysis
by: Robeer, Marcel, et al.
Published: (2024)
by: Robeer, Marcel, et al.
Published: (2024)
Reasoning Provenance for Autonomous AI Agents: Structured Behavioral Analytics Beyond State Checkpoints and Execution Traces
by: Vispute, Neelmani, et al.
Published: (2026)
by: Vispute, Neelmani, et al.
Published: (2026)
On the Soundness and Consistency of LLM Agents for Executing Test Cases Written in Natural Language
by: Salva, Sébastien, et al.
Published: (2025)
by: Salva, Sébastien, et al.
Published: (2025)
ASSERTIFY: Utilizing Large Language Models to Generate Assertions for Production Code
by: Torkamani, Mohammad Jalili, et al.
Published: (2024)
by: Torkamani, Mohammad Jalili, et al.
Published: (2024)
Closed-Loop Autonomous Software Development via Jira-Integrated Backlog Orchestration: A Case Study in Deterministic Control and Safety-Constrained Automation
by: Calboreanu, Elias
Published: (2026)
by: Calboreanu, Elias
Published: (2026)
RMCBench: Benchmarking Large Language Models' Resistance to Malicious Code
by: Chen, Jiachi, et al.
Published: (2024)
by: Chen, Jiachi, et al.
Published: (2024)
Similar Items
-
Augmenting software engineering with AI and developing it further towards AI-assisted model-driven software engineering
by: Schieferdecker, Ina K.
Published: (2024) -
Fuzzing the brain: Automated stress testing for the safety of ML-driven neurostimulation
by: Downing, Mara, et al.
Published: (2025) -
In industrial embedded software, are some compilation errors easier to localize and fix than others?
by: Fu, Han, et al.
Published: (2024) -
Enhancing Differential Testing With LLMs For Testing Deep Learning Libraries
by: Li, Meiziniu, et al.
Published: (2024) -
Demystifying the Silence of Correctness Bugs in PyTorch Compiler
by: Li, Meiziniu, et al.
Published: (2026)