TDAD: Test-Driven Agentic Development - Reducing Code Regressions in AI Coding Agents via Graph-Based Impact Analysis
Fuente:
arXiv
Saved in:
| Main Authors: | Alonso, Pepe, Yovine, Sergio, Braberman, Victor A. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AdvFusion: Adapter-based Knowledge Transfer for Code Summarization on Code Language Models
by: Saberi, Iman, et al.
Published: (2023)
by: Saberi, Iman, et al.
Published: (2023)
A Prompt Learning Framework for Source Code Summarization
by: Xu, Tingting, et al.
Published: (2023)
by: Xu, Tingting, et al.
Published: (2023)
RepoAudit: An Autonomous LLM-Agent for Repository-Level Code Auditing
by: Guo, Jinyao, et al.
Published: (2025)
by: Guo, Jinyao, et al.
Published: (2025)
Context Engineering for Multi-Agent LLM Code Assistants Using Elicit, NotebookLM, ChatGPT, and Claude Code
by: Haseeb, Muhammad
Published: (2025)
by: Haseeb, Muhammad
Published: (2025)
Kodezi Chronos: A Debugging-First Language Model for Repository-Scale Code Understanding
by: Khan, Ishraq, et al.
Published: (2025)
by: Khan, Ishraq, et al.
Published: (2025)
SLEAN: Simple Lightweight Ensemble Analysis Network for Multi-Provider LLM Coordination: Design, Implementation, and Vibe Coding Bug Investigation Case Study
by: Vargas, Matheus J. T.
Published: (2025)
by: Vargas, Matheus J. T.
Published: (2025)
Test-Driven AI Agent Definition (TDAD): Compiling Tool-Using Agents from Behavioral Specifications
by: Rehan, Tzafrir
Published: (2026)
by: Rehan, Tzafrir
Published: (2026)
LLMDFA: Analyzing Dataflow in Code with Large Language Models
by: Wang, Chengpeng, et al.
Published: (2024)
by: Wang, Chengpeng, et al.
Published: (2024)
AI for software engineering: from probable to provable
by: Meyer, Bertrand
Published: (2025)
by: Meyer, Bertrand
Published: (2025)
Generative AI and the Transformation of Software Development Practices
by: Acharya, Vivek
Published: (2025)
by: Acharya, Vivek
Published: (2025)
Feature-Factory: Automating Software Feature Integration Using Generative AI
by: Vsevolodovna, Ruslan Idelfonso Magana
Published: (2024)
by: Vsevolodovna, Ruslan Idelfonso Magana
Published: (2024)
QHackBench: Benchmarking Large Language Models for Quantum Code Generation Using PennyLane Hackathon Challenges
by: Basit, Abdul, et al.
Published: (2025)
by: Basit, Abdul, et al.
Published: (2025)
Utilization of Pre-trained Language Model for Adapter-based Knowledge Transfer in Software Engineering
by: Saberi, Iman, et al.
Published: (2023)
by: Saberi, Iman, et al.
Published: (2023)
DRS-OSS: Practical Diff Risk Scoring with LLMs
by: Sayedsalehi, Ali, et al.
Published: (2025)
by: Sayedsalehi, Ali, et al.
Published: (2025)
See-Saw Generative Mechanism for Scalable Recursive Code Generation with Generative AI
by: Vsevolodovna, Ruslan Idelfonso Magaña
Published: (2024)
by: Vsevolodovna, Ruslan Idelfonso Magaña
Published: (2024)
ABox Abduction for Inconsistent Knowledge Bases under Repair Semantics
by: Haak, Anselm, et al.
Published: (2026)
by: Haak, Anselm, et al.
Published: (2026)
Dual-Process Scaffold Reasoning for Enhancing LLM Code Debugging
by: Hsieh, Po-Chung, et al.
Published: (2025)
by: Hsieh, Po-Chung, et al.
Published: (2025)
Agentic, Context-Aware Risk Intelligence in the Internet of Value
by: Magableh, Basel, et al.
Published: (2026)
by: Magableh, Basel, et al.
Published: (2026)
The Right Prompts for the Job: Repair Code-Review Defects with Large Language Model
by: Zhao, Zelin, et al.
Published: (2023)
by: Zhao, Zelin, et al.
Published: (2023)
Semantic Modeling for World-Centered Architectures
by: Mantsivoda, Andrei, et al.
Published: (2026)
by: Mantsivoda, Andrei, et al.
Published: (2026)
Implementing Knowledge Representation and Reasoning with Object Oriented Design
by: Bassiouny, Abdelrhman, et al.
Published: (2026)
by: Bassiouny, Abdelrhman, et al.
Published: (2026)
Mind the Metrics: Patterns for Telemetry-Aware In-IDE AI Application Development using the Model Context Protocol (MCP)
by: Koc, Vincent, et al.
Published: (2025)
by: Koc, Vincent, et al.
Published: (2025)
Evaluating Developer-written Unit Test Case Reduction for Java -- A Replication Study
by: Le, Tuan D, et al.
Published: (2025)
by: Le, Tuan D, et al.
Published: (2025)
Supporting software engineering tasks with agentic AI: Demonstration on document retrieval and test scenario generation
by: Kica, Marian, et al.
Published: (2026)
by: Kica, Marian, et al.
Published: (2026)
LLM Agents for Generating Microservice-based Applications: how complex is your specification?
by: Yellin, Daniel M.
Published: (2025)
by: Yellin, Daniel M.
Published: (2025)
Large Language Models in Software Documentation and Modeling: A Literature Review and Findings
by: Radosky, Lukas, et al.
Published: (2026)
by: Radosky, Lukas, et al.
Published: (2026)
Bug In the Code Stack: Can LLMs Find Bugs in Large Python Code Stacks
by: Lee, Hokyung, et al.
Published: (2024)
by: Lee, Hokyung, et al.
Published: (2024)
Eliminating Backdoors in Neural Code Models for Secure Code Understanding
by: Sun, Weisong, et al.
Published: (2024)
by: Sun, Weisong, et al.
Published: (2024)
TableMoE: Neuro-Symbolic Routing for Structured Expert Reasoning in Multimodal Table Understanding
by: Zhang, Junwen, et al.
Published: (2025)
by: Zhang, Junwen, et al.
Published: (2025)
GraphSkill: Documentation-Guided Hierarchical Retrieval-Augmented Coding for Complex Graph Reasoning
by: Wang, Fali, et al.
Published: (2026)
by: Wang, Fali, et al.
Published: (2026)
Case Studies and Reflections on Agentic Software Engineering for Rapid Development of Digital Music Instruments
by: Yee-King, Matthew John
Published: (2026)
by: Yee-King, Matthew John
Published: (2026)
CETBench: A Novel Dataset constructed via Transformations over Programs for Benchmarking LLMs for Code-Equivalence Checking
by: Oza, Neeva, et al.
Published: (2025)
by: Oza, Neeva, et al.
Published: (2025)
Optimizing Large Language Models for OpenAPI Code Completion
by: Petryshyn, Bohdan, et al.
Published: (2024)
by: Petryshyn, Bohdan, et al.
Published: (2024)
CodeEvolve: LLM-Driven Evolutionary Optimization with Runtime-Enriched Target Selection for Multi-Language Code Enhancement
by: Borra, Ajay Krishna, et al.
Published: (2026)
by: Borra, Ajay Krishna, et al.
Published: (2026)
Rethinking LLM-Based RTL Code Optimization Via Timing Logic Metamorphosis
by: Xu, Zhihao, et al.
Published: (2025)
by: Xu, Zhihao, et al.
Published: (2025)
OpCode-Based Malware Classification Using Machine Learning and Deep Learning Techniques
by: Saini, Varij, et al.
Published: (2025)
by: Saini, Varij, et al.
Published: (2025)
When Many-Shot Prompting Fails: An Empirical Study of LLM Code Translation
by: Oskooei, Amirkia Rafiei, et al.
Published: (2025)
by: Oskooei, Amirkia Rafiei, et al.
Published: (2025)
NES: An Instruction-Free, Low-Latency Next Edit Suggestion Framework Powered by Learned Historical Editing Trajectories
by: Chen, Xinfang, et al.
Published: (2025)
by: Chen, Xinfang, et al.
Published: (2025)
A Comprehensive Mathematical and System-Level Analysis of Autonomous Vehicle Timelines
by: Perrone, Paul
Published: (2025)
by: Perrone, Paul
Published: (2025)
Query-Aware Flow Diffusion for Graph-Based RAG with Retrieval Guarantees
by: Zhou, Zhuoping, et al.
Published: (2026)
by: Zhou, Zhuoping, et al.
Published: (2026)
Similar Items
-
AdvFusion: Adapter-based Knowledge Transfer for Code Summarization on Code Language Models
by: Saberi, Iman, et al.
Published: (2023) -
A Prompt Learning Framework for Source Code Summarization
by: Xu, Tingting, et al.
Published: (2023) -
RepoAudit: An Autonomous LLM-Agent for Repository-Level Code Auditing
by: Guo, Jinyao, et al.
Published: (2025) -
Context Engineering for Multi-Agent LLM Code Assistants Using Elicit, NotebookLM, ChatGPT, and Claude Code
by: Haseeb, Muhammad
Published: (2025) -
Kodezi Chronos: A Debugging-First Language Model for Repository-Scale Code Understanding
by: Khan, Ishraq, et al.
Published: (2025)