To What Extent Does Agent-generated Code Require Maintenance? An Empirical Study
Fuente:
arXiv
Saved in:
| Main Authors: | Sawada, Shota, Shirai, Tatsuya, Kashiwa, Yutaro, Yamaguchi, Ken'ichi, Iwata, Hiroshi, Iida, Hajimu |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
What to Cut? Predicting Unnecessary Methods in Agentic Code Generation
by: Watanabe, Kan, et al.
Published: (2026)
by: Watanabe, Kan, et al.
Published: (2026)
Does Programming Language Matter? An Empirical Study of Fuzzing Bug Detection
by: Shirai, Tatsuya, et al.
Published: (2026)
by: Shirai, Tatsuya, et al.
Published: (2026)
Large-Scale Empirical Analysis of Continuous Fuzzing: Insights from 1 Million Fuzzing Sessions
by: Shirai, Tatsuya, et al.
Published: (2025)
by: Shirai, Tatsuya, et al.
Published: (2025)
Understanding Self-Admitted Technical Debt in Test Code: An Empirical Study
by: Nakamura, Ibuki, et al.
Published: (2025)
by: Nakamura, Ibuki, et al.
Published: (2025)
Do AI Agents Really Improve Code Readability?
by: Horikawa, Kyogo, et al.
Published: (2026)
by: Horikawa, Kyogo, et al.
Published: (2026)
Agentic Refactoring: An Empirical Study of AI Coding Agents
by: Horikawa, Kosei, et al.
Published: (2025)
by: Horikawa, Kosei, et al.
Published: (2025)
Testing with AI Agents: An Empirical Study of Test Generation Frequency, Quality, and Coverage
by: Yoshimoto, Suzuka, et al.
Published: (2026)
by: Yoshimoto, Suzuka, et al.
Published: (2026)
On the Use of Agentic Coding: An Empirical Study of Pull Requests on GitHub
by: Watanabe, Miku, et al.
Published: (2025)
by: Watanabe, Miku, et al.
Published: (2025)
An Empirical Study of Security-Policy Related Issues in Open Source Projects
by: Kanaji, Rintaro, et al.
Published: (2025)
by: Kanaji, Rintaro, et al.
Published: (2025)
On the Use of Agentic Coding Manifests: An Empirical Study of Claude Code
by: Chatlatanagulchai, Worawalan, et al.
Published: (2025)
by: Chatlatanagulchai, Worawalan, et al.
Published: (2025)
Agent READMEs: An Empirical Study of Context Files for Agentic Coding
by: Chatlatanagulchai, Worawalan, et al.
Published: (2025)
by: Chatlatanagulchai, Worawalan, et al.
Published: (2025)
Detecting and Characterizing Low and No Functionality Packages in the NPM Ecosystem
by: Tevarut, Napasorn, et al.
Published: (2025)
by: Tevarut, Napasorn, et al.
Published: (2025)
Usage, Effects and Requirements for AI Coding Assistants in the Enterprise: An Empirical Study
by: Vukovic, Maja, et al.
Published: (2026)
by: Vukovic, Maja, et al.
Published: (2026)
What Does Explainable AI Mean in Practice? Evaluative Requirements from a Longitudinal Clinical Case Study
by: Sporsem, Tor, et al.
Published: (2025)
by: Sporsem, Tor, et al.
Published: (2025)
Requirements Development and Formalization for Reliable Code Generation: A Multi-Agent Vision
by: Lu, Xu, et al.
Published: (2025)
by: Lu, Xu, et al.
Published: (2025)
Are Coding Agents Generating Over-Mocked Tests? An Empirical Study
by: Hora, Andre, et al.
Published: (2026)
by: Hora, Andre, et al.
Published: (2026)
From Industry Claims to Empirical Reality: An Empirical Study of Code Review Agents in Pull Requests
by: Chowdhury, Kowshik, et al.
Published: (2026)
by: Chowdhury, Kowshik, et al.
Published: (2026)
Moderately Mighty: To What Extent Can Internal Software Metrics Predict App Popularity at Launch?
by: Opu, Md Nahidul Islam, et al.
Published: (2025)
by: Opu, Md Nahidul Islam, et al.
Published: (2025)
What to Retrieve for Effective Retrieval-Augmented Code Generation? An Empirical Study and Beyond
by: Gu, Wenchao, et al.
Published: (2025)
by: Gu, Wenchao, et al.
Published: (2025)
How Does Chunking Affect Retrieval-Augmented Code Completion? A Controlled Empirical Study
by: Wu, Xinjian, et al.
Published: (2026)
by: Wu, Xinjian, et al.
Published: (2026)
Requirements are All You Need: From Requirements to Code with LLMs
by: Wei, Bingyang
Published: (2024)
by: Wei, Bingyang
Published: (2024)
Do AI Coding Agents Log Like Humans? An Empirical Study
by: Ouatiti, Youssef Esseddiq, et al.
Published: (2026)
by: Ouatiti, Youssef Esseddiq, et al.
Published: (2026)
The Hidden Costs of Automation: An Empirical Study on GitHub Actions Workflow Maintenance
by: Valenzuela-Toledo, Pablo, et al.
Published: (2024)
by: Valenzuela-Toledo, Pablo, et al.
Published: (2024)
Agentic Software Engineering: Foundational Pillars and a Research Roadmap
by: Hassan, Ahmed E., et al.
Published: (2025)
by: Hassan, Ahmed E., et al.
Published: (2025)
Requirements Elicitation in Government Projects: A Preliminary Empirical Study
by: Ren, Anqi, et al.
Published: (2024)
by: Ren, Anqi, et al.
Published: (2024)
An Empirical Study on the Amount of Changes Required for Merge Request Acceptance
by: Kansab, Samah, et al.
Published: (2025)
by: Kansab, Samah, et al.
Published: (2025)
Supporting Stakeholder Requirements Expression with LLM Revisions: An Empirical Evaluation
by: Mircea, Michael, et al.
Published: (2026)
by: Mircea, Michael, et al.
Published: (2026)
Does My README File Need To Be Updated? Exploring LLM-Based README Maintenance
by: Gao, Haoyu, et al.
Published: (2026)
by: Gao, Haoyu, et al.
Published: (2026)
Are We All Using Agents the Same Way? An Empirical Study of Core and Peripheral Developers Use of Coding Agents
by: Cynthia, Shamse Tasnim, et al.
Published: (2026)
by: Cynthia, Shamse Tasnim, et al.
Published: (2026)
A Comprehensive Empirical Evaluation of Agent Frameworks on Code-centric Software Engineering Tasks
by: Yin, Zhuowen, et al.
Published: (2025)
by: Yin, Zhuowen, et al.
Published: (2025)
Is Multi-Agent Debate (MAD) the Silver Bullet? An Empirical Analysis of MAD in Code Summarization and Translation
by: Chun, Jina, et al.
Published: (2025)
by: Chun, Jina, et al.
Published: (2025)
Quality Requirements for Code: On the Untapped Potential in Maintainability Specifications
by: Borg, Markus
Published: (2024)
by: Borg, Markus
Published: (2024)
Aligning Requirement for Large Language Model's Code Generation
by: Tian, Zhao, et al.
Published: (2025)
by: Tian, Zhao, et al.
Published: (2025)
Effect of Requirements Analyst Experience on Elicitation Effectiveness: A Family of Empirical Studies
by: Aranda, Alejandrina M., et al.
Published: (2024)
by: Aranda, Alejandrina M., et al.
Published: (2024)
R2Code: A Self-Reflective LLM Framework for Requirements-to-Code Traceability
by: Wang, Yifei, et al.
Published: (2026)
by: Wang, Yifei, et al.
Published: (2026)
How Do Agents Perform Code Optimization? An Empirical Study
by: Peng, Huiyun, et al.
Published: (2025)
by: Peng, Huiyun, et al.
Published: (2025)
An Empirical Study on LLM-based Classification of Requirements-related Provisions in Food-safety Regulations
by: Hassani, Shabnam, et al.
Published: (2025)
by: Hassani, Shabnam, et al.
Published: (2025)
Does Code Cleanliness Affect Coding Agents? A Controlled Minimal-Pair Study
by: Trivedi, Priyansh, et al.
Published: (2026)
by: Trivedi, Priyansh, et al.
Published: (2026)
Human to Document, AI to Code: Comparing GenAI for Notebook Competitions
by: Settewong, Tasha, et al.
Published: (2025)
by: Settewong, Tasha, et al.
Published: (2025)
Beyond Bug Fixes: An Empirical Investigation of Post-Merge Code Quality Issues in Agent-Generated Pull Requests
by: Cynthia, Shamse Tasnim, et al.
Published: (2026)
by: Cynthia, Shamse Tasnim, et al.
Published: (2026)
Similar Items
-
What to Cut? Predicting Unnecessary Methods in Agentic Code Generation
by: Watanabe, Kan, et al.
Published: (2026) -
Does Programming Language Matter? An Empirical Study of Fuzzing Bug Detection
by: Shirai, Tatsuya, et al.
Published: (2026) -
Large-Scale Empirical Analysis of Continuous Fuzzing: Insights from 1 Million Fuzzing Sessions
by: Shirai, Tatsuya, et al.
Published: (2025) -
Understanding Self-Admitted Technical Debt in Test Code: An Empirical Study
by: Nakamura, Ibuki, et al.
Published: (2025) -
Do AI Agents Really Improve Code Readability?
by: Horikawa, Kyogo, et al.
Published: (2026)