Is LLM-Generated Code More Maintainable \& Reliable than Human-Written Code?
Fuente:
arXiv
Saved in:
| Main Authors: | Molison, Alfred Santa, Moraes, Marcia, Melo, Glaucia, Santos, Fabio, Assuncao, Wesley K. G. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond Resolution Rates: Behavioral Drivers of Coding Agent Success and Failure
by: Mehtiyev, Tural, et al.
Published: (2026)
by: Mehtiyev, Tural, et al.
Published: (2026)
How to Compare the Security of Code Written by Humans to LLM-generated Code
by: Balebako, Rebecca, et al.
Published: (2026)
by: Balebako, Rebecca, et al.
Published: (2026)
Test Code Review in the Era of GitHub Actions: A Replication Study
by: Sun, Hui, et al.
Published: (2026)
by: Sun, Hui, et al.
Published: (2026)
BigCodeArena: Unveiling More Reliable Human Preferences in Code Generation via Execution
by: Zhuo, Terry Yue, et al.
Published: (2025)
by: Zhuo, Terry Yue, et al.
Published: (2025)
Brevity is the Soul of Wit: Condensing Code Changes to Improve Commit Message Generation
by: Kuang, Hongyu, et al.
Published: (2025)
by: Kuang, Hongyu, et al.
Published: (2025)
Assessing the Bug-Proneness of Refactored Code: A Longitudinal Multi-Project Study
by: Ferreira, Isabella, et al.
Published: (2025)
by: Ferreira, Isabella, et al.
Published: (2025)
MaintainCoder: Maintainable Code Generation Under Dynamic Requirements
by: Wang, Zhengren, et al.
Published: (2025)
by: Wang, Zhengren, et al.
Published: (2025)
Human-Written vs. AI-Generated Code: A Large-Scale Study of Defects, Vulnerabilities, and Complexity
by: Cotroneo, Domenico, et al.
Published: (2025)
by: Cotroneo, Domenico, et al.
Published: (2025)
Does Co-Development with AI Assistants Lead to More Maintainable Code? A Registered Report
by: Borg, Markus, et al.
Published: (2024)
by: Borg, Markus, et al.
Published: (2024)
Automatic Generation of Benchmarks and Reliable LLM Judgment for Code Tasks
by: Farchi, Eitan, et al.
Published: (2024)
by: Farchi, Eitan, et al.
Published: (2024)
Code Quality Analysis of Translations from C to Rust
by: Tadesse, Biruk, et al.
Published: (2026)
by: Tadesse, Biruk, et al.
Published: (2026)
"I Would Have Written My Code Differently'': Beginners Struggle to Understand LLM-Generated Code
by: Zi, Yangtian, et al.
Published: (2025)
by: Zi, Yangtian, et al.
Published: (2025)
Measuring the Runtime Performance of C++ Code Written by Humans using GitHub Copilot
by: Erhabor, Daniel, et al.
Published: (2023)
by: Erhabor, Daniel, et al.
Published: (2023)
Multi-Agent Code-Orchestrated Generation for Reliable Infrastructure-as-Code
by: Khan, Rana Nameer Hussain, et al.
Published: (2025)
by: Khan, Rana Nameer Hussain, et al.
Published: (2025)
Refactoring $\neq$ Bug-Inducing: Improving Defect Prediction with Code Change Tactics Analysis
by: Niu, Feifei, et al.
Published: (2025)
by: Niu, Feifei, et al.
Published: (2025)
Quality Requirements for Code: On the Untapped Potential in Maintainability Specifications
by: Borg, Markus
Published: (2024)
by: Borg, Markus
Published: (2024)
Increasing, not Diminishing: Investigating the Returns of Highly Maintainable Code
by: Borg, Markus, et al.
Published: (2024)
by: Borg, Markus, et al.
Published: (2024)
Code Less, Align More: Efficient LLM Fine-tuning for Code Generation with Data Pruning
by: Tsai, Yun-Da, et al.
Published: (2024)
by: Tsai, Yun-Da, et al.
Published: (2024)
Inducing Vulnerable Code Generation in LLM Coding Assistants
by: Zeng, Binqi, et al.
Published: (2025)
by: Zeng, Binqi, et al.
Published: (2025)
LLM-as-a-Judge for Human-AI Co-Creation: A Reliability-Aware Evaluation Framework for Coding
by: Amin, Md Faizul Ibne, et al.
Published: (2026)
by: Amin, Md Faizul Ibne, et al.
Published: (2026)
Challenges in Comparing Code Maintainability across Different Programming Languages
by: Ponsard, Christophe, et al.
Published: (2024)
by: Ponsard, Christophe, et al.
Published: (2024)
Industrial Code Quality Benchmarks: Toward Gamification of Software Maintainability
by: Borg, Markus, et al.
Published: (2024)
by: Borg, Markus, et al.
Published: (2024)
Less is More: DocString Compression in Code Generation
by: Yang, Guang, et al.
Published: (2024)
by: Yang, Guang, et al.
Published: (2024)
HumanEvo: An Evolution-aware Benchmark for More Realistic Evaluation of Repository-level Code Generation
by: Zheng, Dewu, et al.
Published: (2024)
by: Zheng, Dewu, et al.
Published: (2024)
Where are the Hidden Gems? Applying Transformer Models for Design Discussion Detection
by: Arkoh, Lawrence, et al.
Published: (2026)
by: Arkoh, Lawrence, et al.
Published: (2026)
Human or LLM? A Comparative Study on Accessible Code Generation Capability
by: Suh, Hyunjae, et al.
Published: (2025)
by: Suh, Hyunjae, et al.
Published: (2025)
Beyond Code Generation: Assessing Code LLM Maturity with Postconditions
by: He, Fusen, et al.
Published: (2024)
by: He, Fusen, et al.
Published: (2024)
On the Reliability of Code Comprehension Proxies
by: Arvan, Erfan, et al.
Published: (2026)
by: Arvan, Erfan, et al.
Published: (2026)
ComplexCodeEval: A Benchmark for Evaluating Large Code Models on More Complex Code
by: Feng, Jia, et al.
Published: (2024)
by: Feng, Jia, et al.
Published: (2024)
CodeArena: A Collective Evaluation Platform for LLM Code Generation
by: Du, Mingzhe, et al.
Published: (2025)
by: Du, Mingzhe, et al.
Published: (2025)
RefModel: Detecting Refactorings using Foundation Models
by: Simões, Pedro, et al.
Published: (2025)
by: Simões, Pedro, et al.
Published: (2025)
When More Retrieval Hurts: Retrieval-Augmented Code Review Generation
by: Meng, Qianru, et al.
Published: (2025)
by: Meng, Qianru, et al.
Published: (2025)
HumanEvalComm: Benchmarking the Communication Competence of Code Generation for LLMs and LLM Agent
by: Wu, Jie JW, et al.
Published: (2024)
by: Wu, Jie JW, et al.
Published: (2024)
An Empirical Study of Interaction Smells in Multi-Turn Human-LLM Collaborative Code Generation
by: Zhang, Binquan, et al.
Published: (2026)
by: Zhang, Binquan, et al.
Published: (2026)
Structured Safety Auditing for Balancing Code Correctness and Content Safety in LLM-Generated Code
by: Tan, Honghao, et al.
Published: (2026)
by: Tan, Honghao, et al.
Published: (2026)
Model-based Maintenance and Evolution with GenAI: A Look into the Future
by: Marchezan, Luciano, et al.
Published: (2024)
by: Marchezan, Luciano, et al.
Published: (2024)
Contemporary Software Modernization: Perspectives and Challenges to Deal with Legacy Systems
by: Assunção, Wesley K. G., et al.
Published: (2024)
by: Assunção, Wesley K. G., et al.
Published: (2024)
How Maintainable is Proficient Code? A Case Study of Three PyPI Libraries
by: Febriyanti, Indira, et al.
Published: (2024)
by: Febriyanti, Indira, et al.
Published: (2024)
Novice Developers Produce Larger Review Overhead for Project Maintainers while Vibe Coding
by: Asdaque, Syed Ammar, et al.
Published: (2026)
by: Asdaque, Syed Ammar, et al.
Published: (2026)
On-the-Fly Input Adaptation for Reliable Code Intelligence
by: Rathnasuriya, Ravishka, et al.
Published: (2026)
by: Rathnasuriya, Ravishka, et al.
Published: (2026)
Similar Items
-
Beyond Resolution Rates: Behavioral Drivers of Coding Agent Success and Failure
by: Mehtiyev, Tural, et al.
Published: (2026) -
How to Compare the Security of Code Written by Humans to LLM-generated Code
by: Balebako, Rebecca, et al.
Published: (2026) -
Test Code Review in the Era of GitHub Actions: A Replication Study
by: Sun, Hui, et al.
Published: (2026) -
BigCodeArena: Unveiling More Reliable Human Preferences in Code Generation via Execution
by: Zhuo, Terry Yue, et al.
Published: (2025) -
Brevity is the Soul of Wit: Condensing Code Changes to Improve Commit Message Generation
by: Kuang, Hongyu, et al.
Published: (2025)