On Iterative Evaluation and Enhancement of Code Quality Using GPT-4o
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Rundong, Frade, Andre, Vaidya, Amal, Labonne, Maxime, Kaiser, Marcus, Chakrabarti, Bismayan, Budd, Jonathan, Moran, Sean |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
When Prompt Under-Specification Improves Code Correctness: An Exploratory Study of Prompt Wording and Structure Effects on LLM-Based Code Generation
by: AKLI, Amal, et al.
Published: (2026)
by: AKLI, Amal, et al.
Published: (2026)
ChatGPT in Introductory Programming: Counterbalanced Evaluation of Code Quality, Conceptual Learning, and Student Perceptions
by: Andleeb, Shiza, et al.
Published: (2025)
by: Andleeb, Shiza, et al.
Published: (2025)
Defective Task Descriptions in LLM-Based Code Generation: Detection and Analysis
by: Akli, Amal, et al.
Published: (2026)
by: Akli, Amal, et al.
Published: (2026)
Using AI/ML to Find and Remediate Enterprise Secrets in Code & Document Sharing Platforms
by: Kerr, Gregor, et al.
Published: (2024)
by: Kerr, Gregor, et al.
Published: (2024)
No Need to Lift a Finger Anymore? Assessing the Quality of Code Generation by ChatGPT
by: Liu, Zhijie, et al.
Published: (2023)
by: Liu, Zhijie, et al.
Published: (2023)
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions
by: Larbi, Maya, et al.
Published: (2025)
by: Larbi, Maya, et al.
Published: (2025)
Automated Generation of High-Quality Bug Reports for Android Applications
by: Saha, Antu, et al.
Published: (2026)
by: Saha, Antu, et al.
Published: (2026)
Source-Code Analysis of iFogSim for Simulating Distributed IoT Architectures: Coverage, Challenges, and Enhancements
by: Ndadji, Milliam Maxime Zekeng
Published: (2026)
by: Ndadji, Milliam Maxime Zekeng
Published: (2026)
CYCLE: Learning to Self-Refine the Code Generation
by: Ding, Yangruibo, et al.
Published: (2024)
by: Ding, Yangruibo, et al.
Published: (2024)
Whodunit: Classifying Code as Human Authored or GPT-4 Generated -- A case study on CodeChef problems
by: Idialu, Oseremen Joy, et al.
Published: (2024)
by: Idialu, Oseremen Joy, et al.
Published: (2024)
One Model, Many Skills: Parameter-Efficient Fine-Tuning for Multitask Code Analysis
by: Akli, Amal, et al.
Published: (2026)
by: Akli, Amal, et al.
Published: (2026)
From Evaluation to Enhancement: Large Language Models for Zero-Knowledge Proof Code Generation
by: Xue, Zhantong, et al.
Published: (2025)
by: Xue, Zhantong, et al.
Published: (2025)
Guiding AI to Fix Its Own Flaws: An Empirical Study on LLM-Driven Secure Code Generation
by: Yan, Hao, et al.
Published: (2025)
by: Yan, Hao, et al.
Published: (2025)
The First Prompt Counts the Most! An Evaluation of Large Language Models on Iterative Example-Based Code Generation
by: Fu, Yingjie, et al.
Published: (2024)
by: Fu, Yingjie, et al.
Published: (2024)
On the Quality of AI-Generated Source Code Comments: A Comprehensive Evaluation
by: Guelman, Ian, et al.
Published: (2024)
by: Guelman, Ian, et al.
Published: (2024)
Evaluating the Effectiveness of GPT-4 Turbo in Creating Defeaters for Assurance Cases
by: Shahandashti, Kimya Khakzad, et al.
Published: (2024)
by: Shahandashti, Kimya Khakzad, et al.
Published: (2024)
From Human to Machine Refactoring: Assessing GPT-4's Impact on Python Class Quality and Readability
by: Midolo, Alessandro, et al.
Published: (2026)
by: Midolo, Alessandro, et al.
Published: (2026)
An Iterative Test-and-Repair Framework for Competitive Code Generation
by: Tang, Lingxiao, et al.
Published: (2026)
by: Tang, Lingxiao, et al.
Published: (2026)
Self-collaboration Code Generation via ChatGPT
by: Dong, Yihong, et al.
Published: (2023)
by: Dong, Yihong, et al.
Published: (2023)
Do Prompt Patterns Affect Code Quality? A First Empirical Assessment of ChatGPT-Generated Code
by: Della Porta, Antonio, et al.
Published: (2025)
by: Della Porta, Antonio, et al.
Published: (2025)
LLMs as Evaluators: A Novel Approach to Evaluate Bug Report Summarization
by: Kumar, Abhishek, et al.
Published: (2024)
by: Kumar, Abhishek, et al.
Published: (2024)
CodeMEM: AST-Guided Adaptive Memory for Repository-Level Iterative Code Generation
by: Wang, Peiding, et al.
Published: (2026)
by: Wang, Peiding, et al.
Published: (2026)
Evaluation of the Code Generation Capabilities of ChatGPT 4: A Comparative Analysis in 19 Programming Languages
by: Gilbert, L. C.
Published: (2025)
by: Gilbert, L. C.
Published: (2025)
NoCodeGPT: A No-Code Interface for Building Web Apps with Language Models
by: Monteiro, Mauricio, et al.
Published: (2023)
by: Monteiro, Mauricio, et al.
Published: (2023)
Evaluating Source Code Quality with Large Language Models: a comparative study
by: Simões, Igor Regis da Silva, et al.
Published: (2024)
by: Simões, Igor Regis da Silva, et al.
Published: (2024)
Are They All Good? Evaluating the Quality of CoTs in LLM-based Code Generation
by: Zhang, Binquan, et al.
Published: (2025)
by: Zhang, Binquan, et al.
Published: (2025)
Reassessing Code Authorship Attribution in the Era of Language Models
by: Dipongkor, Atish Kumar, et al.
Published: (2025)
by: Dipongkor, Atish Kumar, et al.
Published: (2025)
An Empirical Study of the Non-determinism of ChatGPT in Code Generation
by: Ouyang, Shuyin, et al.
Published: (2023)
by: Ouyang, Shuyin, et al.
Published: (2023)
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension
by: Xu, Fangzhou, et al.
Published: (2024)
by: Xu, Fangzhou, et al.
Published: (2024)
Quality Evaluation of COBOL to Java Code Transformation
by: Froimovich, Shmulik, et al.
Published: (2025)
by: Froimovich, Shmulik, et al.
Published: (2025)
Why Do Developers Engage with ChatGPT in Issue-Tracker? Investigating Usage and Reliance on ChatGPT-Generated Code
by: Das, Joy Krishan, et al.
Published: (2024)
by: Das, Joy Krishan, et al.
Published: (2024)
ChatGPT for Code Refactoring: Analyzing Topics, Interaction, and Effective Prompts
by: AlOmar, Eman Abdullah, et al.
Published: (2025)
by: AlOmar, Eman Abdullah, et al.
Published: (2025)
A Formal Verification Approach to Safeguard Controller Variables from Single Event Upset
by: Ganesha, et al.
Published: (2025)
by: Ganesha, et al.
Published: (2025)
Personalization of Code Readability Evaluation Based on LLM Using Collaborative Filtering
by: Hiraki, Buntaro, et al.
Published: (2024)
by: Hiraki, Buntaro, et al.
Published: (2024)
Treating Run-time Execution History as a First-Class Citizen: Co-Versioning Run-time Behavior alongside Code
by: Kessel, Marcus
Published: (2026)
by: Kessel, Marcus
Published: (2026)
On the Possibility of Breaking Copyleft Licenses When Reusing Code Generated by ChatGPT
by: Colombo, Gaia, et al.
Published: (2025)
by: Colombo, Gaia, et al.
Published: (2025)
Studying How Configurations Impact Code Generation in LLMs: the Case of ChatGPT
by: Donato, Benedetta, et al.
Published: (2025)
by: Donato, Benedetta, et al.
Published: (2025)
How to Refactor this Code? An Exploratory Study on Developer-ChatGPT Refactoring Conversations
by: AlOmar, Eman Abdullah, et al.
Published: (2024)
by: AlOmar, Eman Abdullah, et al.
Published: (2024)
DSL or Code? Evaluating the Quality of LLM-Generated Algebraic Specifications: A Case Study in Optimization at Kinaxis
by: Ayoughi, Negin, et al.
Published: (2026)
by: Ayoughi, Negin, et al.
Published: (2026)
Leveraging Design-Aware Context in Large Language Models for Code Comment Generation
by: Mitra, Aritra, et al.
Published: (2025)
by: Mitra, Aritra, et al.
Published: (2025)
Similar Items
-
When Prompt Under-Specification Improves Code Correctness: An Exploratory Study of Prompt Wording and Structure Effects on LLM-Based Code Generation
by: AKLI, Amal, et al.
Published: (2026) -
ChatGPT in Introductory Programming: Counterbalanced Evaluation of Code Quality, Conceptual Learning, and Student Perceptions
by: Andleeb, Shiza, et al.
Published: (2025) -
Defective Task Descriptions in LLM-Based Code Generation: Detection and Analysis
by: Akli, Amal, et al.
Published: (2026) -
Using AI/ML to Find and Remediate Enterprise Secrets in Code & Document Sharing Platforms
by: Kerr, Gregor, et al.
Published: (2024) -
No Need to Lift a Finger Anymore? Assessing the Quality of Code Generation by ChatGPT
by: Liu, Zhijie, et al.
Published: (2023)