On Iterative Evaluation and Enhancement of Code Quality Using GPT-4o
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Rundong, Frade, Andre, Vaidya, Amal, Labonne, Maxime, Kaiser, Marcus, Chakrabarti, Bismayan, Budd, Jonathan, Moran, Sean |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
When Prompt Under-Specification Improves Code Correctness: An Exploratory Study of Prompt Wording and Structure Effects on LLM-Based Code Generation
di: AKLI, Amal, et al.
Pubblicazione: (2026)
di: AKLI, Amal, et al.
Pubblicazione: (2026)
ChatGPT in Introductory Programming: Counterbalanced Evaluation of Code Quality, Conceptual Learning, and Student Perceptions
di: Andleeb, Shiza, et al.
Pubblicazione: (2025)
di: Andleeb, Shiza, et al.
Pubblicazione: (2025)
Defective Task Descriptions in LLM-Based Code Generation: Detection and Analysis
di: Akli, Amal, et al.
Pubblicazione: (2026)
di: Akli, Amal, et al.
Pubblicazione: (2026)
Using AI/ML to Find and Remediate Enterprise Secrets in Code & Document Sharing Platforms
di: Kerr, Gregor, et al.
Pubblicazione: (2024)
di: Kerr, Gregor, et al.
Pubblicazione: (2024)
No Need to Lift a Finger Anymore? Assessing the Quality of Code Generation by ChatGPT
di: Liu, Zhijie, et al.
Pubblicazione: (2023)
di: Liu, Zhijie, et al.
Pubblicazione: (2023)
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions
di: Larbi, Maya, et al.
Pubblicazione: (2025)
di: Larbi, Maya, et al.
Pubblicazione: (2025)
Automated Generation of High-Quality Bug Reports for Android Applications
di: Saha, Antu, et al.
Pubblicazione: (2026)
di: Saha, Antu, et al.
Pubblicazione: (2026)
Source-Code Analysis of iFogSim for Simulating Distributed IoT Architectures: Coverage, Challenges, and Enhancements
di: Ndadji, Milliam Maxime Zekeng
Pubblicazione: (2026)
di: Ndadji, Milliam Maxime Zekeng
Pubblicazione: (2026)
CYCLE: Learning to Self-Refine the Code Generation
di: Ding, Yangruibo, et al.
Pubblicazione: (2024)
di: Ding, Yangruibo, et al.
Pubblicazione: (2024)
Whodunit: Classifying Code as Human Authored or GPT-4 Generated -- A case study on CodeChef problems
di: Idialu, Oseremen Joy, et al.
Pubblicazione: (2024)
di: Idialu, Oseremen Joy, et al.
Pubblicazione: (2024)
One Model, Many Skills: Parameter-Efficient Fine-Tuning for Multitask Code Analysis
di: Akli, Amal, et al.
Pubblicazione: (2026)
di: Akli, Amal, et al.
Pubblicazione: (2026)
From Evaluation to Enhancement: Large Language Models for Zero-Knowledge Proof Code Generation
di: Xue, Zhantong, et al.
Pubblicazione: (2025)
di: Xue, Zhantong, et al.
Pubblicazione: (2025)
Guiding AI to Fix Its Own Flaws: An Empirical Study on LLM-Driven Secure Code Generation
di: Yan, Hao, et al.
Pubblicazione: (2025)
di: Yan, Hao, et al.
Pubblicazione: (2025)
The First Prompt Counts the Most! An Evaluation of Large Language Models on Iterative Example-Based Code Generation
di: Fu, Yingjie, et al.
Pubblicazione: (2024)
di: Fu, Yingjie, et al.
Pubblicazione: (2024)
On the Quality of AI-Generated Source Code Comments: A Comprehensive Evaluation
di: Guelman, Ian, et al.
Pubblicazione: (2024)
di: Guelman, Ian, et al.
Pubblicazione: (2024)
Evaluating the Effectiveness of GPT-4 Turbo in Creating Defeaters for Assurance Cases
di: Shahandashti, Kimya Khakzad, et al.
Pubblicazione: (2024)
di: Shahandashti, Kimya Khakzad, et al.
Pubblicazione: (2024)
From Human to Machine Refactoring: Assessing GPT-4's Impact on Python Class Quality and Readability
di: Midolo, Alessandro, et al.
Pubblicazione: (2026)
di: Midolo, Alessandro, et al.
Pubblicazione: (2026)
An Iterative Test-and-Repair Framework for Competitive Code Generation
di: Tang, Lingxiao, et al.
Pubblicazione: (2026)
di: Tang, Lingxiao, et al.
Pubblicazione: (2026)
Self-collaboration Code Generation via ChatGPT
di: Dong, Yihong, et al.
Pubblicazione: (2023)
di: Dong, Yihong, et al.
Pubblicazione: (2023)
Do Prompt Patterns Affect Code Quality? A First Empirical Assessment of ChatGPT-Generated Code
di: Della Porta, Antonio, et al.
Pubblicazione: (2025)
di: Della Porta, Antonio, et al.
Pubblicazione: (2025)
LLMs as Evaluators: A Novel Approach to Evaluate Bug Report Summarization
di: Kumar, Abhishek, et al.
Pubblicazione: (2024)
di: Kumar, Abhishek, et al.
Pubblicazione: (2024)
CodeMEM: AST-Guided Adaptive Memory for Repository-Level Iterative Code Generation
di: Wang, Peiding, et al.
Pubblicazione: (2026)
di: Wang, Peiding, et al.
Pubblicazione: (2026)
Evaluation of the Code Generation Capabilities of ChatGPT 4: A Comparative Analysis in 19 Programming Languages
di: Gilbert, L. C.
Pubblicazione: (2025)
di: Gilbert, L. C.
Pubblicazione: (2025)
NoCodeGPT: A No-Code Interface for Building Web Apps with Language Models
di: Monteiro, Mauricio, et al.
Pubblicazione: (2023)
di: Monteiro, Mauricio, et al.
Pubblicazione: (2023)
Evaluating Source Code Quality with Large Language Models: a comparative study
di: Simões, Igor Regis da Silva, et al.
Pubblicazione: (2024)
di: Simões, Igor Regis da Silva, et al.
Pubblicazione: (2024)
Are They All Good? Evaluating the Quality of CoTs in LLM-based Code Generation
di: Zhang, Binquan, et al.
Pubblicazione: (2025)
di: Zhang, Binquan, et al.
Pubblicazione: (2025)
Reassessing Code Authorship Attribution in the Era of Language Models
di: Dipongkor, Atish Kumar, et al.
Pubblicazione: (2025)
di: Dipongkor, Atish Kumar, et al.
Pubblicazione: (2025)
An Empirical Study of the Non-determinism of ChatGPT in Code Generation
di: Ouyang, Shuyin, et al.
Pubblicazione: (2023)
di: Ouyang, Shuyin, et al.
Pubblicazione: (2023)
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension
di: Xu, Fangzhou, et al.
Pubblicazione: (2024)
di: Xu, Fangzhou, et al.
Pubblicazione: (2024)
Quality Evaluation of COBOL to Java Code Transformation
di: Froimovich, Shmulik, et al.
Pubblicazione: (2025)
di: Froimovich, Shmulik, et al.
Pubblicazione: (2025)
Why Do Developers Engage with ChatGPT in Issue-Tracker? Investigating Usage and Reliance on ChatGPT-Generated Code
di: Das, Joy Krishan, et al.
Pubblicazione: (2024)
di: Das, Joy Krishan, et al.
Pubblicazione: (2024)
ChatGPT for Code Refactoring: Analyzing Topics, Interaction, and Effective Prompts
di: AlOmar, Eman Abdullah, et al.
Pubblicazione: (2025)
di: AlOmar, Eman Abdullah, et al.
Pubblicazione: (2025)
A Formal Verification Approach to Safeguard Controller Variables from Single Event Upset
di: Ganesha, et al.
Pubblicazione: (2025)
di: Ganesha, et al.
Pubblicazione: (2025)
Personalization of Code Readability Evaluation Based on LLM Using Collaborative Filtering
di: Hiraki, Buntaro, et al.
Pubblicazione: (2024)
di: Hiraki, Buntaro, et al.
Pubblicazione: (2024)
Treating Run-time Execution History as a First-Class Citizen: Co-Versioning Run-time Behavior alongside Code
di: Kessel, Marcus
Pubblicazione: (2026)
di: Kessel, Marcus
Pubblicazione: (2026)
On the Possibility of Breaking Copyleft Licenses When Reusing Code Generated by ChatGPT
di: Colombo, Gaia, et al.
Pubblicazione: (2025)
di: Colombo, Gaia, et al.
Pubblicazione: (2025)
Studying How Configurations Impact Code Generation in LLMs: the Case of ChatGPT
di: Donato, Benedetta, et al.
Pubblicazione: (2025)
di: Donato, Benedetta, et al.
Pubblicazione: (2025)
How to Refactor this Code? An Exploratory Study on Developer-ChatGPT Refactoring Conversations
di: AlOmar, Eman Abdullah, et al.
Pubblicazione: (2024)
di: AlOmar, Eman Abdullah, et al.
Pubblicazione: (2024)
DSL or Code? Evaluating the Quality of LLM-Generated Algebraic Specifications: A Case Study in Optimization at Kinaxis
di: Ayoughi, Negin, et al.
Pubblicazione: (2026)
di: Ayoughi, Negin, et al.
Pubblicazione: (2026)
Leveraging Design-Aware Context in Large Language Models for Code Comment Generation
di: Mitra, Aritra, et al.
Pubblicazione: (2025)
di: Mitra, Aritra, et al.
Pubblicazione: (2025)
Documenti analoghi
-
When Prompt Under-Specification Improves Code Correctness: An Exploratory Study of Prompt Wording and Structure Effects on LLM-Based Code Generation
di: AKLI, Amal, et al.
Pubblicazione: (2026) -
ChatGPT in Introductory Programming: Counterbalanced Evaluation of Code Quality, Conceptual Learning, and Student Perceptions
di: Andleeb, Shiza, et al.
Pubblicazione: (2025) -
Defective Task Descriptions in LLM-Based Code Generation: Detection and Analysis
di: Akli, Amal, et al.
Pubblicazione: (2026) -
Using AI/ML to Find and Remediate Enterprise Secrets in Code & Document Sharing Platforms
di: Kerr, Gregor, et al.
Pubblicazione: (2024) -
No Need to Lift a Finger Anymore? Assessing the Quality of Code Generation by ChatGPT
di: Liu, Zhijie, et al.
Pubblicazione: (2023)