Bugs in Large Language Models Generated Code: An Empirical Study
Fuente:
arXiv
Salvato in:
| Autori principali: | Tambon, Florian, Dakhel, Arghavan Moradi, Nikanjam, Amin, Khomh, Foutse, Desmarais, Michel C., Antoniol, Giuliano |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
TaskEval: Assessing Difficulty of Code Generation Tasks for Large Language Models
di: Tambon, Florian, et al.
Pubblicazione: (2024)
di: Tambon, Florian, et al.
Pubblicazione: (2024)
Chain of Targeted Verification Questions to Improve the Reliability of Code Generated by LLMs
di: Ngassom, Sylvain Kouemo, et al.
Pubblicazione: (2024)
di: Ngassom, Sylvain Kouemo, et al.
Pubblicazione: (2024)
GIST: Generated Inputs Sets Transferability in Deep Learning
di: Tambon, Florian, et al.
Pubblicazione: (2023)
di: Tambon, Florian, et al.
Pubblicazione: (2023)
Common Challenges of Deep Reinforcement Learning Applications Development: An Empirical Study
di: Morovati, Mohammad Mehdi, et al.
Pubblicazione: (2023)
di: Morovati, Mohammad Mehdi, et al.
Pubblicazione: (2023)
Prism: Dynamic and Flexible Benchmarking of LLMs Code Generation with Monte Carlo Tree Search
di: Majdinasab, Vahid, et al.
Pubblicazione: (2025)
di: Majdinasab, Vahid, et al.
Pubblicazione: (2025)
Trained Without My Consent: Detecting Code Inclusion In Language Models Trained on Code
di: Majdinasab, Vahid, et al.
Pubblicazione: (2024)
di: Majdinasab, Vahid, et al.
Pubblicazione: (2024)
DeepCodeProbe: Towards Understanding What Models Trained on Code Learn
di: Majdinasab, Vahid, et al.
Pubblicazione: (2024)
di: Majdinasab, Vahid, et al.
Pubblicazione: (2024)
ReCatcher: Towards LLMs Regression Testing for Code Generation
di: Abbassi, Altaf Allah, et al.
Pubblicazione: (2025)
di: Abbassi, Altaf Allah, et al.
Pubblicazione: (2025)
Continuously Learning Bug Locations
di: Mindom, Paulina Stevia Nouwou, et al.
Pubblicazione: (2024)
di: Mindom, Paulina Stevia Nouwou, et al.
Pubblicazione: (2024)
Refactoring with LLMs: Bridging Human Expertise and Machine Understanding
di: Piao, Yonnel Chen Kuang, et al.
Pubblicazione: (2025)
di: Piao, Yonnel Chen Kuang, et al.
Pubblicazione: (2025)
A Taxonomy of Inefficiencies in LLM-Generated Python Code
di: Abbassi, Altaf Allah, et al.
Pubblicazione: (2025)
di: Abbassi, Altaf Allah, et al.
Pubblicazione: (2025)
A Survey of Bugs in AI-Generated Code
di: Gao, Ruofan, et al.
Pubblicazione: (2025)
di: Gao, Ruofan, et al.
Pubblicazione: (2025)
Fault Localization in Deep Learning-based Software: A System-level Approach
di: Morovati, Mohammad Mehdi, et al.
Pubblicazione: (2024)
di: Morovati, Mohammad Mehdi, et al.
Pubblicazione: (2024)
An Efficient Model Maintenance Approach for MLOps
di: Majidi, Forough, et al.
Pubblicazione: (2024)
di: Majidi, Forough, et al.
Pubblicazione: (2024)
Adversarial Moral Stress Testing of Large Language Models
di: Jamshidi, Saeid, et al.
Pubblicazione: (2026)
di: Jamshidi, Saeid, et al.
Pubblicazione: (2026)
Towards Enhancing the Reproducibility of Deep Learning Bugs: An Empirical Study
di: Shah, Mehil B., et al.
Pubblicazione: (2024)
di: Shah, Mehil B., et al.
Pubblicazione: (2024)
Inferring Code Correctness from Specification
di: Florian, Tambon, et al.
Pubblicazione: (2026)
di: Florian, Tambon, et al.
Pubblicazione: (2026)
Deep Learning Model Reuse in the HuggingFace Community: Challenges, Benefit and Trends
di: Taraghi, Mina, et al.
Pubblicazione: (2024)
di: Taraghi, Mina, et al.
Pubblicazione: (2024)
Imitation Game: Reproducing Deep Learning Bugs Leveraging an Intelligent Agent
di: Shah, Mehil B, et al.
Pubblicazione: (2025)
di: Shah, Mehil B, et al.
Pubblicazione: (2025)
Adversarial Attack Classification and Robustness Testing for Large Language Models for Code
di: Liu, Yang, et al.
Pubblicazione: (2025)
di: Liu, Yang, et al.
Pubblicazione: (2025)
Toward Debugging Deep Reinforcement Learning Programs with RLExplorer
di: Bouchoucha, Rached, et al.
Pubblicazione: (2024)
di: Bouchoucha, Rached, et al.
Pubblicazione: (2024)
PathOCL: Path-Based Prompt Augmentation for OCL Generation with GPT-4
di: Abukhalaf, Seif, et al.
Pubblicazione: (2024)
di: Abukhalaf, Seif, et al.
Pubblicazione: (2024)
Towards Understanding the Impact of Data Bugs on Deep Learning Models in Software Engineering
di: Shah, Mehil B, et al.
Pubblicazione: (2024)
di: Shah, Mehil B, et al.
Pubblicazione: (2024)
An Empirical Study of Policy-as-Code Adoption in Open-Source Software Projects
di: Foalem, Patrick Loic, et al.
Pubblicazione: (2026)
di: Foalem, Patrick Loic, et al.
Pubblicazione: (2026)
Exploring Security Practices in Infrastructure as Code: An Empirical Study
di: Verdet, Alexandre, et al.
Pubblicazione: (2023)
di: Verdet, Alexandre, et al.
Pubblicazione: (2023)
Leveraging Data Characteristics for Bug Localization in Deep Learning Programs
di: Manke, Ruchira, et al.
Pubblicazione: (2024)
di: Manke, Ruchira, et al.
Pubblicazione: (2024)
RefAgent: A Multi-agent LLM-based Framework for Automatic Software Refactoring
di: Oueslati, Khouloud, et al.
Pubblicazione: (2025)
di: Oueslati, Khouloud, et al.
Pubblicazione: (2025)
Improving the Robustness of Large Language Models for Code Tasks via Fine-tuning with Perturbed Data
di: Liu, Yang, et al.
Pubblicazione: (2026)
di: Liu, Yang, et al.
Pubblicazione: (2026)
Machine Learning Robustness: A Primer
di: Braiek, Houssem Ben, et al.
Pubblicazione: (2024)
di: Braiek, Houssem Ben, et al.
Pubblicazione: (2024)
An Empirical Study of Self-Admitted Technical Debt in Machine Learning Software
di: Bhatia, Aaditya, et al.
Pubblicazione: (2023)
di: Bhatia, Aaditya, et al.
Pubblicazione: (2023)
Empirical Characterization of Logging Smells in Machine Learning Code
di: Foalem, Patrick Loic, et al.
Pubblicazione: (2026)
di: Foalem, Patrick Loic, et al.
Pubblicazione: (2026)
Empirical Characterization of Logging Smells in Machine Learning Code
di: Foalem, Patrick Loic, et al.
Pubblicazione: (2026)
di: Foalem, Patrick Loic, et al.
Pubblicazione: (2026)
LLMs and Stack Overflow Discussions: Reliability, Impact, and Challenges
di: Da Silva, Leuson, et al.
Pubblicazione: (2024)
di: Da Silva, Leuson, et al.
Pubblicazione: (2024)
Secure Tool Manifest and Digital Signing Solution for Verifiable MCP and LLM Pipelines
di: Jamshidi, Saeid, et al.
Pubblicazione: (2026)
di: Jamshidi, Saeid, et al.
Pubblicazione: (2026)
The Moral Consistency Pipeline: Continuous Ethical Evaluation for Large Language Models
di: Jamshidi, Saeid, et al.
Pubblicazione: (2025)
di: Jamshidi, Saeid, et al.
Pubblicazione: (2025)
SDLog: A Deep Learning Framework for Detecting Sensitive Information in Software Logs
di: Aghili, Roozbeh, et al.
Pubblicazione: (2025)
di: Aghili, Roozbeh, et al.
Pubblicazione: (2025)
On Fixing Insecure AI-Generated Code through Model Fine-Tuning and Prompting Strategies
di: Jahromi, Ali Soltanian Fard, et al.
Pubblicazione: (2026)
di: Jahromi, Ali Soltanian Fard, et al.
Pubblicazione: (2026)
Quality Issues in Machine Learning Software Systems
di: Côté, Pierre-Olivier, et al.
Pubblicazione: (2023)
di: Côté, Pierre-Olivier, et al.
Pubblicazione: (2023)
What Information Contributes to Log-based Anomaly Detection? Insights from a Configurable Transformer-Based Approach
di: Wu, Xingfang, et al.
Pubblicazione: (2024)
di: Wu, Xingfang, et al.
Pubblicazione: (2024)
Beyond Autoregression: An Empirical Study of Diffusion Large Language Models for Code Generation
di: Li, Chengze, et al.
Pubblicazione: (2025)
di: Li, Chengze, et al.
Pubblicazione: (2025)
Documenti analoghi
-
TaskEval: Assessing Difficulty of Code Generation Tasks for Large Language Models
di: Tambon, Florian, et al.
Pubblicazione: (2024) -
Chain of Targeted Verification Questions to Improve the Reliability of Code Generated by LLMs
di: Ngassom, Sylvain Kouemo, et al.
Pubblicazione: (2024) -
GIST: Generated Inputs Sets Transferability in Deep Learning
di: Tambon, Florian, et al.
Pubblicazione: (2023) -
Common Challenges of Deep Reinforcement Learning Applications Development: An Empirical Study
di: Morovati, Mohammad Mehdi, et al.
Pubblicazione: (2023) -
Prism: Dynamic and Flexible Benchmarking of LLMs Code Generation with Monte Carlo Tree Search
di: Majdinasab, Vahid, et al.
Pubblicazione: (2025)