On Fixing Insecure AI-Generated Code through Model Fine-Tuning and Prompting Strategies
Fuente:
arXiv
Saved in:
| Main Authors: | Jahromi, Ali Soltanian Fard, Tahir, Amjed, Liang, Peng, Khomh, Foutse |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Survey of Bugs in AI-Generated Code
by: Gao, Ruofan, et al.
Published: (2025)
by: Gao, Ruofan, et al.
Published: (2025)
On Developers' Self-Declaration of AI-Generated Code: An Analysis of Practices
by: Kashif, Syed Mohammad, et al.
Published: (2025)
by: Kashif, Syed Mohammad, et al.
Published: (2025)
From Prompting to Verification: How Experience Shapes Vibe Coding Practices
by: Fawzy, Ahmed, et al.
Published: (2026)
by: Fawzy, Ahmed, et al.
Published: (2026)
Improving the Robustness of Large Language Models for Code Tasks via Fine-tuning with Perturbed Data
by: Liu, Yang, et al.
Published: (2026)
by: Liu, Yang, et al.
Published: (2026)
PathOCL: Path-Based Prompt Augmentation for OCL Generation with GPT-4
by: Abukhalaf, Seif, et al.
Published: (2024)
by: Abukhalaf, Seif, et al.
Published: (2024)
Trained Without My Consent: Detecting Code Inclusion In Language Models Trained on Code
by: Majdinasab, Vahid, et al.
Published: (2024)
by: Majdinasab, Vahid, et al.
Published: (2024)
Adversarial Attack Classification and Robustness Testing for Large Language Models for Code
by: Liu, Yang, et al.
Published: (2025)
by: Liu, Yang, et al.
Published: (2025)
A Taxonomy of Inefficiencies in LLM-Generated Python Code
by: Abbassi, Altaf Allah, et al.
Published: (2025)
by: Abbassi, Altaf Allah, et al.
Published: (2025)
Vibe Coding in Practice: Motivations, Challenges, and a Future Outlook -- a Grey Literature Review
by: Fawzy, Ahmed, et al.
Published: (2025)
by: Fawzy, Ahmed, et al.
Published: (2025)
Prism: Dynamic and Flexible Benchmarking of LLMs Code Generation with Monte Carlo Tree Search
by: Majdinasab, Vahid, et al.
Published: (2025)
by: Majdinasab, Vahid, et al.
Published: (2025)
Exploring Data Management Challenges and Solutions in Agile Software Development: A Literature Review and Practitioner Survey
by: Fawzy, Ahmed, et al.
Published: (2024)
by: Fawzy, Ahmed, et al.
Published: (2024)
TaskEval: Assessing Difficulty of Code Generation Tasks for Large Language Models
by: Tambon, Florian, et al.
Published: (2024)
by: Tambon, Florian, et al.
Published: (2024)
Real Faults in Model Context Protocol (MCP) Software: a Comprehensive Taxonomy
by: Taraghi, Mina, et al.
Published: (2026)
by: Taraghi, Mina, et al.
Published: (2026)
GIST: Generated Inputs Sets Transferability in Deep Learning
by: Tambon, Florian, et al.
Published: (2023)
by: Tambon, Florian, et al.
Published: (2023)
Protecting Privacy in Software Logs: What Should Be Anonymized?
by: Aghili, Roozbeh, et al.
Published: (2024)
by: Aghili, Roozbeh, et al.
Published: (2024)
ReCatcher: Towards LLMs Regression Testing for Code Generation
by: Abbassi, Altaf Allah, et al.
Published: (2025)
by: Abbassi, Altaf Allah, et al.
Published: (2025)
An Empirical Study of Policy-as-Code Adoption in Open-Source Software Projects
by: Foalem, Patrick Loic, et al.
Published: (2026)
by: Foalem, Patrick Loic, et al.
Published: (2026)
Structural Anchors and Reasoning Fragility:Understanding CoT Robustness in LLM4Code
by: Liu, Yang, et al.
Published: (2026)
by: Liu, Yang, et al.
Published: (2026)
Chain of Targeted Verification Questions to Improve the Reliability of Code Generated by LLMs
by: Ngassom, Sylvain Kouemo, et al.
Published: (2024)
by: Ngassom, Sylvain Kouemo, et al.
Published: (2024)
Security Weaknesses of Copilot-Generated Code in GitHub Projects: An Empirical Study
by: Fu, Yujia, et al.
Published: (2023)
by: Fu, Yujia, et al.
Published: (2023)
Towards Understanding the Impact of Data Bugs on Deep Learning Models in Software Engineering
by: Shah, Mehil B, et al.
Published: (2024)
by: Shah, Mehil B, et al.
Published: (2024)
Mock Deep Testing: Toward Separate Development of Data and Models for Deep Learning
by: Manke, Ruchira, et al.
Published: (2025)
by: Manke, Ruchira, et al.
Published: (2025)
Exploring Security Practices in Infrastructure as Code: An Empirical Study
by: Verdet, Alexandre, et al.
Published: (2023)
by: Verdet, Alexandre, et al.
Published: (2023)
Bugs in Large Language Models Generated Code: An Empirical Study
by: Tambon, Florian, et al.
Published: (2024)
by: Tambon, Florian, et al.
Published: (2024)
Empirical Characterization of Logging Smells in Machine Learning Code
by: Foalem, Patrick Loic, et al.
Published: (2026)
by: Foalem, Patrick Loic, et al.
Published: (2026)
Empirical Characterization of Logging Smells in Machine Learning Code
by: Foalem, Patrick Loic, et al.
Published: (2026)
by: Foalem, Patrick Loic, et al.
Published: (2026)
Tracing Optimization for Performance Modeling and Regression Detection
by: Shahedi, Kaveh, et al.
Published: (2024)
by: Shahedi, Kaveh, et al.
Published: (2024)
QMon: Monitoring the Execution of Quantum Circuits with Mid-Circuit Measurement and Reset
by: Ma, Ning, et al.
Published: (2025)
by: Ma, Ning, et al.
Published: (2025)
Characterizing Faults in Agentic AI: A Taxonomy of Types, Symptoms, and Root Causes
by: Shah, Mehil B, et al.
Published: (2026)
by: Shah, Mehil B, et al.
Published: (2026)
DeepCodeProbe: Towards Understanding What Models Trained on Code Learn
by: Majdinasab, Vahid, et al.
Published: (2024)
by: Majdinasab, Vahid, et al.
Published: (2024)
On the Effectiveness of Log Representation for Log-based Anomaly Detection
by: Wu, Xingfang, et al.
Published: (2023)
by: Wu, Xingfang, et al.
Published: (2023)
RefAgent: A Multi-agent LLM-based Framework for Automatic Software Refactoring
by: Oueslati, Khouloud, et al.
Published: (2025)
by: Oueslati, Khouloud, et al.
Published: (2025)
Leveraging Data Characteristics for Bug Localization in Deep Learning Programs
by: Manke, Ruchira, et al.
Published: (2024)
by: Manke, Ruchira, et al.
Published: (2024)
Understanding Web Application Workloads and Their Applications: Systematic Literature Review and Characterization
by: Aghili, Roozbeh, et al.
Published: (2024)
by: Aghili, Roozbeh, et al.
Published: (2024)
Reputation Gaming in Stack Overflow
by: Mazloomzadeh, Iren, et al.
Published: (2021)
by: Mazloomzadeh, Iren, et al.
Published: (2021)
Evaluating and Enhancing Segmentation Model Robustness with Metamorphic Testing
by: Mzoughi, Seif, et al.
Published: (2025)
by: Mzoughi, Seif, et al.
Published: (2025)
Beyond Functional Correctness: Design Issues in AI IDE-Generated Large-Scale Projects
by: Kashif, Syed Mohammad, et al.
Published: (2026)
by: Kashif, Syed Mohammad, et al.
Published: (2026)
An Insight into Security Code Review with LLMs: Capabilities, Obstacles, and Influential Factors
by: Yu, Jiaxin, et al.
Published: (2024)
by: Yu, Jiaxin, et al.
Published: (2024)
An Efficient Model Maintenance Approach for MLOps
by: Majidi, Forough, et al.
Published: (2024)
by: Majidi, Forough, et al.
Published: (2024)
Machine Learning Robustness: A Primer
by: Braiek, Houssem Ben, et al.
Published: (2024)
by: Braiek, Houssem Ben, et al.
Published: (2024)
Similar Items
-
A Survey of Bugs in AI-Generated Code
by: Gao, Ruofan, et al.
Published: (2025) -
On Developers' Self-Declaration of AI-Generated Code: An Analysis of Practices
by: Kashif, Syed Mohammad, et al.
Published: (2025) -
From Prompting to Verification: How Experience Shapes Vibe Coding Practices
by: Fawzy, Ahmed, et al.
Published: (2026) -
Improving the Robustness of Large Language Models for Code Tasks via Fine-tuning with Perturbed Data
by: Liu, Yang, et al.
Published: (2026) -
PathOCL: Path-Based Prompt Augmentation for OCL Generation with GPT-4
by: Abukhalaf, Seif, et al.
Published: (2024)