Prompt Engineering or Fine-Tuning: An Empirical Assessment of LLMs for Code
Fuente:
arXiv
Saved in:
| Main Authors: | Shin, Jiho, Tang, Clark, Mohati, Tahmineh, Nayebi, Maleknaz, Wang, Song, Hemmati, Hadi |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Domain Adaptation for Code Model-based Unit Test Case Generation
by: Shin, Jiho, et al.
Published: (2023)
by: Shin, Jiho, et al.
Published: (2023)
Extension Decisions in Open Source Software Ecosystem
by: Onagh, Elmira, et al.
Published: (2025)
by: Onagh, Elmira, et al.
Published: (2025)
Assessing Evaluation Metrics for Neural Test Oracle Generation
by: Shin, Jiho, et al.
Published: (2023)
by: Shin, Jiho, et al.
Published: (2023)
Analysis of Marketed versus Not-marketed Mobile App Releases
by: Nayebi, Maleknaz, et al.
Published: (2024)
by: Nayebi, Maleknaz, et al.
Published: (2024)
Negative Results of Image Processing for Identifying Duplicate Questions on Stack Overflow
by: Ahmed, Faiz, et al.
Published: (2024)
by: Ahmed, Faiz, et al.
Published: (2024)
The Impact of Foundational Models on Patient-Centric e-Health Systems
by: Onagh, Elmira, et al.
Published: (2025)
by: Onagh, Elmira, et al.
Published: (2025)
ImageR: Enhancing Bug Report Clarity by Screenshots
by: Tan, Xuchen, et al.
Published: (2025)
by: Tan, Xuchen, et al.
Published: (2025)
Recommending and Release Planning of User-Driven Functionality Deletion for Mobile Apps
by: Nayebi, Maleknaz, et al.
Published: (2024)
by: Nayebi, Maleknaz, et al.
Published: (2024)
Retrieval-Augmented Test Generation: How Far Are We?
by: Shin, Jiho, et al.
Published: (2024)
by: Shin, Jiho, et al.
Published: (2024)
GitHub Marketplace: Driving Automation and Fostering Innovation in Software Development
by: Saroar, SK. Golam, et al.
Published: (2025)
by: Saroar, SK. Golam, et al.
Published: (2025)
GitHub Marketplace for Automation and Innovation in Software Production
by: Saroar, SK Golam, et al.
Published: (2024)
by: Saroar, SK Golam, et al.
Published: (2024)
More Insight from Being More Focused: Analysis of Clustered Market Apps
by: Nayebi, Maleknaz, et al.
Published: (2024)
by: Nayebi, Maleknaz, et al.
Published: (2024)
Inferring Questions from Programming Screenshots
by: Ahmed, Faiz, et al.
Published: (2025)
by: Ahmed, Faiz, et al.
Published: (2025)
One Documentation Does Not Fit All: Case Study of TensorFlow Documentation
by: Thirimanne, Sharuka Promodya, et al.
Published: (2025)
by: Thirimanne, Sharuka Promodya, et al.
Published: (2025)
Demystifying Errors in LLM Reasoning Traces: An Empirical Study of Code Execution Simulation
by: Abdollahi, Mohammad, et al.
Published: (2025)
by: Abdollahi, Mohammad, et al.
Published: (2025)
Automated Prompt Engineering for Cost-Effective Code Generation Using Evolutionary Algorithm
by: Taherkhani, Hamed, et al.
Published: (2024)
by: Taherkhani, Hamed, et al.
Published: (2024)
An Empirical Study on Bug Severity Estimation using Source Code Metrics and Static Analysis
by: Mashhadi, Ehsan, et al.
Published: (2022)
by: Mashhadi, Ehsan, et al.
Published: (2022)
Enhancing LLM-Based Code Generation with Complexity Metrics: A Feedback-Driven Approach
by: Sepidband, Melika, et al.
Published: (2025)
by: Sepidband, Melika, et al.
Published: (2025)
Fine-Tuning and Prompt Engineering for Large Language Models-based Code Review Automation
by: Pornprasit, Chanathip, et al.
Published: (2024)
by: Pornprasit, Chanathip, et al.
Published: (2024)
Examining Ownership Models in Software Teams: A Systematic Literature Review and a Replication Study
by: Koana, Umme Ayman, et al.
Published: (2024)
by: Koana, Umme Ayman, et al.
Published: (2024)
The Good, the Bad, and the Missing: Neural Code Generation for Machine Learning Tasks
by: Shin, Jiho, et al.
Published: (2023)
by: Shin, Jiho, et al.
Published: (2023)
An Empirical Study on the Effects of System Prompts in Instruction-Tuned Models for Code Generation
by: Cheng, Zaiyu, et al.
Published: (2026)
by: Cheng, Zaiyu, et al.
Published: (2026)
Delving into Parameter-Efficient Fine-Tuning in Code Change Learning: An Empirical Study
by: Liu, Shuo, et al.
Published: (2024)
by: Liu, Shuo, et al.
Published: (2024)
Toward Automated Validation of Language Model Synthesized Test Cases using Semantic Entropy
by: Taherkhani, Hamed, et al.
Published: (2024)
by: Taherkhani, Hamed, et al.
Published: (2024)
A Systematic Mapping Study of Crowd Knowledge Enhanced Software Engineering Research Using Stack Overflow
by: Tanzil, Minaoar, et al.
Published: (2024)
by: Tanzil, Minaoar, et al.
Published: (2024)
On Fixing Insecure AI-Generated Code through Model Fine-Tuning and Prompting Strategies
by: Jahromi, Ali Soltanian Fard, et al.
Published: (2026)
by: Jahromi, Ali Soltanian Fard, et al.
Published: (2026)
Investigating the Role of LLMs Hyperparameter Tuning and Prompt Engineering to Support Domain Modeling
by: Bulhakov, Vladyslav, et al.
Published: (2025)
by: Bulhakov, Vladyslav, et al.
Published: (2025)
SecVulEval: Benchmarking LLMs for Real-World C/C++ Vulnerability Detection
by: Ahmed, Md Basim Uddin, et al.
Published: (2025)
by: Ahmed, Md Basim Uddin, et al.
Published: (2025)
Can ChatGPT Support Developers? An Empirical Evaluation of Large Language Models for Code Generation
by: Jin, Kailun, et al.
Published: (2024)
by: Jin, Kailun, et al.
Published: (2024)
FlakyFix: Using Large Language Models for Predicting Flaky Test Fix Categories and Test Code Repair
by: Fatima, Sakina, et al.
Published: (2023)
by: Fatima, Sakina, et al.
Published: (2023)
Engineering Pitfalls in AI Coding Tools: An Empirical Study of Bugs in Claude Code, Codex, and Gemini CLI
by: Zhang, Ruixin, et al.
Published: (2026)
by: Zhang, Ruixin, et al.
Published: (2026)
Integrating Symbolic Execution into the Fine-Tuning of Code-Generating LLMs
by: Sakharova, Marina, et al.
Published: (2025)
by: Sakharova, Marina, et al.
Published: (2025)
FGIT: Fault-Guided Fine-Tuning for Code Generation
by: Fan, Lishui, et al.
Published: (2025)
by: Fan, Lishui, et al.
Published: (2025)
Program Slicing in the Era of Large Language Models
by: Shahandashti, Kimya Khakzad, et al.
Published: (2024)
by: Shahandashti, Kimya Khakzad, et al.
Published: (2024)
When Fine-Tuning LLMs Meets Data Privacy: An Empirical Study of Federated Learning in LLM-Based Program Repair
by: Luo, Wenqiang, et al.
Published: (2024)
by: Luo, Wenqiang, et al.
Published: (2024)
RGFL: Reasoning Guided Fault Localization for Automated Program Repair Using Large Language Models
by: Sepidband, Melika, et al.
Published: (2026)
by: Sepidband, Melika, et al.
Published: (2026)
Automatic Instantiation of Assurance Cases from Patterns Using Large Language Models
by: Odu, Oluwafemi, et al.
Published: (2024)
by: Odu, Oluwafemi, et al.
Published: (2024)
Automated Prompt Generation for Code Intelligence: An Empirical study and Experience in WeChat
by: Ji, Kexing, et al.
Published: (2025)
by: Ji, Kexing, et al.
Published: (2025)
LLMs are Bug Replicators: An Empirical Study on LLMs' Capability in Completing Bug-prone Code
by: Guo, Liwei, et al.
Published: (2025)
by: Guo, Liwei, et al.
Published: (2025)
Fine-Tuning Models for Automated Code Review Feedback
by: Kumar, Smitha S, et al.
Published: (2026)
by: Kumar, Smitha S, et al.
Published: (2026)
Similar Items
-
Domain Adaptation for Code Model-based Unit Test Case Generation
by: Shin, Jiho, et al.
Published: (2023) -
Extension Decisions in Open Source Software Ecosystem
by: Onagh, Elmira, et al.
Published: (2025) -
Assessing Evaluation Metrics for Neural Test Oracle Generation
by: Shin, Jiho, et al.
Published: (2023) -
Analysis of Marketed versus Not-marketed Mobile App Releases
by: Nayebi, Maleknaz, et al.
Published: (2024) -
Negative Results of Image Processing for Identifying Duplicate Questions on Stack Overflow
by: Ahmed, Faiz, et al.
Published: (2024)