An Exploratory Study of Bayesian Prompt Optimization for Test-Driven Code Generation with Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Tomar, Shlok, Deshwal, Aryan, Villalovoz, Ethan, Fazzini, Mattia, Cai, Haipeng, Doppa, Janardhan Rao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Automatically Removing Unnecessary Stubbings from Test Suites
by: Li, Mengzhen, et al.
Published: (2024)
by: Li, Mengzhen, et al.
Published: (2024)
Improving LLM-Driven Test Generation by Learning from Mocking Information
by: Lee, Jamie, et al.
Published: (2026)
by: Lee, Jamie, et al.
Published: (2026)
ICST Tool Competition 2025 -- Self-Driving Car Testing Track
by: Birchler, Christian, et al.
Published: (2025)
by: Birchler, Christian, et al.
Published: (2025)
When Prompt Under-Specification Improves Code Correctness: An Exploratory Study of Prompt Wording and Structure Effects on LLM-Based Code Generation
by: AKLI, Amal, et al.
Published: (2026)
by: AKLI, Amal, et al.
Published: (2026)
APPATCH: Automated Adaptive Prompting Large Language Models for Real-World Software Vulnerability Patching
by: Nong, Yu, et al.
Published: (2024)
by: Nong, Yu, et al.
Published: (2024)
Evaluating Large Language Models for Code Translation: Effects of Prompt Language and Prompt Design
by: Aljagthami, Aamer, et al.
Published: (2025)
by: Aljagthami, Aamer, et al.
Published: (2025)
Are Large Language Models a Threat to Programming Platforms? An Exploratory Study
by: Billah, Md Mustakim, et al.
Published: (2024)
by: Billah, Md Mustakim, et al.
Published: (2024)
Configuring Agentic AI Coding Tools: An Exploratory Study
by: Galster, Matthias, et al.
Published: (2026)
by: Galster, Matthias, et al.
Published: (2026)
An Exploratory Study on Fine-Tuning Large Language Models for Secure Code Generation
by: Li, Junjie, et al.
Published: (2024)
by: Li, Junjie, et al.
Published: (2024)
Tests as Prompt: A Test-Driven-Development Benchmark for LLM Code Generation
by: Cui, Yi
Published: (2025)
by: Cui, Yi
Published: (2025)
Human-Agent versus Human Pull Requests: A Testing-Focused Characterization and Comparison
by: Milanese, Roberto, et al.
Published: (2026)
by: Milanese, Roberto, et al.
Published: (2026)
Automated Harmfulness Testing for Code Large Language Models
by: Tan, Honghao, et al.
Published: (2025)
by: Tan, Honghao, et al.
Published: (2025)
The Impact of Generative AI on Code Expertise Models: An Exploratory Study
by: Cury, Otávio, et al.
Published: (2025)
by: Cury, Otávio, et al.
Published: (2025)
Guidelines to Prompt Large Language Models for Code Generation: An Empirical Characterization
by: Midolo, Alessandro, et al.
Published: (2026)
by: Midolo, Alessandro, et al.
Published: (2026)
An Exploratory Investigation into Code License Infringements in Large Language Model Training Datasets
by: Katzy, Jonathan, et al.
Published: (2024)
by: Katzy, Jonathan, et al.
Published: (2024)
Scaling Test-Driven Code Generation from Functions to Classes: An Empirical Study
by: Liang, Yunhao, et al.
Published: (2026)
by: Liang, Yunhao, et al.
Published: (2026)
Offline Model-Based Optimization via Policy-Guided Gradient Search
by: Chemingui, Yassine, et al.
Published: (2024)
by: Chemingui, Yassine, et al.
Published: (2024)
Optimizing Large Language Model Hyperparameters for Code Generation
by: Arora, Chetan, et al.
Published: (2024)
by: Arora, Chetan, et al.
Published: (2024)
Early Discoveries of Algorithmist I: Promise of Provable Algorithm Synthesis at Scale
by: Kulkarni, Janardhan
Published: (2026)
by: Kulkarni, Janardhan
Published: (2026)
CodeMorph: Mitigating Data Leakage in Large Language Model Assessment
by: Rao, Hongzhou, et al.
Published: (2025)
by: Rao, Hongzhou, et al.
Published: (2025)
Adversarial Attack Classification and Robustness Testing for Large Language Models for Code
by: Liu, Yang, et al.
Published: (2025)
by: Liu, Yang, et al.
Published: (2025)
Toward Linking Declined Proposals and Source Code: An Exploratory Study on the Go Repository
by: Nakashima, Sota, et al.
Published: (2026)
by: Nakashima, Sota, et al.
Published: (2026)
How to Refactor this Code? An Exploratory Study on Developer-ChatGPT Refactoring Conversations
by: AlOmar, Eman Abdullah, et al.
Published: (2024)
by: AlOmar, Eman Abdullah, et al.
Published: (2024)
Fine-Tuning and Prompt Engineering for Large Language Models-based Code Review Automation
by: Pornprasit, Chanathip, et al.
Published: (2024)
by: Pornprasit, Chanathip, et al.
Published: (2024)
DiffSpec: Differential Testing with LLMs using Natural Language Specifications and Code Artifacts
by: Rao, Nikitha, et al.
Published: (2024)
by: Rao, Nikitha, et al.
Published: (2024)
LLM-Based Test-Driven Interactive Code Generation: User Study and Empirical Evaluation
by: Fakhoury, Sarah, et al.
Published: (2024)
by: Fakhoury, Sarah, et al.
Published: (2024)
Automated Test Suite Enhancement Using Large Language Models with Few-shot Prompting
by: Chudic, Alex, et al.
Published: (2026)
by: Chudic, Alex, et al.
Published: (2026)
Identification and Optimization of Redundant Code Using Large Language Models
by: Cynthia, Shamse Tasnim
Published: (2025)
by: Cynthia, Shamse Tasnim
Published: (2025)
Process-based Indicators of Vulnerability Re-Introducing Code Changes: An Exploratory Case Study
by: Shimmi, Samiha, et al.
Published: (2025)
by: Shimmi, Samiha, et al.
Published: (2025)
ABTest: Behavior-Driven Testing for AI Coding Agents
by: Dai, Wuyang, et al.
Published: (2026)
by: Dai, Wuyang, et al.
Published: (2026)
RippleGUItester: Change-Aware Exploratory Testing
by: Su, Yanqi, et al.
Published: (2026)
by: Su, Yanqi, et al.
Published: (2026)
Bug Priority Change Prediction: An Exploratory Study on Apache Software
by: Cai, Guangzong, et al.
Published: (2025)
by: Cai, Guangzong, et al.
Published: (2025)
Regression Testing in Remote and Hybrid Software Teams: An Exploratory Study of Processes, Tools, and Practices
by: Pascoal, Juliane, et al.
Published: (2026)
by: Pascoal, Juliane, et al.
Published: (2026)
LLMParser: An Exploratory Study on Using Large Language Models for Log Parsing
by: Ma, Zeyang, et al.
Published: (2024)
by: Ma, Zeyang, et al.
Published: (2024)
PromSec: Prompt Optimization for Secure Generation of Functional Source Code with Large Language Models (LLMs)
by: Nazzal, Mahmoud, et al.
Published: (2024)
by: Nazzal, Mahmoud, et al.
Published: (2024)
Mutation Testing via Iterative Large Language Model-Driven Scientific Debugging
by: Straubinger, Philipp, et al.
Published: (2025)
by: Straubinger, Philipp, et al.
Published: (2025)
Prompt Optimization for LLM Code Generation via Reinforcement Learning
by: Esfahani, Ali Mohammadi, et al.
Published: (2026)
by: Esfahani, Ali Mohammadi, et al.
Published: (2026)
An Exploratory Eye Tracking Study on How Developers Classify and Debug Python Code in Different Paradigms
by: Flint, Samuel W., et al.
Published: (2025)
by: Flint, Samuel W., et al.
Published: (2025)
The First Prompt Counts the Most! An Evaluation of Large Language Models on Iterative Example-Based Code Generation
by: Fu, Yingjie, et al.
Published: (2024)
by: Fu, Yingjie, et al.
Published: (2024)
An Empirical Study on the Code Refactoring Capability of Large Language Models
by: Cordeiro, Jonathan, et al.
Published: (2024)
by: Cordeiro, Jonathan, et al.
Published: (2024)
Similar Items
-
Automatically Removing Unnecessary Stubbings from Test Suites
by: Li, Mengzhen, et al.
Published: (2024) -
Improving LLM-Driven Test Generation by Learning from Mocking Information
by: Lee, Jamie, et al.
Published: (2026) -
ICST Tool Competition 2025 -- Self-Driving Car Testing Track
by: Birchler, Christian, et al.
Published: (2025) -
When Prompt Under-Specification Improves Code Correctness: An Exploratory Study of Prompt Wording and Structure Effects on LLM-Based Code Generation
by: AKLI, Amal, et al.
Published: (2026) -
APPATCH: Automated Adaptive Prompting Large Language Models for Real-World Software Vulnerability Patching
by: Nong, Yu, et al.
Published: (2024)