The Prompt Alchemist: Automated LLM-Tailored Prompt Optimization for Test Case Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gao, Shuzheng, Wang, Chaozheng, Gao, Cuiyun, Jiao, Xiaoqian, Chong, Chun Yong, Gao, Shan, Lyu, Michael |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Search-Based LLMs for Code Optimization
von: Gao, Shuzheng, et al.
Veröffentlicht: (2024)
von: Gao, Shuzheng, et al.
Veröffentlicht: (2024)
SEER: Enhancing Chain-of-Thought Code Generation through Self-Exploring Deep Reasoning
von: Gao, Shuzheng, et al.
Veröffentlicht: (2025)
von: Gao, Shuzheng, et al.
Veröffentlicht: (2025)
A Systematic Evaluation of Large Code Models in API Suggestion: When, Which, and How
von: Wang, Chaozheng, et al.
Veröffentlicht: (2024)
von: Wang, Chaozheng, et al.
Veröffentlicht: (2024)
Automated Prompt Generation for Code Intelligence: An Empirical study and Experience in WeChat
von: Ji, Kexing, et al.
Veröffentlicht: (2025)
von: Ji, Kexing, et al.
Veröffentlicht: (2025)
ComplexCodeEval: A Benchmark for Evaluating Large Code Models on More Complex Code
von: Feng, Jia, et al.
Veröffentlicht: (2024)
von: Feng, Jia, et al.
Veröffentlicht: (2024)
RAG or Fine-tuning? A Comparative Study on LCMs-based Code Completion in Industry
von: Wang, Chaozheng, et al.
Veröffentlicht: (2025)
von: Wang, Chaozheng, et al.
Veröffentlicht: (2025)
LLM-Based Test Case Generation in DBMS through Monte Carlo Tree Search
von: Chen, Yujia, et al.
Veröffentlicht: (2026)
von: Chen, Yujia, et al.
Veröffentlicht: (2026)
SR-Eval: Evaluating LLMs on Code Generation under Stepwise Requirement Refinement
von: Zhan, Zexun, et al.
Veröffentlicht: (2025)
von: Zhan, Zexun, et al.
Veröffentlicht: (2025)
Can LLMs Replace Human Evaluators? An Empirical Study of LLM-as-a-Judge in Software Engineering
von: Wang, Ruiqi, et al.
Veröffentlicht: (2025)
von: Wang, Ruiqi, et al.
Veröffentlicht: (2025)
Empirical Study of Code Large Language Models for Binary Security Patch Detection
von: Li, Qingyuan, et al.
Veröffentlicht: (2025)
von: Li, Qingyuan, et al.
Veröffentlicht: (2025)
AXIOM: Benchmarking LLM-as-a-Judge for Code via Rule-Based Perturbation and Multisource Quality Calibration
von: Wang, Ruiqi, et al.
Veröffentlicht: (2025)
von: Wang, Ruiqi, et al.
Veröffentlicht: (2025)
TREAT: A Code LLMs Trustworthiness / Reliability Evaluation and Testing Framework
von: Gao, Shuzheng, et al.
Veröffentlicht: (2025)
von: Gao, Shuzheng, et al.
Veröffentlicht: (2025)
A Taxonomy of Prompt Defects in LLM Systems
von: Tian, Haoye, et al.
Veröffentlicht: (2025)
von: Tian, Haoye, et al.
Veröffentlicht: (2025)
CodeVisionary: An Agent-based Framework for Evaluating Large Language Models in Code Generation
von: Wang, Xinchen, et al.
Veröffentlicht: (2025)
von: Wang, Xinchen, et al.
Veröffentlicht: (2025)
Integrating Rules and Semantics for LLM-Based C-to-Rust Translation
von: Luo, Feng, et al.
Veröffentlicht: (2025)
von: Luo, Feng, et al.
Veröffentlicht: (2025)
Cascaded Code Editing: Large-Small Model Collaboration for Effective and Efficient Code Editing
von: Wang, Chaozheng, et al.
Veröffentlicht: (2026)
von: Wang, Chaozheng, et al.
Veröffentlicht: (2026)
What Makes Good In-context Demonstrations for Code Intelligence Tasks with LLMs?
von: Gao, Shuzheng, et al.
Veröffentlicht: (2023)
von: Gao, Shuzheng, et al.
Veröffentlicht: (2023)
SCALE: Constructing Structured Natural Language Comment Trees for Software Vulnerability Detection
von: Wen, Xin-Cheng, et al.
Veröffentlicht: (2024)
von: Wen, Xin-Cheng, et al.
Veröffentlicht: (2024)
Repo2Run: Automated Building Executable Environment for Code Repository at Scale
von: Hu, Ruida, et al.
Veröffentlicht: (2025)
von: Hu, Ruida, et al.
Veröffentlicht: (2025)
Vul-R2: A Reasoning LLM for Automated Vulnerability Repair
von: Wen, Xin-Cheng, et al.
Veröffentlicht: (2025)
von: Wen, Xin-Cheng, et al.
Veröffentlicht: (2025)
Learning in the Wild: Towards Leveraging Unlabeled Data for Effectively Tuning Pre-trained Code Models
von: Gao, Shuzheng, et al.
Veröffentlicht: (2024)
von: Gao, Shuzheng, et al.
Veröffentlicht: (2024)
Learning to Ask: When LLM Agents Meet Unclear Instruction
von: Wang, Wenxuan, et al.
Veröffentlicht: (2024)
von: Wang, Wenxuan, et al.
Veröffentlicht: (2024)
APRIL: API Synthesis with Automatic Prompt Optimization and Reinforcement Learning
von: Zhong, Hua, et al.
Veröffentlicht: (2025)
von: Zhong, Hua, et al.
Veröffentlicht: (2025)
Evaluating LLM-Based 0-to-1 Software Generation in End-to-End CLI Tool Scenarios
von: Hu, Ruida, et al.
Veröffentlicht: (2026)
von: Hu, Ruida, et al.
Veröffentlicht: (2026)
Validating LLM-Generated Programs with Metamorphic Prompt Testing
von: Wang, Xiaoyin, et al.
Veröffentlicht: (2024)
von: Wang, Xiaoyin, et al.
Veröffentlicht: (2024)
SPENCER: Self-Adaptive Model Distillation for Efficient Code Retrieval
von: Gu, Wenchao, et al.
Veröffentlicht: (2025)
von: Gu, Wenchao, et al.
Veröffentlicht: (2025)
A Deep Dive into Retrieval-Augmented Generation for Code Completion: Experience on WeChat
von: Yang, Zezhou, et al.
Veröffentlicht: (2025)
von: Yang, Zezhou, et al.
Veröffentlicht: (2025)
Tests as Prompt: A Test-Driven-Development Benchmark for LLM Code Generation
von: Cui, Yi
Veröffentlicht: (2025)
von: Cui, Yi
Veröffentlicht: (2025)
What Prompts Don't Say: Understanding and Managing Underspecification in LLM Prompts
von: Yang, Chenyang, et al.
Veröffentlicht: (2025)
von: Yang, Chenyang, et al.
Veröffentlicht: (2025)
Learner-Tailored Program Repair: A Solution Generator with Iterative Edit-Driven Retrieval Enhancement
von: Dai, Zhenlong, et al.
Veröffentlicht: (2026)
von: Dai, Zhenlong, et al.
Veröffentlicht: (2026)
PromptPex: Automatic Test Generation for Language Model Prompts
von: Sharma, Reshabh K, et al.
Veröffentlicht: (2025)
von: Sharma, Reshabh K, et al.
Veröffentlicht: (2025)
Evaluating LLM-Based Goal Extraction in Requirements Engineering: Prompting Strategies and Their Limitations
von: Arnaudo, Anna, et al.
Veröffentlicht: (2026)
von: Arnaudo, Anna, et al.
Veröffentlicht: (2026)
OBsmith: LLM-Powered JavaScript Obfuscator Testing
von: Jiang, Shan, et al.
Veröffentlicht: (2025)
von: Jiang, Shan, et al.
Veröffentlicht: (2025)
Exploring Multi-Lingual Bias of Large Code Models in Code Generation
von: Wang, Chaozheng, et al.
Veröffentlicht: (2024)
von: Wang, Chaozheng, et al.
Veröffentlicht: (2024)
Evaluating Repository-level Software Documentation via Question Answering and Feature-Driven Development
von: Wang, Xinchen, et al.
Veröffentlicht: (2026)
von: Wang, Xinchen, et al.
Veröffentlicht: (2026)
Weakly Supervised Vulnerability Localization via Multiple Instance Learning
von: Gu, Wenchao, et al.
Veröffentlicht: (2025)
von: Gu, Wenchao, et al.
Veröffentlicht: (2025)
LecPrompt: A Prompt-based Approach for Logical Error Correction with CodeBERT
von: Xu, Zhenyu, et al.
Veröffentlicht: (2024)
von: Xu, Zhenyu, et al.
Veröffentlicht: (2024)
ECO: Enhanced Code Optimization via Performance-Aware Prompting for Code-LLMs
von: Kim, Su-Hyeon, et al.
Veröffentlicht: (2025)
von: Kim, Su-Hyeon, et al.
Veröffentlicht: (2025)
CodeCrash: Exposing LLM Fragility to Misleading Natural Language in Code Reasoning
von: Lam, Man Ho, et al.
Veröffentlicht: (2025)
von: Lam, Man Ho, et al.
Veröffentlicht: (2025)
AEGIS: An Agent-based Framework for General Bug Reproduction from Issue Descriptions
von: Wang, Xinchen, et al.
Veröffentlicht: (2024)
von: Wang, Xinchen, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Search-Based LLMs for Code Optimization
von: Gao, Shuzheng, et al.
Veröffentlicht: (2024) -
SEER: Enhancing Chain-of-Thought Code Generation through Self-Exploring Deep Reasoning
von: Gao, Shuzheng, et al.
Veröffentlicht: (2025) -
A Systematic Evaluation of Large Code Models in API Suggestion: When, Which, and How
von: Wang, Chaozheng, et al.
Veröffentlicht: (2024) -
Automated Prompt Generation for Code Intelligence: An Empirical study and Experience in WeChat
von: Ji, Kexing, et al.
Veröffentlicht: (2025) -
ComplexCodeEval: A Benchmark for Evaluating Large Code Models on More Complex Code
von: Feng, Jia, et al.
Veröffentlicht: (2024)