Automated Repair of Ambiguous Problem Descriptions for LLM-Based Code Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Jia, Haoxiang, Morris, Robbie, Ye, He, Sarro, Federica, Mechtaev, Sergey |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Compressing Code Context for LLM-based Issue Resolution
by: Jia, Haoxiang, et al.
Published: (2026)
by: Jia, Haoxiang, et al.
Published: (2026)
Statistical Independence Aware Caching for LLM Workflows
by: Dai, Yihan, et al.
Published: (2025)
by: Dai, Yihan, et al.
Published: (2025)
Defusing Logic Bombs in Symbolic Execution with LLM-Generated Ghost Code
by: Bouras, Dimitrios Stamatios, et al.
Published: (2026)
by: Bouras, Dimitrios Stamatios, et al.
Published: (2026)
The Fact Selection Problem in LLM-Based Program Repair
by: Parasaram, Nikhil, et al.
Published: (2024)
by: Parasaram, Nikhil, et al.
Published: (2024)
Reducing Hallucinations in LLM-Generated Code via Semantic Triangulation
by: Dai, Yihan, et al.
Published: (2025)
by: Dai, Yihan, et al.
Published: (2025)
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions
by: Larbi, Maya, et al.
Published: (2025)
by: Larbi, Maya, et al.
Published: (2025)
Environment-in-the-Loop: Rethinking Code Migration with LLM-based Agents
by: Li, Xiang, et al.
Published: (2026)
by: Li, Xiang, et al.
Published: (2026)
HerAgent: Rethinking the Automated Environment Deployment via Hierarchical Test Pyramid
by: Li, Xiang, et al.
Published: (2026)
by: Li, Xiang, et al.
Published: (2026)
Generative AI for Testing of Autonomous Driving Systems: A Survey
by: Song, Qunying, et al.
Published: (2025)
by: Song, Qunying, et al.
Published: (2025)
Understanding Code Agent Behaviour: An Empirical Study of Success and Failure Trajectories
by: Majgaonkar, Oorja, et al.
Published: (2025)
by: Majgaonkar, Oorja, et al.
Published: (2025)
Comparing AI Coding Agents: A Task-Stratified Analysis of Pull Request Acceptance
by: Pinna, Giovanni, et al.
Published: (2026)
by: Pinna, Giovanni, et al.
Published: (2026)
LLM-Guided Genetic Improvement: Envisioning Semantic Aware Automated Software Evolution
by: Even-Mendoza, Karine, et al.
Published: (2025)
by: Even-Mendoza, Karine, et al.
Published: (2025)
Echo: Graph-Enhanced Retrieval and Execution Feedback for Issue Reproduction Test Generation
by: Fei, Zhiwei, et al.
Published: (2026)
by: Fei, Zhiwei, et al.
Published: (2026)
LLM-Based Misconfiguration Detection for AWS Serverless Computing
by: Wen, Jinfeng, et al.
Published: (2024)
by: Wen, Jinfeng, et al.
Published: (2024)
An Iterative Test-and-Repair Framework for Competitive Code Generation
by: Tang, Lingxiao, et al.
Published: (2026)
by: Tang, Lingxiao, et al.
Published: (2026)
Psychological Safety Framework in Pull-based Open Source Projects
by: Sesari, Emeralda, et al.
Published: (2025)
by: Sesari, Emeralda, et al.
Published: (2025)
Understanding Fairness in Software Engineering: Insights from Stack Exchange
by: Sesari, Emeralda, et al.
Published: (2024)
by: Sesari, Emeralda, et al.
Published: (2024)
BayesInsights: Modelling Software Delivery and Developer Experience with Bayesian Networks at Bloomberg
by: Kirbas, Serkan, et al.
Published: (2026)
by: Kirbas, Serkan, et al.
Published: (2026)
It is Giving Major Satisfaction: Why Fairness Matters for Software Practitioners
by: Sesari, Emeralda, et al.
Published: (2024)
by: Sesari, Emeralda, et al.
Published: (2024)
DebugRepair: Enhancing LLM-Based Automated Program Repair via Self-Directed Debugging
by: Wu, Linhao, et al.
Published: (2026)
by: Wu, Linhao, et al.
Published: (2026)
SCPatcher: Automated Smart Contract Code Repair via Retrieval-Augmented Generation and Knowledge Graph
by: Li, Xiaoqi, et al.
Published: (2026)
by: Li, Xiaoqi, et al.
Published: (2026)
On the Compression of Language Models for Code: An Empirical Study on CodeBERT
by: d'Aloisio, Giordano, et al.
Published: (2024)
by: d'Aloisio, Giordano, et al.
Published: (2024)
HotBugs.jar: A Benchmark of Hot Fixes for Time-Critical Bugs
by: Hanna, Carol, et al.
Published: (2025)
by: Hanna, Carol, et al.
Published: (2025)
Hot Fixing Software: A Comprehensive Review of Terminology, Techniques, and Applications
by: Hanna, Carol, et al.
Published: (2024)
by: Hanna, Carol, et al.
Published: (2024)
Unveiling Overlooked Performance Variance in Serverless Computing
by: Wen, Jinfeng, et al.
Published: (2023)
by: Wen, Jinfeng, et al.
Published: (2023)
From Research to Practice: An Interactive Rapid Review of Autonomous Driving System Testing in Industry
by: Song, Qunying, et al.
Published: (2026)
by: Song, Qunying, et al.
Published: (2026)
On the Influence of Data Resampling for Deep Learning-Based Log Anomaly Detection: Insights and Recommendations
by: Ma, Xiaoxue, et al.
Published: (2024)
by: Ma, Xiaoxue, et al.
Published: (2024)
Prometheus: Towards Long-Horizon Codebase Navigation for Repository-Level Problem Solving
by: Pan, Yue, et al.
Published: (2025)
by: Pan, Yue, et al.
Published: (2025)
Test-based Patch Clustering for Automatically-Generated Patches Assessment
by: Martinez, Matias, et al.
Published: (2022)
by: Martinez, Matias, et al.
Published: (2022)
FairRF: Multi-Objective Search for Single and Intersectional Software Fairness
by: d'Alosio, Giordano, et al.
Published: (2026)
by: d'Alosio, Giordano, et al.
Published: (2026)
GuardRails: Automated Suggestions for Clarifying Ambiguous Purpose Statements
by: Pawagi, Mrigank, et al.
Published: (2023)
by: Pawagi, Mrigank, et al.
Published: (2023)
RulER: Automated Rule-Based Semantic Error Localization and Repair for Code Translation
by: Jin, Shuo, et al.
Published: (2025)
by: Jin, Shuo, et al.
Published: (2025)
HEJ-Robust: A Robustness Benchmark for LLM-Based Automated Program Repair
by: Rabbi, Fazle, et al.
Published: (2026)
by: Rabbi, Fazle, et al.
Published: (2026)
A Survey of Code Review Benchmarks and Evaluation Practices in Pre-LLM and LLM Era
by: Khan, Taufiqul Islam, et al.
Published: (2026)
by: Khan, Taufiqul Islam, et al.
Published: (2026)
Defective Task Descriptions in LLM-Based Code Generation: Detection and Analysis
by: Akli, Amal, et al.
Published: (2026)
by: Akli, Amal, et al.
Published: (2026)
No Man is an Island: Towards Fully Automatic Programming by Code Search, Code Generation and Program Repair
by: Zhang, Quanjun, et al.
Published: (2024)
by: Zhang, Quanjun, et al.
Published: (2024)
Using Semantic Distance to Estimate Uncertainty in LLM-Based Code Generation
by: He, Weilin, et al.
Published: (2026)
by: He, Weilin, et al.
Published: (2026)
HoarePrompt: Structural Reasoning About Program Correctness in Natural Language
by: Bouras, Dimitrios Stamatios, et al.
Published: (2025)
by: Bouras, Dimitrios Stamatios, et al.
Published: (2025)
ArkEval: Benchmarking and Evaluating Automated CodeRepair for ArkTS
by: Xie, Bang, et al.
Published: (2026)
by: Xie, Bang, et al.
Published: (2026)
How Do Generative Models Draw a Software Engineer? A Case Study on Stable Diffusion Bias
by: Fadahunsi, Tosin, et al.
Published: (2025)
by: Fadahunsi, Tosin, et al.
Published: (2025)
Similar Items
-
Compressing Code Context for LLM-based Issue Resolution
by: Jia, Haoxiang, et al.
Published: (2026) -
Statistical Independence Aware Caching for LLM Workflows
by: Dai, Yihan, et al.
Published: (2025) -
Defusing Logic Bombs in Symbolic Execution with LLM-Generated Ghost Code
by: Bouras, Dimitrios Stamatios, et al.
Published: (2026) -
The Fact Selection Problem in LLM-Based Program Repair
by: Parasaram, Nikhil, et al.
Published: (2024) -
Reducing Hallucinations in LLM-Generated Code via Semantic Triangulation
by: Dai, Yihan, et al.
Published: (2025)