Monte Carlo Tree Search for Execution-Guided Program Repair with Large Language Models
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Liang, Yixuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Meta-Prompting Protocol: Orchestrating LLMs via Adversarial Feedback Loops
von: Fu, Fanzhe
Veröffentlicht: (2025)
von: Fu, Fanzhe
Veröffentlicht: (2025)
CoderUJB: An Executable and Unified Java Benchmark for Practical Programming Scenarios
von: Zeng, Zhengran, et al.
Veröffentlicht: (2024)
von: Zeng, Zhengran, et al.
Veröffentlicht: (2024)
Retrieval-augmented code completion for local projects using large language models
von: Hostnik, Marko, et al.
Veröffentlicht: (2024)
von: Hostnik, Marko, et al.
Veröffentlicht: (2024)
Automated Code Generation and Validation for Software Components of Microcontrollers
von: Haug, Sebastian, et al.
Veröffentlicht: (2025)
von: Haug, Sebastian, et al.
Veröffentlicht: (2025)
FEM-Bench: A Structured Scientific Reasoning Benchmark for Evaluating Code-Generating LLMs
von: Mohammadzadeh, Saeed, et al.
Veröffentlicht: (2025)
von: Mohammadzadeh, Saeed, et al.
Veröffentlicht: (2025)
Two Is Better Than One: Rotations Scale LoRAs
von: Guo, Hongcan, et al.
Veröffentlicht: (2025)
von: Guo, Hongcan, et al.
Veröffentlicht: (2025)
On Integrating Large Language Models and Scenario-Based Programming for Improving Software Reliability
von: Berzack, Ayelet, et al.
Veröffentlicht: (2025)
von: Berzack, Ayelet, et al.
Veröffentlicht: (2025)
A transfer learning approach for automatic conflicts detection in software requirement sentence pairs based on dual encoders
von: Wang, Yizheng, et al.
Veröffentlicht: (2025)
von: Wang, Yizheng, et al.
Veröffentlicht: (2025)
AdvFusion: Adapter-based Knowledge Transfer for Code Summarization on Code Language Models
von: Saberi, Iman, et al.
Veröffentlicht: (2023)
von: Saberi, Iman, et al.
Veröffentlicht: (2023)
QHackBench: Benchmarking Large Language Models for Quantum Code Generation Using PennyLane Hackathon Challenges
von: Basit, Abdul, et al.
Veröffentlicht: (2025)
von: Basit, Abdul, et al.
Veröffentlicht: (2025)
Evaluating the Application of SOLID Principles in Modern AI Framework Architectures
von: Shrestha, Jonesh
Veröffentlicht: (2025)
von: Shrestha, Jonesh
Veröffentlicht: (2025)
Preprocessing is All You Need: Boosting the Performance of Log Parsers With a General Preprocessing Framework
von: Qin, Qiaolin, et al.
Veröffentlicht: (2024)
von: Qin, Qiaolin, et al.
Veröffentlicht: (2024)
React-ing to Grace Hopper 200: Five Open-Weights Coding Models, One React Native App, One GH200, One Weekend
von: Potanin, Alex
Veröffentlicht: (2026)
von: Potanin, Alex
Veröffentlicht: (2026)
See-Saw Generative Mechanism for Scalable Recursive Code Generation with Generative AI
von: Vsevolodovna, Ruslan Idelfonso Magaña
Veröffentlicht: (2024)
von: Vsevolodovna, Ruslan Idelfonso Magaña
Veröffentlicht: (2024)
On Augmenting Scenario-Based Modeling with Generative AI
von: Harel, David, et al.
Veröffentlicht: (2024)
von: Harel, David, et al.
Veröffentlicht: (2024)
Formal Methods: From Academia to Industrial Practice. A Travel Guide
von: Huisman, Marieke, et al.
Veröffentlicht: (2020)
von: Huisman, Marieke, et al.
Veröffentlicht: (2020)
Deep Reinforcement Learning Xiangqi Player with Monte Carlo Tree Search
von: Yilmaz, Berk, et al.
Veröffentlicht: (2025)
von: Yilmaz, Berk, et al.
Veröffentlicht: (2025)
Utilization of Pre-trained Language Model for Adapter-based Knowledge Transfer in Software Engineering
von: Saberi, Iman, et al.
Veröffentlicht: (2023)
von: Saberi, Iman, et al.
Veröffentlicht: (2023)
SHAPR: Operationalising Human-AI Collaborative Research Through Structured Knowledge Generation
von: Chan, Ka Ching
Veröffentlicht: (2026)
von: Chan, Ka Ching
Veröffentlicht: (2026)
CELI: Controller-Embedded Language Model Interactions
von: Wagner, Jan-Samuel, et al.
Veröffentlicht: (2024)
von: Wagner, Jan-Samuel, et al.
Veröffentlicht: (2024)
Supporting software engineering tasks with agentic AI: Demonstration on document retrieval and test scenario generation
von: Kica, Marian, et al.
Veröffentlicht: (2026)
von: Kica, Marian, et al.
Veröffentlicht: (2026)
Anchor Attention, Small Cache: Code Generation with Large Language Models
von: Zhang, Xiangyu, et al.
Veröffentlicht: (2024)
von: Zhang, Xiangyu, et al.
Veröffentlicht: (2024)
Kodezi Chronos: A Debugging-First Language Model for Repository-Scale Code Understanding
von: Khan, Ishraq, et al.
Veröffentlicht: (2025)
von: Khan, Ishraq, et al.
Veröffentlicht: (2025)
GenAIOps for GenAI Model-Agility
von: Ueno, Ken, et al.
Veröffentlicht: (2024)
von: Ueno, Ken, et al.
Veröffentlicht: (2024)
The Right Prompts for the Job: Repair Code-Review Defects with Large Language Model
von: Zhao, Zelin, et al.
Veröffentlicht: (2023)
von: Zhao, Zelin, et al.
Veröffentlicht: (2023)
Genomic Language Models: Opportunities and Challenges
von: Benegas, Gonzalo, et al.
Veröffentlicht: (2024)
von: Benegas, Gonzalo, et al.
Veröffentlicht: (2024)
cozy: Comparative Symbolic Execution for Binary Programs
von: Helbling, Caleb, et al.
Veröffentlicht: (2025)
von: Helbling, Caleb, et al.
Veröffentlicht: (2025)
Parameter Tuning of the Firefly Algorithm by Standard Monte Carlo and Quasi-Monte Carlo Methods
von: Joy, Geethu, et al.
Veröffentlicht: (2024)
von: Joy, Geethu, et al.
Veröffentlicht: (2024)
GraphSkill: Documentation-Guided Hierarchical Retrieval-Augmented Coding for Complex Graph Reasoning
von: Wang, Fali, et al.
Veröffentlicht: (2026)
von: Wang, Fali, et al.
Veröffentlicht: (2026)
Compilation of Generalized Matrix Chains with Symbolic Sizes
von: López, Francisco, et al.
Veröffentlicht: (2025)
von: López, Francisco, et al.
Veröffentlicht: (2025)
The Llama 4 Herd: Architecture, Training, Evaluation, and Deployment Notes
von: arXiv, Redacted by
Veröffentlicht: (2026)
von: arXiv, Redacted by
Veröffentlicht: (2026)
Quality Issues in Machine Learning Software Systems
von: Côté, Pierre-Olivier, et al.
Veröffentlicht: (2023)
von: Côté, Pierre-Olivier, et al.
Veröffentlicht: (2023)
Chain-Oriented Objective Logic with Neural Network Feedback Control and Cascade Filtering for Dynamic Multi-DSL Regulation
von: Han, Jipeng
Veröffentlicht: (2024)
von: Han, Jipeng
Veröffentlicht: (2024)
Autonomous QA Agent: A Retrieval-Augmented Framework for Reliable Selenium Script Generation
von: Vali, Dudekula Kasim
Veröffentlicht: (2025)
von: Vali, Dudekula Kasim
Veröffentlicht: (2025)
Unraveling Media Perspectives: A Comprehensive Methodology Combining Large Language Models, Topic Modeling, Sentiment Analysis, and Ontology Learning to Analyse Media Bias
von: Jähde, Orlando, et al.
Veröffentlicht: (2025)
von: Jähde, Orlando, et al.
Veröffentlicht: (2025)
Piloting Copilot, Codex, and StarCoder2: Hot Temperature, Cold Prompts, or Black Magic?
von: Döderlein, Jean-Baptiste, et al.
Veröffentlicht: (2022)
von: Döderlein, Jean-Baptiste, et al.
Veröffentlicht: (2022)
Hallucination Detection in Large Language Models with Metamorphic Relations
von: Yang, Borui, et al.
Veröffentlicht: (2025)
von: Yang, Borui, et al.
Veröffentlicht: (2025)
Building Whitespace-Sensitive Languages Using Whitespace-Insensitive Components
von: Hellwig, Alexander, et al.
Veröffentlicht: (2025)
von: Hellwig, Alexander, et al.
Veröffentlicht: (2025)
Tool-Genesis: A Task-Driven Tool Creation Benchmark for Self-Evolving Language Agent
von: Xia, Bowei, et al.
Veröffentlicht: (2026)
von: Xia, Bowei, et al.
Veröffentlicht: (2026)
Evolutionary Computation as Natural Generative AI
von: Shi, Yaxin, et al.
Veröffentlicht: (2025)
von: Shi, Yaxin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
The Meta-Prompting Protocol: Orchestrating LLMs via Adversarial Feedback Loops
von: Fu, Fanzhe
Veröffentlicht: (2025) -
CoderUJB: An Executable and Unified Java Benchmark for Practical Programming Scenarios
von: Zeng, Zhengran, et al.
Veröffentlicht: (2024) -
Retrieval-augmented code completion for local projects using large language models
von: Hostnik, Marko, et al.
Veröffentlicht: (2024) -
Automated Code Generation and Validation for Software Components of Microcontrollers
von: Haug, Sebastian, et al.
Veröffentlicht: (2025) -
FEM-Bench: A Structured Scientific Reasoning Benchmark for Evaluating Code-Generating LLMs
von: Mohammadzadeh, Saeed, et al.
Veröffentlicht: (2025)