Reinforcement Learning from Compiler and Language Server Feedback
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhang, Yifan, Contributors, Lanser |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Benchmarking Large Language Models for ABAP Code Generation: An Empirical Study on Iterative Improvement by Compiler Feedback
par: Wallraven, Stephan, et autres
Publié: (2026)
par: Wallraven, Stephan, et autres
Publié: (2026)
LEGO-Compiler: Enhancing Neural Compilation Through Translation Composability
par: Zhang, Shuoming, et autres
Publié: (2025)
par: Zhang, Shuoming, et autres
Publié: (2025)
The New Compiler Stack: A Survey on the Synergy of LLMs and Compilers
par: Zhang, Shuoming, et autres
Publié: (2026)
par: Zhang, Shuoming, et autres
Publié: (2026)
PerfCodeGen: Improving Performance of LLM Generated Code with Execution Feedback
par: Peng, Yun, et autres
Publié: (2024)
par: Peng, Yun, et autres
Publié: (2024)
LangGPT: Rethinking Structured Reusable Prompt Design Framework for LLMs from the Programming Language
par: Wang, Ming, et autres
Publié: (2024)
par: Wang, Ming, et autres
Publié: (2024)
LLMs Lean on Priors, Not Programming Language Semantics
par: Thimmaiah, Aditya, et autres
Publié: (2025)
par: Thimmaiah, Aditya, et autres
Publié: (2025)
SLMFix: Leveraging Small Language Models for Error Fixing with Reinforcement Learning
par: Fu, David Jiahao, et autres
Publié: (2025)
par: Fu, David Jiahao, et autres
Publié: (2025)
Benchmarking LLM Code Generation for Audio Programming with Visual Dataflow Languages
par: Zhang, William, et autres
Publié: (2024)
par: Zhang, William, et autres
Publié: (2024)
CodeMind: Evaluating Large Language Models for Code Reasoning
par: Liu, Changshu, et autres
Publié: (2024)
par: Liu, Changshu, et autres
Publié: (2024)
VisCoder2: Building Multi-Language Visualization Coding Agents
par: Ni, Yuansheng, et autres
Publié: (2025)
par: Ni, Yuansheng, et autres
Publié: (2025)
Refactoring Programs Using Large Language Models with Few-Shot Examples
par: Shirafuji, Atsushi, et autres
Publié: (2023)
par: Shirafuji, Atsushi, et autres
Publié: (2023)
BuildBench: Benchmarking LLM Agents on Compiling Real-World Open-Source Software
par: Zhang, Zehua, et autres
Publié: (2025)
par: Zhang, Zehua, et autres
Publié: (2025)
DeepLL: Considering Linear Logic for the Analysis of Deep Learning Experiments
par: Papoulias, Nick
Publié: (2024)
par: Papoulias, Nick
Publié: (2024)
EffiPair: Improving the Efficiency of LLM-generated Code with Relative Contrastive Feedback
par: Hajizadeh, Samira, et autres
Publié: (2026)
par: Hajizadeh, Samira, et autres
Publié: (2026)
A Taxonomy of Prompt Defects in LLM Systems
par: Tian, Haoye, et autres
Publié: (2025)
par: Tian, Haoye, et autres
Publié: (2025)
SuperCoder: Assembly Program Superoptimization with Large Language Models
par: Wei, Anjiang, et autres
Publié: (2025)
par: Wei, Anjiang, et autres
Publié: (2025)
Introduction to Analytical Software Engineering Design Paradigm
par: Houichime, Tarik, et autres
Publié: (2025)
par: Houichime, Tarik, et autres
Publié: (2025)
MoSE: Hierarchical Self-Distillation Enhances Early Layer Embeddings
par: Gurioli, Andrea, et autres
Publié: (2025)
par: Gurioli, Andrea, et autres
Publié: (2025)
debug-gym: A Text-Based Environment for Interactive Debugging
par: Yuan, Xingdi, et autres
Publié: (2025)
par: Yuan, Xingdi, et autres
Publié: (2025)
AutoCode: LLMs as Problem Setters for Competitive Programming
par: Zhou, Shang, et autres
Publié: (2025)
par: Zhou, Shang, et autres
Publié: (2025)
Code Broker: A Multi-Agent System for Automated Code Quality Assessment
par: Attrah, Samer
Publié: (2026)
par: Attrah, Samer
Publié: (2026)
Ranking LLM-Generated Loop Invariants for Program Verification
par: Chakraborty, Saikat, et autres
Publié: (2023)
par: Chakraborty, Saikat, et autres
Publié: (2023)
Comparing large language models and human programmers for generating programming code
par: Hou, Wenpin, et autres
Publié: (2024)
par: Hou, Wenpin, et autres
Publié: (2024)
PPM: Automated Generation of Diverse Programming Problems for Benchmarking Code Generation Models
par: Chen, Simin, et autres
Publié: (2024)
par: Chen, Simin, et autres
Publié: (2024)
Code Repair with LLMs gives an Exploration-Exploitation Tradeoff
par: Tang, Hao, et autres
Publié: (2024)
par: Tang, Hao, et autres
Publié: (2024)
From Code to Correctness: Closing the Last Mile of Code Generation with Hierarchical Debugging
par: Shi, Yuling, et autres
Publié: (2024)
par: Shi, Yuling, et autres
Publié: (2024)
Verus-SpecGym: An Agentic Environment for Evaluating Specification Autoformalization
par: Agarwal, Anmol, et autres
Publié: (2026)
par: Agarwal, Anmol, et autres
Publié: (2026)
A Survey of Neural Code Intelligence: Paradigms, Advances and Beyond
par: Sun, Qiushi, et autres
Publié: (2024)
par: Sun, Qiushi, et autres
Publié: (2024)
Is Self-Repair a Silver Bullet for Code Generation?
par: Olausson, Theo X., et autres
Publié: (2023)
par: Olausson, Theo X., et autres
Publié: (2023)
Analysis of AdvFusion: Adapter-based Multilingual Learning for Code Large Language Models
par: Esmaeili, Amirreza, et autres
Publié: (2025)
par: Esmaeili, Amirreza, et autres
Publié: (2025)
StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
par: Dou, Shihan, et autres
Publié: (2024)
par: Dou, Shihan, et autres
Publié: (2024)
ViC: Virtual Compiler Is All You Need For Assembly Code Search
par: Gao, Zeyu, et autres
Publié: (2024)
par: Gao, Zeyu, et autres
Publié: (2024)
REINFOREST: Reinforcing Semantic Code Similarity for Cross-Lingual Code Search Models
par: Saieva, Anthony, et autres
Publié: (2023)
par: Saieva, Anthony, et autres
Publié: (2023)
$\textbf{PLUM}$: Improving Code LMs with Execution-Guided On-Policy Preference Learning Driven By Synthetic Test Cases
par: Zhang, Dylan, et autres
Publié: (2024)
par: Zhang, Dylan, et autres
Publié: (2024)
EquiBench: Benchmarking Large Language Models' Reasoning about Program Semantics via Equivalence Checking
par: Wei, Anjiang, et autres
Publié: (2025)
par: Wei, Anjiang, et autres
Publié: (2025)
Co-Learning: Code Learning for Multi-Agent Reinforcement Collaborative Framework with Conversational Natural Language Interfaces
par: Yu, Jiapeng, et autres
Publié: (2024)
par: Yu, Jiapeng, et autres
Publié: (2024)
CodeIF-Bench: Evaluating Instruction-Following Capabilities of Large Language Models in Interactive Code Generation
par: Wang, Peiding, et autres
Publié: (2025)
par: Wang, Peiding, et autres
Publié: (2025)
R1-Fuzz: Specializing Language Models for Textual Fuzzing via Reinforcement Learning
par: Lin, Jiayi, et autres
Publié: (2025)
par: Lin, Jiayi, et autres
Publié: (2025)
Can Large Language Models Transform Natural Language Intent into Formal Method Postconditions?
par: Endres, Madeline, et autres
Publié: (2023)
par: Endres, Madeline, et autres
Publié: (2023)
Turn: A Language for Agentic Computation
par: Kizito, Muyukani
Publié: (2026)
par: Kizito, Muyukani
Publié: (2026)
Documents similaires
-
Benchmarking Large Language Models for ABAP Code Generation: An Empirical Study on Iterative Improvement by Compiler Feedback
par: Wallraven, Stephan, et autres
Publié: (2026) -
LEGO-Compiler: Enhancing Neural Compilation Through Translation Composability
par: Zhang, Shuoming, et autres
Publié: (2025) -
The New Compiler Stack: A Survey on the Synergy of LLMs and Compilers
par: Zhang, Shuoming, et autres
Publié: (2026) -
PerfCodeGen: Improving Performance of LLM Generated Code with Execution Feedback
par: Peng, Yun, et autres
Publié: (2024) -
LangGPT: Rethinking Structured Reusable Prompt Design Framework for LLMs from the Programming Language
par: Wang, Ming, et autres
Publié: (2024)