Domain-Adaptable Reinforcement Learning for Code Generation with Dense Rewards
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jolfaei, Erfan Aghadavoodi, Maninger, Daniel, Anand, Abhinav, Tiftikci, Mert, Mezini, Mira |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Trustworthy AI Software Development Assistance
von: Maninger, Daniel, et al.
Veröffentlicht: (2023)
von: Maninger, Daniel, et al.
Veröffentlicht: (2023)
Deep Graph-Language Fusion for Structure-Aware Code Generation
von: Tiftikci, Mert, et al.
Veröffentlicht: (2026)
von: Tiftikci, Mert, et al.
Veröffentlicht: (2026)
Where Do LLMs Still Struggle? An In-Depth Analysis of Code Generation Benchmarks
von: Sharifloo, Amir Molzam, et al.
Veröffentlicht: (2025)
von: Sharifloo, Amir Molzam, et al.
Veröffentlicht: (2025)
Evaluating and Mitigating Errors in LLM-Generated Web API Integrations
von: Maninger, Daniel, et al.
Veröffentlicht: (2025)
von: Maninger, Daniel, et al.
Veröffentlicht: (2025)
Integrating Symbolic Execution into the Fine-Tuning of Code-Generating LLMs
von: Sakharova, Marina, et al.
Veröffentlicht: (2025)
von: Sakharova, Marina, et al.
Veröffentlicht: (2025)
CodeSSM: Towards State Space Models for Code Understanding
von: Verma, Shweta, et al.
Veröffentlicht: (2025)
von: Verma, Shweta, et al.
Veröffentlicht: (2025)
A Critical Study of What Code-LLMs (Do Not) Learn
von: Anand, Abhinav, et al.
Veröffentlicht: (2024)
von: Anand, Abhinav, et al.
Veröffentlicht: (2024)
Exploring Pass-Rate Reward in Reinforcement Learning for Code Generation
von: Li, Xin-Ye, et al.
Veröffentlicht: (2026)
von: Li, Xin-Ye, et al.
Veröffentlicht: (2026)
Benchmarking Reward Hack Detection in Code Environments via Contrastive Analysis
von: Deshpande, Darshan, et al.
Veröffentlicht: (2026)
von: Deshpande, Darshan, et al.
Veröffentlicht: (2026)
FunPRM: Function-as-Step Process Reward Model with Meta Reward Correction for Code Generation
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2026)
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2026)
Leveraging Reward Models for Guiding Code Review Comment Generation
von: Sghaier, Oussama Ben, et al.
Veröffentlicht: (2025)
von: Sghaier, Oussama Ben, et al.
Veröffentlicht: (2025)
ReCode: Reinforcing Code Generation with Reasoning-Process Rewards
von: Fan, Lishui, et al.
Veröffentlicht: (2025)
von: Fan, Lishui, et al.
Veröffentlicht: (2025)
SemRep: Generative Code Representation Learning with Code Transformations
von: Li, Weichen, et al.
Veröffentlicht: (2026)
von: Li, Weichen, et al.
Veröffentlicht: (2026)
SIMCOPILOT: Evaluating Large Language Models for Copilot-Style Code Generation
von: Jiang, Mingchao, et al.
Veröffentlicht: (2025)
von: Jiang, Mingchao, et al.
Veröffentlicht: (2025)
A Survey on LLM-based Code Generation for Low-Resource and Domain-Specific Programming Languages
von: Joel, Sathvik, et al.
Veröffentlicht: (2024)
von: Joel, Sathvik, et al.
Veröffentlicht: (2024)
Themis: Training Robust Multilingual Code Reward Models for Flexible Multi-Criteria Scoring
von: Paul, Indraneil, et al.
Veröffentlicht: (2026)
von: Paul, Indraneil, et al.
Veröffentlicht: (2026)
QiMeng-PRepair: Precise Code Repair via Edit-Aware Reward Optimization
von: Ke, Changxin, et al.
Veröffentlicht: (2026)
von: Ke, Changxin, et al.
Veröffentlicht: (2026)
AgentForge: A Flexible Low-Code Platform for Reinforcement Learning Agent Design
von: Junior, Francisco Erivaldo Fernandes, et al.
Veröffentlicht: (2024)
von: Junior, Francisco Erivaldo Fernandes, et al.
Veröffentlicht: (2024)
QEDCartographer: Automating Formal Verification Using Reward-Free Reinforcement Learning
von: Sanchez-Stern, Alex, et al.
Veröffentlicht: (2024)
von: Sanchez-Stern, Alex, et al.
Veröffentlicht: (2024)
LogRouter: Adaptive Two-Level LLM Routing for Log Question Answering in Big Data Systems
von: Coskuner, Mert, et al.
Veröffentlicht: (2026)
von: Coskuner, Mert, et al.
Veröffentlicht: (2026)
Leveraging LLMs for Legacy Code Modernization: Challenges and Opportunities for LLM-Generated Documentation
von: Diggs, Colin, et al.
Veröffentlicht: (2024)
von: Diggs, Colin, et al.
Veröffentlicht: (2024)
Think Anywhere in Code Generation
von: Jiang, Xue, et al.
Veröffentlicht: (2026)
von: Jiang, Xue, et al.
Veröffentlicht: (2026)
Code Roulette: How Prompt Variability Affects LLM Code Generation
von: Paleyes, Andrei, et al.
Veröffentlicht: (2025)
von: Paleyes, Andrei, et al.
Veröffentlicht: (2025)
Reinforcement Learning from Automatic Feedback for High-Quality Unit Test Generation
von: Steenhoek, Benjamin, et al.
Veröffentlicht: (2023)
von: Steenhoek, Benjamin, et al.
Veröffentlicht: (2023)
Reinforcement Learning from Automatic Feedback for High-Quality Unit Test Generation
von: Steenhoek, Benjamin, et al.
Veröffentlicht: (2024)
von: Steenhoek, Benjamin, et al.
Veröffentlicht: (2024)
On the Usage of Continual Learning for Out-of-Distribution Generalization in Pre-trained Language Models of Code
von: Weyssow, Martin, et al.
Veröffentlicht: (2023)
von: Weyssow, Martin, et al.
Veröffentlicht: (2023)
DeputyDev -- AI Powered Developer Assistant: Breaking the Code Review Logjam through Contextual AI to Boost Developer Productivity
von: Khare, Vishal, et al.
Veröffentlicht: (2025)
von: Khare, Vishal, et al.
Veröffentlicht: (2025)
Improving the Reproducibility of Deep Learning Software: An Initial Investigation through a Case Study Analysis
von: Ravi, Nikita, et al.
Veröffentlicht: (2025)
von: Ravi, Nikita, et al.
Veröffentlicht: (2025)
CodeIF: Benchmarking the Instruction-Following Capabilities of Large Language Models for Code Generation
von: Yan, Kaiwen, et al.
Veröffentlicht: (2025)
von: Yan, Kaiwen, et al.
Veröffentlicht: (2025)
CodeSAM: Source Code Representation Learning by Infusing Self-Attention with Multi-Code-View Graphs
von: Mathai, Alex, et al.
Veröffentlicht: (2024)
von: Mathai, Alex, et al.
Veröffentlicht: (2024)
Enabling Global, Human-Centered Explanations for LLMs:From Tokens to Interpretable Code and Test Generation
von: Khati, Dipin, et al.
Veröffentlicht: (2025)
von: Khati, Dipin, et al.
Veröffentlicht: (2025)
LLM Performance for Code Generation on Noisy Tasks
von: Sendyka, Radzim, et al.
Veröffentlicht: (2025)
von: Sendyka, Radzim, et al.
Veröffentlicht: (2025)
OSS-Bench: Benchmark Generator for Coding LLMs
von: Jiang, Yuancheng, et al.
Veröffentlicht: (2025)
von: Jiang, Yuancheng, et al.
Veröffentlicht: (2025)
Robust Learning of Diverse Code Edits
von: Aggarwal, Tushar, et al.
Veröffentlicht: (2025)
von: Aggarwal, Tushar, et al.
Veröffentlicht: (2025)
Code Less, Align More: Efficient LLM Fine-tuning for Code Generation with Data Pruning
von: Tsai, Yun-Da, et al.
Veröffentlicht: (2024)
von: Tsai, Yun-Da, et al.
Veröffentlicht: (2024)
StructCoder: Structure-Aware Transformer for Code Generation
von: Tipirneni, Sindhu, et al.
Veröffentlicht: (2022)
von: Tipirneni, Sindhu, et al.
Veröffentlicht: (2022)
Leveraging Reviewer Experience in Code Review Comment Generation
von: Lin, Hong Yi, et al.
Veröffentlicht: (2024)
von: Lin, Hong Yi, et al.
Veröffentlicht: (2024)
Leveraging Generative AI for Enhancing Domain-Driven Software Design
von: Wiegand, Götz-Henrik, et al.
Veröffentlicht: (2026)
von: Wiegand, Götz-Henrik, et al.
Veröffentlicht: (2026)
ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinforcement Learning
von: Jiang, Juyong, et al.
Veröffentlicht: (2026)
von: Jiang, Juyong, et al.
Veröffentlicht: (2026)
Design-Specification Tiling for ICL-based CAD Code Generation
von: Du, Yali, et al.
Veröffentlicht: (2026)
von: Du, Yali, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Towards Trustworthy AI Software Development Assistance
von: Maninger, Daniel, et al.
Veröffentlicht: (2023) -
Deep Graph-Language Fusion for Structure-Aware Code Generation
von: Tiftikci, Mert, et al.
Veröffentlicht: (2026) -
Where Do LLMs Still Struggle? An In-Depth Analysis of Code Generation Benchmarks
von: Sharifloo, Amir Molzam, et al.
Veröffentlicht: (2025) -
Evaluating and Mitigating Errors in LLM-Generated Web API Integrations
von: Maninger, Daniel, et al.
Veröffentlicht: (2025) -
Integrating Symbolic Execution into the Fine-Tuning of Code-Generating LLMs
von: Sakharova, Marina, et al.
Veröffentlicht: (2025)