Process-Supervised Reinforcement Learning for Code Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ye, Yufan, Zhang, Ting, Jiang, Wenbin, Huang, Hua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Debug2Fix: Can Interactive Debugging Help Coding Agents Fix More Bugs?
von: Garg, Spandan, et al.
Veröffentlicht: (2026)
von: Garg, Spandan, et al.
Veröffentlicht: (2026)
Klear-CodeTest: Scalable Test Case Generation for Code Reinforcement Learning
von: Fu, Jia, et al.
Veröffentlicht: (2025)
von: Fu, Jia, et al.
Veröffentlicht: (2025)
Exploring Pass-Rate Reward in Reinforcement Learning for Code Generation
von: Li, Xin-Ye, et al.
Veröffentlicht: (2026)
von: Li, Xin-Ye, et al.
Veröffentlicht: (2026)
Schedule-and-Calibrate: Utility-Guided Multi-Task Reinforcement Learning for Code LLMs
von: Chen, Yujia, et al.
Veröffentlicht: (2026)
von: Chen, Yujia, et al.
Veröffentlicht: (2026)
Reinforcement Learning-Guided Chain-of-Draft for Token-Efficient Code Generation
von: Tang, Xunzhu, et al.
Veröffentlicht: (2025)
von: Tang, Xunzhu, et al.
Veröffentlicht: (2025)
Improving LLM Code Generation via Requirement-Aware Curriculum Reinforcement Learning
von: Yin, Shouyu, et al.
Veröffentlicht: (2026)
von: Yin, Shouyu, et al.
Veröffentlicht: (2026)
CLAP: Learning Transferable Binary Code Representations with Natural Language Supervision
von: Wang, Hao, et al.
Veröffentlicht: (2024)
von: Wang, Hao, et al.
Veröffentlicht: (2024)
Autoregressive, Yet Revisable: In Decoding Revision for Secure Code Generation
von: Yang, Chengran, et al.
Veröffentlicht: (2026)
von: Yang, Chengran, et al.
Veröffentlicht: (2026)
CVeDRL: An Efficient Code Verifier via Difficulty-aware Reinforcement Learning
von: Shi, Ji, et al.
Veröffentlicht: (2026)
von: Shi, Ji, et al.
Veröffentlicht: (2026)
Afterburner: Reinforcement Learning Facilitates Self-Improving Code Efficiency Optimization
von: Du, Mingzhe, et al.
Veröffentlicht: (2025)
von: Du, Mingzhe, et al.
Veröffentlicht: (2025)
ReCode: Reinforcing Code Generation with Reasoning-Process Rewards
von: Fan, Lishui, et al.
Veröffentlicht: (2025)
von: Fan, Lishui, et al.
Veröffentlicht: (2025)
Coding with Eyes: Visual Feedback Unlocks Reliable GUI Code Generating and Debugging
von: Liu, Zhilin, et al.
Veröffentlicht: (2026)
von: Liu, Zhilin, et al.
Veröffentlicht: (2026)
Uncovering LLM-Generated Code: A Zero-Shot Synthetic Code Detector via Code Rewriting
von: Ye, Tong, et al.
Veröffentlicht: (2024)
von: Ye, Tong, et al.
Veröffentlicht: (2024)
ZS4C: Zero-Shot Synthesis of Compilable Code for Incomplete Code Snippets using LLMs
von: Kabir, Azmain, et al.
Veröffentlicht: (2024)
von: Kabir, Azmain, et al.
Veröffentlicht: (2024)
Nova: Generative Language Models for Assembly Code with Hierarchical Attention and Contrastive Learning
von: Jiang, Nan, et al.
Veröffentlicht: (2023)
von: Jiang, Nan, et al.
Veröffentlicht: (2023)
Saving SWE-Bench: A Benchmark Mutation Approach for Realistic Agent Evaluation
von: Garg, Spandan, et al.
Veröffentlicht: (2025)
von: Garg, Spandan, et al.
Veröffentlicht: (2025)
Terminus-4B: Can a Smaller Model Replace Frontier LLMs at Agentic Execution Tasks?
von: Garg, Spandan, et al.
Veröffentlicht: (2026)
von: Garg, Spandan, et al.
Veröffentlicht: (2026)
CodeRL+: Improving Code Generation via Reinforcement with Execution Semantics Alignment
von: Jiang, Xue, et al.
Veröffentlicht: (2025)
von: Jiang, Xue, et al.
Veröffentlicht: (2025)
Top Pass: Improve Code Generation by Pass@k-Maximized Code Ranking
von: Lyu, Zhi-Cun, et al.
Veröffentlicht: (2024)
von: Lyu, Zhi-Cun, et al.
Veröffentlicht: (2024)
CodeCoT: Tackling Code Syntax Errors in CoT Reasoning for Code Generation
von: Huang, Dong, et al.
Veröffentlicht: (2023)
von: Huang, Dong, et al.
Veröffentlicht: (2023)
Enhancing High-Quality Code Generation in Large Language Models with Comparative Prefix-Tuning
von: Jiang, Yuan, et al.
Veröffentlicht: (2025)
von: Jiang, Yuan, et al.
Veröffentlicht: (2025)
Synergizing Code Coverage and Gameplay Intent: Coverage-Aware Game Playtesting with LLM-Guided Reinforcement Learning
von: Mu, Enhong, et al.
Veröffentlicht: (2025)
von: Mu, Enhong, et al.
Veröffentlicht: (2025)
LLM-Powered Code Vulnerability Repair with Reinforcement Learning and Semantic Reward
von: Islam, Nafis Tanveer, et al.
Veröffentlicht: (2024)
von: Islam, Nafis Tanveer, et al.
Veröffentlicht: (2024)
DSTC: Direct Preference Learning with Only Self-Generated Tests and Code to Improve Code LMs
von: Liu, Zhihan, et al.
Veröffentlicht: (2024)
von: Liu, Zhihan, et al.
Veröffentlicht: (2024)
Flow2Code: Evaluating Large Language Models for Flowchart-based Code Generation Capability
von: He, Mengliang, et al.
Veröffentlicht: (2025)
von: He, Mengliang, et al.
Veröffentlicht: (2025)
From Patches to Trajectories: Privileged Process Supervision for Software-Engineering Agents
von: Ma, Murong, et al.
Veröffentlicht: (2026)
von: Ma, Murong, et al.
Veröffentlicht: (2026)
Unlock the Correlation between Supervised Fine-Tuning and Reinforcement Learning in Training Code Large Language Models
von: Chen, Jie, et al.
Veröffentlicht: (2024)
von: Chen, Jie, et al.
Veröffentlicht: (2024)
Bias Testing and Mitigation in LLM-based Code Generation
von: Huang, Dong, et al.
Veröffentlicht: (2023)
von: Huang, Dong, et al.
Veröffentlicht: (2023)
Learn to Code Sustainably: An Empirical Study on LLM-based Green Code Generation
von: Vartziotis, Tina, et al.
Veröffentlicht: (2024)
von: Vartziotis, Tina, et al.
Veröffentlicht: (2024)
The Readability Spectrum: Patterns, Issues, and Prompt Effects in LLM-Generated Code
von: Ye, Hengzhi, et al.
Veröffentlicht: (2026)
von: Ye, Hengzhi, et al.
Veröffentlicht: (2026)
Optimizing Code Runtime Performance through Context-Aware Retrieval-Augmented Generation
von: Acharya, Manish, et al.
Veröffentlicht: (2025)
von: Acharya, Manish, et al.
Veröffentlicht: (2025)
ModiGen: A Large Language Model-Based Workflow for Multi-Task Modelica Code Generation
von: Xiang, Jiahui, et al.
Veröffentlicht: (2025)
von: Xiang, Jiahui, et al.
Veröffentlicht: (2025)
Skeleton-Guided-Translation: A Benchmarking Framework for Code Repository Translation with Fine-Grained Quality Evaluation
von: Zhang, Xing, et al.
Veröffentlicht: (2025)
von: Zhang, Xing, et al.
Veröffentlicht: (2025)
Towards Fair Machine Learning Software: Understanding and Addressing Model Bias Through Counterfactual Thinking
von: Wang, Zichong, et al.
Veröffentlicht: (2023)
von: Wang, Zichong, et al.
Veröffentlicht: (2023)
TENET: Leveraging Tests Beyond Validation for Code Generation
von: Hu, Yiran, et al.
Veröffentlicht: (2025)
von: Hu, Yiran, et al.
Veröffentlicht: (2025)
Learning to Generate Unit Test via Adversarial Reinforcement Learning
von: Lee, Dongjun, et al.
Veröffentlicht: (2025)
von: Lee, Dongjun, et al.
Veröffentlicht: (2025)
GNN-Coder: Boosting Semantic Code Retrieval with Combined GNNs and Transformer
von: Ye, Yufan, et al.
Veröffentlicht: (2025)
von: Ye, Yufan, et al.
Veröffentlicht: (2025)
Effective Code Membership Inference for Code Completion Models via Adversarial Prompts
von: Jiang, Yuan, et al.
Veröffentlicht: (2025)
von: Jiang, Yuan, et al.
Veröffentlicht: (2025)
CodeFort: Robust Training for Code Generation Models
von: Zhang, Yuhao, et al.
Veröffentlicht: (2024)
von: Zhang, Yuhao, et al.
Veröffentlicht: (2024)
TransformCode: A Contrastive Learning Framework for Code Embedding via Subtree Transformation
von: Xian, Zixiang, et al.
Veröffentlicht: (2023)
von: Xian, Zixiang, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Debug2Fix: Can Interactive Debugging Help Coding Agents Fix More Bugs?
von: Garg, Spandan, et al.
Veröffentlicht: (2026) -
Klear-CodeTest: Scalable Test Case Generation for Code Reinforcement Learning
von: Fu, Jia, et al.
Veröffentlicht: (2025) -
Exploring Pass-Rate Reward in Reinforcement Learning for Code Generation
von: Li, Xin-Ye, et al.
Veröffentlicht: (2026) -
Schedule-and-Calibrate: Utility-Guided Multi-Task Reinforcement Learning for Code LLMs
von: Chen, Yujia, et al.
Veröffentlicht: (2026) -
Reinforcement Learning-Guided Chain-of-Draft for Token-Efficient Code Generation
von: Tang, Xunzhu, et al.
Veröffentlicht: (2025)