Enhancing Code LLMs with Reinforcement Learning in Code Generation: A Survey
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Junqiao, Zhang, Zeng, He, Yangfan, Zhang, Zihao, Song, Xinyuan, Song, Yuyang, Shi, Tianyu, Li, Yuchen, Xu, Hengyuan, Wu, Kunyu, Yi, Xin, Wan, Zhongwei, Yuan, Xinhang, Wang, Zijun, Lu, Kuan, Huo, Menghao, Jingqun, Tang, Qian, Guangwu, Li, Keqin, Chen, Qiuwu, He, Lewei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FALCON: Feedback-driven Adaptive Long/short-term memory reinforced Coding Optimization system
von: Li, Zeyuan, et al.
Veröffentlicht: (2024)
von: Li, Zeyuan, et al.
Veröffentlicht: (2024)
Enhancing Low-Cost Video Editing with Lightweight Adaptors and Temporal-Aware Inversion
von: He, Yangfan, et al.
Veröffentlicht: (2025)
von: He, Yangfan, et al.
Veröffentlicht: (2025)
Efficient Temporal Consistency in Diffusion-Based Video Editing with Adaptor Modules: A Theoretical Framework
von: Song, Xinyuan, et al.
Veröffentlicht: (2025)
von: Song, Xinyuan, et al.
Veröffentlicht: (2025)
Enhancing Intent Understanding for Ambiguous prompt: A Human-Machine Co-Adaption Strategy
von: He, Yangfan, et al.
Veröffentlicht: (2025)
von: He, Yangfan, et al.
Veröffentlicht: (2025)
SCORE: Story Coherence and Retrieval Enhancement for AI Narratives
von: Yi, Qiang, et al.
Veröffentlicht: (2025)
von: Yi, Qiang, et al.
Veröffentlicht: (2025)
Twin Co-Adaptive Dialogue for Progressive Image Generation
von: Wang, Jianhui, et al.
Veröffentlicht: (2025)
von: Wang, Jianhui, et al.
Veröffentlicht: (2025)
Optimizing Multi-Round Enhanced Training in Diffusion Models for Improved Preference Understanding
von: Li, Kun, et al.
Veröffentlicht: (2025)
von: Li, Kun, et al.
Veröffentlicht: (2025)
AutoCoder: Enhancing Code Large Language Model with \textsc{AIEV-Instruct}
von: Lei, Bin, et al.
Veröffentlicht: (2024)
von: Lei, Bin, et al.
Veröffentlicht: (2024)
Free-Mask: A Novel Paradigm of Integration Between the Segmentation Diffusion Model and Image Editing
von: Gao, Bo, et al.
Veröffentlicht: (2024)
von: Gao, Bo, et al.
Veröffentlicht: (2024)
Self-evolving Agents with reflective and memory-augmented abilities
von: Liang, Xuechen, et al.
Veröffentlicht: (2024)
von: Liang, Xuechen, et al.
Veröffentlicht: (2024)
Low-Cost Test-Time Adaptation for Robust Video Editing
von: Wang, Jianhui, et al.
Veröffentlicht: (2025)
von: Wang, Jianhui, et al.
Veröffentlicht: (2025)
ShieldedCode: Learning Robust Representations for Virtual Machine Protected Code
von: Mo, Mingqiao, et al.
Veröffentlicht: (2026)
von: Mo, Mingqiao, et al.
Veröffentlicht: (2026)
Enhancing Commentary Strategies for Imperfect Information Card Games: A Study of Large Language Models in Guandan Commentary
von: Tao, Meiling, et al.
Veröffentlicht: (2024)
von: Tao, Meiling, et al.
Veröffentlicht: (2024)
TRiCo: Triadic Game-Theoretic Co-Training for Robust Semi-Supervised Learning
von: He, Hongyang, et al.
Veröffentlicht: (2025)
von: He, Hongyang, et al.
Veröffentlicht: (2025)
DDPM-MoCo: Advancing Industrial Surface Defect Generation and Detection with Generative and Contrastive Learning
von: He, Yangfan, et al.
Veröffentlicht: (2024)
von: He, Yangfan, et al.
Veröffentlicht: (2024)
Image-to-Image Translation with Diffusion Transformers and CLIP-Based Image Conditioning
von: Zhu, Qiang, et al.
Veröffentlicht: (2025)
von: Zhu, Qiang, et al.
Veröffentlicht: (2025)
MountainLion: A Multi-Modal LLM-Based Agent System for Interpretable and Adaptive Financial Trading
von: Wu, Siyi, et al.
Veröffentlicht: (2025)
von: Wu, Siyi, et al.
Veröffentlicht: (2025)
PokéAI: A Goal-Generating, Battle-Optimizing Multi-agent System for Pokemon Red
von: Liu, Zihao, et al.
Veröffentlicht: (2025)
von: Liu, Zihao, et al.
Veröffentlicht: (2025)
ReGraP-LLaVA: Reasoning enabled Graph-based Personalized Large Language and Vision Assistant
von: Xiang, Yifan, et al.
Veröffentlicht: (2025)
von: Xiang, Yifan, et al.
Veröffentlicht: (2025)
CT-PatchTST: Channel-Time Patch Time-Series Transformer for Long-Term Renewable Energy Forecasting
von: Lu, Kuan, et al.
Veröffentlicht: (2025)
von: Lu, Kuan, et al.
Veröffentlicht: (2025)
LPC-SM: Local Predictive Coding and Sparse Memory for Long-Context Language Modeling
von: Xie, Keqin
Veröffentlicht: (2026)
von: Xie, Keqin
Veröffentlicht: (2026)
Enhancing Customer Contact Efficiency with Graph Neural Networks in Credit Card Fraud Detection Workflow
von: Huo, Menghao, et al.
Veröffentlicht: (2025)
von: Huo, Menghao, et al.
Veröffentlicht: (2025)
Difficulty-Aware Agentic Orchestration for Query-Specific Multi-Agent Workflows
von: Su, Jinwei, et al.
Veröffentlicht: (2025)
von: Su, Jinwei, et al.
Veröffentlicht: (2025)
DTP: A Simple yet Effective Distracting Token Pruning Framework for Vision-Language Action Models
von: Li, Chenyang, et al.
Veröffentlicht: (2026)
von: Li, Chenyang, et al.
Veröffentlicht: (2026)
Symplectic Wigner Distribution in the Linear Canonical Transform Domain: Theory and Application
von: He, Yangfan, et al.
Veröffentlicht: (2025)
von: He, Yangfan, et al.
Veröffentlicht: (2025)
eSapiens: A Real-World NLP Framework for Multimodal Document Understanding and Enterprise Knowledge Processing
von: Shi, Isaac, et al.
Veröffentlicht: (2025)
von: Shi, Isaac, et al.
Veröffentlicht: (2025)
Texture-guided Coding for Deep Features
von: Xiong, Lei, et al.
Veröffentlicht: (2024)
von: Xiong, Lei, et al.
Veröffentlicht: (2024)
Meta-Learning Reinforcement Learning for Crypto-Return Prediction
von: Wang, Junqiao, et al.
Veröffentlicht: (2025)
von: Wang, Junqiao, et al.
Veröffentlicht: (2025)
LLM-Grounded Explainable AI for Supply Chain Risk Early Warning via Temporal Graph Attention Networks
von: Xue, Zhiming, et al.
Veröffentlicht: (2026)
von: Xue, Zhiming, et al.
Veröffentlicht: (2026)
EAGLE: Edge-Aware Graph Learning for Proactive Delivery Delay Prediction in Smart Logistics Networks
von: Xue, Zhiming, et al.
Veröffentlicht: (2026)
von: Xue, Zhiming, et al.
Veröffentlicht: (2026)
Physics-Informed Deep Recurrent Back-Projection Network for Tunnel Propagation Modeling
von: Wu, Kunyu, et al.
Veröffentlicht: (2026)
von: Wu, Kunyu, et al.
Veröffentlicht: (2026)
Let the Code LLM Edit Itself When You Edit the Code
von: He, Zhenyu, et al.
Veröffentlicht: (2024)
von: He, Zhenyu, et al.
Veröffentlicht: (2024)
Guided Debugging of Auto-Translated Code Using Differential Testing
von: Wu, Shengnan, et al.
Veröffentlicht: (2025)
von: Wu, Shengnan, et al.
Veröffentlicht: (2025)
Optimal Repair of $(k+2, k, 2)$ MDS Array Codes
von: Zhang, Zihao, et al.
Veröffentlicht: (2025)
von: Zhang, Zihao, et al.
Veröffentlicht: (2025)
VibeContract: The Missing Quality Assurance Piece in Vibe Coding
von: Wang, Song
Veröffentlicht: (2026)
von: Wang, Song
Veröffentlicht: (2026)
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code
von: Bao, Keqin, et al.
Veröffentlicht: (2025)
von: Bao, Keqin, et al.
Veröffentlicht: (2025)
Scalable and Precise Synthesis of Structurally Colored Bottlebrush Block Copolymers: Enabling Refined Color Calibration for Sustainable Photonic Pigments
von: Bangbang Wang, et al.
Veröffentlicht: (2025)
von: Bangbang Wang, et al.
Veröffentlicht: (2025)
Type-Constrained Code Generation with Language Models
von: Mündler, Niels, et al.
Veröffentlicht: (2025)
von: Mündler, Niels, et al.
Veröffentlicht: (2025)
CodeCriticBench: A Holistic Code Critique Benchmark for Large Language Models
von: Zhang, Alexander, et al.
Veröffentlicht: (2025)
von: Zhang, Alexander, et al.
Veröffentlicht: (2025)
Superlubric sliding ferroelectricity
von: Yang, Zihao, et al.
Veröffentlicht: (2025)
von: Yang, Zihao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
FALCON: Feedback-driven Adaptive Long/short-term memory reinforced Coding Optimization system
von: Li, Zeyuan, et al.
Veröffentlicht: (2024) -
Enhancing Low-Cost Video Editing with Lightweight Adaptors and Temporal-Aware Inversion
von: He, Yangfan, et al.
Veröffentlicht: (2025) -
Efficient Temporal Consistency in Diffusion-Based Video Editing with Adaptor Modules: A Theoretical Framework
von: Song, Xinyuan, et al.
Veröffentlicht: (2025) -
Enhancing Intent Understanding for Ambiguous prompt: A Human-Machine Co-Adaption Strategy
von: He, Yangfan, et al.
Veröffentlicht: (2025) -
SCORE: Story Coherence and Retrieval Enhancement for AI Narratives
von: Yi, Qiang, et al.
Veröffentlicht: (2025)