Schedule-and-Calibrate: Utility-Guided Multi-Task Reinforcement Learning for Code LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Yujia, Ye, Yang, Chu, Xiao, Ma, Yuchi, Gao, Cuiyun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Benchmarking Multimodal LLMs on Code Generation for Complex Interactive Webpages
di: Wu, Fan, et al.
Pubblicazione: (2026)
di: Wu, Fan, et al.
Pubblicazione: (2026)
An Empirical Study of Knowledge Distillation for Code Understanding Tasks
di: Wang, Ruiqi, et al.
Pubblicazione: (2025)
di: Wang, Ruiqi, et al.
Pubblicazione: (2025)
Boosting Vulnerability Detection of LLMs via Curriculum Preference Optimization with Synthetic Reasoning Data
di: Wen, Xin-Cheng, et al.
Pubblicazione: (2025)
di: Wen, Xin-Cheng, et al.
Pubblicazione: (2025)
Towards Mitigating API Hallucination in Code Generated by LLMs with Hierarchical Dependency Aware
di: Chen, Yujia, et al.
Pubblicazione: (2025)
di: Chen, Yujia, et al.
Pubblicazione: (2025)
LLM-Based Test Case Generation in DBMS through Monte Carlo Tree Search
di: Chen, Yujia, et al.
Pubblicazione: (2026)
di: Chen, Yujia, et al.
Pubblicazione: (2026)
Search-Based LLMs for Code Optimization
di: Gao, Shuzheng, et al.
Pubblicazione: (2024)
di: Gao, Shuzheng, et al.
Pubblicazione: (2024)
Smaller but Better: Self-Paced Knowledge Distillation for Lightweight yet Effective LCMs
di: Chen, Yujia, et al.
Pubblicazione: (2024)
di: Chen, Yujia, et al.
Pubblicazione: (2024)
AXIOM: Benchmarking LLM-as-a-Judge for Code via Rule-Based Perturbation and Multisource Quality Calibration
di: Wang, Ruiqi, et al.
Pubblicazione: (2025)
di: Wang, Ruiqi, et al.
Pubblicazione: (2025)
Benchmarking LLMs for Fine-Grained Code Review with Enriched Context in Practice
di: Hu, Ruida, et al.
Pubblicazione: (2025)
di: Hu, Ruida, et al.
Pubblicazione: (2025)
EvalSVA: Multi-Agent Evaluators for Next-Gen Software Vulnerability Assessment
di: Wen, Xin-Cheng, et al.
Pubblicazione: (2024)
di: Wen, Xin-Cheng, et al.
Pubblicazione: (2024)
Fine-grained Approaches for Confidence Calibration of LLMs in Automated Code Revision
di: Lin, Hong Yi, et al.
Pubblicazione: (2026)
di: Lin, Hong Yi, et al.
Pubblicazione: (2026)
Vul-R2: A Reasoning LLM for Automated Vulnerability Repair
di: Wen, Xin-Cheng, et al.
Pubblicazione: (2025)
di: Wen, Xin-Cheng, et al.
Pubblicazione: (2025)
CodeR: Issue Resolving with Multi-Agent and Task Graphs
di: Chen, Dong, et al.
Pubblicazione: (2024)
di: Chen, Dong, et al.
Pubblicazione: (2024)
SPENCER: Self-Adaptive Model Distillation for Efficient Code Retrieval
di: Gu, Wenchao, et al.
Pubblicazione: (2025)
di: Gu, Wenchao, et al.
Pubblicazione: (2025)
Process-Supervised Reinforcement Learning for Code Generation
di: Ye, Yufan, et al.
Pubblicazione: (2025)
di: Ye, Yufan, et al.
Pubblicazione: (2025)
Empowering AI to Generate Better AI Code: Guided Generation of Deep Learning Projects with LLMs
di: Xie, Chen, et al.
Pubblicazione: (2025)
di: Xie, Chen, et al.
Pubblicazione: (2025)
Reinforcement Learning-Guided Chain-of-Draft for Token-Efficient Code Generation
di: Tang, Xunzhu, et al.
Pubblicazione: (2025)
di: Tang, Xunzhu, et al.
Pubblicazione: (2025)
Beyond Functional Correctness: Exploring Hallucinations in LLM-Generated Code
di: Liu, Fang, et al.
Pubblicazione: (2024)
di: Liu, Fang, et al.
Pubblicazione: (2024)
Can LLMs Replace Human Evaluators? An Empirical Study of LLM-as-a-Judge in Software Engineering
di: Wang, Ruiqi, et al.
Pubblicazione: (2025)
di: Wang, Ruiqi, et al.
Pubblicazione: (2025)
An Insight into Security Code Review with LLMs: Capabilities, Obstacles, and Influential Factors
di: Yu, Jiaxin, et al.
Pubblicazione: (2024)
di: Yu, Jiaxin, et al.
Pubblicazione: (2024)
Empirical Study of Code Large Language Models for Binary Security Patch Detection
di: Li, Qingyuan, et al.
Pubblicazione: (2025)
di: Li, Qingyuan, et al.
Pubblicazione: (2025)
Weakly Supervised Vulnerability Localization via Multiple Instance Learning
di: Gu, Wenchao, et al.
Pubblicazione: (2025)
di: Gu, Wenchao, et al.
Pubblicazione: (2025)
CodeRepoQA: A Large-scale Benchmark for Software Engineering Question Answering
di: Hu, Ruida, et al.
Pubblicazione: (2024)
di: Hu, Ruida, et al.
Pubblicazione: (2024)
LibRec: Benchmarking Retrieval-Augmented LLMs for Library Migration Recommendations
di: Han, Junxiao, et al.
Pubblicazione: (2025)
di: Han, Junxiao, et al.
Pubblicazione: (2025)
Personality-Guided Code Generation Using Large Language Models
di: Guo, Yaoqi, et al.
Pubblicazione: (2024)
di: Guo, Yaoqi, et al.
Pubblicazione: (2024)
SWE-Fuse: Empowering Software Agents via Issue-free Trajectory Learning and Entropy-aware RLVR Training
di: Wen, Xin-Cheng, et al.
Pubblicazione: (2026)
di: Wen, Xin-Cheng, et al.
Pubblicazione: (2026)
AlignCoder: Aligning Retrieval with Target Intent for Repository-Level Code Completion
di: Jiang, Tianyue, et al.
Pubblicazione: (2026)
di: Jiang, Tianyue, et al.
Pubblicazione: (2026)
Beyond Functional Correctness: Investigating Coding Style Inconsistencies in Large Language Models
di: Wang, Yanlin, et al.
Pubblicazione: (2024)
di: Wang, Yanlin, et al.
Pubblicazione: (2024)
Neuron-Guided Interpretation of Code LLMs: Where, Why, and How?
di: Yin, Zhe, et al.
Pubblicazione: (2025)
di: Yin, Zhe, et al.
Pubblicazione: (2025)
Data Dependency-Aware Code Generation from Enhanced UML Sequence Diagrams
di: Mao, Wenxin, et al.
Pubblicazione: (2025)
di: Mao, Wenxin, et al.
Pubblicazione: (2025)
RepoMasterEval: Evaluating Code Completion via Real-World Repositories
di: Wu, Qinyun, et al.
Pubblicazione: (2024)
di: Wu, Qinyun, et al.
Pubblicazione: (2024)
Evaluating Repository-level Software Documentation via Question Answering and Feature-Driven Development
di: Wang, Xinchen, et al.
Pubblicazione: (2026)
di: Wang, Xinchen, et al.
Pubblicazione: (2026)
ModiGen: A Large Language Model-Based Workflow for Multi-Task Modelica Code Generation
di: Xiang, Jiahui, et al.
Pubblicazione: (2025)
di: Xiang, Jiahui, et al.
Pubblicazione: (2025)
Top General Performance = Top Domain Performance? DomainCodeBench: A Multi-domain Code Generation Benchmark
di: Zheng, Dewu, et al.
Pubblicazione: (2024)
di: Zheng, Dewu, et al.
Pubblicazione: (2024)
Klear-CodeTest: Scalable Test Case Generation for Code Reinforcement Learning
di: Fu, Jia, et al.
Pubblicazione: (2025)
di: Fu, Jia, et al.
Pubblicazione: (2025)
Synergizing Code Coverage and Gameplay Intent: Coverage-Aware Game Playtesting with LLM-Guided Reinforcement Learning
di: Mu, Enhong, et al.
Pubblicazione: (2025)
di: Mu, Enhong, et al.
Pubblicazione: (2025)
Task Abstention for Large Language Models in Code Generation
di: Zhou, Yanke, et al.
Pubblicazione: (2026)
di: Zhou, Yanke, et al.
Pubblicazione: (2026)
CodeVisionary: An Agent-based Framework for Evaluating Large Language Models in Code Generation
di: Wang, Xinchen, et al.
Pubblicazione: (2025)
di: Wang, Xinchen, et al.
Pubblicazione: (2025)
RedCoder: Automated Multi-Turn Red Teaming for Code LLMs
di: Mo, Wenjie Jacky, et al.
Pubblicazione: (2025)
di: Mo, Wenjie Jacky, et al.
Pubblicazione: (2025)
CoRe: Benchmarking LLMs Code Reasoning Capabilities through Static Analysis Tasks
di: Xie, Danning, et al.
Pubblicazione: (2025)
di: Xie, Danning, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Benchmarking Multimodal LLMs on Code Generation for Complex Interactive Webpages
di: Wu, Fan, et al.
Pubblicazione: (2026) -
An Empirical Study of Knowledge Distillation for Code Understanding Tasks
di: Wang, Ruiqi, et al.
Pubblicazione: (2025) -
Boosting Vulnerability Detection of LLMs via Curriculum Preference Optimization with Synthetic Reasoning Data
di: Wen, Xin-Cheng, et al.
Pubblicazione: (2025) -
Towards Mitigating API Hallucination in Code Generated by LLMs with Hierarchical Dependency Aware
di: Chen, Yujia, et al.
Pubblicazione: (2025) -
LLM-Based Test Case Generation in DBMS through Monte Carlo Tree Search
di: Chen, Yujia, et al.
Pubblicazione: (2026)