CRScore++: Reinforcement Learning with Verifiable Tool and AI Feedback for Code Review
Fuente:
arXiv
Saved in:
| Main Authors: | Kapadnis, Manav Nitin, Naik, Atharva, Rose, Carolyn |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CRScore: Grounding Automated Evaluation of Code Review Comments in Code Claims and Smells
by: Naik, Atharva, et al.
Published: (2024)
by: Naik, Atharva, et al.
Published: (2024)
An Empirical Study on Strong-Weak Model Collaboration for Repo-level Code Generation
by: Gandhi, Shubham, et al.
Published: (2025)
by: Gandhi, Shubham, et al.
Published: (2025)
Data Augmentation for Code Translation with Comparable Corpora and Multiple References
by: Xie, Yiqing, et al.
Published: (2023)
by: Xie, Yiqing, et al.
Published: (2023)
On the Limitations of Embedding Based Methods for Measuring Functional Correctness for Code Generation
by: Naik, Atharva
Published: (2024)
by: Naik, Atharva
Published: (2024)
ChartEditBench: Evaluating Grounded Multi-Turn Chart Editing in Multimodal Language Models
by: Kapadnis, Manav Nitin, et al.
Published: (2026)
by: Kapadnis, Manav Nitin, et al.
Published: (2026)
MetaLint: Easy-to-Hard Generalization for Code Linting
by: Naik, Atharva, et al.
Published: (2025)
by: Naik, Atharva, et al.
Published: (2025)
Fine-Tuning Models for Automated Code Review Feedback
by: Kumar, Smitha S, et al.
Published: (2026)
by: Kumar, Smitha S, et al.
Published: (2026)
CVeDRL: An Efficient Code Verifier via Difficulty-aware Reinforcement Learning
by: Shi, Ji, et al.
Published: (2026)
by: Shi, Ji, et al.
Published: (2026)
Configuring Agentic AI Coding Tools: An Exploratory Study
by: Galster, Matthias, et al.
Published: (2026)
by: Galster, Matthias, et al.
Published: (2026)
A Dataset of Agentic AI Coding Tool Configurations
by: Galster, Matthias, et al.
Published: (2026)
by: Galster, Matthias, et al.
Published: (2026)
Hold On! Is My Feedback Useful? Evaluating the Usefulness of Code Review Comments
by: Ahmed, Sharif, et al.
Published: (2025)
by: Ahmed, Sharif, et al.
Published: (2025)
Review of Tools for Zero-Code LLM Based Application Development
by: Pattnayak, Priyaranjan, et al.
Published: (2025)
by: Pattnayak, Priyaranjan, et al.
Published: (2025)
An Empirical Study of Static Analysis Tools for Secure Code Review
by: Charoenwet, Wachiraphan, et al.
Published: (2024)
by: Charoenwet, Wachiraphan, et al.
Published: (2024)
Towards Verifiably Safe Tool Use for LLM Agents
by: Doshi, Aarya, et al.
Published: (2026)
by: Doshi, Aarya, et al.
Published: (2026)
CodeDPO: Aligning Code Models with Self Generated and Verified Source Code
by: Zhang, Kechi, et al.
Published: (2024)
by: Zhang, Kechi, et al.
Published: (2024)
ExecVerify: White-Box RL with Verifiable Stepwise Rewards for Code Execution Reasoning
by: Tang, Lingxiao, et al.
Published: (2026)
by: Tang, Lingxiao, et al.
Published: (2026)
Learning to Solve and Verify: A Self-Play Framework for Code and Test Generation
by: Lin, Zi, et al.
Published: (2025)
by: Lin, Zi, et al.
Published: (2025)
CodeBenchGen: Creating Scalable Execution-based Code Generation Benchmarks
by: Xie, Yiqing, et al.
Published: (2024)
by: Xie, Yiqing, et al.
Published: (2024)
StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
by: Dou, Shihan, et al.
Published: (2024)
by: Dou, Shihan, et al.
Published: (2024)
SWE-PRBench: Benchmarking AI Code Review Quality Against Pull Request Feedback
by: Kumar, Deepak
Published: (2026)
by: Kumar, Deepak
Published: (2026)
DeVAIC: A Tool for Security Assessment of AI-generated Code
by: Cotroneo, Domenico, et al.
Published: (2024)
by: Cotroneo, Domenico, et al.
Published: (2024)
Exploring Requirements Elicitation from App Store User Reviews Using Large Language Models
by: Ghosh, Tanmai Kumar, et al.
Published: (2024)
by: Ghosh, Tanmai Kumar, et al.
Published: (2024)
Human-AI Synergy in Agentic Code Review
by: Zhong, Suzhen, et al.
Published: (2026)
by: Zhong, Suzhen, et al.
Published: (2026)
Engineering Pitfalls in AI Coding Tools: An Empirical Study of Bugs in Claude Code, Codex, and Gemini CLI
by: Zhang, Ruixin, et al.
Published: (2026)
by: Zhang, Ruixin, et al.
Published: (2026)
Improving Small Language Models for Code Generation with Reinforcement Learning from Verification Feedback
by: Skopin, Egor, et al.
Published: (2026)
by: Skopin, Egor, et al.
Published: (2026)
Can Code Language Models Learn Clarification-Seeking Behaviors?
by: Wu, Jie JW, et al.
Published: (2025)
by: Wu, Jie JW, et al.
Published: (2025)
CodeReasoner: Enhancing the Code Reasoning Ability with Reinforcement Learning
by: Tang, Lingxiao, et al.
Published: (2025)
by: Tang, Lingxiao, et al.
Published: (2025)
Instructive Code Retriever: Learn from Large Language Model's Feedback for Code Intelligence Tasks
by: Lu, Jiawei, et al.
Published: (2024)
by: Lu, Jiawei, et al.
Published: (2024)
SpecTra: Enhancing the Code Translation Ability of Language Models by Generating Multi-Modal Specifications
by: Nitin, Vikram, et al.
Published: (2024)
by: Nitin, Vikram, et al.
Published: (2024)
AI-powered Code Review with LLMs: Early Results
by: Rasheed, Zeeshan, et al.
Published: (2024)
by: Rasheed, Zeeshan, et al.
Published: (2024)
ATLAS: Automated Toolkit for Large-Scale Verified Code Synthesis
by: Baksys, Mantas, et al.
Published: (2025)
by: Baksys, Mantas, et al.
Published: (2025)
VerifyThisBench: Generating Code, Specifications, and Proofs All at Once
by: Deng, Xun, et al.
Published: (2025)
by: Deng, Xun, et al.
Published: (2025)
SmartPatchLinker: An Open-Source Tool to Linked Changes Detection for Code Review
by: Khemissi, Islem, et al.
Published: (2026)
by: Khemissi, Islem, et al.
Published: (2026)
Why Personalizing Deep Learning-Based Code Completion Tools Matters
by: Giagnorio, Alessandro, et al.
Published: (2025)
by: Giagnorio, Alessandro, et al.
Published: (2025)
Understanding Dominant Themes in Reviewing Agentic AI-authored Code
by: Haider, Md. Asif, et al.
Published: (2026)
by: Haider, Md. Asif, et al.
Published: (2026)
RevMine: An LLM-Assisted Tool for Code Review Mining and Analysis Across Git Platforms
by: Kansab, Samah, et al.
Published: (2025)
by: Kansab, Samah, et al.
Published: (2025)
TransferFuzz: Fuzzing with Historical Trace for Verifying Propagated Vulnerability Code
by: Li, Siyuan, et al.
Published: (2024)
by: Li, Siyuan, et al.
Published: (2024)
RepoST: Scalable Repository-Level Coding Environment Construction with Sandbox Testing
by: Xie, Yiqing, et al.
Published: (2025)
by: Xie, Yiqing, et al.
Published: (2025)
Toward Inclusive AI-Driven Development: Exploring Gender Differences in Code Generation Tool Interactions
by: Basha, Manaal, et al.
Published: (2025)
by: Basha, Manaal, et al.
Published: (2025)
RLCoder: Reinforcement Learning for Repository-Level Code Completion
by: Wang, Yanlin, et al.
Published: (2024)
by: Wang, Yanlin, et al.
Published: (2024)
Similar Items
-
CRScore: Grounding Automated Evaluation of Code Review Comments in Code Claims and Smells
by: Naik, Atharva, et al.
Published: (2024) -
An Empirical Study on Strong-Weak Model Collaboration for Repo-level Code Generation
by: Gandhi, Shubham, et al.
Published: (2025) -
Data Augmentation for Code Translation with Comparable Corpora and Multiple References
by: Xie, Yiqing, et al.
Published: (2023) -
On the Limitations of Embedding Based Methods for Measuring Functional Correctness for Code Generation
by: Naik, Atharva
Published: (2024) -
ChartEditBench: Evaluating Grounded Multi-Turn Chart Editing in Multimodal Language Models
by: Kapadnis, Manav Nitin, et al.
Published: (2026)