Reflective Paper-to-Code Reproduction Enabled by Fine-Grained Verification
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Mingyang, Yao, Quanming, Du, Lun, Wei, Lanning, Zheng, Da |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ReCode: Improving LLM-based Code Repair with Fine-Grained Retrieval-Augmented Generation
by: Zhao, Yicong, et al.
Published: (2025)
by: Zhao, Yicong, et al.
Published: (2025)
Benchmarking LLMs for Fine-Grained Code Review with Enriched Context in Practice
by: Hu, Ruida, et al.
Published: (2025)
by: Hu, Ruida, et al.
Published: (2025)
Turning the Tide: Repository-based Code Reflection
by: Zhang, Wei, et al.
Published: (2025)
by: Zhang, Wei, et al.
Published: (2025)
TransAgent: Enhancing LLM-Based Code Translation via Fine-Grained Execution Alignment
by: Yuan, Zhiqiang, et al.
Published: (2024)
by: Yuan, Zhiqiang, et al.
Published: (2024)
What Makes AI Research Replicable? Executable Knowledge Graphs as Scientific Knowledge Representations
by: Luo, Yujie, et al.
Published: (2025)
by: Luo, Yujie, et al.
Published: (2025)
Skeleton-Guided-Translation: A Benchmarking Framework for Code Repository Translation with Fine-Grained Quality Evaluation
by: Zhang, Xing, et al.
Published: (2025)
by: Zhang, Xing, et al.
Published: (2025)
$T^3$: Multi-level Tree-based Automatic Program Repair with Large Language Models
by: Liu, Quanming, et al.
Published: (2025)
by: Liu, Quanming, et al.
Published: (2025)
LLMs as Continuous Learners: Improving the Reproduction of Defective Code in Software Issues
by: Lin, Yalan, et al.
Published: (2024)
by: Lin, Yalan, et al.
Published: (2024)
Comparative Analysis of Large Language Models for Context-Aware Code Completion using SAFIM Framework
by: Zhang, Hang, et al.
Published: (2025)
by: Zhang, Hang, et al.
Published: (2025)
Goedel-Code-Prover: Hierarchical Proof Search for Open State-of-the-Art Code Verification
by: Li, Zenan, et al.
Published: (2026)
by: Li, Zenan, et al.
Published: (2026)
FastCoder: Accelerating Repository-level Code Generation via Efficient Retrieval and Verification
by: Zhao, Qianhui, et al.
Published: (2025)
by: Zhao, Qianhui, et al.
Published: (2025)
Vibe-Coding: Feedback-Based Automated Verification with no Human Code Inspection, a Feasibility Study
by: Töpfer, Michal, et al.
Published: (2026)
by: Töpfer, Michal, et al.
Published: (2026)
Verifying LLM-Generated Code in the Context of Software Verification with Ada/SPARK
by: Cramer, Marcos, et al.
Published: (2025)
by: Cramer, Marcos, et al.
Published: (2025)
Chain of Targeted Verification Questions to Improve the Reliability of Code Generated by LLMs
by: Ngassom, Sylvain Kouemo, et al.
Published: (2024)
by: Ngassom, Sylvain Kouemo, et al.
Published: (2024)
Automated Repair of AI Code with Large Language Models and Formal Verification
by: Charalambous, Yiannis, et al.
Published: (2024)
by: Charalambous, Yiannis, et al.
Published: (2024)
Perceptual Self-Reflection in Agentic Physics Simulation Code Generation
by: Shende, Prashant, et al.
Published: (2026)
by: Shende, Prashant, et al.
Published: (2026)
From Natural Language to Verified Code: Toward AI Assisted Problem-to-Code Generation with Dafny-Based Formal Verification
by: Erfan, Md, et al.
Published: (2026)
by: Erfan, Md, et al.
Published: (2026)
CLAP: Learning Transferable Binary Code Representations with Natural Language Supervision
by: Wang, Hao, et al.
Published: (2024)
by: Wang, Hao, et al.
Published: (2024)
AutoP2C: An LLM-Based Agent Framework for Code Repository Generation from Multimodal Content in Academic Papers
by: Lin, Zijie, et al.
Published: (2025)
by: Lin, Zijie, et al.
Published: (2025)
CODE-DITING: A Reasoning-Based Metric for Functional Alignment in Code Evaluation
by: Yang, Guang, et al.
Published: (2025)
by: Yang, Guang, et al.
Published: (2025)
Devstral: Fine-tuning Language Models for Coding Agent Applications
by: Rastogi, Abhinav, et al.
Published: (2025)
by: Rastogi, Abhinav, et al.
Published: (2025)
Integrating Symbolic Execution into the Fine-Tuning of Code-Generating LLMs
by: Sakharova, Marina, et al.
Published: (2025)
by: Sakharova, Marina, et al.
Published: (2025)
Towards Automated Formal Verification of Backend Systems with LLMs
by: Xu, Kangping, et al.
Published: (2025)
by: Xu, Kangping, et al.
Published: (2025)
Beyond Function-Level Search: Repository-Aware Dual-Encoder Code Retrieval with Adversarial Verification
by: Liu, Aofan, et al.
Published: (2025)
by: Liu, Aofan, et al.
Published: (2025)
Generating Automotive Code: Large Language Models for Software Development and Verification in Safety-Critical Systems
by: Kirchner, Sven, et al.
Published: (2025)
by: Kirchner, Sven, et al.
Published: (2025)
Exploring the Potential of Large Language Models in Fine-Grained Review Comment Classification
by: Nguyen, Linh, et al.
Published: (2025)
by: Nguyen, Linh, et al.
Published: (2025)
CodeCircuit: Toward Inferring LLM-Generated Code Correctness via Attribution Graphs
by: He, Yicheng, et al.
Published: (2026)
by: He, Yicheng, et al.
Published: (2026)
Hybrid-Code v2: Zero-Hallucination Clinical ICD-10 Coding via Neuro-Symbolic Verification and Automated Knowledge Base Expansion
by: Yu, Yunguo
Published: (2025)
by: Yu, Yunguo
Published: (2025)
Post-Incorporating Code Structural Knowledge into Pretrained Models via ICL for Code Translation
by: Du, Yali, et al.
Published: (2025)
by: Du, Yali, et al.
Published: (2025)
Verification Limits Code LLM Training
by: Gureja, Srishti, et al.
Published: (2025)
by: Gureja, Srishti, et al.
Published: (2025)
Fine-Tuning Code Language Models to Detect Cross-Language Bugs
by: Li, Zengyang, et al.
Published: (2025)
by: Li, Zengyang, et al.
Published: (2025)
Fine-grained Approaches for Confidence Calibration of LLMs in Automated Code Revision
by: Lin, Hong Yi, et al.
Published: (2026)
by: Lin, Hong Yi, et al.
Published: (2026)
Does Your Neural Code Completion Model Use My Code? A Membership Inference Approach
by: Wan, Yao, et al.
Published: (2024)
by: Wan, Yao, et al.
Published: (2024)
Agentic Bug Reproduction for Effective Automated Program Repair at Google
by: Cheng, Runxiang, et al.
Published: (2025)
by: Cheng, Runxiang, et al.
Published: (2025)
Dynamic Cogeneration of Bug Reproduction Test in Agentic Program Repair
by: Cheng, Runxiang, et al.
Published: (2026)
by: Cheng, Runxiang, et al.
Published: (2026)
Code Copycat Conundrum: Demystifying Repetition in LLM-based Code Generation
by: Liu, Mingwei, et al.
Published: (2025)
by: Liu, Mingwei, et al.
Published: (2025)
A New Benchmark for the Appropriate Evaluation of RTL Code Optimization
by: Lu, Yao, et al.
Published: (2026)
by: Lu, Yao, et al.
Published: (2026)
Performance Review on LLM for solving leetcode problems
by: Wang, Lun, et al.
Published: (2025)
by: Wang, Lun, et al.
Published: (2025)
ReCatcher: Towards LLMs Regression Testing for Code Generation
by: Abbassi, Altaf Allah, et al.
Published: (2025)
by: Abbassi, Altaf Allah, et al.
Published: (2025)
Can LLMs Enable Verification in Mainstream Programming?
by: Shefer, Aleksandr, et al.
Published: (2025)
by: Shefer, Aleksandr, et al.
Published: (2025)
Similar Items
-
ReCode: Improving LLM-based Code Repair with Fine-Grained Retrieval-Augmented Generation
by: Zhao, Yicong, et al.
Published: (2025) -
Benchmarking LLMs for Fine-Grained Code Review with Enriched Context in Practice
by: Hu, Ruida, et al.
Published: (2025) -
Turning the Tide: Repository-based Code Reflection
by: Zhang, Wei, et al.
Published: (2025) -
TransAgent: Enhancing LLM-Based Code Translation via Fine-Grained Execution Alignment
by: Yuan, Zhiqiang, et al.
Published: (2024) -
What Makes AI Research Replicable? Executable Knowledge Graphs as Scientific Knowledge Representations
by: Luo, Yujie, et al.
Published: (2025)