CodeScore: Evaluating Code Generation by Learning Code Execution
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dong, Yihong, Ding, Jiazheng, Jiang, Xue, Li, Ge, Li, Zhuo, Jin, Zhi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Self-collaboration Code Generation via ChatGPT
von: Dong, Yihong, et al.
Veröffentlicht: (2023)
von: Dong, Yihong, et al.
Veröffentlicht: (2023)
CodeScore-R: An Automated Robustness Metric for Assessing the FunctionalCorrectness of Code Synthesis
von: Yang, Guang, et al.
Veröffentlicht: (2024)
von: Yang, Guang, et al.
Veröffentlicht: (2024)
Uncertainty-Guided Chain-of-Thought for Code Generation with LLMs
von: Zhu, Yuqi, et al.
Veröffentlicht: (2025)
von: Zhu, Yuqi, et al.
Veröffentlicht: (2025)
EvoCodeBench: An Evolving Code Generation Benchmark Aligned with Real-World Code Repositories
von: Li, Jia, et al.
Veröffentlicht: (2024)
von: Li, Jia, et al.
Veröffentlicht: (2024)
CodeDPO: Aligning Code Models with Self Generated and Verified Source Code
von: Zhang, Kechi, et al.
Veröffentlicht: (2024)
von: Zhang, Kechi, et al.
Veröffentlicht: (2024)
Rethinking Repetition Problems of LLMs in Code Generation
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
Self-planning Code Generation with Large Language Models
von: Jiang, Xue, et al.
Veröffentlicht: (2023)
von: Jiang, Xue, et al.
Veröffentlicht: (2023)
CodeRL+: Improving Code Generation via Reinforcement with Execution Semantics Alignment
von: Jiang, Xue, et al.
Veröffentlicht: (2025)
von: Jiang, Xue, et al.
Veröffentlicht: (2025)
EvoCodeBench: An Evolving Code Generation Benchmark with Domain-Specific Evaluations
von: Li, Jia, et al.
Veröffentlicht: (2024)
von: Li, Jia, et al.
Veröffentlicht: (2024)
IntentCoding: Amplifying User Intent in Code Generation
von: Fang, Zheng, et al.
Veröffentlicht: (2026)
von: Fang, Zheng, et al.
Veröffentlicht: (2026)
ROCODE: Integrating Backtracking Mechanism and Program Analysis in Large Language Models for Code Generation
von: Jiang, Xue, et al.
Veröffentlicht: (2024)
von: Jiang, Xue, et al.
Veröffentlicht: (2024)
Line-level Semantic Structure Learning for Code Vulnerability Detection
von: Wang, Ziliang, et al.
Veröffentlicht: (2024)
von: Wang, Ziliang, et al.
Veröffentlicht: (2024)
Exploring Data-Efficient Adaptation of Large Language Models for Code Generation
von: Jiang, Xue, et al.
Veröffentlicht: (2024)
von: Jiang, Xue, et al.
Veröffentlicht: (2024)
Think Anywhere in Code Generation
von: Jiang, Xue, et al.
Veröffentlicht: (2026)
von: Jiang, Xue, et al.
Veröffentlicht: (2026)
Focused-DPO: Enhancing Code Generation Through Focused Preference Optimization on Error-Prone Points
von: Zhang, Kechi, et al.
Veröffentlicht: (2025)
von: Zhang, Kechi, et al.
Veröffentlicht: (2025)
A Survey on Code Generation with LLM-based Agents
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
CodeAgent: Enhancing Code Generation with Tool-Integrated Agent Systems for Real-World Repo-level Coding Challenges
von: Zhang, Kechi, et al.
Veröffentlicht: (2024)
von: Zhang, Kechi, et al.
Veröffentlicht: (2024)
Sifting through the Chaff: On Utilizing Execution Feedback for Ranking the Generated Code Candidates
von: Sun, Zhihong, et al.
Veröffentlicht: (2024)
von: Sun, Zhihong, et al.
Veröffentlicht: (2024)
Large Language Model Unlearning for Source Code
von: Jiang, Xue, et al.
Veröffentlicht: (2025)
von: Jiang, Xue, et al.
Veröffentlicht: (2025)
DevEval: A Manually-Annotated Code Generation Benchmark Aligned with Real-World Code Repositories
von: Li, Jia, et al.
Veröffentlicht: (2024)
von: Li, Jia, et al.
Veröffentlicht: (2024)
Python Symbolic Execution with LLM-powered Code Generation
von: Wang, Wenhan, et al.
Veröffentlicht: (2024)
von: Wang, Wenhan, et al.
Veröffentlicht: (2024)
DevEval: Evaluating Code Generation in Practical Software Projects
von: Li, Jia, et al.
Veröffentlicht: (2024)
von: Li, Jia, et al.
Veröffentlicht: (2024)
GraphCodeAgent: Dual Graph-Guided LLM Agent for Retrieval-Augmented Repo-Level Code Generation
von: Li, Jia, et al.
Veröffentlicht: (2025)
von: Li, Jia, et al.
Veröffentlicht: (2025)
RedCode: Risky Code Execution and Generation Benchmark for Code Agents
von: Guo, Chengquan, et al.
Veröffentlicht: (2024)
von: Guo, Chengquan, et al.
Veröffentlicht: (2024)
CodeArena: A Collective Evaluation Platform for LLM Code Generation
von: Du, Mingzhe, et al.
Veröffentlicht: (2025)
von: Du, Mingzhe, et al.
Veröffentlicht: (2025)
SemGuard: Real-Time Semantic Evaluator for Correcting LLM-Generated Code
von: Wang, Qinglin, et al.
Veröffentlicht: (2025)
von: Wang, Qinglin, et al.
Veröffentlicht: (2025)
From I/O to Code with Discovery Agent
von: Dong, Yihong, et al.
Veröffentlicht: (2026)
von: Dong, Yihong, et al.
Veröffentlicht: (2026)
aiXcoder-7B-v2: Training LLMs to Fully Utilize the Long Context in Repository-level Code Completion
von: Li, Jia, et al.
Veröffentlicht: (2025)
von: Li, Jia, et al.
Veröffentlicht: (2025)
Showing LLM-Generated Code Selectively Based on Confidence of LLMs
von: Li, Jia, et al.
Veröffentlicht: (2024)
von: Li, Jia, et al.
Veröffentlicht: (2024)
RealBench: A Repo-Level Code Generation Benchmark Aligned with Real-World Software Development Practices
von: Li, Jia, et al.
Veröffentlicht: (2026)
von: Li, Jia, et al.
Veröffentlicht: (2026)
Executing as You Generate: Hiding Execution Latency in LLM Code Generation
von: Sun, Zhensu, et al.
Veröffentlicht: (2026)
von: Sun, Zhensu, et al.
Veröffentlicht: (2026)
TRACE: Evaluating Execution Efficiency of LLM-Based Code Translation
von: Gong, Zhihao, et al.
Veröffentlicht: (2026)
von: Gong, Zhihao, et al.
Veröffentlicht: (2026)
TRACE: Evaluating Execution Efficiency of LLM-Based Code Translation
von: Gong, Zhihao, et al.
Veröffentlicht: (2025)
von: Gong, Zhihao, et al.
Veröffentlicht: (2025)
HiRoPE: Length Extrapolation for Code Models Using Hierarchical Position
von: Zhang, Kechi, et al.
Veröffentlicht: (2024)
von: Zhang, Kechi, et al.
Veröffentlicht: (2024)
OpenCodeInterpreter: Integrating Code Generation with Execution and Refinement
von: Zheng, Tianyu, et al.
Veröffentlicht: (2024)
von: Zheng, Tianyu, et al.
Veröffentlicht: (2024)
ICE-Score: Instructing Large Language Models to Evaluate Code
von: Zhuo, Terry Yue
Veröffentlicht: (2023)
von: Zhuo, Terry Yue
Veröffentlicht: (2023)
Knowledge-Aware Code Generation with Large Language Models
von: Huang, Tao, et al.
Veröffentlicht: (2024)
von: Huang, Tao, et al.
Veröffentlicht: (2024)
SelfPiCo: Self-Guided Partial Code Execution with LLMs
von: Xue, Zhipeng, et al.
Veröffentlicht: (2024)
von: Xue, Zhipeng, et al.
Veröffentlicht: (2024)
Deep Learning Based Code Generation Methods: Literature Review
von: Yang, Zezhou, et al.
Veröffentlicht: (2023)
von: Yang, Zezhou, et al.
Veröffentlicht: (2023)
PACE: Improving Prompt with Actor-Critic Editing for Large Language Model
von: Dong, Yihong, et al.
Veröffentlicht: (2023)
von: Dong, Yihong, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Self-collaboration Code Generation via ChatGPT
von: Dong, Yihong, et al.
Veröffentlicht: (2023) -
CodeScore-R: An Automated Robustness Metric for Assessing the FunctionalCorrectness of Code Synthesis
von: Yang, Guang, et al.
Veröffentlicht: (2024) -
Uncertainty-Guided Chain-of-Thought for Code Generation with LLMs
von: Zhu, Yuqi, et al.
Veröffentlicht: (2025) -
EvoCodeBench: An Evolving Code Generation Benchmark Aligned with Real-World Code Repositories
von: Li, Jia, et al.
Veröffentlicht: (2024) -
CodeDPO: Aligning Code Models with Self Generated and Verified Source Code
von: Zhang, Kechi, et al.
Veröffentlicht: (2024)