OpenCodeInterpreter: Integrating Code Generation with Execution and Refinement
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zheng, Tianyu, Zhang, Ge, Shen, Tianhao, Liu, Xueling, Lin, Bill Yuchen, Fu, Jie, Chen, Wenhu, Yue, Xiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VisCoder: Fine-Tuning LLMs for Executable Python Visualization Code Generation
von: Ni, Yuansheng, et al.
Veröffentlicht: (2025)
von: Ni, Yuansheng, et al.
Veröffentlicht: (2025)
CodeEditorBench: Evaluating Code Editing Capability of Large Language Models
von: Guo, Jiawei, et al.
Veröffentlicht: (2024)
von: Guo, Jiawei, et al.
Veröffentlicht: (2024)
Python Symbolic Execution with LLM-powered Code Generation
von: Wang, Wenhan, et al.
Veröffentlicht: (2024)
von: Wang, Wenhan, et al.
Veröffentlicht: (2024)
CodeHalu: Investigating Code Hallucinations in LLMs via Execution-based Verification
von: Tian, Yuchen, et al.
Veröffentlicht: (2024)
von: Tian, Yuchen, et al.
Veröffentlicht: (2024)
Iterative Refinement of Project-Level Code Context for Precise Code Generation with Compiler Feedback
von: Bi, Zhangqian, et al.
Veröffentlicht: (2024)
von: Bi, Zhangqian, et al.
Veröffentlicht: (2024)
CodeBenchGen: Creating Scalable Execution-based Code Generation Benchmarks
von: Xie, Yiqing, et al.
Veröffentlicht: (2024)
von: Xie, Yiqing, et al.
Veröffentlicht: (2024)
StepCodeReasoner: Aligning Code Reasoning with Stepwise Execution Traces via Reinforcement Learning
von: Wang, Hao, et al.
Veröffentlicht: (2026)
von: Wang, Hao, et al.
Veröffentlicht: (2026)
Executing as You Generate: Hiding Execution Latency in LLM Code Generation
von: Sun, Zhensu, et al.
Veröffentlicht: (2026)
von: Sun, Zhensu, et al.
Veröffentlicht: (2026)
CYCLE: Learning to Self-Refine the Code Generation
von: Ding, Yangruibo, et al.
Veröffentlicht: (2024)
von: Ding, Yangruibo, et al.
Veröffentlicht: (2024)
CodeRL+: Improving Code Generation via Reinforcement with Execution Semantics Alignment
von: Jiang, Xue, et al.
Veröffentlicht: (2025)
von: Jiang, Xue, et al.
Veröffentlicht: (2025)
KCoEvo: A Knowledge Graph Augmented Framework for Evolutionary Code Generation
von: Kang, Jiazhen, et al.
Veröffentlicht: (2026)
von: Kang, Jiazhen, et al.
Veröffentlicht: (2026)
GenX: Mastering Code and Test Generation with Execution Feedback
von: Wang, Nan, et al.
Veröffentlicht: (2024)
von: Wang, Nan, et al.
Veröffentlicht: (2024)
CodeScore: Evaluating Code Generation by Learning Code Execution
von: Dong, Yihong, et al.
Veröffentlicht: (2023)
von: Dong, Yihong, et al.
Veröffentlicht: (2023)
Enhancing Code LLMs with Reinforcement Learning in Code Generation: A Survey
von: Wang, Junqiao, et al.
Veröffentlicht: (2024)
von: Wang, Junqiao, et al.
Veröffentlicht: (2024)
CodeSpecBench: Benchmarking LLMs for Executable Behavioral Specification Generation
von: Chen, Zaoyu, et al.
Veröffentlicht: (2026)
von: Chen, Zaoyu, et al.
Veröffentlicht: (2026)
Defusing Logic Bombs in Symbolic Execution with LLM-Generated Ghost Code
von: Bouras, Dimitrios Stamatios, et al.
Veröffentlicht: (2026)
von: Bouras, Dimitrios Stamatios, et al.
Veröffentlicht: (2026)
IntentCoding: Amplifying User Intent in Code Generation
von: Fang, Zheng, et al.
Veröffentlicht: (2026)
von: Fang, Zheng, et al.
Veröffentlicht: (2026)
BanglaForge: LLM Collaboration with Self-Refinement for Bangla Code Generation
von: Dihan, Mahir Labib, et al.
Veröffentlicht: (2025)
von: Dihan, Mahir Labib, et al.
Veröffentlicht: (2025)
BigCodeArena: Unveiling More Reliable Human Preferences in Code Generation via Execution
von: Zhuo, Terry Yue, et al.
Veröffentlicht: (2025)
von: Zhuo, Terry Yue, et al.
Veröffentlicht: (2025)
DuET: Dual Execution for Test Output Prediction with Generated Code and Pseudocode
von: Han, Hojae, et al.
Veröffentlicht: (2026)
von: Han, Hojae, et al.
Veröffentlicht: (2026)
PerfCodeGen: Improving Performance of LLM Generated Code with Execution Feedback
von: Peng, Yun, et al.
Veröffentlicht: (2024)
von: Peng, Yun, et al.
Veröffentlicht: (2024)
CodeJudge-Eval: Can Large Language Models be Good Judges in Code Understanding?
von: Zhao, Yuwei, et al.
Veröffentlicht: (2024)
von: Zhao, Yuwei, et al.
Veröffentlicht: (2024)
CodeScope: An Execution-based Multilingual Multitask Multidimensional Benchmark for Evaluating LLMs on Code Understanding and Generation
von: Yan, Weixiang, et al.
Veröffentlicht: (2023)
von: Yan, Weixiang, et al.
Veröffentlicht: (2023)
CodeIF-Bench: Evaluating Instruction-Following Capabilities of Large Language Models in Interactive Code Generation
von: Wang, Peiding, et al.
Veröffentlicht: (2025)
von: Wang, Peiding, et al.
Veröffentlicht: (2025)
EvoCodeBench: An Evolving Code Generation Benchmark with Domain-Specific Evaluations
von: Li, Jia, et al.
Veröffentlicht: (2024)
von: Li, Jia, et al.
Veröffentlicht: (2024)
Rethinking Code Refinement: Learning to Judge Code Efficiency
von: Seo, Minju, et al.
Veröffentlicht: (2024)
von: Seo, Minju, et al.
Veröffentlicht: (2024)
RethinkMCTS: Refining Erroneous Thoughts in Monte Carlo Tree Search for Code Generation
von: Li, Qingyao, et al.
Veröffentlicht: (2024)
von: Li, Qingyao, et al.
Veröffentlicht: (2024)
Rethinking Repetition Problems of LLMs in Code Generation
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
LLMigrate: Transforming "Lazy" Large Language Models into Efficient Source Code Migrators
von: Liu, Yuchen, et al.
Veröffentlicht: (2025)
von: Liu, Yuchen, et al.
Veröffentlicht: (2025)
ContextEcho: A Benchmark for Persona Drift in Long Agentic-Coding Sessions
von: Ding, Xianzhong, et al.
Veröffentlicht: (2026)
von: Ding, Xianzhong, et al.
Veröffentlicht: (2026)
EvoCodeBench: An Evolving Code Generation Benchmark Aligned with Real-World Code Repositories
von: Li, Jia, et al.
Veröffentlicht: (2024)
von: Li, Jia, et al.
Veröffentlicht: (2024)
Evaluation of Code LLMs on Geospatial Code Generation
von: Gramacki, Piotr, et al.
Veröffentlicht: (2024)
von: Gramacki, Piotr, et al.
Veröffentlicht: (2024)
Showing LLM-Generated Code Selectively Based on Confidence of LLMs
von: Li, Jia, et al.
Veröffentlicht: (2024)
von: Li, Jia, et al.
Veröffentlicht: (2024)
Measuring the Influence of Incorrect Code on Test Generation
von: Huang, Dong, et al.
Veröffentlicht: (2024)
von: Huang, Dong, et al.
Veröffentlicht: (2024)
DevEval: A Manually-Annotated Code Generation Benchmark Aligned with Real-World Code Repositories
von: Li, Jia, et al.
Veröffentlicht: (2024)
von: Li, Jia, et al.
Veröffentlicht: (2024)
VisCoder2: Building Multi-Language Visualization Coding Agents
von: Ni, Yuansheng, et al.
Veröffentlicht: (2025)
von: Ni, Yuansheng, et al.
Veröffentlicht: (2025)
OpenCodeInstruct: A Large-scale Instruction Tuning Dataset for Code LLMs
von: Ahmad, Wasi Uddin, et al.
Veröffentlicht: (2025)
von: Ahmad, Wasi Uddin, et al.
Veröffentlicht: (2025)
AutoCodeBench: Large Language Models are Automatic Code Benchmark Generators
von: Chou, Jason, et al.
Veröffentlicht: (2025)
von: Chou, Jason, et al.
Veröffentlicht: (2025)
Code Fingerprints: Disentangled Attribution of LLM-Generated Code
von: Guo, Jiaxun, et al.
Veröffentlicht: (2026)
von: Guo, Jiaxun, et al.
Veröffentlicht: (2026)
VersiCode: Towards Version-controllable Code Generation
von: Wu, Tongtong, et al.
Veröffentlicht: (2024)
von: Wu, Tongtong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
VisCoder: Fine-Tuning LLMs for Executable Python Visualization Code Generation
von: Ni, Yuansheng, et al.
Veröffentlicht: (2025) -
CodeEditorBench: Evaluating Code Editing Capability of Large Language Models
von: Guo, Jiawei, et al.
Veröffentlicht: (2024) -
Python Symbolic Execution with LLM-powered Code Generation
von: Wang, Wenhan, et al.
Veröffentlicht: (2024) -
CodeHalu: Investigating Code Hallucinations in LLMs via Execution-based Verification
von: Tian, Yuchen, et al.
Veröffentlicht: (2024) -
Iterative Refinement of Project-Level Code Context for Precise Code Generation with Compiler Feedback
von: Bi, Zhangqian, et al.
Veröffentlicht: (2024)