Clover: Closed-Loop Verifiable Code Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Chuyue, Sheng, Ying, Padon, Oded, Barrett, Clark |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VeriScale: Adversarial Test-Suite Scaling for Verifiable Code Generation
von: Bai, Yifan, et al.
Veröffentlicht: (2026)
von: Bai, Yifan, et al.
Veröffentlicht: (2026)
VeriContest: A Competitive-Programming Benchmark for Verifiable Code Generation
von: Xie, Zichen, et al.
Veröffentlicht: (2026)
von: Xie, Zichen, et al.
Veröffentlicht: (2026)
LLM-Vectorizer: LLM-based Verified Loop Vectorizer
von: Taneja, Jubi, et al.
Veröffentlicht: (2024)
von: Taneja, Jubi, et al.
Veröffentlicht: (2024)
Talking with Verifiers: Automatic Specification Generation for Neural Network Verification
von: Elboher, Yizhak Y., et al.
Veröffentlicht: (2026)
von: Elboher, Yizhak Y., et al.
Veröffentlicht: (2026)
Proving the Coding Interview: A Benchmark for Formally Verified Code Generation
von: Dougherty, Quinn, et al.
Veröffentlicht: (2025)
von: Dougherty, Quinn, et al.
Veröffentlicht: (2025)
DafnyBench: A Benchmark for Formal Software Verification
von: Loughridge, Chloe, et al.
Veröffentlicht: (2024)
von: Loughridge, Chloe, et al.
Veröffentlicht: (2024)
DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
von: DeepSeek-AI, et al.
Veröffentlicht: (2024)
von: DeepSeek-AI, et al.
Veröffentlicht: (2024)
VeriStruct: AI-assisted Automated Verification of Data-Structure Modules in Verus
von: Sun, Chuyue, et al.
Veröffentlicht: (2025)
von: Sun, Chuyue, et al.
Veröffentlicht: (2025)
Exploring Pass-Rate Reward in Reinforcement Learning for Code Generation
von: Li, Xin-Ye, et al.
Veröffentlicht: (2026)
von: Li, Xin-Ye, et al.
Veröffentlicht: (2026)
Scoring Verifiers: Evaluating Synthetic Verification for Code and Reasoning
von: Ficek, Aleksander, et al.
Veröffentlicht: (2025)
von: Ficek, Aleksander, et al.
Veröffentlicht: (2025)
Causal Fuzzing for Verifying Machine Unlearning
von: Mazhar, Anna, et al.
Veröffentlicht: (2025)
von: Mazhar, Anna, et al.
Veröffentlicht: (2025)
CodeTaste: Can LLMs Generate Human-Level Code Refactorings?
von: Thillen, Alex, et al.
Veröffentlicht: (2026)
von: Thillen, Alex, et al.
Veröffentlicht: (2026)
DevBench: A Realistic, Developer-Informed Benchmark for Code Generation Models
von: Kumarappan, Adarsh, et al.
Veröffentlicht: (2026)
von: Kumarappan, Adarsh, et al.
Veröffentlicht: (2026)
Optimizing AI-Assisted Code Generation
von: Torka, Simon, et al.
Veröffentlicht: (2024)
von: Torka, Simon, et al.
Veröffentlicht: (2024)
Operational Robustness of LLMs on Code Generation
von: Paul, Debalina Ghosh, et al.
Veröffentlicht: (2026)
von: Paul, Debalina Ghosh, et al.
Veröffentlicht: (2026)
Can Coding Agents Be General Agents?
von: Ivanov, Maksim, et al.
Veröffentlicht: (2026)
von: Ivanov, Maksim, et al.
Veröffentlicht: (2026)
VERINA: Benchmarking Verifiable Code Generation
von: Ye, Zhe, et al.
Veröffentlicht: (2025)
von: Ye, Zhe, et al.
Veröffentlicht: (2025)
Functional Overlap Reranking for Neural Code Generation
von: To, Hung Quoc, et al.
Veröffentlicht: (2023)
von: To, Hung Quoc, et al.
Veröffentlicht: (2023)
Code Generation by Differential Test Time Scaling
von: He, Yifeng, et al.
Veröffentlicht: (2026)
von: He, Yifeng, et al.
Veröffentlicht: (2026)
JARVIS: A Multi-Agent Code Assistant for High-Quality EDA Script Generation
von: Pasandi, Ghasem, et al.
Veröffentlicht: (2025)
von: Pasandi, Ghasem, et al.
Veröffentlicht: (2025)
SoundnessBench: A Soundness Benchmark for Neural Network Verifiers
von: Zhou, Xingjian, et al.
Veröffentlicht: (2024)
von: Zhou, Xingjian, et al.
Veröffentlicht: (2024)
Protocode: Prototype-Driven Interpretability for Code Generation in LLMs
von: Bodla, Krishna Vamshi, et al.
Veröffentlicht: (2025)
von: Bodla, Krishna Vamshi, et al.
Veröffentlicht: (2025)
A Theoretical Analysis of Test-Driven Code Generation
von: Menet, Nicolas, et al.
Veröffentlicht: (2026)
von: Menet, Nicolas, et al.
Veröffentlicht: (2026)
Sketch-and-Verify: Structured Inference-Time Scaling via Program Sketching
von: Jiang, Shan, et al.
Veröffentlicht: (2026)
von: Jiang, Shan, et al.
Veröffentlicht: (2026)
CodeGeeX: A Pre-Trained Model for Code Generation with Multilingual Benchmarking on HumanEval-X
von: Zheng, Qinkai, et al.
Veröffentlicht: (2023)
von: Zheng, Qinkai, et al.
Veröffentlicht: (2023)
Supersonic: Learning to Generate Source Code Optimizations in C/C++
von: Chen, Zimin, et al.
Veröffentlicht: (2023)
von: Chen, Zimin, et al.
Veröffentlicht: (2023)
Enhancing LLM-Based Test Generation by Eliminating Covered Code
von: Xu, WeiZhe, et al.
Veröffentlicht: (2026)
von: Xu, WeiZhe, et al.
Veröffentlicht: (2026)
Deep-Bench: Deep Learning Benchmark Dataset for Code Generation
von: Daghighfarsoodeh, Alireza, et al.
Veröffentlicht: (2025)
von: Daghighfarsoodeh, Alireza, et al.
Veröffentlicht: (2025)
Drawing Pandas: A Benchmark for LLMs in Generating Plotting Code
von: Galimzyanov, Timur, et al.
Veröffentlicht: (2024)
von: Galimzyanov, Timur, et al.
Veröffentlicht: (2024)
Semantic Voting: Execution-Grounded Consensus for LLM Code Generation
von: Jiang, Shan, et al.
Veröffentlicht: (2026)
von: Jiang, Shan, et al.
Veröffentlicht: (2026)
Co-Located Tests, Better AI Code: How Test Syntax Structure Affects Foundation Model Code Generation
von: Jacopin, Éric
Veröffentlicht: (2026)
von: Jacopin, Éric
Veröffentlicht: (2026)
An Initial Exploration of Contrastive Prompt Tuning to Generate Energy-Efficient Code
von: Weidmann, Sophie, et al.
Veröffentlicht: (2026)
von: Weidmann, Sophie, et al.
Veröffentlicht: (2026)
LiCoEval: Evaluating LLMs on License Compliance in Code Generation
von: Xu, Weiwei, et al.
Veröffentlicht: (2024)
von: Xu, Weiwei, et al.
Veröffentlicht: (2024)
Promise and Peril of Collaborative Code Generation Models: Balancing Effectiveness and Memorization
von: Chen, Zhi, et al.
Veröffentlicht: (2024)
von: Chen, Zhi, et al.
Veröffentlicht: (2024)
Beyond Verifiable Rewards: Rubric-Based GRM for Reinforced Fine-Tuning SWE Agents
von: Huang, Jiawei, et al.
Veröffentlicht: (2026)
von: Huang, Jiawei, et al.
Veröffentlicht: (2026)
The Counterfeit Conundrum: Can Code Language Models Grasp the Nuances of Their Incorrect Generations?
von: Gu, Alex, et al.
Veröffentlicht: (2024)
von: Gu, Alex, et al.
Veröffentlicht: (2024)
How Efficient is LLM-Generated Code? A Rigorous & High-Standard Benchmark
von: Qiu, Ruizhong, et al.
Veröffentlicht: (2024)
von: Qiu, Ruizhong, et al.
Veröffentlicht: (2024)
Large Language Models Should Ask Clarifying Questions to Increase Confidence in Generated Code
von: Wu, Jie JW
Veröffentlicht: (2023)
von: Wu, Jie JW
Veröffentlicht: (2023)
scicode-lint: Detecting Methodology Bugs in Scientific Python Code with LLM-Generated Patterns
von: Samsonau, Sergey V.
Veröffentlicht: (2026)
von: Samsonau, Sergey V.
Veröffentlicht: (2026)
Prism: Dynamic and Flexible Benchmarking of LLMs Code Generation with Monte Carlo Tree Search
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2025)
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
VeriScale: Adversarial Test-Suite Scaling for Verifiable Code Generation
von: Bai, Yifan, et al.
Veröffentlicht: (2026) -
VeriContest: A Competitive-Programming Benchmark for Verifiable Code Generation
von: Xie, Zichen, et al.
Veröffentlicht: (2026) -
LLM-Vectorizer: LLM-based Verified Loop Vectorizer
von: Taneja, Jubi, et al.
Veröffentlicht: (2024) -
Talking with Verifiers: Automatic Specification Generation for Neural Network Verification
von: Elboher, Yizhak Y., et al.
Veröffentlicht: (2026) -
Proving the Coding Interview: A Benchmark for Formally Verified Code Generation
von: Dougherty, Quinn, et al.
Veröffentlicht: (2025)