AutoICE: Automatically Synthesizing Verifiable C Code via LLM-driven Evolution
Fuente:
arXiv
Guardado en:
| Autores principales: | Luo, Weilin, Liang, Xueyi, Deng, Haotian, Liu, Yanan, Wan, Hai |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AutoWebWorld: Synthesizing Infinite Verifiable Web Environments via Finite State Machines
por: Wu, Yifan, et al.
Publicado: (2026)
por: Wu, Yifan, et al.
Publicado: (2026)
irace-evo: Automatic Algorithm Configuration Extended With LLM-Based Code Evolution
por: Sartori, Camilo Chacón, et al.
Publicado: (2025)
por: Sartori, Camilo Chacón, et al.
Publicado: (2025)
Verifying LLM-Generated Code in the Context of Software Verification with Ada/SPARK
por: Cramer, Marcos, et al.
Publicado: (2025)
por: Cramer, Marcos, et al.
Publicado: (2025)
CodeSift: An LLM-Based Reference-Less Framework for Automatic Code Validation
por: Aggarwal, Pooja, et al.
Publicado: (2024)
por: Aggarwal, Pooja, et al.
Publicado: (2024)
AutoCodeBench: Large Language Models are Automatic Code Benchmark Generators
por: Chou, Jason, et al.
Publicado: (2025)
por: Chou, Jason, et al.
Publicado: (2025)
Automatic Building Code Review: A Case Study
por: Wan, Hanlong, et al.
Publicado: (2025)
por: Wan, Hanlong, et al.
Publicado: (2025)
CVeDRL: An Efficient Code Verifier via Difficulty-aware Reinforcement Learning
por: Shi, Ji, et al.
Publicado: (2026)
por: Shi, Ji, et al.
Publicado: (2026)
RustEvo^2: An Evolving Benchmark for API Evolution in LLM-based Rust Code Generation
por: Liang, Linxi, et al.
Publicado: (2025)
por: Liang, Linxi, et al.
Publicado: (2025)
EvoCodeBench: A Human-Performance Benchmark for Self-Evolving LLM-Driven Coding Systems
por: Zhang, Wentao, et al.
Publicado: (2026)
por: Zhang, Wentao, et al.
Publicado: (2026)
PyVeritas: On Verifying Python via LLM-Based Transpilation and Bounded Model Checking for C
por: Orvalho, Pedro, et al.
Publicado: (2025)
por: Orvalho, Pedro, et al.
Publicado: (2025)
SIEVE: Towards Verifiable Certification for Code-datasets
por: Mbodji, Fatou Ndiaye, et al.
Publicado: (2025)
por: Mbodji, Fatou Ndiaye, et al.
Publicado: (2025)
AutoP2C: An LLM-Based Agent Framework for Code Repository Generation from Multimodal Content in Academic Papers
por: Lin, Zijie, et al.
Publicado: (2025)
por: Lin, Zijie, et al.
Publicado: (2025)
AutoMCQ -- Automatically Generate Code Comprehension Questions using GenAI
por: Goodfellow, Martin, et al.
Publicado: (2025)
por: Goodfellow, Martin, et al.
Publicado: (2025)
Does Your Neural Code Completion Model Use My Code? A Membership Inference Approach
por: Wan, Yao, et al.
Publicado: (2024)
por: Wan, Yao, et al.
Publicado: (2024)
Reducing Hallucinations in LLM-Generated Code via Semantic Triangulation
por: Dai, Yihan, et al.
Publicado: (2025)
por: Dai, Yihan, et al.
Publicado: (2025)
FrontendBench: A Benchmark for Evaluating LLMs on Front-End Development via Automatic Evaluation
por: Zhu, Hongda, et al.
Publicado: (2025)
por: Zhu, Hongda, et al.
Publicado: (2025)
DialogAgent: An Auto-engagement Agent for Code Question Answering Data Production
por: Liang, Xiaoyun, et al.
Publicado: (2024)
por: Liang, Xiaoyun, et al.
Publicado: (2024)
WybeCoder: Verified Imperative Code Generation
por: Gloeckle, Fabian, et al.
Publicado: (2026)
por: Gloeckle, Fabian, et al.
Publicado: (2026)
VeriScale: Adversarial Test-Suite Scaling for Verifiable Code Generation
por: Bai, Yifan, et al.
Publicado: (2026)
por: Bai, Yifan, et al.
Publicado: (2026)
AutoDroid: LLM-powered Task Automation in Android
por: Wen, Hao, et al.
Publicado: (2023)
por: Wen, Hao, et al.
Publicado: (2023)
Stingy Context: 18:1 Hierarchical Code Compression for LLM Auto-Coding
por: Ostby, David Linus
Publicado: (2026)
por: Ostby, David Linus
Publicado: (2026)
A Qualitative Investigation into LLM-Generated Multilingual Code Comments and Automatic Evaluation Metrics
por: Katzy, Jonathan, et al.
Publicado: (2025)
por: Katzy, Jonathan, et al.
Publicado: (2025)
Aletheia: What Makes RLVR For Code Verifiers Tick?
por: Venkatkrishna, Vatsal, et al.
Publicado: (2026)
por: Venkatkrishna, Vatsal, et al.
Publicado: (2026)
AutoCodeRover: Autonomous Program Improvement
por: Zhang, Yuntong, et al.
Publicado: (2024)
por: Zhang, Yuntong, et al.
Publicado: (2024)
ICE-Score: Instructing Large Language Models to Evaluate Code
por: Zhuo, Terry Yue
Publicado: (2023)
por: Zhuo, Terry Yue
Publicado: (2023)
Automatically Generating UI Code from Screenshot: A Divide-and-Conquer-Based Approach
por: Wan, Yuxuan, et al.
Publicado: (2024)
por: Wan, Yuxuan, et al.
Publicado: (2024)
Toward Automated and Trustworthy Scientific Analysis and Visualization with LLM-Generated Code
por: Chakroborti, Apu Kumar, et al.
Publicado: (2025)
por: Chakroborti, Apu Kumar, et al.
Publicado: (2025)
Edit, But Verify: An Empirical Audit of Instructed Code-Editing Benchmarks
por: Ebrahimi, Amir M., et al.
Publicado: (2026)
por: Ebrahimi, Amir M., et al.
Publicado: (2026)
An Execution-Verified Multi-Language Benchmark for Code Semantic Reasoning
por: Li, Yikun, et al.
Publicado: (2026)
por: Li, Yikun, et al.
Publicado: (2026)
Auto-SPT: Automating Semantic Preserving Transformations for Code
por: Hooda, Ashish, et al.
Publicado: (2025)
por: Hooda, Ashish, et al.
Publicado: (2025)
Using Semantic Distance to Estimate Uncertainty in LLM-Based Code Generation
por: He, Weilin, et al.
Publicado: (2026)
por: He, Weilin, et al.
Publicado: (2026)
Automated Proof Generation for Rust Code via Self-Evolution
por: Chen, Tianyu, et al.
Publicado: (2024)
por: Chen, Tianyu, et al.
Publicado: (2024)
SWE-Synth: Synthesizing Verifiable Bug-Fix Data to Enable Large Language Models in Resolving Real-World Bugs
por: Pham, Minh V. T., et al.
Publicado: (2025)
por: Pham, Minh V. T., et al.
Publicado: (2025)
AutoSafeCoder: A Multi-Agent Framework for Securing LLM Code Generation through Static Analysis and Fuzz Testing
por: Nunez, Ana, et al.
Publicado: (2024)
por: Nunez, Ana, et al.
Publicado: (2024)
DOMAINEVAL: An Auto-Constructed Benchmark for Multi-Domain Code Generation
por: Zhu, Qiming, et al.
Publicado: (2024)
por: Zhu, Qiming, et al.
Publicado: (2024)
AutoTest: Evolutionary Code Solution Selection with Test Cases
por: Duan, Zhihua, et al.
Publicado: (2024)
por: Duan, Zhihua, et al.
Publicado: (2024)
Impact-driven Context Filtering For Cross-file Code Completion
por: Li, Yanzhou, et al.
Publicado: (2025)
por: Li, Yanzhou, et al.
Publicado: (2025)
AutoFeedback: An LLM-based Framework for Efficient and Accurate API Request Generation
por: Liu, Huanxi, et al.
Publicado: (2024)
por: Liu, Huanxi, et al.
Publicado: (2024)
Uncovering LLM-Generated Code: A Zero-Shot Synthetic Code Detector via Code Rewriting
por: Ye, Tong, et al.
Publicado: (2024)
por: Ye, Tong, et al.
Publicado: (2024)
Verify-Gated Completion as Admission Control in a Governed Multi-Agent Runtime: A Bounded Architecture Case Study
por: Nguyen, Hai-Duong, et al.
Publicado: (2026)
por: Nguyen, Hai-Duong, et al.
Publicado: (2026)
Ejemplares similares
-
AutoWebWorld: Synthesizing Infinite Verifiable Web Environments via Finite State Machines
por: Wu, Yifan, et al.
Publicado: (2026) -
irace-evo: Automatic Algorithm Configuration Extended With LLM-Based Code Evolution
por: Sartori, Camilo Chacón, et al.
Publicado: (2025) -
Verifying LLM-Generated Code in the Context of Software Verification with Ada/SPARK
por: Cramer, Marcos, et al.
Publicado: (2025) -
CodeSift: An LLM-Based Reference-Less Framework for Automatic Code Validation
por: Aggarwal, Pooja, et al.
Publicado: (2024) -
AutoCodeBench: Large Language Models are Automatic Code Benchmark Generators
por: Chou, Jason, et al.
Publicado: (2025)