Can LLMs Enable Verification in Mainstream Programming?
Fuente:
arXiv
Saved in:
| Main Authors: | Shefer, Aleksandr, Engel, Igor, Alekseev, Stanislav, Berezun, Daniil, Verbitskaia, Ekaterina, Podkopaev, Anton |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Can LLMs Reason About Program Semantics? A Comprehensive Evaluation of LLMs on Formal Specification Inference
by: Le-Cong, Thanh, et al.
Published: (2025)
by: Le-Cong, Thanh, et al.
Published: (2025)
dafny-annotator: AI-Assisted Verification of Dafny Programs
by: Poesia, Gabriel, et al.
Published: (2024)
by: Poesia, Gabriel, et al.
Published: (2024)
Towards Repository-Level Program Verification with Large Language Models
by: Zhong, Si Cheng, et al.
Published: (2025)
by: Zhong, Si Cheng, et al.
Published: (2025)
Beyond Postconditions: Can Large Language Models infer Formal Contracts for Automatic Software Verification?
by: Richter, Cedric, et al.
Published: (2025)
by: Richter, Cedric, et al.
Published: (2025)
Ranking LLM-Generated Loop Invariants for Program Verification
by: Chakraborty, Saikat, et al.
Published: (2023)
by: Chakraborty, Saikat, et al.
Published: (2023)
LitmusKt: Concurrency Stress Testing for Kotlin
by: Lochmelis, Denis, et al.
Published: (2025)
by: Lochmelis, Denis, et al.
Published: (2025)
Structured Program Synthesis using LLMs: Results and Insights from the IPARC Challenge
by: Surana, Shraddha, et al.
Published: (2025)
by: Surana, Shraddha, et al.
Published: (2025)
LLMs Lean on Priors, Not Programming Language Semantics
by: Thimmaiah, Aditya, et al.
Published: (2025)
by: Thimmaiah, Aditya, et al.
Published: (2025)
AutoCode: LLMs as Problem Setters for Competitive Programming
by: Zhou, Shang, et al.
Published: (2025)
by: Zhou, Shang, et al.
Published: (2025)
Agentic Proving for Program Verification
by: Sosso, Alessandro, et al.
Published: (2026)
by: Sosso, Alessandro, et al.
Published: (2026)
A Problem-Oriented Perspective and Anchor Verification for Code Optimization
by: Ye, Tong, et al.
Published: (2024)
by: Ye, Tong, et al.
Published: (2024)
Program Skeletons for Automated Program Translation
by: Wang, Bo, et al.
Published: (2025)
by: Wang, Bo, et al.
Published: (2025)
LangGPT: Rethinking Structured Reusable Prompt Design Framework for LLMs from the Programming Language
by: Wang, Ming, et al.
Published: (2024)
by: Wang, Ming, et al.
Published: (2024)
SACTOR: LLM-Driven Correct and Idiomatic C to Rust Translation with Static Analysis and FFI-Based Verification
by: Zhou, Tianyang, et al.
Published: (2025)
by: Zhou, Tianyang, et al.
Published: (2025)
Assessing Code Understanding in LLMs
by: Laneve, Cosimo, et al.
Published: (2025)
by: Laneve, Cosimo, et al.
Published: (2025)
OSVBench: Benchmarking LLMs on Specification Generation Tasks for Operating System Verification
by: Li, Shangyu, et al.
Published: (2025)
by: Li, Shangyu, et al.
Published: (2025)
Is Programming by Example solved by LLMs?
by: Li, Wen-Ding, et al.
Published: (2024)
by: Li, Wen-Ding, et al.
Published: (2024)
Herb.jl: A Unifying Program Synthesis Library
by: Hinnerichs, Tilman, et al.
Published: (2025)
by: Hinnerichs, Tilman, et al.
Published: (2025)
Certified Program Synthesis with a Multi-Modal Verifier
by: Feng, Yueyang, et al.
Published: (2026)
by: Feng, Yueyang, et al.
Published: (2026)
Can Large Language Models Transform Natural Language Intent into Formal Method Postconditions?
by: Endres, Madeline, et al.
Published: (2023)
by: Endres, Madeline, et al.
Published: (2023)
Smaller = Weaker? Benchmarking Robustness of Quantized LLMs in Code Generation
by: Fang, Sen, et al.
Published: (2025)
by: Fang, Sen, et al.
Published: (2025)
Raw Pointer Rewriting with LLMs for Translating C to Safer Rust
by: Gao, Yifei, et al.
Published: (2025)
by: Gao, Yifei, et al.
Published: (2025)
The New Compiler Stack: A Survey on the Synergy of LLMs and Compilers
by: Zhang, Shuoming, et al.
Published: (2026)
by: Zhang, Shuoming, et al.
Published: (2026)
Towards Neural Synthesis for SMT-Assisted Proof-Oriented Programming
by: Chakraborty, Saikat, et al.
Published: (2024)
by: Chakraborty, Saikat, et al.
Published: (2024)
Agentic Interpretation: Lattice-Structured Evidence for LLM-Based Program Analysis
by: Mitchell, Jacqueline L., et al.
Published: (2026)
by: Mitchell, Jacqueline L., et al.
Published: (2026)
Natural Language-Oriented Programming (NLOP): Towards Democratizing Software Creation
by: Beheshti, Amin
Published: (2024)
by: Beheshti, Amin
Published: (2024)
Leveraging LLMs to support co-evolution between definitions and instances of textual DSLs
by: Zhang, Weixing, et al.
Published: (2025)
by: Zhang, Weixing, et al.
Published: (2025)
ECO: Enhanced Code Optimization via Performance-Aware Prompting for Code-LLMs
by: Kim, Su-Hyeon, et al.
Published: (2025)
by: Kim, Su-Hyeon, et al.
Published: (2025)
Reverse Chain: A Generic-Rule for LLMs to Master Multi-API Planning
by: Zhang, Yinger, et al.
Published: (2023)
by: Zhang, Yinger, et al.
Published: (2023)
Enhancing Automated Loop Invariant Generation for Complex Programs with Large Language Models
by: Liu, Ruibang, et al.
Published: (2024)
by: Liu, Ruibang, et al.
Published: (2024)
AbstractBeam: Enhancing Bottom-Up Program Synthesis using Library Learning
by: Zenkner, Janis, et al.
Published: (2024)
by: Zenkner, Janis, et al.
Published: (2024)
Learning to Guarantee Type Correctness in Code Generation through Type-Guided Program Synthesis
by: Huang, Zhechong, et al.
Published: (2025)
by: Huang, Zhechong, et al.
Published: (2025)
Static Program Slicing Using Language Models With Dataflow-Aware Pretraining and Constrained Decoding
by: He, Pengfei, et al.
Published: (2026)
by: He, Pengfei, et al.
Published: (2026)
AI Coders Are Among Us: Rethinking Programming Language Grammar Towards Efficient Code Generation
by: Sun, Zhensu, et al.
Published: (2024)
by: Sun, Zhensu, et al.
Published: (2024)
Perish or Flourish? A Holistic Evaluation of Large Language Models for Code Generation in Functional Programming
by: Lang, Nguyet-Anh H., et al.
Published: (2026)
by: Lang, Nguyet-Anh H., et al.
Published: (2026)
Evaluating the Performance of Large Language Models in Competitive Programming: A Multi-Year, Multi-Grade Analysis
by: Dumitran, Adrian Marius, et al.
Published: (2024)
by: Dumitran, Adrian Marius, et al.
Published: (2024)
Doc2Spec: Synthesizing Formal Programming Specifications from Natural Language via Grammar Induction
by: Xia, Shihao, et al.
Published: (2026)
by: Xia, Shihao, et al.
Published: (2026)
Typed Embedding of miniKanren for Functional Conversion
by: Engel, Igor, et al.
Published: (2025)
by: Engel, Igor, et al.
Published: (2025)
Agentic Program Repair from Test Failures at Scale: A Neuro-symbolic approach with static analysis and test execution feedback
by: Maddila, Chandra, et al.
Published: (2025)
by: Maddila, Chandra, et al.
Published: (2025)
Code Repair with LLMs gives an Exploration-Exploitation Tradeoff
by: Tang, Hao, et al.
Published: (2024)
by: Tang, Hao, et al.
Published: (2024)
Similar Items
-
Can LLMs Reason About Program Semantics? A Comprehensive Evaluation of LLMs on Formal Specification Inference
by: Le-Cong, Thanh, et al.
Published: (2025) -
dafny-annotator: AI-Assisted Verification of Dafny Programs
by: Poesia, Gabriel, et al.
Published: (2024) -
Towards Repository-Level Program Verification with Large Language Models
by: Zhong, Si Cheng, et al.
Published: (2025) -
Beyond Postconditions: Can Large Language Models infer Formal Contracts for Automatic Software Verification?
by: Richter, Cedric, et al.
Published: (2025) -
Ranking LLM-Generated Loop Invariants for Program Verification
by: Chakraborty, Saikat, et al.
Published: (2023)