Towards Repository-Level Program Verification with Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Zhong, Si Cheng, Si, Xujie |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RAG-Verus: Repository-Level Program Verification with LLMs using Retrieval Augmented Generation
by: Zhong, Sicheng, et al.
Published: (2025)
by: Zhong, Sicheng, et al.
Published: (2025)
TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories
by: Dong, Honghua, et al.
Published: (2025)
by: Dong, Honghua, et al.
Published: (2025)
Code Repair with LLMs gives an Exploration-Exploitation Tradeoff
by: Tang, Hao, et al.
Published: (2024)
by: Tang, Hao, et al.
Published: (2024)
Beyond Postconditions: Can Large Language Models infer Formal Contracts for Automatic Software Verification?
by: Richter, Cedric, et al.
Published: (2025)
by: Richter, Cedric, et al.
Published: (2025)
Can LLMs Enable Verification in Mainstream Programming?
by: Shefer, Aleksandr, et al.
Published: (2025)
by: Shefer, Aleksandr, et al.
Published: (2025)
dafny-annotator: AI-Assisted Verification of Dafny Programs
by: Poesia, Gabriel, et al.
Published: (2024)
by: Poesia, Gabriel, et al.
Published: (2024)
Enhancing Automated Loop Invariant Generation for Complex Programs with Large Language Models
by: Liu, Ruibang, et al.
Published: (2024)
by: Liu, Ruibang, et al.
Published: (2024)
Natural Language-Oriented Programming (NLOP): Towards Democratizing Software Creation
by: Beheshti, Amin
Published: (2024)
by: Beheshti, Amin
Published: (2024)
Perish or Flourish? A Holistic Evaluation of Large Language Models for Code Generation in Functional Programming
by: Lang, Nguyet-Anh H., et al.
Published: (2026)
by: Lang, Nguyet-Anh H., et al.
Published: (2026)
Large Language Models to Generate System-Level Test Programs Targeting Non-functional Properties
by: Schwachhofer, Denis, et al.
Published: (2024)
by: Schwachhofer, Denis, et al.
Published: (2024)
Evaluating the Performance of Large Language Models in Competitive Programming: A Multi-Year, Multi-Grade Analysis
by: Dumitran, Adrian Marius, et al.
Published: (2024)
by: Dumitran, Adrian Marius, et al.
Published: (2024)
Refactoring Programs Using Large Language Models with Few-Shot Examples
by: Shirafuji, Atsushi, et al.
Published: (2023)
by: Shirafuji, Atsushi, et al.
Published: (2023)
Ranking LLM-Generated Loop Invariants for Program Verification
by: Chakraborty, Saikat, et al.
Published: (2023)
by: Chakraborty, Saikat, et al.
Published: (2023)
AInsteinBench: Benchmarking Coding Agents on Scientific Repositories
by: Duston, Titouan, et al.
Published: (2025)
by: Duston, Titouan, et al.
Published: (2025)
AI Coders Are Among Us: Rethinking Programming Language Grammar Towards Efficient Code Generation
by: Sun, Zhensu, et al.
Published: (2024)
by: Sun, Zhensu, et al.
Published: (2024)
SuperCoder: Assembly Program Superoptimization with Large Language Models
by: Wei, Anjiang, et al.
Published: (2025)
by: Wei, Anjiang, et al.
Published: (2025)
Towards Neural Synthesis for SMT-Assisted Proof-Oriented Programming
by: Chakraborty, Saikat, et al.
Published: (2024)
by: Chakraborty, Saikat, et al.
Published: (2024)
Assessing the Interpretability of Programmatic Policies with Large Language Models
by: Bashir, Zahra, et al.
Published: (2023)
by: Bashir, Zahra, et al.
Published: (2023)
Static Program Slicing Using Language Models With Dataflow-Aware Pretraining and Constrained Decoding
by: He, Pengfei, et al.
Published: (2026)
by: He, Pengfei, et al.
Published: (2026)
Can Large Language Models Transform Natural Language Intent into Formal Method Postconditions?
by: Endres, Madeline, et al.
Published: (2023)
by: Endres, Madeline, et al.
Published: (2023)
Solving Data-centric Tasks using Large Language Models
by: Barke, Shraddha, et al.
Published: (2024)
by: Barke, Shraddha, et al.
Published: (2024)
ClassInvGen: Class Invariant Synthesis using Large Language Models
by: Sun, Chuyue, et al.
Published: (2025)
by: Sun, Chuyue, et al.
Published: (2025)
Agentic Proving for Program Verification
by: Sosso, Alessandro, et al.
Published: (2026)
by: Sosso, Alessandro, et al.
Published: (2026)
Analysis of AdvFusion: Adapter-based Multilingual Learning for Code Large Language Models
by: Esmaeili, Amirreza, et al.
Published: (2025)
by: Esmaeili, Amirreza, et al.
Published: (2025)
SAT-DIFF: A Tree Diffing Framework Using SAT Solving
by: Geng, Chuqin, et al.
Published: (2024)
by: Geng, Chuqin, et al.
Published: (2024)
CodeIF-Bench: Evaluating Instruction-Following Capabilities of Large Language Models in Interactive Code Generation
by: Wang, Peiding, et al.
Published: (2025)
by: Wang, Peiding, et al.
Published: (2025)
A Problem-Oriented Perspective and Anchor Verification for Code Optimization
by: Ye, Tong, et al.
Published: (2024)
by: Ye, Tong, et al.
Published: (2024)
Benchmarking Large Language Models for ABAP Code Generation: An Empirical Study on Iterative Improvement by Compiler Feedback
by: Wallraven, Stephan, et al.
Published: (2026)
by: Wallraven, Stephan, et al.
Published: (2026)
Program Skeletons for Automated Program Translation
by: Wang, Bo, et al.
Published: (2025)
by: Wang, Bo, et al.
Published: (2025)
LLMs Lean on Priors, Not Programming Language Semantics
by: Thimmaiah, Aditya, et al.
Published: (2025)
by: Thimmaiah, Aditya, et al.
Published: (2025)
CodeMind: Evaluating Large Language Models for Code Reasoning
by: Liu, Changshu, et al.
Published: (2024)
by: Liu, Changshu, et al.
Published: (2024)
Doc2Spec: Synthesizing Formal Programming Specifications from Natural Language via Grammar Induction
by: Xia, Shihao, et al.
Published: (2026)
by: Xia, Shihao, et al.
Published: (2026)
SACTOR: LLM-Driven Correct and Idiomatic C to Rust Translation with Static Analysis and FFI-Based Verification
by: Zhou, Tianyang, et al.
Published: (2025)
by: Zhou, Tianyang, et al.
Published: (2025)
Benchmarking LLM Code Generation for Audio Programming with Visual Dataflow Languages
by: Zhang, William, et al.
Published: (2024)
by: Zhang, William, et al.
Published: (2024)
EquiBench: Benchmarking Large Language Models' Reasoning about Program Semantics via Equivalence Checking
by: Wei, Anjiang, et al.
Published: (2025)
by: Wei, Anjiang, et al.
Published: (2025)
ViScratch: Using Large Language Models and Gameplay Videos for Automated Feedback in Scratch
by: Si, Yuan, et al.
Published: (2025)
by: Si, Yuan, et al.
Published: (2025)
Herb.jl: A Unifying Program Synthesis Library
by: Hinnerichs, Tilman, et al.
Published: (2025)
by: Hinnerichs, Tilman, et al.
Published: (2025)
Certified Program Synthesis with a Multi-Modal Verifier
by: Feng, Yueyang, et al.
Published: (2026)
by: Feng, Yueyang, et al.
Published: (2026)
Large Language Models for Code Summarization
by: Szalontai, Balázs, et al.
Published: (2024)
by: Szalontai, Balázs, et al.
Published: (2024)
CodexGraph: Bridging Large Language Models and Code Repositories via Code Graph Databases
by: Liu, Xiangyan, et al.
Published: (2024)
by: Liu, Xiangyan, et al.
Published: (2024)
Similar Items
-
RAG-Verus: Repository-Level Program Verification with LLMs using Retrieval Augmented Generation
by: Zhong, Sicheng, et al.
Published: (2025) -
TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories
by: Dong, Honghua, et al.
Published: (2025) -
Code Repair with LLMs gives an Exploration-Exploitation Tradeoff
by: Tang, Hao, et al.
Published: (2024) -
Beyond Postconditions: Can Large Language Models infer Formal Contracts for Automatic Software Verification?
by: Richter, Cedric, et al.
Published: (2025) -
Can LLMs Enable Verification in Mainstream Programming?
by: Shefer, Aleksandr, et al.
Published: (2025)