CoqPilot, a plugin for LLM-based generation of proofs
Fuente:
arXiv
Saved in:
| Main Authors: | Kozyrev, Andrei, Solovev, Gleb, Khramov, Nikita, Podkopaev, Anton |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RocqStar: Leveraging Similarity-driven Retrieval and Agentic Systems for Rocq generation
by: Kozyrev, Andrei, et al.
Published: (2025)
by: Kozyrev, Andrei, et al.
Published: (2025)
RocqSmith: Can Automatic Optimization Forge Better Proof Agents?
by: Kozyrev, Andrei, et al.
Published: (2026)
by: Kozyrev, Andrei, et al.
Published: (2026)
Neuro-Symbolic Verification of LLM Outputs for Data-Sensitive Domains (extended preprint)
by: Sigloch, Paul, et al.
Published: (2026)
by: Sigloch, Paul, et al.
Published: (2026)
Towards Automatic Transformations of Coq Proof Scripts
by: Magaud, Nicolas
Published: (2024)
by: Magaud, Nicolas
Published: (2024)
Declarative Scenario-based Testing with RoadLogic
by: Bartocci, Ezio, et al.
Published: (2026)
by: Bartocci, Ezio, et al.
Published: (2026)
MacroSwarm: A Field-based Compositional Framework for Swarm Programming
by: Aguzzi, Gianluca, et al.
Published: (2024)
by: Aguzzi, Gianluca, et al.
Published: (2024)
Lawful and Accountable Personal Data Processing with GDPR-based Access and Usage Control in Distributed Systems
by: van Binsbergen, L. Thomas, et al.
Published: (2025)
by: van Binsbergen, L. Thomas, et al.
Published: (2025)
Layered and Staged Monte Carlo Tree Search for SMT Strategy Synthesis
by: Lu, Zhengyang, et al.
Published: (2024)
by: Lu, Zhengyang, et al.
Published: (2024)
Can Language Models Pretend Solvers? Logic Code Simulation with LLMs
by: Chen, Minyu, et al.
Published: (2024)
by: Chen, Minyu, et al.
Published: (2024)
CFaults: Model-Based Diagnosis for Fault Localization in C Programs with Multiple Test Cases
by: Orvalho, Pedro, et al.
Published: (2024)
by: Orvalho, Pedro, et al.
Published: (2024)
StatWhy: Formal Verification Tool for Statistical Hypothesis Testing Programs
by: Kawamoto, Yusuke, et al.
Published: (2024)
by: Kawamoto, Yusuke, et al.
Published: (2024)
Predictable and Performant Reactive Synthesis Modulo Theories via Functional Synthesis
by: Rodríguez, Andoni, et al.
Published: (2024)
by: Rodríguez, Andoni, et al.
Published: (2024)
Watchdogs and Oracles: Runtime Verification Meets Large Language Models for Autonomous Systems
by: Ferrando, Angelo
Published: (2025)
by: Ferrando, Angelo
Published: (2025)
LLMs and Fuzzing in Tandem: A New Approach to Automatically Generating Weakest Preconditions
by: King, Daragh, et al.
Published: (2025)
by: King, Daragh, et al.
Published: (2025)
LTLGuard: Formalizing LTL Specifications with Compact Language Models and Lightweight Symbolic Reasoning
by: Andresel, Medina, et al.
Published: (2026)
by: Andresel, Medina, et al.
Published: (2026)
Pseudo-Boolean d-DNNF Compilation for Expressive Feature Modeling Constructs
by: Sundermann, Chico, et al.
Published: (2025)
by: Sundermann, Chico, et al.
Published: (2025)
Accelerating Policy Synthesis in Large-Scale MDPs via Hierarchical Adaptive Refinement
by: Evangelidis, Alexandros, et al.
Published: (2025)
by: Evangelidis, Alexandros, et al.
Published: (2025)
Next Steps in LLM-Supported Java Verification
by: Teuber, Samuel, et al.
Published: (2025)
by: Teuber, Samuel, et al.
Published: (2025)
LogicAsker: Evaluating and Improving the Logical Reasoning Ability of Large Language Models
by: Wan, Yuxuan, et al.
Published: (2024)
by: Wan, Yuxuan, et al.
Published: (2024)
Grammars of Formal Uncertainty: When to Trust LLMs in Automated Reasoning Tasks
by: Ganguly, Debargha, et al.
Published: (2025)
by: Ganguly, Debargha, et al.
Published: (2025)
Model-Based Diagnosis with Multiple Observations: A Unified Approach for C Software and Boolean Circuits
by: Orvalho, Pedro, et al.
Published: (2025)
by: Orvalho, Pedro, et al.
Published: (2025)
Evaluating Implicit Regulatory Compliance in LLM Tool Invocation via Logic-Guided Synthesis
by: Song, Da, et al.
Published: (2026)
by: Song, Da, et al.
Published: (2026)
Evaluating the Ability of Large Language Models to Generate Verifiable Specifications in VeriFast
by: Fan, Wen, et al.
Published: (2024)
by: Fan, Wen, et al.
Published: (2024)
Agentic Proving for Program Verification
by: Sosso, Alessandro, et al.
Published: (2026)
by: Sosso, Alessandro, et al.
Published: (2026)
Proving the Coding Interview: A Benchmark for Formally Verified Code Generation
by: Dougherty, Quinn, et al.
Published: (2025)
by: Dougherty, Quinn, et al.
Published: (2025)
Inferring multiple helper Dafny assertions with LLMs
by: Silva, Álvaro, et al.
Published: (2025)
by: Silva, Álvaro, et al.
Published: (2025)
MPBMC: Multi-Property Bounded Model Checking with GNN-guided Clustering
by: Roy, Soumik Guha, et al.
Published: (2026)
by: Roy, Soumik Guha, et al.
Published: (2026)
Viverra: Text-to-Code with Guarantees
by: Wu, Haoze, et al.
Published: (2026)
by: Wu, Haoze, et al.
Published: (2026)
Runtime Monitoring and Enforcement of Conditional Fairness in Generative AIs
by: Cheng, Chih-Hong, et al.
Published: (2024)
by: Cheng, Chih-Hong, et al.
Published: (2024)
Continuous reasoning for adaptive container image distribution in the cloud-edge continuum
by: Azzolini, Damiano, et al.
Published: (2024)
by: Azzolini, Damiano, et al.
Published: (2024)
Dafny as Verification-Aware Intermediate Language for Code Generation
by: Li, Yue Chen, et al.
Published: (2025)
by: Li, Yue Chen, et al.
Published: (2025)
Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search
by: Lu, Jialin, et al.
Published: (2026)
by: Lu, Jialin, et al.
Published: (2026)
VerMCTS: Synthesizing Multi-Step Programs using a Verifier, a Large Language Model, and Tree Search
by: Brandfonbrener, David, et al.
Published: (2024)
by: Brandfonbrener, David, et al.
Published: (2024)
Portus: Linking Alloy with SMT-based Finite Model Finding
by: Dancy, Ryan, et al.
Published: (2024)
by: Dancy, Ryan, et al.
Published: (2024)
Verifying DNN-based Semantic Communication Against Generative Adversarial Noise
by: Le, Thanh, et al.
Published: (2026)
by: Le, Thanh, et al.
Published: (2026)
Taming Silent Failures: A Framework for Verifiable AI Reliability
by: Yang, Guan-Yan, et al.
Published: (2025)
by: Yang, Guan-Yan, et al.
Published: (2025)
VERINA: Benchmarking Verifiable Code Generation
by: Ye, Zhe, et al.
Published: (2025)
by: Ye, Zhe, et al.
Published: (2025)
CLEVER: A Curated Benchmark for Formally Verified Code Generation
by: Thakur, Amitayush, et al.
Published: (2025)
by: Thakur, Amitayush, et al.
Published: (2025)
Intent-aligned Formal Specification Synthesis via Traceable Refinement
by: Ye, Zhe, et al.
Published: (2026)
by: Ye, Zhe, et al.
Published: (2026)
Proceedings 9th edition of Working Formal Methods Symposium
by: Arusoaie, Andrei, et al.
Published: (2025)
by: Arusoaie, Andrei, et al.
Published: (2025)
Similar Items
-
RocqStar: Leveraging Similarity-driven Retrieval and Agentic Systems for Rocq generation
by: Kozyrev, Andrei, et al.
Published: (2025) -
RocqSmith: Can Automatic Optimization Forge Better Proof Agents?
by: Kozyrev, Andrei, et al.
Published: (2026) -
Neuro-Symbolic Verification of LLM Outputs for Data-Sensitive Domains (extended preprint)
by: Sigloch, Paul, et al.
Published: (2026) -
Towards Automatic Transformations of Coq Proof Scripts
by: Magaud, Nicolas
Published: (2024) -
Declarative Scenario-based Testing with RoadLogic
by: Bartocci, Ezio, et al.
Published: (2026)