Towards Automated Formal Verification of Backend Systems with LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Kangping, Luo, Yifan, Yuan, Yang, Yao, Andrew Chi-Chih |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Automated Repair of AI Code with Large Language Models and Formal Verification
by: Charalambous, Yiannis, et al.
Published: (2024)
by: Charalambous, Yiannis, et al.
Published: (2024)
Accessible Smart Contracts Verification: Synthesizing Formal Models with Tamed LLMs
by: Corazza, Jan, et al.
Published: (2025)
by: Corazza, Jan, et al.
Published: (2025)
Constraint Decay: The Fragility of LLM Agents in Backend Code Generation
by: Dente, Francesco, et al.
Published: (2026)
by: Dente, Francesco, et al.
Published: (2026)
From Natural Language to Verified Code: Toward AI Assisted Problem-to-Code Generation with Dafny-Based Formal Verification
by: Erfan, Md, et al.
Published: (2026)
by: Erfan, Md, et al.
Published: (2026)
On the Effectiveness of LLMs for Manual Test Verifications
by: Peixoto, Myron David Lucena Campos, et al.
Published: (2024)
by: Peixoto, Myron David Lucena Campos, et al.
Published: (2024)
PropertyGPT: LLM-driven Formal Verification of Smart Contracts through Retrieval-Augmented Property Generation
by: Liu, Ye, et al.
Published: (2024)
by: Liu, Ye, et al.
Published: (2024)
RefEvo: Agentic Design with Co-Evolutionary Verification for Agile Reference Model Generation
by: Zhang, Yifan, et al.
Published: (2026)
by: Zhang, Yifan, et al.
Published: (2026)
ConVer: Using Contracts and Loop Invariant Synthesis for Scalable Formal Software Verification
by: Pirzada, Muhammad A. A., et al.
Published: (2026)
by: Pirzada, Muhammad A. A., et al.
Published: (2026)
Goedel-Code-Prover: Hierarchical Proof Search for Open State-of-the-Art Code Verification
by: Li, Zenan, et al.
Published: (2026)
by: Li, Zenan, et al.
Published: (2026)
UnitTenX: Generating Tests for Legacy Packages with AI Agents Powered by Formal Verification
by: Charalambous, Yiannis, et al.
Published: (2025)
by: Charalambous, Yiannis, et al.
Published: (2025)
ABC-Bench: Benchmarking Agentic Backend Coding in Real-World Development
by: Yang, Jie, et al.
Published: (2026)
by: Yang, Jie, et al.
Published: (2026)
Vulnerability Detection: From Formal Verification to Large Language Models and Hybrid Approaches: A Comprehensive Overview
by: Tihanyi, Norbert, et al.
Published: (2025)
by: Tihanyi, Norbert, et al.
Published: (2025)
Verification and Validation of Autonomous Systems
by: Shetiya, Sneha Sudhir, et al.
Published: (2024)
by: Shetiya, Sneha Sudhir, et al.
Published: (2024)
Can LLMs Reason Like Automated Theorem Provers for Rust Verification? VCoT-Bench: Evaluating via Verification Chain of Thought
by: Xie, Zichen, et al.
Published: (2026)
by: Xie, Zichen, et al.
Published: (2026)
Utilizing LLMs for Industrial Process Automation
by: Fares, Salim
Published: (2026)
by: Fares, Salim
Published: (2026)
VeriStruct: AI-assisted Automated Verification of Data-Structure Modules in Verus
by: Sun, Chuyue, et al.
Published: (2025)
by: Sun, Chuyue, et al.
Published: (2025)
Chain of Targeted Verification Questions to Improve the Reliability of Code Generated by LLMs
by: Ngassom, Sylvain Kouemo, et al.
Published: (2024)
by: Ngassom, Sylvain Kouemo, et al.
Published: (2024)
EmbedGenius: Towards Automated Software Development for Generic Embedded IoT Systems
by: Yang, Huanqi, et al.
Published: (2024)
by: Yang, Huanqi, et al.
Published: (2024)
Towards Comprehensive Benchmarking Infrastructure for LLMs In Software Engineering
by: Rodriguez-Cardenas, Daniel, et al.
Published: (2026)
by: Rodriguez-Cardenas, Daniel, et al.
Published: (2026)
Skill Discovery for Software Scripting Automation via Offline Simulations with LLMs
by: Xu, Paiheng, et al.
Published: (2025)
by: Xu, Paiheng, et al.
Published: (2025)
Evaluating the Generalizability of LLMs in Automated Program Repair
by: Li, Fengjie, et al.
Published: (2025)
by: Li, Fengjie, et al.
Published: (2025)
Enhancing Large Language Model Efficiencyvia Symbolic Compression: A Formal Approach Towards Interpretability
by: AI, Lumen, et al.
Published: (2025)
by: AI, Lumen, et al.
Published: (2025)
Verification-Guided Context Optimization for Tool Calling via Hierarchical LLMs-as-Editors
by: Li, Henger, et al.
Published: (2025)
by: Li, Henger, et al.
Published: (2025)
Harnessing the Power of LLMs: Automating Unit Test Generation for High-Performance Computing
by: Karanjai, Rabimba, et al.
Published: (2024)
by: Karanjai, Rabimba, et al.
Published: (2024)
Towards Secure Logging: Characterizing and Benchmarking Logging Code Security Issues with LLMs
by: Yuan, He Yang, et al.
Published: (2026)
by: Yuan, He Yang, et al.
Published: (2026)
Vibe-Coding: Feedback-Based Automated Verification with no Human Code Inspection, a Feasibility Study
by: Töpfer, Michal, et al.
Published: (2026)
by: Töpfer, Michal, et al.
Published: (2026)
RAG-Verus: Repository-Level Program Verification with LLMs using Retrieval Augmented Generation
by: Zhong, Sicheng, et al.
Published: (2025)
by: Zhong, Sicheng, et al.
Published: (2025)
Cybernaut: Towards Reliable Web Automation
by: Tomar, Ankur, et al.
Published: (2025)
by: Tomar, Ankur, et al.
Published: (2025)
Reflective Paper-to-Code Reproduction Enabled by Fine-Grained Verification
by: Zhou, Mingyang, et al.
Published: (2025)
by: Zhou, Mingyang, et al.
Published: (2025)
TAM-Eval: Evaluating LLMs for Automated Unit Test Maintenance
by: Bruches, Elena, et al.
Published: (2026)
by: Bruches, Elena, et al.
Published: (2026)
The Potential of LLMs in Automating Software Testing: From Generation to Reporting
by: Sherifi, Betim, et al.
Published: (2024)
by: Sherifi, Betim, et al.
Published: (2024)
Early-Stage Product Line Validation Using LLMs: A Study on Semi-Formal Blueprint Analysis
by: Le, Viet-Man, et al.
Published: (2026)
by: Le, Viet-Man, et al.
Published: (2026)
Explicating Tacit Regulatory Knowledge from LLMs to Auto-Formalize Requirements for Compliance Test Case Generation
by: Xue, Zhiyi, et al.
Published: (2026)
by: Xue, Zhiyi, et al.
Published: (2026)
MR-Coupler: Automated Metamorphic Test Generation via Functional Coupling Analysis
by: Xu, Congying, et al.
Published: (2026)
by: Xu, Congying, et al.
Published: (2026)
CONSTRUCTA: Automating Commercial Construction Schedules in Fabrication Facilities with Large Language Models
by: Zhang, Yifan, et al.
Published: (2025)
by: Zhang, Yifan, et al.
Published: (2025)
SPARC: Scenario Planning and Reasoning for Automated C Unit Test Generation
by: Chowdhury, Jaid Monwar, et al.
Published: (2026)
by: Chowdhury, Jaid Monwar, et al.
Published: (2026)
LogReasoner: Empowering LLMs with Expert-like Coarse-to-Fine Reasoning for Automated Log Analysis
by: Ma, Lipeng, et al.
Published: (2025)
by: Ma, Lipeng, et al.
Published: (2025)
Leveraging LLMs, IDEs, and Semantic Embeddings for Automated Move Method Refactoring
by: Bellur, Abhiram, et al.
Published: (2025)
by: Bellur, Abhiram, et al.
Published: (2025)
RedCoder: Automated Multi-Turn Red Teaming for Code LLMs
by: Mo, Wenjie Jacky, et al.
Published: (2025)
by: Mo, Wenjie Jacky, et al.
Published: (2025)
Fine-grained Approaches for Confidence Calibration of LLMs in Automated Code Revision
by: Lin, Hong Yi, et al.
Published: (2026)
by: Lin, Hong Yi, et al.
Published: (2026)
Similar Items
-
Automated Repair of AI Code with Large Language Models and Formal Verification
by: Charalambous, Yiannis, et al.
Published: (2024) -
Accessible Smart Contracts Verification: Synthesizing Formal Models with Tamed LLMs
by: Corazza, Jan, et al.
Published: (2025) -
Constraint Decay: The Fragility of LLM Agents in Backend Code Generation
by: Dente, Francesco, et al.
Published: (2026) -
From Natural Language to Verified Code: Toward AI Assisted Problem-to-Code Generation with Dafny-Based Formal Verification
by: Erfan, Md, et al.
Published: (2026) -
On the Effectiveness of LLMs for Manual Test Verifications
by: Peixoto, Myron David Lucena Campos, et al.
Published: (2024)