FM-Agent: Scaling Formal Methods to Large Systems via LLM-Based Hoare-Style Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Ding, Haoran, Wang, Zhaoguo, Chen, Haibo |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HoarePrompt: Structural Reasoning About Program Correctness in Natural Language
by: Bouras, Dimitrios Stamatios, et al.
Published: (2025)
by: Bouras, Dimitrios Stamatios, et al.
Published: (2025)
ComplexMCP: Evaluation of LLM Agents in Dynamic, Interdependent, and Large-Scale Tool Sandbox
by: Li, Yuanyang, et al.
Published: (2026)
by: Li, Yuanyang, et al.
Published: (2026)
MAS-FIRE: Fault Injection and Reliability Evaluation for LLM-Based Multi-Agent Systems
by: Jia, Jin, et al.
Published: (2026)
by: Jia, Jin, et al.
Published: (2026)
TransAgent: Enhancing LLM-Based Code Translation via Fine-Grained Execution Alignment
by: Yuan, Zhiqiang, et al.
Published: (2024)
by: Yuan, Zhiqiang, et al.
Published: (2024)
Trae Agent: An LLM-based Agent for Software Engineering with Test-time Scaling
by: Trae Research Team, et al.
Published: (2025)
by: Trae Research Team, et al.
Published: (2025)
Taint-Style Vulnerability Detection and Confirmation for Node.js Packages Using LLM Agent Reasoning
by: Ni, Ronghao, et al.
Published: (2026)
by: Ni, Ronghao, et al.
Published: (2026)
The Fusion of Large Language Models and Formal Methods for Trustworthy AI Agents: A Roadmap
by: Zhang, Yedi, et al.
Published: (2024)
by: Zhang, Yedi, et al.
Published: (2024)
A Large-Scale Study on the Development and Issues of Multi-Agent AI Systems
by: Liu, Daniel, et al.
Published: (2026)
by: Liu, Daniel, et al.
Published: (2026)
Efficient Failure Management for Multi-Agent Systems with Reasoning Trace Representation
by: Zhang, Lingzhe, et al.
Published: (2026)
by: Zhang, Lingzhe, et al.
Published: (2026)
Identifying Performance-Sensitive Configurations in Software Systems through Code Analysis with LLM Agents
by: Wang, Zehao, et al.
Published: (2024)
by: Wang, Zehao, et al.
Published: (2024)
Scaling Coding Agents via Atomic Skills
by: Ma, Yingwei, et al.
Published: (2026)
by: Ma, Yingwei, et al.
Published: (2026)
Rethinking the Value of Agent-Generated Tests for LLM-Based Software Engineering Agents
by: Chen, Zhi, et al.
Published: (2026)
by: Chen, Zhi, et al.
Published: (2026)
Large Language Model-Based Agents for Software Engineering: A Survey
by: Liu, Junwei, et al.
Published: (2024)
by: Liu, Junwei, et al.
Published: (2024)
Correct Code, Vulnerable Dependencies: A Large Scale Measurement Study of LLM-Specified Library Versions
by: Wang, Chengjie, et al.
Published: (2026)
by: Wang, Chengjie, et al.
Published: (2026)
Watson: A Cognitive Observability Framework for the Reasoning of LLM-Powered Agents
by: Rombaut, Benjamin, et al.
Published: (2024)
by: Rombaut, Benjamin, et al.
Published: (2024)
Beyond Functional Correctness: Investigating Coding Style Inconsistencies in Large Language Models
by: Wang, Yanlin, et al.
Published: (2024)
by: Wang, Yanlin, et al.
Published: (2024)
RepairAgent: An Autonomous, LLM-Based Agent for Program Repair
by: Bouzenia, Islem, et al.
Published: (2024)
by: Bouzenia, Islem, et al.
Published: (2024)
Ethics Testing: Proactive Identification of Generative AI System Harms
by: Tan, Shin Hwei, et al.
Published: (2026)
by: Tan, Shin Hwei, et al.
Published: (2026)
Designing Adaptive Digital Nudging Systems with LLM-Driven Reasoning
by: Santilli, Tiziano, et al.
Published: (2026)
by: Santilli, Tiziano, et al.
Published: (2026)
Formal Architecture Descriptors as Navigation Primitives for AI Coding Agents
by: Jin, Ruoqi
Published: (2026)
by: Jin, Ruoqi
Published: (2026)
An Empirical Evaluation of LLM-Based Approaches for Code Vulnerability Detection: RAG, SFT, and Dual-Agent Systems
by: Saju, Md Hasan, et al.
Published: (2026)
by: Saju, Md Hasan, et al.
Published: (2026)
From Flat Logs to Causal Graphs: Hierarchical Failure Attribution for LLM-based Multi-Agent Systems
by: Wang, Yawen, et al.
Published: (2026)
by: Wang, Yawen, et al.
Published: (2026)
MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution
by: Tao, Wei, et al.
Published: (2024)
by: Tao, Wei, et al.
Published: (2024)
S3LLM: Large-Scale Scientific Software Understanding with LLMs using Source, Metadata, and Document
by: Shaik, Kareem, et al.
Published: (2024)
by: Shaik, Kareem, et al.
Published: (2024)
Uncertainty Propagation in LLM-Based Systems
by: Xia, Boming, et al.
Published: (2026)
by: Xia, Boming, et al.
Published: (2026)
OrcaLoca: An LLM Agent Framework for Software Issue Localization
by: Yu, Zhongming, et al.
Published: (2025)
by: Yu, Zhongming, et al.
Published: (2025)
DomAgent: Leveraging Knowledge Graphs and Case-Based Reasoning for Domain-Specific Code Generation
by: Wang, Shuai, et al.
Published: (2026)
by: Wang, Shuai, et al.
Published: (2026)
Towards Automated Formal Verification of Backend Systems with LLMs
by: Xu, Kangping, et al.
Published: (2025)
by: Xu, Kangping, et al.
Published: (2025)
Beyond Local Code Optimization: Multi-Agent Reasoning for Software System Optimization
by: Peng, Huiyun, et al.
Published: (2026)
by: Peng, Huiyun, et al.
Published: (2026)
Multi-Agent Causal Reasoning System for Error Pattern Rule Automation in Vehicles
by: Math, Hugo, et al.
Published: (2026)
by: Math, Hugo, et al.
Published: (2026)
Automating Android Build Repair: Bridging the Reasoning-Execution Gap in LLM Agents with Domain-Specific Tools
by: Son, Ha Min, et al.
Published: (2025)
by: Son, Ha Min, et al.
Published: (2025)
ArgRE: Formal Argumentation for Conflict Resolution in Multi-Agent Requirements Negotiation
by: Cheng, Haowei, et al.
Published: (2026)
by: Cheng, Haowei, et al.
Published: (2026)
RGD: Multi-LLM Based Agent Debugger via Refinement and Generation Guidance
by: Jin, Haolin, et al.
Published: (2024)
by: Jin, Haolin, et al.
Published: (2024)
A Self-Healing Framework for Reliable LLM-Based Autonomous Agents
by: Jeong, Cheonsu, et al.
Published: (2026)
by: Jeong, Cheonsu, et al.
Published: (2026)
LLM Collaboration With Multi-Agent Reinforcement Learning
by: Liu, Shuo, et al.
Published: (2025)
by: Liu, Shuo, et al.
Published: (2025)
Thinking Longer, Not Larger: Enhancing Software Engineering Agents via Scaling Test-Time Compute
by: Ma, Yingwei, et al.
Published: (2025)
by: Ma, Yingwei, et al.
Published: (2025)
ShortcutsBench: A Large-Scale Real-world Benchmark for API-based Agents
by: Shen, Haiyang, et al.
Published: (2024)
by: Shen, Haiyang, et al.
Published: (2024)
DoVer: Intervention-Driven Auto Debugging for LLM Multi-Agent Systems
by: Ma, Ming, et al.
Published: (2025)
by: Ma, Ming, et al.
Published: (2025)
LLM-Rosetta: A Hub-and-Spoke Intermediate Representation for Cross-Provider LLM API Translation
by: Ding, Peng
Published: (2026)
by: Ding, Peng
Published: (2026)
Automated Repair of AI Code with Large Language Models and Formal Verification
by: Charalambous, Yiannis, et al.
Published: (2024)
by: Charalambous, Yiannis, et al.
Published: (2024)
Similar Items
-
HoarePrompt: Structural Reasoning About Program Correctness in Natural Language
by: Bouras, Dimitrios Stamatios, et al.
Published: (2025) -
ComplexMCP: Evaluation of LLM Agents in Dynamic, Interdependent, and Large-Scale Tool Sandbox
by: Li, Yuanyang, et al.
Published: (2026) -
MAS-FIRE: Fault Injection and Reliability Evaluation for LLM-Based Multi-Agent Systems
by: Jia, Jin, et al.
Published: (2026) -
TransAgent: Enhancing LLM-Based Code Translation via Fine-Grained Execution Alignment
by: Yuan, Zhiqiang, et al.
Published: (2024) -
Trae Agent: An LLM-based Agent for Software Engineering with Test-time Scaling
by: Trae Research Team, et al.
Published: (2025)