VeriAct: Beyond Verifiability -- Agentic Synthesis of Correct and Complete Formal Specifications
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Misu, Md Rakib Hossain, Ma, Iris, Lopes, Cristina V. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards AI-Assisted Synthesis of Verified Dafny Methods
von: Misu, Md Rakib Hossain, et al.
Veröffentlicht: (2024)
von: Misu, Md Rakib Hossain, et al.
Veröffentlicht: (2024)
Toward Automated Validation of Language Model Synthesized Test Cases using Semantic Entropy
von: Taherkhani, Hamed, et al.
Veröffentlicht: (2024)
von: Taherkhani, Hamed, et al.
Veröffentlicht: (2024)
From Natural Language to Verified Code: Toward AI Assisted Problem-to-Code Generation with Dafny-Based Formal Verification
von: Erfan, Md, et al.
Veröffentlicht: (2026)
von: Erfan, Md, et al.
Veröffentlicht: (2026)
DafnyBench: A Benchmark for Formal Software Verification
von: Loughridge, Chloe, et al.
Veröffentlicht: (2024)
von: Loughridge, Chloe, et al.
Veröffentlicht: (2024)
Automated Proof Generation for Rust Code via Self-Evolution
von: Chen, Tianyu, et al.
Veröffentlicht: (2024)
von: Chen, Tianyu, et al.
Veröffentlicht: (2024)
AlgoVeri: An Aligned Benchmark for Verified Code Generation on Classical Algorithms
von: Zhao, Haoyu, et al.
Veröffentlicht: (2026)
von: Zhao, Haoyu, et al.
Veröffentlicht: (2026)
Enhancing Formal Software Specification with Artificial Intelligence
von: Nassar, Antonio Abu, et al.
Veröffentlicht: (2026)
von: Nassar, Antonio Abu, et al.
Veröffentlicht: (2026)
Beyond the 'Diff': Addressing Agentic Entropy in Agentic Software Development
von: Casserini, Matteo, et al.
Veröffentlicht: (2026)
von: Casserini, Matteo, et al.
Veröffentlicht: (2026)
Beyond Task Completion: An Assessment Framework for Evaluating Agentic AI Systems
von: Akshathala, Sreemaee, et al.
Veröffentlicht: (2025)
von: Akshathala, Sreemaee, et al.
Veröffentlicht: (2025)
Evaluating the Ability of Large Language Models to Generate Verifiable Specifications in VeriFast
von: Fan, Wen, et al.
Veröffentlicht: (2024)
von: Fan, Wen, et al.
Veröffentlicht: (2024)
BabelCoder: Agentic Code Translation with Specification Alignment
von: Rabbi, Fazle, et al.
Veröffentlicht: (2025)
von: Rabbi, Fazle, et al.
Veröffentlicht: (2025)
VeriScale: Adversarial Test-Suite Scaling for Verifiable Code Generation
von: Bai, Yifan, et al.
Veröffentlicht: (2026)
von: Bai, Yifan, et al.
Veröffentlicht: (2026)
VeriContest: A Competitive-Programming Benchmark for Verifiable Code Generation
von: Xie, Zichen, et al.
Veröffentlicht: (2026)
von: Xie, Zichen, et al.
Veröffentlicht: (2026)
Veri-Sure: A Contract-Aware Multi-Agent Framework with Temporal Tracing and Formal Verification for Correct RTL Code Generation
von: Liu, Jiale, et al.
Veröffentlicht: (2026)
von: Liu, Jiale, et al.
Veröffentlicht: (2026)
Uncovering Systematic Failures of LLMs in Verifying Code Against Natural Language Specifications
von: Jin, Haolin, et al.
Veröffentlicht: (2025)
von: Jin, Haolin, et al.
Veröffentlicht: (2025)
Neuro-Symbolic Generation and Validation of Memory-Aware Formal Function Specifications
von: Zhang, Liao, et al.
Veröffentlicht: (2026)
von: Zhang, Liao, et al.
Veröffentlicht: (2026)
Inferring Code Correctness from Specification
von: Florian, Tambon, et al.
Veröffentlicht: (2026)
von: Florian, Tambon, et al.
Veröffentlicht: (2026)
LiveFMBench: Unveiling the Power and Limits of Agentic Workflows in Specification Generation
von: Xu, Dong, et al.
Veröffentlicht: (2026)
von: Xu, Dong, et al.
Veröffentlicht: (2026)
BLAgent: Agentic RAG for File-Level Bug Localization
von: Mamun, Md Afif Al, et al.
Veröffentlicht: (2026)
von: Mamun, Md Afif Al, et al.
Veröffentlicht: (2026)
Test Smell: A Parasitic Energy Consumer in Software Testing
von: Misu, Md Rakib Hossain, et al.
Veröffentlicht: (2023)
von: Misu, Md Rakib Hossain, et al.
Veröffentlicht: (2023)
Beyond Functional Correctness: Exploring Hallucinations in LLM-Generated Code
von: Liu, Fang, et al.
Veröffentlicht: (2024)
von: Liu, Fang, et al.
Veröffentlicht: (2024)
WebVIA: A Web-based Vision-Language Agentic Framework for Interactive and Verifiable UI-to-Code Generation
von: Xu, Mingde, et al.
Veröffentlicht: (2025)
von: Xu, Mingde, et al.
Veröffentlicht: (2025)
Verify-Gated Completion as Admission Control in a Governed Multi-Agent Runtime: A Bounded Architecture Case Study
von: Nguyen, Hai-Duong, et al.
Veröffentlicht: (2026)
von: Nguyen, Hai-Duong, et al.
Veröffentlicht: (2026)
Artificial Intelligence in Open Source Software Engineering: A Foundation for Sustainability
von: Karim, S M Rakib UI, et al.
Veröffentlicht: (2026)
von: Karim, S M Rakib UI, et al.
Veröffentlicht: (2026)
Beyond Functional Correctness: Investigating Coding Style Inconsistencies in Large Language Models
von: Wang, Yanlin, et al.
Veröffentlicht: (2024)
von: Wang, Yanlin, et al.
Veröffentlicht: (2024)
ARCS: Agentic Retrieval-Augmented Code Synthesis with Iterative Refinement
von: Bhattarai, Manish, et al.
Veröffentlicht: (2025)
von: Bhattarai, Manish, et al.
Veröffentlicht: (2025)
Beyond Human-Readable: Rethinking Software Engineering Conventions for the Agentic Development Era
von: Ustynov, Dmytro
Veröffentlicht: (2026)
von: Ustynov, Dmytro
Veröffentlicht: (2026)
Project Prometheus: Bridging the Intent Gap in Agentic Program Repair via Reverse-Engineered Executable Specifications
von: Wang, Yongchao, et al.
Veröffentlicht: (2026)
von: Wang, Yongchao, et al.
Veröffentlicht: (2026)
DIVE: Scaling Diversity in Agentic Task Synthesis for Generalizable Tool Use
von: Chen, Aili, et al.
Veröffentlicht: (2026)
von: Chen, Aili, et al.
Veröffentlicht: (2026)
Talking with Verifiers: Automatic Specification Generation for Neural Network Verification
von: Elboher, Yizhak Y., et al.
Veröffentlicht: (2026)
von: Elboher, Yizhak Y., et al.
Veröffentlicht: (2026)
A-ProS: Towards Reliable Autonomous Programming Through Multi-Model Feedback
von: Tabassum, Anika, et al.
Veröffentlicht: (2026)
von: Tabassum, Anika, et al.
Veröffentlicht: (2026)
ConVer: Using Contracts and Loop Invariant Synthesis for Scalable Formal Software Verification
von: Pirzada, Muhammad A. A., et al.
Veröffentlicht: (2026)
von: Pirzada, Muhammad A. A., et al.
Veröffentlicht: (2026)
QualityFlow: An Agentic Workflow for Program Synthesis Controlled by LLM Quality Checks
von: Hu, Yaojie, et al.
Veröffentlicht: (2025)
von: Hu, Yaojie, et al.
Veröffentlicht: (2025)
AutoVerus: Automated Proof Generation for Rust Code
von: Yang, Chenyuan, et al.
Veröffentlicht: (2024)
von: Yang, Chenyuan, et al.
Veröffentlicht: (2024)
OpenComputer: Verifiable Software Worlds for Computer-Use Agents
von: Wei, Jinbiao, et al.
Veröffentlicht: (2026)
von: Wei, Jinbiao, et al.
Veröffentlicht: (2026)
PyGen: A Collaborative Human-AI Approach to Python Package Creation
von: Barua, Saikat, et al.
Veröffentlicht: (2024)
von: Barua, Saikat, et al.
Veröffentlicht: (2024)
Agentic Agile-V: From Vibe Coding to Verified Engineering in Software and Hardware Development
von: Koch, Christopher
Veröffentlicht: (2026)
von: Koch, Christopher
Veröffentlicht: (2026)
Certified Program Synthesis with a Multi-Modal Verifier
von: Feng, Yueyang, et al.
Veröffentlicht: (2026)
von: Feng, Yueyang, et al.
Veröffentlicht: (2026)
Demystifying the Lifecycle of Failures in Platform-Orchestrated Agentic Workflows
von: Ma, Xuyan, et al.
Veröffentlicht: (2025)
von: Ma, Xuyan, et al.
Veröffentlicht: (2025)
SWE-Universe: Scale Real-World Verifiable Environments to Millions
von: Chen, Mouxiang, et al.
Veröffentlicht: (2026)
von: Chen, Mouxiang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Towards AI-Assisted Synthesis of Verified Dafny Methods
von: Misu, Md Rakib Hossain, et al.
Veröffentlicht: (2024) -
Toward Automated Validation of Language Model Synthesized Test Cases using Semantic Entropy
von: Taherkhani, Hamed, et al.
Veröffentlicht: (2024) -
From Natural Language to Verified Code: Toward AI Assisted Problem-to-Code Generation with Dafny-Based Formal Verification
von: Erfan, Md, et al.
Veröffentlicht: (2026) -
DafnyBench: A Benchmark for Formal Software Verification
von: Loughridge, Chloe, et al.
Veröffentlicht: (2024) -
Automated Proof Generation for Rust Code via Self-Evolution
von: Chen, Tianyu, et al.
Veröffentlicht: (2024)