Saved in:
| Main Authors: | Mitchell, Jacqueline, Shaaban, Yasser |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2511.00202 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Scientific Texts to Verifiable Code: Automating the Process with Transformers
by: Wang, Changjie, et al.
Published: (2025)
by: Wang, Changjie, et al.
Published: (2025)
Context-Augmented Code Generation: How Product Context Improves AI Coding Agent Decision Compliance by 49%
by: Dillon, Drew, et al.
Published: (2026)
by: Dillon, Drew, et al.
Published: (2026)
Imandra CodeLogician: Neuro-Symbolic Reasoning for Precise Analysis of Software Logic
by: Lin, Hongyu, et al.
Published: (2026)
by: Lin, Hongyu, et al.
Published: (2026)
Reasoning about expression evaluation under interference
by: Hayes, Ian J., et al.
Published: (2024)
by: Hayes, Ian J., et al.
Published: (2024)
Data reification in a concurrent rely-guarantee algebra
by: Meinicke, Larissa A., et al.
Published: (2024)
by: Meinicke, Larissa A., et al.
Published: (2024)
Reasoning about concurrent loops and recursion with rely-guarantee rules
by: Hayes, Ian J., et al.
Published: (2025)
by: Hayes, Ian J., et al.
Published: (2025)
Generative transformations and patterns in LLM-native approaches for software verification and falsification
by: Braberman, Víctor A., et al.
Published: (2024)
by: Braberman, Víctor A., et al.
Published: (2024)
Monitoring Unmanned Aircraft: Specification, Integration, and Lessons-learned
by: Baumeister, Jan, et al.
Published: (2024)
by: Baumeister, Jan, et al.
Published: (2024)
CETBench: A Novel Dataset constructed via Transformations over Programs for Benchmarking LLMs for Code-Equivalence Checking
by: Oza, Neeva, et al.
Published: (2025)
by: Oza, Neeva, et al.
Published: (2025)
Formal Verification of Imperative First-Class Functions in Move
by: Grieskamp, Wolfgang, et al.
Published: (2026)
by: Grieskamp, Wolfgang, et al.
Published: (2026)
Towards Bug-Free Distributed Go Programs
by: Koo, Zhengqun
Published: (2025)
by: Koo, Zhengqun
Published: (2025)
Quantifying Confidence in Assurance 2.0 Arguments
by: Bloomfield, Robin, et al.
Published: (2026)
by: Bloomfield, Robin, et al.
Published: (2026)
DeepCodeProbe: Towards Understanding What Models Trained on Code Learn
by: Majdinasab, Vahid, et al.
Published: (2024)
by: Majdinasab, Vahid, et al.
Published: (2024)
VibeContract: The Missing Quality Assurance Piece in Vibe Coding
by: Wang, Song
Published: (2026)
by: Wang, Song
Published: (2026)
Instruction and Solution Probabilities as Heuristics for Inductive Programming
by: McDaid, Edward, et al.
Published: (2025)
by: McDaid, Edward, et al.
Published: (2025)
Converting BPMN Diagrams to Privacy Calculus
by: Pitsiladis, Georgios V., et al.
Published: (2024)
by: Pitsiladis, Georgios V., et al.
Published: (2024)
A Rust-to-Lean Verification Pipeline with AI Provers: An Experience Report
by: Klaus, Natalia, et al.
Published: (2026)
by: Klaus, Natalia, et al.
Published: (2026)
Vibe Checker: Aligning Code Evaluation with Human Preference
by: Zhong, Ming, et al.
Published: (2025)
by: Zhong, Ming, et al.
Published: (2025)
RefactorBench: Evaluating Stateful Reasoning in Language Agents Through Code
by: Gautam, Dhruv, et al.
Published: (2025)
by: Gautam, Dhruv, et al.
Published: (2025)
Neural Theorem Proving for Verification Conditions: A Real-World Benchmark
by: Xu, Qiyuan, et al.
Published: (2026)
by: Xu, Qiyuan, et al.
Published: (2026)
Talk is Cheap, Logic is Hard: Benchmarking LLMs on Post-Condition Formalization
by: Prasetya, I. S. W. B., et al.
Published: (2026)
by: Prasetya, I. S. W. B., et al.
Published: (2026)
Abductive Vibe Coding (Extended Abstract)
by: Murphy, Logan, et al.
Published: (2026)
by: Murphy, Logan, et al.
Published: (2026)
On the Soundness and Consistency of LLM Agents for Executing Test Cases Written in Natural Language
by: Salva, Sébastien, et al.
Published: (2025)
by: Salva, Sébastien, et al.
Published: (2025)
SLEAN: Simple Lightweight Ensemble Analysis Network for Multi-Provider LLM Coordination: Design, Implementation, and Vibe Coding Bug Investigation Case Study
by: Vargas, Matheus J. T.
Published: (2025)
by: Vargas, Matheus J. T.
Published: (2025)
Scattered Forest Search: Smarter Code Space Exploration with LLMs
by: Light, Jonathan, et al.
Published: (2024)
by: Light, Jonathan, et al.
Published: (2024)
Combining LLM Code Generation with Formal Specifications and Reactive Program Synthesis
by: Murphy, William, et al.
Published: (2024)
by: Murphy, William, et al.
Published: (2024)
A Unit Proofing Framework for Code-level Verification: A Research Agenda
by: Amusuo, Paschal C., et al.
Published: (2024)
by: Amusuo, Paschal C., et al.
Published: (2024)
RepoLaunch: Automating Build&Test Pipeline of Code Repositories on ANY Language and ANY Platform
by: Li, Kenan, et al.
Published: (2026)
by: Li, Kenan, et al.
Published: (2026)
Constrained LTL Specification Learning from Examples
by: Zhang, Changjian, et al.
Published: (2024)
by: Zhang, Changjian, et al.
Published: (2024)
Proof-Carrying Neuro-Symbolic Code
by: Komendantskaya, Ekaterina
Published: (2025)
by: Komendantskaya, Ekaterina
Published: (2025)
Proving the Coding Interview: A Benchmark for Formally Verified Code Generation
by: Dougherty, Quinn, et al.
Published: (2025)
by: Dougherty, Quinn, et al.
Published: (2025)
Non-Ground Congruence Closure
by: Leidinger, Hendrik, et al.
Published: (2024)
by: Leidinger, Hendrik, et al.
Published: (2024)
Utilizing Precise and Complete Code Context to Guide LLM in Automatic False Positive Mitigation
by: Chen, Jinbao, et al.
Published: (2024)
by: Chen, Jinbao, et al.
Published: (2024)
Kajal: Extracting Grammar of a Source Code Using Large Language Models
by: Torkamani, Mohammad Jalili
Published: (2024)
by: Torkamani, Mohammad Jalili
Published: (2024)
From Prompting to Verification: How Experience Shapes Vibe Coding Practices
by: Fawzy, Ahmed, et al.
Published: (2026)
by: Fawzy, Ahmed, et al.
Published: (2026)
Are Users More Willing to Use Formally Verified Password Managers?
by: Carreira, Carolina, et al.
Published: (2025)
by: Carreira, Carolina, et al.
Published: (2025)
ProofBridge: Auto-Formalization of Natural Language Proofs in Lean via Joint Embeddings
by: Jana, Prithwish, et al.
Published: (2025)
by: Jana, Prithwish, et al.
Published: (2025)
Vibe Code Bench: Evaluating AI Models on End-to-End Web Application Development
by: Tran, Hung, et al.
Published: (2026)
by: Tran, Hung, et al.
Published: (2026)
Taming Differentiable Logics with Coq Formalisation
by: Affeldt, Reynald, et al.
Published: (2024)
by: Affeldt, Reynald, et al.
Published: (2024)
Quantitative Assurance and Synthesis of Controllers from Activity Diagrams
by: Ye, Kangfeng, et al.
Published: (2024)
by: Ye, Kangfeng, et al.
Published: (2024)
Similar Items
-
From Scientific Texts to Verifiable Code: Automating the Process with Transformers
by: Wang, Changjie, et al.
Published: (2025) -
Context-Augmented Code Generation: How Product Context Improves AI Coding Agent Decision Compliance by 49%
by: Dillon, Drew, et al.
Published: (2026) -
Imandra CodeLogician: Neuro-Symbolic Reasoning for Precise Analysis of Software Logic
by: Lin, Hongyu, et al.
Published: (2026) -
Reasoning about expression evaluation under interference
by: Hayes, Ian J., et al.
Published: (2024) -
Data reification in a concurrent rely-guarantee algebra
by: Meinicke, Larissa A., et al.
Published: (2024)