Adaptive Proof Refinement with LLM-Guided Strategy Selection
Fuente:
arXiv
Saved in:
| Main Authors: | Lu, Minghai, Zhou, Zhe, Xie, Danning, Jia, Songlin, Delaware, Benjamin, Zhang, Tianyi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Proof Automation with Large Language Models
by: Lu, Minghai, et al.
Published: (2024)
by: Lu, Minghai, et al.
Published: (2024)
RustAssure: Differential Symbolic Testing for LLM-Transpiled C-to-Rust Code
by: Bai, Yubo, et al.
Published: (2025)
by: Bai, Yubo, et al.
Published: (2025)
Scalable Language Agnostic Taint Tracking using Explicit Data Dependencies
by: Effendi, Sedick David Baker, et al.
Published: (2025)
by: Effendi, Sedick David Baker, et al.
Published: (2025)
Analyzing Logs of Large-Scale Software Systems using Time Curves Visualization
by: Borysenkov, Dmytro, et al.
Published: (2024)
by: Borysenkov, Dmytro, et al.
Published: (2024)
Safeguarding DeFi Smart Contracts against Oracle Deviations
by: Deng, Xun, et al.
Published: (2024)
by: Deng, Xun, et al.
Published: (2024)
A Unit Proofing Framework for Code-level Verification: A Research Agenda
by: Amusuo, Paschal C., et al.
Published: (2024)
by: Amusuo, Paschal C., et al.
Published: (2024)
Do Unit Proofs Work? An Empirical Study of Compositional Bounded Model Checking for Memory Safety Verification
by: Amusuo, Paschal C., et al.
Published: (2025)
by: Amusuo, Paschal C., et al.
Published: (2025)
Resilient Microservices: A Systematic Review of Recovery Patterns, Strategies, and Evaluation Frameworks
by: Mohammad, Muzeeb
Published: (2025)
by: Mohammad, Muzeeb
Published: (2025)
SWE-ABS: Adversarial Benchmark Strengthening Exposes Inflated Success Rates on Test-based Benchmark
by: Yu, Boxi, et al.
Published: (2026)
by: Yu, Boxi, et al.
Published: (2026)
ATLAS: A Layered Constraint-Guided Framework for Structured Artifact Generation in LLM-Assisted MDE
by: Ma, Tong, et al.
Published: (2025)
by: Ma, Tong, et al.
Published: (2025)
Validating Formal Specifications with LLM-generated Test Cases
by: Cunha, Alcino, et al.
Published: (2025)
by: Cunha, Alcino, et al.
Published: (2025)
Comparing Human and LLM Generated Code: The Jury is Still Out!
by: Licorish, Sherlock A., et al.
Published: (2025)
by: Licorish, Sherlock A., et al.
Published: (2025)
Testing SSD Firmware with State Data-Aware Fuzzing: Accelerating Coverage in Nondeterministic I/O Environments
by: Yoon, Gangho, et al.
Published: (2025)
by: Yoon, Gangho, et al.
Published: (2025)
Combined Program Analysis Techniques: A Systematic Mapping Study
by: Braione, Pietro, et al.
Published: (2026)
by: Braione, Pietro, et al.
Published: (2026)
A Study of Undefined Behavior Across Foreign Function Boundaries in Rust Libraries
by: McCormack, Ian, et al.
Published: (2024)
by: McCormack, Ian, et al.
Published: (2024)
Automatically Detecting Numerical Instability in Machine Learning Applications via Soft Assertions
by: Sharmin, Shaila, et al.
Published: (2025)
by: Sharmin, Shaila, et al.
Published: (2025)
Ownership in low-level intermediate representation
by: Priya, Siddharth, et al.
Published: (2024)
by: Priya, Siddharth, et al.
Published: (2024)
Rethinking Software Empirical Studies with Structural Causal Models
by: Rodriguez-Cardenas, Daniel, et al.
Published: (2026)
by: Rodriguez-Cardenas, Daniel, et al.
Published: (2026)
A History Equivalence Algorithm for Dynamic Process Migration
by: Bakshi, Gargi, et al.
Published: (2024)
by: Bakshi, Gargi, et al.
Published: (2024)
FORGE: An LLM-driven Framework for Large-Scale Smart Contract Vulnerability Dataset Construction
by: Chen, Jiachi, et al.
Published: (2025)
by: Chen, Jiachi, et al.
Published: (2025)
AutoSOUP: Safety-Oriented Unit Proof Generation for Component-level Memory-Safety Verification
by: Amusuo, Paschal C., et al.
Published: (2026)
by: Amusuo, Paschal C., et al.
Published: (2026)
Synthesizing Test Cases for Narrowing Specification Candidates
by: Cunha, Alcino, et al.
Published: (2025)
by: Cunha, Alcino, et al.
Published: (2025)
The Range Shrinks, the Threat Remains: Re-evaluating LLM Package Hallucinations on the 2026 Frontier-Model Cohort
by: Churilov, Aleksandr
Published: (2026)
by: Churilov, Aleksandr
Published: (2026)
Secure coding for web applications: Frameworks, challenges, and the role of LLMs
by: Kiashemshaki, Kiana, et al.
Published: (2025)
by: Kiashemshaki, Kiana, et al.
Published: (2025)
Towards Comprehensive Sampling of SMT Solutions
by: Lyu, Shuangyu, et al.
Published: (2025)
by: Lyu, Shuangyu, et al.
Published: (2025)
Architectural Transformations and Emerging Verification Demands in AI-Enabled Cyber-Physical Systems
by: Yusuf, Hadiza Umar, et al.
Published: (2025)
by: Yusuf, Hadiza Umar, et al.
Published: (2025)
Evaluating Cryptographic API Misuse Detectors for Go
by: Andersson, Vivi, et al.
Published: (2026)
by: Andersson, Vivi, et al.
Published: (2026)
Tractable Verification of Model Transformations: A Cutoff-Theorem Approach for DSLTrans
by: Lucio, Levi
Published: (2026)
by: Lucio, Levi
Published: (2026)
The Specification as Quality Gate: Three Hypotheses on AI-Assisted Code Review
by: Zietsman, Christo
Published: (2026)
by: Zietsman, Christo
Published: (2026)
Docker-based CI/CD for Rocq/OCaml projects
by: Martin-Dorel, Érik
Published: (2025)
by: Martin-Dorel, Érik
Published: (2025)
QuickCheck for VDM
by: Battle, Nick, et al.
Published: (2024)
by: Battle, Nick, et al.
Published: (2024)
MoCheQoS: Automated Analysis of Quality of Service Properties of Communicating Systems
by: Pombo, Carlos G. Lopez, et al.
Published: (2023)
by: Pombo, Carlos G. Lopez, et al.
Published: (2023)
A Practical Approach to Formal Methods: An Eclipse Integrated Development Environment (IDE) for Security Protocols
by: Garcia, Rémi, et al.
Published: (2024)
by: Garcia, Rémi, et al.
Published: (2024)
Model checking of hyperproperties for high-level relational models
by: Macedo, Nuno, et al.
Published: (2025)
by: Macedo, Nuno, et al.
Published: (2025)
From Separate Compilation to Sound Language Composition
by: Bruzzone, Federico, et al.
Published: (2026)
by: Bruzzone, Federico, et al.
Published: (2026)
Towards a Probabilistic Framework for Analyzing and Improving LLM-Enabled Software
by: Baldonado, Juan Manuel, et al.
Published: (2025)
by: Baldonado, Juan Manuel, et al.
Published: (2025)
Formal Methods: From Academia to Industrial Practice. A Travel Guide
by: Huisman, Marieke, et al.
Published: (2020)
by: Huisman, Marieke, et al.
Published: (2020)
Conflict Essences for Transformation Rules with Nested Application Conditions -- Long Version
by: Lauer, Alexander, et al.
Published: (2026)
by: Lauer, Alexander, et al.
Published: (2026)
Analyzing the Adoption of Database Management Systems Throughout the History of Open Source Projects
by: Paiva, Camila A., et al.
Published: (2026)
by: Paiva, Camila A., et al.
Published: (2026)
Site Reliability Engineering (SRE) and Observations on SRE Process to Make Tasks Easier
by: Puli, Balaram
Published: (2025)
by: Puli, Balaram
Published: (2025)
Similar Items
-
Proof Automation with Large Language Models
by: Lu, Minghai, et al.
Published: (2024) -
RustAssure: Differential Symbolic Testing for LLM-Transpiled C-to-Rust Code
by: Bai, Yubo, et al.
Published: (2025) -
Scalable Language Agnostic Taint Tracking using Explicit Data Dependencies
by: Effendi, Sedick David Baker, et al.
Published: (2025) -
Analyzing Logs of Large-Scale Software Systems using Time Curves Visualization
by: Borysenkov, Dmytro, et al.
Published: (2024) -
Safeguarding DeFi Smart Contracts against Oracle Deviations
by: Deng, Xun, et al.
Published: (2024)