PyVeritas: On Verifying Python via LLM-Based Transpilation and Bounded Model Checking for C
Fuente:
arXiv
Saved in:
| Main Authors: | Orvalho, Pedro, Kwiatkowska, Marta |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Are Large Language Models Robust in Understanding Code Against Semantics-Preserving Mutations?
by: Orvalho, Pedro, et al.
Published: (2025)
by: Orvalho, Pedro, et al.
Published: (2025)
Model-Based Diagnosis with Multiple Observations: A Unified Approach for C Software and Boolean Circuits
by: Orvalho, Pedro, et al.
Published: (2025)
by: Orvalho, Pedro, et al.
Published: (2025)
CFaults: Model-Based Diagnosis for Fault Localization in C Programs with Multiple Test Cases
by: Orvalho, Pedro, et al.
Published: (2024)
by: Orvalho, Pedro, et al.
Published: (2024)
Counterexample Guided Program Repair Using Zero-Shot Learning and MaxSAT-based Fault Localization
by: Orvalho, Pedro, et al.
Published: (2024)
by: Orvalho, Pedro, et al.
Published: (2024)
STELP: Secure Transpilation and Execution of LLM-Generated Programs
by: Shinde, Swapnil, et al.
Published: (2026)
by: Shinde, Swapnil, et al.
Published: (2026)
LLM-Powered Quantum Code Transpilation
by: Siavash, Nazanin, et al.
Published: (2025)
by: Siavash, Nazanin, et al.
Published: (2025)
InvAASTCluster: On Applying Invariant-Based Program Clustering to Introductory Programming Assignments
by: Orvalho, Pedro, et al.
Published: (2022)
by: Orvalho, Pedro, et al.
Published: (2022)
PyGen: A Collaborative Human-AI Approach to Python Package Creation
by: Barua, Saikat, et al.
Published: (2024)
by: Barua, Saikat, et al.
Published: (2024)
AutoICE: Automatically Synthesizing Verifiable C Code via LLM-driven Evolution
by: Luo, Weilin, et al.
Published: (2025)
by: Luo, Weilin, et al.
Published: (2025)
PyResBugs: A Dataset of Residual Python Bugs for Natural Language-Driven Fault Injection
by: Cotroneo, Domenico, et al.
Published: (2025)
by: Cotroneo, Domenico, et al.
Published: (2025)
FasterPy: An LLM-based Code Execution Efficiency Optimization Framework
by: Wu, Yue, et al.
Published: (2025)
by: Wu, Yue, et al.
Published: (2025)
PyBench: Evaluating LLM Agent on various real-world coding tasks
by: Zhang, Yaolun, et al.
Published: (2024)
by: Zhang, Yaolun, et al.
Published: (2024)
QualityFlow: An Agentic Workflow for Program Synthesis Controlled by LLM Quality Checks
by: Hu, Yaojie, et al.
Published: (2025)
by: Hu, Yaojie, et al.
Published: (2025)
Keeping Deep Learning Models in Check: A History-Based Approach to Mitigate Overfitting
by: Li, Hao, et al.
Published: (2024)
by: Li, Hao, et al.
Published: (2024)
Verifying LLM-Generated Code in the Context of Software Verification with Ada/SPARK
by: Cramer, Marcos, et al.
Published: (2025)
by: Cramer, Marcos, et al.
Published: (2025)
Faver: Boosting LLM-based RTL Generation with Function Abstracted Verifiable Middleware
by: Mu, Jianan, et al.
Published: (2025)
by: Mu, Jianan, et al.
Published: (2025)
Verify-Gated Completion as Admission Control in a Governed Multi-Agent Runtime: A Bounded Architecture Case Study
by: Nguyen, Hai-Duong, et al.
Published: (2026)
by: Nguyen, Hai-Duong, et al.
Published: (2026)
Agentic Property-Based Testing: Finding Bugs Across the Python Ecosystem
by: Maaz, Muhammad, et al.
Published: (2025)
by: Maaz, Muhammad, et al.
Published: (2025)
Large Language Model-Driven Code Compliance Checking in Building Information Modeling
by: Madireddy, Soumya, et al.
Published: (2025)
by: Madireddy, Soumya, et al.
Published: (2025)
CVeDRL: An Efficient Code Verifier via Difficulty-aware Reinforcement Learning
by: Shi, Ji, et al.
Published: (2026)
by: Shi, Ji, et al.
Published: (2026)
MIMIC-Py: An Extensible Tool for Personality-Driven Automated Game Testing with Large Language Models
by: Chen, Yifei, et al.
Published: (2026)
by: Chen, Yifei, et al.
Published: (2026)
Evaluating AI-generated code for C++, Fortran, Go, Java, Julia, Matlab, Python, R, and Rust
by: Diehl, Patrick, et al.
Published: (2024)
by: Diehl, Patrick, et al.
Published: (2024)
ChronoLLM: A Framework for Customizing Large Language Model for Digital Twins generalization based on PyChrono
by: Wang, Jingquan, et al.
Published: (2025)
by: Wang, Jingquan, et al.
Published: (2025)
Better Python Programming for all: With the focus on Maintainability
by: Shivashankar, Karthik, et al.
Published: (2024)
by: Shivashankar, Karthik, et al.
Published: (2024)
AutoWebWorld: Synthesizing Infinite Verifiable Web Environments via Finite State Machines
by: Wu, Yifan, et al.
Published: (2026)
by: Wu, Yifan, et al.
Published: (2026)
LLM-Vectorizer: LLM-based Verified Loop Vectorizer
by: Taneja, Jubi, et al.
Published: (2024)
by: Taneja, Jubi, et al.
Published: (2024)
VIBEPASS: Can Vibe Coders Really Pass the Vibe Check?
by: Bansal, Srijan, et al.
Published: (2026)
by: Bansal, Srijan, et al.
Published: (2026)
ARCEAK: An Automated Rule Checking Framework Enhanced with Architectural Knowledge
by: Chen, Junyong, et al.
Published: (2024)
by: Chen, Junyong, et al.
Published: (2024)
SIEVE: Towards Verifiable Certification for Code-datasets
by: Mbodji, Fatou Ndiaye, et al.
Published: (2025)
by: Mbodji, Fatou Ndiaye, et al.
Published: (2025)
WybeCoder: Verified Imperative Code Generation
by: Gloeckle, Fabian, et al.
Published: (2026)
by: Gloeckle, Fabian, et al.
Published: (2026)
BPMN Assistant: An LLM-Based Approach to Business Process Modeling
by: Licardo, Josip Tomo, et al.
Published: (2025)
by: Licardo, Josip Tomo, et al.
Published: (2025)
PyGraft: Configurable Generation of Synthetic Schemas and Knowledge Graphs at Your Fingertips
by: Hubert, Nicolas, et al.
Published: (2023)
by: Hubert, Nicolas, et al.
Published: (2023)
TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories
by: Dong, Honghua, et al.
Published: (2025)
by: Dong, Honghua, et al.
Published: (2025)
The Last Dependency Crusade: Solving Python Dependency Conflicts with LLMs
by: Bartlett, Antony, et al.
Published: (2025)
by: Bartlett, Antony, et al.
Published: (2025)
Machine Learning Techniques for Python Source Code Vulnerability Detection
by: Farasat, Talaya, et al.
Published: (2024)
by: Farasat, Talaya, et al.
Published: (2024)
Aletheia: What Makes RLVR For Code Verifiers Tick?
by: Venkatkrishna, Vatsal, et al.
Published: (2026)
by: Venkatkrishna, Vatsal, et al.
Published: (2026)
From Natural Language to Verified Code: Toward AI Assisted Problem-to-Code Generation with Dafny-Based Formal Verification
by: Erfan, Md, et al.
Published: (2026)
by: Erfan, Md, et al.
Published: (2026)
Studying the Impact of TensorFlow and PyTorch Bindings on Machine Learning Software Quality
by: Li, Hao, et al.
Published: (2024)
by: Li, Hao, et al.
Published: (2024)
Quality and Security Signals in AI-Generated Python Refactoring Pull Requests
by: Almukhtar, Mohamed, et al.
Published: (2026)
by: Almukhtar, Mohamed, et al.
Published: (2026)
Adaptive Hierarchical Evaluation of LLMs and SAST tools for CWE Prediction in Python
by: Adnan, Muntasir, et al.
Published: (2026)
by: Adnan, Muntasir, et al.
Published: (2026)
Similar Items
-
Are Large Language Models Robust in Understanding Code Against Semantics-Preserving Mutations?
by: Orvalho, Pedro, et al.
Published: (2025) -
Model-Based Diagnosis with Multiple Observations: A Unified Approach for C Software and Boolean Circuits
by: Orvalho, Pedro, et al.
Published: (2025) -
CFaults: Model-Based Diagnosis for Fault Localization in C Programs with Multiple Test Cases
by: Orvalho, Pedro, et al.
Published: (2024) -
Counterexample Guided Program Repair Using Zero-Shot Learning and MaxSAT-based Fault Localization
by: Orvalho, Pedro, et al.
Published: (2024) -
STELP: Secure Transpilation and Execution of LLM-Generated Programs
by: Shinde, Swapnil, et al.
Published: (2026)