PyVeritas: On Verifying Python via LLM-Based Transpilation and Bounded Model Checking for C
Fuente:
arXiv
Salvato in:
| Autori principali: | Orvalho, Pedro, Kwiatkowska, Marta |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Are Large Language Models Robust in Understanding Code Against Semantics-Preserving Mutations?
di: Orvalho, Pedro, et al.
Pubblicazione: (2025)
di: Orvalho, Pedro, et al.
Pubblicazione: (2025)
Model-Based Diagnosis with Multiple Observations: A Unified Approach for C Software and Boolean Circuits
di: Orvalho, Pedro, et al.
Pubblicazione: (2025)
di: Orvalho, Pedro, et al.
Pubblicazione: (2025)
CFaults: Model-Based Diagnosis for Fault Localization in C Programs with Multiple Test Cases
di: Orvalho, Pedro, et al.
Pubblicazione: (2024)
di: Orvalho, Pedro, et al.
Pubblicazione: (2024)
Counterexample Guided Program Repair Using Zero-Shot Learning and MaxSAT-based Fault Localization
di: Orvalho, Pedro, et al.
Pubblicazione: (2024)
di: Orvalho, Pedro, et al.
Pubblicazione: (2024)
STELP: Secure Transpilation and Execution of LLM-Generated Programs
di: Shinde, Swapnil, et al.
Pubblicazione: (2026)
di: Shinde, Swapnil, et al.
Pubblicazione: (2026)
LLM-Powered Quantum Code Transpilation
di: Siavash, Nazanin, et al.
Pubblicazione: (2025)
di: Siavash, Nazanin, et al.
Pubblicazione: (2025)
InvAASTCluster: On Applying Invariant-Based Program Clustering to Introductory Programming Assignments
di: Orvalho, Pedro, et al.
Pubblicazione: (2022)
di: Orvalho, Pedro, et al.
Pubblicazione: (2022)
PyGen: A Collaborative Human-AI Approach to Python Package Creation
di: Barua, Saikat, et al.
Pubblicazione: (2024)
di: Barua, Saikat, et al.
Pubblicazione: (2024)
AutoICE: Automatically Synthesizing Verifiable C Code via LLM-driven Evolution
di: Luo, Weilin, et al.
Pubblicazione: (2025)
di: Luo, Weilin, et al.
Pubblicazione: (2025)
PyResBugs: A Dataset of Residual Python Bugs for Natural Language-Driven Fault Injection
di: Cotroneo, Domenico, et al.
Pubblicazione: (2025)
di: Cotroneo, Domenico, et al.
Pubblicazione: (2025)
FasterPy: An LLM-based Code Execution Efficiency Optimization Framework
di: Wu, Yue, et al.
Pubblicazione: (2025)
di: Wu, Yue, et al.
Pubblicazione: (2025)
PyBench: Evaluating LLM Agent on various real-world coding tasks
di: Zhang, Yaolun, et al.
Pubblicazione: (2024)
di: Zhang, Yaolun, et al.
Pubblicazione: (2024)
QualityFlow: An Agentic Workflow for Program Synthesis Controlled by LLM Quality Checks
di: Hu, Yaojie, et al.
Pubblicazione: (2025)
di: Hu, Yaojie, et al.
Pubblicazione: (2025)
Keeping Deep Learning Models in Check: A History-Based Approach to Mitigate Overfitting
di: Li, Hao, et al.
Pubblicazione: (2024)
di: Li, Hao, et al.
Pubblicazione: (2024)
Verifying LLM-Generated Code in the Context of Software Verification with Ada/SPARK
di: Cramer, Marcos, et al.
Pubblicazione: (2025)
di: Cramer, Marcos, et al.
Pubblicazione: (2025)
Faver: Boosting LLM-based RTL Generation with Function Abstracted Verifiable Middleware
di: Mu, Jianan, et al.
Pubblicazione: (2025)
di: Mu, Jianan, et al.
Pubblicazione: (2025)
Verify-Gated Completion as Admission Control in a Governed Multi-Agent Runtime: A Bounded Architecture Case Study
di: Nguyen, Hai-Duong, et al.
Pubblicazione: (2026)
di: Nguyen, Hai-Duong, et al.
Pubblicazione: (2026)
Agentic Property-Based Testing: Finding Bugs Across the Python Ecosystem
di: Maaz, Muhammad, et al.
Pubblicazione: (2025)
di: Maaz, Muhammad, et al.
Pubblicazione: (2025)
Large Language Model-Driven Code Compliance Checking in Building Information Modeling
di: Madireddy, Soumya, et al.
Pubblicazione: (2025)
di: Madireddy, Soumya, et al.
Pubblicazione: (2025)
CVeDRL: An Efficient Code Verifier via Difficulty-aware Reinforcement Learning
di: Shi, Ji, et al.
Pubblicazione: (2026)
di: Shi, Ji, et al.
Pubblicazione: (2026)
MIMIC-Py: An Extensible Tool for Personality-Driven Automated Game Testing with Large Language Models
di: Chen, Yifei, et al.
Pubblicazione: (2026)
di: Chen, Yifei, et al.
Pubblicazione: (2026)
Evaluating AI-generated code for C++, Fortran, Go, Java, Julia, Matlab, Python, R, and Rust
di: Diehl, Patrick, et al.
Pubblicazione: (2024)
di: Diehl, Patrick, et al.
Pubblicazione: (2024)
ChronoLLM: A Framework for Customizing Large Language Model for Digital Twins generalization based on PyChrono
di: Wang, Jingquan, et al.
Pubblicazione: (2025)
di: Wang, Jingquan, et al.
Pubblicazione: (2025)
Better Python Programming for all: With the focus on Maintainability
di: Shivashankar, Karthik, et al.
Pubblicazione: (2024)
di: Shivashankar, Karthik, et al.
Pubblicazione: (2024)
AutoWebWorld: Synthesizing Infinite Verifiable Web Environments via Finite State Machines
di: Wu, Yifan, et al.
Pubblicazione: (2026)
di: Wu, Yifan, et al.
Pubblicazione: (2026)
LLM-Vectorizer: LLM-based Verified Loop Vectorizer
di: Taneja, Jubi, et al.
Pubblicazione: (2024)
di: Taneja, Jubi, et al.
Pubblicazione: (2024)
VIBEPASS: Can Vibe Coders Really Pass the Vibe Check?
di: Bansal, Srijan, et al.
Pubblicazione: (2026)
di: Bansal, Srijan, et al.
Pubblicazione: (2026)
ARCEAK: An Automated Rule Checking Framework Enhanced with Architectural Knowledge
di: Chen, Junyong, et al.
Pubblicazione: (2024)
di: Chen, Junyong, et al.
Pubblicazione: (2024)
SIEVE: Towards Verifiable Certification for Code-datasets
di: Mbodji, Fatou Ndiaye, et al.
Pubblicazione: (2025)
di: Mbodji, Fatou Ndiaye, et al.
Pubblicazione: (2025)
WybeCoder: Verified Imperative Code Generation
di: Gloeckle, Fabian, et al.
Pubblicazione: (2026)
di: Gloeckle, Fabian, et al.
Pubblicazione: (2026)
BPMN Assistant: An LLM-Based Approach to Business Process Modeling
di: Licardo, Josip Tomo, et al.
Pubblicazione: (2025)
di: Licardo, Josip Tomo, et al.
Pubblicazione: (2025)
PyGraft: Configurable Generation of Synthetic Schemas and Knowledge Graphs at Your Fingertips
di: Hubert, Nicolas, et al.
Pubblicazione: (2023)
di: Hubert, Nicolas, et al.
Pubblicazione: (2023)
TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories
di: Dong, Honghua, et al.
Pubblicazione: (2025)
di: Dong, Honghua, et al.
Pubblicazione: (2025)
The Last Dependency Crusade: Solving Python Dependency Conflicts with LLMs
di: Bartlett, Antony, et al.
Pubblicazione: (2025)
di: Bartlett, Antony, et al.
Pubblicazione: (2025)
Machine Learning Techniques for Python Source Code Vulnerability Detection
di: Farasat, Talaya, et al.
Pubblicazione: (2024)
di: Farasat, Talaya, et al.
Pubblicazione: (2024)
Aletheia: What Makes RLVR For Code Verifiers Tick?
di: Venkatkrishna, Vatsal, et al.
Pubblicazione: (2026)
di: Venkatkrishna, Vatsal, et al.
Pubblicazione: (2026)
From Natural Language to Verified Code: Toward AI Assisted Problem-to-Code Generation with Dafny-Based Formal Verification
di: Erfan, Md, et al.
Pubblicazione: (2026)
di: Erfan, Md, et al.
Pubblicazione: (2026)
Studying the Impact of TensorFlow and PyTorch Bindings on Machine Learning Software Quality
di: Li, Hao, et al.
Pubblicazione: (2024)
di: Li, Hao, et al.
Pubblicazione: (2024)
Quality and Security Signals in AI-Generated Python Refactoring Pull Requests
di: Almukhtar, Mohamed, et al.
Pubblicazione: (2026)
di: Almukhtar, Mohamed, et al.
Pubblicazione: (2026)
Adaptive Hierarchical Evaluation of LLMs and SAST tools for CWE Prediction in Python
di: Adnan, Muntasir, et al.
Pubblicazione: (2026)
di: Adnan, Muntasir, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Are Large Language Models Robust in Understanding Code Against Semantics-Preserving Mutations?
di: Orvalho, Pedro, et al.
Pubblicazione: (2025) -
Model-Based Diagnosis with Multiple Observations: A Unified Approach for C Software and Boolean Circuits
di: Orvalho, Pedro, et al.
Pubblicazione: (2025) -
CFaults: Model-Based Diagnosis for Fault Localization in C Programs with Multiple Test Cases
di: Orvalho, Pedro, et al.
Pubblicazione: (2024) -
Counterexample Guided Program Repair Using Zero-Shot Learning and MaxSAT-based Fault Localization
di: Orvalho, Pedro, et al.
Pubblicazione: (2024) -
STELP: Secure Transpilation and Execution of LLM-Generated Programs
di: Shinde, Swapnil, et al.
Pubblicazione: (2026)