Validating Solidity Code Defects using Symbolic and Concrete Execution powered by Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Susan, Ştefan-Claudiu, Arusoaie, Andrei, Lucanu, Dorel |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Specification as Quality Gate: Three Hypotheses on AI-Assisted Code Review
by: Zietsman, Christo
Published: (2026)
by: Zietsman, Christo
Published: (2026)
LLMLOOP: Improving LLM-Generated Code and Tests through Automated Iterative Feedback Loops
by: Ravi, Ravin, et al.
Published: (2026)
by: Ravi, Ravin, et al.
Published: (2026)
Secure coding for web applications: Frameworks, challenges, and the role of LLMs
by: Kiashemshaki, Kiana, et al.
Published: (2025)
by: Kiashemshaki, Kiana, et al.
Published: (2025)
Trace Validation of Unmodified Concurrent Systems with OmniLink
by: Hackett, Finn, et al.
Published: (2026)
by: Hackett, Finn, et al.
Published: (2026)
Adaptive and AI-Augmented Security Testing: A Systematic Survey of Program Analysis, Feedback-Driven Testing, and Hybrid Learning-Based Approaches
by: Wienczkowski, Michael
Published: (2026)
by: Wienczkowski, Michael
Published: (2026)
You Don't Need Public Tests to Generate Correct Code
by: Silva, Kaushitha, et al.
Published: (2026)
by: Silva, Kaushitha, et al.
Published: (2026)
Orion: Fuzzing Workflow Automation
by: Bazalii, Max, et al.
Published: (2025)
by: Bazalii, Max, et al.
Published: (2025)
SeMA: Extending and Analyzing Storyboards to Develop Secure Android Apps
by: Mitra, Joydeep, et al.
Published: (2020)
by: Mitra, Joydeep, et al.
Published: (2020)
Towards Observation Lakehouses: Living, Interactive Archives of Software Behavior
by: Kessel, Marcus
Published: (2025)
by: Kessel, Marcus
Published: (2025)
CovRL: Fuzzing JavaScript Engines with Coverage-Guided Reinforcement Learning for LLM-based Mutation
by: Eom, Jueon, et al.
Published: (2024)
by: Eom, Jueon, et al.
Published: (2024)
NeuroLog: Reasoning You Can Audit -- Neuro-Symbolic Vulnerability Discovery via LLM Facts, Datalog, and SMT
by: Rawat, Sanjay
Published: (2026)
by: Rawat, Sanjay
Published: (2026)
AIRA: AI-Induced Risk Audit: A Structured Inspection Framework for AI-Generated Code
by: Parris, William M.
Published: (2026)
by: Parris, William M.
Published: (2026)
Validating Formal Specifications with LLM-generated Test Cases
by: Cunha, Alcino, et al.
Published: (2025)
by: Cunha, Alcino, et al.
Published: (2025)
LLM-Assisted Translation of Legacy FORTRAN Codes to C++: A Cross-Platform Study
by: Ranasinghe, Nishath Rajiv, et al.
Published: (2025)
by: Ranasinghe, Nishath Rajiv, et al.
Published: (2025)
Comparing Human and LLM Generated Code: The Jury is Still Out!
by: Licorish, Sherlock A., et al.
Published: (2025)
by: Licorish, Sherlock A., et al.
Published: (2025)
ATLAS: A Layered Constraint-Guided Framework for Structured Artifact Generation in LLM-Assisted MDE
by: Ma, Tong, et al.
Published: (2025)
by: Ma, Tong, et al.
Published: (2025)
A History Equivalence Algorithm for Dynamic Process Migration
by: Bakshi, Gargi, et al.
Published: (2024)
by: Bakshi, Gargi, et al.
Published: (2024)
Testing SSD Firmware with State Data-Aware Fuzzing: Accelerating Coverage in Nondeterministic I/O Environments
by: Yoon, Gangho, et al.
Published: (2025)
by: Yoon, Gangho, et al.
Published: (2025)
SWE-ABS: Adversarial Benchmark Strengthening Exposes Inflated Success Rates on Test-based Benchmark
by: Yu, Boxi, et al.
Published: (2026)
by: Yu, Boxi, et al.
Published: (2026)
Combined Program Analysis Techniques: A Systematic Mapping Study
by: Braione, Pietro, et al.
Published: (2026)
by: Braione, Pietro, et al.
Published: (2026)
Automatically Detecting Numerical Instability in Machine Learning Applications via Soft Assertions
by: Sharmin, Shaila, et al.
Published: (2025)
by: Sharmin, Shaila, et al.
Published: (2025)
Monitoring Agentic Systems Before They're Reliable
by: Boston, Marisa Ferrara, et al.
Published: (2026)
by: Boston, Marisa Ferrara, et al.
Published: (2026)
AgentEval: DAG-Structured Step-Level Evaluation for Agentic Workflows with Error Propagation Tracking
by: Guo, Dongxin, et al.
Published: (2026)
by: Guo, Dongxin, et al.
Published: (2026)
Towards a Probabilistic Framework for Analyzing and Improving LLM-Enabled Software
by: Baldonado, Juan Manuel, et al.
Published: (2025)
by: Baldonado, Juan Manuel, et al.
Published: (2025)
Exploring Large Language Models for Access Control Policy Synthesis and Summarization
by: Vatsa, Adarsh, et al.
Published: (2025)
by: Vatsa, Adarsh, et al.
Published: (2025)
Test-driven Software Experimentation with LASSO: an LLM Prompt Benchmarking Example
by: Kessel, Marcus
Published: (2024)
by: Kessel, Marcus
Published: (2024)
N-Version Assessment and Enhancement of Generative AI
by: Kessel, Marcus, et al.
Published: (2024)
by: Kessel, Marcus, et al.
Published: (2024)
Morescient GAI for Software Engineering (Extended Version)
by: Kessel, Marcus, et al.
Published: (2024)
by: Kessel, Marcus, et al.
Published: (2024)
Synergy of Large Language Model and Model Driven Engineering for Automated Development of Centralized Vehicular Systems
by: Petrovic, Nenad, et al.
Published: (2024)
by: Petrovic, Nenad, et al.
Published: (2024)
Iterative Audit Convergence in LLM-Managed Multi-Agent Systems: A Case Study in Prompt Engineering Quality Assurance
by: Calboreanu, Elias
Published: (2026)
by: Calboreanu, Elias
Published: (2026)
GBM Returns the Best Prediction Performance among Regression Approaches: A Case Study of Stack Overflow Code Quality
by: Licorish, Sherlock A., et al.
Published: (2025)
by: Licorish, Sherlock A., et al.
Published: (2025)
Multicalibration for LLM-based Code Generation
by: Campos, Viola, et al.
Published: (2025)
by: Campos, Viola, et al.
Published: (2025)
Synthesizing Test Cases for Narrowing Specification Candidates
by: Cunha, Alcino, et al.
Published: (2025)
by: Cunha, Alcino, et al.
Published: (2025)
A Practical Approach to Formal Methods: An Eclipse Integrated Development Environment (IDE) for Security Protocols
by: Garcia, Rémi, et al.
Published: (2024)
by: Garcia, Rémi, et al.
Published: (2024)
Model checking of hyperproperties for high-level relational models
by: Macedo, Nuno, et al.
Published: (2025)
by: Macedo, Nuno, et al.
Published: (2025)
Automated structural testing of LLM-based agents: methods, framework, and case studies
by: Kohl, Jens, et al.
Published: (2026)
by: Kohl, Jens, et al.
Published: (2026)
A Self-Improving Architecture for Dynamic Safety in Large Language Models
by: Slater, Tyler
Published: (2025)
by: Slater, Tyler
Published: (2025)
Towards Single-System Illusion in Software-Defined Vehicles -- Automated, AI-Powered Workflow
by: Lebioda, Krzysztof, et al.
Published: (2024)
by: Lebioda, Krzysztof, et al.
Published: (2024)
Governed Capability Evolution: Lifecycle-Time Compatibility Checking and Rollback for AI-Component-Based Systems, with Embodied Agents as Case Study
by: Qin, Xue, et al.
Published: (2026)
by: Qin, Xue, et al.
Published: (2026)
Secure and Scalable Blockchain Voting: A Comparative Framework and the Role of Large Language Models
by: Kiashemshaki, Kiana, et al.
Published: (2025)
by: Kiashemshaki, Kiana, et al.
Published: (2025)
Similar Items
-
The Specification as Quality Gate: Three Hypotheses on AI-Assisted Code Review
by: Zietsman, Christo
Published: (2026) -
LLMLOOP: Improving LLM-Generated Code and Tests through Automated Iterative Feedback Loops
by: Ravi, Ravin, et al.
Published: (2026) -
Secure coding for web applications: Frameworks, challenges, and the role of LLMs
by: Kiashemshaki, Kiana, et al.
Published: (2025) -
Trace Validation of Unmodified Concurrent Systems with OmniLink
by: Hackett, Finn, et al.
Published: (2026) -
Adaptive and AI-Augmented Security Testing: A Systematic Survey of Program Analysis, Feedback-Driven Testing, and Hybrid Learning-Based Approaches
by: Wienczkowski, Michael
Published: (2026)