From Natural Language to Verified Code: Toward AI Assisted Problem-to-Code Generation with Dafny-Based Formal Verification
Fuente:
arXiv
Saved in:
| Main Authors: | Erfan, Md, Chowdhury, Md Kamal Hossain, Ryan, Ahmed, Rahman, Md Rayhanur |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Relative Positioning Based Code Chunking Method For Rich Context Retrieval In Repository Level Code Completion Task With Code Language Model
by: Rahman, Imranur, et al.
Published: (2025)
by: Rahman, Imranur, et al.
Published: (2025)
Towards AI-Assisted Synthesis of Verified Dafny Methods
by: Misu, Md Rakib Hossain, et al.
Published: (2024)
by: Misu, Md Rakib Hossain, et al.
Published: (2024)
What Are Adversaries Doing? Automating Tactics, Techniques, and Procedures Extraction: A Systematic Review
by: Tamanna, Mahzabin, et al.
Published: (2026)
by: Tamanna, Mahzabin, et al.
Published: (2026)
Mind the Gap: Evaluating LLMs for High-Level Malicious Package Detection vs. Fine-Grained Indicator Identification
by: Ryan, Ahmed, et al.
Published: (2026)
by: Ryan, Ahmed, et al.
Published: (2026)
DafnyBench: A Benchmark for Formal Software Verification
by: Loughridge, Chloe, et al.
Published: (2024)
by: Loughridge, Chloe, et al.
Published: (2024)
DafnyPro: LLM-Assisted Automated Verification for Dafny Programs
by: Banerjee, Debangshu, et al.
Published: (2026)
by: Banerjee, Debangshu, et al.
Published: (2026)
Do Automatic Comment Generation Techniques Fall Short? Exploring the Influence of Method Dependencies on Code Understanding
by: Billah, Md Mustakim, et al.
Published: (2025)
by: Billah, Md Mustakim, et al.
Published: (2025)
Beyond Single Reports: Evaluating Automated ATT&CK Technique Extraction in Multi-Report Campaign Settings
by: Haque, Md Nazmul, et al.
Published: (2026)
by: Haque, Md Nazmul, et al.
Published: (2026)
Secret Breach Detection in Source Code with Large Language Models
by: Rahman, Md Nafiu, et al.
Published: (2025)
by: Rahman, Md Nafiu, et al.
Published: (2025)
BanglaForge: LLM Collaboration with Self-Refinement for Bangla Code Generation
by: Dihan, Mahir Labib, et al.
Published: (2025)
by: Dihan, Mahir Labib, et al.
Published: (2025)
From Code to Career: Assessing Competitive Programmers for Industry Placement
by: Akib, Md Imranur Rahman, et al.
Published: (2025)
by: Akib, Md Imranur Rahman, et al.
Published: (2025)
Enhanced LLM-Based Framework for Predicting Null Pointer Dereference in Source Code
by: Sultan, Md. Fahim, et al.
Published: (2024)
by: Sultan, Md. Fahim, et al.
Published: (2024)
Code Refactoring with LLM: A Comprehensive Evaluation With Few-Shot Settings
by: Tapader, Md. Raihan, et al.
Published: (2025)
by: Tapader, Md. Raihan, et al.
Published: (2025)
Error Understanding in Program Code With LLM-DL for Multi-label Classification
by: Amin, Md Faizul Ibne, et al.
Published: (2026)
by: Amin, Md Faizul Ibne, et al.
Published: (2026)
Context Before Code: An Experience Report on Vibe Coding in Practice
by: Shuvo, Md Nasir Uddin, et al.
Published: (2026)
by: Shuvo, Md Nasir Uddin, et al.
Published: (2026)
Dafny as Verification-Aware Intermediate Language for Code Generation
by: Li, Yue Chen, et al.
Published: (2025)
by: Li, Yue Chen, et al.
Published: (2025)
ChatGPT vs. DeepSeek: A Comparative Study on AI-Based Code Generation
by: Manik, Md Motaleb Hossen
Published: (2025)
by: Manik, Md Motaleb Hossen
Published: (2025)
LLM-Based Detection of Tangled Code Changes for Higher-Quality Method-Level Bug Datasets
by: Opu, Md Nahidul Islam, et al.
Published: (2025)
by: Opu, Md Nahidul Islam, et al.
Published: (2025)
A Systematic Literature Review of the Use of GenAI Assistants for Code Comprehension: Implications for Computing Education Research and Practice
by: Qiao, Yunhan, et al.
Published: (2025)
by: Qiao, Yunhan, et al.
Published: (2025)
CodeT5-RNN: Reinforcing Contextual Embeddings for Enhanced Code Comprehension
by: Rahman, Md Mostafizer, et al.
Published: (2026)
by: Rahman, Md Mostafizer, et al.
Published: (2026)
VeriAct: Beyond Verifiability -- Agentic Synthesis of Correct and Complete Formal Specifications
by: Misu, Md Rakib Hossain, et al.
Published: (2026)
by: Misu, Md Rakib Hossain, et al.
Published: (2026)
SWE-Shepherd: Advancing PRMs for Reinforcing Code Agents
by: Dihan, Mahir Labib, et al.
Published: (2026)
by: Dihan, Mahir Labib, et al.
Published: (2026)
ConceptCoder: Improve Code Reasoning via Concept Learning
by: Rahman, Md Mahbubur, et al.
Published: (2026)
by: Rahman, Md Mahbubur, et al.
Published: (2026)
dafny-annotator: AI-Assisted Verification of Dafny Programs
by: Poesia, Gabriel, et al.
Published: (2024)
by: Poesia, Gabriel, et al.
Published: (2024)
Understanding Dominant Themes in Reviewing Agentic AI-authored Code
by: Haider, Md. Asif, et al.
Published: (2026)
by: Haider, Md. Asif, et al.
Published: (2026)
LLM-as-a-Judge for Human-AI Co-Creation: A Reliability-Aware Evaluation Framework for Coding
by: Amin, Md Faizul Ibne, et al.
Published: (2026)
by: Amin, Md Faizul Ibne, et al.
Published: (2026)
HistoryFinder: Advancing Method-Level Source Code History Generation with Accurate Oracles and Enhanced Algorithm
by: Islam, Shahidul, et al.
Published: (2025)
by: Islam, Shahidul, et al.
Published: (2025)
AgenticCyOps: Securing Multi-Agentic AI Integration in Enterprise Cyber Operations
by: Mitra, Shaswata, et al.
Published: (2026)
by: Mitra, Shaswata, et al.
Published: (2026)
CODESIM: Multi-Agent Code Generation and Problem Solving through Simulation-Driven Planning and Debugging
by: Islam, Md. Ashraful, et al.
Published: (2025)
by: Islam, Md. Ashraful, et al.
Published: (2025)
Analysis on LLMs Performance for Code Summarization
by: Akib, Md. Ahnaf, et al.
Published: (2024)
by: Akib, Md. Ahnaf, et al.
Published: (2024)
Sentiment Analysis of ML Projects: Bridging Emotional Intelligence and Code Quality
by: Ahmed, Md Shoaib, et al.
Published: (2024)
by: Ahmed, Md Shoaib, et al.
Published: (2024)
Automated Repair of AI Code with Large Language Models and Formal Verification
by: Charalambous, Yiannis, et al.
Published: (2024)
by: Charalambous, Yiannis, et al.
Published: (2024)
Assessing and Improving the Representativeness of Code Generation Benchmarks Using Knowledge Units (KUs) of Programming Languages -- An Empirical Study
by: Ahasanuzzaman, Md, et al.
Published: (2026)
by: Ahasanuzzaman, Md, et al.
Published: (2026)
How Do Agentic AI Systems Deal With Software Energy Concerns? A Pull Request-Based Study
by: Mitul, Tanjum Motin, et al.
Published: (2025)
by: Mitul, Tanjum Motin, et al.
Published: (2025)
A First Look at the Self-Admitted Technical Debt in Test Code: Taxonomy and Detection
by: Islam, Shahidul, et al.
Published: (2025)
by: Islam, Shahidul, et al.
Published: (2025)
VeCoGen: Automating Generation of Formally Verified C Code with Large Language Models
by: Sevenhuijsen, Merlijn, et al.
Published: (2024)
by: Sevenhuijsen, Merlijn, et al.
Published: (2024)
Exploring Sustainability in Scientific Software through Code Quality & Test Coverage Metrics
by: Rahman, Sheikh Md. Mushfiqur, et al.
Published: (2026)
by: Rahman, Sheikh Md. Mushfiqur, et al.
Published: (2026)
On the Adversarial Robustness of Instruction-Tuned Large Language Models for Code
by: Hossen, Md Imran, et al.
Published: (2024)
by: Hossen, Md Imran, et al.
Published: (2024)
Code Comprehension with GitHub Copilot: Performance Gains, Comprehension Trade-offs, and Behavioral Predictors in Brownfield Programming
by: Qiao, Yunhan, et al.
Published: (2025)
by: Qiao, Yunhan, et al.
Published: (2025)
Progressive Code Integration for Abstractive Bug Report Summarization
by: Karim, Shaira Sadia, et al.
Published: (2025)
by: Karim, Shaira Sadia, et al.
Published: (2025)
Similar Items
-
Relative Positioning Based Code Chunking Method For Rich Context Retrieval In Repository Level Code Completion Task With Code Language Model
by: Rahman, Imranur, et al.
Published: (2025) -
Towards AI-Assisted Synthesis of Verified Dafny Methods
by: Misu, Md Rakib Hossain, et al.
Published: (2024) -
What Are Adversaries Doing? Automating Tactics, Techniques, and Procedures Extraction: A Systematic Review
by: Tamanna, Mahzabin, et al.
Published: (2026) -
Mind the Gap: Evaluating LLMs for High-Level Malicious Package Detection vs. Fine-Grained Indicator Identification
by: Ryan, Ahmed, et al.
Published: (2026) -
DafnyBench: A Benchmark for Formal Software Verification
by: Loughridge, Chloe, et al.
Published: (2024)