HoarePrompt: Structural Reasoning About Program Correctness in Natural Language
Fuente:
arXiv
Guardado en:
| Autores principales: | Bouras, Dimitrios Stamatios, Dai, Yihan, Wang, Tairan, Xiong, Yingfei, Mechtaev, Sergey |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Statistical Independence Aware Caching for LLM Workflows
por: Dai, Yihan, et al.
Publicado: (2025)
por: Dai, Yihan, et al.
Publicado: (2025)
Defusing Logic Bombs in Symbolic Execution with LLM-Generated Ghost Code
por: Bouras, Dimitrios Stamatios, et al.
Publicado: (2026)
por: Bouras, Dimitrios Stamatios, et al.
Publicado: (2026)
FM-Agent: Scaling Formal Methods to Large Systems via LLM-Based Hoare-Style Reasoning
por: Ding, Haoran, et al.
Publicado: (2026)
por: Ding, Haoran, et al.
Publicado: (2026)
Reducing Hallucinations in LLM-Generated Code via Semantic Triangulation
por: Dai, Yihan, et al.
Publicado: (2025)
por: Dai, Yihan, et al.
Publicado: (2025)
Reducing Cost of LLM Agents with Trajectory Reduction
por: Xiao, Yuan-An, et al.
Publicado: (2025)
por: Xiao, Yuan-An, et al.
Publicado: (2025)
Learning to Guarantee Type Correctness in Code Generation through Type-Guided Program Synthesis
por: Huang, Zhechong, et al.
Publicado: (2025)
por: Huang, Zhechong, et al.
Publicado: (2025)
CrossPL: Evaluating Large Language Models on Cross Programming Language Code Generation
por: Xiong, Zhanhang, et al.
Publicado: (2025)
por: Xiong, Zhanhang, et al.
Publicado: (2025)
Natural Language-Programming Language Software Traceability Link Recovery Needs More than Textual Similarity
por: Zou, Zhiyuan, et al.
Publicado: (2025)
por: Zou, Zhiyuan, et al.
Publicado: (2025)
CodeCrash: Exposing LLM Fragility to Misleading Natural Language in Code Reasoning
por: Lam, Man Ho, et al.
Publicado: (2025)
por: Lam, Man Ho, et al.
Publicado: (2025)
Validating LLM-Generated Programs with Metamorphic Prompt Testing
por: Wang, Xiaoyin, et al.
Publicado: (2024)
por: Wang, Xiaoyin, et al.
Publicado: (2024)
CodeGRAG: Bridging the Gap between Natural Language and Programming Language via Graphical Retrieval Augmented Generation
por: Du, Kounianhua, et al.
Publicado: (2024)
por: Du, Kounianhua, et al.
Publicado: (2024)
Online Prompt Selection for Program Synthesis
por: Li, Yixuan, et al.
Publicado: (2025)
por: Li, Yixuan, et al.
Publicado: (2025)
Lyra: A Benchmark for Turducken-Style Code Generation
por: Liang, Qingyuan, et al.
Publicado: (2021)
por: Liang, Qingyuan, et al.
Publicado: (2021)
Prompt Driven Development with Claude Code: Building a Complete TUI Framework for the Ring Programming Language
por: Fayed, Mahmoud Samir, et al.
Publicado: (2026)
por: Fayed, Mahmoud Samir, et al.
Publicado: (2026)
Can LLMs Reason About Program Semantics? A Comprehensive Evaluation of LLMs on Formal Specification Inference
por: Le-Cong, Thanh, et al.
Publicado: (2025)
por: Le-Cong, Thanh, et al.
Publicado: (2025)
PromptPex: Automatic Test Generation for Language Model Prompts
por: Sharma, Reshabh K, et al.
Publicado: (2025)
por: Sharma, Reshabh K, et al.
Publicado: (2025)
RePaCA: Leveraging Reasoning Large Language Models for Static Automated Patch Correctness Assessment
por: Fuster-Pena, Marcos, et al.
Publicado: (2025)
por: Fuster-Pena, Marcos, et al.
Publicado: (2025)
Compressing Code Context for LLM-based Issue Resolution
por: Jia, Haoxiang, et al.
Publicado: (2026)
por: Jia, Haoxiang, et al.
Publicado: (2026)
Class-Level Code Generation from Natural Language Using Iterative, Tool-Enhanced Reasoning over Repository
por: Deshpande, Ajinkya, et al.
Publicado: (2024)
por: Deshpande, Ajinkya, et al.
Publicado: (2024)
INTERVENOR: Prompting the Coding Ability of Large Language Models with the Interactive Chain of Repair
por: Wang, Hanbin, et al.
Publicado: (2023)
por: Wang, Hanbin, et al.
Publicado: (2023)
Towards Reliable Evaluation of Neural Program Repair with Natural Robustness Testing
por: Le-Cong, Thanh, et al.
Publicado: (2024)
por: Le-Cong, Thanh, et al.
Publicado: (2024)
Strategic Optimization and Challenges of Large Language Models in Object-Oriented Programming
por: Wang, Zinan
Publicado: (2024)
por: Wang, Zinan
Publicado: (2024)
Planning-Driven Programming: A Large Language Model Programming Workflow
por: Lei, Chao, et al.
Publicado: (2024)
por: Lei, Chao, et al.
Publicado: (2024)
ProgramBench: Can Language Models Rebuild Programs From Scratch?
por: Yang, John, et al.
Publicado: (2026)
por: Yang, John, et al.
Publicado: (2026)
Metamorphic Testing of Large Language Models for Natural Language Processing
por: Cho, Steven, et al.
Publicado: (2025)
por: Cho, Steven, et al.
Publicado: (2025)
LangGPT: Rethinking Structured Reusable Prompt Design Framework for LLMs from the Programming Language
por: Wang, Ming, et al.
Publicado: (2024)
por: Wang, Ming, et al.
Publicado: (2024)
Neurosymbolic Auditing of Natural-Language Software Requirements
por: Hall, Bethel, et al.
Publicado: (2026)
por: Hall, Bethel, et al.
Publicado: (2026)
A Contemporary Survey of Large Language Model Assisted Program Analysis
por: Wang, Jiayimei, et al.
Publicado: (2025)
por: Wang, Jiayimei, et al.
Publicado: (2025)
GeoContra: From Fluent GIS Code to Verifiable Spatial Analysis with Geography-Grounded Repair
por: Xiao, Yinhao, et al.
Publicado: (2026)
por: Xiao, Yinhao, et al.
Publicado: (2026)
When Prompt Engineering Meets Software Engineering: CNL-P as Natural and Robust "APIs'' for Human-AI Interaction
por: Xing, Zhenchang, et al.
Publicado: (2025)
por: Xing, Zhenchang, et al.
Publicado: (2025)
Natural Language-Oriented Programming (NLOP): Towards Democratizing Software Creation
por: Beheshti, Amin
Publicado: (2024)
por: Beheshti, Amin
Publicado: (2024)
CGP-Tuning: Structure-Aware Soft Prompt Tuning for Code Vulnerability Detection
por: Feng, Ruijun, et al.
Publicado: (2025)
por: Feng, Ruijun, et al.
Publicado: (2025)
Beyond Functional Correctness: Investigating Coding Style Inconsistencies in Large Language Models
por: Wang, Yanlin, et al.
Publicado: (2024)
por: Wang, Yanlin, et al.
Publicado: (2024)
CLAP: Learning Transferable Binary Code Representations with Natural Language Supervision
por: Wang, Hao, et al.
Publicado: (2024)
por: Wang, Hao, et al.
Publicado: (2024)
Are LLMs Ready for TOON? Benchmarking Structural Correctness-Sustainability Trade-offs in Novel Structured Output Formats
por: Masciari, Elio, et al.
Publicado: (2026)
por: Masciari, Elio, et al.
Publicado: (2026)
Conventional Commit Classification using Large Language Models and Prompt Engineering
por: Quadir, H. M. Sazzad, et al.
Publicado: (2026)
por: Quadir, H. M. Sazzad, et al.
Publicado: (2026)
LLM4EFFI: Leveraging Large Language Models to Enhance Code Efficiency and Correctness
por: Ye, Tong, et al.
Publicado: (2025)
por: Ye, Tong, et al.
Publicado: (2025)
AutoIOT: LLM-Driven Automated Natural Language Programming for AIoT Applications
por: Shen, Leming, et al.
Publicado: (2025)
por: Shen, Leming, et al.
Publicado: (2025)
Do Current Language Models Support Code Intelligence for R Programming Language?
por: Zhao, ZiXiao, et al.
Publicado: (2024)
por: Zhao, ZiXiao, et al.
Publicado: (2024)
JTPRO: A Joint Tool-Prompt Reflective Optimization Framework for Language Agents
por: Ghoshal, Sandip, et al.
Publicado: (2026)
por: Ghoshal, Sandip, et al.
Publicado: (2026)
Ejemplares similares
-
Statistical Independence Aware Caching for LLM Workflows
por: Dai, Yihan, et al.
Publicado: (2025) -
Defusing Logic Bombs in Symbolic Execution with LLM-Generated Ghost Code
por: Bouras, Dimitrios Stamatios, et al.
Publicado: (2026) -
FM-Agent: Scaling Formal Methods to Large Systems via LLM-Based Hoare-Style Reasoning
por: Ding, Haoran, et al.
Publicado: (2026) -
Reducing Hallucinations in LLM-Generated Code via Semantic Triangulation
por: Dai, Yihan, et al.
Publicado: (2025) -
Reducing Cost of LLM Agents with Trajectory Reduction
por: Xiao, Yuan-An, et al.
Publicado: (2025)