Generating Verifiable Chain of Thoughts from Exection-Traces
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Thakur, Shailja, Saxena, Vaibhav, Kulkarni, Rohan, Singh, Shivdeep, Selvam, Parameswaran, Patel, Hima, Kanayama, Hiroshi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Certified Program Synthesis with a Multi-Modal Verifier
von: Feng, Yueyang, et al.
Veröffentlicht: (2026)
von: Feng, Yueyang, et al.
Veröffentlicht: (2026)
Taming the Hydra: Targeted Control-Flow Transformations for Dynamic Symbolic Execution
von: Saumya, Charitha, et al.
Veröffentlicht: (2023)
von: Saumya, Charitha, et al.
Veröffentlicht: (2023)
Reverse Chain: A Generic-Rule for LLMs to Master Multi-API Planning
von: Zhang, Yinger, et al.
Veröffentlicht: (2023)
von: Zhang, Yinger, et al.
Veröffentlicht: (2023)
EnvTrace: Simulation-Based Semantic Evaluation of LLM Code via Execution Trace Alignment -- Demonstrated at Synchrotron Beamlines
von: van der Vleuten, Noah, et al.
Veröffentlicht: (2025)
von: van der Vleuten, Noah, et al.
Veröffentlicht: (2025)
Beyond Code Pairs: Dialogue-Based Data Generation for LLM Code Translation
von: Chen, Le, et al.
Veröffentlicht: (2025)
von: Chen, Le, et al.
Veröffentlicht: (2025)
Program Skeletons for Automated Program Translation
von: Wang, Bo, et al.
Veröffentlicht: (2025)
von: Wang, Bo, et al.
Veröffentlicht: (2025)
Scaling Granite Code Models to 128K Context
von: Stallone, Matt, et al.
Veröffentlicht: (2024)
von: Stallone, Matt, et al.
Veröffentlicht: (2024)
CLEVER: A Curated Benchmark for Formally Verified Code Generation
von: Thakur, Amitayush, et al.
Veröffentlicht: (2025)
von: Thakur, Amitayush, et al.
Veröffentlicht: (2025)
Shepherd: A Runtime Substrate Empowering Meta-Agents with a Formalized Execution Trace
von: Yu, Simon, et al.
Veröffentlicht: (2026)
von: Yu, Simon, et al.
Veröffentlicht: (2026)
Formally Verifiable Generated ASN.1/ACN Encoders and Decoders: A Case Study
von: Bucev, Mario, et al.
Veröffentlicht: (2024)
von: Bucev, Mario, et al.
Veröffentlicht: (2024)
Verifying a Realistic Mutable Hash Table
von: Chassot, Samuel, et al.
Veröffentlicht: (2021)
von: Chassot, Samuel, et al.
Veröffentlicht: (2021)
Literate Tracing
von: Sotoudeh, Matthew
Veröffentlicht: (2025)
von: Sotoudeh, Matthew
Veröffentlicht: (2025)
Evaluating the Ability of Large Language Models to Generate Verifiable Specifications in VeriFast
von: Fan, Wen, et al.
Veröffentlicht: (2024)
von: Fan, Wen, et al.
Veröffentlicht: (2024)
Verified invertible lexer using regular expressions and DFAs
von: Chassot, Samuel, et al.
Veröffentlicht: (2024)
von: Chassot, Samuel, et al.
Veröffentlicht: (2024)
Towards AI-Assisted Synthesis of Verified Dafny Methods
von: Misu, Md Rakib Hossain, et al.
Veröffentlicht: (2024)
von: Misu, Md Rakib Hossain, et al.
Veröffentlicht: (2024)
IterGen: Iterative Semantic-aware Structured LLM Generation with Backtracking
von: Ugare, Shubham, et al.
Veröffentlicht: (2024)
von: Ugare, Shubham, et al.
Veröffentlicht: (2024)
Executing as You Generate: Hiding Execution Latency in LLM Code Generation
von: Sun, Zhensu, et al.
Veröffentlicht: (2026)
von: Sun, Zhensu, et al.
Veröffentlicht: (2026)
VERINA: Benchmarking Verifiable Code Generation
von: Ye, Zhe, et al.
Veröffentlicht: (2025)
von: Ye, Zhe, et al.
Veröffentlicht: (2025)
Dynamic Stability of LLM-Generated Code
von: Rajput, Prateek, et al.
Veröffentlicht: (2025)
von: Rajput, Prateek, et al.
Veröffentlicht: (2025)
Effective LLM-Driven Code Generation with Pythoness
von: Levin, Kyla H., et al.
Veröffentlicht: (2025)
von: Levin, Kyla H., et al.
Veröffentlicht: (2025)
VERT: Verified Equivalent Rust Transpilation with Large Language Models as Few-Shot Learners
von: Yang, Aidan Z. H., et al.
Veröffentlicht: (2024)
von: Yang, Aidan Z. H., et al.
Veröffentlicht: (2024)
Smaller = Weaker? Benchmarking Robustness of Quantized LLMs in Code Generation
von: Fang, Sen, et al.
Veröffentlicht: (2025)
von: Fang, Sen, et al.
Veröffentlicht: (2025)
SEVerA: Verified Synthesis of Self-Evolving Agents
von: Banerjee, Debangshu, et al.
Veröffentlicht: (2026)
von: Banerjee, Debangshu, et al.
Veröffentlicht: (2026)
Hydra: Efficient, Correct Code Generation via Checkpoint-and-Rollback Support
von: Du, Alexander, et al.
Veröffentlicht: (2026)
von: Du, Alexander, et al.
Veröffentlicht: (2026)
Assessing GPT-4-Vision's Capabilities in UML-Based Code Generation
von: Antal, Gábor, et al.
Veröffentlicht: (2024)
von: Antal, Gábor, et al.
Veröffentlicht: (2024)
Self-Improving Code Generation via Semantic Entropy and Behavioral Consensus
von: Zhang, Huan, et al.
Veröffentlicht: (2026)
von: Zhang, Huan, et al.
Veröffentlicht: (2026)
Enhancing Automated Loop Invariant Generation for Complex Programs with Large Language Models
von: Liu, Ruibang, et al.
Veröffentlicht: (2024)
von: Liu, Ruibang, et al.
Veröffentlicht: (2024)
From Code Generation to Software Testing: AI Copilot with Context-Based RAG
von: Wang, Yuchen, et al.
Veröffentlicht: (2025)
von: Wang, Yuchen, et al.
Veröffentlicht: (2025)
AutoMCQ -- Automatically Generate Code Comprehension Questions using GenAI
von: Goodfellow, Martin, et al.
Veröffentlicht: (2025)
von: Goodfellow, Martin, et al.
Veröffentlicht: (2025)
MapReplay: Trace-Driven Benchmark Generation for Java HashMap
von: Schiavio, Filippo, et al.
Veröffentlicht: (2026)
von: Schiavio, Filippo, et al.
Veröffentlicht: (2026)
A Trace-based Approach for Code Safety Analysis
von: Xu, Hui
Veröffentlicht: (2025)
von: Xu, Hui
Veröffentlicht: (2025)
Validating Traces of Distributed Programs Against TLA+ Specifications
von: Cirstea, Horatiu, et al.
Veröffentlicht: (2024)
von: Cirstea, Horatiu, et al.
Veröffentlicht: (2024)
Once4All: Skeleton-Guided SMT Solver Fuzzing with LLM-Synthesized Generators
von: Sun, Maolin, et al.
Veröffentlicht: (2025)
von: Sun, Maolin, et al.
Veröffentlicht: (2025)
Is Functional Correctness Enough to Evaluate Code Language Models? Exploring Diversity of Generated Codes
von: Chon, Heejae, et al.
Veröffentlicht: (2024)
von: Chon, Heejae, et al.
Veröffentlicht: (2024)
Learning to Guarantee Type Correctness in Code Generation through Type-Guided Program Synthesis
von: Huang, Zhechong, et al.
Veröffentlicht: (2025)
von: Huang, Zhechong, et al.
Veröffentlicht: (2025)
ACCeLLiuM: Supervised Fine-Tuning for Automated OpenACC Pragma Generation
von: Jhaveri, Samyak, et al.
Veröffentlicht: (2025)
von: Jhaveri, Samyak, et al.
Veröffentlicht: (2025)
CodeIF-Bench: Evaluating Instruction-Following Capabilities of Large Language Models in Interactive Code Generation
von: Wang, Peiding, et al.
Veröffentlicht: (2025)
von: Wang, Peiding, et al.
Veröffentlicht: (2025)
AI Coders Are Among Us: Rethinking Programming Language Grammar Towards Efficient Code Generation
von: Sun, Zhensu, et al.
Veröffentlicht: (2024)
von: Sun, Zhensu, et al.
Veröffentlicht: (2024)
Perish or Flourish? A Holistic Evaluation of Large Language Models for Code Generation in Functional Programming
von: Lang, Nguyet-Anh H., et al.
Veröffentlicht: (2026)
von: Lang, Nguyet-Anh H., et al.
Veröffentlicht: (2026)
A Preliminary Study of Multilingual Code Language Models for Code Generation Task Using Translated Benchmarks
von: Dandamudi, Rohit, et al.
Veröffentlicht: (2024)
von: Dandamudi, Rohit, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Certified Program Synthesis with a Multi-Modal Verifier
von: Feng, Yueyang, et al.
Veröffentlicht: (2026) -
Taming the Hydra: Targeted Control-Flow Transformations for Dynamic Symbolic Execution
von: Saumya, Charitha, et al.
Veröffentlicht: (2023) -
Reverse Chain: A Generic-Rule for LLMs to Master Multi-API Planning
von: Zhang, Yinger, et al.
Veröffentlicht: (2023) -
EnvTrace: Simulation-Based Semantic Evaluation of LLM Code via Execution Trace Alignment -- Demonstrated at Synchrotron Beamlines
von: van der Vleuten, Noah, et al.
Veröffentlicht: (2025) -
Beyond Code Pairs: Dialogue-Based Data Generation for LLM Code Translation
von: Chen, Le, et al.
Veröffentlicht: (2025)