Quokka: Accelerating Program Verification with LLMs via Invariant Synthesis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wei, Anjiang, Sun, Tianran, Suresh, Tarun, Wu, Haoze, Wang, Ke, Aiken, Alex |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CodeARC: Benchmarking Reasoning Capabilities of LLM Agents for Inductive Program Synthesis
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
SuperCoder: Assembly Program Superoptimization with Large Language Models
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
SATBench: Benchmarking LLMs' Logical Reasoning via Automated Puzzle Generation from SAT Formulas
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
EquiBench: Benchmarking Large Language Models' Reasoning about Program Semantics via Equivalence Checking
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
VeriCoder: Enhancing LLM-Based RTL Code Generation through Functional Correctness Validation
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
Astra: A Multi-Agent System for GPU Kernel Performance Optimization
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
Improving Parallel Program Performance with LLM Optimizers via Agent-System Interfaces
von: Wei, Anjiang, et al.
Veröffentlicht: (2024)
von: Wei, Anjiang, et al.
Veröffentlicht: (2024)
Hexcute: A Compiler Framework for Automating Layout Synthesis in GPU Programs
von: Zhang, Xiao, et al.
Veröffentlicht: (2025)
von: Zhang, Xiao, et al.
Veröffentlicht: (2025)
Program Synthesis using Inductive Logic Programming for the Abstraction and Reasoning Corpus
von: Rocha, Filipe Marinho, et al.
Veröffentlicht: (2024)
von: Rocha, Filipe Marinho, et al.
Veröffentlicht: (2024)
Bridging the Knowledge Void: Inference-time Acquisition of Unfamiliar Programming Languages for Coding Tasks
von: Shen, Chen, et al.
Veröffentlicht: (2026)
von: Shen, Chen, et al.
Veröffentlicht: (2026)
Improving LLM Code Reasoning via Semantic Equivalence Self-Play with Formal Verification
von: Barone, Antonio Valerio Miceli, et al.
Veröffentlicht: (2026)
von: Barone, Antonio Valerio Miceli, et al.
Veröffentlicht: (2026)
Is Programming by Example solved by LLMs?
von: Li, Wen-Ding, et al.
Veröffentlicht: (2024)
von: Li, Wen-Ding, et al.
Veröffentlicht: (2024)
Program Machine Policy: Addressing Long-Horizon Tasks by Integrating Program Synthesis and State Machines
von: Lin, Yu-An, et al.
Veröffentlicht: (2023)
von: Lin, Yu-An, et al.
Veröffentlicht: (2023)
Neural Task Synthesis for Visual Programming
von: Pădurean, Victor-Alexandru, et al.
Veröffentlicht: (2023)
von: Pădurean, Victor-Alexandru, et al.
Veröffentlicht: (2023)
EcoSearch: A Constant-Delay Best-First Search Algorithm for Program Synthesis
von: Matricon, Théo, et al.
Veröffentlicht: (2024)
von: Matricon, Théo, et al.
Veröffentlicht: (2024)
What I cannot execute, I do not understand: Training and Evaluating LLMs on Program Execution Traces
von: Armengol-Estapé, Jordi, et al.
Veröffentlicht: (2025)
von: Armengol-Estapé, Jordi, et al.
Veröffentlicht: (2025)
Prism: Symbolic Superoptimization of Tensor Programs
von: Wu, Mengdi, et al.
Veröffentlicht: (2026)
von: Wu, Mengdi, et al.
Veröffentlicht: (2026)
Emergent Representations of Program Semantics in Language Models Trained on Programs
von: Jin, Charles, et al.
Veröffentlicht: (2023)
von: Jin, Charles, et al.
Veröffentlicht: (2023)
Lemur: Integrating Large Language Models in Automated Program Verification
von: Wu, Haoze, et al.
Veröffentlicht: (2023)
von: Wu, Haoze, et al.
Veröffentlicht: (2023)
AIOS Compiler: LLM as Interpreter for Natural Language Programming and Flow Programming of AI Agents
von: Xu, Shuyuan, et al.
Veröffentlicht: (2024)
von: Xu, Shuyuan, et al.
Veröffentlicht: (2024)
APPL: A Prompt Programming Language for Harmonious Integration of Programs and Large Language Model Prompts
von: Dong, Honghua, et al.
Veröffentlicht: (2024)
von: Dong, Honghua, et al.
Veröffentlicht: (2024)
Can LLMs Compress (and Decompress)? Evaluating Code Understanding and Execution via Invertibility
von: Maveli, Nickil, et al.
Veröffentlicht: (2026)
von: Maveli, Nickil, et al.
Veröffentlicht: (2026)
Ranking LLM-Generated Loop Invariants for Program Verification
von: Chakraborty, Saikat, et al.
Veröffentlicht: (2023)
von: Chakraborty, Saikat, et al.
Veröffentlicht: (2023)
Mirage: A Multi-Level Superoptimizer for Tensor Programs
von: Wu, Mengdi, et al.
Veröffentlicht: (2024)
von: Wu, Mengdi, et al.
Veröffentlicht: (2024)
The Elements of Differentiable Programming
von: Blondel, Mathieu, et al.
Veröffentlicht: (2024)
von: Blondel, Mathieu, et al.
Veröffentlicht: (2024)
EnCompass: Enhancing Agent Programming with Search Over Program Execution Paths
von: Li, Zhening, et al.
Veröffentlicht: (2025)
von: Li, Zhening, et al.
Veröffentlicht: (2025)
DafnyBench: A Benchmark for Formal Software Verification
von: Loughridge, Chloe, et al.
Veröffentlicht: (2024)
von: Loughridge, Chloe, et al.
Veröffentlicht: (2024)
Probabilistic Programming with Programmable Variational Inference
von: Becker, McCoy R., et al.
Veröffentlicht: (2024)
von: Becker, McCoy R., et al.
Veröffentlicht: (2024)
Towards LLM-based optimization compilers. Can LLMs learn how to apply a single peephole optimization? Reasoning is all LLMs need!
von: Fang, Xiangxin, et al.
Veröffentlicht: (2024)
von: Fang, Xiangxin, et al.
Veröffentlicht: (2024)
Program Semantic Inequivalence Game with Large Language Models
von: Miceli-Barone, Antonio Valerio, et al.
Veröffentlicht: (2025)
von: Miceli-Barone, Antonio Valerio, et al.
Veröffentlicht: (2025)
Generating Pragmatic Examples to Train Neural Program Synthesizers
von: Vaduguru, Saujas, et al.
Veröffentlicht: (2023)
von: Vaduguru, Saujas, et al.
Veröffentlicht: (2023)
From Large to Small: Transferring CUDA Optimization Expertise via Reasoning Graph
von: Gong, Junfeng, et al.
Veröffentlicht: (2025)
von: Gong, Junfeng, et al.
Veröffentlicht: (2025)
Tilus: A Tile-Level GPGPU Programming Language for Low-Precision Computation
von: Ding, Yaoyao, et al.
Veröffentlicht: (2025)
von: Ding, Yaoyao, et al.
Veröffentlicht: (2025)
Synthesizing Programmatic Reinforcement Learning Policies with Large Language Model Guided Search
von: Liu, Max, et al.
Veröffentlicht: (2024)
von: Liu, Max, et al.
Veröffentlicht: (2024)
ProofOptimizer: Training Language Models to Simplify Proofs without Human Demonstrations
von: Gu, Alex, et al.
Veröffentlicht: (2025)
von: Gu, Alex, et al.
Veröffentlicht: (2025)
NLP Verification: Towards a General Methodology for Certifying Robustness
von: Casadio, Marco, et al.
Veröffentlicht: (2024)
von: Casadio, Marco, et al.
Veröffentlicht: (2024)
Evaluating LLMs for Hardware Design and Test
von: Blocklove, Jason, et al.
Veröffentlicht: (2024)
von: Blocklove, Jason, et al.
Veröffentlicht: (2024)
A Joint Learning Model with Variational Interaction for Multilingual Program Translation
von: Du, Yali, et al.
Veröffentlicht: (2024)
von: Du, Yali, et al.
Veröffentlicht: (2024)
Lobster: A GPU-Accelerated Framework for Neurosymbolic Programming
von: Biberstein, Paul, et al.
Veröffentlicht: (2025)
von: Biberstein, Paul, et al.
Veröffentlicht: (2025)
DINGO: Constrained Inference for Diffusion LLMs
von: Suresh, Tarun, et al.
Veröffentlicht: (2025)
von: Suresh, Tarun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
CodeARC: Benchmarking Reasoning Capabilities of LLM Agents for Inductive Program Synthesis
von: Wei, Anjiang, et al.
Veröffentlicht: (2025) -
SuperCoder: Assembly Program Superoptimization with Large Language Models
von: Wei, Anjiang, et al.
Veröffentlicht: (2025) -
SATBench: Benchmarking LLMs' Logical Reasoning via Automated Puzzle Generation from SAT Formulas
von: Wei, Anjiang, et al.
Veröffentlicht: (2025) -
EquiBench: Benchmarking Large Language Models' Reasoning about Program Semantics via Equivalence Checking
von: Wei, Anjiang, et al.
Veröffentlicht: (2025) -
VeriCoder: Enhancing LLM-Based RTL Code Generation through Functional Correctness Validation
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)