Current Practices for Building LLM-Powered Reasoning Tools Are Ad Hoc -- and We Can Do Better
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Bembenek, Aaron |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Declarative Language for Building And Orchestrating LLM-Powered Agent Workflows
von: Daunis, Ivan
Veröffentlicht: (2025)
von: Daunis, Ivan
Veröffentlicht: (2025)
Bit-Vector CHC Solving for Binary Analysis and Binary Analysis for Bit-Vector CHC Solving
von: Bembenek, Aaron, et al.
Veröffentlicht: (2026)
von: Bembenek, Aaron, et al.
Veröffentlicht: (2026)
Oracular Programming: A Modular Foundation for Building LLM-Enabled Software
von: Laurent, Jonathan, et al.
Veröffentlicht: (2025)
von: Laurent, Jonathan, et al.
Veröffentlicht: (2025)
Making Formulog Fast: An Argument for Unconventional Datalog Evaluation (Extended Version)
von: Bembenek, Aaron, et al.
Veröffentlicht: (2024)
von: Bembenek, Aaron, et al.
Veröffentlicht: (2024)
BetterV: Controlled Verilog Generation with Discriminative Guidance
von: Pei, Zehua, et al.
Veröffentlicht: (2024)
von: Pei, Zehua, et al.
Veröffentlicht: (2024)
An LLM-Tool Compiler for Fused Parallel Function Calling
von: Singh, Simranjit, et al.
Veröffentlicht: (2024)
von: Singh, Simranjit, et al.
Veröffentlicht: (2024)
Towards Adaptive, Scalable, and Robust Coordination of LLM Agents: A Dynamic Ad-Hoc Networking Perspective
von: Li, Rui, et al.
Veröffentlicht: (2026)
von: Li, Rui, et al.
Veröffentlicht: (2026)
Can We Trust LLM Detectors?
von: Sandhan, Jivnesh, et al.
Veröffentlicht: (2026)
von: Sandhan, Jivnesh, et al.
Veröffentlicht: (2026)
Towards LLM-based optimization compilers. Can LLMs learn how to apply a single peephole optimization? Reasoning is all LLMs need!
von: Fang, Xiangxin, et al.
Veröffentlicht: (2024)
von: Fang, Xiangxin, et al.
Veröffentlicht: (2024)
Towards Formal Verification of LLM-Generated Code from Natural Language Prompts
von: Councilman, Aaron, et al.
Veröffentlicht: (2025)
von: Councilman, Aaron, et al.
Veröffentlicht: (2025)
Hey Pentti, We Did It!: A Fully Vector-Symbolic Lisp
von: Tomkins-Flanagan, Eilene, et al.
Veröffentlicht: (2025)
von: Tomkins-Flanagan, Eilene, et al.
Veröffentlicht: (2025)
OBsmith: LLM-Powered JavaScript Obfuscator Testing
von: Jiang, Shan, et al.
Veröffentlicht: (2025)
von: Jiang, Shan, et al.
Veröffentlicht: (2025)
LLM-REVal: Can We Trust LLM Reviewers Yet?
von: Li, Rui, et al.
Veröffentlicht: (2025)
von: Li, Rui, et al.
Veröffentlicht: (2025)
Automatically Improving LLM-based Verilog Generation using EDA Tool Feedback
von: Blocklove, Jason, et al.
Veröffentlicht: (2024)
von: Blocklove, Jason, et al.
Veröffentlicht: (2024)
How Do Humans Write Code? Large Models Do It the Same Way Too
von: Li, Long, et al.
Veröffentlicht: (2024)
von: Li, Long, et al.
Veröffentlicht: (2024)
BuildBench: Benchmarking LLM Agents on Compiling Real-World Open-Source Software
von: Zhang, Zehua, et al.
Veröffentlicht: (2025)
von: Zhang, Zehua, et al.
Veröffentlicht: (2025)
Symbol Correctness in Deep Neural Networks Containing Symbolic Layers
von: Bembenek, Aaron, et al.
Veröffentlicht: (2024)
von: Bembenek, Aaron, et al.
Veröffentlicht: (2024)
Can Language Models Solve Olympiad Programming?
von: Shi, Quan, et al.
Veröffentlicht: (2024)
von: Shi, Quan, et al.
Veröffentlicht: (2024)
Can LLMs Reason About Program Semantics? A Comprehensive Evaluation of LLMs on Formal Specification Inference
von: Le-Cong, Thanh, et al.
Veröffentlicht: (2025)
von: Le-Cong, Thanh, et al.
Veröffentlicht: (2025)
CodeARC: Benchmarking Reasoning Capabilities of LLM Agents for Inductive Program Synthesis
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
Improving LLM Code Reasoning via Semantic Equivalence Self-Play with Formal Verification
von: Barone, Antonio Valerio Miceli, et al.
Veröffentlicht: (2026)
von: Barone, Antonio Valerio Miceli, et al.
Veröffentlicht: (2026)
Learning From Mistakes Makes LLM Better Reasoner
von: An, Shengnan, et al.
Veröffentlicht: (2023)
von: An, Shengnan, et al.
Veröffentlicht: (2023)
From Tool Calling to Symbolic Thinking: LLMs in a Persistent Lisp Metaprogramming Loop
von: de la Torre, Jordi
Veröffentlicht: (2025)
von: de la Torre, Jordi
Veröffentlicht: (2025)
Can LLMs Learn by Teaching for Better Reasoning? A Preliminary Study
von: Ning, Xuefei, et al.
Veröffentlicht: (2024)
von: Ning, Xuefei, et al.
Veröffentlicht: (2024)
REAMS: Reasoning Enhanced Algorithm for Maths Solving
von: Singh, Eishkaran, et al.
Veröffentlicht: (2025)
von: Singh, Eishkaran, et al.
Veröffentlicht: (2025)
LiteCoOp: Lightweight Multi-LLM Shared-Tree Reasoning for Model-Serving Compiler Optimizations
von: Tang, Annabelle Sujun, et al.
Veröffentlicht: (2026)
von: Tang, Annabelle Sujun, et al.
Veröffentlicht: (2026)
ToolCaching: Towards Efficient Caching for LLM Tool-calling
von: Zhai, Yi, et al.
Veröffentlicht: (2026)
von: Zhai, Yi, et al.
Veröffentlicht: (2026)
LLMs versus the Halting Problem: Characterizing Program Termination Reasoning
von: Sultan, Oren, et al.
Veröffentlicht: (2026)
von: Sultan, Oren, et al.
Veröffentlicht: (2026)
Provable Coordination for LLM Agents via Message Sequence Charts
von: Bollig, Benedikt, et al.
Veröffentlicht: (2026)
von: Bollig, Benedikt, et al.
Veröffentlicht: (2026)
HYSYNTH: Context-Free LLM Approximation for Guiding Program Synthesis
von: Barke, Shraddha, et al.
Veröffentlicht: (2024)
von: Barke, Shraddha, et al.
Veröffentlicht: (2024)
DriftScript: A Domain-Specific Language for Programming Non-Axiomatic Reasoning Agents
von: Brady, Seamus
Veröffentlicht: (2026)
von: Brady, Seamus
Veröffentlicht: (2026)
Improving LLM Classification of Logical Errors by Integrating Error Relationship into Prompts
von: Lee, Yanggyu, et al.
Veröffentlicht: (2024)
von: Lee, Yanggyu, et al.
Veröffentlicht: (2024)
Better LLM Reasoning via Dual-Play
von: Zhang, Zhengxin, et al.
Veröffentlicht: (2025)
von: Zhang, Zhengxin, et al.
Veröffentlicht: (2025)
Testing and Understanding Erroneous Planning in LLM Agents through Synthesized User Inputs
von: Ji, Zhenlan, et al.
Veröffentlicht: (2024)
von: Ji, Zhenlan, et al.
Veröffentlicht: (2024)
Integrating Reasoning Systems for Trustworthy AI, Proceedings of the 4th Workshop on Logic and Practice of Programming (LPOP)
von: Nerode, Anil, et al.
Veröffentlicht: (2024)
von: Nerode, Anil, et al.
Veröffentlicht: (2024)
HyGenar: An LLM-Driven Hybrid Genetic Algorithm for Few-Shot Grammar Generation
von: Tang, Weizhi, et al.
Veröffentlicht: (2025)
von: Tang, Weizhi, et al.
Veröffentlicht: (2025)
ToolLibGen: Scalable Automatic Tool Creation and Aggregation for LLM Reasoning
von: Yue, Murong, et al.
Veröffentlicht: (2025)
von: Yue, Murong, et al.
Veröffentlicht: (2025)
Hey Pentti, We Did (More of) It!: A Vector-Symbolic Lisp With Residue Arithmetic
von: Hanley, Connor, et al.
Veröffentlicht: (2025)
von: Hanley, Connor, et al.
Veröffentlicht: (2025)
Do Before You Judge: Self-Reference as a Pathway to Better LLM Evaluation
von: Lin, Wei-Hsiang, et al.
Veröffentlicht: (2025)
von: Lin, Wei-Hsiang, et al.
Veröffentlicht: (2025)
Adaptive Inference-Time Compute: LLMs Can Predict if They Can Do Better, Even Mid-Generation
von: Manvi, Rohin, et al.
Veröffentlicht: (2024)
von: Manvi, Rohin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Declarative Language for Building And Orchestrating LLM-Powered Agent Workflows
von: Daunis, Ivan
Veröffentlicht: (2025) -
Bit-Vector CHC Solving for Binary Analysis and Binary Analysis for Bit-Vector CHC Solving
von: Bembenek, Aaron, et al.
Veröffentlicht: (2026) -
Oracular Programming: A Modular Foundation for Building LLM-Enabled Software
von: Laurent, Jonathan, et al.
Veröffentlicht: (2025) -
Making Formulog Fast: An Argument for Unconventional Datalog Evaluation (Extended Version)
von: Bembenek, Aaron, et al.
Veröffentlicht: (2024) -
BetterV: Controlled Verilog Generation with Discriminative Guidance
von: Pei, Zehua, et al.
Veröffentlicht: (2024)