Inference-Time Computations for LLM Reasoning and Planning: A Benchmark and Insights
Fuente:
arXiv
Saved in:
| Main Authors: | Parashar, Shubham, Olson, Blake, Khurana, Sambhav, Li, Eric, Ling, Hongyi, Caverlee, James, Ji, Shuiwang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Complex LLM Planning via Automated Heuristics Discovery
by: Ling, Hongyi, et al.
Published: (2025)
by: Ling, Hongyi, et al.
Published: (2025)
Curriculum Reinforcement Learning from Easy to Hard Tasks Improves LLM Reasoning
by: Parashar, Shubham, et al.
Published: (2025)
by: Parashar, Shubham, et al.
Published: (2025)
A Hierarchical Language Model For Interpretable Graph Reasoning
by: Khurana, Sambhav, et al.
Published: (2024)
by: Khurana, Sambhav, et al.
Published: (2024)
Active Test-Time Adaptation: Theoretical Analyses and An Algorithm
by: Gui, Shurui, et al.
Published: (2024)
by: Gui, Shurui, et al.
Published: (2024)
Dynamic Search for Inference-Time Alignment in Diffusion Models
by: Li, Xiner, et al.
Published: (2025)
by: Li, Xiner, et al.
Published: (2025)
BI-DCGAN: A Theoretically Grounded Bayesian Framework for Efficient and Diverse GANs
by: Valizadeh, Mahsa, et al.
Published: (2025)
by: Valizadeh, Mahsa, et al.
Published: (2025)
Empowering GNNs via Edge-Aware Weisfeiler-Leman Algorithm
by: Liu, Meng, et al.
Published: (2022)
by: Liu, Meng, et al.
Published: (2022)
HEARTS: Benchmarking LLM Reasoning on Health Time Series
by: Li, Sirui, et al.
Published: (2026)
by: Li, Sirui, et al.
Published: (2026)
SpecReason: Fast and Accurate Inference-Time Compute via Speculative Reasoning
by: Pan, Rui, et al.
Published: (2025)
by: Pan, Rui, et al.
Published: (2025)
On the Markov Property of Neural Algorithmic Reasoning: Analyses and Methods
by: Bohde, Montgomery, et al.
Published: (2024)
by: Bohde, Montgomery, et al.
Published: (2024)
Improving LLM Reasoning through Scaling Inference Computation with Collaborative Verification
by: Liang, Zhenwen, et al.
Published: (2024)
by: Liang, Zhenwen, et al.
Published: (2024)
Equivariant Graph Network Approximations of High-Degree Polynomials for Force Field Prediction
by: Xu, Zhao, et al.
Published: (2024)
by: Xu, Zhao, et al.
Published: (2024)
A Survey of Test-Time Compute: From Intuitive Inference to Deliberate Reasoning
by: Ji, Yixin, et al.
Published: (2025)
by: Ji, Yixin, et al.
Published: (2025)
When LLM Meets Time Series: Can LLMs Perform Multi-Step Time Series Reasoning and Inference
by: Ye, Wen, et al.
Published: (2025)
by: Ye, Wen, et al.
Published: (2025)
Comprehensive Study Of Predictive Maintenance In Industries Using Classification Models And LSTM Model
by: Maheshwari, Saket, et al.
Published: (2024)
by: Maheshwari, Saket, et al.
Published: (2024)
"I May Not Have Articulated Myself Clearly": Diagnosing Dynamic Instability in LLM Reasoning at Inference Time
by: Chen, Jinkun, et al.
Published: (2026)
by: Chen, Jinkun, et al.
Published: (2026)
Learning Disentangled Equivariant Representation for Explicitly Controllable 3D Molecule Generation
by: Liu, Haoran, et al.
Published: (2024)
by: Liu, Haoran, et al.
Published: (2024)
QH9: A Quantum Hamiltonian Prediction Benchmark for QM9 Molecules
by: Yu, Haiyang, et al.
Published: (2023)
by: Yu, Haiyang, et al.
Published: (2023)
Few-Shot Recognition via Stage-Wise Retrieval-Augmented Finetuning
by: Liu, Tian, et al.
Published: (2024)
by: Liu, Tian, et al.
Published: (2024)
WebLLM: A High-Performance In-Browser LLM Inference Engine
by: Ruan, Charlie F., et al.
Published: (2024)
by: Ruan, Charlie F., et al.
Published: (2024)
Experience-Guided Adaptation of Inference-Time Reasoning Strategies
by: Stein, Adam, et al.
Published: (2025)
by: Stein, Adam, et al.
Published: (2025)
Fractional Reasoning via Latent Steering Vectors Improves Inference Time Compute
by: Liu, Sheng, et al.
Published: (2025)
by: Liu, Sheng, et al.
Published: (2025)
Evaluating System 1 vs. 2 Reasoning Approaches for Zero-Shot Time Series Forecasting: A Benchmark and Insights
by: Liu, Haoxin, et al.
Published: (2025)
by: Liu, Haoxin, et al.
Published: (2025)
ReasonBENCH: Benchmarking the (In)Stability of LLM Reasoning
by: Potamitis, Nearchos, et al.
Published: (2025)
by: Potamitis, Nearchos, et al.
Published: (2025)
Compute Aligned Training: Optimizing for Test Time Inference
by: Ousherovitch, Adam, et al.
Published: (2026)
by: Ousherovitch, Adam, et al.
Published: (2026)
FlowX: Towards Explainable Graph Neural Networks via Message Flows
by: Gui, Shurui, et al.
Published: (2022)
by: Gui, Shurui, et al.
Published: (2022)
Language Models for Controllable DNA Sequence Design
by: Su, Xingyu, et al.
Published: (2025)
by: Su, Xingyu, et al.
Published: (2025)
TS-Reasoner: Domain-Oriented Time Series Inference Agents for Reasoning and Automated Analysis
by: Ye, Wen, et al.
Published: (2024)
by: Ye, Wen, et al.
Published: (2024)
Reason for Future, Act for Now: A Principled Framework for Autonomous LLM Agents with Provable Sample Efficiency
by: Liu, Zhihan, et al.
Published: (2023)
by: Liu, Zhihan, et al.
Published: (2023)
Ragged Paged Attention: A High-Performance and Flexible LLM Inference Kernel for TPU
by: Jiang, Jevin, et al.
Published: (2026)
by: Jiang, Jevin, et al.
Published: (2026)
InsightBuild: LLM-Powered Causal Reasoning in Smart Building Systems
by: Neogi, Pinaki Prasad Guha, et al.
Published: (2025)
by: Neogi, Pinaki Prasad Guha, et al.
Published: (2025)
Decocted Experience Improves Test-Time Inference in LLM Agents
by: Shen, Maohao, et al.
Published: (2026)
by: Shen, Maohao, et al.
Published: (2026)
Best-of-Tails: Bridging Optimism and Pessimism in Inference-Time Alignment
by: Hsu, Hsiang, et al.
Published: (2026)
by: Hsu, Hsiang, et al.
Published: (2026)
Plan Before You Trade: Inference-Time Optimization for RL Trading Agents
by: Go, Eun, et al.
Published: (2026)
by: Go, Eun, et al.
Published: (2026)
CPL: Critical Plan Step Learning Boosts LLM Generalization in Reasoning Tasks
by: Wang, Tianlong, et al.
Published: (2024)
by: Wang, Tianlong, et al.
Published: (2024)
Enhanced LFTSformer: A Novel Long-Term Financial Time Series Prediction Model Using Advanced Feature Engineering and the DS Encoder Informer Architecture
by: Zhang, Jianan, et al.
Published: (2023)
by: Zhang, Jianan, et al.
Published: (2023)
Sample, Scrutinize and Scale: Effective Inference-Time Search by Scaling Verification
by: Zhao, Eric, et al.
Published: (2025)
by: Zhao, Eric, et al.
Published: (2025)
QuArch: A Benchmark for Evaluating LLM Reasoning in Computer Architecture
by: Prakash, Shvetank, et al.
Published: (2025)
by: Prakash, Shvetank, et al.
Published: (2025)
Leveraging Knowledge Graphs and LLM Reasoning to Identify Operational Bottlenecks for Warehouse Planning Assistance
by: Parekh, Rishi, et al.
Published: (2025)
by: Parekh, Rishi, et al.
Published: (2025)
Random Policy Valuation is Enough for LLM Reasoning with Verifiable Rewards
by: He, Haoran, et al.
Published: (2025)
by: He, Haoran, et al.
Published: (2025)
Similar Items
-
Complex LLM Planning via Automated Heuristics Discovery
by: Ling, Hongyi, et al.
Published: (2025) -
Curriculum Reinforcement Learning from Easy to Hard Tasks Improves LLM Reasoning
by: Parashar, Shubham, et al.
Published: (2025) -
A Hierarchical Language Model For Interpretable Graph Reasoning
by: Khurana, Sambhav, et al.
Published: (2024) -
Active Test-Time Adaptation: Theoretical Analyses and An Algorithm
by: Gui, Shurui, et al.
Published: (2024) -
Dynamic Search for Inference-Time Alignment in Diffusion Models
by: Li, Xiner, et al.
Published: (2025)