Wider or Deeper? Scaling LLM Inference-Time Compute with Adaptive Branching Tree Search
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Inoue, Yuichi, Misaki, Kou, Imajuku, Yuki, Kuroki, So, Nakamura, Taishi, Akiba, Takuya |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
UnMaskFork: Test-Time Scaling for Masked Diffusion via Deterministic Action Branching
von: Misaki, Kou, et al.
Veröffentlicht: (2026)
von: Misaki, Kou, et al.
Veröffentlicht: (2026)
Agent Skill Acquisition for Large Language Models via CycleQD
von: Kuroki, So, et al.
Veröffentlicht: (2024)
von: Kuroki, So, et al.
Veröffentlicht: (2024)
String Seed of Thought: Prompting LLMs for Distribution-Faithful and Diverse Generation
von: Misaki, Kou, et al.
Veröffentlicht: (2025)
von: Misaki, Kou, et al.
Veröffentlicht: (2025)
Reimagining Agent-based Modeling with Large Language Model Agents via Shachi
von: Kuroki, So, et al.
Veröffentlicht: (2025)
von: Kuroki, So, et al.
Veröffentlicht: (2025)
Feedback-to-Rubrics: Can We Learn Expert Criteria from Inline Comments?
von: Yoshida, Kotaro, et al.
Veröffentlicht: (2026)
von: Yoshida, Kotaro, et al.
Veröffentlicht: (2026)
KAME: Tandem Architecture for Enhancing Knowledge in Real-Time Speech-to-Speech Conversational AI
von: Kuroki, So, et al.
Veröffentlicht: (2025)
von: Kuroki, So, et al.
Veröffentlicht: (2025)
TAID: Temporally Adaptive Interpolated Distillation for Efficient Knowledge Transfer in Language Models
von: Shing, Makoto, et al.
Veröffentlicht: (2025)
von: Shing, Makoto, et al.
Veröffentlicht: (2025)
ALE-Bench: A Benchmark for Long-Horizon Objective-Driven Algorithm Engineering
von: Imajuku, Yuki, et al.
Veröffentlicht: (2025)
von: Imajuku, Yuki, et al.
Veröffentlicht: (2025)
The Depth Delusion: Why Transformers Should Be Wider, Not Deeper
von: Fahim, Md Muhtasim Munif, et al.
Veröffentlicht: (2026)
von: Fahim, Md Muhtasim Munif, et al.
Veröffentlicht: (2026)
Optimize Wider, Not Deeper: Consensus Aggregation for Policy Optimization
von: Su, Zelal, et al.
Veröffentlicht: (2026)
von: Su, Zelal, et al.
Veröffentlicht: (2026)
SAIL: Test-Time Scaling for In-Context Imitation Learning with VLM
von: Sato, Makoto, et al.
Veröffentlicht: (2026)
von: Sato, Makoto, et al.
Veröffentlicht: (2026)
Drop-Upcycling: Training Sparse Mixture of Experts with Partial Re-initialization
von: Nakamura, Taishi, et al.
Veröffentlicht: (2025)
von: Nakamura, Taishi, et al.
Veröffentlicht: (2025)
LAPPI: Interactive Optimization with LLM-Assisted Preference-Based Problem Instantiation
von: Kuroki, So, et al.
Veröffentlicht: (2025)
von: Kuroki, So, et al.
Veröffentlicht: (2025)
Adaptive Inference-Time Scaling via Cyclic Diffusion Search
von: Lee, Gyubin, et al.
Veröffentlicht: (2025)
von: Lee, Gyubin, et al.
Veröffentlicht: (2025)
Watch Wider and Think Deeper: Collaborative Cross-modal Chain-of-Thought for Complex Visual Reasoning
von: Lu, Wenting, et al.
Veröffentlicht: (2026)
von: Lu, Wenting, et al.
Veröffentlicht: (2026)
Evolving Deeper LLM Thinking
von: Lee, Kuang-Huei, et al.
Veröffentlicht: (2025)
von: Lee, Kuang-Huei, et al.
Veröffentlicht: (2025)
LLM-Assisted Replication for Quantitative Social Science
von: Kubota, So, et al.
Veröffentlicht: (2026)
von: Kubota, So, et al.
Veröffentlicht: (2026)
Fast Think-on-Graph: Wider, Deeper and Faster Reasoning of Large Language Model on Knowledge Graph
von: Liang, Xujian, et al.
Veröffentlicht: (2025)
von: Liang, Xujian, et al.
Veröffentlicht: (2025)
Inference-Time Budget Control for LLM Search Agents
von: Fang, Zhengru, et al.
Veröffentlicht: (2026)
von: Fang, Zhengru, et al.
Veröffentlicht: (2026)
Sudoku-Bench: Evaluating creative reasoning with Sudoku variants
von: Seely, Jeffrey, et al.
Veröffentlicht: (2025)
von: Seely, Jeffrey, et al.
Veröffentlicht: (2025)
Adaptive Parallel Monte Carlo Tree Search for Efficient Test-time Compute Scaling
von: Kim, Hongbeen, et al.
Veröffentlicht: (2026)
von: Kim, Hongbeen, et al.
Veröffentlicht: (2026)
Memory-Guided Tree Search with Cross-Branch Knowledge Transfer for LLM Solver Synthesis
von: Haji, Fatemeh, et al.
Veröffentlicht: (2026)
von: Haji, Fatemeh, et al.
Veröffentlicht: (2026)
Sample, Scrutinize and Scale: Effective Inference-Time Search by Scaling Verification
von: Zhao, Eric, et al.
Veröffentlicht: (2025)
von: Zhao, Eric, et al.
Veröffentlicht: (2025)
Scaling LLM Inference with Optimized Sample Compute Allocation
von: Zhang, Kexun, et al.
Veröffentlicht: (2024)
von: Zhang, Kexun, et al.
Veröffentlicht: (2024)
On the Optimal Reasoning Length for RL-Trained Language Models
von: Nohara, Daisuke, et al.
Veröffentlicht: (2026)
von: Nohara, Daisuke, et al.
Veröffentlicht: (2026)
A Survey on LLM Test-Time Compute via Search: Tasks, LLM Profiling, Search Algorithms, and Relevant Frameworks
von: Li, Xinzhe
Veröffentlicht: (2025)
von: Li, Xinzhe
Veröffentlicht: (2025)
Dual-Dimensional Consistency: Balancing Budget and Quality in Adaptive Inference-Time Scaling
von: Xu, Rongman, et al.
Veröffentlicht: (2026)
von: Xu, Rongman, et al.
Veröffentlicht: (2026)
DiffusionBlocks: Block-wise Neural Network Training via Diffusion Interpretation
von: Shing, Makoto, et al.
Veröffentlicht: (2025)
von: Shing, Makoto, et al.
Veröffentlicht: (2025)
Adaptive Rectification Sampling for Test-Time Compute Scaling
von: Tan, Zhendong, et al.
Veröffentlicht: (2025)
von: Tan, Zhendong, et al.
Veröffentlicht: (2025)
Improving LLM Reasoning through Scaling Inference Computation with Collaborative Verification
von: Liang, Zhenwen, et al.
Veröffentlicht: (2024)
von: Liang, Zhenwen, et al.
Veröffentlicht: (2024)
From Efficiency to Adaptivity: A Deeper Look at Adaptive Reasoning in Large Language Models
von: Wu, Chao, et al.
Veröffentlicht: (2025)
von: Wu, Chao, et al.
Veröffentlicht: (2025)
Retro-Search: Exploring Untaken Paths for Deeper and Efficient Reasoning
von: Lu, Ximing, et al.
Veröffentlicht: (2025)
von: Lu, Ximing, et al.
Veröffentlicht: (2025)
Inference-Time Computations for LLM Reasoning and Planning: A Benchmark and Insights
von: Parashar, Shubham, et al.
Veröffentlicht: (2025)
von: Parashar, Shubham, et al.
Veröffentlicht: (2025)
Jupiter: Enhancing LLM Data Analysis Capabilities via Notebook and Inference-Time Value-Guided Search
von: Li, Shuocheng, et al.
Veröffentlicht: (2025)
von: Li, Shuocheng, et al.
Veröffentlicht: (2025)
When More Thinking Hurts: Overthinking in LLM Test-Time Compute Scaling
von: Zhou, Shu, et al.
Veröffentlicht: (2026)
von: Zhou, Shu, et al.
Veröffentlicht: (2026)
CNN-based Surface Temperature Forecasts with Ensemble Numerical Weather Prediction
von: Inoue, Takuya, et al.
Veröffentlicht: (2025)
von: Inoue, Takuya, et al.
Veröffentlicht: (2025)
A Practical Two-Stage Recipe for Mathematical LLMs: Maximizing Accuracy with SFT and Efficiency with Reinforcement Learning
von: Yoshihara, Hiroshi, et al.
Veröffentlicht: (2025)
von: Yoshihara, Hiroshi, et al.
Veröffentlicht: (2025)
Demystifying LLM-as-a-Judge: Analytically Tractable Model for Inference-Time Scaling
von: Halder, Indranil, et al.
Veröffentlicht: (2025)
von: Halder, Indranil, et al.
Veröffentlicht: (2025)
UniScale: Adaptive Unified Inference Scaling via Online Joint Optimization of Model Routing and Test-Time Scaling
von: Huang, Kaiyu, et al.
Veröffentlicht: (2026)
von: Huang, Kaiyu, et al.
Veröffentlicht: (2026)
Bag of Tricks for Inference-time Computation of LLM Reasoning
von: Liu, Fan, et al.
Veröffentlicht: (2025)
von: Liu, Fan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
UnMaskFork: Test-Time Scaling for Masked Diffusion via Deterministic Action Branching
von: Misaki, Kou, et al.
Veröffentlicht: (2026) -
Agent Skill Acquisition for Large Language Models via CycleQD
von: Kuroki, So, et al.
Veröffentlicht: (2024) -
String Seed of Thought: Prompting LLMs for Distribution-Faithful and Diverse Generation
von: Misaki, Kou, et al.
Veröffentlicht: (2025) -
Reimagining Agent-based Modeling with Large Language Model Agents via Shachi
von: Kuroki, So, et al.
Veröffentlicht: (2025) -
Feedback-to-Rubrics: Can We Learn Expert Criteria from Inline Comments?
von: Yoshida, Kotaro, et al.
Veröffentlicht: (2026)