Review of Inference-Time Scaling Strategies: Reasoning, Search and RAG
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Zhichao, Wan, Cheng, Nie, Dong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Table-R1: Inference-Time Scaling for Table Reasoning
von: Yang, Zheyuan, et al.
Veröffentlicht: (2025)
von: Yang, Zheyuan, et al.
Veröffentlicht: (2025)
Beyond Factual Accuracy: Evaluating Global Reasoning Integrity in RAG Systems with LogicScore
von: Yan, Zhichao, et al.
Veröffentlicht: (2026)
von: Yan, Zhichao, et al.
Veröffentlicht: (2026)
Diversity Enhances an LLM's Performance in RAG and Long-context Task
von: Wang, Zhichao, et al.
Veröffentlicht: (2025)
von: Wang, Zhichao, et al.
Veröffentlicht: (2025)
Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning
von: Yu, Fei, et al.
Veröffentlicht: (2025)
von: Yu, Fei, et al.
Veröffentlicht: (2025)
Scaling Inference-Time Search with Vision Value Model for Improved Visual Comprehension
von: Wang, Xiyao, et al.
Veröffentlicht: (2024)
von: Wang, Xiyao, et al.
Veröffentlicht: (2024)
Parallel Test-Time Scaling for Latent Reasoning Models
von: You, Runyang, et al.
Veröffentlicht: (2025)
von: You, Runyang, et al.
Veröffentlicht: (2025)
Towards Inference-time Scaling for Continuous Space Reasoning
von: Wang, Minghan, et al.
Veröffentlicht: (2025)
von: Wang, Minghan, et al.
Veröffentlicht: (2025)
GTS: Inference-Time Scaling of Latent Reasoning with a Learnable Gaussian Thought Sampler
von: Wang, Minghan, et al.
Veröffentlicht: (2026)
von: Wang, Minghan, et al.
Veröffentlicht: (2026)
Toward Optimal Search and Retrieval for RAG
von: Leto, Alexandria, et al.
Veröffentlicht: (2024)
von: Leto, Alexandria, et al.
Veröffentlicht: (2024)
Inference Time Alignment with Reward-Guided Tree Search
von: Hung, Chia-Yu, et al.
Veröffentlicht: (2024)
von: Hung, Chia-Yu, et al.
Veröffentlicht: (2024)
One Refiner to Unlock Them All: Inference-Time Reasoning Elicitation via Reinforcement Query Refinement
von: Zhou, Yixiao, et al.
Veröffentlicht: (2026)
von: Zhou, Yixiao, et al.
Veröffentlicht: (2026)
CARE-RAG - Clinical Assessment and Reasoning in RAG
von: Potluri, Deepthi, et al.
Veröffentlicht: (2025)
von: Potluri, Deepthi, et al.
Veröffentlicht: (2025)
GlobalRAG: Enhancing Global Reasoning in Multi-hop Question Answering via Reinforcement Learning
von: Luo, Jinchang, et al.
Veröffentlicht: (2025)
von: Luo, Jinchang, et al.
Veröffentlicht: (2025)
Examining False Positives under Inference Scaling for Mathematical Reasoning
von: Wang, Yu, et al.
Veröffentlicht: (2025)
von: Wang, Yu, et al.
Veröffentlicht: (2025)
Modeling Uncertainty Trends for Timely Retrieval in Dynamic RAG
von: Li, Bo, et al.
Veröffentlicht: (2025)
von: Li, Bo, et al.
Veröffentlicht: (2025)
SyncThink: A Training-Free Strategy to Align Inference Termination with Reasoning Saturation
von: Li, Gengyang, et al.
Veröffentlicht: (2026)
von: Li, Gengyang, et al.
Veröffentlicht: (2026)
MentorCollab: Selective Large-to-Small Inference-Time Guidance for Efficient Reasoning
von: Wang, Haojin, et al.
Veröffentlicht: (2026)
von: Wang, Haojin, et al.
Veröffentlicht: (2026)
Inference Scaled GraphRAG: Improving Multi Hop Question Answering on Knowledge Graphs
von: Thompson, Travis, et al.
Veröffentlicht: (2025)
von: Thompson, Travis, et al.
Veröffentlicht: (2025)
PrismRAG: Boosting RAG Factuality with Distractor Resilience and Strategized Reasoning
von: Kachuee, Mohammad, et al.
Veröffentlicht: (2025)
von: Kachuee, Mohammad, et al.
Veröffentlicht: (2025)
Enhancing Retrieval Systems with Inference-Time Logical Reasoning
von: Faltings, Felix, et al.
Veröffentlicht: (2025)
von: Faltings, Felix, et al.
Veröffentlicht: (2025)
Diffusion Language Model Inference with Monte Carlo Tree Search
von: Huang, Zheng, et al.
Veröffentlicht: (2025)
von: Huang, Zheng, et al.
Veröffentlicht: (2025)
Towards Agentic RAG with Deep Reasoning: A Survey of RAG-Reasoning Systems in LLMs
von: Li, Yangning, et al.
Veröffentlicht: (2025)
von: Li, Yangning, et al.
Veröffentlicht: (2025)
Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning
von: Wang, Zezhong, et al.
Veröffentlicht: (2025)
von: Wang, Zezhong, et al.
Veröffentlicht: (2025)
Relevance Isn't All You Need: Scaling RAG Systems With Inference-Time Compute Via Multi-Criteria Reranking
von: LeVine, Will, et al.
Veröffentlicht: (2025)
von: LeVine, Will, et al.
Veröffentlicht: (2025)
T1: Advancing Language Model Reasoning through Reinforcement Learning and Inference Scaling
von: Hou, Zhenyu, et al.
Veröffentlicht: (2025)
von: Hou, Zhenyu, et al.
Veröffentlicht: (2025)
Causal-Counterfactual RAG: The Integration of Causal-Counterfactual Reasoning into RAG
von: Khadilkar, Harshad, et al.
Veröffentlicht: (2025)
von: Khadilkar, Harshad, et al.
Veröffentlicht: (2025)
LatentRAG: Latent Reasoning and Retrieval for Efficient Agentic RAG
von: Zheng, Yijia, et al.
Veröffentlicht: (2026)
von: Zheng, Yijia, et al.
Veröffentlicht: (2026)
Rethinking Agentic Workflows: Evaluating Inference-Based Test-Time Scaling Strategies in Text2SQL Tasks
von: Guo, Jiajing, et al.
Veröffentlicht: (2025)
von: Guo, Jiajing, et al.
Veröffentlicht: (2025)
Thinking Slow, Fast: Scaling Inference Compute with Distilled Reasoners
von: Paliotta, Daniele, et al.
Veröffentlicht: (2025)
von: Paliotta, Daniele, et al.
Veröffentlicht: (2025)
Scaling Search-Augmented LLM Reasoning via Adaptive Information Control
von: Xiong, Siheng, et al.
Veröffentlicht: (2026)
von: Xiong, Siheng, et al.
Veröffentlicht: (2026)
Adaptive Blockwise Search: Inference-Time Alignment for Large Language Models
von: Quamar, Mohammad Atif, et al.
Veröffentlicht: (2025)
von: Quamar, Mohammad Atif, et al.
Veröffentlicht: (2025)
Retrieval is Not Enough: Enhancing RAG Reasoning through Test-Time Critique and Optimization
von: Wei, Jiaqi, et al.
Veröffentlicht: (2025)
von: Wei, Jiaqi, et al.
Veröffentlicht: (2025)
MarkovScale: Towards Optimal Sequential Scaling at Inference Time
von: Wang, Youkang, et al.
Veröffentlicht: (2026)
von: Wang, Youkang, et al.
Veröffentlicht: (2026)
A Survey of Frontiers in LLM Reasoning: Inference Scaling, Learning to Reason, and Agentic Systems
von: Ke, Zixuan, et al.
Veröffentlicht: (2025)
von: Ke, Zixuan, et al.
Veröffentlicht: (2025)
Towards Hyper-Efficient RAG Systems in VecDBs: Distributed Parallel Multi-Resolution Vector Search
von: Liu, Dong, et al.
Veröffentlicht: (2025)
von: Liu, Dong, et al.
Veröffentlicht: (2025)
Evaluating Social Bias in RAG Systems: When External Context Helps and Reasoning Hurts
von: Parihar, Shweta, et al.
Veröffentlicht: (2026)
von: Parihar, Shweta, et al.
Veröffentlicht: (2026)
StructRAG: Boosting Knowledge Intensive Reasoning of LLMs via Inference-time Hybrid Information Structurization
von: Li, Zhuoqun, et al.
Veröffentlicht: (2024)
von: Li, Zhuoqun, et al.
Veröffentlicht: (2024)
BioRAG: A RAG-LLM Framework for Biological Question Reasoning
von: Wang, Chengrui, et al.
Veröffentlicht: (2024)
von: Wang, Chengrui, et al.
Veröffentlicht: (2024)
DEL-ToM: Inference-Time Scaling for Theory-of-Mind Reasoning via Dynamic Epistemic Logic
von: Wu, Yuheng, et al.
Veröffentlicht: (2025)
von: Wu, Yuheng, et al.
Veröffentlicht: (2025)
VisuoThink: Empowering LVLM Reasoning with Multimodal Tree Search
von: Wang, Yikun, et al.
Veröffentlicht: (2025)
von: Wang, Yikun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Table-R1: Inference-Time Scaling for Table Reasoning
von: Yang, Zheyuan, et al.
Veröffentlicht: (2025) -
Beyond Factual Accuracy: Evaluating Global Reasoning Integrity in RAG Systems with LogicScore
von: Yan, Zhichao, et al.
Veröffentlicht: (2026) -
Diversity Enhances an LLM's Performance in RAG and Long-context Task
von: Wang, Zhichao, et al.
Veröffentlicht: (2025) -
Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning
von: Yu, Fei, et al.
Veröffentlicht: (2025) -
Scaling Inference-Time Search with Vision Value Model for Improved Visual Comprehension
von: Wang, Xiyao, et al.
Veröffentlicht: (2024)