Enhancing Test-Time Scaling of Large Language Models with Hierarchical Retrieval-Augmented MCTS
Fuente:
arXiv
Saved in:
| Main Authors: | Dou, Alex ZH, Wan, Zhongwei, Cui, Dongfei, Wang, Xin, Xiong, Jing, Lin, Haokun, Tao, Chaofan, Yan, Shen, Zhang, Mi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SRPO: Enhancing Multimodal LLM Reasoning via Reflection-Aware Reinforcement Learning
by: Wan, Zhongwei, et al.
Published: (2025)
by: Wan, Zhongwei, et al.
Published: (2025)
DSDR: Dual-Scale Diversity Regularization for Exploration in LLM Reasoning
by: Wan, Zhongwei, et al.
Published: (2026)
by: Wan, Zhongwei, et al.
Published: (2026)
MEIT: Multimodal Electrocardiogram Instruction Tuning on Large Language Models for Report Generation
by: Wan, Zhongwei, et al.
Published: (2024)
by: Wan, Zhongwei, et al.
Published: (2024)
ATTS: Asynchronous Test-Time Scaling via Conformal Prediction
by: Xiong, Jing, et al.
Published: (2025)
by: Xiong, Jing, et al.
Published: (2025)
D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models
by: Wan, Zhongwei, et al.
Published: (2024)
by: Wan, Zhongwei, et al.
Published: (2024)
QuantVLA: Scale-Calibrated Post-Training Quantization for Vision-Language-Action Models
by: Zhang, Jingxuan, et al.
Published: (2026)
by: Zhang, Jingxuan, et al.
Published: (2026)
Scaling Laws with Vocabulary: Larger Models Deserve Larger Vocabularies
by: Tao, Chaofan, et al.
Published: (2024)
by: Tao, Chaofan, et al.
Published: (2024)
SVD-LLM V2: Optimizing Singular Value Truncation for Large Language Model Compression
by: Wang, Xin, et al.
Published: (2025)
by: Wang, Xin, et al.
Published: (2025)
MCTS-RAG: Enhancing Retrieval-Augmented Generation with Monte Carlo Tree Search
by: Hu, Yunhai, et al.
Published: (2025)
by: Hu, Yunhai, et al.
Published: (2025)
Famba-V: Fast Vision Mamba with Cross-Layer Token Fusion
by: Shen, Hui, et al.
Published: (2024)
by: Shen, Hui, et al.
Published: (2024)
SVD-LLM: Truncation-aware Singular Value Decomposition for Large Language Model Compression
by: Wang, Xin, et al.
Published: (2024)
by: Wang, Xin, et al.
Published: (2024)
FREESON: Retriever-Free Retrieval-Augmented Reasoning via Corpus-Traversing MCTS
by: Kim, Chaeeun, et al.
Published: (2025)
by: Kim, Chaeeun, et al.
Published: (2025)
MMDeepResearch-Bench: A Benchmark for Multimodal Deep Research Agents
by: Huang, Peizhou, et al.
Published: (2026)
by: Huang, Peizhou, et al.
Published: (2026)
Method to Overcome Psychological Barriers in Students Learning the German Language
by: A.ZH. AKHMETOVA
Published: (2020)
by: A.ZH. AKHMETOVA
Published: (2020)
Transgene expressions in seabass (Lates calcarifer) following muscular injection of plasmid DNA: a strategy for vaccine development?
by: Sulaiman, Z.H.
Published: (1998)
by: Sulaiman, Z.H.
Published: (1998)
Transgenic fish research
by: Sulaiman, Z.H.
Published: (1995)
by: Sulaiman, Z.H.
Published: (1995)
MCTS-Judge: Test-Time Scaling in LLM-as-a-Judge for Code Correctness Evaluation
by: Wang, Yutong, et al.
Published: (2025)
by: Wang, Yutong, et al.
Published: (2025)
UncertaintyRAG: Span-Level Uncertainty Enhanced Long-Context Modeling for Retrieval-Augmented Generation
by: Li, Zixuan, et al.
Published: (2024)
by: Li, Zixuan, et al.
Published: (2024)
Retrieve-then-Adapt: Retrieval-Augmented Test-Time Adaptation for Sequential Recommendation
by: Tang, Xing, et al.
Published: (2026)
by: Tang, Xing, et al.
Published: (2026)
BiT-MCTS: A Theme-based Bidirectional MCTS Approach to Chinese Fiction Generation
by: Li, Zhaoyi, et al.
Published: (2026)
by: Li, Zhaoyi, et al.
Published: (2026)
Reasoning in Action: MCTS-Driven Knowledge Retrieval for Large Language Models
by: Liu, Shuqi, et al.
Published: (2025)
by: Liu, Shuqi, et al.
Published: (2025)
Efficient Retrieval Scaling with Hierarchical Indexing for Large Scale Recommendation
by: Fu, Dongqi, et al.
Published: (2026)
by: Fu, Dongqi, et al.
Published: (2026)
Socratic-MCTS: Test-Time Visual Reasoning by Asking the Right Questions
by: Acuna, David, et al.
Published: (2025)
by: Acuna, David, et al.
Published: (2025)
MEDA: Dynamic KV Cache Allocation for Efficient Multimodal Long-Context Inference
by: Wan, Zhongwei, et al.
Published: (2025)
by: Wan, Zhongwei, et al.
Published: (2025)
Hierarchical Consistency Learning for Test-time Adaptation in Camouflage Perception
by: Zha, Mingfeng, et al.
Published: (2026)
by: Zha, Mingfeng, et al.
Published: (2026)
Single‐Anion Conductive Solid‐State Electrolytes with Hierarchical Ionic Highways for Flexible Zinc‐Air Battery
by: Mi Xu, et al.
Published: (2024)
by: Mi Xu, et al.
Published: (2024)
Single‐Anion Conductive Solid‐State Electrolytes with Hierarchical Ionic Highways for Flexible Zinc‐Air Battery
by: Mi Xu, et al.
Published: (2024)
by: Mi Xu, et al.
Published: (2024)
Efficient Diffusion Models: A Survey
by: Shen, Hui, et al.
Published: (2025)
by: Shen, Hui, et al.
Published: (2025)
Metacognitive Retrieval-Augmented Large Language Models
by: Zhou, Yujia, et al.
Published: (2024)
by: Zhou, Yujia, et al.
Published: (2024)
MIRAGE: Scaling Test-Time Inference with Parallel Graph-Retrieval-Augmented Reasoning Chains
by: Wei, Kaiwen, et al.
Published: (2025)
by: Wei, Kaiwen, et al.
Published: (2025)
Efficient Test-Time Retrieval Augmented Generation
by: Yin, Hailong, et al.
Published: (2025)
by: Yin, Hailong, et al.
Published: (2025)
Hierarchical Structured Neural Network: Efficient Retrieval Scaling for Large Scale Recommendation
by: Rangadurai, Kaushik, et al.
Published: (2024)
by: Rangadurai, Kaushik, et al.
Published: (2024)
Prompting Test-Time Scaling Is A Strong LLM Reasoning Data Augmentation
by: Bsharat, Sondos Mahmoud, et al.
Published: (2025)
by: Bsharat, Sondos Mahmoud, et al.
Published: (2025)
MCTS-Reasoning: Monte Carlo Tree Search for LLM Reasoning
by: Towell, Alex
Published: (2026)
by: Towell, Alex
Published: (2026)
A-RAG: Scaling Agentic Retrieval-Augmented Generation via Hierarchical Retrieval Interfaces
by: Du, Mingxuan, et al.
Published: (2026)
by: Du, Mingxuan, et al.
Published: (2026)
UNComp: Can Matrix Entropy Uncover Sparsity? -- A Compressor Design from an Uncertainty-Aware Perspective
by: Xiong, Jing, et al.
Published: (2024)
by: Xiong, Jing, et al.
Published: (2024)
Dienoic‐Acid Coupling Effect Induced Hierarchical Interface for High‐Performance Zinc Metal Batteries
by: Tianyi Yang, et al.
Published: (2025)
by: Tianyi Yang, et al.
Published: (2025)
DSADF: Thinking Fast and Slow for Decision Making
by: Dou, Zhihao, et al.
Published: (2025)
by: Dou, Zhihao, et al.
Published: (2025)
Evaluating Self-Generated Documents for Enhancing Retrieval-Augmented Generation with Large Language Models
by: Li, Jiatao, et al.
Published: (2024)
by: Li, Jiatao, et al.
Published: (2024)
Retrieval, Refinement, and Ranking for Text-to-Video Generation via Prompt Optimization and Test-Time Scaling
by: Rahman, Zillur, et al.
Published: (2026)
by: Rahman, Zillur, et al.
Published: (2026)
Similar Items
-
SRPO: Enhancing Multimodal LLM Reasoning via Reflection-Aware Reinforcement Learning
by: Wan, Zhongwei, et al.
Published: (2025) -
DSDR: Dual-Scale Diversity Regularization for Exploration in LLM Reasoning
by: Wan, Zhongwei, et al.
Published: (2026) -
MEIT: Multimodal Electrocardiogram Instruction Tuning on Large Language Models for Report Generation
by: Wan, Zhongwei, et al.
Published: (2024) -
ATTS: Asynchronous Test-Time Scaling via Conformal Prediction
by: Xiong, Jing, et al.
Published: (2025) -
D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models
by: Wan, Zhongwei, et al.
Published: (2024)