Hierarchical Reasoning Model
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Guan, Li, Jin, Sun, Yuhao, Chen, Xing, Liu, Changling, Wu, Yue, Lu, Meng, Song, Sen, Yadkori, Yasin Abbasi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
To Believe or Not to Believe Your LLM
by: Yadkori, Yasin Abbasi, et al.
Published: (2024)
by: Yadkori, Yasin Abbasi, et al.
Published: (2024)
Pointwise confidence estimation in the non-linear $\ell^2$-regularized least squares
by: Kuzborskij, Ilja, et al.
Published: (2025)
by: Kuzborskij, Ilja, et al.
Published: (2025)
Low-rank bias, weight decay, and model merging in neural networks
by: Kuzborskij, Ilja, et al.
Published: (2025)
by: Kuzborskij, Ilja, et al.
Published: (2025)
Why Attend to Everything? Focus is the Key
by: Yao, Hengshuai, et al.
Published: (2026)
by: Yao, Hengshuai, et al.
Published: (2026)
Signal-Adaptive Trust Regions for Gradient-Free Optimization of Recurrent Spiking Neural Networks
by: Li, Jinhao, et al.
Published: (2026)
by: Li, Jinhao, et al.
Published: (2026)
Are Your Reasoning Models Reasoning or Guessing? A Mechanistic Analysis of Hierarchical Reasoning Models
by: Ren, Zirui, et al.
Published: (2026)
by: Ren, Zirui, et al.
Published: (2026)
Mitigating LLM Hallucinations via Conformal Abstention
by: Yadkori, Yasin Abbasi, et al.
Published: (2024)
by: Yadkori, Yasin Abbasi, et al.
Published: (2024)
A Hierarchical Language Model For Interpretable Graph Reasoning
by: Khurana, Sambhav, et al.
Published: (2024)
by: Khurana, Sambhav, et al.
Published: (2024)
Multi-Path Collaborative Reasoning via Reinforcement Learning
by: Lv, Jindi, et al.
Published: (2025)
by: Lv, Jindi, et al.
Published: (2025)
Hierarchical Reasoning Models: Perspectives and Misconceptions
by: Ge, Renee, et al.
Published: (2025)
by: Ge, Renee, et al.
Published: (2025)
CHARMS: A Cognitive Hierarchical Agent for Reasoning and Motion Stylization in Autonomous Driving
by: Wang, Jingyi, et al.
Published: (2025)
by: Wang, Jingyi, et al.
Published: (2025)
Scaling up Energy-Aware Multi-Agent Reinforcement Learning for Mission-Oriented Drone Networks with Individual Reward
by: Li, Changling, et al.
Published: (2026)
by: Li, Changling, et al.
Published: (2026)
TimeOmni-1: Incentivizing Complex Reasoning with Time Series in Large Language Models
by: Guan, Tong, et al.
Published: (2025)
by: Guan, Tong, et al.
Published: (2025)
Cog-Rethinker: Hierarchical Metacognitive Reinforcement Learning for LLM Reasoning
by: Sun, Zexu, et al.
Published: (2025)
by: Sun, Zexu, et al.
Published: (2025)
Native Reasoning Models: Training Language Models to Reason on Unverifiable Data
by: Wang, Yuanfu, et al.
Published: (2026)
by: Wang, Yuanfu, et al.
Published: (2026)
Divide-Then-Rule: A Cluster-Driven Hierarchical Interpolator for Attribute-Missing Graphs
by: Hu, Yaowen, et al.
Published: (2025)
by: Hu, Yaowen, et al.
Published: (2025)
Learning Discrete Concepts in Latent Hierarchical Models
by: Kong, Lingjing, et al.
Published: (2024)
by: Kong, Lingjing, et al.
Published: (2024)
Hierarchical Reinforcement Learning for Swarm Confrontation with High Uncertainty
by: Wu, Qizhen, et al.
Published: (2024)
by: Wu, Qizhen, et al.
Published: (2024)
Learning Cell-Aware Hierarchical Multi-Modal Representations for Robust Molecular Modeling
by: Li, Mengran, et al.
Published: (2025)
by: Li, Mengran, et al.
Published: (2025)
REMA: A Unified Reasoning Manifold Framework for Interpreting Large Language Model
by: Li, Bo, et al.
Published: (2025)
by: Li, Bo, et al.
Published: (2025)
Look Globally and Reason: Two-stage Path Reasoning over Sparse Knowledge Graphs
by: Guan, Saiping, et al.
Published: (2024)
by: Guan, Saiping, et al.
Published: (2024)
WaveHiTS: Wavelet-Enhanced Hierarchical Time Series Modeling for Wind Direction Nowcasting in Eastern Inner Mongolia
by: Shu, Hailong, et al.
Published: (2025)
by: Shu, Hailong, et al.
Published: (2025)
HISA: Efficient Hierarchical Indexing for Fine-Grained Sparse Attention
by: Xu, Yufei, et al.
Published: (2026)
by: Xu, Yufei, et al.
Published: (2026)
Interaction Locality in Hierarchical Recursive Reasoning
by: Miyanishi, Yosuke, et al.
Published: (2026)
by: Miyanishi, Yosuke, et al.
Published: (2026)
Best of both worlds: Stochastic & adversarial best-arm identification
by: Abbasi-Yadkori, Yasin, et al.
Published: (2026)
by: Abbasi-Yadkori, Yasin, et al.
Published: (2026)
Locally Interpretable Individualized Treatment Rules for Black-Box Decision Models
by: Charvadeh, Yasin Khadem, et al.
Published: (2026)
by: Charvadeh, Yasin Khadem, et al.
Published: (2026)
Decomposing and Measuring Evaluation Awareness
by: Li, Changling, et al.
Published: (2026)
by: Li, Changling, et al.
Published: (2026)
Lightning OPD: Efficient Post-Training for Large Reasoning Models with Offline On-Policy Distillation
by: Wu, Yecheng, et al.
Published: (2026)
by: Wu, Yecheng, et al.
Published: (2026)
The Curse of Depth in Large Language Models
by: Sun, Wenfang, et al.
Published: (2025)
by: Sun, Wenfang, et al.
Published: (2025)
ROER: Regularized Optimal Experience Replay
by: Li, Changling, et al.
Published: (2024)
by: Li, Changling, et al.
Published: (2024)
HER: Human-like Reasoning and Reinforcement Learning for LLM Role-playing
by: Du, Chengyu, et al.
Published: (2026)
by: Du, Chengyu, et al.
Published: (2026)
A Hierarchical Language Model with Predictable Scaling Laws and Provable Benefits of Reasoning
by: Gaitonde, Jason, et al.
Published: (2026)
by: Gaitonde, Jason, et al.
Published: (2026)
Reasoning Language Model Inference Serving Unveiled: An Empirical Study
by: Li, Qi, et al.
Published: (2025)
by: Li, Qi, et al.
Published: (2025)
A Discordance-Aware Multimodal Framework with Multi-Agent Clinical Reasoning
by: Ahadian, Pegah, et al.
Published: (2026)
by: Ahadian, Pegah, et al.
Published: (2026)
Train at Moving Edge: Online-Verified Prompt Selection for Efficient RL Training of Large Reasoning Model
by: Wu, Jiahao, et al.
Published: (2026)
by: Wu, Jiahao, et al.
Published: (2026)
No Free Lunch: Rethinking Internal Feedback for LLM Reasoning
by: Zhang, Yanzhi, et al.
Published: (2025)
by: Zhang, Yanzhi, et al.
Published: (2025)
Can LLMs Learn to Reason Robustly under Noisy Supervision?
by: Yang, Shenzhi, et al.
Published: (2026)
by: Yang, Shenzhi, et al.
Published: (2026)
HC-GAE: The Hierarchical Cluster-based Graph Auto-Encoder for Graph Representation Learning
by: Xu, Zhuo, et al.
Published: (2024)
by: Xu, Zhuo, et al.
Published: (2024)
TimeMKG: Knowledge-Infused Causal Reasoning for Multivariate Time Series Modeling
by: Sun, Yifei, et al.
Published: (2025)
by: Sun, Yifei, et al.
Published: (2025)
Efficient and Interpretable Traffic Destination Prediction using Explainable Boosting Machines
by: Yousif, Yasin, et al.
Published: (2024)
by: Yousif, Yasin, et al.
Published: (2024)
Similar Items
-
To Believe or Not to Believe Your LLM
by: Yadkori, Yasin Abbasi, et al.
Published: (2024) -
Pointwise confidence estimation in the non-linear $\ell^2$-regularized least squares
by: Kuzborskij, Ilja, et al.
Published: (2025) -
Low-rank bias, weight decay, and model merging in neural networks
by: Kuzborskij, Ilja, et al.
Published: (2025) -
Why Attend to Everything? Focus is the Key
by: Yao, Hengshuai, et al.
Published: (2026) -
Signal-Adaptive Trust Regions for Gradient-Free Optimization of Recurrent Spiking Neural Networks
by: Li, Jinhao, et al.
Published: (2026)