A State-Transition Framework for Efficient LLM Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Liang, Zhao, Yu, Wang, Longyue, Shi, Tianqi, Luo, Weihua, Zhang, Kaifu, Su, Jinsong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Marco-o1 v2: Towards Widening The Distillation Bottleneck for Reasoning Models
von: Yin, Huifeng, et al.
Veröffentlicht: (2025)
von: Yin, Huifeng, et al.
Veröffentlicht: (2025)
Marco-MoE: Open Multilingual Mixture-of-Expert Language Models with Efficient Upcycling
von: Jiang, Fan, et al.
Veröffentlicht: (2026)
von: Jiang, Fan, et al.
Veröffentlicht: (2026)
Finding the Translation Switch: Discovering and Exploiting the Task-Initiation Features in LLMs
von: Wu, Xinwei, et al.
Veröffentlicht: (2026)
von: Wu, Xinwei, et al.
Veröffentlicht: (2026)
ComplexMCP: Evaluation of LLM Agents in Dynamic, Interdependent, and Large-Scale Tool Sandbox
von: Li, Yuanyang, et al.
Veröffentlicht: (2026)
von: Li, Yuanyang, et al.
Veröffentlicht: (2026)
Building Decision Making Models Through Language Model Regime
von: Zhang, Yu, et al.
Veröffentlicht: (2024)
von: Zhang, Yu, et al.
Veröffentlicht: (2024)
HSCodeComp: A Realistic and Expert-level Benchmark for Deep Search Agents in Hierarchical Rule Application
von: Yang, Yiqian, et al.
Veröffentlicht: (2025)
von: Yang, Yiqian, et al.
Veröffentlicht: (2025)
VIDA: A dataset for Visually Dependent Ambiguity in Multimodal Machine Translation
von: Pan, Jingheng, et al.
Veröffentlicht: (2026)
von: Pan, Jingheng, et al.
Veröffentlicht: (2026)
Difficulty-Estimated Policy Optimization
von: Zhao, Yu, et al.
Veröffentlicht: (2026)
von: Zhao, Yu, et al.
Veröffentlicht: (2026)
From Insight to Action: A Novel Framework for Interpretability-Guided Data Selection in Large Language Models
von: Shi, Ling, et al.
Veröffentlicht: (2026)
von: Shi, Ling, et al.
Veröffentlicht: (2026)
Syn-Diag: An LLM-based Synergistic Framework for Generalizable Few-shot Fault Diagnosis on the Edge
von: Jia, Zijun, et al.
Veröffentlicht: (2025)
von: Jia, Zijun, et al.
Veröffentlicht: (2025)
Marco-o1: Towards Open Reasoning Models for Open-Ended Solutions
von: Zhao, Yu, et al.
Veröffentlicht: (2024)
von: Zhao, Yu, et al.
Veröffentlicht: (2024)
Marco DeepResearch: Unlocking Efficient Deep Research Agents via Verification-Centric Design
von: Zhu, Bin, et al.
Veröffentlicht: (2026)
von: Zhu, Bin, et al.
Veröffentlicht: (2026)
CaptchaMind: Training CAPTCHA Solvers via Reinforcement Learning with Explicit Reasoning Supervision
von: Wang, Pengcheng, et al.
Veröffentlicht: (2026)
von: Wang, Pengcheng, et al.
Veröffentlicht: (2026)
Distilling LLM Reasoning into Graph of Concept Predictors
von: Yu, Ziyang, et al.
Veröffentlicht: (2026)
von: Yu, Ziyang, et al.
Veröffentlicht: (2026)
Logical Phase Transitions: Understanding Collapse in LLM Logical Reasoning
von: Zhang, Xinglang, et al.
Veröffentlicht: (2026)
von: Zhang, Xinglang, et al.
Veröffentlicht: (2026)
A Multi-Agent Framework with Automated Decision Rule Optimization for Cross-Domain Misinformation Detection
von: Li, Hui, et al.
Veröffentlicht: (2025)
von: Li, Hui, et al.
Veröffentlicht: (2025)
Advancing Tool-Augmented Large Language Models: Integrating Insights from Errors in Inference Trees
von: Chen, Sijia, et al.
Veröffentlicht: (2024)
von: Chen, Sijia, et al.
Veröffentlicht: (2024)
VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
von: Li, Yunxin, et al.
Veröffentlicht: (2024)
von: Li, Yunxin, et al.
Veröffentlicht: (2024)
RA-Rec: An Efficient ID Representation Alignment Framework for LLM-based Recommendation
von: Yu, Xiaohan, et al.
Veröffentlicht: (2024)
von: Yu, Xiaohan, et al.
Veröffentlicht: (2024)
LUMIR: an LLM-Driven Unified Agent Framework for Multi-task Infrared Spectroscopy Reasoning
von: Xie, Zujie, et al.
Veröffentlicht: (2025)
von: Xie, Zujie, et al.
Veröffentlicht: (2025)
PMMT: Preference Alignment in Multilingual Machine Translation via LLM Distillation
von: Sun, Shuqiao, et al.
Veröffentlicht: (2024)
von: Sun, Shuqiao, et al.
Veröffentlicht: (2024)
SABER: Switchable and Balanced Training for Efficient LLM Reasoning
von: Zhao, Kai, et al.
Veröffentlicht: (2025)
von: Zhao, Kai, et al.
Veröffentlicht: (2025)
BearLLM: A Prior Knowledge-Enhanced Bearing Health Management Framework with Unified Vibration Signal Representation
von: Peng, Haotian, et al.
Veröffentlicht: (2024)
von: Peng, Haotian, et al.
Veröffentlicht: (2024)
Multimodal Tabular Reasoning with Privileged Structured Information
von: Jiang, Jun-Peng, et al.
Veröffentlicht: (2025)
von: Jiang, Jun-Peng, et al.
Veröffentlicht: (2025)
Beyond ReAct: A Planner-Centric Framework for Complex Tool-Augmented LLM Reasoning
von: Wei, Xiaolong, et al.
Veröffentlicht: (2025)
von: Wei, Xiaolong, et al.
Veröffentlicht: (2025)
Accordion-Thinking: Self-Regulated Step Summaries for Efficient and Readable LLM Reasoning
von: Yang, Zhicheng, et al.
Veröffentlicht: (2026)
von: Yang, Zhicheng, et al.
Veröffentlicht: (2026)
MemOCR: Layout-Aware Visual Memory for Efficient Long-Horizon Reasoning
von: Shi, Yaorui, et al.
Veröffentlicht: (2026)
von: Shi, Yaorui, et al.
Veröffentlicht: (2026)
Dynamic Parallel Tree Search for Efficient LLM Reasoning
von: Ding, Yifu, et al.
Veröffentlicht: (2025)
von: Ding, Yifu, et al.
Veröffentlicht: (2025)
Beyond Safe Answers: A Benchmark for Evaluating True Risk Awareness in Large Reasoning Models
von: Zheng, Baihui, et al.
Veröffentlicht: (2025)
von: Zheng, Baihui, et al.
Veröffentlicht: (2025)
STARec: An Efficient Agent Framework for Recommender Systems via Autonomous Deliberate Reasoning
von: Wu, Chenghao, et al.
Veröffentlicht: (2025)
von: Wu, Chenghao, et al.
Veröffentlicht: (2025)
Closing Reasoning Gaps in Clinical Agents with Differential Reasoning Learning
von: Liu, Jinsong, et al.
Veröffentlicht: (2026)
von: Liu, Jinsong, et al.
Veröffentlicht: (2026)
Tracing LLM Reasoning Processes with Strategic Games: A Framework for Planning, Revision, and Resource-Constrained Decision Making
von: Yuan, Xiaopeng, et al.
Veröffentlicht: (2025)
von: Yuan, Xiaopeng, et al.
Veröffentlicht: (2025)
A Smart Multimodal Healthcare Copilot with Powerful LLM Reasoning
von: Zhao, Xuejiao, et al.
Veröffentlicht: (2025)
von: Zhao, Xuejiao, et al.
Veröffentlicht: (2025)
MMCR: Advancing Visual Language Model in Multimodal Multi-Turn Contextual Reasoning
von: Yan, Dawei, et al.
Veröffentlicht: (2025)
von: Yan, Dawei, et al.
Veröffentlicht: (2025)
Efficient Reinforcement Learning with Semantic and Token Entropy for LLM Reasoning
von: Cao, Hongye, et al.
Veröffentlicht: (2025)
von: Cao, Hongye, et al.
Veröffentlicht: (2025)
Ovis: Structural Embedding Alignment for Multimodal Large Language Model
von: Lu, Shiyin, et al.
Veröffentlicht: (2024)
von: Lu, Shiyin, et al.
Veröffentlicht: (2024)
Searching Meta Reasoning Skeleton to Guide LLM Reasoning
von: Zhang, Ziying, et al.
Veröffentlicht: (2025)
von: Zhang, Ziying, et al.
Veröffentlicht: (2025)
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning
von: Wang, Shaojie, et al.
Veröffentlicht: (2026)
von: Wang, Shaojie, et al.
Veröffentlicht: (2026)
From Implicit to Explicit: Token-Efficient Logical Supervision for Mathematical Reasoning in LLMs
von: Wang, Shaojie, et al.
Veröffentlicht: (2026)
von: Wang, Shaojie, et al.
Veröffentlicht: (2026)
Deconstructing Long Chain-of-Thought: A Structured Reasoning Optimization Framework for Long CoT Distillation
von: Luo, Yijia, et al.
Veröffentlicht: (2025)
von: Luo, Yijia, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Marco-o1 v2: Towards Widening The Distillation Bottleneck for Reasoning Models
von: Yin, Huifeng, et al.
Veröffentlicht: (2025) -
Marco-MoE: Open Multilingual Mixture-of-Expert Language Models with Efficient Upcycling
von: Jiang, Fan, et al.
Veröffentlicht: (2026) -
Finding the Translation Switch: Discovering and Exploiting the Task-Initiation Features in LLMs
von: Wu, Xinwei, et al.
Veröffentlicht: (2026) -
ComplexMCP: Evaluation of LLM Agents in Dynamic, Interdependent, and Large-Scale Tool Sandbox
von: Li, Yuanyang, et al.
Veröffentlicht: (2026) -
Building Decision Making Models Through Language Model Regime
von: Zhang, Yu, et al.
Veröffentlicht: (2024)