ShorterBetter: Guiding Reasoning Models to Find Optimal Inference Length for Efficient Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yi, Jingyang, Wang, Jiazheng, Li, Sida |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Shorter but not Worse: Frugal Reasoning via Easy Samples as Length Regularizers in Math RLVR
von: Bounhar, Abdelaziz, et al.
Veröffentlicht: (2025)
von: Bounhar, Abdelaziz, et al.
Veröffentlicht: (2025)
LEAD: Length-Efficient Adaptive and Dynamic Reasoning for Large Language Models
von: Wei, Songtao, et al.
Veröffentlicht: (2026)
von: Wei, Songtao, et al.
Veröffentlicht: (2026)
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training
von: Wang, Chen, et al.
Veröffentlicht: (2026)
von: Wang, Chen, et al.
Veröffentlicht: (2026)
Anti-Length Shift: Dynamic Outlier Truncation for Training Efficient Reasoning Models
von: Wu, Wei, et al.
Veröffentlicht: (2026)
von: Wu, Wei, et al.
Veröffentlicht: (2026)
On the Optimal Reasoning Length for RL-Trained Language Models
von: Nohara, Daisuke, et al.
Veröffentlicht: (2026)
von: Nohara, Daisuke, et al.
Veröffentlicht: (2026)
Don't Overthink it. Preferring Shorter Thinking Chains for Improved LLM Reasoning
von: Hassid, Michael, et al.
Veröffentlicht: (2025)
von: Hassid, Michael, et al.
Veröffentlicht: (2025)
Leash: Adaptive Length Penalty and Reward Shaping for Efficient Large Reasoning Model
von: Li, Yanhao, et al.
Veröffentlicht: (2025)
von: Li, Yanhao, et al.
Veröffentlicht: (2025)
Boosting Inference with Guided Reasoning: Stochastic Exploration for Recursive Models
von: Corbett, Andrew, et al.
Veröffentlicht: (2026)
von: Corbett, Andrew, et al.
Veröffentlicht: (2026)
Pseudocode-Guided Structured Reasoning for Automating Reliable Inference in Vision-Language Models
von: Ni, Weicong, et al.
Veröffentlicht: (2026)
von: Ni, Weicong, et al.
Veröffentlicht: (2026)
Critical Thinking: Which Kinds of Complexity Govern Optimal Reasoning Length?
von: Lee, Celine, et al.
Veröffentlicht: (2025)
von: Lee, Celine, et al.
Veröffentlicht: (2025)
Towards Interpretable and Inference-Optimal COT Reasoning with Sparse Autoencoder-Guided Generation
von: Zhao, Daniel, et al.
Veröffentlicht: (2025)
von: Zhao, Daniel, et al.
Veröffentlicht: (2025)
Mixed Distillation Helps Smaller Language Model Better Reasoning
von: Li, Chenglin, et al.
Veröffentlicht: (2023)
von: Li, Chenglin, et al.
Veröffentlicht: (2023)
Stepwise Penalization for Length-Efficient Chain-of-Thought Reasoning
von: Li, Xintong, et al.
Veröffentlicht: (2026)
von: Li, Xintong, et al.
Veröffentlicht: (2026)
Learning to Self-Verify Makes Language Models Better Reasoners
von: Chen, Yuxin, et al.
Veröffentlicht: (2026)
von: Chen, Yuxin, et al.
Veröffentlicht: (2026)
Reasoning Models Better Express Their Confidence
von: Yoon, Dongkeun, et al.
Veröffentlicht: (2025)
von: Yoon, Dongkeun, et al.
Veröffentlicht: (2025)
EntroCut: Entropy-Guided Adaptive Truncation for Efficient Chain-of-Thought Reasoning in Small-scale Large Reasoning Models
von: Yan, Hongxi, et al.
Veröffentlicht: (2026)
von: Yan, Hongxi, et al.
Veröffentlicht: (2026)
An Empirical Study of LLM Reasoning Ability Under Strict Output Length Constraint
von: Sun, Yi, et al.
Veröffentlicht: (2025)
von: Sun, Yi, et al.
Veröffentlicht: (2025)
Entropy-Guided Data-Efficient Training for Multimodal Reasoning Reward Models
von: Yang, Shidong, et al.
Veröffentlicht: (2026)
von: Yang, Shidong, et al.
Veröffentlicht: (2026)
Done Is Better than Perfect: Unlocking Efficient Reasoning by Structured Multi-Turn Decomposition
von: Zeng, Zihao, et al.
Veröffentlicht: (2025)
von: Zeng, Zihao, et al.
Veröffentlicht: (2025)
Are Machines Better at Complex Reasoning? Unveiling Human-Machine Inference Gaps in Entailment Verification
von: Sanyal, Soumya, et al.
Veröffentlicht: (2024)
von: Sanyal, Soumya, et al.
Veröffentlicht: (2024)
Predictive Scheduling for Efficient Inference-Time Reasoning in Large Language Models
von: Brown, Katrina, et al.
Veröffentlicht: (2026)
von: Brown, Katrina, et al.
Veröffentlicht: (2026)
CoLD: Counterfactually-Guided Length Debiasing for Process Reward Models in Mathematical Reasoning
von: Zheng, Congmin, et al.
Veröffentlicht: (2025)
von: Zheng, Congmin, et al.
Veröffentlicht: (2025)
Optimizing Length Compression in Large Reasoning Models
von: Cheng, Zhengxiang, et al.
Veröffentlicht: (2025)
von: Cheng, Zhengxiang, et al.
Veröffentlicht: (2025)
ESTAR: Early-Stopping Token-Aware Reasoning For Efficient Inference
von: Wang, Junda, et al.
Veröffentlicht: (2026)
von: Wang, Junda, et al.
Veröffentlicht: (2026)
Timo: Towards Better Temporal Reasoning for Language Models
von: Su, Zhaochen, et al.
Veröffentlicht: (2024)
von: Su, Zhaochen, et al.
Veröffentlicht: (2024)
Trace Length is a Simple Uncertainty Signal in Reasoning Models
von: Devic, Siddartha, et al.
Veröffentlicht: (2025)
von: Devic, Siddartha, et al.
Veröffentlicht: (2025)
Efficient Reasoning via Reward Model
von: Wang, Yuhao, et al.
Veröffentlicht: (2025)
von: Wang, Yuhao, et al.
Veröffentlicht: (2025)
ThreadWeaver: Adaptive Threading for Efficient Parallel Reasoning in Language Models
von: Lian, Long, et al.
Veröffentlicht: (2025)
von: Lian, Long, et al.
Veröffentlicht: (2025)
Searching Meta Reasoning Skeleton to Guide LLM Reasoning
von: Zhang, Ziying, et al.
Veröffentlicht: (2025)
von: Zhang, Ziying, et al.
Veröffentlicht: (2025)
Hawkeye:Efficient Reasoning with Model Collaboration
von: She, Jianshu, et al.
Veröffentlicht: (2025)
von: She, Jianshu, et al.
Veröffentlicht: (2025)
Beyond Token Length: Step Pruner for Efficient and Accurate Reasoning in Large Language Models
von: Wu, Canhui, et al.
Veröffentlicht: (2025)
von: Wu, Canhui, et al.
Veröffentlicht: (2025)
From Table to Cell: Attention for Better Reasoning with TABALIGN
von: Kwok, Tung Sum Thomas, et al.
Veröffentlicht: (2026)
von: Kwok, Tung Sum Thomas, et al.
Veröffentlicht: (2026)
Experience-Guided Adaptation of Inference-Time Reasoning Strategies
von: Stein, Adam, et al.
Veröffentlicht: (2025)
von: Stein, Adam, et al.
Veröffentlicht: (2025)
A Theory for Length Generalization in Learning to Reason
von: Xiao, Changnan, et al.
Veröffentlicht: (2024)
von: Xiao, Changnan, et al.
Veröffentlicht: (2024)
The Impact of Reasoning Step Length on Large Language Models
von: Jin, Mingyu, et al.
Veröffentlicht: (2024)
von: Jin, Mingyu, et al.
Veröffentlicht: (2024)
RankGuide: Tensor-Rank-Guided Routing and Steering for Efficient Reasoning
von: Tian, Jiayi, et al.
Veröffentlicht: (2026)
von: Tian, Jiayi, et al.
Veröffentlicht: (2026)
Optimal Self-Consistency for Efficient Reasoning with Large Language Models
von: Feng, Austin, et al.
Veröffentlicht: (2025)
von: Feng, Austin, et al.
Veröffentlicht: (2025)
CausalEval: Towards Better Causal Reasoning in Language Models
von: Yu, Longxuan, et al.
Veröffentlicht: (2024)
von: Yu, Longxuan, et al.
Veröffentlicht: (2024)
Abstraction-of-Thought Makes Language Models Better Reasoners
von: Hong, Ruixin, et al.
Veröffentlicht: (2024)
von: Hong, Ruixin, et al.
Veröffentlicht: (2024)
Harnessing the Reasoning Economy: A Survey of Efficient Reasoning for Large Language Models
von: Wang, Rui, et al.
Veröffentlicht: (2025)
von: Wang, Rui, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Shorter but not Worse: Frugal Reasoning via Easy Samples as Length Regularizers in Math RLVR
von: Bounhar, Abdelaziz, et al.
Veröffentlicht: (2025) -
LEAD: Length-Efficient Adaptive and Dynamic Reasoning for Large Language Models
von: Wei, Songtao, et al.
Veröffentlicht: (2026) -
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training
von: Wang, Chen, et al.
Veröffentlicht: (2026) -
Anti-Length Shift: Dynamic Outlier Truncation for Training Efficient Reasoning Models
von: Wu, Wei, et al.
Veröffentlicht: (2026) -
On the Optimal Reasoning Length for RL-Trained Language Models
von: Nohara, Daisuke, et al.
Veröffentlicht: (2026)