Exploring and Exploiting the Inherent Efficiency within Large Reasoning Models for Self-Guided Efficiency Enhancement

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Zhao, Weixiang, Guo, Jiahe, Deng, Yang, Sui, Xingyu, Hu, Yulin, Zhao, Yanyan, Che, Wanxiang, Qin, Bing, Chua, Tat-Seng, Liu, Ting
Natura: Preprint
Pubblicazione: 2025
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866918063128969216
author Zhao, Weixiang
Guo, Jiahe
Deng, Yang
Sui, Xingyu
Hu, Yulin
Zhao, Yanyan
Che, Wanxiang
Qin, Bing
Chua, Tat-Seng
Liu, Ting
author_facet Zhao, Weixiang
Guo, Jiahe
Deng, Yang
Sui, Xingyu
Hu, Yulin
Zhao, Yanyan
Che, Wanxiang
Qin, Bing
Chua, Tat-Seng
Liu, Ting
contents Recent advancements in large reasoning models (LRMs) have significantly enhanced language models' capabilities in complex problem-solving by emulating human-like deliberative thinking. However, these models often exhibit overthinking (i.e., the generation of unnecessarily verbose and redundant content), which hinders efficiency and inflates inference cost. In this work, we explore the representational and behavioral origins of this inefficiency, revealing that LRMs inherently possess the capacity for more concise reasoning. Empirical analyses show that correct reasoning paths vary significantly in length, and the shortest correct responses often suffice, indicating untapped efficiency potential. Exploiting these findings, we propose two lightweight methods to enhance LRM efficiency. First, we introduce Efficiency Steering, a training-free activation steering technique that modulates reasoning behavior via a single direction in the model's representation space. Second, we develop Self-Rewarded Efficiency RL, a reinforcement learning framework that dynamically balances task accuracy and brevity by rewarding concise correct solutions. Extensive experiments on seven LRM backbones across multiple mathematical reasoning benchmarks demonstrate that our methods significantly reduce reasoning length while preserving or improving task performance. Our results highlight that reasoning efficiency can be improved by leveraging and guiding the intrinsic capabilities of existing models in a self-guided manner.
format Preprint
id arxiv_https___arxiv_org_abs_2506_15647
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Exploring and Exploiting the Inherent Efficiency within Large Reasoning Models for Self-Guided Efficiency Enhancement
Zhao, Weixiang
Guo, Jiahe
Deng, Yang
Sui, Xingyu
Hu, Yulin
Zhao, Yanyan
Che, Wanxiang
Qin, Bing
Chua, Tat-Seng
Liu, Ting
Artificial Intelligence
Recent advancements in large reasoning models (LRMs) have significantly enhanced language models' capabilities in complex problem-solving by emulating human-like deliberative thinking. However, these models often exhibit overthinking (i.e., the generation of unnecessarily verbose and redundant content), which hinders efficiency and inflates inference cost. In this work, we explore the representational and behavioral origins of this inefficiency, revealing that LRMs inherently possess the capacity for more concise reasoning. Empirical analyses show that correct reasoning paths vary significantly in length, and the shortest correct responses often suffice, indicating untapped efficiency potential. Exploiting these findings, we propose two lightweight methods to enhance LRM efficiency. First, we introduce Efficiency Steering, a training-free activation steering technique that modulates reasoning behavior via a single direction in the model's representation space. Second, we develop Self-Rewarded Efficiency RL, a reinforcement learning framework that dynamically balances task accuracy and brevity by rewarding concise correct solutions. Extensive experiments on seven LRM backbones across multiple mathematical reasoning benchmarks demonstrate that our methods significantly reduce reasoning length while preserving or improving task performance. Our results highlight that reasoning efficiency can be improved by leveraging and guiding the intrinsic capabilities of existing models in a self-guided manner.
title Exploring and Exploiting the Inherent Efficiency within Large Reasoning Models for Self-Guided Efficiency Enhancement
topic Artificial Intelligence
url https://arxiv.org/abs/2506.15647