Reconsidering Overthinking: Penalizing Internal and External Redundancy in CoT Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hong, Jialiang, Zhen, Taihang, Chen, Kai, Liu, Jiaheng, Feng, Junlan, Zhu, Wenpeng, Huo, Jing, Gao, Yang, Wang, Depeng, Wan, Haitao, Yang, Xi, Wang, Boyan, Meng, Fanyu, Zhang, Yuyao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Thinking-Based Non-Thinking: Solving the Reward Hacking Problem in Training Hybrid Reasoning Models via Reinforcement Learning
von: Gan, Siyuan, et al.
Veröffentlicht: (2026)
von: Gan, Siyuan, et al.
Veröffentlicht: (2026)
From Explicit CoT to Implicit CoT: Learning to Internalize CoT Step by Step
von: Deng, Yuntian, et al.
Veröffentlicht: (2024)
von: Deng, Yuntian, et al.
Veröffentlicht: (2024)
LayerCraft: Enhancing Text-to-Image Generation with CoT Reasoning and Layered Object Integration
von: Zhang, Yuyao, et al.
Veröffentlicht: (2025)
von: Zhang, Yuyao, et al.
Veröffentlicht: (2025)
Efficient Reinforcement Learning with Semantic and Token Entropy for LLM Reasoning
von: Cao, Hongye, et al.
Veröffentlicht: (2025)
von: Cao, Hongye, et al.
Veröffentlicht: (2025)
Innate Reasoning is Not Enough: In-Context Learning Enhances Reasoning Large Language Models with Less Overthinking
von: Ge, Yuyao, et al.
Veröffentlicht: (2025)
von: Ge, Yuyao, et al.
Veröffentlicht: (2025)
Investigating CoT Monitorability in Large Reasoning Models
von: Yang, Shu, et al.
Veröffentlicht: (2025)
von: Yang, Shu, et al.
Veröffentlicht: (2025)
CoT Vectors: Transferring and Probing the Reasoning Mechanisms of LLMs
von: Li, Li, et al.
Veröffentlicht: (2025)
von: Li, Li, et al.
Veröffentlicht: (2025)
GRASP: Grounded CoT Reasoning with Dual-Stage Optimization for Multimodal Sarcasm Target Identification
von: Wan, Faxian, et al.
Veröffentlicht: (2026)
von: Wan, Faxian, et al.
Veröffentlicht: (2026)
ToxiFrench: Benchmarking and Enhancing Language Models via CoT Fine-Tuning for French Toxicity Detection
von: Delaval, Axel, et al.
Veröffentlicht: (2025)
von: Delaval, Axel, et al.
Veröffentlicht: (2025)
Parrot: A Training Pipeline Enhances Both Program CoT and Natural Language CoT for Reasoning
von: Jin, Senjie, et al.
Veröffentlicht: (2025)
von: Jin, Senjie, et al.
Veröffentlicht: (2025)
SafeDialBench: A Fine-Grained Safety Evaluation Benchmark for Large Language Models in Multi-Turn Dialogues with Diverse Jailbreak Attacks
von: Cao, Hongye, et al.
Veröffentlicht: (2025)
von: Cao, Hongye, et al.
Veröffentlicht: (2025)
THOUGHTTERMINATOR: Benchmarking, Calibrating, and Mitigating Overthinking in Reasoning Models
von: Pu, Xiao, et al.
Veröffentlicht: (2025)
von: Pu, Xiao, et al.
Veröffentlicht: (2025)
The First Impression Problem: Internal Bias Triggers Overthinking in Reasoning Models
von: Dang, Renfei, et al.
Veröffentlicht: (2025)
von: Dang, Renfei, et al.
Veröffentlicht: (2025)
How Likely Do LLMs with CoT Mimic Human Reasoning?
von: Bao, Guangsheng, et al.
Veröffentlicht: (2024)
von: Bao, Guangsheng, et al.
Veröffentlicht: (2024)
Think-with-Rubrics: From External Evaluator to Internal Reasoning Guidance
von: Yu, Jiachen, et al.
Veröffentlicht: (2026)
von: Yu, Jiachen, et al.
Veröffentlicht: (2026)
Hindsight Hint Distillation: Scaffolded Reasoning for SWE Agents from CoT-free Answers
von: Wang, Shengjie, et al.
Veröffentlicht: (2026)
von: Wang, Shengjie, et al.
Veröffentlicht: (2026)
APR: Penalizing Structural Redundancy in Large Reasoning Models via Anchor-based Process Rewards
von: Chang, Kaiyan, et al.
Veröffentlicht: (2026)
von: Chang, Kaiyan, et al.
Veröffentlicht: (2026)
R-CoT: A Reasoning-Layer Watermark via Redundant Chain-of-Thought in Large Language Models
von: Zhang, Ziming, et al.
Veröffentlicht: (2026)
von: Zhang, Ziming, et al.
Veröffentlicht: (2026)
Deconstructing Long Chain-of-Thought: A Structured Reasoning Optimization Framework for Long CoT Distillation
von: Luo, Yijia, et al.
Veröffentlicht: (2025)
von: Luo, Yijia, et al.
Veröffentlicht: (2025)
Mitigating Overthinking through Reasoning Shaping
von: Song, Feifan, et al.
Veröffentlicht: (2025)
von: Song, Feifan, et al.
Veröffentlicht: (2025)
CoT is Not the Chain of Truth: An Empirical Internal Analysis of Reasoning LLMs for Fake News Generation
von: Tong, Zhao, et al.
Veröffentlicht: (2026)
von: Tong, Zhao, et al.
Veröffentlicht: (2026)
The Quest for Efficient Reasoning: A Data-Centric Benchmark to CoT Distillation
von: Zhang, Ruichen, et al.
Veröffentlicht: (2025)
von: Zhang, Ruichen, et al.
Veröffentlicht: (2025)
Empowering Lightweight MLLMs with Reasoning via Long CoT SFT
von: Ou, Linyu, et al.
Veröffentlicht: (2025)
von: Ou, Linyu, et al.
Veröffentlicht: (2025)
Reasoning as an Attack Surface: Adaptive Evolutionary CoT Jailbreaks for LLMs
von: Li, Jianan, et al.
Veröffentlicht: (2026)
von: Li, Jianan, et al.
Veröffentlicht: (2026)
Can Pruning Improve Reasoning? Revisiting Long-CoT Compression with Capability in Mind for Better Reasoning
von: Zhao, Shangziqi, et al.
Veröffentlicht: (2025)
von: Zhao, Shangziqi, et al.
Veröffentlicht: (2025)
CoT2-Meta: Budgeted Metacognitive Control for Test-Time Reasoning
von: Ma, Siyuan, et al.
Veröffentlicht: (2026)
von: Ma, Siyuan, et al.
Veröffentlicht: (2026)
CoT-Kinetics: A Theoretical Modeling Assessing LRM Reasoning Process
von: Bi, Jinhe, et al.
Veröffentlicht: (2025)
von: Bi, Jinhe, et al.
Veröffentlicht: (2025)
Efficient Long CoT Reasoning in Small Language Models
von: Wang, Zhaoyang, et al.
Veröffentlicht: (2025)
von: Wang, Zhaoyang, et al.
Veröffentlicht: (2025)
Self-Verifying Reflection Helps Transformers with CoT Reasoning
von: Yu, Zhongwei, et al.
Veröffentlicht: (2025)
von: Yu, Zhongwei, et al.
Veröffentlicht: (2025)
To CoT or not to CoT? Chain-of-thought helps mainly on math and symbolic reasoning
von: Sprague, Zayne, et al.
Veröffentlicht: (2024)
von: Sprague, Zayne, et al.
Veröffentlicht: (2024)
Distribution-Aligned Sequence Distillation for Superior Long-CoT Reasoning
von: Yan, Shaotian, et al.
Veröffentlicht: (2026)
von: Yan, Shaotian, et al.
Veröffentlicht: (2026)
CoT-Valve: Length-Compressible Chain-of-Thought Tuning
von: Ma, Xinyin, et al.
Veröffentlicht: (2025)
von: Ma, Xinyin, et al.
Veröffentlicht: (2025)
ImageGen-CoT: Enhancing Text-to-Image In-context Learning with Chain-of-Thought Reasoning
von: Liao, Jiaqi, et al.
Veröffentlicht: (2025)
von: Liao, Jiaqi, et al.
Veröffentlicht: (2025)
Beyond Isolated Capabilities: Bridging Long CoT Reasoning and Long-Context Understanding
von: Wang, Yifei
Veröffentlicht: (2025)
von: Wang, Yifei
Veröffentlicht: (2025)
Zebra-CoT: A Dataset for Interleaved Vision Language Reasoning
von: Li, Ang, et al.
Veröffentlicht: (2025)
von: Li, Ang, et al.
Veröffentlicht: (2025)
ReasonGen-R1: CoT for Autoregressive Image generation models through SFT and RL
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
Batch Prompting Suppresses Overthinking Reasoning Under Constraint: How Batch Prompting Suppresses Overthinking in Reasoning Models
von: Srivastava, Saurabh, et al.
Veröffentlicht: (2025)
von: Srivastava, Saurabh, et al.
Veröffentlicht: (2025)
Exploring the Limitations of Mamba in COPY and CoT Reasoning
von: Ren, Ruifeng, et al.
Veröffentlicht: (2024)
von: Ren, Ruifeng, et al.
Veröffentlicht: (2024)
Evaluating LLM Reasoning Beyond Correctness and CoT
von: Abbasloo, Soheil
Veröffentlicht: (2025)
von: Abbasloo, Soheil
Veröffentlicht: (2025)
Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models
von: Sui, Yang, et al.
Veröffentlicht: (2025)
von: Sui, Yang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Thinking-Based Non-Thinking: Solving the Reward Hacking Problem in Training Hybrid Reasoning Models via Reinforcement Learning
von: Gan, Siyuan, et al.
Veröffentlicht: (2026) -
From Explicit CoT to Implicit CoT: Learning to Internalize CoT Step by Step
von: Deng, Yuntian, et al.
Veröffentlicht: (2024) -
LayerCraft: Enhancing Text-to-Image Generation with CoT Reasoning and Layered Object Integration
von: Zhang, Yuyao, et al.
Veröffentlicht: (2025) -
Efficient Reinforcement Learning with Semantic and Token Entropy for LLM Reasoning
von: Cao, Hongye, et al.
Veröffentlicht: (2025) -
Innate Reasoning is Not Enough: In-Context Learning Enhances Reasoning Large Language Models with Less Overthinking
von: Ge, Yuyao, et al.
Veröffentlicht: (2025)