When More is Less: Understanding Chain-of-Thought Length in LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Wu, Yuyang, Wang, Yifei, Ye, Ziyu, Du, Tianqi, Jegelka, Stefanie, Wang, Yisen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Theoretical Understanding of Self-Correction through In-context Alignment
di: Wang, Yifei, et al.
Pubblicazione: (2024)
di: Wang, Yifei, et al.
Pubblicazione: (2024)
Understanding the Role of Equivariance in Self-supervised Learning
di: Wang, Yifei, et al.
Pubblicazione: (2024)
di: Wang, Yifei, et al.
Pubblicazione: (2024)
LANPO: Bootstrapping Language and Numerical Feedback for Reinforcement Learning in LLMs
di: Li, Ang, et al.
Pubblicazione: (2025)
di: Li, Ang, et al.
Pubblicazione: (2025)
G1: Teaching LLMs to Reason on Graphs with Reinforcement Learning
di: Guo, Xiaojun, et al.
Pubblicazione: (2025)
di: Guo, Xiaojun, et al.
Pubblicazione: (2025)
Understanding Chain-of-Thought in LLMs through Information Theory
di: Ton, Jean-Francois, et al.
Pubblicazione: (2024)
di: Ton, Jean-Francois, et al.
Pubblicazione: (2024)
Less is More: Extreme Gradient Boost Rank-1 Adaption for Efficient Finetuning of LLMs
di: Zhang, Yifei, et al.
Pubblicazione: (2024)
di: Zhang, Yifei, et al.
Pubblicazione: (2024)
Long-Short Alignment for Effective Long-Context Modeling in LLMs
di: Du, Tianqi, et al.
Pubblicazione: (2025)
di: Du, Tianqi, et al.
Pubblicazione: (2025)
Explain Less, Understand More: Jargon Detection via Personalized Parameter-Efficient Fine-tuning
di: Wu, Bohao, et al.
Pubblicazione: (2025)
di: Wu, Bohao, et al.
Pubblicazione: (2025)
How to Craft Backdoors with Unlabeled Data Alone?
di: Wang, Yifei, et al.
Pubblicazione: (2024)
di: Wang, Yifei, et al.
Pubblicazione: (2024)
Stepwise Penalization for Length-Efficient Chain-of-Thought Reasoning
di: Li, Xintong, et al.
Pubblicazione: (2026)
di: Li, Xintong, et al.
Pubblicazione: (2026)
Understanding Reasoning in Chain-of-Thought from the Hopfieldian View
di: Hu, Lijie, et al.
Pubblicazione: (2024)
di: Hu, Lijie, et al.
Pubblicazione: (2024)
When Chain-of-Thought Fails, the Solution Hides in the Hidden States
di: Mehrafarin, Houman, et al.
Pubblicazione: (2026)
di: Mehrafarin, Houman, et al.
Pubblicazione: (2026)
More Thinking, More Bias: Length-Driven Position Bias in Reasoning Models
di: Wang, Xiao
Pubblicazione: (2026)
di: Wang, Xiao
Pubblicazione: (2026)
Less is More for Improving Automatic Evaluation of Factual Consistency
di: Wang, Tong, et al.
Pubblicazione: (2024)
di: Wang, Tong, et al.
Pubblicazione: (2024)
Diffusion of Thoughts: Chain-of-Thought Reasoning in Diffusion Language Models
di: Ye, Jiacheng, et al.
Pubblicazione: (2024)
di: Ye, Jiacheng, et al.
Pubblicazione: (2024)
Think When You Need: Self-Adaptive Chain-of-Thought Learning
di: Yang, Junjie, et al.
Pubblicazione: (2025)
di: Yang, Junjie, et al.
Pubblicazione: (2025)
Is Chain-of-Thought Reasoning of LLMs a Mirage? A Data Distribution Lens
di: Zhao, Chengshuai, et al.
Pubblicazione: (2025)
di: Zhao, Chengshuai, et al.
Pubblicazione: (2025)
Crystal-KV: Efficient KV Cache Management for Chain-of-Thought LLMs via Answer-First Principle
di: Wang, Zihan, et al.
Pubblicazione: (2026)
di: Wang, Zihan, et al.
Pubblicazione: (2026)
LIMR: Less is More for RL Scaling
di: Li, Xuefeng, et al.
Pubblicazione: (2025)
di: Li, Xuefeng, et al.
Pubblicazione: (2025)
Disambiguation-Centric Finetuning Makes Enterprise Tool-Calling LLMs More Realistic and Less Risky
di: Hathidara, Ashutosh, et al.
Pubblicazione: (2025)
di: Hathidara, Ashutosh, et al.
Pubblicazione: (2025)
Learn More, Forget Less: A Gradient-Aware Data Selection Approach for LLM
di: Liu, Yibai, et al.
Pubblicazione: (2025)
di: Liu, Yibai, et al.
Pubblicazione: (2025)
Less is More: Improving LLM Alignment via Preference Data Selection
di: Deng, Xun, et al.
Pubblicazione: (2025)
di: Deng, Xun, et al.
Pubblicazione: (2025)
Supernova: Achieving More with Less in Transformer Architectures
di: Tanase, Andrei-Valentin, et al.
Pubblicazione: (2025)
di: Tanase, Andrei-Valentin, et al.
Pubblicazione: (2025)
When Should LLMs Be Less Specific? Selective Abstraction for Reliable Long-Form Text Generation
di: Goren, Shani, et al.
Pubblicazione: (2026)
di: Goren, Shani, et al.
Pubblicazione: (2026)
A Theoretical Understanding of Chain-of-Thought: Coherent Reasoning and Error-Aware Demonstration
di: Cui, Yingqian, et al.
Pubblicazione: (2024)
di: Cui, Yingqian, et al.
Pubblicazione: (2024)
Scalable Chain of Thoughts via Elastic Reasoning
di: Xu, Yuhui, et al.
Pubblicazione: (2025)
di: Xu, Yuhui, et al.
Pubblicazione: (2025)
Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning
di: Yue, Murong, et al.
Pubblicazione: (2023)
di: Yue, Murong, et al.
Pubblicazione: (2023)
Evolving Demonstration Optimization for Chain-of-Thought Feature Transformation
di: Wang, Xinyuan, et al.
Pubblicazione: (2026)
di: Wang, Xinyuan, et al.
Pubblicazione: (2026)
Understanding and Improving Length Generalization in Hierarchical Sparse Attention Models
di: Leng, Jiaqi, et al.
Pubblicazione: (2025)
di: Leng, Jiaqi, et al.
Pubblicazione: (2025)
When Models Know More Than They Say: Probing Analogical Reasoning in LLMs
di: McGovern, Hope, et al.
Pubblicazione: (2026)
di: McGovern, Hope, et al.
Pubblicazione: (2026)
Is Chain-of-Thought Really Not Explainability? Chain-of-Thought Can Be Faithful without Hint Verbalization
di: Zaman, Kerem, et al.
Pubblicazione: (2025)
di: Zaman, Kerem, et al.
Pubblicazione: (2025)
Enhancing Generalization in Chain of Thought Reasoning for Smaller Models
di: Yin, Maxwell J., et al.
Pubblicazione: (2025)
di: Yin, Maxwell J., et al.
Pubblicazione: (2025)
Less is More: Understanding Word-level Textual Adversarial Attack via n-gram Frequency Descend
di: Lu, Ning, et al.
Pubblicazione: (2023)
di: Lu, Ning, et al.
Pubblicazione: (2023)
Fractured Chain-of-Thought Reasoning
di: Liao, Baohao, et al.
Pubblicazione: (2025)
di: Liao, Baohao, et al.
Pubblicazione: (2025)
Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
di: Wei, Zeming, et al.
Pubblicazione: (2023)
di: Wei, Zeming, et al.
Pubblicazione: (2023)
Beyond Interpretability: The Gains of Feature Monosemanticity on Model Robustness
di: Zhang, Qi, et al.
Pubblicazione: (2024)
di: Zhang, Qi, et al.
Pubblicazione: (2024)
Value-Guided Search for Efficient Chain-of-Thought Reasoning
di: Wang, Kaiwen, et al.
Pubblicazione: (2025)
di: Wang, Kaiwen, et al.
Pubblicazione: (2025)
A Formal Comparison Between Chain of Thought and Latent Thought
di: Xu, Kevin, et al.
Pubblicazione: (2025)
di: Xu, Kevin, et al.
Pubblicazione: (2025)
Less is More: Denoising Knowledge Graphs For Retrieval Augmented Generation
di: Zheng, Yilun, et al.
Pubblicazione: (2025)
di: Zheng, Yilun, et al.
Pubblicazione: (2025)
Less is More: Local Intrinsic Dimensions of Contextual Language Models
di: Ruppik, Benjamin Matthias, et al.
Pubblicazione: (2025)
di: Ruppik, Benjamin Matthias, et al.
Pubblicazione: (2025)
Documenti analoghi
-
A Theoretical Understanding of Self-Correction through In-context Alignment
di: Wang, Yifei, et al.
Pubblicazione: (2024) -
Understanding the Role of Equivariance in Self-supervised Learning
di: Wang, Yifei, et al.
Pubblicazione: (2024) -
LANPO: Bootstrapping Language and Numerical Feedback for Reinforcement Learning in LLMs
di: Li, Ang, et al.
Pubblicazione: (2025) -
G1: Teaching LLMs to Reason on Graphs with Reinforcement Learning
di: Guo, Xiaojun, et al.
Pubblicazione: (2025) -
Understanding Chain-of-Thought in LLMs through Information Theory
di: Ton, Jean-Francois, et al.
Pubblicazione: (2024)