Fractured Chain-of-Thought Reasoning
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Liao, Baohao, Dong, Hanze, Xu, Yuhui, Sahoo, Doyen, Monz, Christof, Li, Junnan, Xiong, Caiming |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Scalable Chain of Thoughts via Elastic Reasoning
par: Xu, Yuhui, et autres
Publié: (2025)
par: Xu, Yuhui, et autres
Publié: (2025)
Reward-Guided Speculative Decoding for Efficient LLM Reasoning
par: Liao, Baohao, et autres
Publié: (2025)
par: Liao, Baohao, et autres
Publié: (2025)
Self-Hinting Language Models Enhance Reinforcement Learning
par: Liao, Baohao, et autres
Publié: (2026)
par: Liao, Baohao, et autres
Publié: (2026)
3-in-1: 2D Rotary Adaptation for Efficient Finetuning, Efficient Batching and Composability
par: Liao, Baohao, et autres
Publié: (2024)
par: Liao, Baohao, et autres
Publié: (2024)
Automatic Curriculum Expert Iteration for Reliable LLM Reasoning
par: Zhao, Zirui, et autres
Publié: (2024)
par: Zhao, Zirui, et autres
Publié: (2024)
Reward Models Identify Consistency, Not Causality
par: Xu, Yuhui, et autres
Publié: (2025)
par: Xu, Yuhui, et autres
Publié: (2025)
A Minimalist Approach to LLM Reasoning: from Rejection Sampling to Reinforce
par: Xiong, Wei, et autres
Publié: (2025)
par: Xiong, Wei, et autres
Publié: (2025)
Reinforce-Ada: An Adaptive Sampling Framework under Non-linear RL Objectives
par: Xiong, Wei, et autres
Publié: (2025)
par: Xiong, Wei, et autres
Publié: (2025)
Is It a Free Lunch for Removing Outliers during Pretraining?
par: Liao, Baohao, et autres
Publié: (2024)
par: Liao, Baohao, et autres
Publié: (2024)
RLHF Workflow: From Reward Modeling to Online RLHF
par: Dong, Hanze, et autres
Publié: (2024)
par: Dong, Hanze, et autres
Publié: (2024)
Optimizing Chain-of-Thought Reasoners via Gradient Variance Minimization in Rejection Sampling and RL
par: Yao, Jiarui, et autres
Publié: (2025)
par: Yao, Jiarui, et autres
Publié: (2025)
LiveMathematicianBench: A Live Benchmark for Mathematician-Level Reasoning with Proof Sketches
par: He, Linyang, et autres
Publié: (2026)
par: He, Linyang, et autres
Publié: (2026)
ApiQ: Finetuning of 2-Bit Quantized Large Language Model
par: Liao, Baohao, et autres
Publié: (2024)
par: Liao, Baohao, et autres
Publié: (2024)
Analyzing the Evaluation of Cross-Lingual Knowledge Transfer in Multilingual Language Models
par: Rajaee, Sara, et autres
Publié: (2024)
par: Rajaee, Sara, et autres
Publié: (2024)
Disentangling the Roles of Target-Side Transfer and Regularization in Multilingual Machine Translation
par: Meng, Yan, et autres
Publié: (2024)
par: Meng, Yan, et autres
Publié: (2024)
ThinK: Thinner Key Cache by Query-Driven Pruning
par: Xu, Yuhui, et autres
Publié: (2024)
par: Xu, Yuhui, et autres
Publié: (2024)
Diffusion of Thoughts: Chain-of-Thought Reasoning in Diffusion Language Models
par: Ye, Jiacheng, et autres
Publié: (2024)
par: Ye, Jiacheng, et autres
Publié: (2024)
Stepwise Penalization for Length-Efficient Chain-of-Thought Reasoning
par: Li, Xintong, et autres
Publié: (2026)
par: Li, Xintong, et autres
Publié: (2026)
Enhancing Auto-regressive Chain-of-Thought through Loop-Aligned Reasoning
par: Yu, Qifan, et autres
Publié: (2025)
par: Yu, Qifan, et autres
Publié: (2025)
Understanding Reasoning in Chain-of-Thought from the Hopfieldian View
par: Hu, Lijie, et autres
Publié: (2024)
par: Hu, Lijie, et autres
Publié: (2024)
Chain-of-Thought Reasoning In The Wild Is Not Always Faithful
par: Arcuschin, Iván, et autres
Publié: (2025)
par: Arcuschin, Iván, et autres
Publié: (2025)
Compositional Reasoning with Transformers, RNNs, and Chain of Thought
par: Yehudai, Gilad, et autres
Publié: (2025)
par: Yehudai, Gilad, et autres
Publié: (2025)
Long Chain-of-Thought Reasoning Across Languages
par: Barua, Josh, et autres
Publié: (2025)
par: Barua, Josh, et autres
Publié: (2025)
FinChain: A Symbolic Benchmark for Verifiable Chain-of-Thought Financial Reasoning
par: Xie, Zhuohan, et autres
Publié: (2025)
par: Xie, Zhuohan, et autres
Publié: (2025)
Value-Guided Search for Efficient Chain-of-Thought Reasoning
par: Wang, Kaiwen, et autres
Publié: (2025)
par: Wang, Kaiwen, et autres
Publié: (2025)
Enhancing Generalization in Chain of Thought Reasoning for Smaller Models
par: Yin, Maxwell J., et autres
Publié: (2025)
par: Yin, Maxwell J., et autres
Publié: (2025)
Long-Context Reasoning Through Proxy-Based Chain-of-Thought Tuning
par: Li, Miao, et autres
Publié: (2026)
par: Li, Miao, et autres
Publié: (2026)
ClusComp: A Simple Paradigm for Model Compression and Efficient Finetuning
par: Liao, Baohao, et autres
Publié: (2025)
par: Liao, Baohao, et autres
Publié: (2025)
A Formal Comparison Between Chain of Thought and Latent Thought
par: Xu, Kevin, et autres
Publié: (2025)
par: Xu, Kevin, et autres
Publié: (2025)
Verifying Chain-of-Thought Reasoning via Its Computational Graph
par: Zhao, Zheng, et autres
Publié: (2025)
par: Zhao, Zheng, et autres
Publié: (2025)
Reasoning Theater: Disentangling Model Beliefs from Chain-of-Thought
par: Boppana, Siddharth, et autres
Publié: (2026)
par: Boppana, Siddharth, et autres
Publié: (2026)
Entropy-Based Block Pruning for Efficient Large Language Models
par: Yang, Liangwei, et autres
Publié: (2025)
par: Yang, Liangwei, et autres
Publié: (2025)
Is Chain-of-Thought Reasoning of LLMs a Mirage? A Data Distribution Lens
par: Zhao, Chengshuai, et autres
Publié: (2025)
par: Zhao, Chengshuai, et autres
Publié: (2025)
Rethinking Chain-of-Thought Reasoning for Videos
par: Zhong, Yiwu, et autres
Publié: (2025)
par: Zhong, Yiwu, et autres
Publié: (2025)
NPG-Muse: Scaling Long Chain-of-Thought Reasoning with NP-Hard Graph Problems
par: Wang, Yuyao, et autres
Publié: (2025)
par: Wang, Yuyao, et autres
Publié: (2025)
Hidden Error Awareness in Chain-of-Thought Reasoning: The Signal Is Diagnostic, Not Causal
par: Yuan, Aojie, et autres
Publié: (2026)
par: Yuan, Aojie, et autres
Publié: (2026)
CoTAR: Chain-of-Thought Attribution Reasoning with Multi-level Granularity
par: Berchansky, Moshe, et autres
Publié: (2024)
par: Berchansky, Moshe, et autres
Publié: (2024)
Stepwise Perplexity-Guided Refinement for Efficient Chain-of-Thought Reasoning in Large Language Models
par: Cui, Yingqian, et autres
Publié: (2025)
par: Cui, Yingqian, et autres
Publié: (2025)
Control the Temperature: Selective Sampling for Diverse and High-Quality LLM Outputs
par: Troshin, Sergey, et autres
Publié: (2025)
par: Troshin, Sergey, et autres
Publié: (2025)
Think Consistently, Reason Efficiently: Energy-Based Calibration for Implicit Chain-of-Thought
par: Chen, Zhikang, et autres
Publié: (2025)
par: Chen, Zhikang, et autres
Publié: (2025)
Documents similaires
-
Scalable Chain of Thoughts via Elastic Reasoning
par: Xu, Yuhui, et autres
Publié: (2025) -
Reward-Guided Speculative Decoding for Efficient LLM Reasoning
par: Liao, Baohao, et autres
Publié: (2025) -
Self-Hinting Language Models Enhance Reinforcement Learning
par: Liao, Baohao, et autres
Publié: (2026) -
3-in-1: 2D Rotary Adaptation for Efficient Finetuning, Efficient Batching and Composability
par: Liao, Baohao, et autres
Publié: (2024) -
Automatic Curriculum Expert Iteration for Reliable LLM Reasoning
par: Zhao, Zirui, et autres
Publié: (2024)