Fast Chain-of-Thought: A Glance of Future from Parallel Decoding Leads to Answers Faster
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhang, Hongxuan, Liu, Zhining, Zhao, Yao, Zheng, Jiaqi, Zhuang, Chenyi, Gu, Jinjie, Chen, Guihai |
|---|---|
| Format: | Preprint |
| Publié: |
2023
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
CSR:Achieving 1 Bit Key-Value Cache via Sparse Representation
par: Zhang, Hongxuan, et autres
Publié: (2024)
par: Zhang, Hongxuan, et autres
Publié: (2024)
RAG-R1: Incentivizing the Search and Reasoning Capabilities of LLMs through Multi-query Parallelism
par: Tan, Zhiwen, et autres
Publié: (2025)
par: Tan, Zhiwen, et autres
Publié: (2025)
Mitigate Position Bias with Coupled Ranking Bias on CTR Prediction
par: Zhao, Yao, et autres
Publié: (2024)
par: Zhao, Yao, et autres
Publié: (2024)
CharPoet: A Chinese Classical Poetry Generation System Based on Token-free LLM
par: Yu, Chengyue, et autres
Publié: (2024)
par: Yu, Chengyue, et autres
Publié: (2024)
From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought
par: Tan, Wentao, et autres
Publié: (2025)
par: Tan, Wentao, et autres
Publié: (2025)
Preemptive Answer "Attacks" on Chain-of-Thought Reasoning
par: Xu, Rongwu, et autres
Publié: (2024)
par: Xu, Rongwu, et autres
Publié: (2024)
Beyond Chain-of-Thought, Effective Graph-of-Thought Reasoning in Language Models
par: Yao, Yao, et autres
Publié: (2023)
par: Yao, Yao, et autres
Publié: (2023)
Decoding at the Speed of Thought: Harnessing Parallel Decoding of Lexical Units for LLMs
par: Sun, Chenxi, et autres
Publié: (2024)
par: Sun, Chenxi, et autres
Publié: (2024)
Parallel Continuous Chain-of-Thought with Jacobi Iteration
par: Wu, Haoyi, et autres
Publié: (2025)
par: Wu, Haoyi, et autres
Publié: (2025)
Glancing Future for Simultaneous Machine Translation
par: Guo, Shoutao, et autres
Publié: (2023)
par: Guo, Shoutao, et autres
Publié: (2023)
Mitigating Spurious Correlations Between Question and Answer via Chain-of-Thought Correctness Perception Distillation
par: Xie, Hongyan, et autres
Publié: (2025)
par: Xie, Hongyan, et autres
Publié: (2025)
SmartThinker: Progressive Chain-of-Thought Length Calibration for Efficient Large Language Model Reasoning
par: Hu, Chenzhi, et autres
Publié: (2026)
par: Hu, Chenzhi, et autres
Publié: (2026)
Skip-Thinking: Chunk-wise Chain-of-Thought Distillation Enable Smaller Language Models to Reason Better and Faster
par: Chen, Xiao, et autres
Publié: (2025)
par: Chen, Xiao, et autres
Publié: (2025)
Unlocking the Capabilities of Thought: A Reasoning Boundary Framework to Quantify and Optimize Chain-of-Thought
par: Chen, Qiguang, et autres
Publié: (2024)
par: Chen, Qiguang, et autres
Publié: (2024)
Lookahead: An Inference Acceleration Framework for Large Language Model with Lossless Generation Accuracy
par: Zhao, Yao, et autres
Publié: (2023)
par: Zhao, Yao, et autres
Publié: (2023)
Latent Chain-of-Thought? Decoding the Depth-Recurrent Transformer
par: Lu, Wenquan, et autres
Publié: (2025)
par: Lu, Wenquan, et autres
Publié: (2025)
Rethinking Dense Sequential Chains: Reasoning Language Models Can Extract Answers from Sparse, Order-Shuffling Chain-of-Thoughts
par: Chen, Yi-Chang, et autres
Publié: (2026)
par: Chen, Yi-Chang, et autres
Publié: (2026)
Explainable Behavior Cloning: Teaching Large Language Model Agents through Learning by Demonstration
par: Guan, Yanchu, et autres
Publié: (2024)
par: Guan, Yanchu, et autres
Publié: (2024)
Expediting and Elevating Large Language Model Reasoning via Hidden Chain-of-Thought Decoding
par: Liu, Tianqiao, et autres
Publié: (2024)
par: Liu, Tianqiao, et autres
Publié: (2024)
Navigate through Enigmatic Labyrinth A Survey of Chain of Thought Reasoning: Advances, Frontiers and Future
par: Chu, Zheng, et autres
Publié: (2023)
par: Chu, Zheng, et autres
Publié: (2023)
MoDE: A Mixture-of-Experts Model with Mutual Distillation among the Experts
par: Xie, Zhitian, et autres
Publié: (2024)
par: Xie, Zhitian, et autres
Publié: (2024)
Context Over Compute Human-in-the-Loop Outperforms Iterative Chain-of-Thought Prompting in Interview Answer Quality
par: Zhu, Kewen, et autres
Publié: (2026)
par: Zhu, Kewen, et autres
Publié: (2026)
NPG-Muse: Scaling Long Chain-of-Thought Reasoning with NP-Hard Graph Problems
par: Wang, Yuyao, et autres
Publié: (2025)
par: Wang, Yuyao, et autres
Publié: (2025)
Deductive Beam Search: Decoding Deducible Rationale for Chain-of-Thought Reasoning
par: Zhu, Tinghui, et autres
Publié: (2024)
par: Zhu, Tinghui, et autres
Publié: (2024)
SIM-CoT: Supervised Implicit Chain-of-Thought
par: Wei, Xilin, et autres
Publié: (2025)
par: Wei, Xilin, et autres
Publié: (2025)
Fast and Accurate Causal Parallel Decoding using Jacobi Forcing
par: Hu, Lanxiang, et autres
Publié: (2025)
par: Hu, Lanxiang, et autres
Publié: (2025)
Pre$^3$: Enabling Deterministic Pushdown Automata for Faster Structured LLM Generation
par: Chen, Junyi, et autres
Publié: (2025)
par: Chen, Junyi, et autres
Publié: (2025)
Cheaper, Better, Faster, Stronger: Robust Text-to-SQL without Chain-of-Thought or Fine-Tuning
par: Dönder, Yusuf Denizay, et autres
Publié: (2025)
par: Dönder, Yusuf Denizay, et autres
Publié: (2025)
Chain-of-Conceptual-Thought Elicits Daily Conversation in Large Language Models
par: Gu, Qingqing, et autres
Publié: (2025)
par: Gu, Qingqing, et autres
Publié: (2025)
Reassessing the Role of Chain-of-Thought in Sentiment Analysis: Insights and Limitations
par: Zheng, Kaiyuan, et autres
Publié: (2025)
par: Zheng, Kaiyuan, et autres
Publié: (2025)
Visual Thoughts: A Unified Perspective of Understanding Multimodal Chain-of-Thought
par: Cheng, Zihui, et autres
Publié: (2025)
par: Cheng, Zihui, et autres
Publié: (2025)
Merge-of-Thought Distillation
par: Shen, Zhanming, et autres
Publié: (2025)
par: Shen, Zhanming, et autres
Publié: (2025)
Upfront Chain-of-Thought: A Cooperative Framework for Chain-of-Thought Compression
par: Li, Chengzhengxu, et autres
Publié: (2025)
par: Li, Chengzhengxu, et autres
Publié: (2025)
LogitsCoder: Towards Efficient Chain-of-Thought Path Search via Logits Preference Decoding for Code Generation
par: Chen, Jizheng, et autres
Publié: (2026)
par: Chen, Jizheng, et autres
Publié: (2026)
Token Signature: Predicting Chain-of-Thought Gains with Token Decoding Feature in Large Language Models
par: Liu, Peijie, et autres
Publié: (2025)
par: Liu, Peijie, et autres
Publié: (2025)
Mind the Gap: Bridging Thought Leap for Improved Chain-of-Thought Tuning
par: Xu, Haolei, et autres
Publié: (2025)
par: Xu, Haolei, et autres
Publié: (2025)
Dialogue Ontology Relation Extraction via Constrained Chain-of-Thought Decoding
par: Vukovic, Renato, et autres
Publié: (2024)
par: Vukovic, Renato, et autres
Publié: (2024)
dParallel: Learnable Parallel Decoding for dLLMs
par: Chen, Zigeng, et autres
Publié: (2025)
par: Chen, Zigeng, et autres
Publié: (2025)
Dynamic Depth Decoding: Faster Speculative Decoding for LLMs
par: Brown, Oscar, et autres
Publié: (2024)
par: Brown, Oscar, et autres
Publié: (2024)
Chain of Draft: Thinking Faster by Writing Less
par: Xu, Silei, et autres
Publié: (2025)
par: Xu, Silei, et autres
Publié: (2025)
Documents similaires
-
CSR:Achieving 1 Bit Key-Value Cache via Sparse Representation
par: Zhang, Hongxuan, et autres
Publié: (2024) -
RAG-R1: Incentivizing the Search and Reasoning Capabilities of LLMs through Multi-query Parallelism
par: Tan, Zhiwen, et autres
Publié: (2025) -
Mitigate Position Bias with Coupled Ranking Bias on CTR Prediction
par: Zhao, Yao, et autres
Publié: (2024) -
CharPoet: A Chinese Classical Poetry Generation System Based on Token-free LLM
par: Yu, Chengyue, et autres
Publié: (2024) -
From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought
par: Tan, Wentao, et autres
Publié: (2025)