Fast Chain-of-Thought: A Glance of Future from Parallel Decoding Leads to Answers Faster
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Hongxuan, Liu, Zhining, Zhao, Yao, Zheng, Jiaqi, Zhuang, Chenyi, Gu, Jinjie, Chen, Guihai |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CSR:Achieving 1 Bit Key-Value Cache via Sparse Representation
por: Zhang, Hongxuan, et al.
Publicado: (2024)
por: Zhang, Hongxuan, et al.
Publicado: (2024)
RAG-R1: Incentivizing the Search and Reasoning Capabilities of LLMs through Multi-query Parallelism
por: Tan, Zhiwen, et al.
Publicado: (2025)
por: Tan, Zhiwen, et al.
Publicado: (2025)
Mitigate Position Bias with Coupled Ranking Bias on CTR Prediction
por: Zhao, Yao, et al.
Publicado: (2024)
por: Zhao, Yao, et al.
Publicado: (2024)
CharPoet: A Chinese Classical Poetry Generation System Based on Token-free LLM
por: Yu, Chengyue, et al.
Publicado: (2024)
por: Yu, Chengyue, et al.
Publicado: (2024)
From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought
por: Tan, Wentao, et al.
Publicado: (2025)
por: Tan, Wentao, et al.
Publicado: (2025)
Preemptive Answer "Attacks" on Chain-of-Thought Reasoning
por: Xu, Rongwu, et al.
Publicado: (2024)
por: Xu, Rongwu, et al.
Publicado: (2024)
Beyond Chain-of-Thought, Effective Graph-of-Thought Reasoning in Language Models
por: Yao, Yao, et al.
Publicado: (2023)
por: Yao, Yao, et al.
Publicado: (2023)
Decoding at the Speed of Thought: Harnessing Parallel Decoding of Lexical Units for LLMs
por: Sun, Chenxi, et al.
Publicado: (2024)
por: Sun, Chenxi, et al.
Publicado: (2024)
Parallel Continuous Chain-of-Thought with Jacobi Iteration
por: Wu, Haoyi, et al.
Publicado: (2025)
por: Wu, Haoyi, et al.
Publicado: (2025)
Glancing Future for Simultaneous Machine Translation
por: Guo, Shoutao, et al.
Publicado: (2023)
por: Guo, Shoutao, et al.
Publicado: (2023)
Mitigating Spurious Correlations Between Question and Answer via Chain-of-Thought Correctness Perception Distillation
por: Xie, Hongyan, et al.
Publicado: (2025)
por: Xie, Hongyan, et al.
Publicado: (2025)
SmartThinker: Progressive Chain-of-Thought Length Calibration for Efficient Large Language Model Reasoning
por: Hu, Chenzhi, et al.
Publicado: (2026)
por: Hu, Chenzhi, et al.
Publicado: (2026)
Skip-Thinking: Chunk-wise Chain-of-Thought Distillation Enable Smaller Language Models to Reason Better and Faster
por: Chen, Xiao, et al.
Publicado: (2025)
por: Chen, Xiao, et al.
Publicado: (2025)
Unlocking the Capabilities of Thought: A Reasoning Boundary Framework to Quantify and Optimize Chain-of-Thought
por: Chen, Qiguang, et al.
Publicado: (2024)
por: Chen, Qiguang, et al.
Publicado: (2024)
Lookahead: An Inference Acceleration Framework for Large Language Model with Lossless Generation Accuracy
por: Zhao, Yao, et al.
Publicado: (2023)
por: Zhao, Yao, et al.
Publicado: (2023)
Latent Chain-of-Thought? Decoding the Depth-Recurrent Transformer
por: Lu, Wenquan, et al.
Publicado: (2025)
por: Lu, Wenquan, et al.
Publicado: (2025)
Rethinking Dense Sequential Chains: Reasoning Language Models Can Extract Answers from Sparse, Order-Shuffling Chain-of-Thoughts
por: Chen, Yi-Chang, et al.
Publicado: (2026)
por: Chen, Yi-Chang, et al.
Publicado: (2026)
Explainable Behavior Cloning: Teaching Large Language Model Agents through Learning by Demonstration
por: Guan, Yanchu, et al.
Publicado: (2024)
por: Guan, Yanchu, et al.
Publicado: (2024)
Expediting and Elevating Large Language Model Reasoning via Hidden Chain-of-Thought Decoding
por: Liu, Tianqiao, et al.
Publicado: (2024)
por: Liu, Tianqiao, et al.
Publicado: (2024)
Navigate through Enigmatic Labyrinth A Survey of Chain of Thought Reasoning: Advances, Frontiers and Future
por: Chu, Zheng, et al.
Publicado: (2023)
por: Chu, Zheng, et al.
Publicado: (2023)
MoDE: A Mixture-of-Experts Model with Mutual Distillation among the Experts
por: Xie, Zhitian, et al.
Publicado: (2024)
por: Xie, Zhitian, et al.
Publicado: (2024)
Context Over Compute Human-in-the-Loop Outperforms Iterative Chain-of-Thought Prompting in Interview Answer Quality
por: Zhu, Kewen, et al.
Publicado: (2026)
por: Zhu, Kewen, et al.
Publicado: (2026)
NPG-Muse: Scaling Long Chain-of-Thought Reasoning with NP-Hard Graph Problems
por: Wang, Yuyao, et al.
Publicado: (2025)
por: Wang, Yuyao, et al.
Publicado: (2025)
Deductive Beam Search: Decoding Deducible Rationale for Chain-of-Thought Reasoning
por: Zhu, Tinghui, et al.
Publicado: (2024)
por: Zhu, Tinghui, et al.
Publicado: (2024)
SIM-CoT: Supervised Implicit Chain-of-Thought
por: Wei, Xilin, et al.
Publicado: (2025)
por: Wei, Xilin, et al.
Publicado: (2025)
Fast and Accurate Causal Parallel Decoding using Jacobi Forcing
por: Hu, Lanxiang, et al.
Publicado: (2025)
por: Hu, Lanxiang, et al.
Publicado: (2025)
Pre$^3$: Enabling Deterministic Pushdown Automata for Faster Structured LLM Generation
por: Chen, Junyi, et al.
Publicado: (2025)
por: Chen, Junyi, et al.
Publicado: (2025)
Cheaper, Better, Faster, Stronger: Robust Text-to-SQL without Chain-of-Thought or Fine-Tuning
por: Dönder, Yusuf Denizay, et al.
Publicado: (2025)
por: Dönder, Yusuf Denizay, et al.
Publicado: (2025)
Chain-of-Conceptual-Thought Elicits Daily Conversation in Large Language Models
por: Gu, Qingqing, et al.
Publicado: (2025)
por: Gu, Qingqing, et al.
Publicado: (2025)
Reassessing the Role of Chain-of-Thought in Sentiment Analysis: Insights and Limitations
por: Zheng, Kaiyuan, et al.
Publicado: (2025)
por: Zheng, Kaiyuan, et al.
Publicado: (2025)
Visual Thoughts: A Unified Perspective of Understanding Multimodal Chain-of-Thought
por: Cheng, Zihui, et al.
Publicado: (2025)
por: Cheng, Zihui, et al.
Publicado: (2025)
Merge-of-Thought Distillation
por: Shen, Zhanming, et al.
Publicado: (2025)
por: Shen, Zhanming, et al.
Publicado: (2025)
Upfront Chain-of-Thought: A Cooperative Framework for Chain-of-Thought Compression
por: Li, Chengzhengxu, et al.
Publicado: (2025)
por: Li, Chengzhengxu, et al.
Publicado: (2025)
LogitsCoder: Towards Efficient Chain-of-Thought Path Search via Logits Preference Decoding for Code Generation
por: Chen, Jizheng, et al.
Publicado: (2026)
por: Chen, Jizheng, et al.
Publicado: (2026)
Token Signature: Predicting Chain-of-Thought Gains with Token Decoding Feature in Large Language Models
por: Liu, Peijie, et al.
Publicado: (2025)
por: Liu, Peijie, et al.
Publicado: (2025)
Mind the Gap: Bridging Thought Leap for Improved Chain-of-Thought Tuning
por: Xu, Haolei, et al.
Publicado: (2025)
por: Xu, Haolei, et al.
Publicado: (2025)
Dialogue Ontology Relation Extraction via Constrained Chain-of-Thought Decoding
por: Vukovic, Renato, et al.
Publicado: (2024)
por: Vukovic, Renato, et al.
Publicado: (2024)
dParallel: Learnable Parallel Decoding for dLLMs
por: Chen, Zigeng, et al.
Publicado: (2025)
por: Chen, Zigeng, et al.
Publicado: (2025)
Dynamic Depth Decoding: Faster Speculative Decoding for LLMs
por: Brown, Oscar, et al.
Publicado: (2024)
por: Brown, Oscar, et al.
Publicado: (2024)
Chain of Draft: Thinking Faster by Writing Less
por: Xu, Silei, et al.
Publicado: (2025)
por: Xu, Silei, et al.
Publicado: (2025)
Ejemplares similares
-
CSR:Achieving 1 Bit Key-Value Cache via Sparse Representation
por: Zhang, Hongxuan, et al.
Publicado: (2024) -
RAG-R1: Incentivizing the Search and Reasoning Capabilities of LLMs through Multi-query Parallelism
por: Tan, Zhiwen, et al.
Publicado: (2025) -
Mitigate Position Bias with Coupled Ranking Bias on CTR Prediction
por: Zhao, Yao, et al.
Publicado: (2024) -
CharPoet: A Chinese Classical Poetry Generation System Based on Token-free LLM
por: Yu, Chengyue, et al.
Publicado: (2024) -
From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought
por: Tan, Wentao, et al.
Publicado: (2025)