MiCoTA: Bridging the Learnability Gap with Intermediate CoT and Teacher Assistants
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ding, Dongyi, Wang, Tiannan, Zhu, Chenghao, Tao, Meiling, Jiang, Yuchen Eleanor, Zhou, Wangchunshu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization
von: Tao, Meiling, et al.
Veröffentlicht: (2025)
von: Tao, Meiling, et al.
Veröffentlicht: (2025)
Towards Faithful and Controllable Personalization via Critique-Post-Edit Reinforcement Learning
von: Zhu, Chenghao, et al.
Veröffentlicht: (2025)
von: Zhu, Chenghao, et al.
Veröffentlicht: (2025)
AI PERSONA: Towards Life-long Personalization of LLMs
von: Wang, Tiannan, et al.
Veröffentlicht: (2024)
von: Wang, Tiannan, et al.
Veröffentlicht: (2024)
Syzygy of Thoughts: Improving LLM CoT with the Minimal Free Resolution
von: Li, Chenghao, et al.
Veröffentlicht: (2025)
von: Li, Chenghao, et al.
Veröffentlicht: (2025)
Parrot: A Training Pipeline Enhances Both Program CoT and Natural Language CoT for Reasoning
von: Jin, Senjie, et al.
Veröffentlicht: (2025)
von: Jin, Senjie, et al.
Veröffentlicht: (2025)
From Explicit CoT to Implicit CoT: Learning to Internalize CoT Step by Step
von: Deng, Yuntian, et al.
Veröffentlicht: (2024)
von: Deng, Yuntian, et al.
Veröffentlicht: (2024)
To CoT or not to CoT? Chain-of-thought helps mainly on math and symbolic reasoning
von: Sprague, Zayne, et al.
Veröffentlicht: (2024)
von: Sprague, Zayne, et al.
Veröffentlicht: (2024)
Beyond Isolated Capabilities: Bridging Long CoT Reasoning and Long-Context Understanding
von: Wang, Yifei
Veröffentlicht: (2025)
von: Wang, Yifei
Veröffentlicht: (2025)
O-Mem: Omni Memory System for Personalized, Long Horizon, Self-Evolving Agents
von: Wang, Piaohong, et al.
Veröffentlicht: (2025)
von: Wang, Piaohong, et al.
Veröffentlicht: (2025)
CoT-Evo: Evolutionary Distillation of Chain-of-Thought for Scientific Reasoning
von: Feng, Kehua, et al.
Veröffentlicht: (2025)
von: Feng, Kehua, et al.
Veröffentlicht: (2025)
D-SCoRE: Document-Centric Segmentation and CoT Reasoning with Structured Export for QA-CoT Data Generation
von: Zhou, Weibo, et al.
Veröffentlicht: (2025)
von: Zhou, Weibo, et al.
Veröffentlicht: (2025)
Symbolic Learning Enables Self-Evolving Agents
von: Zhou, Wangchunshu, et al.
Veröffentlicht: (2024)
von: Zhou, Wangchunshu, et al.
Veröffentlicht: (2024)
Efficient Long CoT Reasoning in Small Language Models
von: Wang, Zhaoyang, et al.
Veröffentlicht: (2025)
von: Wang, Zhaoyang, et al.
Veröffentlicht: (2025)
Chain-of-Probe: Examining the Necessity and Accuracy of CoT Step-by-Step
von: Wang, Zezhong, et al.
Veröffentlicht: (2024)
von: Wang, Zezhong, et al.
Veröffentlicht: (2024)
Investigating Mysteries of CoT-Augmented Distillation
von: Wadhwa, Somin, et al.
Veröffentlicht: (2024)
von: Wadhwa, Somin, et al.
Veröffentlicht: (2024)
Investigating CoT Monitorability in Large Reasoning Models
von: Yang, Shu, et al.
Veröffentlicht: (2025)
von: Yang, Shu, et al.
Veröffentlicht: (2025)
CoT Vectors: Transferring and Probing the Reasoning Mechanisms of LLMs
von: Li, Li, et al.
Veröffentlicht: (2025)
von: Li, Li, et al.
Veröffentlicht: (2025)
SIM-CoT: Supervised Implicit Chain-of-Thought
von: Wei, Xilin, et al.
Veröffentlicht: (2025)
von: Wei, Xilin, et al.
Veröffentlicht: (2025)
CoT Referring: Improving Referring Expression Tasks with Grounded Reasoning
von: Dong, Qihua, et al.
Veröffentlicht: (2025)
von: Dong, Qihua, et al.
Veröffentlicht: (2025)
"The Whole Is Greater Than the Sum of Its Parts": A Compatibility-Aware Multi-Teacher CoT Distillation Framework
von: Cui, Jin, et al.
Veröffentlicht: (2026)
von: Cui, Jin, et al.
Veröffentlicht: (2026)
CoT-Valve: Length-Compressible Chain-of-Thought Tuning
von: Ma, Xinyin, et al.
Veröffentlicht: (2025)
von: Ma, Xinyin, et al.
Veröffentlicht: (2025)
Instruction Tuning and CoT Prompting for Contextual Medical QA with LLMs
von: Le, Chenqian, et al.
Veröffentlicht: (2025)
von: Le, Chenqian, et al.
Veröffentlicht: (2025)
AS-ES Learning: Towards Efficient CoT Learning in Small Models
von: Xi, Nuwa, et al.
Veröffentlicht: (2024)
von: Xi, Nuwa, et al.
Veröffentlicht: (2024)
Effectiveness of Zero-shot-CoT in Japanese Prompts
von: Takayama, Shusuke, et al.
Veröffentlicht: (2025)
von: Takayama, Shusuke, et al.
Veröffentlicht: (2025)
The Curse of CoT: On the Limitations of Chain-of-Thought in In-Context Learning
von: Zheng, Tianshi, et al.
Veröffentlicht: (2025)
von: Zheng, Tianshi, et al.
Veröffentlicht: (2025)
Generating Effective CoT Traces for Mitigating Causal Hallucination
von: Zhao, Yiheng, et al.
Veröffentlicht: (2026)
von: Zhao, Yiheng, et al.
Veröffentlicht: (2026)
Co-CoT: A Prompt-Based Framework for Collaborative Chain-of-Thought Reasoning
von: Yoo, Seunghyun
Veröffentlicht: (2025)
von: Yoo, Seunghyun
Veröffentlicht: (2025)
Distribution-Aligned Sequence Distillation for Superior Long-CoT Reasoning
von: Yan, Shaotian, et al.
Veröffentlicht: (2026)
von: Yan, Shaotian, et al.
Veröffentlicht: (2026)
Privacy-preserved LLM Cascade via CoT-enhanced Policy Learning
von: Zhang, Kai, et al.
Veröffentlicht: (2024)
von: Zhang, Kai, et al.
Veröffentlicht: (2024)
Exploring the Limitations of Mamba in COPY and CoT Reasoning
von: Ren, Ruifeng, et al.
Veröffentlicht: (2024)
von: Ren, Ruifeng, et al.
Veröffentlicht: (2024)
Zebra-CoT: A Dataset for Interleaved Vision Language Reasoning
von: Li, Ang, et al.
Veröffentlicht: (2025)
von: Li, Ang, et al.
Veröffentlicht: (2025)
DraCo: Draft as CoT for Text-to-Image Preview and Rare Concept Generation
von: Jiang, Dongzhi, et al.
Veröffentlicht: (2025)
von: Jiang, Dongzhi, et al.
Veröffentlicht: (2025)
PM-KVQ: Progressive Mixed-precision KV Cache Quantization for Long-CoT LLMs
von: Liu, Tengxuan, et al.
Veröffentlicht: (2025)
von: Liu, Tengxuan, et al.
Veröffentlicht: (2025)
Fine-Tuning on Diverse Reasoning Chains Drives Within-Inference CoT Refinement in LLMs
von: Puerto, Haritz, et al.
Veröffentlicht: (2024)
von: Puerto, Haritz, et al.
Veröffentlicht: (2024)
The Synergy Dilemma of Long-CoT SFT and RL: Investigating Post-Training Techniques for Reasoning VLMs
von: Chen, Jierun, et al.
Veröffentlicht: (2025)
von: Chen, Jierun, et al.
Veröffentlicht: (2025)
Ada-R1: Hybrid-CoT via Bi-Level Adaptive Reasoning Optimization
von: Luo, Haotian, et al.
Veröffentlicht: (2025)
von: Luo, Haotian, et al.
Veröffentlicht: (2025)
Improve Student's Reasoning Generalizability through Cascading Decomposed CoTs Distillation
von: Dai, Chengwei, et al.
Veröffentlicht: (2024)
von: Dai, Chengwei, et al.
Veröffentlicht: (2024)
Nash CoT: Multi-Path Inference with Preference Equilibrium
von: Zhang, Ziqi, et al.
Veröffentlicht: (2024)
von: Zhang, Ziqi, et al.
Veröffentlicht: (2024)
How Likely Do LLMs with CoT Mimic Human Reasoning?
von: Bao, Guangsheng, et al.
Veröffentlicht: (2024)
von: Bao, Guangsheng, et al.
Veröffentlicht: (2024)
Beyond In-Distribution Success: Scaling Curves of CoT Granularity for Language Model Generalization
von: Wang, Ru, et al.
Veröffentlicht: (2025)
von: Wang, Ru, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization
von: Tao, Meiling, et al.
Veröffentlicht: (2025) -
Towards Faithful and Controllable Personalization via Critique-Post-Edit Reinforcement Learning
von: Zhu, Chenghao, et al.
Veröffentlicht: (2025) -
AI PERSONA: Towards Life-long Personalization of LLMs
von: Wang, Tiannan, et al.
Veröffentlicht: (2024) -
Syzygy of Thoughts: Improving LLM CoT with the Minimal Free Resolution
von: Li, Chenghao, et al.
Veröffentlicht: (2025) -
Parrot: A Training Pipeline Enhances Both Program CoT and Natural Language CoT for Reasoning
von: Jin, Senjie, et al.
Veröffentlicht: (2025)