Facilitating Long Context Understanding via Supervised Chain-of-Thought Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lin, Jingyang, Wong, Andy, Xia, Tian, He, Shenghua, Wei, Hui, Han, Mei, Luo, Jiebo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Systematic Evaluation of LLM-as-a-Judge in LLM Alignment Tasks: Explainable Metrics and Diverse Prompt Templates
von: Wei, Hui, et al.
Veröffentlicht: (2024)
von: Wei, Hui, et al.
Veröffentlicht: (2024)
Chain-of-Thought Matters: Improving Long-Context Language Models with Reasoning Path Supervision
von: Zhu, Dawei, et al.
Veröffentlicht: (2025)
von: Zhu, Dawei, et al.
Veröffentlicht: (2025)
Latent Chain-of-Thought for Visual Reasoning
von: Sun, Guohao, et al.
Veröffentlicht: (2025)
von: Sun, Guohao, et al.
Veröffentlicht: (2025)
On the Role of Reasoning Patterns in the Generalization Discrepancy of Long Chain-of-Thought Supervised Fine-Tuning
von: Li, Zhaoyi, et al.
Veröffentlicht: (2026)
von: Li, Zhaoyi, et al.
Veröffentlicht: (2026)
Recall with Reasoning: Chain-of-Thought Distillation for Mamba's Long-Context Memory and Extrapolation
von: Ma, Junyu, et al.
Veröffentlicht: (2025)
von: Ma, Junyu, et al.
Veröffentlicht: (2025)
Long-Short Chain-of-Thought Mixture Supervised Fine-Tuning Eliciting Efficient Reasoning in Large Language Models
von: Yu, Bin, et al.
Veröffentlicht: (2025)
von: Yu, Bin, et al.
Veröffentlicht: (2025)
Long-Context Reasoning Through Proxy-Based Chain-of-Thought Tuning
von: Li, Miao, et al.
Veröffentlicht: (2026)
von: Li, Miao, et al.
Veröffentlicht: (2026)
A Theoretical Understanding of Chain-of-Thought: Coherent Reasoning and Error-Aware Demonstration
von: Cui, Yingqian, et al.
Veröffentlicht: (2024)
von: Cui, Yingqian, et al.
Veröffentlicht: (2024)
Response-Level Rewards Are All You Need for Online Reinforcement Learning in LLMs: A Mathematical Perspective
von: He, Shenghua, et al.
Veröffentlicht: (2025)
von: He, Shenghua, et al.
Veröffentlicht: (2025)
Understanding Before Reasoning: Enhancing Chain-of-Thought with Iterative Summarization Pre-Prompting
von: Zhu, Dong-Hai, et al.
Veröffentlicht: (2025)
von: Zhu, Dong-Hai, et al.
Veröffentlicht: (2025)
DRT: Deep Reasoning Translation via Long Chain-of-Thought
von: Wang, Jiaan, et al.
Veröffentlicht: (2024)
von: Wang, Jiaan, et al.
Veröffentlicht: (2024)
Efficient Reasoning via Chain of Unconscious Thought
von: Gong, Ruihan, et al.
Veröffentlicht: (2025)
von: Gong, Ruihan, et al.
Veröffentlicht: (2025)
ReasonFlux-PRM: Trajectory-Aware PRMs for Long Chain-of-Thought Reasoning in LLMs
von: Zou, Jiaru, et al.
Veröffentlicht: (2025)
von: Zou, Jiaru, et al.
Veröffentlicht: (2025)
Beyond Fine-Tuning: In-Context Learning and Chain-of-Thought for Reasoned Distractor Generation
von: Alhazmi, Elaf, et al.
Veröffentlicht: (2026)
von: Alhazmi, Elaf, et al.
Veröffentlicht: (2026)
Mechanistic Evidence for Faithfulness Decay in Chain-of-Thought Reasoning
von: Ye, Donald, et al.
Veröffentlicht: (2026)
von: Ye, Donald, et al.
Veröffentlicht: (2026)
Understanding Hidden Computations in Chain-of-Thought Reasoning
von: Bharadwaj, Aryasomayajula Ram
Veröffentlicht: (2024)
von: Bharadwaj, Aryasomayajula Ram
Veröffentlicht: (2024)
Process Supervision for Chain-of-Thought Reasoning via Monte Carlo Net Information Gain
von: Royer, Corentin, et al.
Veröffentlicht: (2026)
von: Royer, Corentin, et al.
Veröffentlicht: (2026)
Demystifying Long Chain-of-Thought Reasoning in LLMs
von: Yeo, Edward, et al.
Veröffentlicht: (2025)
von: Yeo, Edward, et al.
Veröffentlicht: (2025)
LongReason: A Synthetic Long-Context Reasoning Benchmark via Context Expansion
von: Ling, Zhan, et al.
Veröffentlicht: (2025)
von: Ling, Zhan, et al.
Veröffentlicht: (2025)
SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities
von: Jiang, Fengqing, et al.
Veröffentlicht: (2025)
von: Jiang, Fengqing, et al.
Veröffentlicht: (2025)
DiffCoT: Diffusion-styled Chain-of-Thought Reasoning in LLMs
von: Cao, Shidong, et al.
Veröffentlicht: (2026)
von: Cao, Shidong, et al.
Veröffentlicht: (2026)
Expediting and Elevating Large Language Model Reasoning via Hidden Chain-of-Thought Decoding
von: Liu, Tianqiao, et al.
Veröffentlicht: (2024)
von: Liu, Tianqiao, et al.
Veröffentlicht: (2024)
SIM-CoT: Supervised Implicit Chain-of-Thought
von: Wei, Xilin, et al.
Veröffentlicht: (2025)
von: Wei, Xilin, et al.
Veröffentlicht: (2025)
Unleashing Hour-Scale Video Training for Long Video-Language Understanding
von: Lin, Jingyang, et al.
Veröffentlicht: (2025)
von: Lin, Jingyang, et al.
Veröffentlicht: (2025)
Supervised Chain of Thought
von: Zhang, Xiang, et al.
Veröffentlicht: (2024)
von: Zhang, Xiang, et al.
Veröffentlicht: (2024)
The Molecular Structure of Thought: Mapping the Topology of Long Chain-of-Thought Reasoning
von: Chen, Qiguang, et al.
Veröffentlicht: (2026)
von: Chen, Qiguang, et al.
Veröffentlicht: (2026)
Scaling Code-Assisted Chain-of-Thoughts and Instructions for Model Reasoning
von: Lin, Honglin, et al.
Veröffentlicht: (2025)
von: Lin, Honglin, et al.
Veröffentlicht: (2025)
Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?
von: He, Yancheng, et al.
Veröffentlicht: (2025)
von: He, Yancheng, et al.
Veröffentlicht: (2025)
Chain of Preference Optimization: Improving Chain-of-Thought Reasoning in LLMs
von: Zhang, Xuan, et al.
Veröffentlicht: (2024)
von: Zhang, Xuan, et al.
Veröffentlicht: (2024)
The Curse of CoT: On the Limitations of Chain-of-Thought in In-Context Learning
von: Zheng, Tianshi, et al.
Veröffentlicht: (2025)
von: Zheng, Tianshi, et al.
Veröffentlicht: (2025)
Teaching Models to Verbalize Reward Hacking in Chain-of-Thought Reasoning
von: Turpin, Miles, et al.
Veröffentlicht: (2025)
von: Turpin, Miles, et al.
Veröffentlicht: (2025)
Diffusion of Thoughts: Chain-of-Thought Reasoning in Diffusion Language Models
von: Ye, Jiacheng, et al.
Veröffentlicht: (2024)
von: Ye, Jiacheng, et al.
Veröffentlicht: (2024)
CSRP: Chain-of-Thought Reasoning for Chinese Text Correction via Reinforcement Learning with Efficiency-Aware Rewards
von: Tian, Wei, et al.
Veröffentlicht: (2026)
von: Tian, Wei, et al.
Veröffentlicht: (2026)
Draft-Thinking: Learning Efficient Reasoning in Long Chain-of-Thought LLMs
von: Cao, Jie, et al.
Veröffentlicht: (2026)
von: Cao, Jie, et al.
Veröffentlicht: (2026)
Long Chain-of-Thought Reasoning Across Languages
von: Barua, Josh, et al.
Veröffentlicht: (2025)
von: Barua, Josh, et al.
Veröffentlicht: (2025)
Reasoning Beyond Language: A Comprehensive Survey on Latent Chain-of-Thought Reasoning
von: Chen, Xinghao, et al.
Veröffentlicht: (2025)
von: Chen, Xinghao, et al.
Veröffentlicht: (2025)
R1-Compress: Long Chain-of-Thought Compression via Chunk Compression and Search
von: Wang, Yibo, et al.
Veröffentlicht: (2025)
von: Wang, Yibo, et al.
Veröffentlicht: (2025)
Faithful Logical Reasoning via Symbolic Chain-of-Thought
von: Xu, Jundong, et al.
Veröffentlicht: (2024)
von: Xu, Jundong, et al.
Veröffentlicht: (2024)
Chain-of-Thought Reasoning Improves Context-Aware Translation with Large Language Models
von: Ataee, Shabnam, et al.
Veröffentlicht: (2025)
von: Ataee, Shabnam, et al.
Veröffentlicht: (2025)
Long Grounded Thoughts: Synthesizing Visual Problems and Reasoning Chains at Scale
von: Acuna, David, et al.
Veröffentlicht: (2025)
von: Acuna, David, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Systematic Evaluation of LLM-as-a-Judge in LLM Alignment Tasks: Explainable Metrics and Diverse Prompt Templates
von: Wei, Hui, et al.
Veröffentlicht: (2024) -
Chain-of-Thought Matters: Improving Long-Context Language Models with Reasoning Path Supervision
von: Zhu, Dawei, et al.
Veröffentlicht: (2025) -
Latent Chain-of-Thought for Visual Reasoning
von: Sun, Guohao, et al.
Veröffentlicht: (2025) -
On the Role of Reasoning Patterns in the Generalization Discrepancy of Long Chain-of-Thought Supervised Fine-Tuning
von: Li, Zhaoyi, et al.
Veröffentlicht: (2026) -
Recall with Reasoning: Chain-of-Thought Distillation for Mamba's Long-Context Memory and Extrapolation
von: Ma, Junyu, et al.
Veröffentlicht: (2025)