Long-Chain Reasoning Distillation via Adaptive Prefix Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Zhenghao, Wu, Zhuoyang, Li, Xinze, Yan, Yukun, Wang, Shuo, Chen, Zulong, Gu, Yu, Yu, Ge, Sun, Maosong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Enhancing Long-Chain Reasoning Distillation through Error-Aware Self-Reflection
von: Wu, Zhuoyang, et al.
Veröffentlicht: (2025)
von: Wu, Zhuoyang, et al.
Veröffentlicht: (2025)
Chunks as Arms: Multi-Armed Bandit-Guided Sampling for Long-Context LLM Preference Optimization
von: Duan, Shaohua, et al.
Veröffentlicht: (2025)
von: Duan, Shaohua, et al.
Veröffentlicht: (2025)
MetaMem: Evolving Meta-Memory for Knowledge Utilization through Self-Reflective Symbolic Optimization
von: Xin, Haidong, et al.
Veröffentlicht: (2026)
von: Xin, Haidong, et al.
Veröffentlicht: (2026)
RankCoT: Refining Knowledge for Retrieval-Augmented Generation through Ranking Chain-of-Thoughts
von: Wu, Mingyan, et al.
Veröffentlicht: (2025)
von: Wu, Mingyan, et al.
Veröffentlicht: (2025)
Mixture-of-Retrieval Experts for Reasoning-Guided Multimodal Knowledge Exploitation
von: Peng, Chunyi, et al.
Veröffentlicht: (2025)
von: Peng, Chunyi, et al.
Veröffentlicht: (2025)
Structured Knowledge Representation through Contextual Pages for Retrieval-Augmented Generation
von: Li, Xinze, et al.
Veröffentlicht: (2026)
von: Li, Xinze, et al.
Veröffentlicht: (2026)
Building A Coding Assistant via the Retrieval-Augmented Language Model
von: Li, Xinze, et al.
Veröffentlicht: (2024)
von: Li, Xinze, et al.
Veröffentlicht: (2024)
Empirical Analysis of Decoding Biases in Masked Diffusion Models
von: Huang, Pengcheng, et al.
Veröffentlicht: (2025)
von: Huang, Pengcheng, et al.
Veröffentlicht: (2025)
Say More with Less: Understanding Prompt Learning Behaviors through Gist Compression
von: Li, Xinze, et al.
Veröffentlicht: (2024)
von: Li, Xinze, et al.
Veröffentlicht: (2024)
Judge as A Judge: Improving the Evaluation of Retrieval-Augmented Generation through the Judge-Consistency of Large Language Models
von: Liu, Shuliang, et al.
Veröffentlicht: (2025)
von: Liu, Shuliang, et al.
Veröffentlicht: (2025)
Legal$Δ$: Enhancing Legal Reasoning in LLMs via Reinforcement Learning with Chain-of-Thought Guided Information Gain
von: Dai, Xin, et al.
Veröffentlicht: (2025)
von: Dai, Xin, et al.
Veröffentlicht: (2025)
Finding What Matters: Anchoring Context Knowledge with Evolving Indices for Iterative Retrieval
von: Wu, Mingyan, et al.
Veröffentlicht: (2026)
von: Wu, Mingyan, et al.
Veröffentlicht: (2026)
RAG-DDR: Optimizing Retrieval-Augmented Generation Using Differentiable Data Rewards
von: Li, Xinze, et al.
Veröffentlicht: (2024)
von: Li, Xinze, et al.
Veröffentlicht: (2024)
ReCUT: Balancing Reasoning Length and Accuracy in LLMs via Stepwise Trails and Preference Optimization
von: Jin, Zhensheng, et al.
Veröffentlicht: (2025)
von: Jin, Zhensheng, et al.
Veröffentlicht: (2025)
ReAlign: Optimizing the Visual Document Retriever with Reasoning-Guided Fine-Grained Alignment
von: Yang, Hao, et al.
Veröffentlicht: (2026)
von: Yang, Hao, et al.
Veröffentlicht: (2026)
KG-Infused RAG: Augmenting Corpus-Based RAG with External Knowledge Graphs
von: Wu, Dingjun, et al.
Veröffentlicht: (2025)
von: Wu, Dingjun, et al.
Veröffentlicht: (2025)
VisRAG 2.0: Evidence-Guided Multi-Image Reasoning in Visual Retrieval-Augmented Generation
von: Sun, Yubo, et al.
Veröffentlicht: (2025)
von: Sun, Yubo, et al.
Veröffentlicht: (2025)
UNIKIE-BENCH: Benchmarking Large Multimodal Models for Key Information Extraction in Visual Documents
von: Ji, Yifan, et al.
Veröffentlicht: (2026)
von: Ji, Yifan, et al.
Veröffentlicht: (2026)
HIPPO: Enhancing the Table Understanding Capability of LLMs through Hybrid-Modal Preference Optimization
von: Wang, Haolan, et al.
Veröffentlicht: (2025)
von: Wang, Haolan, et al.
Veröffentlicht: (2025)
ClueAnchor: Clue-Anchored Knowledge Reasoning Exploration and Optimization for Retrieval-Augmented Generation
von: Chen, Hao, et al.
Veröffentlicht: (2025)
von: Chen, Hao, et al.
Veröffentlicht: (2025)
Lang2Act: Fine-Grained Visual Reasoning through Self-Emergent Linguistic Toolchains
von: Xiong, Yuqi, et al.
Veröffentlicht: (2026)
von: Xiong, Yuqi, et al.
Veröffentlicht: (2026)
NaviRAG: Towards Active Knowledge Navigation for Retrieval-Augmented Generation
von: Dai, Jihao, et al.
Veröffentlicht: (2026)
von: Dai, Jihao, et al.
Veröffentlicht: (2026)
A*-Thought: Efficient Reasoning via Bidirectional Compression for Low-Resource Settings
von: Xu, Xiaoang, et al.
Veröffentlicht: (2025)
von: Xu, Xiaoang, et al.
Veröffentlicht: (2025)
LegalDuet: Learning Fine-grained Representations for Legal Judgment Prediction via a Dual-View Contrastive Learning
von: Xu, Buqiang, et al.
Veröffentlicht: (2024)
von: Xu, Buqiang, et al.
Veröffentlicht: (2024)
KBAlign: Efficient Self Adaptation on Specific Knowledge Bases
von: Zeng, Zheni, et al.
Veröffentlicht: (2024)
von: Zeng, Zheni, et al.
Veröffentlicht: (2024)
Mitigating Judgment Preference Bias in Large Language Models through Group-Based Polling
von: Liu, Shuliang, et al.
Veröffentlicht: (2025)
von: Liu, Shuliang, et al.
Veröffentlicht: (2025)
Teaching LLMs to Learn Tool Trialing and Execution through Environment Interaction
von: Gao, Xingjie, et al.
Veröffentlicht: (2026)
von: Gao, Xingjie, et al.
Veröffentlicht: (2026)
ThinkNote: Enhancing Knowledge Integration and Utilization of Large Language Models via Constructivist Cognition Modeling
von: Xu, Zhipeng, et al.
Veröffentlicht: (2024)
von: Xu, Zhipeng, et al.
Veröffentlicht: (2024)
Cleaner Pretraining Corpus Curation with Neural Web Scraping
von: Xu, Zhipeng, et al.
Veröffentlicht: (2024)
von: Xu, Zhipeng, et al.
Veröffentlicht: (2024)
ParamMute: Suppressing Knowledge-Critical FFNs for Faithful Retrieval-Augmented Generation
von: Huang, Pengcheng, et al.
Veröffentlicht: (2025)
von: Huang, Pengcheng, et al.
Veröffentlicht: (2025)
DeepNote: Note-Centric Deep Retrieval-Augmented Generation
von: Wang, Ruobing, et al.
Veröffentlicht: (2024)
von: Wang, Ruobing, et al.
Veröffentlicht: (2024)
PersLLM: A Personified Training Approach for Large Language Models
von: Zeng, Zheni, et al.
Veröffentlicht: (2024)
von: Zeng, Zheni, et al.
Veröffentlicht: (2024)
Mixed Distillation Helps Smaller Language Model Better Reasoning
von: Li, Chenglin, et al.
Veröffentlicht: (2023)
von: Li, Chenglin, et al.
Veröffentlicht: (2023)
Know More, Know Clearer: A Meta-Cognitive Framework for Knowledge Augmentation in Large Language Models
von: Chen, Hao, et al.
Veröffentlicht: (2026)
von: Chen, Hao, et al.
Veröffentlicht: (2026)
KARE-RAG: Knowledge-Aware Refinement and Enhancement for RAG
von: Li, Yongjian, et al.
Veröffentlicht: (2025)
von: Li, Yongjian, et al.
Veröffentlicht: (2025)
LISRec: Modeling User Preferences with Learned Item Shortcuts for Sequential Recommendation
von: Xin, Haidong, et al.
Veröffentlicht: (2025)
von: Xin, Haidong, et al.
Veröffentlicht: (2025)
Recall with Reasoning: Chain-of-Thought Distillation for Mamba's Long-Context Memory and Extrapolation
von: Ma, Junyu, et al.
Veröffentlicht: (2025)
von: Ma, Junyu, et al.
Veröffentlicht: (2025)
Save the Good Prefix: Precise Error Penalization via Process-Supervised RL to Enhance LLM Reasoning
von: Liu, Haolin, et al.
Veröffentlicht: (2026)
von: Liu, Haolin, et al.
Veröffentlicht: (2026)
Revealing the Attention Floating Mechanism in Masked Diffusion Models
von: Dai, Xin, et al.
Veröffentlicht: (2026)
von: Dai, Xin, et al.
Veröffentlicht: (2026)
MatPlotAgent: Method and Evaluation for LLM-Based Agentic Scientific Data Visualization
von: Yang, Zhiyu, et al.
Veröffentlicht: (2024)
von: Yang, Zhiyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Enhancing Long-Chain Reasoning Distillation through Error-Aware Self-Reflection
von: Wu, Zhuoyang, et al.
Veröffentlicht: (2025) -
Chunks as Arms: Multi-Armed Bandit-Guided Sampling for Long-Context LLM Preference Optimization
von: Duan, Shaohua, et al.
Veröffentlicht: (2025) -
MetaMem: Evolving Meta-Memory for Knowledge Utilization through Self-Reflective Symbolic Optimization
von: Xin, Haidong, et al.
Veröffentlicht: (2026) -
RankCoT: Refining Knowledge for Retrieval-Augmented Generation through Ranking Chain-of-Thoughts
von: Wu, Mingyan, et al.
Veröffentlicht: (2025) -
Mixture-of-Retrieval Experts for Reasoning-Guided Multimodal Knowledge Exploitation
von: Peng, Chunyi, et al.
Veröffentlicht: (2025)