Where Did This Sentence Come From? Tracing Provenance in LLM Reasoning Distillation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Kaiyuan, Yan, Shaotian, Miao, Rui, Wang, Bing, Shen, Chen, Zhang, Jun, Ye, Jieping |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On the Step Length Confounding in LLM Reasoning Data Selection
von: Wang, Bing, et al.
Veröffentlicht: (2026)
von: Wang, Bing, et al.
Veröffentlicht: (2026)
Distribution-Aligned Sequence Distillation for Superior Long-CoT Reasoning
von: Yan, Shaotian, et al.
Veröffentlicht: (2026)
von: Yan, Shaotian, et al.
Veröffentlicht: (2026)
Backtracking When It Strays: Mitigating Dual Exposure Biases in LLM Reasoning Distillation
von: Wang, Bing, et al.
Veröffentlicht: (2026)
von: Wang, Bing, et al.
Veröffentlicht: (2026)
Are Rationales Necessary and Sufficient? Tuning LLMs for Explainable Misinformation Detection
von: Wang, Bing, et al.
Veröffentlicht: (2026)
von: Wang, Bing, et al.
Veröffentlicht: (2026)
Don't Take Things Out of Context: Attention Intervention for Enhancing Chain-of-Thought Reasoning in Large Language Models
von: Yan, Shaotian, et al.
Veröffentlicht: (2025)
von: Yan, Shaotian, et al.
Veröffentlicht: (2025)
Concise and Organized Perception Facilitates Reasoning in Large Language Models
von: Liu, Junjie, et al.
Veröffentlicht: (2023)
von: Liu, Junjie, et al.
Veröffentlicht: (2023)
Prefix Teach, Suffix Fade: Local Teachability Collapse in Strong-to-Weak On-Policy Distillation
von: Liu, Kaiyuan, et al.
Veröffentlicht: (2026)
von: Liu, Kaiyuan, et al.
Veröffentlicht: (2026)
SalaMAnder: Shapley-based Mathematical Expression Attribution and Metric for Chain-of-Thought Reasoning
von: Xin, Yue, et al.
Veröffentlicht: (2025)
von: Xin, Yue, et al.
Veröffentlicht: (2025)
From Redundancy to Relevance: Information Flow in LVLMs Across Reasoning Tasks
von: Zhang, Xiaofeng, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaofeng, et al.
Veröffentlicht: (2024)
Efficient Reasoning Through Suppression of Self-Affirmation Reflections in Large Reasoning Models
von: Liu, Kaiyuan, et al.
Veröffentlicht: (2025)
von: Liu, Kaiyuan, et al.
Veröffentlicht: (2025)
Enhancing Chain-of-Thought Reasoning with Critical Representation Fine-tuning
von: Huang, Chenxi, et al.
Veröffentlicht: (2025)
von: Huang, Chenxi, et al.
Veröffentlicht: (2025)
Through the Lens of Contrast: Self-Improving Visual Reasoning in VLMs
von: Pan, Zhiyu, et al.
Veröffentlicht: (2026)
von: Pan, Zhiyu, et al.
Veröffentlicht: (2026)
Where Did It Go Wrong? Attributing Undesirable LLM Behaviors via Representation Gradient Tracing
von: Li, Zhe, et al.
Veröffentlicht: (2025)
von: Li, Zhe, et al.
Veröffentlicht: (2025)
Instance-adaptive Zero-shot Chain-of-Thought Prompting
von: Yuan, Xiaosong, et al.
Veröffentlicht: (2024)
von: Yuan, Xiaosong, et al.
Veröffentlicht: (2024)
TROVE: A Challenge for Fine-Grained Text Provenance via Source Sentence Tracing and Relationship Classification
von: Zhu, Junnan, et al.
Veröffentlicht: (2025)
von: Zhu, Junnan, et al.
Veröffentlicht: (2025)
Where Did the Amaterasu Particle Come From?
von: Unger, Michael, et al.
Veröffentlicht: (2023)
von: Unger, Michael, et al.
Veröffentlicht: (2023)
Where Did Development Economics Come From?
von: Eric Helleiner
Veröffentlicht: (2026)
von: Eric Helleiner
Veröffentlicht: (2026)
Beamforming-LLM: What, Where and When Did I Miss?
von: Choudhari, Vishal
Veröffentlicht: (2025)
von: Choudhari, Vishal
Veröffentlicht: (2025)
Consistent Paths Lead to Truth: Self-Rewarding Reinforcement Learning for LLM Reasoning
von: Zhang, Kongcheng, et al.
Veröffentlicht: (2025)
von: Zhang, Kongcheng, et al.
Veröffentlicht: (2025)
Understanding LLMs' Cross-Lingual Context Retrieval: How Good It Is And Where It Comes From
von: Gao, Changjiang, et al.
Veröffentlicht: (2025)
von: Gao, Changjiang, et al.
Veröffentlicht: (2025)
Refining Sentence Embedding Model through Ranking Sentences Generation with Large Language Models
von: He, Liyang, et al.
Veröffentlicht: (2025)
von: He, Liyang, et al.
Veröffentlicht: (2025)
From Where Words Come: Efficient Regularization of Code Tokenizers Through Source Attribution
von: Chizhov, Pavel, et al.
Veröffentlicht: (2026)
von: Chizhov, Pavel, et al.
Veröffentlicht: (2026)
Enhancing Large Language Models with Reward-guided Tree Search for Knowledge Graph Question and Answering
von: Long, Xiao, et al.
Veröffentlicht: (2025)
von: Long, Xiao, et al.
Veröffentlicht: (2025)
NaturalThoughts: Selecting and Distilling Reasoning Traces for General Reasoning Tasks
von: Li, Yang, et al.
Veröffentlicht: (2025)
von: Li, Yang, et al.
Veröffentlicht: (2025)
Learning to Reason via Self-Iterative Process Feedback for Small Language Models
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2024)
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2024)
TextSeal: A Localized LLM Watermark for Provenance & Distillation Protection
von: Sander, Tom, et al.
Veröffentlicht: (2026)
von: Sander, Tom, et al.
Veröffentlicht: (2026)
Towards Multidisciplinary Summarization of Hospital Stays: Efficient Sentence-Level Clinical Provenance Categorization
von: Karacan, Baris, et al.
Veröffentlicht: (2026)
von: Karacan, Baris, et al.
Veröffentlicht: (2026)
LLM-Guided Knowledge Distillation for Temporal Knowledge Graph Reasoning
von: Xing, Wang, et al.
Veröffentlicht: (2026)
von: Xing, Wang, et al.
Veröffentlicht: (2026)
Reason-Align-Respond: Aligning LLM Reasoning with Knowledge Graphs for KGQA
von: Shen, Xiangqing, et al.
Veröffentlicht: (2025)
von: Shen, Xiangqing, et al.
Veröffentlicht: (2025)
Enhancing LLM's Cognition via Structurization
von: Liu, Kai, et al.
Veröffentlicht: (2024)
von: Liu, Kai, et al.
Veröffentlicht: (2024)
Thinking with DistilQwen: A Tale of Four Distilled Reasoning and Reward Model Series
von: Cai, Wenrui, et al.
Veröffentlicht: (2025)
von: Cai, Wenrui, et al.
Veröffentlicht: (2025)
When and Where Did it Happen? An Encoder-Decoder Model to Identify Scenario Context
von: Noriega-Atala, Enrique, et al.
Veröffentlicht: (2024)
von: Noriega-Atala, Enrique, et al.
Veröffentlicht: (2024)
Controlling Thinking Speed in Reasoning Models
von: Lin, Zhengkai, et al.
Veröffentlicht: (2025)
von: Lin, Zhengkai, et al.
Veröffentlicht: (2025)
Improving Complex Reasoning with Dynamic Prompt Corruption: A soft prompt Optimization Approach
von: Fan, Sinan, et al.
Veröffentlicht: (2025)
von: Fan, Sinan, et al.
Veröffentlicht: (2025)
OmniThoughtVis: A Scalable Distillation Pipeline for Deployable Multimodal Reasoning Models
von: Yue, Yuanhao, et al.
Veröffentlicht: (2026)
von: Yue, Yuanhao, et al.
Veröffentlicht: (2026)
Self-Enhanced Reasoning Training: Activating Latent Reasoning in Small Models for Enhanced Reasoning Distillation
von: Zhang, Yong, et al.
Veröffentlicht: (2025)
von: Zhang, Yong, et al.
Veröffentlicht: (2025)
ReasonOps: Operator Segmentation for LLM Reasoning Traces
von: Lee, Daniel, et al.
Veröffentlicht: (2026)
von: Lee, Daniel, et al.
Veröffentlicht: (2026)
SuperCorrect: Advancing Small LLM Reasoning with Thought Template Distillation and Self-Correction
von: Yang, Ling, et al.
Veröffentlicht: (2024)
von: Yang, Ling, et al.
Veröffentlicht: (2024)
HypER: Literature-grounded Hypothesis Generation and Distillation with Provenance
von: Vasu, Rosni, et al.
Veröffentlicht: (2025)
von: Vasu, Rosni, et al.
Veröffentlicht: (2025)
SAPO: Self-Adaptive Process Optimization Makes Small Reasoners Stronger
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2026)
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
On the Step Length Confounding in LLM Reasoning Data Selection
von: Wang, Bing, et al.
Veröffentlicht: (2026) -
Distribution-Aligned Sequence Distillation for Superior Long-CoT Reasoning
von: Yan, Shaotian, et al.
Veröffentlicht: (2026) -
Backtracking When It Strays: Mitigating Dual Exposure Biases in LLM Reasoning Distillation
von: Wang, Bing, et al.
Veröffentlicht: (2026) -
Are Rationales Necessary and Sufficient? Tuning LLMs for Explainable Misinformation Detection
von: Wang, Bing, et al.
Veröffentlicht: (2026) -
Don't Take Things Out of Context: Attention Intervention for Enhancing Chain-of-Thought Reasoning in Large Language Models
von: Yan, Shaotian, et al.
Veröffentlicht: (2025)