Enhancing Chain-of-Thought Reasoning with Critical Representation Fine-tuning
Fuente:
arXiv
Salvato in:
| Autori principali: | Huang, Chenxi, Yan, Shaotian, Xie, Liang, Lin, Binbin, Fan, Sinan, Xin, Yue, Cai, Deng, Shen, Chen, Ye, Jieping |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SalaMAnder: Shapley-based Mathematical Expression Attribution and Metric for Chain-of-Thought Reasoning
di: Xin, Yue, et al.
Pubblicazione: (2025)
di: Xin, Yue, et al.
Pubblicazione: (2025)
Don't Take Things Out of Context: Attention Intervention for Enhancing Chain-of-Thought Reasoning in Large Language Models
di: Yan, Shaotian, et al.
Pubblicazione: (2025)
di: Yan, Shaotian, et al.
Pubblicazione: (2025)
Instance-adaptive Zero-shot Chain-of-Thought Prompting
di: Yuan, Xiaosong, et al.
Pubblicazione: (2024)
di: Yuan, Xiaosong, et al.
Pubblicazione: (2024)
Distribution-Aligned Sequence Distillation for Superior Long-CoT Reasoning
di: Yan, Shaotian, et al.
Pubblicazione: (2026)
di: Yan, Shaotian, et al.
Pubblicazione: (2026)
Concise and Organized Perception Facilitates Reasoning in Large Language Models
di: Liu, Junjie, et al.
Pubblicazione: (2023)
di: Liu, Junjie, et al.
Pubblicazione: (2023)
Enhancing Spatial Reasoning through Visual and Textual Thinking
di: Liang, Xun, et al.
Pubblicazione: (2025)
di: Liang, Xun, et al.
Pubblicazione: (2025)
Backtracking When It Strays: Mitigating Dual Exposure Biases in LLM Reasoning Distillation
di: Wang, Bing, et al.
Pubblicazione: (2026)
di: Wang, Bing, et al.
Pubblicazione: (2026)
On the Step Length Confounding in LLM Reasoning Data Selection
di: Wang, Bing, et al.
Pubblicazione: (2026)
di: Wang, Bing, et al.
Pubblicazione: (2026)
Improving Complex Reasoning with Dynamic Prompt Corruption: A soft prompt Optimization Approach
di: Fan, Sinan, et al.
Pubblicazione: (2025)
di: Fan, Sinan, et al.
Pubblicazione: (2025)
Where Did This Sentence Come From? Tracing Provenance in LLM Reasoning Distillation
di: Liu, Kaiyuan, et al.
Pubblicazione: (2025)
di: Liu, Kaiyuan, et al.
Pubblicazione: (2025)
From Redundancy to Relevance: Information Flow in LVLMs Across Reasoning Tasks
di: Zhang, Xiaofeng, et al.
Pubblicazione: (2024)
di: Zhang, Xiaofeng, et al.
Pubblicazione: (2024)
Through the Lens of Contrast: Self-Improving Visual Reasoning in VLMs
di: Pan, Zhiyu, et al.
Pubblicazione: (2026)
di: Pan, Zhiyu, et al.
Pubblicazione: (2026)
Delving into the Reversal Curse: How Far Can Large Language Models Generalize?
di: Lin, Zhengkai, et al.
Pubblicazione: (2024)
di: Lin, Zhengkai, et al.
Pubblicazione: (2024)
TokenSqueeze: Performance-Preserving Compression for Reasoning LLMs
di: Zhang, Yuxiang, et al.
Pubblicazione: (2025)
di: Zhang, Yuxiang, et al.
Pubblicazione: (2025)
Critical or Compliant? The Double-Edged Sword of Reasoning in Chain-of-Thought Explanations
di: Park, Eunkyu, et al.
Pubblicazione: (2025)
di: Park, Eunkyu, et al.
Pubblicazione: (2025)
Putting on the Thinking Hats: A Survey on Chain of Thought Fine-tuning from the Perspective of Human Reasoning Mechanism
di: Chen, Xiaoshu, et al.
Pubblicazione: (2025)
di: Chen, Xiaoshu, et al.
Pubblicazione: (2025)
TRACT: Regression-Aware Fine-tuning Meets Chain-of-Thought Reasoning for LLM-as-a-Judge
di: Chiang, Cheng-Han, et al.
Pubblicazione: (2025)
di: Chiang, Cheng-Han, et al.
Pubblicazione: (2025)
Attention-guided Fine-tuning of Multimodal Large Language Models Improves Chain-of-Thought Reasoning
di: Sinha, Sanchit, et al.
Pubblicazione: (2026)
di: Sinha, Sanchit, et al.
Pubblicazione: (2026)
Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs
di: Zhang, Xiaofeng, et al.
Pubblicazione: (2024)
di: Zhang, Xiaofeng, et al.
Pubblicazione: (2024)
CoAT: Chain-of-Associated-Thoughts Framework for Enhancing Large Language Models Reasoning
di: Pan, Jianfeng, et al.
Pubblicazione: (2025)
di: Pan, Jianfeng, et al.
Pubblicazione: (2025)
When Visualizing is the First Step to Reasoning: MIRA, a Benchmark for Visual Chain-of-Thought
di: Zhou, Yiyang, et al.
Pubblicazione: (2025)
di: Zhou, Yiyang, et al.
Pubblicazione: (2025)
Unlocking General Long Chain-of-Thought Reasoning Capabilities of Large Language Models via Representation Engineering
di: Tang, Xinyu, et al.
Pubblicazione: (2025)
di: Tang, Xinyu, et al.
Pubblicazione: (2025)
Enhancing Spatial Reasoning in Vision-Language Models via Chain-of-Thought Prompting and Reinforcement Learning
di: Ji, Binbin, et al.
Pubblicazione: (2025)
di: Ji, Binbin, et al.
Pubblicazione: (2025)
Fine-Tuned Thoughts: Leveraging Chain-of-Thought Reasoning for Industrial Asset Health Monitoring
di: Lin, Shuxin, et al.
Pubblicazione: (2025)
di: Lin, Shuxin, et al.
Pubblicazione: (2025)
MM-Verify: Enhancing Multimodal Reasoning with Chain-of-Thought Verification
di: Sun, Linzhuang, et al.
Pubblicazione: (2025)
di: Sun, Linzhuang, et al.
Pubblicazione: (2025)
On the Impact of Fine-Tuning on Chain-of-Thought Reasoning
di: Lobo, Elita, et al.
Pubblicazione: (2024)
di: Lobo, Elita, et al.
Pubblicazione: (2024)
Diffusion of Thoughts: Chain-of-Thought Reasoning in Diffusion Language Models
di: Ye, Jiacheng, et al.
Pubblicazione: (2024)
di: Ye, Jiacheng, et al.
Pubblicazione: (2024)
Neural Chain-of-Thought Search: Searching the Optimal Reasoning Path to Enhance Large Language Models
di: Ling, Guoming, et al.
Pubblicazione: (2026)
di: Ling, Guoming, et al.
Pubblicazione: (2026)
Controlling Thinking Speed in Reasoning Models
di: Lin, Zhengkai, et al.
Pubblicazione: (2025)
di: Lin, Zhengkai, et al.
Pubblicazione: (2025)
When the Chain Breaks: Interactive Diagnosis of LLM Chain-of-Thought Reasoning Errors
di: Chen, Shiwei, et al.
Pubblicazione: (2026)
di: Chen, Shiwei, et al.
Pubblicazione: (2026)
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
di: Jiang, Chaoya, et al.
Pubblicazione: (2025)
di: Jiang, Chaoya, et al.
Pubblicazione: (2025)
Reason from Future: Reverse Thought Chain Enhances LLM Reasoning
di: Xu, Yinlong, et al.
Pubblicazione: (2025)
di: Xu, Yinlong, et al.
Pubblicazione: (2025)
From Yes-Men to Truth-Tellers: Addressing Sycophancy in Large Language Models with Pinpoint Tuning
di: Chen, Wei, et al.
Pubblicazione: (2024)
di: Chen, Wei, et al.
Pubblicazione: (2024)
VITAL: Visual-Semantic Dual Supervision for Enhanced and Interpretable Latent Reasoning in Medical MLLMs
di: Li, Qiaoru, et al.
Pubblicazione: (2026)
di: Li, Qiaoru, et al.
Pubblicazione: (2026)
Legal$Δ$: Enhancing Legal Reasoning in LLMs via Reinforcement Learning with Chain-of-Thought Guided Information Gain
di: Dai, Xin, et al.
Pubblicazione: (2025)
di: Dai, Xin, et al.
Pubblicazione: (2025)
Enhancing Interpretability in Software Change Management with Chain-of-Thought Reasoning
di: Sun, Yongqian, et al.
Pubblicazione: (2025)
di: Sun, Yongqian, et al.
Pubblicazione: (2025)
Argus: Vision-Centric Reasoning with Grounded Chain-of-Thought
di: Man, Yunze, et al.
Pubblicazione: (2025)
di: Man, Yunze, et al.
Pubblicazione: (2025)
Are Rationales Necessary and Sufficient? Tuning LLMs for Explainable Misinformation Detection
di: Wang, Bing, et al.
Pubblicazione: (2026)
di: Wang, Bing, et al.
Pubblicazione: (2026)
Chain-of-Thought Enhanced Shallow Transformers for Wireless Symbol Detection
di: Fan, Li, et al.
Pubblicazione: (2025)
di: Fan, Li, et al.
Pubblicazione: (2025)
Unveiling Confirmation Bias in Chain-of-Thought Reasoning
di: Wan, Yue, et al.
Pubblicazione: (2025)
di: Wan, Yue, et al.
Pubblicazione: (2025)
Documenti analoghi
-
SalaMAnder: Shapley-based Mathematical Expression Attribution and Metric for Chain-of-Thought Reasoning
di: Xin, Yue, et al.
Pubblicazione: (2025) -
Don't Take Things Out of Context: Attention Intervention for Enhancing Chain-of-Thought Reasoning in Large Language Models
di: Yan, Shaotian, et al.
Pubblicazione: (2025) -
Instance-adaptive Zero-shot Chain-of-Thought Prompting
di: Yuan, Xiaosong, et al.
Pubblicazione: (2024) -
Distribution-Aligned Sequence Distillation for Superior Long-CoT Reasoning
di: Yan, Shaotian, et al.
Pubblicazione: (2026) -
Concise and Organized Perception Facilitates Reasoning in Large Language Models
di: Liu, Junjie, et al.
Pubblicazione: (2023)