STEPER: Step-wise Knowledge Distillation for Enhancing Reasoning Ability in Multi-Step Retrieval-Augmented Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Lee, Kyumin, Jeon, Minjin, Jang, Sanghwan, Yu, Hwanjo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Rectifying Demonstration Shortcut in In-Context Learning
di: Jang, Joonwon, et al.
Pubblicazione: (2024)
di: Jang, Joonwon, et al.
Pubblicazione: (2024)
Exploring Language Model's Code Generation Ability with Auxiliary Functions
di: Lee, Seonghyeon, et al.
Pubblicazione: (2024)
di: Lee, Seonghyeon, et al.
Pubblicazione: (2024)
REFIND at SemEval-2025 Task 3: Retrieval-Augmented Factuality Hallucination Detection in Large Language Models
di: Lee, DongGeon, et al.
Pubblicazione: (2025)
di: Lee, DongGeon, et al.
Pubblicazione: (2025)
Step-by-Step Mastery: Enhancing Soft Constraint Following Ability of Large Language Models
di: Ren, Qingyu, et al.
Pubblicazione: (2025)
di: Ren, Qingyu, et al.
Pubblicazione: (2025)
SOD: Step-wise On-policy Distillation for Small Language Model Agents
di: Zhong, Qiyong, et al.
Pubblicazione: (2026)
di: Zhong, Qiyong, et al.
Pubblicazione: (2026)
Enhancing Multi-Step Reasoning Abilities of Language Models through Direct Q-Function Optimization
di: Ji, Kaixuan, et al.
Pubblicazione: (2024)
di: Ji, Kaixuan, et al.
Pubblicazione: (2024)
Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models
di: Patel, Nisarg, et al.
Pubblicazione: (2024)
di: Patel, Nisarg, et al.
Pubblicazione: (2024)
Filling the Gaps: Selective Knowledge Augmentation for LLM Recommenders
di: Lee, Jaehyun, et al.
Pubblicazione: (2026)
di: Lee, Jaehyun, et al.
Pubblicazione: (2026)
Omanic: Towards Step-wise Evaluation of Multi-hop Reasoning in Large Language Models
di: Gu, Xiaojie, et al.
Pubblicazione: (2026)
di: Gu, Xiaojie, et al.
Pubblicazione: (2026)
Empowering Large Language Models for Textual Data Augmentation
di: Li, Yichuan, et al.
Pubblicazione: (2024)
di: Li, Yichuan, et al.
Pubblicazione: (2024)
Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs
di: Lai, Xin, et al.
Pubblicazione: (2024)
di: Lai, Xin, et al.
Pubblicazione: (2024)
Towards Hierarchical Multi-Step Reward Models for Enhanced Reasoning in Large Language Models
di: Wang, Teng, et al.
Pubblicazione: (2025)
di: Wang, Teng, et al.
Pubblicazione: (2025)
Supervised Reinforcement Learning: From Expert Trajectories to Step-wise Reasoning
di: Deng, Yihe, et al.
Pubblicazione: (2025)
di: Deng, Yihe, et al.
Pubblicazione: (2025)
LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models
di: Hao, Shibo, et al.
Pubblicazione: (2024)
di: Hao, Shibo, et al.
Pubblicazione: (2024)
The Impact of Reasoning Step Length on Large Language Models
di: Jin, Mingyu, et al.
Pubblicazione: (2024)
di: Jin, Mingyu, et al.
Pubblicazione: (2024)
How Diversely Can Language Models Solve Problems? Exploring the Algorithmic Diversity of Model-Generated Code
di: Lee, Seonghyeon, et al.
Pubblicazione: (2025)
di: Lee, Seonghyeon, et al.
Pubblicazione: (2025)
Retrieval-Augmented Language Model for Extreme Multi-Label Knowledge Graph Link Prediction
di: Lin, Yu-Hsiang, et al.
Pubblicazione: (2024)
di: Lin, Yu-Hsiang, et al.
Pubblicazione: (2024)
DeepRAG: Thinking to Retrieve Step by Step for Large Language Models
di: Guan, Xinyan, et al.
Pubblicazione: (2025)
di: Guan, Xinyan, et al.
Pubblicazione: (2025)
R3-RAG: Learning Step-by-Step Reasoning and Retrieval for LLMs via Reinforcement Learning
di: Li, Yuan, et al.
Pubblicazione: (2025)
di: Li, Yuan, et al.
Pubblicazione: (2025)
A Multi-faceted Analysis of Cognitive Abilities: Evaluating Prompt Methods with Large Language Models on the CONSORT Checklist
di: Jeon, Sohyeon, et al.
Pubblicazione: (2025)
di: Jeon, Sohyeon, et al.
Pubblicazione: (2025)
Large Language Models for Single-Step and Multi-Step Flight Trajectory Prediction
di: Luo, Kaiwei, et al.
Pubblicazione: (2025)
di: Luo, Kaiwei, et al.
Pubblicazione: (2025)
Infinite Mask Diffusion for Few-Step Distillation
di: Yoo, Jaehoon, et al.
Pubblicazione: (2026)
di: Yoo, Jaehoon, et al.
Pubblicazione: (2026)
DSG-KD: Knowledge Distillation from Domain-Specific to General Language Models
di: Cho, Sangyeon, et al.
Pubblicazione: (2024)
di: Cho, Sangyeon, et al.
Pubblicazione: (2024)
ReaRAG: Knowledge-guided Reasoning Enhances Factuality of Large Reasoning Models with Iterative Retrieval Augmented Generation
di: Lee, Zhicheng, et al.
Pubblicazione: (2025)
di: Lee, Zhicheng, et al.
Pubblicazione: (2025)
Tracking the Limits of Knowledge Propagation: How LLMs Fail at Multi-Step Reasoning with Conflicting Knowledge
di: Feng, Yiyang, et al.
Pubblicazione: (2026)
di: Feng, Yiyang, et al.
Pubblicazione: (2026)
Multi-Step Reasoning with Large Language Models, a Survey
di: Plaat, Aske, et al.
Pubblicazione: (2024)
di: Plaat, Aske, et al.
Pubblicazione: (2024)
Exploring the Role of Reasoning Structures for Constructing Proofs in Multi-Step Natural Language Reasoning with Large Language Models
di: Zheng, Zi'ou, et al.
Pubblicazione: (2024)
di: Zheng, Zi'ou, et al.
Pubblicazione: (2024)
Cross-lingual Collapse: How Language-Centric Foundation Models Shape Reasoning in Large Language Models
di: Park, Cheonbok, et al.
Pubblicazione: (2025)
di: Park, Cheonbok, et al.
Pubblicazione: (2025)
Graph Elicitation for Guiding Multi-Step Reasoning in Large Language Models
di: Park, Jinyoung, et al.
Pubblicazione: (2023)
di: Park, Jinyoung, et al.
Pubblicazione: (2023)
Eliciting Instruction-tuned Code Language Models' Capabilities to Utilize Auxiliary Function for Code Generation
di: Lee, Seonghyeon, et al.
Pubblicazione: (2024)
di: Lee, Seonghyeon, et al.
Pubblicazione: (2024)
A Survey on Enhancing Causal Reasoning Ability of Large Language Models
di: Li, Xin, et al.
Pubblicazione: (2025)
di: Li, Xin, et al.
Pubblicazione: (2025)
Opening the Black Box: A Survey on the Mechanisms of Multi-Step Reasoning in Large Language Models
di: Pan, Liangming, et al.
Pubblicazione: (2026)
di: Pan, Liangming, et al.
Pubblicazione: (2026)
Beyond Imitation: Learning Key Reasoning Steps from Dual Chain-of-Thoughts in Reasoning Distillation
di: Dai, Chengwei, et al.
Pubblicazione: (2024)
di: Dai, Chengwei, et al.
Pubblicazione: (2024)
Step Back to Leap Forward: Self-Backtracking for Boosting Reasoning of Language Models
di: Yang, Xiao-Wen, et al.
Pubblicazione: (2025)
di: Yang, Xiao-Wen, et al.
Pubblicazione: (2025)
Assessing LLM Reasoning Steps via Principal Knowledge Grounding
di: Hwang, Hyeon, et al.
Pubblicazione: (2025)
di: Hwang, Hyeon, et al.
Pubblicazione: (2025)
Learning to Reason for Multi-Step Retrieval of Personal Context in Personalized Question Answering
di: Amirizaniani, Maryam, et al.
Pubblicazione: (2026)
di: Amirizaniani, Maryam, et al.
Pubblicazione: (2026)
Retrieval-Augmented Fine-Tuning With Preference Optimization For Visual Program Generation
di: Kang, Deokhyung, et al.
Pubblicazione: (2025)
di: Kang, Deokhyung, et al.
Pubblicazione: (2025)
MedReason: Eliciting Factual Medical Reasoning Steps in LLMs via Knowledge Graphs
di: Wu, Juncheng, et al.
Pubblicazione: (2025)
di: Wu, Juncheng, et al.
Pubblicazione: (2025)
Step-by-Step Fact Verification System for Medical Claims with Explainable Reasoning
di: Vladika, Juraj, et al.
Pubblicazione: (2025)
di: Vladika, Juraj, et al.
Pubblicazione: (2025)
Verbosity-Aware Rationale Reduction: Effective Reduction of Redundant Rationale via Principled Criteria
di: Jang, Joonwon, et al.
Pubblicazione: (2024)
di: Jang, Joonwon, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Rectifying Demonstration Shortcut in In-Context Learning
di: Jang, Joonwon, et al.
Pubblicazione: (2024) -
Exploring Language Model's Code Generation Ability with Auxiliary Functions
di: Lee, Seonghyeon, et al.
Pubblicazione: (2024) -
REFIND at SemEval-2025 Task 3: Retrieval-Augmented Factuality Hallucination Detection in Large Language Models
di: Lee, DongGeon, et al.
Pubblicazione: (2025) -
Step-by-Step Mastery: Enhancing Soft Constraint Following Ability of Large Language Models
di: Ren, Qingyu, et al.
Pubblicazione: (2025) -
SOD: Step-wise On-policy Distillation for Small Language Model Agents
di: Zhong, Qiyong, et al.
Pubblicazione: (2026)