Gespeichert in:
| Hauptverfasser: | Shi, Boyu, Jiang, YiCheng, Liu, Chang, Wang, Qiufeng, Yang, Xu, Geng, Xin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2605.07783 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
XPERT: Expert Knowledge Transfer for Effective Training of Language Models
von: Liu, Chang, et al.
Veröffentlicht: (2026)
von: Liu, Chang, et al.
Veröffentlicht: (2026)
Learngene Search Across Multiple Datasets for Building Variable-Sized Models
von: Shi, Boyu, et al.
Veröffentlicht: (2026)
von: Shi, Boyu, et al.
Veröffentlicht: (2026)
Understanding Performance Collapse in Layer-Pruned Large Language Models via Decision Representation Transitions
von: Shi, Boyu, et al.
Veröffentlicht: (2026)
von: Shi, Boyu, et al.
Veröffentlicht: (2026)
Thinking-Free Policy Initialization Makes Distilled Reasoning Models More Effective and Efficient Reasoners
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
Reionization History and Neutrino Mass
von: Dai, YiCheng, et al.
Veröffentlicht: (2026)
von: Dai, YiCheng, et al.
Veröffentlicht: (2026)
Constraint on Neutrino Statistics from Cosmological Data
von: Dai, YiCheng, et al.
Veröffentlicht: (2025)
von: Dai, YiCheng, et al.
Veröffentlicht: (2025)
SOD: Step-wise On-policy Distillation for Small Language Model Agents
von: Zhong, Qiyong, et al.
Veröffentlicht: (2026)
von: Zhong, Qiyong, et al.
Veröffentlicht: (2026)
Cross-Cultural Expert-Level Art Critique Evaluation with Vision-Language Models
von: Yu, Haorui, et al.
Veröffentlicht: (2026)
von: Yu, Haorui, et al.
Veröffentlicht: (2026)
Effectiveness of Chain-of-Thought in Distilling Reasoning Capability from Large Language Models
von: Do, Cong-Thanh, et al.
Veröffentlicht: (2025)
von: Do, Cong-Thanh, et al.
Veröffentlicht: (2025)
Symbolic Chain-of-Thought Distillation: Small Models Can Also "Think" Step-by-Step
von: Li, Liunian Harold, et al.
Veröffentlicht: (2023)
von: Li, Liunian Harold, et al.
Veröffentlicht: (2023)
An Analysis for Reasoning Bias of Language Models with Small Initialization
von: Yao, Junjie, et al.
Veröffentlicht: (2025)
von: Yao, Junjie, et al.
Veröffentlicht: (2025)
Seeing Symbols, Missing Cultures: Probing Vision-Language Models' Reasoning on Fire Imagery and Cultural Meaning
von: Yu, Haorui, et al.
Veröffentlicht: (2025)
von: Yu, Haorui, et al.
Veröffentlicht: (2025)
Balanced Actor Initialization: Stable RLHF Training of Distillation-Based Reasoning Models
von: Zheng, Chen, et al.
Veröffentlicht: (2025)
von: Zheng, Chen, et al.
Veröffentlicht: (2025)
Distilling Mathematical Reasoning Capabilities into Small Language Models
von: Zhu, Xunyu, et al.
Veröffentlicht: (2024)
von: Zhu, Xunyu, et al.
Veröffentlicht: (2024)
SLMRec: Distilling Large Language Models into Small for Sequential Recommendation
von: Xu, Wujiang, et al.
Veröffentlicht: (2024)
von: Xu, Wujiang, et al.
Veröffentlicht: (2024)
InfiR : Crafting Effective Small Language Models and Multimodal Small Language Models in Reasoning
von: Xie, Congkai, et al.
Veröffentlicht: (2025)
von: Xie, Congkai, et al.
Veröffentlicht: (2025)
Adam: Dense Retrieval Distillation with Adaptive Dark Examples
von: Tao, Chongyang, et al.
Veröffentlicht: (2022)
von: Tao, Chongyang, et al.
Veröffentlicht: (2022)
Learning to Rank Chain-of-Thought: Using a Small Model
von: Jiang, Eric Hanchen, et al.
Veröffentlicht: (2025)
von: Jiang, Eric Hanchen, et al.
Veröffentlicht: (2025)
Progressively Label Enhancement for Large Language Model Alignment
von: Liu, Biao, et al.
Veröffentlicht: (2024)
von: Liu, Biao, et al.
Veröffentlicht: (2024)
CANDLE: Iterative Conceptualization and Instantiation Distillation from Large Language Models for Commonsense Reasoning
von: Wang, Weiqi, et al.
Veröffentlicht: (2024)
von: Wang, Weiqi, et al.
Veröffentlicht: (2024)
Preference Orchestrator: Prompt-Aware Multi-Objective Alignment for Large Language Models
von: Liu, Biao, et al.
Veröffentlicht: (2025)
von: Liu, Biao, et al.
Veröffentlicht: (2025)
Learning to Maximize Mutual Information for Chain-of-Thought Distillation
von: Chen, Xin, et al.
Veröffentlicht: (2024)
von: Chen, Xin, et al.
Veröffentlicht: (2024)
Self and Cross-Model Distillation for LLMs: Effective Methods for Refusal Pattern Alignment
von: Li, Jie, et al.
Veröffentlicht: (2024)
von: Li, Jie, et al.
Veröffentlicht: (2024)
BOLT: Bootstrap Long Chain-of-Thought in Language Models without Distillation
von: Pang, Bo, et al.
Veröffentlicht: (2025)
von: Pang, Bo, et al.
Veröffentlicht: (2025)
Enhancing Generalization in Chain of Thought Reasoning for Smaller Models
von: Yin, Maxwell J., et al.
Veröffentlicht: (2025)
von: Yin, Maxwell J., et al.
Veröffentlicht: (2025)
Capturing Nuanced Preferences: Preference-Aligned Distillation for Small Language Models
von: Gu, Yanggan, et al.
Veröffentlicht: (2025)
von: Gu, Yanggan, et al.
Veröffentlicht: (2025)
Neural-Symbolic Collaborative Distillation: Advancing Small Language Models for Complex Reasoning Tasks
von: Liao, Huanxuan, et al.
Veröffentlicht: (2024)
von: Liao, Huanxuan, et al.
Veröffentlicht: (2024)
Can Small Language Models Help Large Language Models Reason Better?: LM-Guided Chain-of-Thought
von: Lee, Jooyoung, et al.
Veröffentlicht: (2024)
von: Lee, Jooyoung, et al.
Veröffentlicht: (2024)
Advantage-Guided Distillation for Preference Alignment in Small Language Models
von: Gao, Shiping, et al.
Veröffentlicht: (2025)
von: Gao, Shiping, et al.
Veröffentlicht: (2025)
VULCA-Bench: A Multicultural Vision-Language Benchmark for Evaluating Cultural Understanding
von: Yu, Haorui, et al.
Veröffentlicht: (2026)
von: Yu, Haorui, et al.
Veröffentlicht: (2026)
ChainLM: Empowering Large Language Models with Improved Chain-of-Thought Prompting
von: Cheng, Xiaoxue, et al.
Veröffentlicht: (2024)
von: Cheng, Xiaoxue, et al.
Veröffentlicht: (2024)
Scaling Law for Language Models Training Considering Batch Size
von: Shuai, Xian, et al.
Veröffentlicht: (2024)
von: Shuai, Xian, et al.
Veröffentlicht: (2024)
MeTA-LoRA: Data-Efficient Multi-Task Fine-Tuning for Large Language Models
von: Cheng, Bo, et al.
Veröffentlicht: (2025)
von: Cheng, Bo, et al.
Veröffentlicht: (2025)
Reasoning-preserved Efficient Distillation of Large Language Models via Activation-aware Initialization
von: He, Junlin, et al.
Veröffentlicht: (2026)
von: He, Junlin, et al.
Veröffentlicht: (2026)
ROSD: Reflective On-Policy Self-Distillation for Language Model Reasoning across Domains
von: Zhao, Ziqi, et al.
Veröffentlicht: (2026)
von: Zhao, Ziqi, et al.
Veröffentlicht: (2026)
EasyDistill: A Comprehensive Toolkit for Effective Knowledge Distillation of Large Language Models
von: Wang, Chengyu, et al.
Veröffentlicht: (2025)
von: Wang, Chengyu, et al.
Veröffentlicht: (2025)
Exploring Learngene via Stage-wise Weight Sharing for Initializing Variable-sized Models
von: Xia, Shi-Yu, et al.
Veröffentlicht: (2024)
von: Xia, Shi-Yu, et al.
Veröffentlicht: (2024)
FedCoT: Federated Chain-of-Thought Distillation for Large Language Models
von: Fan, Tao, et al.
Veröffentlicht: (2024)
von: Fan, Tao, et al.
Veröffentlicht: (2024)
Small Language Models as Effective Guides for Large Language Models in Chinese Relation Extraction
von: Tang, Xuemei, et al.
Veröffentlicht: (2024)
von: Tang, Xuemei, et al.
Veröffentlicht: (2024)
A Structured Framework for Evaluating and Enhancing Interpretive Capabilities of Multimodal LLMs in Culturally Situated Tasks
von: Yu, Haorui, et al.
Veröffentlicht: (2025)
von: Yu, Haorui, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
XPERT: Expert Knowledge Transfer for Effective Training of Language Models
von: Liu, Chang, et al.
Veröffentlicht: (2026) -
Learngene Search Across Multiple Datasets for Building Variable-Sized Models
von: Shi, Boyu, et al.
Veröffentlicht: (2026) -
Understanding Performance Collapse in Layer-Pruned Large Language Models via Decision Representation Transitions
von: Shi, Boyu, et al.
Veröffentlicht: (2026) -
Thinking-Free Policy Initialization Makes Distilled Reasoning Models More Effective and Efficient Reasoners
von: Xu, Xin, et al.
Veröffentlicht: (2025) -
Reionization History and Neutrino Mass
von: Dai, YiCheng, et al.
Veröffentlicht: (2026)