Distilling Instruction-following Abilities of Large Language Models with Task-aware Curriculum Planning
Fuente:
arXiv
Salvato in:
| Autori principali: | Yue, Yuanhao, Wang, Chengyu, Huang, Jun, Wang, Peng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DistilQwen2.5: Industrial Practices of Training Distilled Open Lightweight Language Models
di: Wang, Chengyu, et al.
Pubblicazione: (2025)
di: Wang, Chengyu, et al.
Pubblicazione: (2025)
EasyDistill: A Comprehensive Toolkit for Effective Knowledge Distillation of Large Language Models
di: Wang, Chengyu, et al.
Pubblicazione: (2025)
di: Wang, Chengyu, et al.
Pubblicazione: (2025)
Building a Family of Data Augmentation Models for Low-cost LLM Fine-tuning on the Cloud
di: Yue, Yuanhao, et al.
Pubblicazione: (2024)
di: Yue, Yuanhao, et al.
Pubblicazione: (2024)
OmniThoughtVis: A Scalable Distillation Pipeline for Deployable Multimodal Reasoning Models
di: Yue, Yuanhao, et al.
Pubblicazione: (2026)
di: Yue, Yuanhao, et al.
Pubblicazione: (2026)
From Correction to Mastery: Reinforced Distillation of Large Language Model Agents
di: Lyu, Yuanjie, et al.
Pubblicazione: (2025)
di: Lyu, Yuanjie, et al.
Pubblicazione: (2025)
AgenticQwen: Training Small Agentic Language Models with Dual Data Flywheels for Industrial-Scale Tool Use
di: Lyu, Yuanjie, et al.
Pubblicazione: (2026)
di: Lyu, Yuanjie, et al.
Pubblicazione: (2026)
Empowering Cross-lingual Abilities of Instruction-tuned Large Language Models by Translation-following demonstrations
di: Ranaldi, Leonardo, et al.
Pubblicazione: (2023)
di: Ranaldi, Leonardo, et al.
Pubblicazione: (2023)
Eliciting Causal Abilities in Large Language Models for Reasoning Tasks
di: Wang, Yajing, et al.
Pubblicazione: (2024)
di: Wang, Yajing, et al.
Pubblicazione: (2024)
WarriorMath: Enhancing the Mathematical Ability of Large Language Models with a Defect-aware Framework
di: Chen, Yue, et al.
Pubblicazione: (2025)
di: Chen, Yue, et al.
Pubblicazione: (2025)
Constraint-aware Path Planning from Natural Language Instructions Using Large Language Models
di: Shim, Dylan, et al.
Pubblicazione: (2026)
di: Shim, Dylan, et al.
Pubblicazione: (2026)
DELIA: Diversity-Enhanced Learning for Instruction Adaptation in Large Language Models
di: Zeng, Yuanhao, et al.
Pubblicazione: (2024)
di: Zeng, Yuanhao, et al.
Pubblicazione: (2024)
Do Large Language Models Understand Logic or Just Mimick Context?
di: Yan, Junbing, et al.
Pubblicazione: (2024)
di: Yan, Junbing, et al.
Pubblicazione: (2024)
Thinking with DistilQwen: A Tale of Four Distilled Reasoning and Reward Model Series
di: Cai, Wenrui, et al.
Pubblicazione: (2025)
di: Cai, Wenrui, et al.
Pubblicazione: (2025)
Eliciting the Translation Ability of Large Language Models via Multilingual Finetuning with Translation Instructions
di: Li, Jiahuan, et al.
Pubblicazione: (2023)
di: Li, Jiahuan, et al.
Pubblicazione: (2023)
Mock Worlds, Real Skills: Building Small Agentic Language Models with Synthetic Tasks, Simulated Environments, and Rubric-Based Rewards
di: Lyu, Yuanjie, et al.
Pubblicazione: (2026)
di: Lyu, Yuanjie, et al.
Pubblicazione: (2026)
Conifer: Improving Complex Constrained Instruction-Following Ability of Large Language Models
di: Sun, Haoran, et al.
Pubblicazione: (2024)
di: Sun, Haoran, et al.
Pubblicazione: (2024)
Enhancing Reasoning Abilities of Small LLMs with Cognitive Alignment
di: Cai, Wenrui, et al.
Pubblicazione: (2025)
di: Cai, Wenrui, et al.
Pubblicazione: (2025)
Distilling Reasoning Ability from Large Language Models with Adaptive Thinking
di: Chen, Xiaoshu, et al.
Pubblicazione: (2024)
di: Chen, Xiaoshu, et al.
Pubblicazione: (2024)
InFoBench: Evaluating Instruction Following Ability in Large Language Models
di: Qin, Yiwei, et al.
Pubblicazione: (2024)
di: Qin, Yiwei, et al.
Pubblicazione: (2024)
The SIFo Benchmark: Investigating the Sequential Instruction Following Ability of Large Language Models
di: Chen, Xinyi, et al.
Pubblicazione: (2024)
di: Chen, Xinyi, et al.
Pubblicazione: (2024)
Evaluation of Instruction-Following Ability for Large Language Models on Story-Ending Generation
di: Hida, Rem, et al.
Pubblicazione: (2024)
di: Hida, Rem, et al.
Pubblicazione: (2024)
Large Language Models Explore by Latent Distilling
di: Zeng, Yuanhao, et al.
Pubblicazione: (2026)
di: Zeng, Yuanhao, et al.
Pubblicazione: (2026)
DAFNet: Dynamic Auxiliary Fusion for Sequential Model Editing in Large Language Models
di: Zhang, Taolin, et al.
Pubblicazione: (2024)
di: Zhang, Taolin, et al.
Pubblicazione: (2024)
Disentangling Memory and Reasoning Ability in Large Language Models
di: Jin, Mingyu, et al.
Pubblicazione: (2024)
di: Jin, Mingyu, et al.
Pubblicazione: (2024)
InsTALL: Context-aware Instructional Task Assistance with Multi-modal Large Language Models
di: Nguyen, Pha, et al.
Pubblicazione: (2025)
di: Nguyen, Pha, et al.
Pubblicazione: (2025)
Decomposing the Basic Abilities of Large Language Models: Mitigating Cross-Task Interference in Multi-Task Instruct-Tuning
di: Wang, Bing, et al.
Pubblicazione: (2026)
di: Wang, Bing, et al.
Pubblicazione: (2026)
Model Editing Harms General Abilities of Large Language Models: Regularization to the Rescue
di: Gu, Jia-Chen, et al.
Pubblicazione: (2024)
di: Gu, Jia-Chen, et al.
Pubblicazione: (2024)
Psychological Counseling Ability of Large Language Models
di: Peng, Fangyu, et al.
Pubblicazione: (2025)
di: Peng, Fangyu, et al.
Pubblicazione: (2025)
R4: Reinforced Retriever-Reorder-Responder for Retrieval-Augmented Large Language Models
di: Zhang, Taolin, et al.
Pubblicazione: (2024)
di: Zhang, Taolin, et al.
Pubblicazione: (2024)
Extracting and Combining Abilities For Building Multi-lingual Ability-enhanced Large Language Models
di: Chen, Zhipeng, et al.
Pubblicazione: (2024)
di: Chen, Zhipeng, et al.
Pubblicazione: (2024)
LLM-Coordination: Evaluating and Analyzing Multi-agent Coordination Abilities in Large Language Models
di: Agashe, Saaket, et al.
Pubblicazione: (2023)
di: Agashe, Saaket, et al.
Pubblicazione: (2023)
Revisiting Compositional Generalization Capability of Large Language Models Considering Instruction Following Ability
di: Sakai, Yusuke, et al.
Pubblicazione: (2025)
di: Sakai, Yusuke, et al.
Pubblicazione: (2025)
KITE: A Benchmark for Evaluating Korean Instruction-Following Abilities in Large Language Models
di: Kim, Dongjun, et al.
Pubblicazione: (2025)
di: Kim, Dongjun, et al.
Pubblicazione: (2025)
AgentGen: Enhancing Planning Abilities for Large Language Model based Agent via Environment and Task Generation
di: Hu, Mengkang, et al.
Pubblicazione: (2024)
di: Hu, Mengkang, et al.
Pubblicazione: (2024)
From Complex to Simple: Enhancing Multi-Constraint Complex Instruction Following Ability of Large Language Models
di: He, Qianyu, et al.
Pubblicazione: (2024)
di: He, Qianyu, et al.
Pubblicazione: (2024)
Optimizing Language Model's Reasoning Abilities with Weak Supervision
di: Tong, Yongqi, et al.
Pubblicazione: (2024)
di: Tong, Yongqi, et al.
Pubblicazione: (2024)
Token-Efficient Leverage Learning in Large Language Models
di: Zeng, Yuanhao, et al.
Pubblicazione: (2024)
di: Zeng, Yuanhao, et al.
Pubblicazione: (2024)
BrailleLLM: Braille Instruction Tuning with Large Language Models for Braille Domain Tasks
di: Huang, Tianyuan, et al.
Pubblicazione: (2025)
di: Huang, Tianyuan, et al.
Pubblicazione: (2025)
AutoLogi: Automated Generation of Logic Puzzles for Evaluating Reasoning Abilities of Large Language Models
di: Zhu, Qin, et al.
Pubblicazione: (2025)
di: Zhu, Qin, et al.
Pubblicazione: (2025)
D2LLM: Decomposed and Distilled Large Language Models for Semantic Search
di: Liao, Zihan, et al.
Pubblicazione: (2024)
di: Liao, Zihan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
DistilQwen2.5: Industrial Practices of Training Distilled Open Lightweight Language Models
di: Wang, Chengyu, et al.
Pubblicazione: (2025) -
EasyDistill: A Comprehensive Toolkit for Effective Knowledge Distillation of Large Language Models
di: Wang, Chengyu, et al.
Pubblicazione: (2025) -
Building a Family of Data Augmentation Models for Low-cost LLM Fine-tuning on the Cloud
di: Yue, Yuanhao, et al.
Pubblicazione: (2024) -
OmniThoughtVis: A Scalable Distillation Pipeline for Deployable Multimodal Reasoning Models
di: Yue, Yuanhao, et al.
Pubblicazione: (2026) -
From Correction to Mastery: Reinforced Distillation of Large Language Model Agents
di: Lyu, Yuanjie, et al.
Pubblicazione: (2025)