Gespeichert in:
| Hauptverfasser: | Huang, Haiduo, Song, Jiangcheng, Zhang, Yadong, Ren, Pengju |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2510.24021 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DeepKD: A Deeply Decoupled and Denoised Knowledge Distillation Trainer
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
KernelDNA: Dynamic Kernel Sharing via Decoupled Naive Adapters
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
"The Whole Is Greater Than the Sum of Its Parts": A Compatibility-Aware Multi-Teacher CoT Distillation Framework
von: Cui, Jin, et al.
Veröffentlicht: (2026)
von: Cui, Jin, et al.
Veröffentlicht: (2026)
Partial Channel Network: Compute Fewer, Perform Better
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
A Token is Worth over 1,000 Tokens: Efficient Knowledge Distillation through Low-Rank Clone
von: Hao, Jitai, et al.
Veröffentlicht: (2025)
von: Hao, Jitai, et al.
Veröffentlicht: (2025)
Jakiro: Boosting Speculative Decoding with Decoupled Multi-Head via MoE
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
LLM-Oriented Token-Adaptive Knowledge Distillation
von: Xie, Xurong, et al.
Veröffentlicht: (2025)
von: Xie, Xurong, et al.
Veröffentlicht: (2025)
FastEagle: Cascaded Drafting for Accelerating Speculative Decoding
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
Nearly Lossless Adaptive Bit Switching
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
GeGS-PCR: Effective and Robust 3D Point Cloud Registration with Two-Stage Color-Enhanced Geometric-3DGS Fusion
von: Tian, Jiayi, et al.
Veröffentlicht: (2026)
von: Tian, Jiayi, et al.
Veröffentlicht: (2026)
Gumiho: A Hybrid Architecture to Prioritize Early Tokens in Speculative Decoding
von: Li, Jinze, et al.
Veröffentlicht: (2025)
von: Li, Jinze, et al.
Veröffentlicht: (2025)
MedDialBench: Benchmarking LLM Diagnostic Robustness under Parametric Adversarial Patient Behaviors
von: Luo, Xiaotian, et al.
Veröffentlicht: (2026)
von: Luo, Xiaotian, et al.
Veröffentlicht: (2026)
Explain in Your Own Words: Improving Reasoning via Token-Selective Dual Knowledge Distillation
von: Kim, Minsang, et al.
Veröffentlicht: (2026)
von: Kim, Minsang, et al.
Veröffentlicht: (2026)
Pair-In, Pair-Out: Latent Multi-Token Prediction for Efficient LLMs
von: Tan, Wenhui, et al.
Veröffentlicht: (2026)
von: Tan, Wenhui, et al.
Veröffentlicht: (2026)
Less is More: Selective Reflection for Compatible and Efficient Knowledge Distillation in Large Language Models
von: Liu, Lingyuan, et al.
Veröffentlicht: (2025)
von: Liu, Lingyuan, et al.
Veröffentlicht: (2025)
SpecVLM: Fast Speculative Decoding in Vision-Language Models
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
Few-Shot Knowledge Distillation of LLMs With Counterfactual Explanations
von: Hamman, Faisal, et al.
Veröffentlicht: (2025)
von: Hamman, Faisal, et al.
Veröffentlicht: (2025)
AlignDistil: Token-Level Language Model Alignment as Adaptive Policy Distillation
von: Zhang, Songming, et al.
Veröffentlicht: (2025)
von: Zhang, Songming, et al.
Veröffentlicht: (2025)
EasyDistill: A Comprehensive Toolkit for Effective Knowledge Distillation of Large Language Models
von: Wang, Chengyu, et al.
Veröffentlicht: (2025)
von: Wang, Chengyu, et al.
Veröffentlicht: (2025)
AtlasKV: Augmenting LLMs with Billion-Scale Knowledge Graphs in 20GB VRAM
von: Huang, Haoyu, et al.
Veröffentlicht: (2025)
von: Huang, Haoyu, et al.
Veröffentlicht: (2025)
Distilling Reasoning Without Knowledge: A Framework for Reliable LLMs
von: Kietkajornrit, Auksarapak, et al.
Veröffentlicht: (2026)
von: Kietkajornrit, Auksarapak, et al.
Veröffentlicht: (2026)
Unlearning Backdoor Attacks for LLMs with Weak-to-Strong Knowledge Distillation
von: Zhao, Shuai, et al.
Veröffentlicht: (2024)
von: Zhao, Shuai, et al.
Veröffentlicht: (2024)
An Expert is Worth One Token: Synergizing Multiple Expert LLMs as Generalist via Expert Token Routing
von: Chai, Ziwei, et al.
Veröffentlicht: (2024)
von: Chai, Ziwei, et al.
Veröffentlicht: (2024)
RankLLM: Weighted Ranking of LLMs by Quantifying Question Difficulty
von: Zhang, Ziqian, et al.
Veröffentlicht: (2026)
von: Zhang, Ziqian, et al.
Veröffentlicht: (2026)
TokenSelect: Efficient Long-Context Inference and Length Extrapolation for LLMs via Dynamic Token-Level KV Cache Selection
von: Wu, Wei, et al.
Veröffentlicht: (2024)
von: Wu, Wei, et al.
Veröffentlicht: (2024)
Efficient Intent-Based Filtering for Multi-Party Conversations Using Knowledge Distillation from LLMs
von: Gody, Reem, et al.
Veröffentlicht: (2025)
von: Gody, Reem, et al.
Veröffentlicht: (2025)
AdaSPEC: Selective Knowledge Distillation for Efficient Speculative Decoders
von: Hu, Yuezhou, et al.
Veröffentlicht: (2025)
von: Hu, Yuezhou, et al.
Veröffentlicht: (2025)
Hybrid Policy Distillation for LLMs
von: Zhu, Wenhong, et al.
Veröffentlicht: (2026)
von: Zhu, Wenhong, et al.
Veröffentlicht: (2026)
Reliable Reasoning Path: Distilling Effective Guidance for LLM Reasoning with Knowledge Graphs
von: Xiao, Yilin, et al.
Veröffentlicht: (2025)
von: Xiao, Yilin, et al.
Veröffentlicht: (2025)
Dual-Space Knowledge Distillation for Large Language Models
von: Zhang, Songming, et al.
Veröffentlicht: (2024)
von: Zhang, Songming, et al.
Veröffentlicht: (2024)
ConTextual: Improving Clinical Text Summarization in LLMs with Context-preserving Token Filtering and Knowledge Graphs
von: Piya, Fahmida Liza, et al.
Veröffentlicht: (2025)
von: Piya, Fahmida Liza, et al.
Veröffentlicht: (2025)
Multi-Stage Balanced Distillation: Addressing Long-Tail Challenges in Sequence-Level Knowledge Distillation
von: Zhou, Yuhang, et al.
Veröffentlicht: (2024)
von: Zhou, Yuhang, et al.
Veröffentlicht: (2024)
A Diversity-Enhanced Knowledge Distillation Model for Practical Math Word Problem Solving
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency
von: Cai, Aichen, et al.
Veröffentlicht: (2026)
von: Cai, Aichen, et al.
Veröffentlicht: (2026)
LLMs are Not Just Next Token Predictors
von: Downes, Stephen M., et al.
Veröffentlicht: (2024)
von: Downes, Stephen M., et al.
Veröffentlicht: (2024)
Partial Convolution Meets Visual Attention
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
Cornerstones or Stumbling Blocks? Deciphering the Rock Tokens in On-Policy Distillation
von: Jiang, Yuxuan, et al.
Veröffentlicht: (2026)
von: Jiang, Yuxuan, et al.
Veröffentlicht: (2026)
Incorporating Domain Knowledge into Materials Tokenization
von: Oh, Yerim, et al.
Veröffentlicht: (2025)
von: Oh, Yerim, et al.
Veröffentlicht: (2025)
Self-Distillation for Multi-Token Prediction
von: Zhao, Guoliang, et al.
Veröffentlicht: (2026)
von: Zhao, Guoliang, et al.
Veröffentlicht: (2026)
Can LLMs be Good Graph Judge for Knowledge Graph Construction?
von: Huang, Haoyu, et al.
Veröffentlicht: (2024)
von: Huang, Haoyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
DeepKD: A Deeply Decoupled and Denoised Knowledge Distillation Trainer
von: Huang, Haiduo, et al.
Veröffentlicht: (2025) -
KernelDNA: Dynamic Kernel Sharing via Decoupled Naive Adapters
von: Huang, Haiduo, et al.
Veröffentlicht: (2025) -
"The Whole Is Greater Than the Sum of Its Parts": A Compatibility-Aware Multi-Teacher CoT Distillation Framework
von: Cui, Jin, et al.
Veröffentlicht: (2026) -
Partial Channel Network: Compute Fewer, Perform Better
von: Huang, Haiduo, et al.
Veröffentlicht: (2025) -
A Token is Worth over 1,000 Tokens: Efficient Knowledge Distillation through Low-Rank Clone
von: Hao, Jitai, et al.
Veröffentlicht: (2025)