TopKD: Top-scaled Knowledge Distillation
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Qi, Zhou, Jinjia |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-perspective Contrastive Logit Distillation
by: Wang, Qi, et al.
Published: (2024)
by: Wang, Qi, et al.
Published: (2024)
KD-DETR: Knowledge Distillation for Detection Transformer with Consistent Distillation Points Sampling
by: Wang, Yu, et al.
Published: (2022)
by: Wang, Yu, et al.
Published: (2022)
CrossKD: Cross-Head Knowledge Distillation for Object Detection
by: Wang, Jiabao, et al.
Published: (2023)
by: Wang, Jiabao, et al.
Published: (2023)
MoKD: Multi-Task Optimization for Knowledge Distillation
by: Hayder, Zeeshan, et al.
Published: (2025)
by: Hayder, Zeeshan, et al.
Published: (2025)
EA-KD: Entropy-based Adaptive Knowledge Distillation
by: Su, Chi-Ping, et al.
Published: (2023)
by: Su, Chi-Ping, et al.
Published: (2023)
BD-KD: Balancing the Divergences for Online Knowledge Distillation
by: Amara, Ibtihel, et al.
Published: (2022)
by: Amara, Ibtihel, et al.
Published: (2022)
Switch-KD: Visual-Switch Knowledge Distillation for Vision-Language Models
by: Sun, Haoyi, et al.
Published: (2026)
by: Sun, Haoyi, et al.
Published: (2026)
ACAM-KD: Adaptive and Cooperative Attention Masking for Knowledge Distillation
by: Lan, Qizhen, et al.
Published: (2025)
by: Lan, Qizhen, et al.
Published: (2025)
FreeKD: Knowledge Distillation via Semantic Frequency Prompt
by: Zhang, Yuan, et al.
Published: (2023)
by: Zhang, Yuan, et al.
Published: (2023)
ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images
by: Ge, Hongyu, et al.
Published: (2025)
by: Ge, Hongyu, et al.
Published: (2025)
VLM-KD: Knowledge Distillation from VLM for Long-Tail Visual Recognition
by: Zhang, Zaiwei, et al.
Published: (2024)
by: Zhang, Zaiwei, et al.
Published: (2024)
MST-KD: Multiple Specialized Teachers Knowledge Distillation for Fair Face Recognition
by: Caldeira, Eduarda, et al.
Published: (2024)
by: Caldeira, Eduarda, et al.
Published: (2024)
TIE-KD: Teacher-Independent and Explainable Knowledge Distillation for Monocular Depth Estimation
by: Choi, Sangwon, et al.
Published: (2024)
by: Choi, Sangwon, et al.
Published: (2024)
StableKD: Breaking Inter-block Optimization Entanglement for Stable Knowledge Distillation
by: Kao, Shiu-hong, et al.
Published: (2023)
by: Kao, Shiu-hong, et al.
Published: (2023)
Top2Pano: Learning to Generate Indoor Panoramas from Top-Down View
by: Zhang, Zitong, et al.
Published: (2025)
by: Zhang, Zitong, et al.
Published: (2025)
SpargeAttention2: Trainable Sparse Attention via Hybrid Top-k+Top-p Masking and Distillation Fine-Tuning
by: Zhang, Jintao, et al.
Published: (2026)
by: Zhang, Jintao, et al.
Published: (2026)
Adaptive Sampling Scheduler
by: Wang, Qi, et al.
Published: (2025)
by: Wang, Qi, et al.
Published: (2025)
TopViewRS: Vision-Language Models as Top-View Spatial Reasoners
by: Li, Chengzu, et al.
Published: (2024)
by: Li, Chengzu, et al.
Published: (2024)
CustomKD: Customizing Large Vision Foundation for Edge Model Improvement via Knowledge Distillation
by: Lee, Jungsoo, et al.
Published: (2025)
by: Lee, Jungsoo, et al.
Published: (2025)
CanKD: Cross-Attention-based Non-local operation for Feature-based Knowledge Distillation
by: Sun, Shizhe, et al.
Published: (2025)
by: Sun, Shizhe, et al.
Published: (2025)
AuG-KD: Anchor-Based Mixup Generation for Out-of-Domain Knowledge Distillation
by: Tang, Zihao, et al.
Published: (2024)
by: Tang, Zihao, et al.
Published: (2024)
AI-KD: Towards Alignment Invariant Face Image Quality Assessment Using Knowledge Distillation
by: Babnik, Žiga, et al.
Published: (2024)
by: Babnik, Žiga, et al.
Published: (2024)
ComKD-CLIP: Comprehensive Knowledge Distillation for Contrastive Language-Image Pre-traning Model
by: Chen, Yifan, et al.
Published: (2024)
by: Chen, Yifan, et al.
Published: (2024)
DiffKD-DCIS: Predicting Upgrade of Ductal Carcinoma In Situ with Diffusion Augmentation and Knowledge Distillation
by: Li, Tao, et al.
Published: (2026)
by: Li, Tao, et al.
Published: (2026)
Leveraging Bottom-Up and Top-Down Attention for Few-Shot Object Detection
by: Chen, Xianyu, et al.
Published: (2020)
by: Chen, Xianyu, et al.
Published: (2020)
m2mKD: Module-to-Module Knowledge Distillation for Modular Transformers
by: Lo, Ka Man, et al.
Published: (2024)
by: Lo, Ka Man, et al.
Published: (2024)
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders
by: Cao, Jiajun, et al.
Published: (2025)
by: Cao, Jiajun, et al.
Published: (2025)
AI-KD: Adversarial learning and Implicit regularization for self-Knowledge Distillation
by: Kim, Hyungmin, et al.
Published: (2022)
by: Kim, Hyungmin, et al.
Published: (2022)
GaitKD: A Universal Decoupled Distillation Framework for Efficient Gait Recognition
by: Li, Yuqi, et al.
Published: (2026)
by: Li, Yuqi, et al.
Published: (2026)
DocKD: Knowledge Distillation from LLMs for Open-World Document Understanding Models
by: Kim, Sungnyun, et al.
Published: (2024)
by: Kim, Sungnyun, et al.
Published: (2024)
PromptKD: Unsupervised Prompt Distillation for Vision-Language Models
by: Li, Zheng, et al.
Published: (2024)
by: Li, Zheng, et al.
Published: (2024)
TransKD: Transformer Knowledge Distillation for Efficient Semantic Segmentation
by: Liu, Ruiping, et al.
Published: (2022)
by: Liu, Ruiping, et al.
Published: (2022)
CLIP-KD: An Empirical Study of CLIP Model Distillation
by: Yang, Chuanguang, et al.
Published: (2023)
by: Yang, Chuanguang, et al.
Published: (2023)
MapKD: Unlocking Prior Knowledge with Cross-Modal Distillation for Efficient Online HD Map Construction
by: Yan, Ziyang, et al.
Published: (2025)
by: Yan, Ziyang, et al.
Published: (2025)
Top-Down Framework for Weakly-supervised Grounded Image Captioning
by: Cai, Chen, et al.
Published: (2023)
by: Cai, Chen, et al.
Published: (2023)
AgriKD: Cross-Architecture Knowledge Distillation for Efficient Leaf Disease Classification
by: Le, Minh-Dung, et al.
Published: (2026)
by: Le, Minh-Dung, et al.
Published: (2026)
Drive-KD: Multi-Teacher Distillation for VLMs in Autonomous Driving
by: Lian, Weitong, et al.
Published: (2026)
by: Lian, Weitong, et al.
Published: (2026)
ReCo-KD: Region- and Context-Aware Knowledge Distillation for Efficient 3D Medical Image Segmentation
by: Lan, Qizhen, et al.
Published: (2026)
by: Lan, Qizhen, et al.
Published: (2026)
Coarse Semantic Injection for LLM-Conditioned Structured Indoor Prediction
by: Zhu, Shuliang, et al.
Published: (2026)
by: Zhu, Shuliang, et al.
Published: (2026)
Block based Adaptive Compressive Sensing with Sampling Rate Control
by: Iwama, Kosuke, et al.
Published: (2024)
by: Iwama, Kosuke, et al.
Published: (2024)
Similar Items
-
Multi-perspective Contrastive Logit Distillation
by: Wang, Qi, et al.
Published: (2024) -
KD-DETR: Knowledge Distillation for Detection Transformer with Consistent Distillation Points Sampling
by: Wang, Yu, et al.
Published: (2022) -
CrossKD: Cross-Head Knowledge Distillation for Object Detection
by: Wang, Jiabao, et al.
Published: (2023) -
MoKD: Multi-Task Optimization for Knowledge Distillation
by: Hayder, Zeeshan, et al.
Published: (2025) -
EA-KD: Entropy-based Adaptive Knowledge Distillation
by: Su, Chi-Ping, et al.
Published: (2023)