DeepKD: A Deeply Decoupled and Denoised Knowledge Distillation Trainer
Fuente:
arXiv
Salvato in:
| Autori principali: | Huang, Haiduo, Song, Jiangcheng, Zhang, Yadong, Ren, Pengju |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
KernelDNA: Dynamic Kernel Sharing via Decoupled Naive Adapters
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
SelecTKD: Selective Token-Weighted Knowledge Distillation for LLMs
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
Partial Channel Network: Compute Fewer, Perform Better
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
KD-OCT: Efficient Knowledge Distillation for Clinical-Grade Retinal OCT Classification
di: Nourbakhsh, Erfan, et al.
Pubblicazione: (2025)
di: Nourbakhsh, Erfan, et al.
Pubblicazione: (2025)
Adv-KD: Adversarial Knowledge Distillation for Faster Diffusion Sampling
di: Mekonnen, Kidist Amde, et al.
Pubblicazione: (2024)
di: Mekonnen, Kidist Amde, et al.
Pubblicazione: (2024)
DisCoM-KD: Cross-Modal Knowledge Distillation via Disentanglement Representation and Adversarial Learning
di: Ienco, Dino, et al.
Pubblicazione: (2024)
di: Ienco, Dino, et al.
Pubblicazione: (2024)
Nearly Lossless Adaptive Bit Switching
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
GeGS-PCR: Effective and Robust 3D Point Cloud Registration with Two-Stage Color-Enhanced Geometric-3DGS Fusion
di: Tian, Jiayi, et al.
Pubblicazione: (2026)
di: Tian, Jiayi, et al.
Pubblicazione: (2026)
SAFE-KD: Risk-Controlled Early-Exit Distillation for Vision Backbones
di: Khazem, Salim
Pubblicazione: (2026)
di: Khazem, Salim
Pubblicazione: (2026)
Towards Personalized Federated Learning via Comprehensive Knowledge Distillation
di: Wang, Pengju, et al.
Pubblicazione: (2024)
di: Wang, Pengju, et al.
Pubblicazione: (2024)
Privacy-Preserving Model Transcription with Differentially Private Synthetic Distillation
di: Liu, Bochao, et al.
Pubblicazione: (2026)
di: Liu, Bochao, et al.
Pubblicazione: (2026)
PRISM: Diversifying Dataset Distillation by Decoupling Architectural Priors
di: Moser, Brian B., et al.
Pubblicazione: (2025)
di: Moser, Brian B., et al.
Pubblicazione: (2025)
Multi-Label Knowledge Distillation
di: Yang, Penghui, et al.
Pubblicazione: (2023)
di: Yang, Penghui, et al.
Pubblicazione: (2023)
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders
di: Cao, Jiajun, et al.
Pubblicazione: (2025)
di: Cao, Jiajun, et al.
Pubblicazione: (2025)
FerKD: Surgical Label Adaptation for Efficient Distillation
di: Shen, Zhiqiang
Pubblicazione: (2023)
di: Shen, Zhiqiang
Pubblicazione: (2023)
SpecVLM: Fast Speculative Decoding in Vision-Language Models
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
Denoising Score Distillation: From Noisy Diffusion Pretraining to One-Step High-Quality Generation
di: Chen, Tianyu, et al.
Pubblicazione: (2025)
di: Chen, Tianyu, et al.
Pubblicazione: (2025)
Dual-Model Weight Selection and Self-Knowledge Distillation for Medical Image Classification
di: Tsutsumi, Ayaka, et al.
Pubblicazione: (2025)
di: Tsutsumi, Ayaka, et al.
Pubblicazione: (2025)
Dynamic Temperature Knowledge Distillation
di: Wei, Yukang, et al.
Pubblicazione: (2024)
di: Wei, Yukang, et al.
Pubblicazione: (2024)
DistilDoc: Knowledge Distillation for Visually-Rich Document Applications
di: Van Landeghem, Jordy, et al.
Pubblicazione: (2024)
di: Van Landeghem, Jordy, et al.
Pubblicazione: (2024)
Neural Tangent Knowledge Distillation for Optical Convolutional Networks
di: Xiang, Jinlin, et al.
Pubblicazione: (2025)
di: Xiang, Jinlin, et al.
Pubblicazione: (2025)
Personalized Federated Learning via Backbone Self-Distillation
di: Wang, Pengju, et al.
Pubblicazione: (2024)
di: Wang, Pengju, et al.
Pubblicazione: (2024)
The Role of Teacher Calibration in Knowledge Distillation
di: Kim, Suyoung, et al.
Pubblicazione: (2025)
di: Kim, Suyoung, et al.
Pubblicazione: (2025)
On Explaining Knowledge Distillation: Measuring and Visualising the Knowledge Transfer Process
di: Adhane, Gereziher, et al.
Pubblicazione: (2024)
di: Adhane, Gereziher, et al.
Pubblicazione: (2024)
ScaleKD: Strong Vision Transformers Could Be Excellent Teachers
di: Fan, Jiawei, et al.
Pubblicazione: (2024)
di: Fan, Jiawei, et al.
Pubblicazione: (2024)
Generative Dataset Distillation Based on Diffusion Model
di: Su, Duo, et al.
Pubblicazione: (2024)
di: Su, Duo, et al.
Pubblicazione: (2024)
Small Scale Data-Free Knowledge Distillation
di: Liu, He, et al.
Pubblicazione: (2024)
di: Liu, He, et al.
Pubblicazione: (2024)
Self-Supervised Quantization-Aware Knowledge Distillation
di: Zhao, Kaiqi, et al.
Pubblicazione: (2024)
di: Zhao, Kaiqi, et al.
Pubblicazione: (2024)
Efficient Training with Denoised Neural Weights
di: Gong, Yifan, et al.
Pubblicazione: (2024)
di: Gong, Yifan, et al.
Pubblicazione: (2024)
Partial Convolution Meets Visual Attention
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
DKDM: Data-Free Knowledge Distillation for Diffusion Models with Any Architecture
di: Xiang, Qianlong, et al.
Pubblicazione: (2024)
di: Xiang, Qianlong, et al.
Pubblicazione: (2024)
Generative Dataset Distillation Based on Self-knowledge Distillation
di: Li, Longzhen, et al.
Pubblicazione: (2025)
di: Li, Longzhen, et al.
Pubblicazione: (2025)
Privacy-Preserving Student Learning with Differentially Private Data-Free Distillation
di: Liu, Bochao, et al.
Pubblicazione: (2024)
di: Liu, Bochao, et al.
Pubblicazione: (2024)
Good Teachers Explain: Explanation-Enhanced Knowledge Distillation
di: Parchami-Araghi, Amin, et al.
Pubblicazione: (2024)
di: Parchami-Araghi, Amin, et al.
Pubblicazione: (2024)
Improving the Transferability of Adversarial Examples by Inverse Knowledge Distillation
di: Wu, Wenyuan, et al.
Pubblicazione: (2025)
di: Wu, Wenyuan, et al.
Pubblicazione: (2025)
Frequency-Aligned Knowledge Distillation for Lightweight Spatiotemporal Forecasting
di: Li, Yuqi, et al.
Pubblicazione: (2025)
di: Li, Yuqi, et al.
Pubblicazione: (2025)
Improving Diffusion Inverse Problem Solving with Decoupled Noise Annealing
di: Zhang, Bingliang, et al.
Pubblicazione: (2024)
di: Zhang, Bingliang, et al.
Pubblicazione: (2024)
Robust Representation Consistency Model via Contrastive Denoising
di: Lei, Jiachen, et al.
Pubblicazione: (2025)
di: Lei, Jiachen, et al.
Pubblicazione: (2025)
PLD: A Choice-Theoretic List-Wise Knowledge Distillation
di: Bassam, Ejafa, et al.
Pubblicazione: (2025)
di: Bassam, Ejafa, et al.
Pubblicazione: (2025)
Aligning Logits Generatively for Principled Black-Box Knowledge Distillation
di: Ma, Jing, et al.
Pubblicazione: (2022)
di: Ma, Jing, et al.
Pubblicazione: (2022)
Documenti analoghi
-
KernelDNA: Dynamic Kernel Sharing via Decoupled Naive Adapters
di: Huang, Haiduo, et al.
Pubblicazione: (2025) -
SelecTKD: Selective Token-Weighted Knowledge Distillation for LLMs
di: Huang, Haiduo, et al.
Pubblicazione: (2025) -
Partial Channel Network: Compute Fewer, Perform Better
di: Huang, Haiduo, et al.
Pubblicazione: (2025) -
KD-OCT: Efficient Knowledge Distillation for Clinical-Grade Retinal OCT Classification
di: Nourbakhsh, Erfan, et al.
Pubblicazione: (2025) -
Adv-KD: Adversarial Knowledge Distillation for Faster Diffusion Sampling
di: Mekonnen, Kidist Amde, et al.
Pubblicazione: (2024)