Saved in:
| Main Author: | Ku, Eugene |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2403.10807 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FedeKD: Energy-Based Gating for Robust Federated Knowledge Distillation under Heterogeneous Settings
by: Nguyen, Quang-Huy, et al.
Published: (2026)
by: Nguyen, Quang-Huy, et al.
Published: (2026)
Stronger Graph Transformer with Regularized Attention Scores
by: Ku, Eugene
Published: (2023)
by: Ku, Eugene
Published: (2023)
On-the-Fly Adaptive Distillation of Transformer to Dual-State Linear Attention
by: Ro, Yeonju, et al.
Published: (2025)
by: Ro, Yeonju, et al.
Published: (2025)
BicKD: Bilateral Contrastive Knowledge Distillation
by: Zhu, Jiangnan, et al.
Published: (2026)
by: Zhu, Jiangnan, et al.
Published: (2026)
Learning to Fly in Seconds
by: Eschmann, Jonas, et al.
Published: (2023)
by: Eschmann, Jonas, et al.
Published: (2023)
MemFly: On-the-Fly Memory Optimization via Information Bottleneck
by: Zhang, Zhenyuan, et al.
Published: (2026)
by: Zhang, Zhenyuan, et al.
Published: (2026)
Sim-Anchored Learning for On-the-Fly Adaptation
by: Mabsout, Bassel El, et al.
Published: (2023)
by: Mabsout, Bassel El, et al.
Published: (2023)
Learning Aerodynamics for the Control of Flying Humanoid Robots
by: Paolino, Antonello, et al.
Published: (2025)
by: Paolino, Antonello, et al.
Published: (2025)
ShiftKD: Benchmarking Knowledge Distillation under Distribution Shift
by: Zhang, Songming, et al.
Published: (2023)
by: Zhang, Songming, et al.
Published: (2023)
RMT-KD: Random Matrix Theoretic Causal Knowledge Distillation
by: Ettori, Davide, et al.
Published: (2025)
by: Ettori, Davide, et al.
Published: (2025)
Adversarial Curriculum Graph-Free Knowledge Distillation for Graph Neural Networks
by: Jia, Yuang, et al.
Published: (2025)
by: Jia, Yuang, et al.
Published: (2025)
GraphKD: Exploring Knowledge Distillation Towards Document Object Detection with Structured Graph Creation
by: Banerjee, Ayan, et al.
Published: (2024)
by: Banerjee, Ayan, et al.
Published: (2024)
SFedKD: Sequential Federated Learning with Discrepancy-Aware Multi-Teacher Knowledge Distillation
by: Xu, Haotian, et al.
Published: (2025)
by: Xu, Haotian, et al.
Published: (2025)
TuneShift-KD: Knowledge Distillation and Transfer for Fine-tuned Models
by: Guan, Yushi, et al.
Published: (2026)
by: Guan, Yushi, et al.
Published: (2026)
S^2-KD: Semantic-Spectral Knowledge Distillation Spatiotemporal Forecasting
by: Wang, Wenshuo, et al.
Published: (2025)
by: Wang, Wenshuo, et al.
Published: (2025)
TabKD: Tabular Knowledge Distillation through Interaction Diversity of Learned Feature Bins
by: Pereira, Shovon Niverd, et al.
Published: (2026)
by: Pereira, Shovon Niverd, et al.
Published: (2026)
MTL-KD: Multi-Task Learning Via Knowledge Distillation for Generalizable Neural Vehicle Routing Solver
by: Zheng, Yuepeng, et al.
Published: (2025)
by: Zheng, Yuepeng, et al.
Published: (2025)
OptiSeq: Ordering Examples On-The-Fly for In-Context Learning
by: Bhope, Rahul Atul, et al.
Published: (2025)
by: Bhope, Rahul Atul, et al.
Published: (2025)
Agile Interception of a Flying Target using Competitive Reinforcement Learning
by: Gavin, Timothée, et al.
Published: (2026)
by: Gavin, Timothée, et al.
Published: (2026)
CLIP-Embed-KD: Computationally Efficient Knowledge Distillation Using Embeddings as Teachers
by: Nair, Lakshmi
Published: (2024)
by: Nair, Lakshmi
Published: (2024)
AI-KD: Adversarial learning and Implicit regularization for self-Knowledge Distillation
by: Kim, Hyungmin, et al.
Published: (2022)
by: Kim, Hyungmin, et al.
Published: (2022)
UNIDEAL: Curriculum Knowledge Distillation Federated Learning
by: Yang, Yuwen, et al.
Published: (2023)
by: Yang, Yuwen, et al.
Published: (2023)
Adv-KD: Adversarial Knowledge Distillation for Faster Diffusion Sampling
by: Mekonnen, Kidist Amde, et al.
Published: (2024)
by: Mekonnen, Kidist Amde, et al.
Published: (2024)
KD-GAT: Combining Knowledge Distillation and Graph Attention Transformer for a Controller Area Network Intrusion Detection System
by: Frenken, Robert, et al.
Published: (2025)
by: Frenken, Robert, et al.
Published: (2025)
Online Adaptation for Flying Quadrotors in Tight Formations
by: Hsieh, Pei-An, et al.
Published: (2025)
by: Hsieh, Pei-An, et al.
Published: (2025)
Whole-Brain Connectomic Graph Model Enables Whole-Body Locomotion Control in Fruit Fly
by: Jin, Zehao, et al.
Published: (2026)
by: Jin, Zehao, et al.
Published: (2026)
AdaKD: Dynamic Knowledge Distillation of ASR models using Adaptive Loss Weighting
by: Ganguly, Shreyan, et al.
Published: (2024)
by: Ganguly, Shreyan, et al.
Published: (2024)
C2G-KD: PCA-Constrained Generator for Data-Free Knowledge Distillation
by: Bengtsson, Magnus, et al.
Published: (2025)
by: Bengtsson, Magnus, et al.
Published: (2025)
DeepKD: A Deeply Decoupled and Denoised Knowledge Distillation Trainer
by: Huang, Haiduo, et al.
Published: (2025)
by: Huang, Haiduo, et al.
Published: (2025)
m2mKD: Module-to-Module Knowledge Distillation for Modular Transformers
by: Lo, Ka Man, et al.
Published: (2024)
by: Lo, Ka Man, et al.
Published: (2024)
MemKD: Memory-Discrepancy Knowledge Distillation for Efficient Time Series Classification
by: Udayangani, Nilushika, et al.
Published: (2026)
by: Udayangani, Nilushika, et al.
Published: (2026)
Collaborative Adaptive Curriculum for Progressive Knowledge Distillation
by: Liu, Jing, et al.
Published: (2026)
by: Liu, Jing, et al.
Published: (2026)
HPM-KD: Hierarchical Progressive Multi-Teacher Framework for Knowledge Distillation and Efficient Model Compression
by: Haase, Gustavo Coelho, et al.
Published: (2025)
by: Haase, Gustavo Coelho, et al.
Published: (2025)
Can a MISL Fly? Analysis and Ingredients for Mutual Information Skill Learning
by: Zheng, Chongyi, et al.
Published: (2024)
by: Zheng, Chongyi, et al.
Published: (2024)
Discovering Global False Negatives On the Fly for Self-supervised Contrastive Learning
by: Balmaseda, Vicente, et al.
Published: (2025)
by: Balmaseda, Vicente, et al.
Published: (2025)
Distribution Transformers: Fast Approximate Bayesian Inference With On-The-Fly Prior Adaptation
by: Whittle, George, et al.
Published: (2025)
by: Whittle, George, et al.
Published: (2025)
FloE: On-the-Fly MoE Inference on Memory-constrained GPU
by: Zhou, Yuxin, et al.
Published: (2025)
by: Zhou, Yuxin, et al.
Published: (2025)
Optimizing Resources for On-the-Fly Label Estimation with Multiple Unknown Medical Experts
by: Bary, Tim, et al.
Published: (2025)
by: Bary, Tim, et al.
Published: (2025)
FedSiKD: Clients Similarity and Knowledge Distillation: Addressing Non-i.i.d. and Constraints in Federated Learning
by: Alsenani, Yousef, et al.
Published: (2024)
by: Alsenani, Yousef, et al.
Published: (2024)
AuG-KD: Anchor-Based Mixup Generation for Out-of-Domain Knowledge Distillation
by: Tang, Zihao, et al.
Published: (2024)
by: Tang, Zihao, et al.
Published: (2024)
Similar Items
-
FedeKD: Energy-Based Gating for Robust Federated Knowledge Distillation under Heterogeneous Settings
by: Nguyen, Quang-Huy, et al.
Published: (2026) -
Stronger Graph Transformer with Regularized Attention Scores
by: Ku, Eugene
Published: (2023) -
On-the-Fly Adaptive Distillation of Transformer to Dual-State Linear Attention
by: Ro, Yeonju, et al.
Published: (2025) -
BicKD: Bilateral Contrastive Knowledge Distillation
by: Zhu, Jiangnan, et al.
Published: (2026) -
Learning to Fly in Seconds
by: Eschmann, Jonas, et al.
Published: (2023)