Knowledge Distillation via the Target-aware Transformer
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, Sihao, Xie, Hongwei, Wang, Bing, Yu, Kaicheng, Chang, Xiaojun, Liang, Xiaodan, Wang, Gang |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MLP Can Be A Good Transformer Learner
by: Lin, Sihao, et al.
Published: (2024)
by: Lin, Sihao, et al.
Published: (2024)
Learning A Zero-shot Occupancy Network from Vision Foundation Models via Self-supervised Adaptation
by: Lin, Sihao, et al.
Published: (2025)
by: Lin, Sihao, et al.
Published: (2025)
Efficient Training of Large Vision Models via Advanced Automated Progressive Learning
by: Li, Changlin, et al.
Published: (2024)
by: Li, Changlin, et al.
Published: (2024)
KD-DETR: Knowledge Distillation for Detection Transformer with Consistent Distillation Points Sampling
by: Wang, Yu, et al.
Published: (2022)
by: Wang, Yu, et al.
Published: (2022)
MiLA: Multi-view Intensive-fidelity Long-term Video Generation World Model for Autonomous Driving
by: Wang, Haiguang, et al.
Published: (2025)
by: Wang, Haiguang, et al.
Published: (2025)
Optimizing Knowledge Distillation in Transformers: Enabling Multi-Head Attention without Alignment Barriers
by: Bing, Zhaodong, et al.
Published: (2025)
by: Bing, Zhaodong, et al.
Published: (2025)
Predicting Genetic Mutation from Whole Slide Images via Biomedical-Linguistic Knowledge Enhanced Multi-label Classification
by: Huang, Gexin, et al.
Published: (2024)
by: Huang, Gexin, et al.
Published: (2024)
Self-Supervised Multi-Frame Neural Scene Flow
by: Liu, Dongrui, et al.
Published: (2024)
by: Liu, Dongrui, et al.
Published: (2024)
AlignMiF: Geometry-Aligned Multimodal Implicit Field for LiDAR-Camera Joint Synthesis
by: Tang, Tao, et al.
Published: (2024)
by: Tang, Tao, et al.
Published: (2024)
Knowledge Distillation via Query Selection for Detection Transformer
by: Liu, Yi, et al.
Published: (2024)
by: Liu, Yi, et al.
Published: (2024)
Lightweight Model Pre-training via Language Guided Knowledge Distillation
by: Li, Mingsheng, et al.
Published: (2024)
by: Li, Mingsheng, et al.
Published: (2024)
Teaching with Uncertainty: Unleashing the Potential of Knowledge Distillation in Object Detection
by: Yi, Junfei, et al.
Published: (2024)
by: Yi, Junfei, et al.
Published: (2024)
Ground-R1: Incentivizing Grounded Visual Reasoning via Reinforcement Learning
by: Cao, Meng, et al.
Published: (2025)
by: Cao, Meng, et al.
Published: (2025)
DNA Family: Boosting Weight-Sharing NAS with Block-Wise Supervisions
by: Wang, Guangrun, et al.
Published: (2024)
by: Wang, Guangrun, et al.
Published: (2024)
Making Large Language Models Better Planners with Reasoning-Decision Alignment
by: Huang, Zhijian, et al.
Published: (2024)
by: Huang, Zhijian, et al.
Published: (2024)
Target-aware Bidirectional Fusion Transformer for Aerial Object Tracking
by: Sun, Xinglong, et al.
Published: (2025)
by: Sun, Xinglong, et al.
Published: (2025)
Temporal Separation with Entropy Regularization for Knowledge Distillation in Spiking Neural Networks
by: Yu, Kairong, et al.
Published: (2025)
by: Yu, Kairong, et al.
Published: (2025)
Contrastive Learning with Counterfactual Explanations for Radiology Report Generation
by: Li, Mingjie, et al.
Published: (2024)
by: Li, Mingjie, et al.
Published: (2024)
TransMamba: Fast Universal Architecture Adaption from Transformers to Mamba
by: Chen, Xiuwei, et al.
Published: (2025)
by: Chen, Xiuwei, et al.
Published: (2025)
3D Visibility-aware Generalizable Neural Radiance Fields for Interacting Hands
by: Huang, Xuan, et al.
Published: (2024)
by: Huang, Xuan, et al.
Published: (2024)
Target-aware Image Editing via Cycle-consistent Constraints
by: Wang, Yanghao, et al.
Published: (2025)
by: Wang, Yanghao, et al.
Published: (2025)
Infrared UAV Target Tracking with Dynamic Feature Refinement and Global Contextual Attention Knowledge Distillation
by: Fang, Houzhang, et al.
Published: (2025)
by: Fang, Houzhang, et al.
Published: (2025)
Suppressing Prior-Comparison Hallucinations in Radiology Report Generation via Semantically Decoupled Latent Steering
by: Li, Ao, et al.
Published: (2026)
by: Li, Ao, et al.
Published: (2026)
Robust MLLM Unlearning via Visual Knowledge Distillation
by: Wang, Yuhang, et al.
Published: (2025)
by: Wang, Yuhang, et al.
Published: (2025)
Sitcom-Crafter: A Plot-Driven Human Motion Generation System in 3D Scenes
by: Chen, Jianqi, et al.
Published: (2024)
by: Chen, Jianqi, et al.
Published: (2024)
Learning with Semantic Priors: Stabilizing Point-Supervised Infrared Small Target Detection via Hierarchical Knowledge Distillation
by: Yao, Yuanhang, et al.
Published: (2026)
by: Yao, Yuanhang, et al.
Published: (2026)
VRM: Knowledge Distillation via Virtual Relation Matching
by: Zhang, Weijia, et al.
Published: (2025)
by: Zhang, Weijia, et al.
Published: (2025)
BEVLM: Distilling Semantic Knowledge from LLMs into Bird's-Eye View Representations
by: Monninger, Thomas, et al.
Published: (2026)
by: Monninger, Thomas, et al.
Published: (2026)
Rethinking Knowledge in Distillation: An In-context Sample Retrieval Perspective
by: Zhu, Jinjing, et al.
Published: (2025)
by: Zhu, Jinjing, et al.
Published: (2025)
SAMKD: Spatial-aware Adaptive Masking Knowledge Distillation for Object Detection
by: Zhang, Zhourui, et al.
Published: (2025)
by: Zhang, Zhourui, et al.
Published: (2025)
CLoCKDistill: Consistent Location-and-Context-aware Knowledge Distillation for DETRs
by: Lan, Qizhen, et al.
Published: (2025)
by: Lan, Qizhen, et al.
Published: (2025)
360$^\circ$ High-Resolution Depth Estimation via Uncertainty-aware Structural Knowledge Transfer
by: Cao, Zidong, et al.
Published: (2023)
by: Cao, Zidong, et al.
Published: (2023)
Structure-Centric Robust Monocular Depth Estimation via Knowledge Distillation
by: Chen, Runze, et al.
Published: (2024)
by: Chen, Runze, et al.
Published: (2024)
Target-Driven Distillation: Consistency Distillation with Target Timestep Selection and Decoupled Guidance
by: Wang, Cunzheng, et al.
Published: (2024)
by: Wang, Cunzheng, et al.
Published: (2024)
CLIP-CID: Efficient CLIP Distillation via Cluster-Instance Discrimination
by: Yang, Kaicheng, et al.
Published: (2024)
by: Yang, Kaicheng, et al.
Published: (2024)
VideoDistill: Language-aware Vision Distillation for Video Question Answering
by: Zou, Bo, et al.
Published: (2024)
by: Zou, Bo, et al.
Published: (2024)
Enhancing Adaptive Deep Networks for Image Classification via Uncertainty-aware Decision Fusion
by: Zhang, Xu, et al.
Published: (2024)
by: Zhang, Xu, et al.
Published: (2024)
SRKD: Towards Efficient 3D Point Cloud Segmentation via Structure- and Relation-aware Knowledge Distillation
by: Li, Yuqi, et al.
Published: (2025)
by: Li, Yuqi, et al.
Published: (2025)
Improving Facial Landmark Detection Accuracy and Efficiency with Knowledge Distillation
by: Hong, Zong-Wei, et al.
Published: (2024)
by: Hong, Zong-Wei, et al.
Published: (2024)
A Transformer-in-Transformer Network Utilizing Knowledge Distillation for Image Recognition
by: Rahman, Dewan Tauhid, et al.
Published: (2025)
by: Rahman, Dewan Tauhid, et al.
Published: (2025)
Similar Items
-
MLP Can Be A Good Transformer Learner
by: Lin, Sihao, et al.
Published: (2024) -
Learning A Zero-shot Occupancy Network from Vision Foundation Models via Self-supervised Adaptation
by: Lin, Sihao, et al.
Published: (2025) -
Efficient Training of Large Vision Models via Advanced Automated Progressive Learning
by: Li, Changlin, et al.
Published: (2024) -
KD-DETR: Knowledge Distillation for Detection Transformer with Consistent Distillation Points Sampling
by: Wang, Yu, et al.
Published: (2022) -
MiLA: Multi-view Intensive-fidelity Long-term Video Generation World Model for Autonomous Driving
by: Wang, Haiguang, et al.
Published: (2025)