AdaKD: Dynamic Knowledge Distillation of ASR models using Adaptive Loss Weighting
Fuente:
arXiv
Salvato in:
| Autori principali: | Ganguly, Shreyan, Nayak, Roshan, Rao, Rakshith, Deb, Ujan, AP, Prathosh |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FAST: Feature Aware Similarity Thresholding for Weak Unlearning in Black-Box Generative Models
di: Panda, Subhodip, et al.
Pubblicazione: (2023)
di: Panda, Subhodip, et al.
Pubblicazione: (2023)
Spectral Discovery of Continuous Symmetries via Generalized Fourier Transforms
di: Karjol, Pavan, et al.
Pubblicazione: (2026)
di: Karjol, Pavan, et al.
Pubblicazione: (2026)
AdaGMLP: AdaBoosting GNN-to-MLP Knowledge Distillation
di: Lu, Weigang, et al.
Pubblicazione: (2024)
di: Lu, Weigang, et al.
Pubblicazione: (2024)
S^2-KD: Semantic-Spectral Knowledge Distillation Spatiotemporal Forecasting
di: Wang, Wenshuo, et al.
Pubblicazione: (2025)
di: Wang, Wenshuo, et al.
Pubblicazione: (2025)
Leveraging Data Symmetries to Select an Optimal Subset of Training Data under Label Noise
di: Shubham, Kumar, et al.
Pubblicazione: (2026)
di: Shubham, Kumar, et al.
Pubblicazione: (2026)
KD-EKF: Knowledge-Distilled Adaptive Covariance EKF for Robust UWB/PDR Indoor Localization
di: Yoo, Kyeonghyun, et al.
Pubblicazione: (2026)
di: Yoo, Kyeonghyun, et al.
Pubblicazione: (2026)
CLIP-Embed-KD: Computationally Efficient Knowledge Distillation Using Embeddings as Teachers
di: Nair, Lakshmi
Pubblicazione: (2024)
di: Nair, Lakshmi
Pubblicazione: (2024)
C2G-KD: PCA-Constrained Generator for Data-Free Knowledge Distillation
di: Bengtsson, Magnus, et al.
Pubblicazione: (2025)
di: Bengtsson, Magnus, et al.
Pubblicazione: (2025)
TabKD: Tabular Knowledge Distillation through Interaction Diversity of Learned Feature Bins
di: Pereira, Shovon Niverd, et al.
Pubblicazione: (2026)
di: Pereira, Shovon Niverd, et al.
Pubblicazione: (2026)
AdaSPEC: Selective Knowledge Distillation for Efficient Speculative Decoders
di: Hu, Yuezhou, et al.
Pubblicazione: (2025)
di: Hu, Yuezhou, et al.
Pubblicazione: (2025)
AdaTKG: Adaptive Memory for Temporal Knowledge Graph Reasoning
di: Lee, Seunghan, et al.
Pubblicazione: (2026)
di: Lee, Seunghan, et al.
Pubblicazione: (2026)
LoMOE: Localized Multi-Object Editing via Multi-Diffusion
di: Chakrabarty, Goirik, et al.
Pubblicazione: (2024)
di: Chakrabarty, Goirik, et al.
Pubblicazione: (2024)
Adv-KD: Adversarial Knowledge Distillation for Faster Diffusion Sampling
di: Mekonnen, Kidist Amde, et al.
Pubblicazione: (2024)
di: Mekonnen, Kidist Amde, et al.
Pubblicazione: (2024)
WISER: Weak supervISion and supErvised Representation learning to improve drug response prediction in cancer
di: Shubham, Kumar, et al.
Pubblicazione: (2024)
di: Shubham, Kumar, et al.
Pubblicazione: (2024)
DeepKD: A Deeply Decoupled and Denoised Knowledge Distillation Trainer
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
ReMOVE: A Reference-free Metric for Object Erasure
di: Chandrasekar, Aditya, et al.
Pubblicazione: (2024)
di: Chandrasekar, Aditya, et al.
Pubblicazione: (2024)
PromptKD: Distilling Student-Friendly Knowledge for Generative Language Models via Prompt Tuning
di: Kim, Gyeongman, et al.
Pubblicazione: (2024)
di: Kim, Gyeongman, et al.
Pubblicazione: (2024)
KD-OCT: Efficient Knowledge Distillation for Clinical-Grade Retinal OCT Classification
di: Nourbakhsh, Erfan, et al.
Pubblicazione: (2025)
di: Nourbakhsh, Erfan, et al.
Pubblicazione: (2025)
KD-GAT: Combining Knowledge Distillation and Graph Attention Transformer for a Controller Area Network Intrusion Detection System
di: Frenken, Robert, et al.
Pubblicazione: (2025)
di: Frenken, Robert, et al.
Pubblicazione: (2025)
AdaResNet: Enhancing Residual Networks with Dynamic Weight Adjustment for Improved Feature Integration
di: Su, Hong
Pubblicazione: (2024)
di: Su, Hong
Pubblicazione: (2024)
Adaptive Weighted Loss for Sequential Recommendations on Sparse Domains
di: Mittal, Akshay, et al.
Pubblicazione: (2025)
di: Mittal, Akshay, et al.
Pubblicazione: (2025)
KD-LoRA: A Hybrid Approach to Efficient Fine-Tuning with LoRA and Knowledge Distillation
di: Azimi, Rambod, et al.
Pubblicazione: (2024)
di: Azimi, Rambod, et al.
Pubblicazione: (2024)
A Note on Knowledge Distillation Loss Function for Object Classification
di: Chen, Defang
Pubblicazione: (2021)
di: Chen, Defang
Pubblicazione: (2021)
AdaFRUGAL: Adaptive Memory-Efficient Training with Dynamic Control
di: Bui, Quang-Hung, et al.
Pubblicazione: (2025)
di: Bui, Quang-Hung, et al.
Pubblicazione: (2025)
Collaborative Adaptive Curriculum for Progressive Knowledge Distillation
di: Liu, Jing, et al.
Pubblicazione: (2026)
di: Liu, Jing, et al.
Pubblicazione: (2026)
Dynamic Temperature Scheduler for Knowledge Distillation
di: Islam, Sibgat Ul, et al.
Pubblicazione: (2025)
di: Islam, Sibgat Ul, et al.
Pubblicazione: (2025)
Ada-RS: Adaptive Rejection Sampling for Selective Thinking
di: Ge, Yirou, et al.
Pubblicazione: (2026)
di: Ge, Yirou, et al.
Pubblicazione: (2026)
AdaSwitch: Balancing Exploration and Guidance in Knowledge Distillation via Adaptive Switching
di: Peng, Jingyu, et al.
Pubblicazione: (2025)
di: Peng, Jingyu, et al.
Pubblicazione: (2025)
DisCoM-KD: Cross-Modal Knowledge Distillation via Disentanglement Representation and Adversarial Learning
di: Ienco, Dino, et al.
Pubblicazione: (2024)
di: Ienco, Dino, et al.
Pubblicazione: (2024)
AdaQAT: Adaptive Bit-Width Quantization-Aware Training
di: Gernigon, Cédric, et al.
Pubblicazione: (2024)
di: Gernigon, Cédric, et al.
Pubblicazione: (2024)
AdaProb: Efficient Machine Unlearning via Adaptive Probability
di: Zhao, Zihao, et al.
Pubblicazione: (2024)
di: Zhao, Zihao, et al.
Pubblicazione: (2024)
Ada-Diffuser: Latent-Aware Adaptive Diffusion for Decision-Making
di: Feng, Fan, et al.
Pubblicazione: (2026)
di: Feng, Fan, et al.
Pubblicazione: (2026)
AdaDim: Dimensionality Adaptation for SSL Representational Dynamics
di: Kokilepersaud, Kiran, et al.
Pubblicazione: (2025)
di: Kokilepersaud, Kiran, et al.
Pubblicazione: (2025)
MARK: Memory Augmented Refinement of Knowledge
di: Ganguli, Anish, et al.
Pubblicazione: (2025)
di: Ganguli, Anish, et al.
Pubblicazione: (2025)
AdaFlow: Imitation Learning with Variance-Adaptive Flow-Based Policies
di: Hu, Xixi, et al.
Pubblicazione: (2024)
di: Hu, Xixi, et al.
Pubblicazione: (2024)
AdaCS: Adaptive Normalization for Enhanced Code-Switching ASR
di: Chu, The Chuong, et al.
Pubblicazione: (2025)
di: Chu, The Chuong, et al.
Pubblicazione: (2025)
SAFE-KD: Risk-Controlled Early-Exit Distillation for Vision Backbones
di: Khazem, Salim
Pubblicazione: (2026)
di: Khazem, Salim
Pubblicazione: (2026)
AdaBoN: Adaptive Best-of-N Alignment
di: Raman, Vinod, et al.
Pubblicazione: (2025)
di: Raman, Vinod, et al.
Pubblicazione: (2025)
AdaFed: Fair Federated Learning via Adaptive Common Descent Direction
di: Hamidi, Shayan Mohajer, et al.
Pubblicazione: (2024)
di: Hamidi, Shayan Mohajer, et al.
Pubblicazione: (2024)
KD-MARL: Resource-Aware Knowledge Distillation in Multi-Agent Reinforcement Learning
di: Pavel, Monirul Islam, et al.
Pubblicazione: (2026)
di: Pavel, Monirul Islam, et al.
Pubblicazione: (2026)
Documenti analoghi
-
FAST: Feature Aware Similarity Thresholding for Weak Unlearning in Black-Box Generative Models
di: Panda, Subhodip, et al.
Pubblicazione: (2023) -
Spectral Discovery of Continuous Symmetries via Generalized Fourier Transforms
di: Karjol, Pavan, et al.
Pubblicazione: (2026) -
AdaGMLP: AdaBoosting GNN-to-MLP Knowledge Distillation
di: Lu, Weigang, et al.
Pubblicazione: (2024) -
S^2-KD: Semantic-Spectral Knowledge Distillation Spatiotemporal Forecasting
di: Wang, Wenshuo, et al.
Pubblicazione: (2025) -
Leveraging Data Symmetries to Select an Optimal Subset of Training Data under Label Noise
di: Shubham, Kumar, et al.
Pubblicazione: (2026)