Data-Efficient Low-Complexity Acoustic Scene Classification via Distilling and Progressive Pruning
Fuente:
arXiv
Guardado en:
| Autores principales: | Han, Bing, Huang, Wen, Chen, Zhengyang, Jiang, Anbai, Fan, Pingyi, Lu, Cheng, Lv, Zhiqiang, Liu, Jia, Zhang, Wei-Qiang, Qian, Yanmin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AnoPatch: Towards Better Consistency in Machine Anomalous Sound Detection
por: Jiang, Anbai, et al.
Publicado: (2024)
por: Jiang, Anbai, et al.
Publicado: (2024)
Exploring Self-Supervised Audio Models for Generalized Anomalous Sound Detection
por: Han, Bing, et al.
Publicado: (2025)
por: Han, Bing, et al.
Publicado: (2025)
Improving Anomalous Sound Detection via Low-Rank Adaptation Fine-Tuning of Pre-Trained Audio Models
por: Zheng, Xinhu, et al.
Publicado: (2024)
por: Zheng, Xinhu, et al.
Publicado: (2024)
Data-Efficient Low-Complexity Acoustic Scene Classification in the DCASE 2024 Challenge
por: Schmid, Florian, et al.
Publicado: (2024)
por: Schmid, Florian, et al.
Publicado: (2024)
Improving Acoustic Scene Classification in Low-Resource Conditions
por: Chen, Zhi, et al.
Publicado: (2024)
por: Chen, Zhi, et al.
Publicado: (2024)
Prototype and Instance Contrastive Learning for Unsupervised Domain Adaptation in Speaker Verification
por: Huang, Wen, et al.
Publicado: (2024)
por: Huang, Wen, et al.
Publicado: (2024)
Low-Complexity Acoustic Scene Classification with Device Information in the DCASE 2025 Challenge
por: Schmid, Florian, et al.
Publicado: (2025)
por: Schmid, Florian, et al.
Publicado: (2025)
Low-Complexity Acoustic Scene Classification Using Parallel Attention-Convolution Network
por: Li, Yanxiong, et al.
Publicado: (2024)
por: Li, Yanxiong, et al.
Publicado: (2024)
TF-SepNet: An Efficient 1D Kernel Design in CNNs for Low-Complexity Acoustic Scene Classification
por: Cai, Yiqiang, et al.
Publicado: (2023)
por: Cai, Yiqiang, et al.
Publicado: (2023)
Flow-TSVAD: Target-Speaker Voice Activity Detection via Latent Flow Matching
por: Chen, Zhengyang, et al.
Publicado: (2024)
por: Chen, Zhengyang, et al.
Publicado: (2024)
CoopASD: Cooperative Machine Anomalous Sound Detection with Privacy Concerns
por: Jiang, Anbai, et al.
Publicado: (2024)
por: Jiang, Anbai, et al.
Publicado: (2024)
Advanced Zero-Shot Text-to-Speech for Background Removal and Preservation with Controllable Masked Speech Prediction
por: Zhang, Leying, et al.
Publicado: (2025)
por: Zhang, Leying, et al.
Publicado: (2025)
Joint Feature and Output Distillation for Low-complexity Acoustic Scene Classification
por: Li, Haowen, et al.
Publicado: (2025)
por: Li, Haowen, et al.
Publicado: (2025)
Leveraging Self-supervised Audio Representations for Data-Efficient Acoustic Scene Classification
por: Cai, Yiqiang, et al.
Publicado: (2024)
por: Cai, Yiqiang, et al.
Publicado: (2024)
Deep Space Separable Distillation for Lightweight Acoustic Scene Classification
por: Ye, ShuQi, et al.
Publicado: (2024)
por: Ye, ShuQi, et al.
Publicado: (2024)
Generating Speakers by Prompting Listener Impressions for Pre-trained Multi-Speaker Text-to-Speech Systems
por: Chen, Zhengyang, et al.
Publicado: (2024)
por: Chen, Zhengyang, et al.
Publicado: (2024)
Creating a Good Teacher for Knowledge Distillation in Acoustic Scene Classification
por: Morocutti, Tobias, et al.
Publicado: (2025)
por: Morocutti, Tobias, et al.
Publicado: (2025)
Overview of Speaker Modeling and Its Applications: From the Lens of Deep Speaker Representation Learning
por: Wang, Shuai, et al.
Publicado: (2024)
por: Wang, Shuai, et al.
Publicado: (2024)
Structural and Statistical Audio Texture Knowledge Distillation for Acoustic Classification
por: Ritu, Jarin, et al.
Publicado: (2025)
por: Ritu, Jarin, et al.
Publicado: (2025)
Disentangling the Prosody and Semantic Information with Pre-trained Model for In-Context Learning based Zero-Shot Voice Conversion
por: Chen, Zhengyang, et al.
Publicado: (2024)
por: Chen, Zhengyang, et al.
Publicado: (2024)
SPADE: Structured Pruning and Adaptive Distillation for Efficient LLM-TTS
por: Nguyen, Tan Dat, et al.
Publicado: (2025)
por: Nguyen, Tan Dat, et al.
Publicado: (2025)
Scale This, Not That: Investigating Key Dataset Attributes for Efficient Speech Enhancement Scaling
por: Zhang, Leying, et al.
Publicado: (2024)
por: Zhang, Leying, et al.
Publicado: (2024)
DQ-Whisper: Joint Distillation and Quantization for Efficient Multilingual Speech Recognition
por: Shao, Hang, et al.
Publicado: (2023)
por: Shao, Hang, et al.
Publicado: (2023)
Memory-Efficient Training for Deep Speaker Embedding Learning in Speaker Verification
por: Liu, Bei, et al.
Publicado: (2024)
por: Liu, Bei, et al.
Publicado: (2024)
SpeechFake: A Large-Scale Multilingual Speech Deepfake Dataset Incorporating Cutting-Edge Generation Methods
por: Huang, Wen, et al.
Publicado: (2025)
por: Huang, Wen, et al.
Publicado: (2025)
Generalizable Audio Deepfake Detection via Latent Space Refinement and Augmentation
por: Huang, Wen, et al.
Publicado: (2025)
por: Huang, Wen, et al.
Publicado: (2025)
From Sharpness to Better Generalization for Speech Deepfake Detection
por: Huang, Wen, et al.
Publicado: (2025)
por: Huang, Wen, et al.
Publicado: (2025)
Towards Lightweight Speaker Verification via Adaptive Neural Network Quantization
por: Liu, Bei, et al.
Publicado: (2024)
por: Liu, Bei, et al.
Publicado: (2024)
A Hybrid Approach for Low-Complexity Joint Acoustic Echo and Noise Reduction
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
Data Efficient Acoustic Scene Classification using Teacher-Informed Confusing Class Instruction
por: Yeo, Jin Jie Sean, et al.
Publicado: (2024)
por: Yeo, Jin Jie Sean, et al.
Publicado: (2024)
BR-ASR: Efficient and Scalable Bias Retrieval Framework for Contextual Biasing ASR in Speech LLM
por: Gong, Xun, et al.
Publicado: (2025)
por: Gong, Xun, et al.
Publicado: (2025)
Enhancing Speaker Verification with w2v-BERT 2.0 and Knowledge Distillation guided Structured Pruning
por: Li, Ze, et al.
Publicado: (2025)
por: Li, Ze, et al.
Publicado: (2025)
Dual-Branch Knowledge Distillation for Noise-Robust Synthetic Speech Detection
por: Fan, Cunhang, et al.
Publicado: (2023)
por: Fan, Cunhang, et al.
Publicado: (2023)
Continual Learning for Acoustic Event Classification
por: Xiao, Yang
Publicado: (2025)
por: Xiao, Yang
Publicado: (2025)
Diffusion-based Generative Modeling with Discriminative Guidance for Streamable Speech Enhancement
por: Li, Chenda, et al.
Publicado: (2024)
por: Li, Chenda, et al.
Publicado: (2024)
Self-Distillation Prototypes Network: Learning Robust Speaker Representations without Supervision
por: Chen, Yafeng, et al.
Publicado: (2024)
por: Chen, Yafeng, et al.
Publicado: (2024)
Distillation and Pruning for Scalable Self-Supervised Representation-Based Speech Quality Assessment
por: Stahl, Benjamin, et al.
Publicado: (2025)
por: Stahl, Benjamin, et al.
Publicado: (2025)
Frequency-mix Knowledge Distillation for Fake Speech Detection
por: Fan, Cunhang, et al.
Publicado: (2024)
por: Fan, Cunhang, et al.
Publicado: (2024)
Improving Design of Input Condition Invariant Speech Enhancement
por: Zhang, Wangyou, et al.
Publicado: (2024)
por: Zhang, Wangyou, et al.
Publicado: (2024)
Online Domain-Incremental Learning Approach to Classify Acoustic Scenes in All Locations
por: Mulimani, Manjunath, et al.
Publicado: (2024)
por: Mulimani, Manjunath, et al.
Publicado: (2024)
Ejemplares similares
-
AnoPatch: Towards Better Consistency in Machine Anomalous Sound Detection
por: Jiang, Anbai, et al.
Publicado: (2024) -
Exploring Self-Supervised Audio Models for Generalized Anomalous Sound Detection
por: Han, Bing, et al.
Publicado: (2025) -
Improving Anomalous Sound Detection via Low-Rank Adaptation Fine-Tuning of Pre-Trained Audio Models
por: Zheng, Xinhu, et al.
Publicado: (2024) -
Data-Efficient Low-Complexity Acoustic Scene Classification in the DCASE 2024 Challenge
por: Schmid, Florian, et al.
Publicado: (2024) -
Improving Acoustic Scene Classification in Low-Resource Conditions
por: Chen, Zhi, et al.
Publicado: (2024)