BS-PLCNet 2: Two-stage Band-split Packet Loss Concealment Network with Intra-model Knowledge Distillation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Zihan, Xia, Xianjun, Huang, Chuanzeng, Xiao, Yijian, Xie, Lei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BS-PLCNet: Band-split Packet Loss Concealment Network with Multi-task Learning Framework and Multi-discriminators
von: Zhang, Zihan, et al.
Veröffentlicht: (2024)
von: Zhang, Zihan, et al.
Veröffentlicht: (2024)
RaD-Net 2: A causal two-stage repairing and denoising speech enhancement network with knowledge distillation and complex axial self-attention
von: Liu, Mingshuai, et al.
Veröffentlicht: (2024)
von: Liu, Mingshuai, et al.
Veröffentlicht: (2024)
RaD-Net: A Repairing and Denoising Network for Speech Signal Improvement
von: Liu, Mingshuai, et al.
Veröffentlicht: (2024)
von: Liu, Mingshuai, et al.
Veröffentlicht: (2024)
S$^2$Voice: Style-Aware Autoregressive Modeling with Enhanced Conditioning for Singing Style Conversion
von: Wang, Ziqian, et al.
Veröffentlicht: (2026)
von: Wang, Ziqian, et al.
Veröffentlicht: (2026)
The IEEE-IS2 2024 Music Packet Loss Concealment Challenge
von: Mezza, Alessandro Ilic, et al.
Veröffentlicht: (2024)
von: Mezza, Alessandro Ilic, et al.
Veröffentlicht: (2024)
The ICASSP 2024 Audio Deep Packet Loss Concealment Challenge
von: Diener, Lorenz, et al.
Veröffentlicht: (2024)
von: Diener, Lorenz, et al.
Veröffentlicht: (2024)
UniFlow: Unifying Speech Front-End Tasks via Continuous Generative Modeling
von: Wang, Ziqian, et al.
Veröffentlicht: (2025)
von: Wang, Ziqian, et al.
Veröffentlicht: (2025)
LDCodec: A high quality neural audio codec with low-complexity decoder
von: Jiang, Jiawei, et al.
Veröffentlicht: (2025)
von: Jiang, Jiawei, et al.
Veröffentlicht: (2025)
An Intra-BRNN and GB-RVQ Based END-TO-END Neural Audio Codec
von: Xu, Linping, et al.
Veröffentlicht: (2024)
von: Xu, Linping, et al.
Veröffentlicht: (2024)
Distil-DCCRN: A Small-footprint DCCRN Leveraging Feature-based Knowledge Distillation in Speech Enhancement
von: Han, Runduo, et al.
Veröffentlicht: (2024)
von: Han, Runduo, et al.
Veröffentlicht: (2024)
U-SAM: An audio language Model for Unified Speech, Audio, and Music Understanding
von: Wang, Ziqian, et al.
Veröffentlicht: (2025)
von: Wang, Ziqian, et al.
Veröffentlicht: (2025)
SenSE: Semantic-Aware High-Fidelity Universal Speech Enhancement
von: Li, Xingchen, et al.
Veröffentlicht: (2025)
von: Li, Xingchen, et al.
Veröffentlicht: (2025)
Delayed-KD: Delayed Knowledge Distillation based CTC for Low-Latency Streaming ASR
von: Li, Longhao, et al.
Veröffentlicht: (2025)
von: Li, Longhao, et al.
Veröffentlicht: (2025)
Integrated Multi-Level Knowledge Distillation for Enhanced Speaker Verification
von: Yang, Wenhao, et al.
Veröffentlicht: (2024)
von: Yang, Wenhao, et al.
Veröffentlicht: (2024)
A High-Quality and Low-Complexity Streamable Neural Speech Codec with Knowledge Distillation
von: Zhang, En-Wei, et al.
Veröffentlicht: (2025)
von: Zhang, En-Wei, et al.
Veröffentlicht: (2025)
A Two-Stage Band-Split Mamba-2 Network For Music Separation
von: Bai, Jinglin, et al.
Veröffentlicht: (2024)
von: Bai, Jinglin, et al.
Veröffentlicht: (2024)
Enhanced ASR Robustness to Packet Loss with a Front-End Adaptation Network
von: Dissen, Yehoshua, et al.
Veröffentlicht: (2024)
von: Dissen, Yehoshua, et al.
Veröffentlicht: (2024)
Moises-Light: Resource-efficient Band-split U-Net For Music Source Separation
von: Yun-Ning, et al.
Veröffentlicht: (2025)
von: Yun-Ning, et al.
Veröffentlicht: (2025)
Emotion Recognition in Multi-Speaker Conversations through Speaker Identification, Knowledge Distillation, and Hierarchical Fusion
von: Li, Xiao, et al.
Veröffentlicht: (2025)
von: Li, Xiao, et al.
Veröffentlicht: (2025)
EASY: Emotion-aware Speaker Anonymization via Factorized Distillation
von: Yao, Jixun, et al.
Veröffentlicht: (2025)
von: Yao, Jixun, et al.
Veröffentlicht: (2025)
Réduire le bruit grâce à la réalité augmentée sonore -- Auditory Concealer
von: Boukhemia, Clara
Veröffentlicht: (2025)
von: Boukhemia, Clara
Veröffentlicht: (2025)
Unseen but not Unknown: Using Dataset Concealment to Robustly Evaluate Speech Quality Estimation Models
von: Pieper, Jaden, et al.
Veröffentlicht: (2026)
von: Pieper, Jaden, et al.
Veröffentlicht: (2026)
Two-stage Audio-Visual Target Speaker Extraction System for Real-Time Processing On Edge Device
von: Li, Zixuan, et al.
Veröffentlicht: (2025)
von: Li, Zixuan, et al.
Veröffentlicht: (2025)
Efficient Speech Watermarking for Speech Synthesis via Progressive Knowledge Distillation
von: Cui, Yang, et al.
Veröffentlicht: (2025)
von: Cui, Yang, et al.
Veröffentlicht: (2025)
ERVQ: Enhanced Residual Vector Quantization with Intra-and-Inter-Codebook Optimization for Neural Audio Codecs
von: Zheng, Rui-Chen, et al.
Veröffentlicht: (2024)
von: Zheng, Rui-Chen, et al.
Veröffentlicht: (2024)
Synergistic Effects of Knowledge Distillation and Structured Pruning for Self-Supervised Speech Models
von: C, Shiva Kumar, et al.
Veröffentlicht: (2025)
von: C, Shiva Kumar, et al.
Veröffentlicht: (2025)
Acoustic Scene Classification Using CNN-GRU Model Without Knowledge Distillation
von: Tan, Ee-Leng, et al.
Veröffentlicht: (2025)
von: Tan, Ee-Leng, et al.
Veröffentlicht: (2025)
Emphasized Non-Target Speaker Knowledge in Knowledge Distillation for Automatic Speaker Verification
von: Truong, Duc-Tuan, et al.
Veröffentlicht: (2023)
von: Truong, Duc-Tuan, et al.
Veröffentlicht: (2023)
Leveraging ASR Pretrained Conformers for Speaker Verification through Transfer Learning and Knowledge Distillation
von: Cai, Danwei, et al.
Veröffentlicht: (2023)
von: Cai, Danwei, et al.
Veröffentlicht: (2023)
Beyond Two-stage Diffusion TTS: Joint Structure and Content Refinement via Jump Diffusion
von: Ai, Jiabao, et al.
Veröffentlicht: (2026)
von: Ai, Jiabao, et al.
Veröffentlicht: (2026)
To Distill or Not to Distill? On the Robustness of Robust Knowledge Distillation
von: Waheed, Abdul, et al.
Veröffentlicht: (2024)
von: Waheed, Abdul, et al.
Veröffentlicht: (2024)
Study on Inter and Intra Speaker Variability in Speaker Recognition
von: Okhotnikov, Anton, et al.
Veröffentlicht: (2024)
von: Okhotnikov, Anton, et al.
Veröffentlicht: (2024)
EchoFree: Towards Ultra Lightweight and Efficient Neural Acoustic Echo Cancellation
von: Li, Xingchen, et al.
Veröffentlicht: (2025)
von: Li, Xingchen, et al.
Veröffentlicht: (2025)
Efficient Audio Captioning with Encoder-Level Knowledge Distillation
von: Xu, Xuenan, et al.
Veröffentlicht: (2024)
von: Xu, Xuenan, et al.
Veröffentlicht: (2024)
Dynamic Frequency-Adaptive Knowledge Distillation for Speech Enhancement
von: Yuan, Xihao, et al.
Veröffentlicht: (2025)
von: Yuan, Xihao, et al.
Veröffentlicht: (2025)
Frequency-mix Knowledge Distillation for Fake Speech Detection
von: Fan, Cunhang, et al.
Veröffentlicht: (2024)
von: Fan, Cunhang, et al.
Veröffentlicht: (2024)
Affine Modulation-based Audiogram Fusion Network for Joint Noise Reduction and Hearing Loss Compensation
von: Ni, Ye, et al.
Veröffentlicht: (2025)
von: Ni, Ye, et al.
Veröffentlicht: (2025)
Structural and Statistical Audio Texture Knowledge Distillation for Acoustic Classification
von: Ritu, Jarin, et al.
Veröffentlicht: (2025)
von: Ritu, Jarin, et al.
Veröffentlicht: (2025)
Leave No Knowledge Behind During Knowledge Distillation: Towards Practical and Effective Knowledge Distillation for Code-Switching ASR Using Realistic Data
von: Tseng, Liang-Hsuan, et al.
Veröffentlicht: (2024)
von: Tseng, Liang-Hsuan, et al.
Veröffentlicht: (2024)
Multilingual Speech Recognition Using Discrete Tokens with a Two-step Training Strategy
von: Li, Zehan, et al.
Veröffentlicht: (2025)
von: Li, Zehan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
BS-PLCNet: Band-split Packet Loss Concealment Network with Multi-task Learning Framework and Multi-discriminators
von: Zhang, Zihan, et al.
Veröffentlicht: (2024) -
RaD-Net 2: A causal two-stage repairing and denoising speech enhancement network with knowledge distillation and complex axial self-attention
von: Liu, Mingshuai, et al.
Veröffentlicht: (2024) -
RaD-Net: A Repairing and Denoising Network for Speech Signal Improvement
von: Liu, Mingshuai, et al.
Veröffentlicht: (2024) -
S$^2$Voice: Style-Aware Autoregressive Modeling with Enhanced Conditioning for Singing Style Conversion
von: Wang, Ziqian, et al.
Veröffentlicht: (2026) -
The IEEE-IS2 2024 Music Packet Loss Concealment Challenge
von: Mezza, Alessandro Ilic, et al.
Veröffentlicht: (2024)