Triage knowledge distillation for speaker verification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Ju-ho, Jung, Youngmoon, Yang, Joon-Young, Roh, Jaeyoung, Han, Chang Woo, Cho, Hoon-Young |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DAME: Duration-Aware Matryoshka Embedding for Duration-Robust Speaker Verification
von: Jung, Youngmoon, et al.
Veröffentlicht: (2026)
von: Jung, Youngmoon, et al.
Veröffentlicht: (2026)
Relational Proxy Loss for Audio-Text based Keyword Spotting
von: Jung, Youngmoon, et al.
Veröffentlicht: (2024)
von: Jung, Youngmoon, et al.
Veröffentlicht: (2024)
MATE: Matryoshka Audio-Text Embeddings for Open-Vocabulary Keyword Spotting
von: Jung, Youngmoon, et al.
Veröffentlicht: (2026)
von: Jung, Youngmoon, et al.
Veröffentlicht: (2026)
Adversarial Deep Metric Learning for Cross-Modal Audio-Text Alignment in Open-Vocabulary Keyword Spotting
von: Jung, Youngmoon, et al.
Veröffentlicht: (2025)
von: Jung, Youngmoon, et al.
Veröffentlicht: (2025)
Curriculum learning for self-supervised speaker verification
von: Heo, Hee-Soo, et al.
Veröffentlicht: (2022)
von: Heo, Hee-Soo, et al.
Veröffentlicht: (2022)
CTC-aligned Audio-Text Embedding for Streaming Open-vocabulary Keyword Spotting
von: Jin, Sichen, et al.
Veröffentlicht: (2024)
von: Jin, Sichen, et al.
Veröffentlicht: (2024)
Text-Aware Adapter for Few-Shot Keyword Spotting
von: Jung, Youngmoon, et al.
Veröffentlicht: (2024)
von: Jung, Youngmoon, et al.
Veröffentlicht: (2024)
Text adaptation for speaker verification with speaker-text factorized embeddings
von: Yang, Yexin, et al.
Veröffentlicht: (2025)
von: Yang, Yexin, et al.
Veröffentlicht: (2025)
Improving fairness in speaker verification via Group-adapted Fusion Network
von: Shen, Hua, et al.
Veröffentlicht: (2022)
von: Shen, Hua, et al.
Veröffentlicht: (2022)
Improving speaker verification robustness with synthetic emotional utterances
von: Koditala, Nikhil Kumar, et al.
Veröffentlicht: (2024)
von: Koditala, Nikhil Kumar, et al.
Veröffentlicht: (2024)
Tandem spoofing-robust automatic speaker verification based on time-domain embeddings
von: Weizman, Avishai, et al.
Veröffentlicht: (2024)
von: Weizman, Avishai, et al.
Veröffentlicht: (2024)
On the influence of language similarity in non-target speaker verification trials
von: Reuter, Paul M., et al.
Veröffentlicht: (2025)
von: Reuter, Paul M., et al.
Veröffentlicht: (2025)
Hierarchical speaker representation for target speaker extraction
von: He, Shulin, et al.
Veröffentlicht: (2022)
von: He, Shulin, et al.
Veröffentlicht: (2022)
a-DCF: an architecture agnostic metric with application to spoofing-robust speaker verification
von: Shim, Hye-jin, et al.
Veröffentlicht: (2024)
von: Shim, Hye-jin, et al.
Veröffentlicht: (2024)
Gradient weighting for speaker verification in extremely low Signal-to-Noise Ratio
von: Ma, Yi, et al.
Veröffentlicht: (2024)
von: Ma, Yi, et al.
Veröffentlicht: (2024)
A framework of text-dependent speaker verification for chinese numerical string corpus
von: Zheng, Litong, et al.
Veröffentlicht: (2024)
von: Zheng, Litong, et al.
Veröffentlicht: (2024)
Optimization of DNN-based speaker verification model through efficient quantization technique
von: Hong, Yeona, et al.
Veröffentlicht: (2024)
von: Hong, Yeona, et al.
Veröffentlicht: (2024)
InfiniteAudio: Infinite-Length Audio Generation with Consistency
von: Jung, Chaeyoung, et al.
Veröffentlicht: (2025)
von: Jung, Chaeyoung, et al.
Veröffentlicht: (2025)
Period Singer: Integrating Periodic and Aperiodic Variational Autoencoders for Natural-Sounding End-to-End Singing Voice Synthesis
von: Kim, Taewoo, et al.
Veröffentlicht: (2024)
von: Kim, Taewoo, et al.
Veröffentlicht: (2024)
MR-RawNet: Speaker verification system with multiple temporal resolutions for variable duration utterances using raw waveforms
von: Kim, Seung-bin, et al.
Veröffentlicht: (2024)
von: Kim, Seung-bin, et al.
Veröffentlicht: (2024)
High Fidelity Text-to-Speech Via Discrete Tokens Using Token Transducer and Group Masked Language Model
von: Lee, Joun Yeop, et al.
Veröffentlicht: (2024)
von: Lee, Joun Yeop, et al.
Veröffentlicht: (2024)
Naturalness-Aware Curriculum Learning with Dynamic Temperature for Speech Deepfake Detection
von: Kim, Taewoo, et al.
Veröffentlicht: (2025)
von: Kim, Taewoo, et al.
Veröffentlicht: (2025)
FlowAVSE: Efficient Audio-Visual Speech Enhancement with Conditional Flow Matching
von: Jung, Chaeyoung, et al.
Veröffentlicht: (2024)
von: Jung, Chaeyoung, et al.
Veröffentlicht: (2024)
Probing Cross-modal Information Hubs in Audio-Visual LLMs
von: Jung, Jihoo, et al.
Veröffentlicht: (2026)
von: Jung, Jihoo, et al.
Veröffentlicht: (2026)
Clustering-based hard negative sampling for supervised contrastive speaker verification
von: Masztalski, Piotr, et al.
Veröffentlicht: (2025)
von: Masztalski, Piotr, et al.
Veröffentlicht: (2025)
Latent Filling: Latent Space Data Augmentation for Zero-shot Speech Synthesis
von: Bae, Jae-Sung, et al.
Veröffentlicht: (2023)
von: Bae, Jae-Sung, et al.
Veröffentlicht: (2023)
Quantifying the effect of speech pathology on automatic and human speaker verification
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2024)
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2024)
MamTra: A Hybrid Mamba-Transformer Backbone for Speech Synthesis
von: Nguyen, Tan Dat, et al.
Veröffentlicht: (2026)
von: Nguyen, Tan Dat, et al.
Veröffentlicht: (2026)
Privacy-oriented manipulation of speaker representations
von: Teixeira, Francisco, et al.
Veröffentlicht: (2023)
von: Teixeira, Francisco, et al.
Veröffentlicht: (2023)
Improving curriculum learning for target speaker extraction with synthetic speakers
von: Liu, Yun, et al.
Veröffentlicht: (2024)
von: Liu, Yun, et al.
Veröffentlicht: (2024)
Single-Channel Distance-Based Source Separation for Mobile GPU in Outdoor and Indoor Environments
von: Bae, Hanbin, et al.
Veröffentlicht: (2025)
von: Bae, Hanbin, et al.
Veröffentlicht: (2025)
MathReader : Text-to-Speech for Mathematical Documents
von: Hyeon, Sieun, et al.
Veröffentlicht: (2025)
von: Hyeon, Sieun, et al.
Veröffentlicht: (2025)
UNMIXX: Untangling Highly Correlated Singing Voices Mixtures
von: Jung, Jihoo, et al.
Veröffentlicht: (2026)
von: Jung, Jihoo, et al.
Veröffentlicht: (2026)
SCORE: Scaling audio generation using Standardized COmposite REwards
von: Jung, Jaemin, et al.
Veröffentlicht: (2025)
von: Jung, Jaemin, et al.
Veröffentlicht: (2025)
X-CrossNet: A complex spectral mapping approach to target speaker extraction with cross attention speaker embedding fusion
von: Sun, Chang, et al.
Veröffentlicht: (2024)
von: Sun, Chang, et al.
Veröffentlicht: (2024)
DISPATCH: Distilling Selective Patches for Speech Enhancement
von: Kim, Dohwan, et al.
Veröffentlicht: (2025)
von: Kim, Dohwan, et al.
Veröffentlicht: (2025)
Instance-Specific Test-Time Training for Speech Editing in the Wild
von: Kim, Taewoo, et al.
Veröffentlicht: (2025)
von: Kim, Taewoo, et al.
Veröffentlicht: (2025)
Bridging the Gap between Audio and Text using Parallel-attention for User-defined Keyword Spotting
von: Kim, Youkyum, et al.
Veröffentlicht: (2024)
von: Kim, Youkyum, et al.
Veröffentlicht: (2024)
Trainable Adaptive Score Normalization for Automatic Speaker Verification
von: Choi, Jeong-Hwan, et al.
Veröffentlicht: (2025)
von: Choi, Jeong-Hwan, et al.
Veröffentlicht: (2025)
RepAugment: Input-Agnostic Representation-Level Augmentation for Respiratory Sound Classification
von: Kim, June-Woo, et al.
Veröffentlicht: (2024)
von: Kim, June-Woo, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
DAME: Duration-Aware Matryoshka Embedding for Duration-Robust Speaker Verification
von: Jung, Youngmoon, et al.
Veröffentlicht: (2026) -
Relational Proxy Loss for Audio-Text based Keyword Spotting
von: Jung, Youngmoon, et al.
Veröffentlicht: (2024) -
MATE: Matryoshka Audio-Text Embeddings for Open-Vocabulary Keyword Spotting
von: Jung, Youngmoon, et al.
Veröffentlicht: (2026) -
Adversarial Deep Metric Learning for Cross-Modal Audio-Text Alignment in Open-Vocabulary Keyword Spotting
von: Jung, Youngmoon, et al.
Veröffentlicht: (2025) -
Curriculum learning for self-supervised speaker verification
von: Heo, Hee-Soo, et al.
Veröffentlicht: (2022)