Improving fairness in speaker verification via Group-adapted Fusion Network
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Shen, Hua, Yang, Yuguang, Sun, Guoli, Langman, Ryan, Han, Eunjung, Droppo, Jasha, Stolcke, Andreas |
|---|---|
| Format: | Preprint |
| Publié: |
2022
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Reducing Geographic Disparities in Automatic Speech Recognition via Elastic Weight Consolidation
par: Trinh, Viet Anh, et autres
Publié: (2022)
par: Trinh, Viet Anh, et autres
Publié: (2022)
Adversarial Reweighting for Speaker Verification Fairness
par: Jin, Minho, et autres
Publié: (2022)
par: Jin, Minho, et autres
Publié: (2022)
Improving speaker verification robustness with synthetic emotional utterances
par: Koditala, Nikhil Kumar, et autres
Publié: (2024)
par: Koditala, Nikhil Kumar, et autres
Publié: (2024)
Text adaptation for speaker verification with speaker-text factorized embeddings
par: Yang, Yexin, et autres
Publié: (2025)
par: Yang, Yexin, et autres
Publié: (2025)
On the influence of language similarity in non-target speaker verification trials
par: Reuter, Paul M., et autres
Publié: (2025)
par: Reuter, Paul M., et autres
Publié: (2025)
Gradient weighting for speaker verification in extremely low Signal-to-Noise Ratio
par: Ma, Yi, et autres
Publié: (2024)
par: Ma, Yi, et autres
Publié: (2024)
A framework of text-dependent speaker verification for chinese numerical string corpus
par: Zheng, Litong, et autres
Publié: (2024)
par: Zheng, Litong, et autres
Publié: (2024)
Improving curriculum learning for target speaker extraction with synthetic speakers
par: Liu, Yun, et autres
Publié: (2024)
par: Liu, Yun, et autres
Publié: (2024)
Clustering-based hard negative sampling for supervised contrastive speaker verification
par: Masztalski, Piotr, et autres
Publié: (2025)
par: Masztalski, Piotr, et autres
Publié: (2025)
Hierarchical speaker representation for target speaker extraction
par: He, Shulin, et autres
Publié: (2022)
par: He, Shulin, et autres
Publié: (2022)
Quantifying the effect of speech pathology on automatic and human speaker verification
par: Halpern, Bence Mark, et autres
Publié: (2024)
par: Halpern, Bence Mark, et autres
Publié: (2024)
PROCTER: PROnunciation-aware ConTextual adaptER for personalized speech recognition in neural transducers
par: Pandey, Rahul, et autres
Publié: (2023)
par: Pandey, Rahul, et autres
Publié: (2023)
Learning When to Trust Which Teacher for Weakly Supervised ASR
par: Agrawal, Aakriti, et autres
Publié: (2023)
par: Agrawal, Aakriti, et autres
Publié: (2023)
An audio-quality-based multi-strategy approach for target speaker extraction in the MISP 2023 Challenge
par: Han, Runduo, et autres
Publié: (2024)
par: Han, Runduo, et autres
Publié: (2024)
X-CrossNet: A complex spectral mapping approach to target speaker extraction with cross attention speaker embedding fusion
par: Sun, Chang, et autres
Publié: (2024)
par: Sun, Chang, et autres
Publié: (2024)
Voice Biomarker Analysis and Automated Severity Classification of Dysarthric Speech in a Multilingual Context
par: Yeo, Eunjung
Publié: (2024)
par: Yeo, Eunjung
Publié: (2024)
How phonemes contribute to deep speaker models?
par: Li, Pengqi, et autres
Publié: (2024)
par: Li, Pengqi, et autres
Publié: (2024)
The importance of spatial and spectral information in multiple speaker tracking
par: Beit-On, Hanan, et autres
Publié: (2024)
par: Beit-On, Hanan, et autres
Publié: (2024)
Audio-visual child-adult speaker classification in dyadic interactions
par: Xu, Anfeng, et autres
Publié: (2023)
par: Xu, Anfeng, et autres
Publié: (2023)
Spoken language change detection inspired by speaker change detection
par: Mishra, Jagabandhu, et autres
Publié: (2023)
par: Mishra, Jagabandhu, et autres
Publié: (2023)
Efficient Long-Form Speech Recognition for General Speech In-Context Learning
par: Yen, Hao, et autres
Publié: (2024)
par: Yen, Hao, et autres
Publié: (2024)
Spectral or spatial? Leveraging both for speaker extraction in challenging data conditions
par: Eisenberg, Aviad, et autres
Publié: (2025)
par: Eisenberg, Aviad, et autres
Publié: (2025)
Why disentanglement-based speaker anonymization systems fail at preserving emotions?
par: Gaznepoglu, Ünal Ege, et autres
Publié: (2025)
par: Gaznepoglu, Ünal Ege, et autres
Publié: (2025)
Speaker-agnostic Emotion Vector for Cross-speaker Emotion Intensity Control
par: Murata, Masato, et autres
Publié: (2025)
par: Murata, Masato, et autres
Publié: (2025)
HeightCeleb - an enrichment of VoxCeleb dataset with speaker height information
par: Kacprzak, Stanisław, et autres
Publié: (2024)
par: Kacprzak, Stanisław, et autres
Publié: (2024)
EEND-M2F: Masked-attention mask transformers for speaker diarization
par: Härkönen, Marc, et autres
Publié: (2024)
par: Härkönen, Marc, et autres
Publié: (2024)
Exploring synthetic data for cross-speaker style transfer in style representation based TTS
par: Ueda, Lucas H., et autres
Publié: (2024)
par: Ueda, Lucas H., et autres
Publié: (2024)
Fine-grained Preference Optimization Improves Zero-shot Text-to-Speech
par: Yao, Jixun, et autres
Publié: (2025)
par: Yao, Jixun, et autres
Publié: (2025)
PrimeK-Net: Multi-scale Spectral Learning via Group Prime-Kernel Convolutional Neural Networks for Single Channel Speech Enhancement
par: Lin, Zizhen, et autres
Publié: (2025)
par: Lin, Zizhen, et autres
Publié: (2025)
Thinking in cocktail party: Chain-of-Thought and reinforcement learning for target speaker automatic speech recognition
par: Zhang, Yiru, et autres
Publié: (2025)
par: Zhang, Yiru, et autres
Publié: (2025)
An Exploration of ECAPA-TDNN and x-vector Speaker Representations in Zero-shot Multi-speaker TTS
par: Kunešová, Marie, et autres
Publié: (2025)
par: Kunešová, Marie, et autres
Publié: (2025)
Spatio-spectral diarization of meetings by combining TDOA-based segmentation and speaker embedding-based clustering
par: Cord-Landwehr, Tobias, et autres
Publié: (2025)
par: Cord-Landwehr, Tobias, et autres
Publié: (2025)
Malacopula: adversarial automatic speaker verification attacks using a neural-based generalised Hammerstein model
par: Todisco, Massimiliano, et autres
Publié: (2024)
par: Todisco, Massimiliano, et autres
Publié: (2024)
SPGISpeech 2.0: Transcribed multi-speaker financial audio for speaker-tagged transcription
par: Grossman, Raymond, et autres
Publié: (2025)
par: Grossman, Raymond, et autres
Publié: (2025)
DualStream Contextual Fusion Network: Efficient Target Speaker Extraction by Leveraging Mixture and Enrollment Interactions
par: Xue, Ke, et autres
Publié: (2025)
par: Xue, Ke, et autres
Publié: (2025)
F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization
par: Sun, Xiaohui, et autres
Publié: (2025)
par: Sun, Xiaohui, et autres
Publié: (2025)
SelfTTS: cross-speaker style transfer through explicit embedding disentanglement and self-refinement using self-augmentation
par: Ueda, Lucas H., et autres
Publié: (2026)
par: Ueda, Lucas H., et autres
Publié: (2026)
Post-Training Embedding Alignment for Decoupling Enrollment and Runtime Speaker Recognition Models
par: Gao, Chenyang, et autres
Publié: (2024)
par: Gao, Chenyang, et autres
Publié: (2024)
Visual-based spatial audio generation system for multi-speaker environments
par: Liu, Xiaojing, et autres
Publié: (2025)
par: Liu, Xiaojing, et autres
Publié: (2025)
On the calibration of powerset speaker diarization models
par: Plaquet, Alexis, et autres
Publié: (2024)
par: Plaquet, Alexis, et autres
Publié: (2024)
Documents similaires
-
Reducing Geographic Disparities in Automatic Speech Recognition via Elastic Weight Consolidation
par: Trinh, Viet Anh, et autres
Publié: (2022) -
Adversarial Reweighting for Speaker Verification Fairness
par: Jin, Minho, et autres
Publié: (2022) -
Improving speaker verification robustness with synthetic emotional utterances
par: Koditala, Nikhil Kumar, et autres
Publié: (2024) -
Text adaptation for speaker verification with speaker-text factorized embeddings
par: Yang, Yexin, et autres
Publié: (2025) -
On the influence of language similarity in non-target speaker verification trials
par: Reuter, Paul M., et autres
Publié: (2025)