Gradient weighting for speaker verification in extremely low Signal-to-Noise Ratio
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ma, Yi, Lee, Kong Aik, Hautamäki, Ville, Ge, Meng, Li, Haizhou |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Golden Gemini is All You Need: Finding the Sweet Spots for Speaker Verification
von: Liu, Tianchi, et al.
Veröffentlicht: (2023)
von: Liu, Tianchi, et al.
Veröffentlicht: (2023)
Text adaptation for speaker verification with speaker-text factorized embeddings
von: Yang, Yexin, et al.
Veröffentlicht: (2025)
von: Yang, Yexin, et al.
Veröffentlicht: (2025)
Overview of Speaker Modeling and Its Applications: From the Lens of Deep Speaker Representation Learning
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
Malacopula: adversarial automatic speaker verification attacks using a neural-based generalised Hammerstein model
von: Todisco, Massimiliano, et al.
Veröffentlicht: (2024)
von: Todisco, Massimiliano, et al.
Veröffentlicht: (2024)
On the effectiveness of enrollment speech augmentation for Target Speaker Extraction
von: Li, Junjie, et al.
Veröffentlicht: (2024)
von: Li, Junjie, et al.
Veröffentlicht: (2024)
On the influence of language similarity in non-target speaker verification trials
von: Reuter, Paul M., et al.
Veröffentlicht: (2025)
von: Reuter, Paul M., et al.
Veröffentlicht: (2025)
Improving fairness in speaker verification via Group-adapted Fusion Network
von: Shen, Hua, et al.
Veröffentlicht: (2022)
von: Shen, Hua, et al.
Veröffentlicht: (2022)
A framework of text-dependent speaker verification for chinese numerical string corpus
von: Zheng, Litong, et al.
Veröffentlicht: (2024)
von: Zheng, Litong, et al.
Veröffentlicht: (2024)
Generalizable speech deepfake detection via meta-learned LoRA
von: Laakkonen, Janne, et al.
Veröffentlicht: (2025)
von: Laakkonen, Janne, et al.
Veröffentlicht: (2025)
Mixture of Low-Rank Adapter Experts in Generalizable Audio Deepfake Detection
von: Laakkonen, Janne, et al.
Veröffentlicht: (2025)
von: Laakkonen, Janne, et al.
Veröffentlicht: (2025)
Improving speaker verification robustness with synthetic emotional utterances
von: Koditala, Nikhil Kumar, et al.
Veröffentlicht: (2024)
von: Koditala, Nikhil Kumar, et al.
Veröffentlicht: (2024)
Nes2Net: A Lightweight Nested Architecture for Foundation Model Driven Speech Anti-spoofing
von: Liu, Tianchi, et al.
Veröffentlicht: (2025)
von: Liu, Tianchi, et al.
Veröffentlicht: (2025)
Cosine Scoring with Uncertainty for Neural Speaker Embedding
von: Wang, Qiongqiong, et al.
Veröffentlicht: (2024)
von: Wang, Qiongqiong, et al.
Veröffentlicht: (2024)
Meta-Learning Approaches for Improving Detection of Unseen Speech Deepfakes
von: Kukanov, Ivan, et al.
Veröffentlicht: (2024)
von: Kukanov, Ivan, et al.
Veröffentlicht: (2024)
Improved Remixing Process for Domain Adaptation-Based Speech Enhancement by Mitigating Data Imbalance in Signal-to-Noise Ratio
von: Li, Li, et al.
Veröffentlicht: (2024)
von: Li, Li, et al.
Veröffentlicht: (2024)
Hierarchical speaker representation for target speaker extraction
von: He, Shulin, et al.
Veröffentlicht: (2022)
von: He, Shulin, et al.
Veröffentlicht: (2022)
LlamaPartialSpoof: An LLM-Driven Fake Speech Dataset Simulating Disinformation Generation
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2024)
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2024)
Xi+: Uncertainty Supervision for Robust Speaker Embedding
von: Li, Junjie, et al.
Veröffentlicht: (2025)
von: Li, Junjie, et al.
Veröffentlicht: (2025)
Clustering-based hard negative sampling for supervised contrastive speaker verification
von: Masztalski, Piotr, et al.
Veröffentlicht: (2025)
von: Masztalski, Piotr, et al.
Veröffentlicht: (2025)
PhiNet: Speaker Verification with Phonetic Interpretability
von: Ma, Yi, et al.
Veröffentlicht: (2026)
von: Ma, Yi, et al.
Veröffentlicht: (2026)
Revisiting and Improving Scoring Fusion for Spoofing-aware Speaker Verification Using Compositional Data Analysis
von: Wang, Xin, et al.
Veröffentlicht: (2024)
von: Wang, Xin, et al.
Veröffentlicht: (2024)
Room Impulse Responses help attackers to evade Deep Fake Detection
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2024)
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2024)
Improving curriculum learning for target speaker extraction with synthetic speakers
von: Liu, Yun, et al.
Veröffentlicht: (2024)
von: Liu, Yun, et al.
Veröffentlicht: (2024)
Adversarial speech for voice privacy protection from Personalized Speech generation
von: Chen, Shihao, et al.
Veröffentlicht: (2024)
von: Chen, Shihao, et al.
Veröffentlicht: (2024)
Robust Localization of Partially Fake Speech: Metrics and Out-of-Domain Evaluation
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2025)
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2025)
How phonemes contribute to deep speaker models?
von: Li, Pengqi, et al.
Veröffentlicht: (2024)
von: Li, Pengqi, et al.
Veröffentlicht: (2024)
Adaptive Per-Channel Energy Normalization Front-end for Robust Audio Signal Processing
von: Meng, Hanyu, et al.
Veröffentlicht: (2025)
von: Meng, Hanyu, et al.
Veröffentlicht: (2025)
Quantifying the effect of speech pathology on automatic and human speaker verification
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2024)
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2024)
Emphasized Non-Target Speaker Knowledge in Knowledge Distillation for Automatic Speaker Verification
von: Truong, Duc-Tuan, et al.
Veröffentlicht: (2023)
von: Truong, Duc-Tuan, et al.
Veröffentlicht: (2023)
An Empirical Study on the Impact of Positional Encoding in Transformer-based Monaural Speech Enhancement
von: Zhang, Qiquan, et al.
Veröffentlicht: (2024)
von: Zhang, Qiquan, et al.
Veröffentlicht: (2024)
Disentangling Speaker Traits for Deepfake Source Verification via Chebyshev Polynomial and Riemannian Metric Learning
von: Xuan, Xi, et al.
Veröffentlicht: (2026)
von: Xuan, Xi, et al.
Veröffentlicht: (2026)
USED: Universal Speaker Extraction and Diarization
von: Ao, Junyi, et al.
Veröffentlicht: (2023)
von: Ao, Junyi, et al.
Veröffentlicht: (2023)
Text-dependent Speaker Verification (TdSV) Challenge 2024: Challenge Evaluation Plan
von: Hossein, Zeinali, et al.
Veröffentlicht: (2024)
von: Hossein, Zeinali, et al.
Veröffentlicht: (2024)
Target Speech Extraction with Pre-trained AV-HuBERT and Mask-And-Recover Strategy
von: Wu, Wenxuan, et al.
Veröffentlicht: (2024)
von: Wu, Wenxuan, et al.
Veröffentlicht: (2024)
WeSep: A Scalable and Flexible Toolkit Towards Generalizable Target Speaker Extraction
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
The importance of spatial and spectral information in multiple speaker tracking
von: Beit-On, Hanan, et al.
Veröffentlicht: (2024)
von: Beit-On, Hanan, et al.
Veröffentlicht: (2024)
Noise-to-mask Ratio Loss for Deep Neural Network based Audio Watermarking
von: Moritz, Martin, et al.
Veröffentlicht: (2024)
von: Moritz, Martin, et al.
Veröffentlicht: (2024)
Multi-Scale Accent Modeling and Disentangling for Multi-Speaker Multi-Accent Text-to-Speech Synthesis
von: Zhou, Xuehao, et al.
Veröffentlicht: (2024)
von: Zhou, Xuehao, et al.
Veröffentlicht: (2024)
Chain-Talker: Chain Understanding and Rendering for Empathetic Conversational Speech Synthesis
von: Hu, Yifan, et al.
Veröffentlicht: (2025)
von: Hu, Yifan, et al.
Veröffentlicht: (2025)
Audio-visual child-adult speaker classification in dyadic interactions
von: Xu, Anfeng, et al.
Veröffentlicht: (2023)
von: Xu, Anfeng, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Golden Gemini is All You Need: Finding the Sweet Spots for Speaker Verification
von: Liu, Tianchi, et al.
Veröffentlicht: (2023) -
Text adaptation for speaker verification with speaker-text factorized embeddings
von: Yang, Yexin, et al.
Veröffentlicht: (2025) -
Overview of Speaker Modeling and Its Applications: From the Lens of Deep Speaker Representation Learning
von: Wang, Shuai, et al.
Veröffentlicht: (2024) -
Malacopula: adversarial automatic speaker verification attacks using a neural-based generalised Hammerstein model
von: Todisco, Massimiliano, et al.
Veröffentlicht: (2024) -
On the effectiveness of enrollment speech augmentation for Target Speaker Extraction
von: Li, Junjie, et al.
Veröffentlicht: (2024)