Disentangling Age and Identity with a Mutual Information Minimization Approach for Cross-Age Speaker Verification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Fengrun, Zhou, Wangjin, Liu, Yiming, Geng, Wang, Shan, Yahui, Zhang, Chen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Zero-Shot Sing Voice Conversion: built upon clustering-based phoneme representations
von: Zhou, Wangjin, et al.
Veröffentlicht: (2024)
von: Zhou, Wangjin, et al.
Veröffentlicht: (2024)
Boosting Code-Switching ASR with Mixture of Experts Enhanced Speech-Conditioned LLM
von: Zhang, Fengrun, et al.
Veröffentlicht: (2024)
von: Zhang, Fengrun, et al.
Veröffentlicht: (2024)
Enhancing Age-Related Robustness in Children Speaker Verification
von: Shetty, Vishwas M., et al.
Veröffentlicht: (2025)
von: Shetty, Vishwas M., et al.
Veröffentlicht: (2025)
ASD-Diffusion: Anomalous Sound Detection with Diffusion Models
von: Zhang, Fengrun, et al.
Veröffentlicht: (2024)
von: Zhang, Fengrun, et al.
Veröffentlicht: (2024)
Can Audio Large Language Models Verify Speaker Identity?
von: Ren, Yiming, et al.
Veröffentlicht: (2025)
von: Ren, Yiming, et al.
Veröffentlicht: (2025)
Memory-Efficient Training for Deep Speaker Embedding Learning in Speaker Verification
von: Liu, Bei, et al.
Veröffentlicht: (2024)
von: Liu, Bei, et al.
Veröffentlicht: (2024)
Disentangling Speakers in Multi-Talker Speech Recognition with Speaker-Aware CTC
von: Kang, Jiawen, et al.
Veröffentlicht: (2024)
von: Kang, Jiawen, et al.
Veröffentlicht: (2024)
A Joint Noise Disentanglement and Adversarial Training Framework for Robust Speaker Verification
von: Xing, Xujiang, et al.
Veröffentlicht: (2024)
von: Xing, Xujiang, et al.
Veröffentlicht: (2024)
Emotional Text-To-Speech Based on Mutual-Information-Guided Emotion-Timbre Disentanglement
von: Yang, Jianing, et al.
Veröffentlicht: (2025)
von: Yang, Jianing, et al.
Veröffentlicht: (2025)
Explainable Attribute-Based Speaker Verification
von: Wu, Xiaoliang, et al.
Veröffentlicht: (2024)
von: Wu, Xiaoliang, et al.
Veröffentlicht: (2024)
Adversarial Reweighting for Speaker Verification Fairness
von: Jin, Minho, et al.
Veröffentlicht: (2022)
von: Jin, Minho, et al.
Veröffentlicht: (2022)
An Investigation of Reprogramming for Cross-Language Adaptation in Speaker Verification Systems
von: Li, Jingyu, et al.
Veröffentlicht: (2024)
von: Li, Jingyu, et al.
Veröffentlicht: (2024)
Diffusion-Based Adversarial Purification for Speaker Verification
von: Bai, Yibo, et al.
Veröffentlicht: (2023)
von: Bai, Yibo, et al.
Veröffentlicht: (2023)
Multi-Scale Accent Modeling and Disentangling for Multi-Speaker Multi-Accent Text-to-Speech Synthesis
von: Zhou, Xuehao, et al.
Veröffentlicht: (2024)
von: Zhou, Xuehao, et al.
Veröffentlicht: (2024)
ExPO: Explainable Phonetic Trait-Oriented Network for Speaker Verification
von: Ma, Yi, et al.
Veröffentlicht: (2025)
von: Ma, Yi, et al.
Veröffentlicht: (2025)
Adaptive Data Augmentation with NaturalSpeech3 for Far-field Speaker Verification
von: Zhang, Li, et al.
Veröffentlicht: (2025)
von: Zhang, Li, et al.
Veröffentlicht: (2025)
DART: Disentanglement of Accent and Speaker Representation in Multispeaker Text-to-Speech
von: Melechovsky, Jan, et al.
Veröffentlicht: (2024)
von: Melechovsky, Jan, et al.
Veröffentlicht: (2024)
Learning Emotion-Invariant Speaker Representations for Speaker Verification
von: Tian, Jingguang, et al.
Veröffentlicht: (2025)
von: Tian, Jingguang, et al.
Veröffentlicht: (2025)
Phone Duration Modeling for Speaker Age Estimation in Children
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2021)
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2021)
A Probabilistic Fusion Framework for Spoofing Aware Speaker Verification
von: Zhang, You, et al.
Veröffentlicht: (2022)
von: Zhang, You, et al.
Veröffentlicht: (2022)
PhiNet: Speaker Verification with Phonetic Interpretability
von: Ma, Yi, et al.
Veröffentlicht: (2026)
von: Ma, Yi, et al.
Veröffentlicht: (2026)
Emotional Styles Hide in Deep Speaker Embeddings: Disentangle Deep Speaker Embeddings for Speaker Clustering
von: Lin, Chaohao, et al.
Veröffentlicht: (2025)
von: Lin, Chaohao, et al.
Veröffentlicht: (2025)
SONAR: Self-Distilled Continual Pre-training for Domain Adaptive Audio Representation
von: Zhang, Yizhou, et al.
Veröffentlicht: (2025)
von: Zhang, Yizhou, et al.
Veröffentlicht: (2025)
DiEmo-TTS: Disentangled Emotion Representations via Self-Supervised Distillation for Cross-Speaker Emotion Transfer in Text-to-Speech
von: Cho, Deok-Hyeon, et al.
Veröffentlicht: (2025)
von: Cho, Deok-Hyeon, et al.
Veröffentlicht: (2025)
Evaluating Speaker Identity Coding in Self-supervised Models and Humans
von: Elbanna, Gasser
Veröffentlicht: (2024)
von: Elbanna, Gasser
Veröffentlicht: (2024)
Adapting General Disentanglement-Based Speaker Anonymization for Enhanced Emotion Preservation
von: Miao, Xiaoxiao, et al.
Veröffentlicht: (2024)
von: Miao, Xiaoxiao, et al.
Veröffentlicht: (2024)
MASV: Speaker Verification with Global and Local Context Mamba
von: Liu, Yang, et al.
Veröffentlicht: (2024)
von: Liu, Yang, et al.
Veröffentlicht: (2024)
Unispeaker: A Unified Approach for Multimodality-driven Speaker Generation
von: Sheng, Zhengyan, et al.
Veröffentlicht: (2025)
von: Sheng, Zhengyan, et al.
Veröffentlicht: (2025)
Emphasized Non-Target Speaker Knowledge in Knowledge Distillation for Automatic Speaker Verification
von: Truong, Duc-Tuan, et al.
Veröffentlicht: (2023)
von: Truong, Duc-Tuan, et al.
Veröffentlicht: (2023)
Whisper-SV: Adapting Whisper for Low-data-resource Speaker Verification
von: Zhang, Li, et al.
Veröffentlicht: (2024)
von: Zhang, Li, et al.
Veröffentlicht: (2024)
Disentangled Representation Learning for Environment-agnostic Speaker Recognition
von: Nam, KiHyun, et al.
Veröffentlicht: (2024)
von: Nam, KiHyun, et al.
Veröffentlicht: (2024)
Phonetic Richness for Improved Automatic Speaker Verification
von: Klein, Nicholas, et al.
Veröffentlicht: (2024)
von: Klein, Nicholas, et al.
Veröffentlicht: (2024)
Attacking Voice Anonymization Systems with Augmented Feature and Speaker Identity Difference
von: Zhang, Yanzhe, et al.
Veröffentlicht: (2024)
von: Zhang, Yanzhe, et al.
Veröffentlicht: (2024)
Neural Codec-based Adversarial Sample Detection for Speaker Verification
von: Chen, Xuanjun, et al.
Veröffentlicht: (2024)
von: Chen, Xuanjun, et al.
Veröffentlicht: (2024)
MDD: a Mask Diffusion Detector to Protect Speaker Verification Systems from Adversarial Perturbations
von: Bai, Yibo, et al.
Veröffentlicht: (2025)
von: Bai, Yibo, et al.
Veröffentlicht: (2025)
Towards Lightweight Speaker Verification via Adaptive Neural Network Quantization
von: Liu, Bei, et al.
Veröffentlicht: (2024)
von: Liu, Bei, et al.
Veröffentlicht: (2024)
Disentangling Speaker Traits for Deepfake Source Verification via Chebyshev Polynomial and Riemannian Metric Learning
von: Xuan, Xi, et al.
Veröffentlicht: (2026)
von: Xuan, Xi, et al.
Veröffentlicht: (2026)
The SVASR System for Text-dependent Speaker Verification (TdSV) AAIC Challenge 2024
von: Molavi, Mohammadreza, et al.
Veröffentlicht: (2024)
von: Molavi, Mohammadreza, et al.
Veröffentlicht: (2024)
Text-dependent Speaker Verification (TdSV) Challenge 2024: Challenge Evaluation Plan
von: Hossein, Zeinali, et al.
Veröffentlicht: (2024)
von: Hossein, Zeinali, et al.
Veröffentlicht: (2024)
Whisper-PMFA: Partial Multi-Scale Feature Aggregation for Speaker Verification using Whisper Models
von: Zhao, Yiyang, et al.
Veröffentlicht: (2024)
von: Zhao, Yiyang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Zero-Shot Sing Voice Conversion: built upon clustering-based phoneme representations
von: Zhou, Wangjin, et al.
Veröffentlicht: (2024) -
Boosting Code-Switching ASR with Mixture of Experts Enhanced Speech-Conditioned LLM
von: Zhang, Fengrun, et al.
Veröffentlicht: (2024) -
Enhancing Age-Related Robustness in Children Speaker Verification
von: Shetty, Vishwas M., et al.
Veröffentlicht: (2025) -
ASD-Diffusion: Anomalous Sound Detection with Diffusion Models
von: Zhang, Fengrun, et al.
Veröffentlicht: (2024) -
Can Audio Large Language Models Verify Speaker Identity?
von: Ren, Yiming, et al.
Veröffentlicht: (2025)