IDMap: A Pseudo-Speaker Generator Framework Based on Speaker Identity Index to Vector Mapping
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Zeyan, Chen, Liping, Lee, Kong Aik, Ling, Zhenhua |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Pinhole Effect on Linkability and Dispersion in Speaker Anonymization
di: Lee, Kong Aik, et al.
Pubblicazione: (2025)
di: Lee, Kong Aik, et al.
Pubblicazione: (2025)
A Study of the Removability of Speaker-Adversarial Perturbations
di: Chen, Liping, et al.
Pubblicazione: (2025)
di: Chen, Liping, et al.
Pubblicazione: (2025)
On the Generation and Removal of Speaker Adversarial Perturbation for Voice-Privacy Protection
di: Guo, Chenyang, et al.
Pubblicazione: (2024)
di: Guo, Chenyang, et al.
Pubblicazione: (2024)
Speaker Privacy and Security in the Big Data Era: Protection and Defense against Deepfake
di: Chen, Liping, et al.
Pubblicazione: (2025)
di: Chen, Liping, et al.
Pubblicazione: (2025)
Overview of Speaker Modeling and Its Applications: From the Lens of Deep Speaker Representation Learning
di: Wang, Shuai, et al.
Pubblicazione: (2024)
di: Wang, Shuai, et al.
Pubblicazione: (2024)
Cosine Scoring with Uncertainty for Neural Speaker Embedding
di: Wang, Qiongqiong, et al.
Pubblicazione: (2024)
di: Wang, Qiongqiong, et al.
Pubblicazione: (2024)
Emphasized Non-Target Speaker Knowledge in Knowledge Distillation for Automatic Speaker Verification
di: Truong, Duc-Tuan, et al.
Pubblicazione: (2023)
di: Truong, Duc-Tuan, et al.
Pubblicazione: (2023)
Investigation of perception inconsistency in speaker embedding for asynchronous voice anonymization
di: Wang, Rui, et al.
Pubblicazione: (2025)
di: Wang, Rui, et al.
Pubblicazione: (2025)
Golden Gemini is All You Need: Finding the Sweet Spots for Speaker Verification
di: Liu, Tianchi, et al.
Pubblicazione: (2023)
di: Liu, Tianchi, et al.
Pubblicazione: (2023)
Bayesian Learning for Domain-Invariant Speaker Verification and Anti-Spoofing
di: Li, Jin, et al.
Pubblicazione: (2025)
di: Li, Jin, et al.
Pubblicazione: (2025)
Xi+: Uncertainty Supervision for Robust Speaker Embedding
di: Li, Junjie, et al.
Pubblicazione: (2025)
di: Li, Junjie, et al.
Pubblicazione: (2025)
Speaker-IPL: Unsupervised Learning of Speaker Characteristics with i-Vector based Pseudo-Labels
di: Aldeneh, Zakaria, et al.
Pubblicazione: (2024)
di: Aldeneh, Zakaria, et al.
Pubblicazione: (2024)
Vo-Ve: An Explainable Voice-Vector for Speaker Identity Evaluation
di: Lee, Jaejun, et al.
Pubblicazione: (2025)
di: Lee, Jaejun, et al.
Pubblicazione: (2025)
Generalizing Speaker Verification for Spoof Awareness in the Embedding Space
di: Liu, Xuechen, et al.
Pubblicazione: (2024)
di: Liu, Xuechen, et al.
Pubblicazione: (2024)
UNet-Based Fusion and Exponential Moving Average Adaptation for Noise-Robust Speaker Recognition
di: Gan, Chong-Xin, et al.
Pubblicazione: (2026)
di: Gan, Chong-Xin, et al.
Pubblicazione: (2026)
On the effectiveness of enrollment speech augmentation for Target Speaker Extraction
di: Li, Junjie, et al.
Pubblicazione: (2024)
di: Li, Junjie, et al.
Pubblicazione: (2024)
Adversarial speech for voice privacy protection from Personalized Speech generation
di: Chen, Shihao, et al.
Pubblicazione: (2024)
di: Chen, Shihao, et al.
Pubblicazione: (2024)
Revisiting and Improving Scoring Fusion for Spoofing-aware Speaker Verification Using Compositional Data Analysis
di: Wang, Xin, et al.
Pubblicazione: (2024)
di: Wang, Xin, et al.
Pubblicazione: (2024)
StreamVoiceAnon+: Emotion-Preserving Streaming Speaker Anonymization via Frame-Level Acoustic Distillation
di: Kuzmin, Nikita, et al.
Pubblicazione: (2026)
di: Kuzmin, Nikita, et al.
Pubblicazione: (2026)
Asynchronous Voice Anonymization Using Adversarial Perturbation On Speaker Embedding
di: Wang, Rui, et al.
Pubblicazione: (2024)
di: Wang, Rui, et al.
Pubblicazione: (2024)
Stream-Voice-Anon: Enhancing Utility of Real-Time Speaker Anonymization via Neural Audio Codec and Language Models
di: Kuzmin, Nikita, et al.
Pubblicazione: (2026)
di: Kuzmin, Nikita, et al.
Pubblicazione: (2026)
Text-dependent Speaker Verification (TdSV) Challenge 2024: Challenge Evaluation Plan
di: Hossein, Zeinali, et al.
Pubblicazione: (2024)
di: Hossein, Zeinali, et al.
Pubblicazione: (2024)
VoiceVector: Multimodal Enrolment Vectors for Speaker Separation
di: Rahimi, Akam, et al.
Pubblicazione: (2025)
di: Rahimi, Akam, et al.
Pubblicazione: (2025)
VoxGenesis: Unsupervised Discovery of Latent Speaker Manifold for Speech Synthesis
di: Lin, Weiwei, et al.
Pubblicazione: (2024)
di: Lin, Weiwei, et al.
Pubblicazione: (2024)
Beyond Speaker Identity: Text Guided Target Speech Extraction
di: Huo, Mingyue, et al.
Pubblicazione: (2025)
di: Huo, Mingyue, et al.
Pubblicazione: (2025)
MC-LExt: Multi-Channel Target Speaker Extraction with Onset-Prompted Speaker Conditioning Mechanism
di: Ling, Tongtao, et al.
Pubblicazione: (2025)
di: Ling, Tongtao, et al.
Pubblicazione: (2025)
Disentangling Pitch and Creak for Speaker Identity Preservation in Speech Synthesis
di: Rautenberg, Frederik, et al.
Pubblicazione: (2026)
di: Rautenberg, Frederik, et al.
Pubblicazione: (2026)
Study on Inter and Intra Speaker Variability in Speaker Recognition
di: Okhotnikov, Anton, et al.
Pubblicazione: (2024)
di: Okhotnikov, Anton, et al.
Pubblicazione: (2024)
Generating Speakers by Prompting Listener Impressions for Pre-trained Multi-Speaker Text-to-Speech Systems
di: Chen, Zhengyang, et al.
Pubblicazione: (2024)
di: Chen, Zhengyang, et al.
Pubblicazione: (2024)
Enhancing Zero-Shot Multi-Speaker TTS with Negated Speaker Representations
di: Jeon, Yejin, et al.
Pubblicazione: (2024)
di: Jeon, Yejin, et al.
Pubblicazione: (2024)
A Comprehensive Investigation on Speaker Augmentation for Speaker Recognition
di: Zhou, Zhenyu, et al.
Pubblicazione: (2024)
di: Zhou, Zhenyu, et al.
Pubblicazione: (2024)
Channel Adaptation for Speaker Verification Using Optimal Transport with Pseudo Label
di: Yang, Wenhao, et al.
Pubblicazione: (2024)
di: Yang, Wenhao, et al.
Pubblicazione: (2024)
Coherence-Based Frequency Subset Selection For Binaural RTF-Vector-Based Direction of Arrival Estimation for Multiple Speakers
di: Fejgin, Daniel, et al.
Pubblicazione: (2022)
di: Fejgin, Daniel, et al.
Pubblicazione: (2022)
NanoVoice: Efficient Speaker-Adaptive Text-to-Speech for Multiple Speakers
di: Park, Nohil, et al.
Pubblicazione: (2024)
di: Park, Nohil, et al.
Pubblicazione: (2024)
Speaker Contrastive Learning for Source Speaker Tracing
di: Wang, Qing, et al.
Pubblicazione: (2024)
di: Wang, Qing, et al.
Pubblicazione: (2024)
Can Audio Large Language Models Verify Speaker Identity?
di: Ren, Yiming, et al.
Pubblicazione: (2025)
di: Ren, Yiming, et al.
Pubblicazione: (2025)
Interpolating Speaker Identities in Embedding Space for Data Expansion
di: Liu, Tianchi, et al.
Pubblicazione: (2025)
di: Liu, Tianchi, et al.
Pubblicazione: (2025)
Target Speaker Lipreading by Audio-Visual Self-Distillation Pretraining and Speaker Adaptation
di: Zhang, Jing-Xuan, et al.
Pubblicazione: (2025)
di: Zhang, Jing-Xuan, et al.
Pubblicazione: (2025)
Joint Optimization of Speaker and Spoof Detectors for Spoofing-Robust Automatic Speaker Verification
di: Kurnaz, Oğuzhan, et al.
Pubblicazione: (2025)
di: Kurnaz, Oğuzhan, et al.
Pubblicazione: (2025)
Multi-Channel Multi-Speaker ASR Using Target Speaker's Solo Segment
di: Shao, Yiwen, et al.
Pubblicazione: (2024)
di: Shao, Yiwen, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Pinhole Effect on Linkability and Dispersion in Speaker Anonymization
di: Lee, Kong Aik, et al.
Pubblicazione: (2025) -
A Study of the Removability of Speaker-Adversarial Perturbations
di: Chen, Liping, et al.
Pubblicazione: (2025) -
On the Generation and Removal of Speaker Adversarial Perturbation for Voice-Privacy Protection
di: Guo, Chenyang, et al.
Pubblicazione: (2024) -
Speaker Privacy and Security in the Big Data Era: Protection and Defense against Deepfake
di: Chen, Liping, et al.
Pubblicazione: (2025) -
Overview of Speaker Modeling and Its Applications: From the Lens of Deep Speaker Representation Learning
di: Wang, Shuai, et al.
Pubblicazione: (2024)