Investigation of perception inconsistency in speaker embedding for asynchronous voice anonymization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Rui, Chen, Liping, Lee, Kong Aik, Zha, Zhengpeng, Ling, Zhenhua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adversarial speech for voice privacy protection from Personalized Speech generation
von: Chen, Shihao, et al.
Veröffentlicht: (2024)
von: Chen, Shihao, et al.
Veröffentlicht: (2024)
Pinhole Effect on Linkability and Dispersion in Speaker Anonymization
von: Lee, Kong Aik, et al.
Veröffentlicht: (2025)
von: Lee, Kong Aik, et al.
Veröffentlicht: (2025)
IDMap: A Pseudo-Speaker Generator Framework Based on Speaker Identity Index to Vector Mapping
von: Liu, Zeyan, et al.
Veröffentlicht: (2025)
von: Liu, Zeyan, et al.
Veröffentlicht: (2025)
Introducing voice timbre attribute detection
von: He, Jinghao, et al.
Veröffentlicht: (2025)
von: He, Jinghao, et al.
Veröffentlicht: (2025)
A Study of the Removability of Speaker-Adversarial Perturbations
von: Chen, Liping, et al.
Veröffentlicht: (2025)
von: Chen, Liping, et al.
Veröffentlicht: (2025)
Gradient weighting for speaker verification in extremely low Signal-to-Noise Ratio
von: Ma, Yi, et al.
Veröffentlicht: (2024)
von: Ma, Yi, et al.
Veröffentlicht: (2024)
Target speaker anonymization in multi-speaker recordings
von: Tomashenko, Natalia, et al.
Veröffentlicht: (2025)
von: Tomashenko, Natalia, et al.
Veröffentlicht: (2025)
Why disentanglement-based speaker anonymization systems fail at preserving emotions?
von: Gaznepoglu, Ünal Ege, et al.
Veröffentlicht: (2025)
von: Gaznepoglu, Ünal Ege, et al.
Veröffentlicht: (2025)
Text adaptation for speaker verification with speaker-text factorized embeddings
von: Yang, Yexin, et al.
Veröffentlicht: (2025)
von: Yang, Yexin, et al.
Veröffentlicht: (2025)
Speaker Privacy and Security in the Big Data Era: Protection and Defense against Deepfake
von: Chen, Liping, et al.
Veröffentlicht: (2025)
von: Chen, Liping, et al.
Veröffentlicht: (2025)
The Voice Timbre Attribute Detection 2025 Challenge Evaluation Plan
von: Sheng, Zhengyan, et al.
Veröffentlicht: (2025)
von: Sheng, Zhengyan, et al.
Veröffentlicht: (2025)
Geodesic interpolation of frame-wise speaker embeddings for the diarization of meeting scenarios
von: Cord-Landwehr, Tobias, et al.
Veröffentlicht: (2024)
von: Cord-Landwehr, Tobias, et al.
Veröffentlicht: (2024)
Cosine Scoring with Uncertainty for Neural Speaker Embedding
von: Wang, Qiongqiong, et al.
Veröffentlicht: (2024)
von: Wang, Qiongqiong, et al.
Veröffentlicht: (2024)
Challenging margin-based speaker embedding extractors by using the variational information bottleneck
von: Stafylakis, Themos, et al.
Veröffentlicht: (2024)
von: Stafylakis, Themos, et al.
Veröffentlicht: (2024)
Tandem spoofing-robust automatic speaker verification based on time-domain embeddings
von: Weizman, Avishai, et al.
Veröffentlicht: (2024)
von: Weizman, Avishai, et al.
Veröffentlicht: (2024)
On the Generation and Removal of Speaker Adversarial Perturbation for Voice-Privacy Protection
von: Guo, Chenyang, et al.
Veröffentlicht: (2024)
von: Guo, Chenyang, et al.
Veröffentlicht: (2024)
Golden Gemini is All You Need: Finding the Sweet Spots for Speaker Verification
von: Liu, Tianchi, et al.
Veröffentlicht: (2023)
von: Liu, Tianchi, et al.
Veröffentlicht: (2023)
Overview of Speaker Modeling and Its Applications: From the Lens of Deep Speaker Representation Learning
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
Malacopula: adversarial automatic speaker verification attacks using a neural-based generalised Hammerstein model
von: Todisco, Massimiliano, et al.
Veröffentlicht: (2024)
von: Todisco, Massimiliano, et al.
Veröffentlicht: (2024)
Bayesian Learning for Domain-Invariant Speaker Verification and Anti-Spoofing
von: Li, Jin, et al.
Veröffentlicht: (2025)
von: Li, Jin, et al.
Veröffentlicht: (2025)
ProSDD: Learning Prosodic Representations for Speech Deepfake Detection against Expressive and Emotional Attacks
von: Mahapatra, Aurosweta, et al.
Veröffentlicht: (2026)
von: Mahapatra, Aurosweta, et al.
Veröffentlicht: (2026)
Investigating self-supervised features for expressive, multilingual voice conversion
von: Martín-Cortinas, Álvaro, et al.
Veröffentlicht: (2025)
von: Martín-Cortinas, Álvaro, et al.
Veröffentlicht: (2025)
RADAR Challenge 2026: Robust Audio Deepfake Recognition under Media Transformations
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2026)
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2026)
Asynchronous Voice Anonymization Using Adversarial Perturbation On Speaker Embedding
von: Wang, Rui, et al.
Veröffentlicht: (2024)
von: Wang, Rui, et al.
Veröffentlicht: (2024)
Spatio-spectral diarization of meetings by combining TDOA-based segmentation and speaker embedding-based clustering
von: Cord-Landwehr, Tobias, et al.
Veröffentlicht: (2025)
von: Cord-Landwehr, Tobias, et al.
Veröffentlicht: (2025)
Revisiting and Improving Scoring Fusion for Spoofing-aware Speaker Verification Using Compositional Data Analysis
von: Wang, Xin, et al.
Veröffentlicht: (2024)
von: Wang, Xin, et al.
Veröffentlicht: (2024)
Curriculum learning for self-supervised speaker verification
von: Heo, Hee-Soo, et al.
Veröffentlicht: (2022)
von: Heo, Hee-Soo, et al.
Veröffentlicht: (2022)
X-CrossNet: A complex spectral mapping approach to target speaker extraction with cross attention speaker embedding fusion
von: Sun, Chang, et al.
Veröffentlicht: (2024)
von: Sun, Chang, et al.
Veröffentlicht: (2024)
Hierarchical speaker representation for target speaker extraction
von: He, Shulin, et al.
Veröffentlicht: (2022)
von: He, Shulin, et al.
Veröffentlicht: (2022)
StreamVoiceAnon+: Emotion-Preserving Streaming Speaker Anonymization via Frame-Level Acoustic Distillation
von: Kuzmin, Nikita, et al.
Veröffentlicht: (2026)
von: Kuzmin, Nikita, et al.
Veröffentlicht: (2026)
Room Impulse Responses help attackers to evade Deep Fake Detection
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2024)
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2024)
Triage knowledge distillation for speaker verification
von: Kim, Ju-ho, et al.
Veröffentlicht: (2026)
von: Kim, Ju-ho, et al.
Veröffentlicht: (2026)
Privacy-oriented manipulation of speaker representations
von: Teixeira, Francisco, et al.
Veröffentlicht: (2023)
von: Teixeira, Francisco, et al.
Veröffentlicht: (2023)
Improving curriculum learning for target speaker extraction with synthetic speakers
von: Liu, Yun, et al.
Veröffentlicht: (2024)
von: Liu, Yun, et al.
Veröffentlicht: (2024)
On the effectiveness of enrollment speech augmentation for Target Speaker Extraction
von: Li, Junjie, et al.
Veröffentlicht: (2024)
von: Li, Junjie, et al.
Veröffentlicht: (2024)
Stream-Voice-Anon: Enhancing Utility of Real-Time Speaker Anonymization via Neural Audio Codec and Language Models
von: Kuzmin, Nikita, et al.
Veröffentlicht: (2026)
von: Kuzmin, Nikita, et al.
Veröffentlicht: (2026)
A multi-speaker multi-lingual voice cloning system based on vits2 for limmits 2024 challenge
von: Wang, Xiaopeng, et al.
Veröffentlicht: (2024)
von: Wang, Xiaopeng, et al.
Veröffentlicht: (2024)
SelfTTS: cross-speaker style transfer through explicit embedding disentanglement and self-refinement using self-augmentation
von: Ueda, Lucas H., et al.
Veröffentlicht: (2026)
von: Ueda, Lucas H., et al.
Veröffentlicht: (2026)
Subband Architecture Aided Selective Fixed-Filter Active Noise Control
von: Liang, Hong-Cheng, et al.
Veröffentlicht: (2025)
von: Liang, Hong-Cheng, et al.
Veröffentlicht: (2025)
Robust Localization of Partially Fake Speech: Metrics and Out-of-Domain Evaluation
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2025)
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Adversarial speech for voice privacy protection from Personalized Speech generation
von: Chen, Shihao, et al.
Veröffentlicht: (2024) -
Pinhole Effect on Linkability and Dispersion in Speaker Anonymization
von: Lee, Kong Aik, et al.
Veröffentlicht: (2025) -
IDMap: A Pseudo-Speaker Generator Framework Based on Speaker Identity Index to Vector Mapping
von: Liu, Zeyan, et al.
Veröffentlicht: (2025) -
Introducing voice timbre attribute detection
von: He, Jinghao, et al.
Veröffentlicht: (2025) -
A Study of the Removability of Speaker-Adversarial Perturbations
von: Chen, Liping, et al.
Veröffentlicht: (2025)