Preserving spoken content in voice anonymisation with character-level vocoder conditioning
Fuente:
arXiv
Saved in:
| Main Authors: | Panariello, Michele, Todisco, Massimiliano, Evans, Nicholas |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating voice anonymisation using similarity rank disclosure
by: Chandra, Shilpa, et al.
Published: (2026)
by: Chandra, Shilpa, et al.
Published: (2026)
Speaker anonymization using neural audio codec language models
by: Panariello, Michele, et al.
Published: (2023)
by: Panariello, Michele, et al.
Published: (2023)
Reference-free Adversarial Sex Obfuscation in Speech
by: Qu, Yangyang, et al.
Published: (2025)
by: Qu, Yangyang, et al.
Published: (2025)
Latent Secret Spin: Keyed Orthogonal Rotations for Blind Speech Watermarking in Anisotropic Latent Spaces
by: Coletta, Emma, et al.
Published: (2026)
by: Coletta, Emma, et al.
Published: (2026)
MDD: a Mask Diffusion Detector to Protect Speaker Verification Systems from Adversarial Perturbations
by: Bai, Yibo, et al.
Published: (2025)
by: Bai, Yibo, et al.
Published: (2025)
The Risks and Detection of Overestimated Privacy Protection in Voice Anonymisation
by: Panariello, Michele, et al.
Published: (2025)
by: Panariello, Michele, et al.
Published: (2025)
The VoicePrivacy 2022 Challenge: Progress and Perspectives in Voice Anonymisation
by: Panariello, Michele, et al.
Published: (2024)
by: Panariello, Michele, et al.
Published: (2024)
Explaining deep learning models for spoofing and deepfake detection with SHapley Additive exPlanations
by: Ge, Wanying, et al.
Published: (2021)
by: Ge, Wanying, et al.
Published: (2021)
The Third VoicePrivacy Challenge: Preserving Emotional Expressiveness and Linguistic Content in Voice Anonymization
by: Tomashenko, Natalia, et al.
Published: (2026)
by: Tomashenko, Natalia, et al.
Published: (2026)
Malacopula: adversarial automatic speaker verification attacks using a neural-based generalised Hammerstein model
by: Todisco, Massimiliano, et al.
Published: (2024)
by: Todisco, Massimiliano, et al.
Published: (2024)
Fair-Gate: Fairness-Aware Interpretable Risk Gating for Sex-Fair Voice Biometrics
by: Qu, Yangyang, et al.
Published: (2026)
by: Qu, Yangyang, et al.
Published: (2026)
A Comparison of Differential Performance Metrics for the Evaluation of Automatic Speaker Verification Fairness
by: Chouchane, Oubaida, et al.
Published: (2024)
by: Chouchane, Oubaida, et al.
Published: (2024)
The VoicePrivacy 2024 Challenge Evaluation Plan
by: Tomashenko, Natalia, et al.
Published: (2024)
by: Tomashenko, Natalia, et al.
Published: (2024)
Rethinking the joint estimation of magnitude and phase for time-frequency domain neural vocoders
by: Dai, Lingling, et al.
Published: (2025)
by: Dai, Lingling, et al.
Published: (2025)
Spoofing attack augmentation: can differently-trained attack models improve generalisation?
by: Ge, Wanying, et al.
Published: (2023)
by: Ge, Wanying, et al.
Published: (2023)
Synthesizing speech with selected perceptual voice qualities - A case study with creaky voice
by: Rautenberg, Frederik, et al.
Published: (2025)
by: Rautenberg, Frederik, et al.
Published: (2025)
Vocos: Closing the gap between time-domain and Fourier-based neural vocoders for high-quality audio synthesis
by: Siuzdak, Hubert
Published: (2023)
by: Siuzdak, Hubert
Published: (2023)
Identity leakage through accent cues in voice anonymisation
by: Bakari, Rayane, et al.
Published: (2026)
by: Bakari, Rayane, et al.
Published: (2026)
VoxSim: A perceptual voice similarity dataset
by: Ahn, Junseok, et al.
Published: (2024)
by: Ahn, Junseok, et al.
Published: (2024)
Encoding of lexical tone in self-supervised models of spoken language
by: Shen, Gaofei, et al.
Published: (2024)
by: Shen, Gaofei, et al.
Published: (2024)
Gender-ambiguous voice generation through feminine speaking style transfer in male voices
by: Koutsogiannaki, Maria, et al.
Published: (2024)
by: Koutsogiannaki, Maria, et al.
Published: (2024)
Investigation of perception inconsistency in speaker embedding for asynchronous voice anonymization
by: Wang, Rui, et al.
Published: (2025)
by: Wang, Rui, et al.
Published: (2025)
Investigating self-supervised features for expressive, multilingual voice conversion
by: Martín-Cortinas, Álvaro, et al.
Published: (2025)
by: Martín-Cortinas, Álvaro, et al.
Published: (2025)
Extracting accent features in spoken Brazilian Portuguese without sociolinguistic labels
by: Leite, Pedro H. L., et al.
Published: (2026)
by: Leite, Pedro H. L., et al.
Published: (2026)
Out-of-distribution generalisation in spoken language understanding
by: Porjazovski, Dejan, et al.
Published: (2024)
by: Porjazovski, Dejan, et al.
Published: (2024)
Wanna hear your voice? A sample is all we need!
by: Pham, The Hieu, et al.
Published: (2024)
by: Pham, The Hieu, et al.
Published: (2024)
Vclip: Face-based Speaker Generation by Face-voice Association Learning
by: Shi, Yao, et al.
Published: (2026)
by: Shi, Yao, et al.
Published: (2026)
Comparison of fundamental frequency estimators with subharmonic voice signals
by: Ikuma, Takeshi, et al.
Published: (2025)
by: Ikuma, Takeshi, et al.
Published: (2025)
Subjective quality evaluation of personalized own voice reconstruction systems
by: Ohlenbusch, Mattes, et al.
Published: (2025)
by: Ohlenbusch, Mattes, et al.
Published: (2025)
DNN-based ensemble singing voice synthesis with interactions between singers
by: Hyodo, Hiroaki, et al.
Published: (2024)
by: Hyodo, Hiroaki, et al.
Published: (2024)
Using voice analysis as an early indicator of risk for depression in young adults
by: Scherer, Klaus R., et al.
Published: (2024)
by: Scherer, Klaus R., et al.
Published: (2024)
Adversarial speech for voice privacy protection from Personalized Speech generation
by: Chen, Shihao, et al.
Published: (2024)
by: Chen, Shihao, et al.
Published: (2024)
Towards detecting the pathological subharmonic voicing with fully convolutional neural networks
by: Ikuma, Takeshi, et al.
Published: (2025)
by: Ikuma, Takeshi, et al.
Published: (2025)
ASVspoof 5: Design, Collection and Validation of Resources for Spoofing, Deepfake, and Adversarial Attack Detection Using Crowdsourced Speech
by: Wang, Xin, et al.
Published: (2025)
by: Wang, Xin, et al.
Published: (2025)
ASVspoof 5: Crowdsourced Speech Data, Deepfakes, and Adversarial Attacks at Scale
by: Wang, Xin, et al.
Published: (2024)
by: Wang, Xin, et al.
Published: (2024)
Building speech corpus with diverse voice characteristics for its prompt-based representation
by: Watanabe, Aya, et al.
Published: (2024)
by: Watanabe, Aya, et al.
Published: (2024)
Anonymization, Not Elimination: Utility-Preserved Speech Anonymization
by: Xiao, Yunchong, et al.
Published: (2026)
by: Xiao, Yunchong, et al.
Published: (2026)
Effects of automotive microphone frequency response characteristics and noise conditions on speech and ASR quality -- an experimental evaluation
by: Buccoli, Michele, et al.
Published: (2025)
by: Buccoli, Michele, et al.
Published: (2025)
Audiovisual angle and voice incongruence do not affect audiovisual verbal short-term memory in virtual reality
by: Ermert, Cosima A., et al.
Published: (2024)
by: Ermert, Cosima A., et al.
Published: (2024)
Detecting the terminality of speech-turn boundary for spoken interactions in French TV and Radio content
by: Uro, Rémi, et al.
Published: (2024)
by: Uro, Rémi, et al.
Published: (2024)
Similar Items
-
Evaluating voice anonymisation using similarity rank disclosure
by: Chandra, Shilpa, et al.
Published: (2026) -
Speaker anonymization using neural audio codec language models
by: Panariello, Michele, et al.
Published: (2023) -
Reference-free Adversarial Sex Obfuscation in Speech
by: Qu, Yangyang, et al.
Published: (2025) -
Latent Secret Spin: Keyed Orthogonal Rotations for Blind Speech Watermarking in Anisotropic Latent Spaces
by: Coletta, Emma, et al.
Published: (2026) -
MDD: a Mask Diffusion Detector to Protect Speaker Verification Systems from Adversarial Perturbations
by: Bai, Yibo, et al.
Published: (2025)