Preserving spoken content in voice anonymisation with character-level vocoder conditioning
Fuente:
arXiv
Salvato in:
| Autori principali: | Panariello, Michele, Todisco, Massimiliano, Evans, Nicholas |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Evaluating voice anonymisation using similarity rank disclosure
di: Chandra, Shilpa, et al.
Pubblicazione: (2026)
di: Chandra, Shilpa, et al.
Pubblicazione: (2026)
Speaker anonymization using neural audio codec language models
di: Panariello, Michele, et al.
Pubblicazione: (2023)
di: Panariello, Michele, et al.
Pubblicazione: (2023)
Reference-free Adversarial Sex Obfuscation in Speech
di: Qu, Yangyang, et al.
Pubblicazione: (2025)
di: Qu, Yangyang, et al.
Pubblicazione: (2025)
Latent Secret Spin: Keyed Orthogonal Rotations for Blind Speech Watermarking in Anisotropic Latent Spaces
di: Coletta, Emma, et al.
Pubblicazione: (2026)
di: Coletta, Emma, et al.
Pubblicazione: (2026)
MDD: a Mask Diffusion Detector to Protect Speaker Verification Systems from Adversarial Perturbations
di: Bai, Yibo, et al.
Pubblicazione: (2025)
di: Bai, Yibo, et al.
Pubblicazione: (2025)
The Risks and Detection of Overestimated Privacy Protection in Voice Anonymisation
di: Panariello, Michele, et al.
Pubblicazione: (2025)
di: Panariello, Michele, et al.
Pubblicazione: (2025)
The VoicePrivacy 2022 Challenge: Progress and Perspectives in Voice Anonymisation
di: Panariello, Michele, et al.
Pubblicazione: (2024)
di: Panariello, Michele, et al.
Pubblicazione: (2024)
Explaining deep learning models for spoofing and deepfake detection with SHapley Additive exPlanations
di: Ge, Wanying, et al.
Pubblicazione: (2021)
di: Ge, Wanying, et al.
Pubblicazione: (2021)
The Third VoicePrivacy Challenge: Preserving Emotional Expressiveness and Linguistic Content in Voice Anonymization
di: Tomashenko, Natalia, et al.
Pubblicazione: (2026)
di: Tomashenko, Natalia, et al.
Pubblicazione: (2026)
Malacopula: adversarial automatic speaker verification attacks using a neural-based generalised Hammerstein model
di: Todisco, Massimiliano, et al.
Pubblicazione: (2024)
di: Todisco, Massimiliano, et al.
Pubblicazione: (2024)
Fair-Gate: Fairness-Aware Interpretable Risk Gating for Sex-Fair Voice Biometrics
di: Qu, Yangyang, et al.
Pubblicazione: (2026)
di: Qu, Yangyang, et al.
Pubblicazione: (2026)
A Comparison of Differential Performance Metrics for the Evaluation of Automatic Speaker Verification Fairness
di: Chouchane, Oubaida, et al.
Pubblicazione: (2024)
di: Chouchane, Oubaida, et al.
Pubblicazione: (2024)
The VoicePrivacy 2024 Challenge Evaluation Plan
di: Tomashenko, Natalia, et al.
Pubblicazione: (2024)
di: Tomashenko, Natalia, et al.
Pubblicazione: (2024)
Rethinking the joint estimation of magnitude and phase for time-frequency domain neural vocoders
di: Dai, Lingling, et al.
Pubblicazione: (2025)
di: Dai, Lingling, et al.
Pubblicazione: (2025)
Spoofing attack augmentation: can differently-trained attack models improve generalisation?
di: Ge, Wanying, et al.
Pubblicazione: (2023)
di: Ge, Wanying, et al.
Pubblicazione: (2023)
Synthesizing speech with selected perceptual voice qualities - A case study with creaky voice
di: Rautenberg, Frederik, et al.
Pubblicazione: (2025)
di: Rautenberg, Frederik, et al.
Pubblicazione: (2025)
Vocos: Closing the gap between time-domain and Fourier-based neural vocoders for high-quality audio synthesis
di: Siuzdak, Hubert
Pubblicazione: (2023)
di: Siuzdak, Hubert
Pubblicazione: (2023)
Identity leakage through accent cues in voice anonymisation
di: Bakari, Rayane, et al.
Pubblicazione: (2026)
di: Bakari, Rayane, et al.
Pubblicazione: (2026)
VoxSim: A perceptual voice similarity dataset
di: Ahn, Junseok, et al.
Pubblicazione: (2024)
di: Ahn, Junseok, et al.
Pubblicazione: (2024)
Encoding of lexical tone in self-supervised models of spoken language
di: Shen, Gaofei, et al.
Pubblicazione: (2024)
di: Shen, Gaofei, et al.
Pubblicazione: (2024)
Gender-ambiguous voice generation through feminine speaking style transfer in male voices
di: Koutsogiannaki, Maria, et al.
Pubblicazione: (2024)
di: Koutsogiannaki, Maria, et al.
Pubblicazione: (2024)
Investigation of perception inconsistency in speaker embedding for asynchronous voice anonymization
di: Wang, Rui, et al.
Pubblicazione: (2025)
di: Wang, Rui, et al.
Pubblicazione: (2025)
Investigating self-supervised features for expressive, multilingual voice conversion
di: Martín-Cortinas, Álvaro, et al.
Pubblicazione: (2025)
di: Martín-Cortinas, Álvaro, et al.
Pubblicazione: (2025)
Extracting accent features in spoken Brazilian Portuguese without sociolinguistic labels
di: Leite, Pedro H. L., et al.
Pubblicazione: (2026)
di: Leite, Pedro H. L., et al.
Pubblicazione: (2026)
Out-of-distribution generalisation in spoken language understanding
di: Porjazovski, Dejan, et al.
Pubblicazione: (2024)
di: Porjazovski, Dejan, et al.
Pubblicazione: (2024)
Wanna hear your voice? A sample is all we need!
di: Pham, The Hieu, et al.
Pubblicazione: (2024)
di: Pham, The Hieu, et al.
Pubblicazione: (2024)
Vclip: Face-based Speaker Generation by Face-voice Association Learning
di: Shi, Yao, et al.
Pubblicazione: (2026)
di: Shi, Yao, et al.
Pubblicazione: (2026)
Comparison of fundamental frequency estimators with subharmonic voice signals
di: Ikuma, Takeshi, et al.
Pubblicazione: (2025)
di: Ikuma, Takeshi, et al.
Pubblicazione: (2025)
Subjective quality evaluation of personalized own voice reconstruction systems
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2025)
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2025)
DNN-based ensemble singing voice synthesis with interactions between singers
di: Hyodo, Hiroaki, et al.
Pubblicazione: (2024)
di: Hyodo, Hiroaki, et al.
Pubblicazione: (2024)
Using voice analysis as an early indicator of risk for depression in young adults
di: Scherer, Klaus R., et al.
Pubblicazione: (2024)
di: Scherer, Klaus R., et al.
Pubblicazione: (2024)
Adversarial speech for voice privacy protection from Personalized Speech generation
di: Chen, Shihao, et al.
Pubblicazione: (2024)
di: Chen, Shihao, et al.
Pubblicazione: (2024)
Towards detecting the pathological subharmonic voicing with fully convolutional neural networks
di: Ikuma, Takeshi, et al.
Pubblicazione: (2025)
di: Ikuma, Takeshi, et al.
Pubblicazione: (2025)
ASVspoof 5: Design, Collection and Validation of Resources for Spoofing, Deepfake, and Adversarial Attack Detection Using Crowdsourced Speech
di: Wang, Xin, et al.
Pubblicazione: (2025)
di: Wang, Xin, et al.
Pubblicazione: (2025)
ASVspoof 5: Crowdsourced Speech Data, Deepfakes, and Adversarial Attacks at Scale
di: Wang, Xin, et al.
Pubblicazione: (2024)
di: Wang, Xin, et al.
Pubblicazione: (2024)
Building speech corpus with diverse voice characteristics for its prompt-based representation
di: Watanabe, Aya, et al.
Pubblicazione: (2024)
di: Watanabe, Aya, et al.
Pubblicazione: (2024)
Anonymization, Not Elimination: Utility-Preserved Speech Anonymization
di: Xiao, Yunchong, et al.
Pubblicazione: (2026)
di: Xiao, Yunchong, et al.
Pubblicazione: (2026)
Effects of automotive microphone frequency response characteristics and noise conditions on speech and ASR quality -- an experimental evaluation
di: Buccoli, Michele, et al.
Pubblicazione: (2025)
di: Buccoli, Michele, et al.
Pubblicazione: (2025)
Audiovisual angle and voice incongruence do not affect audiovisual verbal short-term memory in virtual reality
di: Ermert, Cosima A., et al.
Pubblicazione: (2024)
di: Ermert, Cosima A., et al.
Pubblicazione: (2024)
Detecting the terminality of speech-turn boundary for spoken interactions in French TV and Radio content
di: Uro, Rémi, et al.
Pubblicazione: (2024)
di: Uro, Rémi, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Evaluating voice anonymisation using similarity rank disclosure
di: Chandra, Shilpa, et al.
Pubblicazione: (2026) -
Speaker anonymization using neural audio codec language models
di: Panariello, Michele, et al.
Pubblicazione: (2023) -
Reference-free Adversarial Sex Obfuscation in Speech
di: Qu, Yangyang, et al.
Pubblicazione: (2025) -
Latent Secret Spin: Keyed Orthogonal Rotations for Blind Speech Watermarking in Anisotropic Latent Spaces
di: Coletta, Emma, et al.
Pubblicazione: (2026) -
MDD: a Mask Diffusion Detector to Protect Speaker Verification Systems from Adversarial Perturbations
di: Bai, Yibo, et al.
Pubblicazione: (2025)