Voice Conversion-based Privacy through Adversarial Information Hiding
Fuente:
arXiv
Guardado en:
| Autores principales: | Webber, Jacob J, Watts, Oliver, Henter, Gustav Eje, Williams, Jennifer, King, Simon |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
When Voice Matters: Evidence of Gender Disparity in Positional Bias of SpeechLLMs
por: Satish, Shree Harsha Bokkahalli, et al.
Publicado: (2025)
por: Satish, Shree Harsha Bokkahalli, et al.
Publicado: (2025)
Do Bias Benchmarks Generalise? Evidence from Voice-based Evaluation of Gender Bias in SpeechLLMs
por: Satish, Shree Harsha Bokkahalli, et al.
Publicado: (2025)
por: Satish, Shree Harsha Bokkahalli, et al.
Publicado: (2025)
Speak Your Mind: The Speech Continuation Task as a Probe of Voice-Based Model Bias
por: Satish, Shree Harsha Bokkahalli, et al.
Publicado: (2025)
por: Satish, Shree Harsha Bokkahalli, et al.
Publicado: (2025)
HiFi-Glot: High-Fidelity Neural Formant Synthesis with Differentiable Resonant Filters
por: Gu, Yicheng, et al.
Publicado: (2024)
por: Gu, Yicheng, et al.
Publicado: (2024)
The Voice Behind the Words: Quantifying Intersectional Bias in SpeechLLMs
por: Satish, Shree Harsha Bokkahalli, et al.
Publicado: (2026)
por: Satish, Shree Harsha Bokkahalli, et al.
Publicado: (2026)
VoXtream: Full-Stream Text-to-Speech with Extremely Low Latency
por: Torgashov, Nikita, et al.
Publicado: (2025)
por: Torgashov, Nikita, et al.
Publicado: (2025)
VoXtream2: Full-stream TTS with dynamic speaking rate control
por: Torgashov, Nikita, et al.
Publicado: (2026)
por: Torgashov, Nikita, et al.
Publicado: (2026)
Comparator Loss: An Ordinal Contrastive Loss to Derive a Severity Score for Speech-based Health Monitoring
por: Webber, Jacob J, et al.
Publicado: (2025)
por: Webber, Jacob J, et al.
Publicado: (2025)
Spatial Voice Conversion: Voice Conversion Preserving Spatial Information and Non-target Signals
por: Seki, Kentaro, et al.
Publicado: (2024)
por: Seki, Kentaro, et al.
Publicado: (2024)
RobustSVC: HuBERT-based Melody Extractor and Adversarial Learning for Robust Singing Voice Conversion
por: Chen, Wei, et al.
Publicado: (2024)
por: Chen, Wei, et al.
Publicado: (2024)
RAVE for Speech: Efficient Voice Conversion at High Sampling Rates
por: Bargum, Anders R., et al.
Publicado: (2024)
por: Bargum, Anders R., et al.
Publicado: (2024)
Voice-ENHANCE: Speech Restoration using a Diffusion-based Voice Conversion Framework
por: Byun, Kyungguen, et al.
Publicado: (2025)
por: Byun, Kyungguen, et al.
Publicado: (2025)
ZSVC: Zero-shot Style Voice Conversion with Disentangled Latent Diffusion Models and Adversarial Training
por: Zhu, Xinfa, et al.
Publicado: (2025)
por: Zhu, Xinfa, et al.
Publicado: (2025)
Gelina: Unified Speech and Gesture Synthesis via Interleaved Token Prediction
por: Guichoux, Téo, et al.
Publicado: (2025)
por: Guichoux, Téo, et al.
Publicado: (2025)
Disentangling the Prosody and Semantic Information with Pre-trained Model for In-Context Learning based Zero-Shot Voice Conversion
por: Chen, Zhengyang, et al.
Publicado: (2024)
por: Chen, Zhengyang, et al.
Publicado: (2024)
Neural Concatenative Singing Voice Conversion: Rethinking Concatenation-Based Approach for One-Shot Singing Voice Conversion
por: Sha, Binzhu, et al.
Publicado: (2023)
por: Sha, Binzhu, et al.
Publicado: (2023)
StreamVoice+: Evolving into End-to-end Streaming Zero-shot Voice Conversion
por: Wang, Zhichao, et al.
Publicado: (2024)
por: Wang, Zhichao, et al.
Publicado: (2024)
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant
por: Hou, Yixuan, et al.
Publicado: (2025)
por: Hou, Yixuan, et al.
Publicado: (2025)
Leveraging Diverse Semantic-based Audio Pretrained Models for Singing Voice Conversion
por: Zhang, Xueyao, et al.
Publicado: (2023)
por: Zhang, Xueyao, et al.
Publicado: (2023)
Generative Adversarial Network based Voice Conversion: Techniques, Challenges, and Recent Advancements
por: Dhar, Sandipan, et al.
Publicado: (2025)
por: Dhar, Sandipan, et al.
Publicado: (2025)
Generating Novel and Realistic Speakers for Voice Conversion
por: Chen, Meiying Melissa, et al.
Publicado: (2025)
por: Chen, Meiying Melissa, et al.
Publicado: (2025)
LatentVoiceGrad: Nonparallel Voice Conversion with Latent Diffusion/Flow-Matching Models
por: Kameoka, Hirokazu, et al.
Publicado: (2025)
por: Kameoka, Hirokazu, et al.
Publicado: (2025)
VoiceGrad: Non-Parallel Any-to-Many Voice Conversion with Annealed Langevin Dynamics
por: Kameoka, Hirokazu, et al.
Publicado: (2020)
por: Kameoka, Hirokazu, et al.
Publicado: (2020)
Zero-Shot Sing Voice Conversion: built upon clustering-based phoneme representations
por: Zhou, Wangjin, et al.
Publicado: (2024)
por: Zhou, Wangjin, et al.
Publicado: (2024)
On the Generation and Removal of Speaker Adversarial Perturbation for Voice-Privacy Protection
por: Guo, Chenyang, et al.
Publicado: (2024)
por: Guo, Chenyang, et al.
Publicado: (2024)
Converting Anyone's Voice: End-to-End Expressive Voice Conversion with a Conditional Diffusion Model
por: Du, Zongyang, et al.
Publicado: (2024)
por: Du, Zongyang, et al.
Publicado: (2024)
Enhancing Polyglot Voices by Leveraging Cross-Lingual Fine-Tuning in Any-to-One Voice Conversion
por: Ruggiero, Giuseppe, et al.
Publicado: (2024)
por: Ruggiero, Giuseppe, et al.
Publicado: (2024)
OneVoice: One Model, Triple Scenarios-Towards Unified Zero-shot Voice Conversion
por: Wang, Zhichao, et al.
Publicado: (2026)
por: Wang, Zhichao, et al.
Publicado: (2026)
Residual Speaker Representation for One-Shot Voice Conversion
por: Xu, Le, et al.
Publicado: (2023)
por: Xu, Le, et al.
Publicado: (2023)
StreamVoice: Streamable Context-Aware Language Modeling for Real-time Zero-Shot Voice Conversion
por: Wang, Zhichao, et al.
Publicado: (2024)
por: Wang, Zhichao, et al.
Publicado: (2024)
Does Your Voice Assistant Remember? Analyzing Conversational Context Recall and Utilization in Voice Interaction Models
por: Kim, Heeseung, et al.
Publicado: (2025)
por: Kim, Heeseung, et al.
Publicado: (2025)
VC-ENHANCE: Speech Restoration with Integrated Noise Suppression and Voice Conversion
por: Byun, Kyungguen, et al.
Publicado: (2024)
por: Byun, Kyungguen, et al.
Publicado: (2024)
SRC4VC: Smartphone-Recorded Corpus for Voice Conversion Benchmark
por: Saito, Yuki, et al.
Publicado: (2024)
por: Saito, Yuki, et al.
Publicado: (2024)
An Extensive Analysis of the Singing Voice Conversion Challenge 2025 Evaluation Results
por: Violeta, Lester Phillip, et al.
Publicado: (2025)
por: Violeta, Lester Phillip, et al.
Publicado: (2025)
End-to-End Zero-Shot Voice Conversion with Location-Variable Convolutions
por: Kang, Wonjune, et al.
Publicado: (2022)
por: Kang, Wonjune, et al.
Publicado: (2022)
FreeSVC: Towards Zero-shot Multilingual Singing Voice Conversion
por: Ferreira, Alef Iury Siqueira, et al.
Publicado: (2025)
por: Ferreira, Alef Iury Siqueira, et al.
Publicado: (2025)
Collective Learning Mechanism based Optimal Transport Generative Adversarial Network for Non-parallel Voice Conversion
por: Dhar, Sandipan, et al.
Publicado: (2025)
por: Dhar, Sandipan, et al.
Publicado: (2025)
Adversarial Multi-Task Learning for Disentangling Timbre and Pitch in Singing Voice Synthesis
por: Kim, Tae-Woo, et al.
Publicado: (2022)
por: Kim, Tae-Woo, et al.
Publicado: (2022)
REWIND: Speech Time Reversal for Enhancing Speaker Representations in Diffusion-based Voice Conversion
por: Biyani, Ishan D., et al.
Publicado: (2025)
por: Biyani, Ishan D., et al.
Publicado: (2025)
Self-Supervised Singing Voice Pre-Training towards Speech-to-Singing Conversion
por: Li, Ruiqi, et al.
Publicado: (2024)
por: Li, Ruiqi, et al.
Publicado: (2024)
Ejemplares similares
-
When Voice Matters: Evidence of Gender Disparity in Positional Bias of SpeechLLMs
por: Satish, Shree Harsha Bokkahalli, et al.
Publicado: (2025) -
Do Bias Benchmarks Generalise? Evidence from Voice-based Evaluation of Gender Bias in SpeechLLMs
por: Satish, Shree Harsha Bokkahalli, et al.
Publicado: (2025) -
Speak Your Mind: The Speech Continuation Task as a Probe of Voice-Based Model Bias
por: Satish, Shree Harsha Bokkahalli, et al.
Publicado: (2025) -
HiFi-Glot: High-Fidelity Neural Formant Synthesis with Differentiable Resonant Filters
por: Gu, Yicheng, et al.
Publicado: (2024) -
The Voice Behind the Words: Quantifying Intersectional Bias in SpeechLLMs
por: Satish, Shree Harsha Bokkahalli, et al.
Publicado: (2026)