Low-Cost Detection of Degraded Voice Clones via Source-Output Acoustic Consistency
Fuente:
arXiv
Guardado en:
| Autores principales: | Shokr, Jana, Papadopoulos, Minos, Cooperstock, Jeremy, Orepic, Pavo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Pronunciation Deviation Analysis Through Voice Cloning and Acoustic Comparison
por: Valdivia, Andrew, et al.
Publicado: (2025)
por: Valdivia, Andrew, et al.
Publicado: (2025)
Spoken Language Corpora Augmentation with Domain-Specific Voice-Cloned Speech
por: Czyżnikiewicz, Mateusz, et al.
Publicado: (2024)
por: Czyżnikiewicz, Mateusz, et al.
Publicado: (2024)
Voice Timbre Attribute Detection with Compact and Interpretable Training-Free Acoustic Parameters
por: Chiu, Aemon Yat Fei, et al.
Publicado: (2026)
por: Chiu, Aemon Yat Fei, et al.
Publicado: (2026)
Open-Source System for Multilingual Translation and Cloned Speech Synthesis
por: Cámara, Mateo, et al.
Publicado: (2025)
por: Cámara, Mateo, et al.
Publicado: (2025)
SF-Speech: Straightened Flow for Zero-Shot Voice Clone
por: Li, Xuyuan, et al.
Publicado: (2024)
por: Li, Xuyuan, et al.
Publicado: (2024)
OpenVoice: Versatile Instant Voice Cloning
por: Qin, Zengyi, et al.
Publicado: (2023)
por: Qin, Zengyi, et al.
Publicado: (2023)
Voice Cloning: Comprehensive Survey
por: Azzuni, Hussam, et al.
Publicado: (2025)
por: Azzuni, Hussam, et al.
Publicado: (2025)
One Voice, Many Tongues: Cross-Lingual Voice Cloning for Scientific Speech
por: Abebe, Amanuel Gizachew, et al.
Publicado: (2026)
por: Abebe, Amanuel Gizachew, et al.
Publicado: (2026)
Multi-Input Multi-Output Target-Speaker Voice Activity Detection For Unified, Flexible, and Robust Audio-Visual Speaker Diarization
por: Cheng, Ming, et al.
Publicado: (2024)
por: Cheng, Ming, et al.
Publicado: (2024)
Audio Avatar Fingerprinting: An Approach for Authorized Use of Voice Cloning in the Era of Synthetic Audio
por: Gerstner, Candice R.
Publicado: (2026)
por: Gerstner, Candice R.
Publicado: (2026)
CAVEMOVE: An Acoustic Database for the Study of Voice-enabled Technologies inside Moving Vehicles
por: Stefanakis, Nikolaos, et al.
Publicado: (2025)
por: Stefanakis, Nikolaos, et al.
Publicado: (2025)
CloneShield: A Framework for Universal Perturbation Against Zero-Shot Voice Cloning
por: Li, Renyuan, et al.
Publicado: (2025)
por: Li, Renyuan, et al.
Publicado: (2025)
DiffVQE: Hybrid Diffusion Voice Quality Enhancement Under Acoustic Echo and Noise
por: Girao, Haljan Lugo, et al.
Publicado: (2026)
por: Girao, Haljan Lugo, et al.
Publicado: (2026)
Fed-PISA: Federated Voice Cloning via Personalized Identity-Style Adaptation
por: Wang, Qi, et al.
Publicado: (2025)
por: Wang, Qi, et al.
Publicado: (2025)
X-Voice: Enabling Everyone to Speak 30 Languages via Zero-Shot Cross-Lingual Voice Cloning
por: Xu, Rixi, et al.
Publicado: (2026)
por: Xu, Rixi, et al.
Publicado: (2026)
RVCBench: Benchmarking the Robustness of Voice Cloning Across Modern Audio Generation Models
por: Jin, Ruinan, et al.
Publicado: (2026)
por: Jin, Ruinan, et al.
Publicado: (2026)
VoiceCraft-X: Unifying Multilingual, Voice-Cloning Speech Synthesis and Speech Editing
por: Zheng, Zhisheng, et al.
Publicado: (2025)
por: Zheng, Zhisheng, et al.
Publicado: (2025)
The THU-HCSI Multi-Speaker Multi-Lingual Few-Shot Voice Cloning System for LIMMITS'24 Challenge
por: Zhou, Yixuan, et al.
Publicado: (2024)
por: Zhou, Yixuan, et al.
Publicado: (2024)
JoyTTS: LLM-based Spoken Chatbot With Voice Cloning
por: Zhou, Fangru, et al.
Publicado: (2025)
por: Zhou, Fangru, et al.
Publicado: (2025)
Multi-modal Adversarial Training for Zero-Shot Voice Cloning
por: Janiczek, John, et al.
Publicado: (2024)
por: Janiczek, John, et al.
Publicado: (2024)
Spatial Reverberation and Dereverberation using an Acoustic Multiple-Input Multiple-Output System
por: Morgenstern, Hai, et al.
Publicado: (2024)
por: Morgenstern, Hai, et al.
Publicado: (2024)
Low-Complexity Own Voice Reconstruction for Hearables with an In-Ear Microphone
por: Ohlenbusch, Mattes, et al.
Publicado: (2024)
por: Ohlenbusch, Mattes, et al.
Publicado: (2024)
Auto-Landmark: Acoustic Landmark Dataset and Open-Source Toolkit for Landmark Extraction
por: Zhang, Xiangyu, et al.
Publicado: (2024)
por: Zhang, Xiangyu, et al.
Publicado: (2024)
Robust Speech Activity Detection in the Presence of Singing Voice
por: Grundhuber, Philipp, et al.
Publicado: (2025)
por: Grundhuber, Philipp, et al.
Publicado: (2025)
The Risks and Detection of Overestimated Privacy Protection in Voice Anonymisation
por: Panariello, Michele, et al.
Publicado: (2025)
por: Panariello, Michele, et al.
Publicado: (2025)
Edit Content, Preserve Acoustics: Imperceptible Text-Based Speech Editing via Self-Consistency Rewards
por: Ren, Yong, et al.
Publicado: (2026)
por: Ren, Yong, et al.
Publicado: (2026)
Joint Feature and Output Distillation for Low-complexity Acoustic Scene Classification
por: Li, Haowen, et al.
Publicado: (2025)
por: Li, Haowen, et al.
Publicado: (2025)
Interpreting the Role of Visemes in Audio-Visual Speech Recognition
por: Papadopoulos, Aristeidis, et al.
Publicado: (2025)
por: Papadopoulos, Aristeidis, et al.
Publicado: (2025)
Where's That Voice Coming? Continual Learning for Sound Source Localization
por: Xiao, Yang, et al.
Publicado: (2024)
por: Xiao, Yang, et al.
Publicado: (2024)
FastTurn: Unifying Acoustic and Streaming Semantic Cues for Low-Latency and Robust Turn Detection
por: Wang, Chengyou, et al.
Publicado: (2026)
por: Wang, Chengyou, et al.
Publicado: (2026)
VoiceSculptor: Your Voice, Designed By You
por: Hu, Jingbin, et al.
Publicado: (2026)
por: Hu, Jingbin, et al.
Publicado: (2026)
DreamVoice: Text-Guided Voice Conversion
por: Hai, Jiarui, et al.
Publicado: (2024)
por: Hai, Jiarui, et al.
Publicado: (2024)
LCM-SVC: Latent Diffusion Model Based Singing Voice Conversion with Inference Acceleration via Latent Consistency Distillation
por: Chen, Shihao, et al.
Publicado: (2024)
por: Chen, Shihao, et al.
Publicado: (2024)
Speech Synthesis along Perceptual Voice Quality Dimensions
por: Rautenberg, Frederik, et al.
Publicado: (2025)
por: Rautenberg, Frederik, et al.
Publicado: (2025)
OmniCodec: Low Frame Rate Universal Audio Codec with Semantic-Acoustic Disentanglement
por: Hu, Jingbin, et al.
Publicado: (2026)
por: Hu, Jingbin, et al.
Publicado: (2026)
Data-Efficient Low-Complexity Acoustic Scene Classification via Distilling and Progressive Pruning
por: Han, Bing, et al.
Publicado: (2024)
por: Han, Bing, et al.
Publicado: (2024)
Fundamentals of Data-Driven Approaches to Acoustic Signal Detection, Filtering, and Transformation
por: Pan, Chao
Publicado: (2025)
por: Pan, Chao
Publicado: (2025)
Voice Cloning for Dysarthric Speech Synthesis: Addressing Data Scarcity in Speech-Language Pathology
por: Moell, Birger, et al.
Publicado: (2025)
por: Moell, Birger, et al.
Publicado: (2025)
RRP-Voice: A Longitudinal Dataset and Benchmark for Recurrent Respiratory Papillomatosis Detection
por: Ren, Wenze, et al.
Publicado: (2026)
por: Ren, Wenze, et al.
Publicado: (2026)
VS-Singer: Vision-Guided Stereo Singing Voice Synthesis with Consistency Schrödinger Bridge
por: Zhao, Zijing, et al.
Publicado: (2025)
por: Zhao, Zijing, et al.
Publicado: (2025)
Ejemplares similares
-
Pronunciation Deviation Analysis Through Voice Cloning and Acoustic Comparison
por: Valdivia, Andrew, et al.
Publicado: (2025) -
Spoken Language Corpora Augmentation with Domain-Specific Voice-Cloned Speech
por: Czyżnikiewicz, Mateusz, et al.
Publicado: (2024) -
Voice Timbre Attribute Detection with Compact and Interpretable Training-Free Acoustic Parameters
por: Chiu, Aemon Yat Fei, et al.
Publicado: (2026) -
Open-Source System for Multilingual Translation and Cloned Speech Synthesis
por: Cámara, Mateo, et al.
Publicado: (2025) -
SF-Speech: Straightened Flow for Zero-Shot Voice Clone
por: Li, Xuyuan, et al.
Publicado: (2024)