OpenVoice: Versatile Instant Voice Cloning
Fuente:
arXiv
Guardado en:
| Autores principales: | Qin, Zengyi, Zhao, Wenliang, Yu, Xumin, Sun, Xin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multi-modal Adversarial Training for Zero-Shot Voice Cloning
por: Janiczek, John, et al.
Publicado: (2024)
por: Janiczek, John, et al.
Publicado: (2024)
Improving Generalization for AI-Synthesized Voice Detection
por: Ren, Hainan, et al.
Publicado: (2024)
por: Ren, Hainan, et al.
Publicado: (2024)
BiSinger: Bilingual Singing Voice Synthesis
por: Zhou, Huali, et al.
Publicado: (2023)
por: Zhou, Huali, et al.
Publicado: (2023)
NaturalVoices: A Large-Scale, Spontaneous and Emotional Podcast Dataset for Voice Conversion
por: Du, Zongyang, et al.
Publicado: (2025)
por: Du, Zongyang, et al.
Publicado: (2025)
VANPY: Voice Analysis Framework
por: Koushnir, Gregory, et al.
Publicado: (2025)
por: Koushnir, Gregory, et al.
Publicado: (2025)
Systematic FAIRness Assessment of Open Voice Biomarker Datasets for Mental Health and Neurodegenerative Diseases
por: Mahapatra, Ishaan, et al.
Publicado: (2025)
por: Mahapatra, Ishaan, et al.
Publicado: (2025)
Compact Neural TTS Voices for Accessibility
por: Jain, Kunal, et al.
Publicado: (2025)
por: Jain, Kunal, et al.
Publicado: (2025)
Speech to Speech Synthesis for Voice Impersonation
por: Johnson, Bjorn, et al.
Publicado: (2026)
por: Johnson, Bjorn, et al.
Publicado: (2026)
Discrete Optimal Transport and Voice Conversion
por: Selitskiy, Anton, et al.
Publicado: (2025)
por: Selitskiy, Anton, et al.
Publicado: (2025)
Pronunciation Deviation Analysis Through Voice Cloning and Acoustic Comparison
por: Valdivia, Andrew, et al.
Publicado: (2025)
por: Valdivia, Andrew, et al.
Publicado: (2025)
Zero-shot Voice Conversion with Diffusion Transformers
por: Liu, Songting
Publicado: (2024)
por: Liu, Songting
Publicado: (2024)
Optimal Transport Maps are Good Voice Converters
por: Asadulaev, Arip, et al.
Publicado: (2024)
por: Asadulaev, Arip, et al.
Publicado: (2024)
SF-Speech: Straightened Flow for Zero-Shot Voice Clone
por: Li, Xuyuan, et al.
Publicado: (2024)
por: Li, Xuyuan, et al.
Publicado: (2024)
SEF-MK: Speaker-Embedding-Free Voice Anonymization through Multi-k-means Quantization
por: Tang, Beilong, et al.
Publicado: (2025)
por: Tang, Beilong, et al.
Publicado: (2025)
ConSinger: Efficient High-Fidelity Singing Voice Generation with Minimal Steps
por: Song, Yulin, et al.
Publicado: (2024)
por: Song, Yulin, et al.
Publicado: (2024)
A Concept-based approach to Voice Disorder Detection
por: Ghia, Davide, et al.
Publicado: (2025)
por: Ghia, Davide, et al.
Publicado: (2025)
Tessellated Linear Model for Age Prediction from Voice
por: Alharthi, Dareen, et al.
Publicado: (2025)
por: Alharthi, Dareen, et al.
Publicado: (2025)
Voice Cloning: Comprehensive Survey
por: Azzuni, Hussam, et al.
Publicado: (2025)
por: Azzuni, Hussam, et al.
Publicado: (2025)
DiffAnon: Diffusion-based Prosody Control for Voice Anonymization
por: Ulgen, Ismail Rasim, et al.
Publicado: (2026)
por: Ulgen, Ismail Rasim, et al.
Publicado: (2026)
Voice Signal Processing for Machine Learning. The Case of Speaker Isolation
por: Ganchev, Radan
Publicado: (2024)
por: Ganchev, Radan
Publicado: (2024)
On the Generation and Removal of Speaker Adversarial Perturbation for Voice-Privacy Protection
por: Guo, Chenyang, et al.
Publicado: (2024)
por: Guo, Chenyang, et al.
Publicado: (2024)
StreamVC: Real-Time Low-Latency Voice Conversion
por: Yang, Yang, et al.
Publicado: (2024)
por: Yang, Yang, et al.
Publicado: (2024)
Discrete Unit based Masking for Improving Disentanglement in Voice Conversion
por: Lee, Philip H., et al.
Publicado: (2024)
por: Lee, Philip H., et al.
Publicado: (2024)
Towards Robust Assessment of Pathological Voices via Combined Low-Level Descriptors and Foundation Model Representations
por: Ariyanti, Whenty, et al.
Publicado: (2025)
por: Ariyanti, Whenty, et al.
Publicado: (2025)
SVSNet+: Enhancing Speaker Voice Similarity Assessment Models with Representations from Speech Foundation Models
por: Yin, Chun, et al.
Publicado: (2024)
por: Yin, Chun, et al.
Publicado: (2024)
De-AntiFake: Rethinking the Protective Perturbations Against Voice Cloning Attacks
por: Fan, Wei, et al.
Publicado: (2025)
por: Fan, Wei, et al.
Publicado: (2025)
Phoneme Hallucinator: One-shot Voice Conversion via Set Expansion
por: Shan, Siyuan, et al.
Publicado: (2023)
por: Shan, Siyuan, et al.
Publicado: (2023)
Voice Conversion with Diverse Intonation using Conditional Variational Auto-Encoder
por: Suh, Soobin, et al.
Publicado: (2025)
por: Suh, Soobin, et al.
Publicado: (2025)
VoiceCraft-X: Unifying Multilingual, Voice-Cloning Speech Synthesis and Speech Editing
por: Zheng, Zhisheng, et al.
Publicado: (2025)
por: Zheng, Zhisheng, et al.
Publicado: (2025)
EchoVoices: Preserving Generational Voices and Memories for Seniors and Children
por: Xu, Haiying, et al.
Publicado: (2025)
por: Xu, Haiying, et al.
Publicado: (2025)
Evaluating Echo State Network for Parkinson's Disease Prediction using Voice Features
por: Hosseininian, Seyedeh Zahra Seyedi, et al.
Publicado: (2024)
por: Hosseininian, Seyedeh Zahra Seyedi, et al.
Publicado: (2024)
Voice Quality Dimensions as Interpretable Primitives for Speaking Style for Atypical Speech and Affect
por: Narain, Jaya, et al.
Publicado: (2025)
por: Narain, Jaya, et al.
Publicado: (2025)
Challenges in Automated Processing of Speech from Child Wearables: The Case of Voice Type Classifier
por: Kunze, Tarek, et al.
Publicado: (2025)
por: Kunze, Tarek, et al.
Publicado: (2025)
Singing Voice Conversion with Accompaniment Using Self-Supervised Representation-Based Melody Features
por: Chen, Wei, et al.
Publicado: (2025)
por: Chen, Wei, et al.
Publicado: (2025)
SiFiSinger: A High-Fidelity End-to-End Singing Voice Synthesizer based on Source-filter Model
por: Cui, Jianwei, et al.
Publicado: (2024)
por: Cui, Jianwei, et al.
Publicado: (2024)
Audio Avatar Fingerprinting: An Approach for Authorized Use of Voice Cloning in the Era of Synthetic Audio
por: Gerstner, Candice R.
Publicado: (2026)
por: Gerstner, Candice R.
Publicado: (2026)
MeanVoiceFlow: One-step Nonparallel Voice Conversion with Mean Flows
por: Kaneko, Takuhiro, et al.
Publicado: (2026)
por: Kaneko, Takuhiro, et al.
Publicado: (2026)
RVCBench: Benchmarking the Robustness of Voice Cloning Across Modern Audio Generation Models
por: Jin, Ruinan, et al.
Publicado: (2026)
por: Jin, Ruinan, et al.
Publicado: (2026)
JoyTTS: LLM-based Spoken Chatbot With Voice Cloning
por: Zhou, Fangru, et al.
Publicado: (2025)
por: Zhou, Fangru, et al.
Publicado: (2025)
RoVo: Robust Voice Protection Against Unauthorized Speech Synthesis with Embedding-Level Perturbations
por: Kim, Seungmin, et al.
Publicado: (2025)
por: Kim, Seungmin, et al.
Publicado: (2025)
Ejemplares similares
-
Multi-modal Adversarial Training for Zero-Shot Voice Cloning
por: Janiczek, John, et al.
Publicado: (2024) -
Improving Generalization for AI-Synthesized Voice Detection
por: Ren, Hainan, et al.
Publicado: (2024) -
BiSinger: Bilingual Singing Voice Synthesis
por: Zhou, Huali, et al.
Publicado: (2023) -
NaturalVoices: A Large-Scale, Spontaneous and Emotional Podcast Dataset for Voice Conversion
por: Du, Zongyang, et al.
Publicado: (2025) -
VANPY: Voice Analysis Framework
por: Koushnir, Gregory, et al.
Publicado: (2025)