OpenVoice: Versatile Instant Voice Cloning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qin, Zengyi, Zhao, Wenliang, Yu, Xumin, Sun, Xin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multi-modal Adversarial Training for Zero-Shot Voice Cloning
von: Janiczek, John, et al.
Veröffentlicht: (2024)
von: Janiczek, John, et al.
Veröffentlicht: (2024)
Improving Generalization for AI-Synthesized Voice Detection
von: Ren, Hainan, et al.
Veröffentlicht: (2024)
von: Ren, Hainan, et al.
Veröffentlicht: (2024)
BiSinger: Bilingual Singing Voice Synthesis
von: Zhou, Huali, et al.
Veröffentlicht: (2023)
von: Zhou, Huali, et al.
Veröffentlicht: (2023)
NaturalVoices: A Large-Scale, Spontaneous and Emotional Podcast Dataset for Voice Conversion
von: Du, Zongyang, et al.
Veröffentlicht: (2025)
von: Du, Zongyang, et al.
Veröffentlicht: (2025)
VANPY: Voice Analysis Framework
von: Koushnir, Gregory, et al.
Veröffentlicht: (2025)
von: Koushnir, Gregory, et al.
Veröffentlicht: (2025)
Systematic FAIRness Assessment of Open Voice Biomarker Datasets for Mental Health and Neurodegenerative Diseases
von: Mahapatra, Ishaan, et al.
Veröffentlicht: (2025)
von: Mahapatra, Ishaan, et al.
Veröffentlicht: (2025)
Compact Neural TTS Voices for Accessibility
von: Jain, Kunal, et al.
Veröffentlicht: (2025)
von: Jain, Kunal, et al.
Veröffentlicht: (2025)
Speech to Speech Synthesis for Voice Impersonation
von: Johnson, Bjorn, et al.
Veröffentlicht: (2026)
von: Johnson, Bjorn, et al.
Veröffentlicht: (2026)
Discrete Optimal Transport and Voice Conversion
von: Selitskiy, Anton, et al.
Veröffentlicht: (2025)
von: Selitskiy, Anton, et al.
Veröffentlicht: (2025)
Pronunciation Deviation Analysis Through Voice Cloning and Acoustic Comparison
von: Valdivia, Andrew, et al.
Veröffentlicht: (2025)
von: Valdivia, Andrew, et al.
Veröffentlicht: (2025)
Zero-shot Voice Conversion with Diffusion Transformers
von: Liu, Songting
Veröffentlicht: (2024)
von: Liu, Songting
Veröffentlicht: (2024)
Optimal Transport Maps are Good Voice Converters
von: Asadulaev, Arip, et al.
Veröffentlicht: (2024)
von: Asadulaev, Arip, et al.
Veröffentlicht: (2024)
SF-Speech: Straightened Flow for Zero-Shot Voice Clone
von: Li, Xuyuan, et al.
Veröffentlicht: (2024)
von: Li, Xuyuan, et al.
Veröffentlicht: (2024)
SEF-MK: Speaker-Embedding-Free Voice Anonymization through Multi-k-means Quantization
von: Tang, Beilong, et al.
Veröffentlicht: (2025)
von: Tang, Beilong, et al.
Veröffentlicht: (2025)
ConSinger: Efficient High-Fidelity Singing Voice Generation with Minimal Steps
von: Song, Yulin, et al.
Veröffentlicht: (2024)
von: Song, Yulin, et al.
Veröffentlicht: (2024)
A Concept-based approach to Voice Disorder Detection
von: Ghia, Davide, et al.
Veröffentlicht: (2025)
von: Ghia, Davide, et al.
Veröffentlicht: (2025)
Tessellated Linear Model for Age Prediction from Voice
von: Alharthi, Dareen, et al.
Veröffentlicht: (2025)
von: Alharthi, Dareen, et al.
Veröffentlicht: (2025)
Voice Cloning: Comprehensive Survey
von: Azzuni, Hussam, et al.
Veröffentlicht: (2025)
von: Azzuni, Hussam, et al.
Veröffentlicht: (2025)
DiffAnon: Diffusion-based Prosody Control for Voice Anonymization
von: Ulgen, Ismail Rasim, et al.
Veröffentlicht: (2026)
von: Ulgen, Ismail Rasim, et al.
Veröffentlicht: (2026)
Voice Signal Processing for Machine Learning. The Case of Speaker Isolation
von: Ganchev, Radan
Veröffentlicht: (2024)
von: Ganchev, Radan
Veröffentlicht: (2024)
On the Generation and Removal of Speaker Adversarial Perturbation for Voice-Privacy Protection
von: Guo, Chenyang, et al.
Veröffentlicht: (2024)
von: Guo, Chenyang, et al.
Veröffentlicht: (2024)
StreamVC: Real-Time Low-Latency Voice Conversion
von: Yang, Yang, et al.
Veröffentlicht: (2024)
von: Yang, Yang, et al.
Veröffentlicht: (2024)
Discrete Unit based Masking for Improving Disentanglement in Voice Conversion
von: Lee, Philip H., et al.
Veröffentlicht: (2024)
von: Lee, Philip H., et al.
Veröffentlicht: (2024)
Towards Robust Assessment of Pathological Voices via Combined Low-Level Descriptors and Foundation Model Representations
von: Ariyanti, Whenty, et al.
Veröffentlicht: (2025)
von: Ariyanti, Whenty, et al.
Veröffentlicht: (2025)
SVSNet+: Enhancing Speaker Voice Similarity Assessment Models with Representations from Speech Foundation Models
von: Yin, Chun, et al.
Veröffentlicht: (2024)
von: Yin, Chun, et al.
Veröffentlicht: (2024)
De-AntiFake: Rethinking the Protective Perturbations Against Voice Cloning Attacks
von: Fan, Wei, et al.
Veröffentlicht: (2025)
von: Fan, Wei, et al.
Veröffentlicht: (2025)
Phoneme Hallucinator: One-shot Voice Conversion via Set Expansion
von: Shan, Siyuan, et al.
Veröffentlicht: (2023)
von: Shan, Siyuan, et al.
Veröffentlicht: (2023)
Voice Conversion with Diverse Intonation using Conditional Variational Auto-Encoder
von: Suh, Soobin, et al.
Veröffentlicht: (2025)
von: Suh, Soobin, et al.
Veröffentlicht: (2025)
VoiceCraft-X: Unifying Multilingual, Voice-Cloning Speech Synthesis and Speech Editing
von: Zheng, Zhisheng, et al.
Veröffentlicht: (2025)
von: Zheng, Zhisheng, et al.
Veröffentlicht: (2025)
EchoVoices: Preserving Generational Voices and Memories for Seniors and Children
von: Xu, Haiying, et al.
Veröffentlicht: (2025)
von: Xu, Haiying, et al.
Veröffentlicht: (2025)
Evaluating Echo State Network for Parkinson's Disease Prediction using Voice Features
von: Hosseininian, Seyedeh Zahra Seyedi, et al.
Veröffentlicht: (2024)
von: Hosseininian, Seyedeh Zahra Seyedi, et al.
Veröffentlicht: (2024)
Voice Quality Dimensions as Interpretable Primitives for Speaking Style for Atypical Speech and Affect
von: Narain, Jaya, et al.
Veröffentlicht: (2025)
von: Narain, Jaya, et al.
Veröffentlicht: (2025)
Challenges in Automated Processing of Speech from Child Wearables: The Case of Voice Type Classifier
von: Kunze, Tarek, et al.
Veröffentlicht: (2025)
von: Kunze, Tarek, et al.
Veröffentlicht: (2025)
Singing Voice Conversion with Accompaniment Using Self-Supervised Representation-Based Melody Features
von: Chen, Wei, et al.
Veröffentlicht: (2025)
von: Chen, Wei, et al.
Veröffentlicht: (2025)
SiFiSinger: A High-Fidelity End-to-End Singing Voice Synthesizer based on Source-filter Model
von: Cui, Jianwei, et al.
Veröffentlicht: (2024)
von: Cui, Jianwei, et al.
Veröffentlicht: (2024)
Audio Avatar Fingerprinting: An Approach for Authorized Use of Voice Cloning in the Era of Synthetic Audio
von: Gerstner, Candice R.
Veröffentlicht: (2026)
von: Gerstner, Candice R.
Veröffentlicht: (2026)
MeanVoiceFlow: One-step Nonparallel Voice Conversion with Mean Flows
von: Kaneko, Takuhiro, et al.
Veröffentlicht: (2026)
von: Kaneko, Takuhiro, et al.
Veröffentlicht: (2026)
RVCBench: Benchmarking the Robustness of Voice Cloning Across Modern Audio Generation Models
von: Jin, Ruinan, et al.
Veröffentlicht: (2026)
von: Jin, Ruinan, et al.
Veröffentlicht: (2026)
JoyTTS: LLM-based Spoken Chatbot With Voice Cloning
von: Zhou, Fangru, et al.
Veröffentlicht: (2025)
von: Zhou, Fangru, et al.
Veröffentlicht: (2025)
RoVo: Robust Voice Protection Against Unauthorized Speech Synthesis with Embedding-Level Perturbations
von: Kim, Seungmin, et al.
Veröffentlicht: (2025)
von: Kim, Seungmin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Multi-modal Adversarial Training for Zero-Shot Voice Cloning
von: Janiczek, John, et al.
Veröffentlicht: (2024) -
Improving Generalization for AI-Synthesized Voice Detection
von: Ren, Hainan, et al.
Veröffentlicht: (2024) -
BiSinger: Bilingual Singing Voice Synthesis
von: Zhou, Huali, et al.
Veröffentlicht: (2023) -
NaturalVoices: A Large-Scale, Spontaneous and Emotional Podcast Dataset for Voice Conversion
von: Du, Zongyang, et al.
Veröffentlicht: (2025) -
VANPY: Voice Analysis Framework
von: Koushnir, Gregory, et al.
Veröffentlicht: (2025)