Voice Cloning: Comprehensive Survey

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Azzuni, Hussam, Saddik, Abdulmotaleb El
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866916716073713664
author Azzuni, Hussam
Saddik, Abdulmotaleb El
author_facet Azzuni, Hussam
Saddik, Abdulmotaleb El
contents Voice Cloning has rapidly advanced in today's digital world, with many researchers and corporations working to improve these algorithms for various applications. This article aims to establish a standardized terminology for voice cloning and explore its different variations. It will cover speaker adaptation as the fundamental concept and then delve deeper into topics such as few-shot, zero-shot, and multilingual TTS within that context. Finally, we will explore the evaluation metrics commonly used in voice cloning research and related datasets. This survey compiles the available voice cloning algorithms to encourage research toward its generation and detection to limit its misuse.
format Preprint
id arxiv_https___arxiv_org_abs_2505_00579
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Voice Cloning: Comprehensive Survey
Azzuni, Hussam
Saddik, Abdulmotaleb El
Sound
Artificial Intelligence
Audio and Speech Processing
Voice Cloning has rapidly advanced in today's digital world, with many researchers and corporations working to improve these algorithms for various applications. This article aims to establish a standardized terminology for voice cloning and explore its different variations. It will cover speaker adaptation as the fundamental concept and then delve deeper into topics such as few-shot, zero-shot, and multilingual TTS within that context. Finally, we will explore the evaluation metrics commonly used in voice cloning research and related datasets. This survey compiles the available voice cloning algorithms to encourage research toward its generation and detection to limit its misuse.
title Voice Cloning: Comprehensive Survey
topic Sound
Artificial Intelligence
Audio and Speech Processing
url https://arxiv.org/abs/2505.00579