High Fidelity Text-to-Speech Via Discrete Tokens Using Token Transducer and Group Masked Language Model

Fuente: arXiv
Enregistré dans:
Détails bibliographiques
Auteurs principaux: Lee, Joun Yeop, Jeong, Myeonghun, Kim, Minchan, Lee, Ji-Hyun, Cho, Hoon-Young, Kim, Nam Soo
Format: Preprint
Publié: 2024
Sujets:
Accès en ligne:
Tags: Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!