Gespeichert in:
| Hauptverfasser: | Chivereanu, Radu-Gabriel, Boros, Tiberiu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2512.12297 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EE-TTS: Emphatic Expressive TTS with Linguistic Information
von: Zhong, Yi, et al.
Veröffentlicht: (2023)
von: Zhong, Yi, et al.
Veröffentlicht: (2023)
A2TTS: TTS for Low Resource Indian Languages
von: Bhadoriya, Ayush Singh, et al.
Veröffentlicht: (2025)
von: Bhadoriya, Ayush Singh, et al.
Veröffentlicht: (2025)
TouchTTS: An Embarrassingly Simple TTS Framework that Everyone Can Touch
von: Song, Xingchen, et al.
Veröffentlicht: (2024)
von: Song, Xingchen, et al.
Veröffentlicht: (2024)
Unsupervised TTS Acoustic Modeling for TTS with Conditional Disentangled Sequential VAE
von: Lian, Jiachen, et al.
Veröffentlicht: (2022)
von: Lian, Jiachen, et al.
Veröffentlicht: (2022)
TTS-1 Technical Report
von: Atamanenko, Oleg, et al.
Veröffentlicht: (2025)
von: Atamanenko, Oleg, et al.
Veröffentlicht: (2025)
MOSS-TTS Technical Report
von: Gong, Yitian, et al.
Veröffentlicht: (2026)
von: Gong, Yitian, et al.
Veröffentlicht: (2026)
BnTTS: Few-Shot Speaker Adaptation in Low-Resource Setting
von: Basher, Mohammad Jahid Ibna, et al.
Veröffentlicht: (2025)
von: Basher, Mohammad Jahid Ibna, et al.
Veröffentlicht: (2025)
Assessing the Ability of Neural TTS Systems to Model Consonant-Induced F0 Perturbation
von: Yang, Tianle, et al.
Veröffentlicht: (2026)
von: Yang, Tianle, et al.
Veröffentlicht: (2026)
DiaMoE-TTS: A Unified IPA-Based Dialect TTS Framework with Mixture-of-Experts and Parameter-Efficient Zero-Shot Adaptation
von: Chen, Ziqi, et al.
Veröffentlicht: (2025)
von: Chen, Ziqi, et al.
Veröffentlicht: (2025)
Qwen3-TTS Technical Report
von: Hu, Hangrui, et al.
Veröffentlicht: (2026)
von: Hu, Hangrui, et al.
Veröffentlicht: (2026)
Tibetan-TTS:Low-Resource Tibetan Speech Synthesis with Large Model Adaptation
von: He, Jiaxu, et al.
Veröffentlicht: (2026)
von: He, Jiaxu, et al.
Veröffentlicht: (2026)
More Data, Fewer Diacritics: Scaling Arabic TTS
von: Musleh, Ahmed, et al.
Veröffentlicht: (2026)
von: Musleh, Ahmed, et al.
Veröffentlicht: (2026)
JaiTTS: A Thai Voice Cloning Model
von: Karnjanaekarin, Jullajak, et al.
Veröffentlicht: (2026)
von: Karnjanaekarin, Jullajak, et al.
Veröffentlicht: (2026)
HyperTTS: Parameter Efficient Adaptation in Text to Speech using Hypernetworks
von: Li, Yingting, et al.
Veröffentlicht: (2024)
von: Li, Yingting, et al.
Veröffentlicht: (2024)
An Initial Investigation of Language Adaptation for TTS Systems under Low-resource Scenarios
von: Gong, Cheng, et al.
Veröffentlicht: (2024)
von: Gong, Cheng, et al.
Veröffentlicht: (2024)
Preference Alignment Improves Language Model-Based TTS
von: Tian, Jinchuan, et al.
Veröffentlicht: (2024)
von: Tian, Jinchuan, et al.
Veröffentlicht: (2024)
Voice Impression Control in Zero-Shot TTS
von: Fujita, Kenichi, et al.
Veröffentlicht: (2025)
von: Fujita, Kenichi, et al.
Veröffentlicht: (2025)
VisualSpeech: Enhancing Prosody Modeling in TTS Using Video
von: Que, Shumin, et al.
Veröffentlicht: (2025)
von: Que, Shumin, et al.
Veröffentlicht: (2025)
Multi-Task Learning for Front-End Text Processing in TTS
von: Kang, Wonjune, et al.
Veröffentlicht: (2024)
von: Kang, Wonjune, et al.
Veröffentlicht: (2024)
Multi-interaction TTS toward professional recording reproduction
von: Kanagawa, Hiroki, et al.
Veröffentlicht: (2025)
von: Kanagawa, Hiroki, et al.
Veröffentlicht: (2025)
RWKVTTS: Yet another TTS based on RWKV-7
von: yueyu, Lin, et al.
Veröffentlicht: (2025)
von: yueyu, Lin, et al.
Veröffentlicht: (2025)
MunTTS: A Text-to-Speech System for Mundari
von: Gumma, Varun, et al.
Veröffentlicht: (2024)
von: Gumma, Varun, et al.
Veröffentlicht: (2024)
Llama-VITS: Enhancing TTS Synthesis with Semantic Awareness
von: Feng, Xincan, et al.
Veröffentlicht: (2024)
von: Feng, Xincan, et al.
Veröffentlicht: (2024)
VoXtream2: Full-stream TTS with dynamic speaking rate control
von: Torgashov, Nikita, et al.
Veröffentlicht: (2026)
von: Torgashov, Nikita, et al.
Veröffentlicht: (2026)
An Exhaustive Evaluation of TTS- and VC-based Data Augmentation for ASR
von: Ogun, Sewade, et al.
Veröffentlicht: (2025)
von: Ogun, Sewade, et al.
Veröffentlicht: (2025)
JoyTTS: LLM-based Spoken Chatbot With Voice Cloning
von: Zhou, Fangru, et al.
Veröffentlicht: (2025)
von: Zhou, Fangru, et al.
Veröffentlicht: (2025)
SP-MCQA: Evaluating Intelligibility of TTS Beyond the Word Level
von: Tee, Hitomi Jin Ling, et al.
Veröffentlicht: (2025)
von: Tee, Hitomi Jin Ling, et al.
Veröffentlicht: (2025)
Word-wise intonation model for cross-language TTS systems
von: A., Tomilov A., et al.
Veröffentlicht: (2024)
von: A., Tomilov A., et al.
Veröffentlicht: (2024)
An investigation of phrase break prediction in an End-to-End TTS system
von: Vadapalli, Anandaswarup
Veröffentlicht: (2023)
von: Vadapalli, Anandaswarup
Veröffentlicht: (2023)
A Language Modeling Approach to Diacritic-Free Hebrew TTS
von: Roth, Amit, et al.
Veröffentlicht: (2024)
von: Roth, Amit, et al.
Veröffentlicht: (2024)
Maximizing Data Efficiency for Cross-Lingual TTS Adaptation by Self-Supervised Representation Mixing and Embedding Initialization
von: Huang, Wei-Ping, et al.
Veröffentlicht: (2024)
von: Huang, Wei-Ping, et al.
Veröffentlicht: (2024)
Daisy-TTS: Simulating Wider Spectrum of Emotions via Prosody Embedding Decomposition
von: Chevi, Rendi, et al.
Veröffentlicht: (2024)
von: Chevi, Rendi, et al.
Veröffentlicht: (2024)
LatinX: Aligning a Multilingual TTS Model with Direct Preference Optimization
von: Chary, Luis Felipe, et al.
Veröffentlicht: (2025)
von: Chary, Luis Felipe, et al.
Veröffentlicht: (2025)
Accent Vector: Controllable Accent Manipulation for Multilingual TTS Without Accented Data
von: Lertpetchpun, Thanathai, et al.
Veröffentlicht: (2026)
von: Lertpetchpun, Thanathai, et al.
Veröffentlicht: (2026)
TTS-Transducer: End-to-End Speech Synthesis with Neural Transducer
von: Bataev, Vladimir, et al.
Veröffentlicht: (2025)
von: Bataev, Vladimir, et al.
Veröffentlicht: (2025)
The State Of TTS: A Case Study with Human Fooling Rates
von: Varadhan, Praveen Srinivasa, et al.
Veröffentlicht: (2025)
von: Varadhan, Praveen Srinivasa, et al.
Veröffentlicht: (2025)
MahaTTS: A Unified Framework for Multilingual Text-to-Speech Synthesis
von: Singh, Jaskaran, et al.
Veröffentlicht: (2025)
von: Singh, Jaskaran, et al.
Veröffentlicht: (2025)
Creating an African American-Sounding TTS: Guidelines, Technical Challenges,and Surprising Evaluations
von: Pinhanez, Claudio, et al.
Veröffentlicht: (2024)
von: Pinhanez, Claudio, et al.
Veröffentlicht: (2024)
GOAT-TTS: Expressive and Realistic Speech Generation via A Dual-Branch LLM
von: Song, Yaodong, et al.
Veröffentlicht: (2025)
von: Song, Yaodong, et al.
Veröffentlicht: (2025)
Cross-modal Consistency Guidance for Robust Emotion Control in Auto-Regressive TTS Models
von: Peng, Yizhou, et al.
Veröffentlicht: (2025)
von: Peng, Yizhou, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
EE-TTS: Emphatic Expressive TTS with Linguistic Information
von: Zhong, Yi, et al.
Veröffentlicht: (2023) -
A2TTS: TTS for Low Resource Indian Languages
von: Bhadoriya, Ayush Singh, et al.
Veröffentlicht: (2025) -
TouchTTS: An Embarrassingly Simple TTS Framework that Everyone Can Touch
von: Song, Xingchen, et al.
Veröffentlicht: (2024) -
Unsupervised TTS Acoustic Modeling for TTS with Conditional Disentangled Sequential VAE
von: Lian, Jiachen, et al.
Veröffentlicht: (2022) -
TTS-1 Technical Report
von: Atamanenko, Oleg, et al.
Veröffentlicht: (2025)