SoulX-Podcast: Towards Realistic Long-form Podcasts with Dialectal and Paralinguistic Diversity
Fuente:
arXiv
Saved in:
| Main Authors: | Xie, Hanke, Lin, Haopeng, Cao, Wenxiao, Guo, Dake, Tian, Wenjie, Wu, Jun, Wen, Hanlin, Shang, Ruixuan, Liu, Hongmei, Jiang, Zhiqi, Jiang, Yuepeng, Chen, Wenxi, Yan, Ruiqi, Qian, Jiale, Yan, Yichao, Yin, Shunshun, Tao, Ming, Chen, Xie, Xie, Lei, Wang, Xinsheng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SoulX-Singer: Towards High-Quality Zero-Shot Singing Voice Synthesis
by: Qian, Jiale, et al.
Published: (2026)
by: Qian, Jiale, et al.
Published: (2026)
SoulX-Duplug: Plug-and-Play Streaming State Prediction Module for Realtime Full-Duplex Speech Conversation
by: Yan, Ruiqi, et al.
Published: (2026)
by: Yan, Ruiqi, et al.
Published: (2026)
SoulX-Transcriber: A Robust End-to-End Framework for Multi-Speaker Speech Transcription
by: Dai, Yuhang, et al.
Published: (2026)
by: Dai, Yuhang, et al.
Published: (2026)
SoulX-LiveAct: Towards Hour-Scale Real-Time Human Animation with Neighbor Forcing and ConvKV Memory
by: Zhen, Dingcheng, et al.
Published: (2026)
by: Zhen, Dingcheng, et al.
Published: (2026)
Joint Learning Global-Local Speaker Classification to Enhance End-to-End Speaker Diarization and Recognition
by: Dai, Yuhang, et al.
Published: (2026)
by: Dai, Yuhang, et al.
Published: (2026)
SAC: Neural Speech Codec with Semantic-Acoustic Dual-Stream Quantization
by: Chen, Wenxi, et al.
Published: (2025)
by: Chen, Wenxi, et al.
Published: (2025)
Podcast Outcasts: Understanding Rumble's Podcast Dynamics
by: Balci, Utkucan, et al.
Published: (2024)
by: Balci, Utkucan, et al.
Published: (2024)
PodAgent: A Comprehensive Framework for Podcast Generation
by: Xiao, Yujia, et al.
Published: (2025)
by: Xiao, Yujia, et al.
Published: (2025)
SoulX-FlashHead: Oracle-guided Generation of Infinite Real-time Streaming Talking Heads
by: Yu, Tan, et al.
Published: (2026)
by: Yu, Tan, et al.
Published: (2026)
Mapping the Podcast Ecosystem with the Structured Podcast Research Corpus
by: Litterer, Benjamin, et al.
Published: (2024)
by: Litterer, Benjamin, et al.
Published: (2024)
FireRedTTS-2: Towards Long Conversational Speech Generation for Podcast and Chatbot
by: Xie, Kun, et al.
Published: (2025)
by: Xie, Kun, et al.
Published: (2025)
Chemie‐Podcast
Published: (2024)
Published: (2024)
Chemie‐Podcast
Published: (2024)
Published: (2024)
Chemie‐Podcast
Published: (2024)
Published: (2024)
Chemie‐Podcast
Published: (2024)
Published: (2024)
Chemie‐Podcast
Published: (2024)
Published: (2024)
Chemie‐Podcast
Published: (2025)
Published: (2025)
Chemie‐Podcast
Published: (2025)
Published: (2025)
Chemie‐Podcast
Published: (2025)
Published: (2025)
AAAP Podcast
Published: (2024)
Published: (2024)
Chemie‐Podcast
Published: (2025)
Published: (2025)
SoulX-FlashTalk: Real-Time Infinite Streaming of Audio-Driven Avatars via Self-Correcting Bidirectional Distillation
by: Shen, Le, et al.
Published: (2025)
by: Shen, Le, et al.
Published: (2025)
Podcast Transcripts (BOW)
by: Verreyen, Loren
Published: (2025)
by: Verreyen, Loren
Published: (2025)
Podcasting as an Intimate Medium
by: Euritt, Alyn
Published: (2025)
by: Euritt, Alyn
Published: (2025)
The MSP-Podcast Corpus
by: Busso, Carlos, et al.
Published: (2025)
by: Busso, Carlos, et al.
Published: (2025)
AAAP Podcast EP
Published: (2024)
Published: (2024)
Creating Communities with Podcasting
by: Jowitt, Angela L.
Published: (2008)
by: Jowitt, Angela L.
Published: (2008)
Launching into the Podcast/Vodcast Universe
by: Sampson, Jo Ann
Published: (2006)
by: Sampson, Jo Ann
Published: (2006)
Podcast 1 2 3
by: Griffey, Jason
Published: (2007)
by: Griffey, Jason
Published: (2007)
Is Podcasting Your Best Alternative?
Published: (2024)
Published: (2024)
MINT-Bench: A Comprehensive Multilingual Benchmark for Instruction-Following Text-to-Speech
by: Chen, Huakang, et al.
Published: (2026)
by: Chen, Huakang, et al.
Published: (2026)
The Routledge Companion to Radio and Podcast Studies
Published: (2022)
Published: (2022)
Podcasting and RSS: The Current State of Affairs
by: Clark, John R.
Published: (2007)
by: Clark, John R.
Published: (2007)
Listen Up! Podcasting for Schools and Libraries
by: Braun, Linda W.
Published: (2007)
by: Braun, Linda W.
Published: (2007)
Podcasts in Support of Experiential Field Learning
by: Jarvis, Claire, et al.
Published: (2010)
by: Jarvis, Claire, et al.
Published: (2010)
Student Satisfaction with Educational Podcasts Questionnaire
by: Rafael Alarcón
Published: (2017)
by: Rafael Alarcón
Published: (2017)
Rhapsody: A Dataset for Highlight Detection in Podcasts
by: Park, Younghan, et al.
Published: (2025)
by: Park, Younghan, et al.
Published: (2025)
Exploring the Use of Podcasting in the Medical Radiation Sciences
by: Elio Arruzza, et al.
Published: (2026)
by: Elio Arruzza, et al.
Published: (2026)
Annotation Tool and Dataset for Fact-Checking Podcasts
by: Setty, Vinay, et al.
Published: (2025)
by: Setty, Vinay, et al.
Published: (2025)
Podcasting as an Educational Building Block in Academic Libraries
by: Ralph, Jaya, et al.
Published: (2007)
by: Ralph, Jaya, et al.
Published: (2007)
Similar Items
-
SoulX-Singer: Towards High-Quality Zero-Shot Singing Voice Synthesis
by: Qian, Jiale, et al.
Published: (2026) -
SoulX-Duplug: Plug-and-Play Streaming State Prediction Module for Realtime Full-Duplex Speech Conversation
by: Yan, Ruiqi, et al.
Published: (2026) -
SoulX-Transcriber: A Robust End-to-End Framework for Multi-Speaker Speech Transcription
by: Dai, Yuhang, et al.
Published: (2026) -
SoulX-LiveAct: Towards Hour-Scale Real-Time Human Animation with Neighbor Forcing and ConvKV Memory
by: Zhen, Dingcheng, et al.
Published: (2026) -
Joint Learning Global-Local Speaker Classification to Enhance End-to-End Speaker Diarization and Recognition
by: Dai, Yuhang, et al.
Published: (2026)