Saved in:
| Main Authors: | Liu, Rui, Jia, Zhenqi, Bao, Feilong, Li, Haizhou |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2501.06467 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multimodal Fine-grained Context Interaction Graph Modeling for Conversational Speech Synthesis
by: Jia, Zhenqi, et al.
Published: (2025)
by: Jia, Zhenqi, et al.
Published: (2025)
Generative Expressive Conversational Speech Synthesis
by: Liu, Rui, et al.
Published: (2024)
by: Liu, Rui, et al.
Published: (2024)
Intra- and Inter-modal Context Interaction Modeling for Conversational Speech Synthesis
by: Jia, Zhenqi, et al.
Published: (2024)
by: Jia, Zhenqi, et al.
Published: (2024)
Emphasis Rendering for Conversational Text-to-Speech with Multi-modal Multi-scale Context Modeling
by: Liu, Rui, et al.
Published: (2024)
by: Liu, Rui, et al.
Published: (2024)
Towards Authentic Movie Dubbing with Retrieve-Augmented Director-Actor Interaction Learning
by: Liu, Rui, et al.
Published: (2025)
by: Liu, Rui, et al.
Published: (2025)
MCDubber: Multimodal Context-Aware Expressive Video Dubbing
by: Zhao, Yuan, et al.
Published: (2024)
by: Zhao, Yuan, et al.
Published: (2024)
NE-PADD: Leveraging Named Entity Knowledge for Robust Partial Audio Deepfake Detection via Attention Aggregation
by: Xian, Huhong, et al.
Published: (2025)
by: Xian, Huhong, et al.
Published: (2025)
ReZG: Retrieval-Augmented Zero-Shot Counter Narrative Generation for Hate Speech
by: Jiang, Shuyu, et al.
Published: (2023)
by: Jiang, Shuyu, et al.
Published: (2023)
Expressivity and Speech Synthesis
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
EmoNews: A Spoken Dialogue System for Expressive News Conversations
by: Matsuura, Ryuki, et al.
Published: (2025)
by: Matsuura, Ryuki, et al.
Published: (2025)
Enhancing Speech-to-Speech Dialogue Modeling with End-to-End Retrieval-Augmented Generation
by: Feng, Pengchao, et al.
Published: (2025)
by: Feng, Pengchao, et al.
Published: (2025)
RAGCap-Bench: Benchmarking Capabilities of LLMs in Agentic Retrieval Augmented Generation Systems
by: Lin, Jingru, et al.
Published: (2025)
by: Lin, Jingru, et al.
Published: (2025)
Incorporating External Knowledge and Goal Guidance for LLM-based Conversational Recommender Systems
by: Li, Chuang, et al.
Published: (2024)
by: Li, Chuang, et al.
Published: (2024)
Uncovering the Potential of ChatGPT for Discourse Analysis in Dialogue: An Empirical Study
by: Fan, Yaxin, et al.
Published: (2023)
by: Fan, Yaxin, et al.
Published: (2023)
FluentEditor2: Text-based Speech Editing by Modeling Multi-Scale Acoustic and Prosody Consistency
by: Liu, Rui, et al.
Published: (2024)
by: Liu, Rui, et al.
Published: (2024)
SpeechCraft: A Fine-grained Expressive Speech Dataset with Natural Language Description
by: Jin, Zeyu, et al.
Published: (2024)
by: Jin, Zeyu, et al.
Published: (2024)
Dialogue Benchmark Generation from Knowledge Graphs with Cost-Effective Retrieval-Augmented LLMs
by: Omar, Reham, et al.
Published: (2025)
by: Omar, Reham, et al.
Published: (2025)
Learning Retrieval Augmentation for Personalized Dialogue Generation
by: Huang, Qiushi, et al.
Published: (2024)
by: Huang, Qiushi, et al.
Published: (2024)
Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
by: Liu, Zhijun, et al.
Published: (2024)
by: Liu, Zhijun, et al.
Published: (2024)
UNO-DST: Leveraging Unlabelled Data in Zero-Shot Dialogue State Tracking
by: Li, Chuang, et al.
Published: (2023)
by: Li, Chuang, et al.
Published: (2023)
MTalk-Bench: Evaluating Speech-to-Speech Models in Multi-Turn Dialogues via Arena-style and Rubrics Protocols
by: Du, Yuhao, et al.
Published: (2025)
by: Du, Yuhao, et al.
Published: (2025)
Instructive Dialogue Summarization with Query Aggregations
by: Wang, Bin, et al.
Published: (2023)
by: Wang, Bin, et al.
Published: (2023)
DialogueAgents: A Hybrid Agent-Based Speech Synthesis Framework for Multi-Party Dialogue
by: Li, Xiang, et al.
Published: (2025)
by: Li, Xiang, et al.
Published: (2025)
Emotion and Intent Joint Understanding in Multimodal Conversation: A Benchmarking Dataset
by: Liu, Rui, et al.
Published: (2024)
by: Liu, Rui, et al.
Published: (2024)
Adaptive Retrieval-Augmented Generation for Conversational Systems
by: Wang, Xi, et al.
Published: (2024)
by: Wang, Xi, et al.
Published: (2024)
Knowledge Graph-Guided Retrieval Augmented Generation
by: Zhu, Xiangrong, et al.
Published: (2025)
by: Zhu, Xiangrong, et al.
Published: (2025)
Graph Retrieval-Augmented Generation: A Survey
by: Peng, Boci, et al.
Published: (2024)
by: Peng, Boci, et al.
Published: (2024)
CADGE: Context-Aware Dialogue Generation Enhanced with Graph-Structured Knowledge Aggregation
by: Zhang, Hongbo, et al.
Published: (2023)
by: Zhang, Hongbo, et al.
Published: (2023)
SAGE: Scalable Automated Robustness Augmentation for LLM Knowledge Evaluation
by: Li, Xiaoyuan, et al.
Published: (2026)
by: Li, Xiaoyuan, et al.
Published: (2026)
Unsupervised Mutual Learning of Discourse Parsing and Topic Segmentation in Dialogue
by: Xu, Jiahui, et al.
Published: (2024)
by: Xu, Jiahui, et al.
Published: (2024)
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style
by: Kang, Wonjune, et al.
Published: (2025)
by: Kang, Wonjune, et al.
Published: (2025)
EmoShift: Lightweight Activation Steering for Enhanced Emotion-Aware Speech Synthesis
by: Zhou, Li, et al.
Published: (2026)
by: Zhou, Li, et al.
Published: (2026)
SeaKR: Self-aware Knowledge Retrieval for Adaptive Retrieval Augmented Generation
by: Yao, Zijun, et al.
Published: (2024)
by: Yao, Zijun, et al.
Published: (2024)
Effective and Efficient Conversation Retrieval for Dialogue State Tracking with Implicit Text Summaries
by: Lee, Seanie, et al.
Published: (2024)
by: Lee, Seanie, et al.
Published: (2024)
Enhancing Personality Recognition in Dialogue by Data Augmentation and Heterogeneous Conversational Graph Networks
by: Fu, Yahui, et al.
Published: (2024)
by: Fu, Yahui, et al.
Published: (2024)
GTR-Voice: Articulatory Phonetics Informed Controllable Expressive Speech Synthesis
by: Li, Zehua Kcriss, et al.
Published: (2024)
by: Li, Zehua Kcriss, et al.
Published: (2024)
Aggregated Knowledge Model: Enhancing Domain-Specific QA with Fine-Tuned and Retrieval-Augmented Generation Models
by: Liu, Fengchen, et al.
Published: (2024)
by: Liu, Fengchen, et al.
Published: (2024)
Style Mixture of Experts for Expressive Text-To-Speech Synthesis
by: Jawaid, Ahad, et al.
Published: (2024)
by: Jawaid, Ahad, et al.
Published: (2024)
Knowledge-driven Augmentation and Retrieval for Integrative Temporal Adaptation
by: Liu, Weisi, et al.
Published: (2026)
by: Liu, Weisi, et al.
Published: (2026)
PK-ICR: Persona-Knowledge Interactive Context Retrieval for Grounded Dialogue
by: Oh, Minsik, et al.
Published: (2023)
by: Oh, Minsik, et al.
Published: (2023)
Similar Items
-
Multimodal Fine-grained Context Interaction Graph Modeling for Conversational Speech Synthesis
by: Jia, Zhenqi, et al.
Published: (2025) -
Generative Expressive Conversational Speech Synthesis
by: Liu, Rui, et al.
Published: (2024) -
Intra- and Inter-modal Context Interaction Modeling for Conversational Speech Synthesis
by: Jia, Zhenqi, et al.
Published: (2024) -
Emphasis Rendering for Conversational Text-to-Speech with Multi-modal Multi-scale Context Modeling
by: Liu, Rui, et al.
Published: (2024) -
Towards Authentic Movie Dubbing with Retrieve-Augmented Director-Actor Interaction Learning
by: Liu, Rui, et al.
Published: (2025)