Retrieval-Augmented Dialogue Knowledge Aggregation for Expressive Conversational Speech Synthesis
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Rui, Jia, Zhenqi, Bao, Feilong, Li, Haizhou |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multimodal Fine-grained Context Interaction Graph Modeling for Conversational Speech Synthesis
por: Jia, Zhenqi, et al.
Publicado: (2025)
por: Jia, Zhenqi, et al.
Publicado: (2025)
Generative Expressive Conversational Speech Synthesis
por: Liu, Rui, et al.
Publicado: (2024)
por: Liu, Rui, et al.
Publicado: (2024)
Intra- and Inter-modal Context Interaction Modeling for Conversational Speech Synthesis
por: Jia, Zhenqi, et al.
Publicado: (2024)
por: Jia, Zhenqi, et al.
Publicado: (2024)
Emphasis Rendering for Conversational Text-to-Speech with Multi-modal Multi-scale Context Modeling
por: Liu, Rui, et al.
Publicado: (2024)
por: Liu, Rui, et al.
Publicado: (2024)
Towards Authentic Movie Dubbing with Retrieve-Augmented Director-Actor Interaction Learning
por: Liu, Rui, et al.
Publicado: (2025)
por: Liu, Rui, et al.
Publicado: (2025)
ReZG: Retrieval-Augmented Zero-Shot Counter Narrative Generation for Hate Speech
por: Jiang, Shuyu, et al.
Publicado: (2023)
por: Jiang, Shuyu, et al.
Publicado: (2023)
NE-PADD: Leveraging Named Entity Knowledge for Robust Partial Audio Deepfake Detection via Attention Aggregation
por: Xian, Huhong, et al.
Publicado: (2025)
por: Xian, Huhong, et al.
Publicado: (2025)
MCDubber: Multimodal Context-Aware Expressive Video Dubbing
por: Zhao, Yuan, et al.
Publicado: (2024)
por: Zhao, Yuan, et al.
Publicado: (2024)
Expressivity and Speech Synthesis
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2024)
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2024)
EmoNews: A Spoken Dialogue System for Expressive News Conversations
por: Matsuura, Ryuki, et al.
Publicado: (2025)
por: Matsuura, Ryuki, et al.
Publicado: (2025)
Enhancing Speech-to-Speech Dialogue Modeling with End-to-End Retrieval-Augmented Generation
por: Feng, Pengchao, et al.
Publicado: (2025)
por: Feng, Pengchao, et al.
Publicado: (2025)
RAGCap-Bench: Benchmarking Capabilities of LLMs in Agentic Retrieval Augmented Generation Systems
por: Lin, Jingru, et al.
Publicado: (2025)
por: Lin, Jingru, et al.
Publicado: (2025)
Incorporating External Knowledge and Goal Guidance for LLM-based Conversational Recommender Systems
por: Li, Chuang, et al.
Publicado: (2024)
por: Li, Chuang, et al.
Publicado: (2024)
SpeechCraft: A Fine-grained Expressive Speech Dataset with Natural Language Description
por: Jin, Zeyu, et al.
Publicado: (2024)
por: Jin, Zeyu, et al.
Publicado: (2024)
Dialogue Benchmark Generation from Knowledge Graphs with Cost-Effective Retrieval-Augmented LLMs
por: Omar, Reham, et al.
Publicado: (2025)
por: Omar, Reham, et al.
Publicado: (2025)
Uncovering the Potential of ChatGPT for Discourse Analysis in Dialogue: An Empirical Study
por: Fan, Yaxin, et al.
Publicado: (2023)
por: Fan, Yaxin, et al.
Publicado: (2023)
Learning Retrieval Augmentation for Personalized Dialogue Generation
por: Huang, Qiushi, et al.
Publicado: (2024)
por: Huang, Qiushi, et al.
Publicado: (2024)
Instructive Dialogue Summarization with Query Aggregations
por: Wang, Bin, et al.
Publicado: (2023)
por: Wang, Bin, et al.
Publicado: (2023)
FluentEditor2: Text-based Speech Editing by Modeling Multi-Scale Acoustic and Prosody Consistency
por: Liu, Rui, et al.
Publicado: (2024)
por: Liu, Rui, et al.
Publicado: (2024)
UNO-DST: Leveraging Unlabelled Data in Zero-Shot Dialogue State Tracking
por: Li, Chuang, et al.
Publicado: (2023)
por: Li, Chuang, et al.
Publicado: (2023)
DialogueAgents: A Hybrid Agent-Based Speech Synthesis Framework for Multi-Party Dialogue
por: Li, Xiang, et al.
Publicado: (2025)
por: Li, Xiang, et al.
Publicado: (2025)
CADGE: Context-Aware Dialogue Generation Enhanced with Graph-Structured Knowledge Aggregation
por: Zhang, Hongbo, et al.
Publicado: (2023)
por: Zhang, Hongbo, et al.
Publicado: (2023)
Adaptive Retrieval-Augmented Generation for Conversational Systems
por: Wang, Xi, et al.
Publicado: (2024)
por: Wang, Xi, et al.
Publicado: (2024)
MTalk-Bench: Evaluating Speech-to-Speech Models in Multi-Turn Dialogues via Arena-style and Rubrics Protocols
por: Du, Yuhao, et al.
Publicado: (2025)
por: Du, Yuhao, et al.
Publicado: (2025)
Knowledge Graph-Guided Retrieval Augmented Generation
por: Zhu, Xiangrong, et al.
Publicado: (2025)
por: Zhu, Xiangrong, et al.
Publicado: (2025)
SeaKR: Self-aware Knowledge Retrieval for Adaptive Retrieval Augmented Generation
por: Yao, Zijun, et al.
Publicado: (2024)
por: Yao, Zijun, et al.
Publicado: (2024)
Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
por: Liu, Zhijun, et al.
Publicado: (2024)
por: Liu, Zhijun, et al.
Publicado: (2024)
Effective and Efficient Conversation Retrieval for Dialogue State Tracking with Implicit Text Summaries
por: Lee, Seanie, et al.
Publicado: (2024)
por: Lee, Seanie, et al.
Publicado: (2024)
Enhancing Personality Recognition in Dialogue by Data Augmentation and Heterogeneous Conversational Graph Networks
por: Fu, Yahui, et al.
Publicado: (2024)
por: Fu, Yahui, et al.
Publicado: (2024)
Knowledge-driven Augmentation and Retrieval for Integrative Temporal Adaptation
por: Liu, Weisi, et al.
Publicado: (2026)
por: Liu, Weisi, et al.
Publicado: (2026)
SAGE: Scalable Automated Robustness Augmentation for LLM Knowledge Evaluation
por: Li, Xiaoyuan, et al.
Publicado: (2026)
por: Li, Xiaoyuan, et al.
Publicado: (2026)
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style
por: Kang, Wonjune, et al.
Publicado: (2025)
por: Kang, Wonjune, et al.
Publicado: (2025)
RASST: Fast Cross-modal Retrieval-Augmented Simultaneous Speech Translation
por: Luo, Jiaxuan, et al.
Publicado: (2026)
por: Luo, Jiaxuan, et al.
Publicado: (2026)
Aggregated Knowledge Model: Enhancing Domain-Specific QA with Fine-Tuned and Retrieval-Augmented Generation Models
por: Liu, Fengchen, et al.
Publicado: (2024)
por: Liu, Fengchen, et al.
Publicado: (2024)
PK-ICR: Persona-Knowledge Interactive Context Retrieval for Grounded Dialogue
por: Oh, Minsik, et al.
Publicado: (2023)
por: Oh, Minsik, et al.
Publicado: (2023)
RAD-Bench: Evaluating Large Language Models Capabilities in Retrieval Augmented Dialogues
por: Kuo, Tzu-Lin, et al.
Publicado: (2024)
por: Kuo, Tzu-Lin, et al.
Publicado: (2024)
Rethinking Evaluation in Retrieval-Augmented Personalized Dialogue: A Cognitive and Linguistic Perspective
por: Zhang, Tianyi, et al.
Publicado: (2026)
por: Zhang, Tianyi, et al.
Publicado: (2026)
CHARP: Conversation History AwaReness Probing for Knowledge-grounded Dialogue Systems
por: Ghaddar, Abbas, et al.
Publicado: (2024)
por: Ghaddar, Abbas, et al.
Publicado: (2024)
A Survey on Knowledge-Oriented Retrieval-Augmented Generation
por: Cheng, Mingyue, et al.
Publicado: (2025)
por: Cheng, Mingyue, et al.
Publicado: (2025)
The StatCan Dialogue Dataset: Retrieving Data Tables through Conversations with Genuine Intents
por: Lu, Xing Han, et al.
Publicado: (2023)
por: Lu, Xing Han, et al.
Publicado: (2023)
Ejemplares similares
-
Multimodal Fine-grained Context Interaction Graph Modeling for Conversational Speech Synthesis
por: Jia, Zhenqi, et al.
Publicado: (2025) -
Generative Expressive Conversational Speech Synthesis
por: Liu, Rui, et al.
Publicado: (2024) -
Intra- and Inter-modal Context Interaction Modeling for Conversational Speech Synthesis
por: Jia, Zhenqi, et al.
Publicado: (2024) -
Emphasis Rendering for Conversational Text-to-Speech with Multi-modal Multi-scale Context Modeling
por: Liu, Rui, et al.
Publicado: (2024) -
Towards Authentic Movie Dubbing with Retrieve-Augmented Director-Actor Interaction Learning
por: Liu, Rui, et al.
Publicado: (2025)