Towards Authentic Movie Dubbing with Retrieve-Augmented Director-Actor Interaction Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Rui, Zhao, Yuan, Jia, Zhenqi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
StyleDubber: Towards Multi-Scale Style Learning for Movie Dubbing
por: Cong, Gaoxiang, et al.
Publicado: (2024)
por: Cong, Gaoxiang, et al.
Publicado: (2024)
Towards Expressive Video Dubbing with Multiscale Multimodal Context Interaction
por: Zhao, Yuan, et al.
Publicado: (2024)
por: Zhao, Yuan, et al.
Publicado: (2024)
Retrieval-Augmented Dialogue Knowledge Aggregation for Expressive Conversational Speech Synthesis
por: Liu, Rui, et al.
Publicado: (2025)
por: Liu, Rui, et al.
Publicado: (2025)
Intra- and Inter-modal Context Interaction Modeling for Conversational Speech Synthesis
por: Jia, Zhenqi, et al.
Publicado: (2024)
por: Jia, Zhenqi, et al.
Publicado: (2024)
Multimodal Fine-grained Context Interaction Graph Modeling for Conversational Speech Synthesis
por: Jia, Zhenqi, et al.
Publicado: (2025)
por: Jia, Zhenqi, et al.
Publicado: (2025)
MCDubber: Multimodal Context-Aware Expressive Video Dubbing
por: Zhao, Yuan, et al.
Publicado: (2024)
por: Zhao, Yuan, et al.
Publicado: (2024)
MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing
por: Zheng, Junjie, et al.
Publicado: (2025)
por: Zheng, Junjie, et al.
Publicado: (2025)
Dubbing in Practice: A Large Scale Study of Human Localization With Insights for Automatic Dubbing
por: Brannon, William, et al.
Publicado: (2022)
por: Brannon, William, et al.
Publicado: (2022)
IBSEN: Director-Actor Agent Collaboration for Controllable and Interactive Drama Script Generation
por: Han, Senyu, et al.
Publicado: (2024)
por: Han, Senyu, et al.
Publicado: (2024)
SyncVoice: Towards Video Dubbing with Vision-Augmented Pretrained TTS Model
por: Wang, Kaidi, et al.
Publicado: (2025)
por: Wang, Kaidi, et al.
Publicado: (2025)
Dub-S2ST: Textless Speech-to-Speech Translation for Seamless Dubbing
por: Choi, Jeongsoo, et al.
Publicado: (2025)
por: Choi, Jeongsoo, et al.
Publicado: (2025)
MovieCORE: COgnitive REasoning in Movies
por: Faure, Gueter Josmy, et al.
Publicado: (2025)
por: Faure, Gueter Josmy, et al.
Publicado: (2025)
Logic-Oriented Retriever Enhancement via Contrastive Learning
por: Zhang, Wenxuan, et al.
Publicado: (2026)
por: Zhang, Wenxuan, et al.
Publicado: (2026)
Towards Visually-Guided Movie Subtitle Translation for Indic Languages
por: Chintada, Tarun, et al.
Publicado: (2026)
por: Chintada, Tarun, et al.
Publicado: (2026)
Emphasis Rendering for Conversational Text-to-Speech with Multi-modal Multi-scale Context Modeling
por: Liu, Rui, et al.
Publicado: (2024)
por: Liu, Rui, et al.
Publicado: (2024)
Corrective Retrieval Augmented Generation
por: Yan, Shi-Qi, et al.
Publicado: (2024)
por: Yan, Shi-Qi, et al.
Publicado: (2024)
Curriculum Guided Reinforcement Learning for Efficient Multi Hop Retrieval Augmented Generation
por: Ji, Yuelyu, et al.
Publicado: (2025)
por: Ji, Yuelyu, et al.
Publicado: (2025)
GIP-RAG: An Evidence-Grounded Retrieval-Augmented Framework for Interpretable Gene Interaction and Pathway Impact Analysis
por: Jia, Fujian, et al.
Publicado: (2026)
por: Jia, Fujian, et al.
Publicado: (2026)
On Retrieval Augmentation and the Limitations of Language Model Training
por: Chiang, Ting-Rui, et al.
Publicado: (2023)
por: Chiang, Ting-Rui, et al.
Publicado: (2023)
Lighter And Better: Towards Flexible Context Adaptation For Retrieval Augmented Generation
por: Liu, Zheng, et al.
Publicado: (2024)
por: Liu, Zheng, et al.
Publicado: (2024)
Toward Faithful Retrieval-Augmented Generation with Sparse Autoencoders
por: Xiong, Guangzhi, et al.
Publicado: (2025)
por: Xiong, Guangzhi, et al.
Publicado: (2025)
NaviRAG: Towards Active Knowledge Navigation for Retrieval-Augmented Generation
por: Dai, Jihao, et al.
Publicado: (2026)
por: Dai, Jihao, et al.
Publicado: (2026)
RQ-RAG: Learning to Refine Queries for Retrieval Augmented Generation
por: Chan, Chi-Min, et al.
Publicado: (2024)
por: Chan, Chi-Min, et al.
Publicado: (2024)
Language-Coupled Reinforcement Learning for Multilingual Retrieval-Augmented Generation
por: Qi, Rui, et al.
Publicado: (2026)
por: Qi, Rui, et al.
Publicado: (2026)
MovieSum: An Abstractive Summarization Dataset for Movie Screenplays
por: Saxena, Rohit, et al.
Publicado: (2024)
por: Saxena, Rohit, et al.
Publicado: (2024)
Toward Robust RALMs: Revealing the Impact of Imperfect Retrieval on Retrieval-Augmented Language Models
por: Park, Seong-Il, et al.
Publicado: (2024)
por: Park, Seong-Il, et al.
Publicado: (2024)
Fine-grained Video Dubbing Duration Alignment with Segment Supervised Preference Optimization
por: Cui, Chaoqun, et al.
Publicado: (2025)
por: Cui, Chaoqun, et al.
Publicado: (2025)
Detecting Hallucinations in Authentic LLM-Human Interactions
por: Ren, Yujie, et al.
Publicado: (2025)
por: Ren, Yujie, et al.
Publicado: (2025)
Steering Over-refusals Towards Safety in Retrieval Augmented Generation
por: Maskey, Utsav, et al.
Publicado: (2025)
por: Maskey, Utsav, et al.
Publicado: (2025)
Bridging Relevance and Reasoning: Rationale Distillation in Retrieval-Augmented Generation
por: Jia, Pengyue, et al.
Publicado: (2024)
por: Jia, Pengyue, et al.
Publicado: (2024)
Learning to Extract Rational Evidence via Reinforcement Learning for Retrieval-Augmented Generation
por: Zhao, Xinping, et al.
Publicado: (2025)
por: Zhao, Xinping, et al.
Publicado: (2025)
DocCHA: Towards LLM-Augmented Interactive Online diagnosis System
por: Liu, Xinyi, et al.
Publicado: (2025)
por: Liu, Xinyi, et al.
Publicado: (2025)
Enhancing Retrieval-Augmented LMs with a Two-stage Consistency Learning Compressor
por: Xu, Chuankai, et al.
Publicado: (2024)
por: Xu, Chuankai, et al.
Publicado: (2024)
An Empirical Study of Retrieval Augmented Generation with Chain-of-Thought
por: Zhao, Yuetong, et al.
Publicado: (2024)
por: Zhao, Yuetong, et al.
Publicado: (2024)
Movie101v2: Improved Movie Narration Benchmark
por: Yue, Zihao, et al.
Publicado: (2024)
por: Yue, Zihao, et al.
Publicado: (2024)
Toward Structured Knowledge Reasoning: Contrastive Retrieval-Augmented Generation on Experience
por: Gu, Jiawei, et al.
Publicado: (2025)
por: Gu, Jiawei, et al.
Publicado: (2025)
Zero-RAG: Towards Retrieval-Augmented Generation with Zero Redundant Knowledge
por: Luo, Qi, et al.
Publicado: (2025)
por: Luo, Qi, et al.
Publicado: (2025)
Towards Comprehensive Vietnamese Retrieval-Augmented Generation and Large Language Models
por: Duc, Nguyen Quang, et al.
Publicado: (2024)
por: Duc, Nguyen Quang, et al.
Publicado: (2024)
MSRS: Evaluating Multi-Source Retrieval-Augmented Generation
por: Phanse, Rohan, et al.
Publicado: (2025)
por: Phanse, Rohan, et al.
Publicado: (2025)
Towards Trustworthy Retrieval Augmented Generation for Large Language Models: A Survey
por: Ni, Bo, et al.
Publicado: (2025)
por: Ni, Bo, et al.
Publicado: (2025)
Ejemplares similares
-
StyleDubber: Towards Multi-Scale Style Learning for Movie Dubbing
por: Cong, Gaoxiang, et al.
Publicado: (2024) -
Towards Expressive Video Dubbing with Multiscale Multimodal Context Interaction
por: Zhao, Yuan, et al.
Publicado: (2024) -
Retrieval-Augmented Dialogue Knowledge Aggregation for Expressive Conversational Speech Synthesis
por: Liu, Rui, et al.
Publicado: (2025) -
Intra- and Inter-modal Context Interaction Modeling for Conversational Speech Synthesis
por: Jia, Zhenqi, et al.
Publicado: (2024) -
Multimodal Fine-grained Context Interaction Graph Modeling for Conversational Speech Synthesis
por: Jia, Zhenqi, et al.
Publicado: (2025)