MoCoTalk: Multi-Conditional Diffusion with Adaptive Router for Controllable Talking Head Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ye, Xinyan, Deng, Jiankang, Edalat, Abbas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MoDA: Multi-modal Diffusion Architecture for Talking Head Generation
von: Li, Xinyang, et al.
Veröffentlicht: (2025)
von: Li, Xinyang, et al.
Veröffentlicht: (2025)
DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models
von: Ma, Yifeng, et al.
Veröffentlicht: (2023)
von: Ma, Yifeng, et al.
Veröffentlicht: (2023)
ConsistTalk: Intensity Controllable Temporally Consistent Talking Head Generation with Diffusion Noise Search
von: Liu, Zhenjie, et al.
Veröffentlicht: (2025)
von: Liu, Zhenjie, et al.
Veröffentlicht: (2025)
StyleTalk++: A Unified Framework for Controlling the Speaking Styles of Talking Heads
von: Wang, Suzhen, et al.
Veröffentlicht: (2024)
von: Wang, Suzhen, et al.
Veröffentlicht: (2024)
EMOdiffhead: Continuously Emotional Control in Talking Head Generation via Diffusion
von: Zhang, Jian, et al.
Veröffentlicht: (2024)
von: Zhang, Jian, et al.
Veröffentlicht: (2024)
TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles
von: Ma, Yifeng, et al.
Veröffentlicht: (2023)
von: Ma, Yifeng, et al.
Veröffentlicht: (2023)
FixTalk: Taming Identity Leakage for High-Quality Talking Head Generation in Extreme Cases
von: Tan, Shuai, et al.
Veröffentlicht: (2025)
von: Tan, Shuai, et al.
Veröffentlicht: (2025)
FD2Talk: Towards Generalized Talking Head Generation with Facial Decoupled Diffusion Model
von: Yao, Ziyu, et al.
Veröffentlicht: (2024)
von: Yao, Ziyu, et al.
Veröffentlicht: (2024)
MoDiTalker: Motion-Disentangled Diffusion Model for High-Fidelity Talking Head Generation
von: Kim, Seyeon, et al.
Veröffentlicht: (2024)
von: Kim, Seyeon, et al.
Veröffentlicht: (2024)
SyncTalk: The Devil is in the Synchronization for Talking Head Synthesis
von: Peng, Ziqiao, et al.
Veröffentlicht: (2023)
von: Peng, Ziqiao, et al.
Veröffentlicht: (2023)
TalkingHeadBench: A Multi-Modal Benchmark & Analysis of Talking-Head DeepFake Detection
von: Xiong, Xinqi, et al.
Veröffentlicht: (2025)
von: Xiong, Xinqi, et al.
Veröffentlicht: (2025)
EmoDiffTalk:Emotion-aware Diffusion for Editable 3D Gaussian Talking Head
von: Liu, Chang, et al.
Veröffentlicht: (2025)
von: Liu, Chang, et al.
Veröffentlicht: (2025)
MoDiT: Learning Highly Consistent 3D Motion Coefficients with Diffusion Transformer for Talking Head Generation
von: Wang, Yucheng, et al.
Veröffentlicht: (2025)
von: Wang, Yucheng, et al.
Veröffentlicht: (2025)
Dimitra: Audio-driven Diffusion model for Expressive Talking Head Generation
von: Chopin, Baptiste, et al.
Veröffentlicht: (2025)
von: Chopin, Baptiste, et al.
Veröffentlicht: (2025)
MultiTalk: Enhancing 3D Talking Head Generation Across Languages with Multilingual Video Dataset
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2024)
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2024)
EmotiveTalk: Expressive Talking Head Generation through Audio Information Decoupling and Emotional Video Diffusion
von: Wang, Haotian, et al.
Veröffentlicht: (2024)
von: Wang, Haotian, et al.
Veröffentlicht: (2024)
Taming Transformer for Emotion-Controllable Talking Face Generation
von: Zhang, Ziqi, et al.
Veröffentlicht: (2025)
von: Zhang, Ziqi, et al.
Veröffentlicht: (2025)
FreeTalk: Emotional Topology-Free 3D Talking Heads
von: Nocentini, Federico, et al.
Veröffentlicht: (2026)
von: Nocentini, Federico, et al.
Veröffentlicht: (2026)
EmoTalkingGaussian: Continuous Emotion-conditioned Talking Head Synthesis
von: Cha, Junuk, et al.
Veröffentlicht: (2025)
von: Cha, Junuk, et al.
Veröffentlicht: (2025)
ScanTalk: 3D Talking Heads from Unregistered Scans
von: Nocentini, Federico, et al.
Veröffentlicht: (2024)
von: Nocentini, Federico, et al.
Veröffentlicht: (2024)
Test-Time Self-Adaptive Conditioning for Stable Audio-Driven Talking-Head Generation
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2026)
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2026)
EDTalk++: Full Disentanglement for Controllable Talking Head Synthesis
von: Tan, Shuai, et al.
Veröffentlicht: (2025)
von: Tan, Shuai, et al.
Veröffentlicht: (2025)
EDTalk: Efficient Disentanglement for Emotional Talking Head Synthesis
von: Tan, Shuai, et al.
Veröffentlicht: (2024)
von: Tan, Shuai, et al.
Veröffentlicht: (2024)
TalkVid: A Large-Scale Diversified Dataset for Audio-Driven Talking Head Synthesis
von: Chen, Shunian, et al.
Veröffentlicht: (2025)
von: Chen, Shunian, et al.
Veröffentlicht: (2025)
GaussianHeadTalk: Wobble-Free 3D Talking Heads with Audio Driven Gaussian Splatting
von: Agarwal, Madhav, et al.
Veröffentlicht: (2025)
von: Agarwal, Madhav, et al.
Veröffentlicht: (2025)
Splat-Portrait: Generalizing Talking Heads with Gaussian Splatting
von: Shi, Tong, et al.
Veröffentlicht: (2026)
von: Shi, Tong, et al.
Veröffentlicht: (2026)
THEval. Evaluation Framework for Talking Head Video Generation
von: Quignon, Nabyl, et al.
Veröffentlicht: (2025)
von: Quignon, Nabyl, et al.
Veröffentlicht: (2025)
AUHead: Realistic Emotional Talking Head Generation via Action Units Control
von: Lyu, Jiayi, et al.
Veröffentlicht: (2026)
von: Lyu, Jiayi, et al.
Veröffentlicht: (2026)
Toward Fine-Grained Facial Control in 3D Talking Head Generation
von: Xie, Shaoyang, et al.
Veröffentlicht: (2026)
von: Xie, Shaoyang, et al.
Veröffentlicht: (2026)
RSATalker: Realistic Socially-Aware Talking Head Generation for Multi-Turn Conversation
von: Chen, Peng, et al.
Veröffentlicht: (2026)
von: Chen, Peng, et al.
Veröffentlicht: (2026)
OT-Talk: Animating 3D Talking Head with Optimal Transportation
von: Wang, Xinmu, et al.
Veröffentlicht: (2025)
von: Wang, Xinmu, et al.
Veröffentlicht: (2025)
Jump Cut Smoothing for Talking Heads
von: Wang, Xiaojuan, et al.
Veröffentlicht: (2024)
von: Wang, Xiaojuan, et al.
Veröffentlicht: (2024)
Audio-visual Controlled Video Diffusion with Masked Selective State Spaces Modeling for Natural Talking Head Generation
von: Hong, Fa-Ting, et al.
Veröffentlicht: (2025)
von: Hong, Fa-Ting, et al.
Veröffentlicht: (2025)
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router
von: Huang, Yubo, et al.
Veröffentlicht: (2025)
von: Huang, Yubo, et al.
Veröffentlicht: (2025)
JambaTalk: Speech-Driven 3D Talking Head Generation Based on Hybrid Transformer-Mamba Model
von: Jafari, Farzaneh, et al.
Veröffentlicht: (2024)
von: Jafari, Farzaneh, et al.
Veröffentlicht: (2024)
SVP: Style-Enhanced Vivid Portrait Talking Head Diffusion Model
von: Tan, Weipeng, et al.
Veröffentlicht: (2024)
von: Tan, Weipeng, et al.
Veröffentlicht: (2024)
Talking Head Generation via AU-Guided Landmark Prediction
von: Chang, Shao-Yu, et al.
Veröffentlicht: (2025)
von: Chang, Shao-Yu, et al.
Veröffentlicht: (2025)
Style2Talker: High-Resolution Talking Head Generation with Emotion Style and Art Style
von: Tan, Shuai, et al.
Veröffentlicht: (2024)
von: Tan, Shuai, et al.
Veröffentlicht: (2024)
Synergizing Motion and Appearance: Multi-Scale Compensatory Codebooks for Talking Head Video Generation
von: Zhao, Shuling, et al.
Veröffentlicht: (2024)
von: Zhao, Shuling, et al.
Veröffentlicht: (2024)
IF-MDM: Implicit Face Motion Diffusion Model for High-Fidelity Realtime Talking Head Generation
von: Yang, Sejong, et al.
Veröffentlicht: (2024)
von: Yang, Sejong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MoDA: Multi-modal Diffusion Architecture for Talking Head Generation
von: Li, Xinyang, et al.
Veröffentlicht: (2025) -
DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models
von: Ma, Yifeng, et al.
Veröffentlicht: (2023) -
ConsistTalk: Intensity Controllable Temporally Consistent Talking Head Generation with Diffusion Noise Search
von: Liu, Zhenjie, et al.
Veröffentlicht: (2025) -
StyleTalk++: A Unified Framework for Controlling the Speaking Styles of Talking Heads
von: Wang, Suzhen, et al.
Veröffentlicht: (2024) -
EMOdiffhead: Continuously Emotional Control in Talking Head Generation via Diffusion
von: Zhang, Jian, et al.
Veröffentlicht: (2024)