TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ma, Yifeng, Wang, Suzhen, Ding, Yu, Ma, Bowen, Lv, Tangjie, Fan, Changjie, Hu, Zhipeng, Deng, Zhidong, Yu, Xin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
StyleTalk++: A Unified Framework for Controlling the Speaking Styles of Talking Heads
von: Wang, Suzhen, et al.
Veröffentlicht: (2024)
von: Wang, Suzhen, et al.
Veröffentlicht: (2024)
DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models
von: Ma, Yifeng, et al.
Veröffentlicht: (2023)
von: Ma, Yifeng, et al.
Veröffentlicht: (2023)
Towards a Simultaneous and Granular Identity-Expression Control in Personalized Face Generation
von: Liu, Renshuai, et al.
Veröffentlicht: (2024)
von: Liu, Renshuai, et al.
Veröffentlicht: (2024)
EmotiveTalk: Expressive Talking Head Generation through Audio Information Decoupling and Emotional Video Diffusion
von: Wang, Haotian, et al.
Veröffentlicht: (2024)
von: Wang, Haotian, et al.
Veröffentlicht: (2024)
Talking Head Generation via AU-Guided Landmark Prediction
von: Chang, Shao-Yu, et al.
Veröffentlicht: (2025)
von: Chang, Shao-Yu, et al.
Veröffentlicht: (2025)
Dimitra: Audio-driven Diffusion model for Expressive Talking Head Generation
von: Chopin, Baptiste, et al.
Veröffentlicht: (2025)
von: Chopin, Baptiste, et al.
Veröffentlicht: (2025)
SyncTalk: The Devil is in the Synchronization for Talking Head Synthesis
von: Peng, Ziqiao, et al.
Veröffentlicht: (2023)
von: Peng, Ziqiao, et al.
Veröffentlicht: (2023)
MEMO: Memory-Guided Diffusion for Expressive Talking Video Generation
von: Zheng, Longtao, et al.
Veröffentlicht: (2024)
von: Zheng, Longtao, et al.
Veröffentlicht: (2024)
PointTalk: Audio-Driven Dynamic Lip Point Cloud for 3D Gaussian-based Talking Head Synthesis
von: Xie, Yifan, et al.
Veröffentlicht: (2024)
von: Xie, Yifan, et al.
Veröffentlicht: (2024)
Style2Talker: High-Resolution Talking Head Generation with Emotion Style and Art Style
von: Tan, Shuai, et al.
Veröffentlicht: (2024)
von: Tan, Shuai, et al.
Veröffentlicht: (2024)
Think-Before-Draw: Decomposing Emotion Semantics & Fine-Grained Controllable Expressive Talking Head Generation
von: Shi, Hanlei, et al.
Veröffentlicht: (2025)
von: Shi, Hanlei, et al.
Veröffentlicht: (2025)
SyncTalk++: High-Fidelity and Efficient Synchronized Talking Heads Synthesis Using Gaussian Splatting
von: Peng, Ziqiao, et al.
Veröffentlicht: (2025)
von: Peng, Ziqiao, et al.
Veröffentlicht: (2025)
MoCoTalk: Multi-Conditional Diffusion with Adaptive Router for Controllable Talking Head Generation
von: Ye, Xinyan, et al.
Veröffentlicht: (2026)
von: Ye, Xinyan, et al.
Veröffentlicht: (2026)
StyleTalker: One-shot Style-based Audio-driven Talking Head Video Generation
von: Min, Dongchan, et al.
Veröffentlicht: (2022)
von: Min, Dongchan, et al.
Veröffentlicht: (2022)
Toward Fine-Grained Facial Control in 3D Talking Head Generation
von: Xie, Shaoyang, et al.
Veröffentlicht: (2026)
von: Xie, Shaoyang, et al.
Veröffentlicht: (2026)
OT-Talk: Animating 3D Talking Head with Optimal Transportation
von: Wang, Xinmu, et al.
Veröffentlicht: (2025)
von: Wang, Xinmu, et al.
Veröffentlicht: (2025)
Prioritized Trajectory Replay: A Replay Memory for Data-driven Reinforcement Learning
von: Liu, Jinyi, et al.
Veröffentlicht: (2023)
von: Liu, Jinyi, et al.
Veröffentlicht: (2023)
DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations
von: Peng, Ziqiao, et al.
Veröffentlicht: (2025)
von: Peng, Ziqiao, et al.
Veröffentlicht: (2025)
PoseTalk: Text-and-Audio-based Pose Control and Motion Refinement for One-Shot Talking Head Generation
von: Ling, Jun, et al.
Veröffentlicht: (2024)
von: Ling, Jun, et al.
Veröffentlicht: (2024)
Audio-driven High-resolution Seamless Talking Head Video Editing via StyleGAN
von: Su, Jiacheng, et al.
Veröffentlicht: (2024)
von: Su, Jiacheng, et al.
Veröffentlicht: (2024)
AVI-Talking: Learning Audio-Visual Instructions for Expressive 3D Talking Face Generation
von: Sun, Yasheng, et al.
Veröffentlicht: (2024)
von: Sun, Yasheng, et al.
Veröffentlicht: (2024)
SVP: Style-Enhanced Vivid Portrait Talking Head Diffusion Model
von: Tan, Weipeng, et al.
Veröffentlicht: (2024)
von: Tan, Weipeng, et al.
Veröffentlicht: (2024)
Straight Talk: Speaking the Language of Administration.
von: Miller, Rachel M.
Veröffentlicht: (1988)
von: Miller, Rachel M.
Veröffentlicht: (1988)
TalkingHeadBench: A Multi-Modal Benchmark & Analysis of Talking-Head DeepFake Detection
von: Xiong, Xinqi, et al.
Veröffentlicht: (2025)
von: Xiong, Xinqi, et al.
Veröffentlicht: (2025)
EmoDiffTalk:Emotion-aware Diffusion for Editable 3D Gaussian Talking Head
von: Liu, Chang, et al.
Veröffentlicht: (2025)
von: Liu, Chang, et al.
Veröffentlicht: (2025)
Faces that Speak: Jointly Synthesising Talking Face and Speech from Text
von: Jang, Youngjoon, et al.
Veröffentlicht: (2024)
von: Jang, Youngjoon, et al.
Veröffentlicht: (2024)
CapTalk: Text-Guided Stylization and Speech-Driven 3D Head Animation
von: Chu, Xuangeng, et al.
Veröffentlicht: (2026)
von: Chu, Xuangeng, et al.
Veröffentlicht: (2026)
MoDA: Multi-modal Diffusion Architecture for Talking Head Generation
von: Li, Xinyang, et al.
Veröffentlicht: (2025)
von: Li, Xinyang, et al.
Veröffentlicht: (2025)
FD2Talk: Towards Generalized Talking Head Generation with Facial Decoupled Diffusion Model
von: Yao, Ziyu, et al.
Veröffentlicht: (2024)
von: Yao, Ziyu, et al.
Veröffentlicht: (2024)
VAST: Vivify Your Talking Avatar via Zero-Shot Expressive Facial Style Transfer
von: Chen, Liyang, et al.
Veröffentlicht: (2023)
von: Chen, Liyang, et al.
Veröffentlicht: (2023)
FaceEditTalker: Controllable Talking Head Generation with Facial Attribute Editing
von: Feng, Guanwen, et al.
Veröffentlicht: (2025)
von: Feng, Guanwen, et al.
Veröffentlicht: (2025)
SoulX-FlashHead: Oracle-guided Generation of Infinite Real-time Streaming Talking Heads
von: Yu, Tan, et al.
Veröffentlicht: (2026)
von: Yu, Tan, et al.
Veröffentlicht: (2026)
EmoTalkingGaussian: Continuous Emotion-conditioned Talking Head Synthesis
von: Cha, Junuk, et al.
Veröffentlicht: (2025)
von: Cha, Junuk, et al.
Veröffentlicht: (2025)
FreeTalk: Emotional Topology-Free 3D Talking Heads
von: Nocentini, Federico, et al.
Veröffentlicht: (2026)
von: Nocentini, Federico, et al.
Veröffentlicht: (2026)
RealTalk: Realistic Emotion-Aware Lifelike Talking-Head Synthesis
von: Wang, Wenqing, et al.
Veröffentlicht: (2025)
von: Wang, Wenqing, et al.
Veröffentlicht: (2025)
ScanTalk: 3D Talking Heads from Unregistered Scans
von: Nocentini, Federico, et al.
Veröffentlicht: (2024)
von: Nocentini, Federico, et al.
Veröffentlicht: (2024)
TalkingGaussian: Structure-Persistent 3D Talking Head Synthesis via Gaussian Splatting
von: Li, Jiahe, et al.
Veröffentlicht: (2024)
von: Li, Jiahe, et al.
Veröffentlicht: (2024)
FixTalk: Taming Identity Leakage for High-Quality Talking Head Generation in Extreme Cases
von: Tan, Shuai, et al.
Veröffentlicht: (2025)
von: Tan, Shuai, et al.
Veröffentlicht: (2025)
ConsistTalk: Intensity Controllable Temporally Consistent Talking Head Generation with Diffusion Noise Search
von: Liu, Zhenjie, et al.
Veröffentlicht: (2025)
von: Liu, Zhenjie, et al.
Veröffentlicht: (2025)
LLM4GEN: Leveraging Semantic Representation of LLMs for Text-to-Image Generation
von: Liu, Mushui, et al.
Veröffentlicht: (2024)
von: Liu, Mushui, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
StyleTalk++: A Unified Framework for Controlling the Speaking Styles of Talking Heads
von: Wang, Suzhen, et al.
Veröffentlicht: (2024) -
DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models
von: Ma, Yifeng, et al.
Veröffentlicht: (2023) -
Towards a Simultaneous and Granular Identity-Expression Control in Personalized Face Generation
von: Liu, Renshuai, et al.
Veröffentlicht: (2024) -
EmotiveTalk: Expressive Talking Head Generation through Audio Information Decoupling and Emotional Video Diffusion
von: Wang, Haotian, et al.
Veröffentlicht: (2024) -
Talking Head Generation via AU-Guided Landmark Prediction
von: Chang, Shao-Yu, et al.
Veröffentlicht: (2025)