DiffPoseTalk: Speech-Driven Stylistic 3D Facial Animation and Head Pose Generation via Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Zhiyao, Lv, Tian, Ye, Sheng, Lin, Matthieu, Sheng, Jenny, Wen, Yu-Hui, Yu, Minjing, Liu, Yong-Jin |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Text-to-3D Framework for Joint Generation of CG-Ready Humans and Compatible Garments
by: Sun, Zhiyao, et al.
Published: (2025)
by: Sun, Zhiyao, et al.
Published: (2025)
OT-Talk: Animating 3D Talking Head with Optimal Transportation
by: Wang, Xinmu, et al.
Published: (2025)
by: Wang, Xinmu, et al.
Published: (2025)
LSF-Animation: Label-Free Speech-Driven Facial Animation via Implicit Feature Representation
by: Lu, Xin, et al.
Published: (2025)
by: Lu, Xin, et al.
Published: (2025)
PESTalk: Speech-Driven 3D Facial Animation with Personalized Emotional Styles
by: Han, Tianshun, et al.
Published: (2025)
by: Han, Tianshun, et al.
Published: (2025)
AniGaussian: Animatable Gaussian Avatar with Pose-guided Deformation
by: Li, Mengtian, et al.
Published: (2025)
by: Li, Mengtian, et al.
Published: (2025)
DiffBody: Diffusion-based Pose and Shape Editing of Human Images
by: Okuyama, Yuta, et al.
Published: (2024)
by: Okuyama, Yuta, et al.
Published: (2024)
KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation
by: Lyu, Tianle, et al.
Published: (2025)
by: Lyu, Tianle, et al.
Published: (2025)
AniGaussian: Animatable Gaussian Avatar With Pose‐Guided Deformation
by: M. Li, et al.
Published: (2026)
by: M. Li, et al.
Published: (2026)
Text‐Guided Diffusion with Spectral Convolution for 3D Human Pose Estimation
by: Liyuan Shi, et al.
Published: (2025)
by: Liyuan Shi, et al.
Published: (2025)
PVP-Recon: Progressive View Planning via Warping Consistency for Sparse-View Surface Reconstruction
by: Ye, Sheng, et al.
Published: (2024)
by: Ye, Sheng, et al.
Published: (2024)
Pose Representations for Deep Skeletal Animation
by: Andreou, Nefeli, et al.
Published: (2021)
by: Andreou, Nefeli, et al.
Published: (2021)
Learning Disentangled Speech- and Expression-Driven Blendshapes for 3D Talking Face Animation
by: Mao, Yuxiang, et al.
Published: (2025)
by: Mao, Yuxiang, et al.
Published: (2025)
AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
by: Guo, Yuwei, et al.
Published: (2023)
by: Guo, Yuwei, et al.
Published: (2023)
SAiD: Speech-driven Blendshape Facial Animation with Diffusion
by: Park, Inkyu, et al.
Published: (2023)
by: Park, Inkyu, et al.
Published: (2023)
Content and Style Aware Audio-Driven Facial Animation
by: Liu, Qingju, et al.
Published: (2024)
by: Liu, Qingju, et al.
Published: (2024)
Ponimator: Unfolding Interactive Pose for Versatile Human-human Interaction Animation
by: Liu, Shaowei, et al.
Published: (2025)
by: Liu, Shaowei, et al.
Published: (2025)
Enhancing Speech-Driven 3D Facial Animation with Audio-Visual Guidance from Lip Reading Expert
by: EunGi, Han, et al.
Published: (2024)
by: EunGi, Han, et al.
Published: (2024)
Model See Model Do: Speech-Driven Facial Animation with Style Control
by: Pan, Yifang, et al.
Published: (2025)
by: Pan, Yifang, et al.
Published: (2025)
PoseGaussian: Pose-Driven Novel View Synthesis for Robust 3D Human Reconstruction
by: Shen, Ju, et al.
Published: (2026)
by: Shen, Ju, et al.
Published: (2026)
FreeAvatar: Robust 3D Facial Animation Transfer by Learning an Expression Foundation Model
by: Qiu, Feng, et al.
Published: (2024)
by: Qiu, Feng, et al.
Published: (2024)
Generating Detailed Character Motion from Blocking Poses
by: Goel, Purvi, et al.
Published: (2025)
by: Goel, Purvi, et al.
Published: (2025)
Democratizing the Creation of Animatable Facial Avatars
by: Zhu, Yilin, et al.
Published: (2024)
by: Zhu, Yilin, et al.
Published: (2024)
PointTalk: Audio-Driven Dynamic Lip Point Cloud for 3D Gaussian-based Talking Head Synthesis
by: Xie, Yifan, et al.
Published: (2024)
by: Xie, Yifan, et al.
Published: (2024)
Lightweight Self-Driven Deformable Organ Animations
by: Kenwright, Benjamnin, et al.
Published: (2024)
by: Kenwright, Benjamnin, et al.
Published: (2024)
Audio Driven Real-Time Facial Animation for Social Telepresence
by: Lee, Jiye, et al.
Published: (2025)
by: Lee, Jiye, et al.
Published: (2025)
MoDA: Multi-modal Diffusion Architecture for Talking Head Generation
by: Li, Xinyang, et al.
Published: (2025)
by: Li, Xinyang, et al.
Published: (2025)
EditYourself: Audio-Driven Generation and Manipulation of Talking Head Videos with Diffusion Transformers
by: Flynn, John, et al.
Published: (2026)
by: Flynn, John, et al.
Published: (2026)
Generalizable and Animatable Gaussian Head Avatar
by: Chu, Xuangeng, et al.
Published: (2024)
by: Chu, Xuangeng, et al.
Published: (2024)
Neural Pose Representation Learning for Generating and Transferring Non-Rigid Object Poses
by: Yoo, Seungwoo, et al.
Published: (2024)
by: Yoo, Seungwoo, et al.
Published: (2024)
CAG-Avatar: Cross-Attention Guided Gaussian Avatars for High-Fidelity Head Reconstruction
by: Chang, Zhe, et al.
Published: (2026)
by: Chang, Zhe, et al.
Published: (2026)
PASE: Phoneme-Aware Speech Encoder to Improve Lip Sync Accuracy for Talking Head Synthesis
by: Huang, Yihuan, et al.
Published: (2025)
by: Huang, Yihuan, et al.
Published: (2025)
Near-realtime Facial Animation by Deep 3D Simulation Super-Resolution
by: Park, Hyojoon, et al.
Published: (2023)
by: Park, Hyojoon, et al.
Published: (2023)
Think2Sing: Orchestrating Structured Motion Subtitles for Singing-Driven 3D Head Animation
by: Huang, Zikai, et al.
Published: (2025)
by: Huang, Zikai, et al.
Published: (2025)
Media2Face: Co-speech Facial Animation Generation With Multi-Modality Guidance
by: Zhao, Qingcheng, et al.
Published: (2024)
by: Zhao, Qingcheng, et al.
Published: (2024)
Llanimation: Llama Driven Gesture Animation
by: J. Windle, et al.
Published: (2024)
by: J. Windle, et al.
Published: (2024)
FaceTalk: Audio-Driven Motion Diffusion for Neural Parametric Head Models
by: Aneja, Shivangi, et al.
Published: (2023)
by: Aneja, Shivangi, et al.
Published: (2023)
EmoDiffGes: Emotion‐Aware Co‐Speech Holistic Gesture Generation with Progressive Synergistic Diffusion
by: Xinru Li, et al.
Published: (2025)
by: Xinru Li, et al.
Published: (2025)
Animating Childlike Drawings with 2.5D Character Rigs
by: Smith, Harrison Jesse, et al.
Published: (2025)
by: Smith, Harrison Jesse, et al.
Published: (2025)
Instant Facial Gaussians Translator for Relightable and Interactable Facial Rendering
by: Qin, Dafei, et al.
Published: (2024)
by: Qin, Dafei, et al.
Published: (2024)
Learning Uniformly Distributed Embedding Clusters of Stylistic Skills for Physically Simulated Characters
by: Liu, Nian, et al.
Published: (2024)
by: Liu, Nian, et al.
Published: (2024)
Similar Items
-
A Text-to-3D Framework for Joint Generation of CG-Ready Humans and Compatible Garments
by: Sun, Zhiyao, et al.
Published: (2025) -
OT-Talk: Animating 3D Talking Head with Optimal Transportation
by: Wang, Xinmu, et al.
Published: (2025) -
LSF-Animation: Label-Free Speech-Driven Facial Animation via Implicit Feature Representation
by: Lu, Xin, et al.
Published: (2025) -
PESTalk: Speech-Driven 3D Facial Animation with Personalized Emotional Styles
by: Han, Tianshun, et al.
Published: (2025) -
AniGaussian: Animatable Gaussian Avatar with Pose-guided Deformation
by: Li, Mengtian, et al.
Published: (2025)