DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer
Fuente:
arXiv
Salvato in:
| Autori principali: | Ma, Zhiyuan, Zhu, Xiangyu, Qi, Guojun, Qian, Chen, Zhang, Zhaoxiang, Lei, Zhen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
GLDiTalker: Speech-Driven 3D Facial Animation with Graph Latent Diffusion Transformer
di: Lin, Yihong, et al.
Pubblicazione: (2024)
di: Lin, Yihong, et al.
Pubblicazione: (2024)
EmoDiffusion: Enhancing Emotional 3D Facial Animation with Latent Diffusion Models
di: Zhang, Yixuan, et al.
Pubblicazione: (2025)
di: Zhang, Yixuan, et al.
Pubblicazione: (2025)
DiffPoseTalk: Speech-Driven Stylistic 3D Facial Animation and Head Pose Generation via Diffusion Models
di: Sun, Zhiyao, et al.
Pubblicazione: (2023)
di: Sun, Zhiyao, et al.
Pubblicazione: (2023)
PC-Talk: Precise Facial Animation Control for Audio-Driven Talking Face Generation
di: Wang, Baiqin, et al.
Pubblicazione: (2025)
di: Wang, Baiqin, et al.
Pubblicazione: (2025)
AdaMesh: Personalized Facial Expressions and Head Poses for Adaptive Speech-Driven 3D Facial Animation
di: Chen, Liyang, et al.
Pubblicazione: (2023)
di: Chen, Liyang, et al.
Pubblicazione: (2023)
KMTalk: Speech-Driven 3D Facial Animation with Key Motion Embedding
di: Xu, Zhihao, et al.
Pubblicazione: (2024)
di: Xu, Zhihao, et al.
Pubblicazione: (2024)
UniTalker: Scaling up Audio-Driven 3D Facial Animation through A Unified Model
di: Fan, Xiangyu, et al.
Pubblicazione: (2024)
di: Fan, Xiangyu, et al.
Pubblicazione: (2024)
EditEmoTalk: Controllable Speech-Driven 3D Facial Animation with Continuous Expression Editing
di: Jiang, Diqiong, et al.
Pubblicazione: (2026)
di: Jiang, Diqiong, et al.
Pubblicazione: (2026)
UltrAvatar: A Realistic Animatable 3D Avatar Diffusion Model with Authenticity Guided Textures
di: Zhou, Mingyuan, et al.
Pubblicazione: (2024)
di: Zhou, Mingyuan, et al.
Pubblicazione: (2024)
Polyglot: Multilingual Style Preserving Speech-Driven Facial Animation
di: Nocentini, Federico, et al.
Pubblicazione: (2026)
di: Nocentini, Federico, et al.
Pubblicazione: (2026)
Diverse Code Query Learning for Speech-Driven Facial Animation
di: Gu, Chunzhi, et al.
Pubblicazione: (2024)
di: Gu, Chunzhi, et al.
Pubblicazione: (2024)
CSTalk: Correlation Supervised Speech-driven 3D Emotional Facial Animation Generation
di: Liang, Xiangyu, et al.
Pubblicazione: (2024)
di: Liang, Xiangyu, et al.
Pubblicazione: (2024)
Seek for Incantations: Towards Accurate Text-to-Image Diffusion Synthesis through Prompt Engineering
di: Yu, Chang, et al.
Pubblicazione: (2024)
di: Yu, Chang, et al.
Pubblicazione: (2024)
MVBoost: Boost 3D Reconstruction with Multi-View Refinement
di: Liu, Xiangyu, et al.
Pubblicazione: (2024)
di: Liu, Xiangyu, et al.
Pubblicazione: (2024)
3D Face Reconstruction with the Geometric Guidance of Facial Part Segmentation
di: Wang, Zidu, et al.
Pubblicazione: (2023)
di: Wang, Zidu, et al.
Pubblicazione: (2023)
StreamingTalker: Audio-driven 3D Facial Animation with Autoregressive Diffusion Model
di: Yang, Yifan, et al.
Pubblicazione: (2025)
di: Yang, Yifan, et al.
Pubblicazione: (2025)
Learning Phonetic Context-Dependent Viseme for Enhancing Speech-Driven 3D Facial Animation
di: Kim, Hyung Kyu, et al.
Pubblicazione: (2025)
di: Kim, Hyung Kyu, et al.
Pubblicazione: (2025)
LSF-Animation: Label-Free Speech-Driven Facial Animation via Implicit Feature Representation
di: Lu, Xin, et al.
Pubblicazione: (2025)
di: Lu, Xin, et al.
Pubblicazione: (2025)
Expressive Speech-driven Facial Animation with controllable emotions
di: Chen, Yutong, et al.
Pubblicazione: (2023)
di: Chen, Yutong, et al.
Pubblicazione: (2023)
ReactDiff: Latent Diffusion for Facial Reaction Generation
di: Li, Jiaming, et al.
Pubblicazione: (2025)
di: Li, Jiaming, et al.
Pubblicazione: (2025)
Progressive Rendering Distillation: Adapting Stable Diffusion for Instant Text-to-Mesh Generation without 3D Data
di: Ma, Zhiyuan, et al.
Pubblicazione: (2025)
di: Ma, Zhiyuan, et al.
Pubblicazione: (2025)
Top-Down Guidance for Learning Object-Centric Representations
di: Zou, Junhong, et al.
Pubblicazione: (2024)
di: Zou, Junhong, et al.
Pubblicazione: (2024)
Revisiting Marr in Face: The Building of 2D--2.5D--3D Representations in Deep Neural Networks
di: Zhu, Xiangyu, et al.
Pubblicazione: (2024)
di: Zhu, Xiangyu, et al.
Pubblicazione: (2024)
Enhancing Speech-Driven 3D Facial Animation with Audio-Visual Guidance from Lip Reading Expert
di: EunGi, Han, et al.
Pubblicazione: (2024)
di: EunGi, Han, et al.
Pubblicazione: (2024)
MemoryTalker: Personalized Speech-Driven 3D Facial Animation via Audio-Guided Stylization
di: Kim, Hyung Kyu, et al.
Pubblicazione: (2025)
di: Kim, Hyung Kyu, et al.
Pubblicazione: (2025)
Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation
di: Li, Hao, et al.
Pubblicazione: (2025)
di: Li, Hao, et al.
Pubblicazione: (2025)
SAiD: Speech-driven Blendshape Facial Animation with Diffusion
di: Park, Inkyu, et al.
Pubblicazione: (2023)
di: Park, Inkyu, et al.
Pubblicazione: (2023)
TalkingEyes: Pluralistic Speech-Driven 3D Eye Gaze Animation
di: Zhuang, Yixiang, et al.
Pubblicazione: (2025)
di: Zhuang, Yixiang, et al.
Pubblicazione: (2025)
ESGaussianFace: Emotional and Stylized Audio-Driven Facial Animation via 3D Gaussian Splatting
di: Ma, Chuhang, et al.
Pubblicazione: (2026)
di: Ma, Chuhang, et al.
Pubblicazione: (2026)
AudioFace: Language-Assisted Speech-Driven Facial Animation with Multimodal Language Models
di: Zheng, Kai, et al.
Pubblicazione: (2026)
di: Zheng, Kai, et al.
Pubblicazione: (2026)
Learning Semantic Facial Descriptors for Accurate Face Animation
di: Zhu, Lei, et al.
Pubblicazione: (2025)
di: Zhu, Lei, et al.
Pubblicazione: (2025)
AnimateMe: 4D Facial Expressions via Diffusion Models
di: Gerogiannis, Dimitrios, et al.
Pubblicazione: (2024)
di: Gerogiannis, Dimitrios, et al.
Pubblicazione: (2024)
SEDTalker: Emotion-Aware 3D Facial Animation Using Frame-Level Speech Emotion Diarization
di: Jafari, Farzaneh, et al.
Pubblicazione: (2026)
di: Jafari, Farzaneh, et al.
Pubblicazione: (2026)
EmoFace: Emotion-Content Disentangled Speech-Driven 3D Talking Face Animation
di: Lin, Yihong, et al.
Pubblicazione: (2024)
di: Lin, Yihong, et al.
Pubblicazione: (2024)
AnimateAnything: Consistent and Controllable Animation for Video Generation
di: Lei, Guojun, et al.
Pubblicazione: (2024)
di: Lei, Guojun, et al.
Pubblicazione: (2024)
ScaleDreamer: Scalable Text-to-3D Synthesis with Asynchronous Score Distillation
di: Ma, Zhiyuan, et al.
Pubblicazione: (2024)
di: Ma, Zhiyuan, et al.
Pubblicazione: (2024)
ProbTalk3D: Non-Deterministic Emotion Controllable Speech-Driven 3D Facial Animation Synthesis Using VQ-VAE
di: Wu, Sichun, et al.
Pubblicazione: (2024)
di: Wu, Sichun, et al.
Pubblicazione: (2024)
ARTalk: Speech-Driven 3D Head Animation via Autoregressive Model
di: Chu, Xuangeng, et al.
Pubblicazione: (2025)
di: Chu, Xuangeng, et al.
Pubblicazione: (2025)
EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditions
di: Chen, Zhiyuan, et al.
Pubblicazione: (2024)
di: Chen, Zhiyuan, et al.
Pubblicazione: (2024)
DEEPTalk: Dynamic Emotion Embedding for Probabilistic Speech-Driven 3D Face Animation
di: Kim, Jisoo, et al.
Pubblicazione: (2024)
di: Kim, Jisoo, et al.
Pubblicazione: (2024)
Documenti analoghi
-
GLDiTalker: Speech-Driven 3D Facial Animation with Graph Latent Diffusion Transformer
di: Lin, Yihong, et al.
Pubblicazione: (2024) -
EmoDiffusion: Enhancing Emotional 3D Facial Animation with Latent Diffusion Models
di: Zhang, Yixuan, et al.
Pubblicazione: (2025) -
DiffPoseTalk: Speech-Driven Stylistic 3D Facial Animation and Head Pose Generation via Diffusion Models
di: Sun, Zhiyao, et al.
Pubblicazione: (2023) -
PC-Talk: Precise Facial Animation Control for Audio-Driven Talking Face Generation
di: Wang, Baiqin, et al.
Pubblicazione: (2025) -
AdaMesh: Personalized Facial Expressions and Head Poses for Adaptive Speech-Driven 3D Facial Animation
di: Chen, Liyang, et al.
Pubblicazione: (2023)