FaceTalk: Audio-Driven Motion Diffusion for Neural Parametric Head Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Aneja, Shivangi, Thies, Justus, Dai, Angela, Nießner, Matthias |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GaussianSpeech: Audio-Driven Gaussian Avatars
von: Aneja, Shivangi, et al.
Veröffentlicht: (2024)
von: Aneja, Shivangi, et al.
Veröffentlicht: (2024)
READ: Real-time and Efficient Asynchronous Diffusion for Audio-driven Talking Head Generation
von: Wang, Haotian, et al.
Veröffentlicht: (2025)
von: Wang, Haotian, et al.
Veröffentlicht: (2025)
Text-driven Talking Face Synthesis by Reprogramming Audio-driven Models
von: Choi, Jeongsoo, et al.
Veröffentlicht: (2023)
von: Choi, Jeongsoo, et al.
Veröffentlicht: (2023)
Audio-Plane: Audio Factorization Plane Gaussian Splatting for Real-Time Talking Head Synthesis
von: Shen, Shuai, et al.
Veröffentlicht: (2025)
von: Shen, Shuai, et al.
Veröffentlicht: (2025)
NeRF-3DTalker: Neural Radiance Field with 3D Prior Aided Audio Disentanglement for Talking Head Synthesis
von: Liu, Xiaoxing, et al.
Veröffentlicht: (2025)
von: Liu, Xiaoxing, et al.
Veröffentlicht: (2025)
Silence is Golden: Leveraging Adversarial Examples to Nullify Audio Control in LDM-based Talking-Head Generation
von: Gan, Yuan, et al.
Veröffentlicht: (2025)
von: Gan, Yuan, et al.
Veröffentlicht: (2025)
Meta-Learning Empowered Meta-Face: Personalized Speaking Style Adaptation for Audio-Driven 3D Talking Face Animation
von: Zhou, Xukun, et al.
Veröffentlicht: (2024)
von: Zhou, Xukun, et al.
Veröffentlicht: (2024)
ELGAR: Expressive Cello Performance Motion Generation for Audio Rendition
von: Qiu, Zhiping, et al.
Veröffentlicht: (2025)
von: Qiu, Zhiping, et al.
Veröffentlicht: (2025)
TCDiff++: An End-to-end Trajectory-Controllable Diffusion Model for Harmonious Music-Driven Group Choreography
von: Dai, Yuqin, et al.
Veröffentlicht: (2025)
von: Dai, Yuqin, et al.
Veröffentlicht: (2025)
RAP: Real-time Audio-driven Portrait Animation with Video Diffusion Transformer
von: Du, Fangyu, et al.
Veröffentlicht: (2025)
von: Du, Fangyu, et al.
Veröffentlicht: (2025)
PointTalk: Audio-Driven Dynamic Lip Point Cloud for 3D Gaussian-based Talking Head Synthesis
von: Xie, Yifan, et al.
Veröffentlicht: (2024)
von: Xie, Yifan, et al.
Veröffentlicht: (2024)
MotionRAG-Diff: A Retrieval-Augmented Diffusion Framework for Long-Term Music-to-Dance Generation
von: Huang, Mingyang, et al.
Veröffentlicht: (2025)
von: Huang, Mingyang, et al.
Veröffentlicht: (2025)
Content and Style Aware Audio-Driven Facial Animation
von: Liu, Qingju, et al.
Veröffentlicht: (2024)
von: Liu, Qingju, et al.
Veröffentlicht: (2024)
A Comprehensive Multi-scale Approach for Speech and Dynamics Synchrony in Talking Head Generation
von: Airale, Louis, et al.
Veröffentlicht: (2023)
von: Airale, Louis, et al.
Veröffentlicht: (2023)
Seeing Sound: Assembling Sounds from Visuals for Audio-to-Image Generation
von: Petermann, Darius, et al.
Veröffentlicht: (2025)
von: Petermann, Darius, et al.
Veröffentlicht: (2025)
MACS: Multi-source Audio-to-image Generation with Contextual Significance and Semantic Alignment
von: Zhou, Hao, et al.
Veröffentlicht: (2025)
von: Zhou, Hao, et al.
Veröffentlicht: (2025)
Inter-Diffusion Generation Model of Speakers and Listeners for Effective Communication
von: Huang, Jinhe, et al.
Veröffentlicht: (2025)
von: Huang, Jinhe, et al.
Veröffentlicht: (2025)
GCDance: Genre-Controlled Music-Driven 3D Full Body Dance Generation
von: Liu, Xinran, et al.
Veröffentlicht: (2025)
von: Liu, Xinran, et al.
Veröffentlicht: (2025)
DuetGen: Music Driven Two-Person Dance Generation via Hierarchical Masked Modeling
von: Ghosh, Anindita, et al.
Veröffentlicht: (2025)
von: Ghosh, Anindita, et al.
Veröffentlicht: (2025)
DIDiffGes: Decoupled Semi-Implicit Diffusion Models for Real-time Gesture Generation from Speech
von: Cheng, Yongkang, et al.
Veröffentlicht: (2025)
von: Cheng, Yongkang, et al.
Veröffentlicht: (2025)
Lodge: A Coarse to Fine Diffusion Network for Long Dance Generation Guided by the Characteristic Dance Primitives
von: Li, Ronghui, et al.
Veröffentlicht: (2024)
von: Li, Ronghui, et al.
Veröffentlicht: (2024)
Sonic: Shifting Focus to Global Audio Perception in Portrait Animation
von: Ji, Xiaozhong, et al.
Veröffentlicht: (2024)
von: Ji, Xiaozhong, et al.
Veröffentlicht: (2024)
Conditional GAN for Enhancing Diffusion Models in Efficient and Authentic Global Gesture Generation from Audios
von: Cheng, Yongkang, et al.
Veröffentlicht: (2024)
von: Cheng, Yongkang, et al.
Veröffentlicht: (2024)
SyncViolinist: Music-Oriented Violin Motion Generation Based on Bowing and Fingering
von: Nishizawa, Hiroki, et al.
Veröffentlicht: (2024)
von: Nishizawa, Hiroki, et al.
Veröffentlicht: (2024)
Dual Audio-Centric Modality Coupling for Talking Head Generation
von: Fu, Ao, et al.
Veröffentlicht: (2025)
von: Fu, Ao, et al.
Veröffentlicht: (2025)
EnchantDance: Unveiling the Potential of Music-Driven Dance Movement
von: Han, Bo, et al.
Veröffentlicht: (2023)
von: Han, Bo, et al.
Veröffentlicht: (2023)
Text-Driven Voice Conversion via Latent State-Space Modeling
von: Li, Wen, et al.
Veröffentlicht: (2025)
von: Li, Wen, et al.
Veröffentlicht: (2025)
DGFM: Full Body Dance Generation Driven by Music Foundation Models
von: Liu, Xinran, et al.
Veröffentlicht: (2025)
von: Liu, Xinran, et al.
Veröffentlicht: (2025)
NAT: Neural Acoustic Transfer for Interactive Scenes in Real Time
von: Jin, Xutong, et al.
Veröffentlicht: (2025)
von: Jin, Xutong, et al.
Veröffentlicht: (2025)
MIDGET: Music Conditioned 3D Dance Generation
von: Wang, Jinwu, et al.
Veröffentlicht: (2024)
von: Wang, Jinwu, et al.
Veröffentlicht: (2024)
Semantic Gesticulator: Semantics-Aware Co-Speech Gesture Synthesis
von: Zhang, Zeyi, et al.
Veröffentlicht: (2024)
von: Zhang, Zeyi, et al.
Veröffentlicht: (2024)
Duolando: Follower GPT with Off-Policy Reinforcement Learning for Dance Accompaniment
von: Siyao, Li, et al.
Veröffentlicht: (2024)
von: Siyao, Li, et al.
Veröffentlicht: (2024)
It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model
von: Shi, Mingyi, et al.
Veröffentlicht: (2024)
von: Shi, Mingyi, et al.
Veröffentlicht: (2024)
LAV: Audio-Driven Dynamic Visual Generation with Neural Compression and StyleGAN2
von: Jung, Jongmin, et al.
Veröffentlicht: (2025)
von: Jung, Jongmin, et al.
Veröffentlicht: (2025)
M3G: Multi-Granular Gesture Generator for Audio-Driven Full-Body Human Motion Synthesis
von: Yin, Zhizhuo, et al.
Veröffentlicht: (2025)
von: Yin, Zhizhuo, et al.
Veröffentlicht: (2025)
Ditto: Motion-Space Diffusion for Controllable Realtime Talking Head Synthesis
von: Li, Tianqi, et al.
Veröffentlicht: (2024)
von: Li, Tianqi, et al.
Veröffentlicht: (2024)
FADA: Fast Diffusion Avatar Synthesis with Mixed-Supervised Multi-CFG Distillation
von: Zhong, Tianyun, et al.
Veröffentlicht: (2024)
von: Zhong, Tianyun, et al.
Veröffentlicht: (2024)
MusicScore: A Dataset for Music Score Modeling and Generation
von: Lin, Yuheng, et al.
Veröffentlicht: (2024)
von: Lin, Yuheng, et al.
Veröffentlicht: (2024)
EmoTalker: Emotionally Editable Talking Face Generation via Diffusion Model
von: Zhang, Bingyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Bingyuan, et al.
Veröffentlicht: (2024)
AsynFusion: Towards Asynchronous Latent Consistency Models for Decoupled Whole-Body Audio-Driven Avatars
von: Zhang, Tianbao, et al.
Veröffentlicht: (2025)
von: Zhang, Tianbao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
GaussianSpeech: Audio-Driven Gaussian Avatars
von: Aneja, Shivangi, et al.
Veröffentlicht: (2024) -
READ: Real-time and Efficient Asynchronous Diffusion for Audio-driven Talking Head Generation
von: Wang, Haotian, et al.
Veröffentlicht: (2025) -
Text-driven Talking Face Synthesis by Reprogramming Audio-driven Models
von: Choi, Jeongsoo, et al.
Veröffentlicht: (2023) -
Audio-Plane: Audio Factorization Plane Gaussian Splatting for Real-Time Talking Head Synthesis
von: Shen, Shuai, et al.
Veröffentlicht: (2025) -
NeRF-3DTalker: Neural Radiance Field with 3D Prior Aided Audio Disentanglement for Talking Head Synthesis
von: Liu, Xiaoxing, et al.
Veröffentlicht: (2025)