Meta-Learning Empowered Meta-Face: Personalized Speaking Style Adaptation for Audio-Driven 3D Talking Face Animation
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhou, Xukun, Li, Fengxin, Peng, Ziqiao, Wu, Kejian, He, Jun, Qin, Biao, Fan, Zhaoxin, Liu, Hongyan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Content and Style Aware Audio-Driven Facial Animation
di: Liu, Qingju, et al.
Pubblicazione: (2024)
di: Liu, Qingju, et al.
Pubblicazione: (2024)
MATHDance: Mamba-Transformer Architecture with Uniform Tokenization for High-Quality 3D Dance Generation
di: Yang, Kaixing, et al.
Pubblicazione: (2025)
di: Yang, Kaixing, et al.
Pubblicazione: (2025)
Text-driven Talking Face Synthesis by Reprogramming Audio-driven Models
di: Choi, Jeongsoo, et al.
Pubblicazione: (2023)
di: Choi, Jeongsoo, et al.
Pubblicazione: (2023)
FaceTalk: Audio-Driven Motion Diffusion for Neural Parametric Head Models
di: Aneja, Shivangi, et al.
Pubblicazione: (2023)
di: Aneja, Shivangi, et al.
Pubblicazione: (2023)
Audio2Face-3D: Audio-driven Realistic Facial Animation For Digital Avatars
di: NVIDIA, et al.
Pubblicazione: (2025)
di: NVIDIA, et al.
Pubblicazione: (2025)
PointTalk: Audio-Driven Dynamic Lip Point Cloud for 3D Gaussian-based Talking Head Synthesis
di: Xie, Yifan, et al.
Pubblicazione: (2024)
di: Xie, Yifan, et al.
Pubblicazione: (2024)
LAV: Audio-Driven Dynamic Visual Generation with Neural Compression and StyleGAN2
di: Jung, Jongmin, et al.
Pubblicazione: (2025)
di: Jung, Jongmin, et al.
Pubblicazione: (2025)
ELGAR: Expressive Cello Performance Motion Generation for Audio Rendition
di: Qiu, Zhiping, et al.
Pubblicazione: (2025)
di: Qiu, Zhiping, et al.
Pubblicazione: (2025)
DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations
di: Peng, Ziqiao, et al.
Pubblicazione: (2025)
di: Peng, Ziqiao, et al.
Pubblicazione: (2025)
EnchantDance: Unveiling the Potential of Music-Driven Dance Movement
di: Han, Bo, et al.
Pubblicazione: (2023)
di: Han, Bo, et al.
Pubblicazione: (2023)
Text-Driven Voice Conversion via Latent State-Space Modeling
di: Li, Wen, et al.
Pubblicazione: (2025)
di: Li, Wen, et al.
Pubblicazione: (2025)
DGFM: Full Body Dance Generation Driven by Music Foundation Models
di: Liu, Xinran, et al.
Pubblicazione: (2025)
di: Liu, Xinran, et al.
Pubblicazione: (2025)
Sonic: Shifting Focus to Global Audio Perception in Portrait Animation
di: Ji, Xiaozhong, et al.
Pubblicazione: (2024)
di: Ji, Xiaozhong, et al.
Pubblicazione: (2024)
READ: Real-time and Efficient Asynchronous Diffusion for Audio-driven Talking Head Generation
di: Wang, Haotian, et al.
Pubblicazione: (2025)
di: Wang, Haotian, et al.
Pubblicazione: (2025)
Audio-Plane: Audio Factorization Plane Gaussian Splatting for Real-Time Talking Head Synthesis
di: Shen, Shuai, et al.
Pubblicazione: (2025)
di: Shen, Shuai, et al.
Pubblicazione: (2025)
RAP: Real-time Audio-driven Portrait Animation with Video Diffusion Transformer
di: Du, Fangyu, et al.
Pubblicazione: (2025)
di: Du, Fangyu, et al.
Pubblicazione: (2025)
Beat-It: Beat-Synchronized Multi-Condition 3D Dance Generation
di: Huang, Zikai, et al.
Pubblicazione: (2024)
di: Huang, Zikai, et al.
Pubblicazione: (2024)
CoheDancers: Enhancing Interactive Group Dance Generation through Music-Driven Coherence Decomposition
di: Yang, Kaixing, et al.
Pubblicazione: (2024)
di: Yang, Kaixing, et al.
Pubblicazione: (2024)
Listen and Move: Improving GANs Coherency in Agnostic Sound-to-Video Generation
di: Redondo, Rafael
Pubblicazione: (2024)
di: Redondo, Rafael
Pubblicazione: (2024)
NAT: Neural Acoustic Transfer for Interactive Scenes in Real Time
di: Jin, Xutong, et al.
Pubblicazione: (2025)
di: Jin, Xutong, et al.
Pubblicazione: (2025)
SyncViolinist: Music-Oriented Violin Motion Generation Based on Bowing and Fingering
di: Nishizawa, Hiroki, et al.
Pubblicazione: (2024)
di: Nishizawa, Hiroki, et al.
Pubblicazione: (2024)
FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles
di: Zhang, Tian-Hao, et al.
Pubblicazione: (2025)
di: Zhang, Tian-Hao, et al.
Pubblicazione: (2025)
Audio is all in one: speech-driven gesture synthetics using WavLM pre-trained model
di: Zhang, Fan, et al.
Pubblicazione: (2023)
di: Zhang, Fan, et al.
Pubblicazione: (2023)
DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation
di: Meng, Ming, et al.
Pubblicazione: (2025)
di: Meng, Ming, et al.
Pubblicazione: (2025)
AsynFusion: Towards Asynchronous Latent Consistency Models for Decoupled Whole-Body Audio-Driven Avatars
di: Zhang, Tianbao, et al.
Pubblicazione: (2025)
di: Zhang, Tianbao, et al.
Pubblicazione: (2025)
NeRF-3DTalker: Neural Radiance Field with 3D Prior Aided Audio Disentanglement for Talking Head Synthesis
di: Liu, Xiaoxing, et al.
Pubblicazione: (2025)
di: Liu, Xiaoxing, et al.
Pubblicazione: (2025)
DanceAnyWay: Synthesizing Beat-Guided 3D Dances with Randomized Temporal Contrastive Learning
di: Bhattacharya, Aneesh, et al.
Pubblicazione: (2023)
di: Bhattacharya, Aneesh, et al.
Pubblicazione: (2023)
Combining Genre Classification and Harmonic-Percussive Features with Diffusion Models for Music-Video Generation
di: Pina, Leonardo, et al.
Pubblicazione: (2024)
di: Pina, Leonardo, et al.
Pubblicazione: (2024)
MusicScore: A Dataset for Music Score Modeling and Generation
di: Lin, Yuheng, et al.
Pubblicazione: (2024)
di: Lin, Yuheng, et al.
Pubblicazione: (2024)
DuetGen: Music Driven Two-Person Dance Generation via Hierarchical Masked Modeling
di: Ghosh, Anindita, et al.
Pubblicazione: (2025)
di: Ghosh, Anindita, et al.
Pubblicazione: (2025)
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval
di: Sun, Haoqin, et al.
Pubblicazione: (2025)
di: Sun, Haoqin, et al.
Pubblicazione: (2025)
Silence is Golden: Leveraging Adversarial Examples to Nullify Audio Control in LDM-based Talking-Head Generation
di: Gan, Yuan, et al.
Pubblicazione: (2025)
di: Gan, Yuan, et al.
Pubblicazione: (2025)
GaussianSpeech: Audio-Driven Gaussian Avatars
di: Aneja, Shivangi, et al.
Pubblicazione: (2024)
di: Aneja, Shivangi, et al.
Pubblicazione: (2024)
Leveraging Mamba with Full-Face Vision for Audio-Visual Speech Enhancement
di: Chao, Rong, et al.
Pubblicazione: (2025)
di: Chao, Rong, et al.
Pubblicazione: (2025)
Seeing Sound: Assembling Sounds from Visuals for Audio-to-Image Generation
di: Petermann, Darius, et al.
Pubblicazione: (2025)
di: Petermann, Darius, et al.
Pubblicazione: (2025)
TCDiff++: An End-to-end Trajectory-Controllable Diffusion Model for Harmonious Music-Driven Group Choreography
di: Dai, Yuqin, et al.
Pubblicazione: (2025)
di: Dai, Yuqin, et al.
Pubblicazione: (2025)
MACS: Multi-source Audio-to-image Generation with Contextual Significance and Semantic Alignment
di: Zhou, Hao, et al.
Pubblicazione: (2025)
di: Zhou, Hao, et al.
Pubblicazione: (2025)
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation
di: Kang, Fang, et al.
Pubblicazione: (2025)
di: Kang, Fang, et al.
Pubblicazione: (2025)
FürElise: Capturing and Physically Synthesizing Hand Motions of Piano Performance
di: Wang, Ruocheng, et al.
Pubblicazione: (2024)
di: Wang, Ruocheng, et al.
Pubblicazione: (2024)
Gaunt coefficients for complex and real spherical harmonics with applications to spherical array processing and Ambisonics
di: Politis, Archontis
Pubblicazione: (2024)
di: Politis, Archontis
Pubblicazione: (2024)
Documenti analoghi
-
Content and Style Aware Audio-Driven Facial Animation
di: Liu, Qingju, et al.
Pubblicazione: (2024) -
MATHDance: Mamba-Transformer Architecture with Uniform Tokenization for High-Quality 3D Dance Generation
di: Yang, Kaixing, et al.
Pubblicazione: (2025) -
Text-driven Talking Face Synthesis by Reprogramming Audio-driven Models
di: Choi, Jeongsoo, et al.
Pubblicazione: (2023) -
FaceTalk: Audio-Driven Motion Diffusion for Neural Parametric Head Models
di: Aneja, Shivangi, et al.
Pubblicazione: (2023) -
Audio2Face-3D: Audio-driven Realistic Facial Animation For Digital Avatars
di: NVIDIA, et al.
Pubblicazione: (2025)