3DXTalker: Unifying Identity, Lip Sync, Emotion, and Spatial Dynamics in Expressive 3D Talking Avatars
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Zhongju, Sun, Zhenhong, Wang, Beier, Wang, Yifu, Dong, Daoyi, Mo, Huadong, Li, Hongdong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
StoryBlender: Inter-Shot Consistent and Editable 3D Storyboard with Spatial-temporal Dynamics
di: Li, Bingliang, et al.
Pubblicazione: (2026)
di: Li, Bingliang, et al.
Pubblicazione: (2026)
T$^3$-S2S: Training-free Triplet Tuning for Sketch to Scene Synthesis in Controllable Concept Art Generation
di: Sun, Zhenhong, et al.
Pubblicazione: (2024)
di: Sun, Zhenhong, et al.
Pubblicazione: (2024)
JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync
di: Park, Sungjoon, et al.
Pubblicazione: (2025)
di: Park, Sungjoon, et al.
Pubblicazione: (2025)
Hierarchical and Step-Layer-Wise Tuning of Attention Specialty for Multi-Instance Synthesis in Diffusion Transformers
di: Zhang, Chunyang, et al.
Pubblicazione: (2025)
di: Zhang, Chunyang, et al.
Pubblicazione: (2025)
Text2Lip: Progressive Lip-Synced Talking Face Generation from Text via Viseme-Guided Rendering
di: Wang, Xu, et al.
Pubblicazione: (2025)
di: Wang, Xu, et al.
Pubblicazione: (2025)
GenSync: A Generalized Talking Head Framework for Audio-driven Multi-Subject Lip-Sync using 3D Gaussian Splatting
di: Agarwal, Anushka, et al.
Pubblicazione: (2025)
di: Agarwal, Anushka, et al.
Pubblicazione: (2025)
Make Your Actor Talk: Generalizable and High-Fidelity Lip Sync with Motion and Appearance Disentanglement
di: Yu, Runyi, et al.
Pubblicazione: (2024)
di: Yu, Runyi, et al.
Pubblicazione: (2024)
PASE: Phoneme-Aware Speech Encoder to Improve Lip Sync Accuracy for Talking Head Synthesis
di: Huang, Yihuan, et al.
Pubblicazione: (2025)
di: Huang, Yihuan, et al.
Pubblicazione: (2025)
3D Gaussian Head Avatars with Expressive Dynamic Appearances by Compact Tensorial Representations
di: Wang, Yating, et al.
Pubblicazione: (2025)
di: Wang, Yating, et al.
Pubblicazione: (2025)
AVI-Talking: Learning Audio-Visual Instructions for Expressive 3D Talking Face Generation
di: Sun, Yasheng, et al.
Pubblicazione: (2024)
di: Sun, Yasheng, et al.
Pubblicazione: (2024)
Exploring Phonetic Context-Aware Lip-Sync For Talking Face Generation
di: Park, Se Jin, et al.
Pubblicazione: (2023)
di: Park, Se Jin, et al.
Pubblicazione: (2023)
Sketch2Scene: Automatic Generation of Interactive 3D Game Scenes from User's Casual Sketches
di: Xu, Yongzhi, et al.
Pubblicazione: (2024)
di: Xu, Yongzhi, et al.
Pubblicazione: (2024)
SyncAnyone: Implicit Disentanglement via Progressive Self-Correction for Lip-Syncing in the wild
di: Zhang, Xindi, et al.
Pubblicazione: (2025)
di: Zhang, Xindi, et al.
Pubblicazione: (2025)
Removing Averaging: Personalized Lip-Sync Driven Characters Based on Identity Adapter
di: Zhu, Yanyu, et al.
Pubblicazione: (2025)
di: Zhu, Yanyu, et al.
Pubblicazione: (2025)
Towards High-fidelity 3D Talking Avatar with Personalized Dynamic Texture
di: Li, Xuanchen, et al.
Pubblicazione: (2025)
di: Li, Xuanchen, et al.
Pubblicazione: (2025)
Lips Are Lying: Spotting the Temporal Inconsistency between Audio and Visual in Lip-Syncing DeepFakes
di: Liu, Weifeng, et al.
Pubblicazione: (2024)
di: Liu, Weifeng, et al.
Pubblicazione: (2024)
LatentSync: Taming Audio-Conditioned Latent Diffusion Models for Lip Sync with SyncNet Supervision
di: Li, Chunyu, et al.
Pubblicazione: (2024)
di: Li, Chunyu, et al.
Pubblicazione: (2024)
Expressive Whole-Body 3D Gaussian Avatar
di: Moon, Gyeongsik, et al.
Pubblicazione: (2024)
di: Moon, Gyeongsik, et al.
Pubblicazione: (2024)
VisualSpeaker: Visually-Guided 3D Avatar Lip Synthesis
di: Symeonidis-Herzig, Alexandre, et al.
Pubblicazione: (2025)
di: Symeonidis-Herzig, Alexandre, et al.
Pubblicazione: (2025)
UniSync: Towards Generalizable and High-Fidelity Lip Synchronization for Challenging Scenarios
di: Fan, Ruidi, et al.
Pubblicazione: (2026)
di: Fan, Ruidi, et al.
Pubblicazione: (2026)
StyleLipSync: Style-based Personalized Lip-sync Video Generation
di: Ki, Taekyung, et al.
Pubblicazione: (2023)
di: Ki, Taekyung, et al.
Pubblicazione: (2023)
Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router
di: Huang, Yubo, et al.
Pubblicazione: (2025)
di: Huang, Yubo, et al.
Pubblicazione: (2025)
UNICA: A Unified Neural Framework for Controllable 3D Avatars
di: Zhu, Jiahe, et al.
Pubblicazione: (2026)
di: Zhu, Jiahe, et al.
Pubblicazione: (2026)
SyncTalk: The Devil is in the Synchronization for Talking Head Synthesis
di: Peng, Ziqiao, et al.
Pubblicazione: (2023)
di: Peng, Ziqiao, et al.
Pubblicazione: (2023)
Integrated Energy Management for Operational Cost Optimization in Community Microgrids
di: Uddin, Moslem, et al.
Pubblicazione: (2025)
di: Uddin, Moslem, et al.
Pubblicazione: (2025)
A review on modelling, evaluation, and optimization of cyber-physical system reliability
di: Uddin, Moslem, et al.
Pubblicazione: (2025)
di: Uddin, Moslem, et al.
Pubblicazione: (2025)
Real-Time Energy Management Strategies for Community Microgrids
di: Uddin, Moslem, et al.
Pubblicazione: (2025)
di: Uddin, Moslem, et al.
Pubblicazione: (2025)
Cost-Effective Design of Grid-tied Community Microgrid
di: Uddin, Moslem, et al.
Pubblicazione: (2025)
di: Uddin, Moslem, et al.
Pubblicazione: (2025)
Online Planning of Power Flows for Power Systems Against Bushfires Using Spatial Context
di: Xu, Jianyu, et al.
Pubblicazione: (2024)
di: Xu, Jianyu, et al.
Pubblicazione: (2024)
TaoAvatar: Real-Time Lifelike Full-Body Talking Avatars for Augmented Reality via 3D Gaussian Splatting
di: Chen, Jianchuan, et al.
Pubblicazione: (2025)
di: Chen, Jianchuan, et al.
Pubblicazione: (2025)
BioLip: Language-Generalizable Lip-Sync Deepfake Detection via Biomechanical Constraint Violation Modeling
di: Chen, Hao, et al.
Pubblicazione: (2026)
di: Chen, Hao, et al.
Pubblicazione: (2026)
Disentangle Identity, Cooperate Emotion: Correlation-Aware Emotional Talking Portrait Generation
di: Tan, Weipeng, et al.
Pubblicazione: (2025)
di: Tan, Weipeng, et al.
Pubblicazione: (2025)
EmotiveTalk: Expressive Talking Head Generation through Audio Information Decoupling and Emotional Video Diffusion
di: Wang, Haotian, et al.
Pubblicazione: (2024)
di: Wang, Haotian, et al.
Pubblicazione: (2024)
FreeTalk: Emotional Topology-Free 3D Talking Heads
di: Nocentini, Federico, et al.
Pubblicazione: (2026)
di: Nocentini, Federico, et al.
Pubblicazione: (2026)
Adaptive BESS and Grid Setpoints Optimization: A Model-Free Framework for Efficient Battery Management under Dynamic Tariff Pricing
di: Selim, Alaa, et al.
Pubblicazione: (2024)
di: Selim, Alaa, et al.
Pubblicazione: (2024)
FlashLips: 100-FPS Mask-Free Latent Lip-Sync using Reconstruction Instead of Diffusion or GANs
di: Zinonos, Andreas, et al.
Pubblicazione: (2025)
di: Zinonos, Andreas, et al.
Pubblicazione: (2025)
Supervising 3D Talking Head Avatars with Analysis-by-Audio-Synthesis
di: Daněček, Radek, et al.
Pubblicazione: (2025)
di: Daněček, Radek, et al.
Pubblicazione: (2025)
OmniSync: Towards Universal Lip Synchronization via Diffusion Transformers
di: Peng, Ziqiao, et al.
Pubblicazione: (2025)
di: Peng, Ziqiao, et al.
Pubblicazione: (2025)
VAST: Vivify Your Talking Avatar via Zero-Shot Expressive Facial Style Transfer
di: Chen, Liyang, et al.
Pubblicazione: (2023)
di: Chen, Liyang, et al.
Pubblicazione: (2023)
DreamWaltz-G: Expressive 3D Gaussian Avatars from Skeleton-Guided 2D Diffusion
di: Huang, Yukun, et al.
Pubblicazione: (2024)
di: Huang, Yukun, et al.
Pubblicazione: (2024)
Documenti analoghi
-
StoryBlender: Inter-Shot Consistent and Editable 3D Storyboard with Spatial-temporal Dynamics
di: Li, Bingliang, et al.
Pubblicazione: (2026) -
T$^3$-S2S: Training-free Triplet Tuning for Sketch to Scene Synthesis in Controllable Concept Art Generation
di: Sun, Zhenhong, et al.
Pubblicazione: (2024) -
JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync
di: Park, Sungjoon, et al.
Pubblicazione: (2025) -
Hierarchical and Step-Layer-Wise Tuning of Attention Specialty for Multi-Instance Synthesis in Diffusion Transformers
di: Zhang, Chunyang, et al.
Pubblicazione: (2025) -
Text2Lip: Progressive Lip-Synced Talking Face Generation from Text via Viseme-Guided Rendering
di: Wang, Xu, et al.
Pubblicazione: (2025)