UniAvatar: Taming Lifelike Audio-Driven Talking Head Generation with Comprehensive Motion and Lighting Control
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Wenzhang, Li, Xiang, Di, Donglin, Liang, Zhuding, Zhang, Qiyuan, Li, Hao, Chen, Wei, Cui, Jianxun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
UniCP: A Unified Caching and Pruning Framework for Efficient Video Generation
von: Sun, Wenzhang, et al.
Veröffentlicht: (2025)
von: Sun, Wenzhang, et al.
Veröffentlicht: (2025)
VASA-3D: Lifelike Audio-Driven Gaussian Head Avatars from a Single Image
von: Xu, Sicheng, et al.
Veröffentlicht: (2025)
von: Xu, Sicheng, et al.
Veröffentlicht: (2025)
VASA-1: Lifelike Audio-Driven Talking Faces Generated in Real Time
von: Xu, Sicheng, et al.
Veröffentlicht: (2024)
von: Xu, Sicheng, et al.
Veröffentlicht: (2024)
Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion Dependency
von: Jiang, Jianwen, et al.
Veröffentlicht: (2024)
von: Jiang, Jianwen, et al.
Veröffentlicht: (2024)
RealTalk: Realistic Emotion-Aware Lifelike Talking-Head Synthesis
von: Wang, Wenqing, et al.
Veröffentlicht: (2025)
von: Wang, Wenqing, et al.
Veröffentlicht: (2025)
Hi-VAE: Efficient Video Autoencoding with Global and Detailed Motion
von: Liu, Huaize, et al.
Veröffentlicht: (2025)
von: Liu, Huaize, et al.
Veröffentlicht: (2025)
A Self-supervised Motion Representation for Portrait Video Generation
von: Zhang, Qiyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Qiyuan, et al.
Veröffentlicht: (2025)
MoEE: Mixture of Emotion Experts for Audio-Driven Portrait Animation
von: Liu, Huaize, et al.
Veröffentlicht: (2025)
von: Liu, Huaize, et al.
Veröffentlicht: (2025)
UniTalking: A Unified Audio-Video Framework for Talking Portrait Generation
von: Li, Hebeizi, et al.
Veröffentlicht: (2026)
von: Li, Hebeizi, et al.
Veröffentlicht: (2026)
TaoAvatar: Real-Time Lifelike Full-Body Talking Avatars for Augmented Reality via 3D Gaussian Splatting
von: Chen, Jianchuan, et al.
Veröffentlicht: (2025)
von: Chen, Jianchuan, et al.
Veröffentlicht: (2025)
PoseTalk: Text-and-Audio-based Pose Control and Motion Refinement for One-Shot Talking Head Generation
von: Ling, Jun, et al.
Veröffentlicht: (2024)
von: Ling, Jun, et al.
Veröffentlicht: (2024)
Supervising 3D Talking Head Avatars with Analysis-by-Audio-Synthesis
von: Daněček, Radek, et al.
Veröffentlicht: (2025)
von: Daněček, Radek, et al.
Veröffentlicht: (2025)
STGAtt: A Spatial-Temporal Unified Graph Attention Network for Traffic Flow Forecasting
von: Liang, Zhuding, et al.
Veröffentlicht: (2025)
von: Liang, Zhuding, et al.
Veröffentlicht: (2025)
FaceTalk: Audio-Driven Motion Diffusion for Neural Parametric Head Models
von: Aneja, Shivangi, et al.
Veröffentlicht: (2023)
von: Aneja, Shivangi, et al.
Veröffentlicht: (2023)
Audio-Driven Universal Gaussian Head Avatars
von: Teotia, Kartik, et al.
Veröffentlicht: (2025)
von: Teotia, Kartik, et al.
Veröffentlicht: (2025)
Development of the Lifelike Head Unit for a Humanoid Cybernetic Avatar `Yui' and Its Operation Interface
von: Nakajima, Mizuki, et al.
Veröffentlicht: (2023)
von: Nakajima, Mizuki, et al.
Veröffentlicht: (2023)
THGS: Lifelike Talking Human Avatar Synthesis From Monocular Video Via 3D Gaussian Splatting
von: Chuang Chen, et al.
Veröffentlicht: (2025)
von: Chuang Chen, et al.
Veröffentlicht: (2025)
Making Avatars Interact: Towards Text-Driven Human-Object Interaction for Controllable Talking Avatars
von: Zhang, Youliang, et al.
Veröffentlicht: (2026)
von: Zhang, Youliang, et al.
Veröffentlicht: (2026)
TurboTalk: Progressive Distillation for One-Step Audio-Driven Talking Avatar Generation
von: Liu, Xiangyu, et al.
Veröffentlicht: (2026)
von: Liu, Xiangyu, et al.
Veröffentlicht: (2026)
DrivingScene: A Multi-Task Online Feed-Forward 3D Gaussian Splatting Method for Dynamic Driving Scenes
von: Hou, Qirui, et al.
Veröffentlicht: (2025)
von: Hou, Qirui, et al.
Veröffentlicht: (2025)
PAGS: Priority-Adaptive Gaussian Splatting for Dynamic Driving Scenes
von: A, Ying, et al.
Veröffentlicht: (2025)
von: A, Ying, et al.
Veröffentlicht: (2025)
EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditions
von: Chen, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Chen, Zhiyuan, et al.
Veröffentlicht: (2024)
GeoDiff4D: Geometry-Aware Diffusion for 4D Head Avatar Reconstruction
von: Xu, Chao, et al.
Veröffentlicht: (2026)
von: Xu, Chao, et al.
Veröffentlicht: (2026)
ConsistentAvatar: Learning to Diffuse Fully Consistent Talking Head Avatar with Temporal Guidance
von: Yang, Haijie, et al.
Veröffentlicht: (2024)
von: Yang, Haijie, et al.
Veröffentlicht: (2024)
MAGIC-Talk: Motion-aware Audio-Driven Talking Face Generation with Customizable Identity Control
von: Nazarieh, Fatemeh, et al.
Veröffentlicht: (2025)
von: Nazarieh, Fatemeh, et al.
Veröffentlicht: (2025)
ALIVE: Animate Your World with Lifelike Audio-Video Generation
von: Guo, Ying, et al.
Veröffentlicht: (2026)
von: Guo, Ying, et al.
Veröffentlicht: (2026)
GaussianHeadTalk: Wobble-Free 3D Talking Heads with Audio Driven Gaussian Splatting
von: Agarwal, Madhav, et al.
Veröffentlicht: (2025)
von: Agarwal, Madhav, et al.
Veröffentlicht: (2025)
TalkVid: A Large-Scale Diversified Dataset for Audio-Driven Talking Head Synthesis
von: Chen, Shunian, et al.
Veröffentlicht: (2025)
von: Chen, Shunian, et al.
Veröffentlicht: (2025)
UniLS: End-to-End Audio-Driven Avatars for Unified Listening and Speaking
von: Chu, Xuangeng, et al.
Veröffentlicht: (2025)
von: Chu, Xuangeng, et al.
Veröffentlicht: (2025)
Total-Editing: Head Avatar with Editable Appearance, Motion, and Lighting
von: Zhao, Yizhou, et al.
Veröffentlicht: (2025)
von: Zhao, Yizhou, et al.
Veröffentlicht: (2025)
Exploiting Temporal Audio-Visual Correlation Embedding for Audio-Driven One-Shot Talking Head Animation
von: Xu, Zhihua, et al.
Veröffentlicht: (2025)
von: Xu, Zhihua, et al.
Veröffentlicht: (2025)
Ditto: Motion-Space Diffusion for Controllable Realtime Talking Head Synthesis
von: Li, Tianqi, et al.
Veröffentlicht: (2024)
von: Li, Tianqi, et al.
Veröffentlicht: (2024)
PC-Talk: Precise Facial Animation Control for Audio-Driven Talking Face Generation
von: Wang, Baiqin, et al.
Veröffentlicht: (2025)
von: Wang, Baiqin, et al.
Veröffentlicht: (2025)
EGSTalker: Real-Time Audio-Driven Talking Head Generation with Efficient Gaussian Deformation
von: Zhu, Tianheng, et al.
Veröffentlicht: (2025)
von: Zhu, Tianheng, et al.
Veröffentlicht: (2025)
Audio-Visual Driven Compression for Low-Bitrate Talking Head Videos
von: Takahashi, Riku, et al.
Veröffentlicht: (2025)
von: Takahashi, Riku, et al.
Veröffentlicht: (2025)
CyberHost: Taming Audio-driven Avatar Diffusion Model with Region Codebook Attention
von: Lin, Gaojie, et al.
Veröffentlicht: (2024)
von: Lin, Gaojie, et al.
Veröffentlicht: (2024)
LightAvatar: Efficient Head Avatar as Dynamic Neural Light Field
von: Wang, Huan, et al.
Veröffentlicht: (2024)
von: Wang, Huan, et al.
Veröffentlicht: (2024)
Avatar Fingerprinting for Authorized Use of Synthetic Talking-Head Videos
von: Prashnani, Ekta, et al.
Veröffentlicht: (2023)
von: Prashnani, Ekta, et al.
Veröffentlicht: (2023)
ActAvatar: Temporally-Aware Precise Action Control for Talking Avatars
von: Peng, Ziqiao, et al.
Veröffentlicht: (2025)
von: Peng, Ziqiao, et al.
Veröffentlicht: (2025)
FixTalk: Taming Identity Leakage for High-Quality Talking Head Generation in Extreme Cases
von: Tan, Shuai, et al.
Veröffentlicht: (2025)
von: Tan, Shuai, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
UniCP: A Unified Caching and Pruning Framework for Efficient Video Generation
von: Sun, Wenzhang, et al.
Veröffentlicht: (2025) -
VASA-3D: Lifelike Audio-Driven Gaussian Head Avatars from a Single Image
von: Xu, Sicheng, et al.
Veröffentlicht: (2025) -
VASA-1: Lifelike Audio-Driven Talking Faces Generated in Real Time
von: Xu, Sicheng, et al.
Veröffentlicht: (2024) -
Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion Dependency
von: Jiang, Jianwen, et al.
Veröffentlicht: (2024) -
RealTalk: Realistic Emotion-Aware Lifelike Talking-Head Synthesis
von: Wang, Wenqing, et al.
Veröffentlicht: (2025)