Audio-Driven Talking Face Generation with Blink Embedding and Hash Grid Landmarks Encoding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yuhui, Yu, Hui, Liang, Wei, Zhang, Sunjie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SwapTalk: Audio-Driven Talking Face Generation with One-Shot Customization in Latent Space
von: Zhang, Zeren, et al.
Veröffentlicht: (2024)
von: Zhang, Zeren, et al.
Veröffentlicht: (2024)
VASA-1: Lifelike Audio-Driven Talking Faces Generated in Real Time
von: Xu, Sicheng, et al.
Veröffentlicht: (2024)
von: Xu, Sicheng, et al.
Veröffentlicht: (2024)
ATL-Diff: Audio-Driven Talking Head Generation with Early Landmarks-Guide Noise Diffusion
von: Vo, Hoang-Son, et al.
Veröffentlicht: (2025)
von: Vo, Hoang-Son, et al.
Veröffentlicht: (2025)
MAGIC-Talk: Motion-aware Audio-Driven Talking Face Generation with Customizable Identity Control
von: Nazarieh, Fatemeh, et al.
Veröffentlicht: (2025)
von: Nazarieh, Fatemeh, et al.
Veröffentlicht: (2025)
Audio-Driven Talking Face Video Generation with Joint Uncertainty Learning
von: Xie, Yifan, et al.
Veröffentlicht: (2025)
von: Xie, Yifan, et al.
Veröffentlicht: (2025)
PC-Talk: Precise Facial Animation Control for Audio-Driven Talking Face Generation
von: Wang, Baiqin, et al.
Veröffentlicht: (2025)
von: Wang, Baiqin, et al.
Veröffentlicht: (2025)
Exploiting Temporal Audio-Visual Correlation Embedding for Audio-Driven One-Shot Talking Head Animation
von: Xu, Zhihua, et al.
Veröffentlicht: (2025)
von: Xu, Zhihua, et al.
Veröffentlicht: (2025)
TurboTalk: Progressive Distillation for One-Step Audio-Driven Talking Avatar Generation
von: Liu, Xiangyu, et al.
Veröffentlicht: (2026)
von: Liu, Xiangyu, et al.
Veröffentlicht: (2026)
GSTalker: Real-time Audio-Driven Talking Face Generation via Deformable Gaussian Splatting
von: Chen, Bo, et al.
Veröffentlicht: (2024)
von: Chen, Bo, et al.
Veröffentlicht: (2024)
Audio-driven Talking Face Generation with Stabilized Synchronization Loss
von: Yaman, Dogucan, et al.
Veröffentlicht: (2023)
von: Yaman, Dogucan, et al.
Veröffentlicht: (2023)
AVI-Talking: Learning Audio-Visual Instructions for Expressive 3D Talking Face Generation
von: Sun, Yasheng, et al.
Veröffentlicht: (2024)
von: Sun, Yasheng, et al.
Veröffentlicht: (2024)
Talking Head Generation via AU-Guided Landmark Prediction
von: Chang, Shao-Yu, et al.
Veröffentlicht: (2025)
von: Chang, Shao-Yu, et al.
Veröffentlicht: (2025)
Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation
von: Kong, Zhe, et al.
Veröffentlicht: (2025)
von: Kong, Zhe, et al.
Veröffentlicht: (2025)
UniAvatar: Taming Lifelike Audio-Driven Talking Head Generation with Comprehensive Motion and Lighting Control
von: Sun, Wenzhang, et al.
Veröffentlicht: (2024)
von: Sun, Wenzhang, et al.
Veröffentlicht: (2024)
Text-Driven Emotionally Continuous Talking Face Generation
von: Yang, Hao, et al.
Veröffentlicht: (2026)
von: Yang, Hao, et al.
Veröffentlicht: (2026)
JoyGen: Audio-Driven 3D Depth-Aware Talking-Face Video Editing
von: Wang, Qili, et al.
Veröffentlicht: (2025)
von: Wang, Qili, et al.
Veröffentlicht: (2025)
Taming Transformer for Emotion-Controllable Talking Face Generation
von: Zhang, Ziqi, et al.
Veröffentlicht: (2025)
von: Zhang, Ziqi, et al.
Veröffentlicht: (2025)
Test-Time Self-Adaptive Conditioning for Stable Audio-Driven Talking-Head Generation
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2026)
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2026)
Audio-Visual Speech Representation Expert for Enhanced Talking Face Video Generation and Evaluation
von: Yaman, Dogucan, et al.
Veröffentlicht: (2024)
von: Yaman, Dogucan, et al.
Veröffentlicht: (2024)
JEAN: Joint Expression and Audio-guided NeRF-based Talking Face Generation
von: Chakkera, Sai Tanmay Reddy, et al.
Veröffentlicht: (2024)
von: Chakkera, Sai Tanmay Reddy, et al.
Veröffentlicht: (2024)
High-fidelity and Lip-synced Talking Face Synthesis via Landmark-based Diffusion Model
von: Zhong, Weizhi, et al.
Veröffentlicht: (2024)
von: Zhong, Weizhi, et al.
Veröffentlicht: (2024)
GaussianHeadTalk: Wobble-Free 3D Talking Heads with Audio Driven Gaussian Splatting
von: Agarwal, Madhav, et al.
Veröffentlicht: (2025)
von: Agarwal, Madhav, et al.
Veröffentlicht: (2025)
Takin-ADA: Emotion Controllable Audio-Driven Animation with Canonical and Landmark Loss Optimization
von: Lin, Bin, et al.
Veröffentlicht: (2024)
von: Lin, Bin, et al.
Veröffentlicht: (2024)
EmoFace: Emotion-Content Disentangled Speech-Driven 3D Talking Face Animation
von: Lin, Yihong, et al.
Veröffentlicht: (2024)
von: Lin, Yihong, et al.
Veröffentlicht: (2024)
KAN-Based Fusion of Dual-Domain for Audio-Driven Facial Landmarks Generation
von: Vo-Thanh, Hoang-Son, et al.
Veröffentlicht: (2024)
von: Vo-Thanh, Hoang-Son, et al.
Veröffentlicht: (2024)
Mask-Free Audio-driven Talking Face Generation for Enhanced Visual Quality and Identity Preservation
von: Yaman, Dogucan, et al.
Veröffentlicht: (2025)
von: Yaman, Dogucan, et al.
Veröffentlicht: (2025)
PortraitTalk: Towards Customizable One-Shot Audio-to-Talking Face Generation
von: Nazarieh, Fatemeh, et al.
Veröffentlicht: (2024)
von: Nazarieh, Fatemeh, et al.
Veröffentlicht: (2024)
Generalizable Face Landmarking Guided by Conditional Face Warping
von: Liang, Jiayi, et al.
Veröffentlicht: (2024)
von: Liang, Jiayi, et al.
Veröffentlicht: (2024)
Grid4D: 4D Decomposed Hash Encoding for High-Fidelity Dynamic Gaussian Splatting
von: Xu, Jiawei, et al.
Veröffentlicht: (2024)
von: Xu, Jiawei, et al.
Veröffentlicht: (2024)
SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation
von: Zhang, Wenli, et al.
Veröffentlicht: (2026)
von: Zhang, Wenli, et al.
Veröffentlicht: (2026)
EmbedTalk: Triplane-Free Talking Head Synthesis using Embedding-Driven Gaussian Deformation
von: Saggar, Arpita, et al.
Veröffentlicht: (2026)
von: Saggar, Arpita, et al.
Veröffentlicht: (2026)
Emotional Conversation: Empowering Talking Faces with Cohesive Expression, Gaze and Pose Generation
von: Liang, Jiadong, et al.
Veröffentlicht: (2024)
von: Liang, Jiadong, et al.
Veröffentlicht: (2024)
IMTalker: Efficient Audio-driven Talking Face Generation with Implicit Motion Transfer
von: Chen, Bo, et al.
Veröffentlicht: (2025)
von: Chen, Bo, et al.
Veröffentlicht: (2025)
Superior and Pragmatic Talking Face Generation with Teacher-Student Framework
von: Liang, Chao, et al.
Veröffentlicht: (2024)
von: Liang, Chao, et al.
Veröffentlicht: (2024)
DAE-Talker: High Fidelity Speech-Driven Talking Face Generation with Diffusion Autoencoder
von: Du, Chenpeng, et al.
Veröffentlicht: (2023)
von: Du, Chenpeng, et al.
Veröffentlicht: (2023)
OmniTalker: One-shot Real-time Text-Driven Talking Audio-Video Generation With Multimodal Style Mimicking
von: Wang, Zhongjian, et al.
Veröffentlicht: (2025)
von: Wang, Zhongjian, et al.
Veröffentlicht: (2025)
FaceChain-ImagineID: Freely Crafting High-Fidelity Diverse Talking Faces from Disentangled Audio
von: Xu, Chao, et al.
Veröffentlicht: (2024)
von: Xu, Chao, et al.
Veröffentlicht: (2024)
Context-aware Talking Face Video Generation
von: Xuanyuan, Meidai, et al.
Veröffentlicht: (2024)
von: Xuanyuan, Meidai, et al.
Veröffentlicht: (2024)
STGV: Spatio-Temporal Hash Encoding for Gaussian-based Video Representation
von: Lin, Jierun, et al.
Veröffentlicht: (2026)
von: Lin, Jierun, et al.
Veröffentlicht: (2026)
LaTo: Landmark-tokenized Diffusion Transformer for Fine-grained Human Face Editing
von: Zhang, Zhenghao, et al.
Veröffentlicht: (2025)
von: Zhang, Zhenghao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SwapTalk: Audio-Driven Talking Face Generation with One-Shot Customization in Latent Space
von: Zhang, Zeren, et al.
Veröffentlicht: (2024) -
VASA-1: Lifelike Audio-Driven Talking Faces Generated in Real Time
von: Xu, Sicheng, et al.
Veröffentlicht: (2024) -
ATL-Diff: Audio-Driven Talking Head Generation with Early Landmarks-Guide Noise Diffusion
von: Vo, Hoang-Son, et al.
Veröffentlicht: (2025) -
MAGIC-Talk: Motion-aware Audio-Driven Talking Face Generation with Customizable Identity Control
von: Nazarieh, Fatemeh, et al.
Veröffentlicht: (2025) -
Audio-Driven Talking Face Video Generation with Joint Uncertainty Learning
von: Xie, Yifan, et al.
Veröffentlicht: (2025)