FaceChain-ImagineID: Freely Crafting High-Fidelity Diverse Talking Faces from Disentangled Audio
Fuente:
arXiv
Salvato in:
| Autori principali: | Xu, Chao, Liu, Yang, Xing, Jiazheng, Wang, Weida, Sun, Mingze, Dan, Jun, Huang, Tianxin, Li, Siyuan, Cheng, Zhi-Qi, Tai, Ying, Sun, Baigui |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FaceChain-FACT: Face Adapter with Decoupled Training for Identity-preserved Personalization
di: Yu, Cheng, et al.
Pubblicazione: (2024)
di: Yu, Cheng, et al.
Pubblicazione: (2024)
FaceChain-SuDe: Building Derived Class to Inherit Category Attributes for One-shot Subject-Driven Generation
di: Qiao, Pengchong, et al.
Pubblicazione: (2024)
di: Qiao, Pengchong, et al.
Pubblicazione: (2024)
TransFace++: Rethinking the Face Recognition Paradigm with a Focus on Accuracy, Efficiency, and Security
di: Dan, Jun, et al.
Pubblicazione: (2023)
di: Dan, Jun, et al.
Pubblicazione: (2023)
Beyond Talking -- Generating Holistic 3D Human Dyadic Motion for Communication
di: Sun, Mingze, et al.
Pubblicazione: (2024)
di: Sun, Mingze, et al.
Pubblicazione: (2024)
TopoFR: A Closer Look at Topology Alignment on Face Recognition
di: Dan, Jun, et al.
Pubblicazione: (2024)
di: Dan, Jun, et al.
Pubblicazione: (2024)
AVI-Talking: Learning Audio-Visual Instructions for Expressive 3D Talking Face Generation
di: Sun, Yasheng, et al.
Pubblicazione: (2024)
di: Sun, Yasheng, et al.
Pubblicazione: (2024)
TryOn-Adapter: Efficient Fine-Grained Clothing Identity Adaptation for High-Fidelity Virtual Try-On
di: Xing, Jiazheng, et al.
Pubblicazione: (2024)
di: Xing, Jiazheng, et al.
Pubblicazione: (2024)
EmoFace: Emotion-Content Disentangled Speech-Driven 3D Talking Face Animation
di: Lin, Yihong, et al.
Pubblicazione: (2024)
di: Lin, Yihong, et al.
Pubblicazione: (2024)
Learning to Decouple the Lights for 3D Face Texture Modeling
di: Huang, Tianxin, et al.
Pubblicazione: (2024)
di: Huang, Tianxin, et al.
Pubblicazione: (2024)
PortraitTalk: Towards Customizable One-Shot Audio-to-Talking Face Generation
di: Nazarieh, Fatemeh, et al.
Pubblicazione: (2024)
di: Nazarieh, Fatemeh, et al.
Pubblicazione: (2024)
Audio-driven Talking Face Generation with Stabilized Synchronization Loss
di: Yaman, Dogucan, et al.
Pubblicazione: (2023)
di: Yaman, Dogucan, et al.
Pubblicazione: (2023)
Spatially and Temporally Optimized Audio‐Driven Talking Face Generation
di: Biao Dong, et al.
Pubblicazione: (2024)
di: Biao Dong, et al.
Pubblicazione: (2024)
DisentTalk: Cross-lingual Talking Face Generation via Semantic Disentangled Diffusion Model
di: Liu, Kangwei, et al.
Pubblicazione: (2025)
di: Liu, Kangwei, et al.
Pubblicazione: (2025)
DreamID: High-Fidelity and Fast diffusion-based Face Swapping via Triplet ID Group Learning
di: Ye, Fulong, et al.
Pubblicazione: (2025)
di: Ye, Fulong, et al.
Pubblicazione: (2025)
PC-Talk: Precise Facial Animation Control for Audio-Driven Talking Face Generation
di: Wang, Baiqin, et al.
Pubblicazione: (2025)
di: Wang, Baiqin, et al.
Pubblicazione: (2025)
Faces that Speak: Jointly Synthesising Talking Face and Speech from Text
di: Jang, Youngjoon, et al.
Pubblicazione: (2024)
di: Jang, Youngjoon, et al.
Pubblicazione: (2024)
Text-Driven Emotionally Continuous Talking Face Generation
di: Yang, Hao, et al.
Pubblicazione: (2026)
di: Yang, Hao, et al.
Pubblicazione: (2026)
Text-driven Talking Face Synthesis by Reprogramming Audio-driven Models
di: Choi, Jeongsoo, et al.
Pubblicazione: (2023)
di: Choi, Jeongsoo, et al.
Pubblicazione: (2023)
Audio-Driven Talking Face Video Generation with Joint Uncertainty Learning
di: Xie, Yifan, et al.
Pubblicazione: (2025)
di: Xie, Yifan, et al.
Pubblicazione: (2025)
SwapTalk: Audio-Driven Talking Face Generation with One-Shot Customization in Latent Space
di: Zhang, Zeren, et al.
Pubblicazione: (2024)
di: Zhang, Zeren, et al.
Pubblicazione: (2024)
MAGIC-Talk: Motion-aware Audio-Driven Talking Face Generation with Customizable Identity Control
di: Nazarieh, Fatemeh, et al.
Pubblicazione: (2025)
di: Nazarieh, Fatemeh, et al.
Pubblicazione: (2025)
Learn2Talk: 3D Talking Face Learns from 2D Talking Face
di: Zhuang, Yixiang, et al.
Pubblicazione: (2024)
di: Zhuang, Yixiang, et al.
Pubblicazione: (2024)
High-Fidelity Diffusion Face Swapping with ID-Constrained Facial Conditioning
di: He, Dailan, et al.
Pubblicazione: (2025)
di: He, Dailan, et al.
Pubblicazione: (2025)
AlphaFace: High Fidelity and Real-time Face Swapper Robust to Facial Pose
di: Yu, Jongmin, et al.
Pubblicazione: (2026)
di: Yu, Jongmin, et al.
Pubblicazione: (2026)
FaceID-6M: A Large-Scale, Open-Source FaceID Customization Dataset
di: Wang, Shuhe, et al.
Pubblicazione: (2025)
di: Wang, Shuhe, et al.
Pubblicazione: (2025)
Learning Disentangled Speech- and Expression-Driven Blendshapes for 3D Talking Face Animation
di: Mao, Yuxiang, et al.
Pubblicazione: (2025)
di: Mao, Yuxiang, et al.
Pubblicazione: (2025)
FaceTalk: Audio-Driven Motion Diffusion for Neural Parametric Head Models
di: Aneja, Shivangi, et al.
Pubblicazione: (2023)
di: Aneja, Shivangi, et al.
Pubblicazione: (2023)
IMTalker: Efficient Audio-driven Talking Face Generation with Implicit Motion Transfer
di: Chen, Bo, et al.
Pubblicazione: (2025)
di: Chen, Bo, et al.
Pubblicazione: (2025)
VASA-1: Lifelike Audio-Driven Talking Faces Generated in Real Time
di: Xu, Sicheng, et al.
Pubblicazione: (2024)
di: Xu, Sicheng, et al.
Pubblicazione: (2024)
See the Speaker: Crafting High-Resolution Talking Faces from Speech with Prior Guidance and Region Refinement
di: Wang, Jinting, et al.
Pubblicazione: (2025)
di: Wang, Jinting, et al.
Pubblicazione: (2025)
DAE-Talker: High Fidelity Speech-Driven Talking Face Generation with Diffusion Autoencoder
di: Du, Chenpeng, et al.
Pubblicazione: (2023)
di: Du, Chenpeng, et al.
Pubblicazione: (2023)
Combo: Co-speech holistic 3D human motion generation and efficient customizable adaptation in harmony
di: Xu, Chao, et al.
Pubblicazione: (2024)
di: Xu, Chao, et al.
Pubblicazione: (2024)
Meta-Learning Empowered Meta-Face: Personalized Speaking Style Adaptation for Audio-Driven 3D Talking Face Animation
di: Zhou, Xukun, et al.
Pubblicazione: (2024)
di: Zhou, Xukun, et al.
Pubblicazione: (2024)
DreamID-V:Bridging the Image-to-Video Gap for High-Fidelity Face Swapping via Diffusion Transformer
di: Guo, Xu, et al.
Pubblicazione: (2026)
di: Guo, Xu, et al.
Pubblicazione: (2026)
NeRF-AD: Neural Radiance Field with Attention-based Disentanglement for Talking Face Synthesis
di: Bi, Chongke, et al.
Pubblicazione: (2024)
di: Bi, Chongke, et al.
Pubblicazione: (2024)
Audio-Visual Speech Representation Expert for Enhanced Talking Face Video Generation and Evaluation
di: Yaman, Dogucan, et al.
Pubblicazione: (2024)
di: Yaman, Dogucan, et al.
Pubblicazione: (2024)
Audio-Driven Talking Face Generation with Blink Embedding and Hash Grid Landmarks Encoding
di: Zhang, Yuhui, et al.
Pubblicazione: (2026)
di: Zhang, Yuhui, et al.
Pubblicazione: (2026)
SyncLipMAE: Contrastive Masked Pretraining for Audio-Visual Talking-Face Representation
di: Ling, Zeyu, et al.
Pubblicazione: (2025)
di: Ling, Zeyu, et al.
Pubblicazione: (2025)
JEAN: Joint Expression and Audio-guided NeRF-based Talking Face Generation
di: Chakkera, Sai Tanmay Reddy, et al.
Pubblicazione: (2024)
di: Chakkera, Sai Tanmay Reddy, et al.
Pubblicazione: (2024)
Lightweight High-Fidelity Low-Bitrate Talking Face Compression for 3D Video Conference
di: Li, Jianglong, et al.
Pubblicazione: (2026)
di: Li, Jianglong, et al.
Pubblicazione: (2026)
Documenti analoghi
-
FaceChain-FACT: Face Adapter with Decoupled Training for Identity-preserved Personalization
di: Yu, Cheng, et al.
Pubblicazione: (2024) -
FaceChain-SuDe: Building Derived Class to Inherit Category Attributes for One-shot Subject-Driven Generation
di: Qiao, Pengchong, et al.
Pubblicazione: (2024) -
TransFace++: Rethinking the Face Recognition Paradigm with a Focus on Accuracy, Efficiency, and Security
di: Dan, Jun, et al.
Pubblicazione: (2023) -
Beyond Talking -- Generating Holistic 3D Human Dyadic Motion for Communication
di: Sun, Mingze, et al.
Pubblicazione: (2024) -
TopoFR: A Closer Look at Topology Alignment on Face Recognition
di: Dan, Jun, et al.
Pubblicazione: (2024)