Memories are One-to-Many Mapping Alleviators in Talking Face Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tang, Anni, He, Tianyu, Tan, Xu, Ling, Jun, Song, Li |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DAE-Talker: High Fidelity Speech-Driven Talking Face Generation with Diffusion Autoencoder
von: Du, Chenpeng, et al.
Veröffentlicht: (2023)
von: Du, Chenpeng, et al.
Veröffentlicht: (2023)
PoseTalk: Text-and-Audio-based Pose Control and Motion Refinement for One-Shot Talking Head Generation
von: Ling, Jun, et al.
Veröffentlicht: (2024)
von: Ling, Jun, et al.
Veröffentlicht: (2024)
SegTalker: Segmentation-based Talking Face Generation with Mask-guided Local Editing
von: Xiong, Lingyu, et al.
Veröffentlicht: (2024)
von: Xiong, Lingyu, et al.
Veröffentlicht: (2024)
GAIA: Zero-shot Talking Avatar Generation
von: He, Tianyu, et al.
Veröffentlicht: (2023)
von: He, Tianyu, et al.
Veröffentlicht: (2023)
OpFlowTalker: Realistic and Natural Talking Face Generation via Optical Flow Guidance
von: Ge, Shuheng, et al.
Veröffentlicht: (2024)
von: Ge, Shuheng, et al.
Veröffentlicht: (2024)
TurboTalk: Progressive Distillation for One-Step Audio-Driven Talking Avatar Generation
von: Liu, Xiangyu, et al.
Veröffentlicht: (2026)
von: Liu, Xiangyu, et al.
Veröffentlicht: (2026)
G4G:A Generic Framework for High Fidelity Talking Face Generation with Fine-grained Intra-modal Alignment
von: Zhang, Juan, et al.
Veröffentlicht: (2024)
von: Zhang, Juan, et al.
Veröffentlicht: (2024)
Rate-aware Compression for NeRF-based Volumetric Video
von: Zhang, Zhiyu, et al.
Veröffentlicht: (2024)
von: Zhang, Zhiyu, et al.
Veröffentlicht: (2024)
FD2Talk: Towards Generalized Talking Head Generation with Facial Decoupled Diffusion Model
von: Yao, Ziyu, et al.
Veröffentlicht: (2024)
von: Yao, Ziyu, et al.
Veröffentlicht: (2024)
NeRF-AD: Neural Radiance Field with Attention-based Disentanglement for Talking Face Synthesis
von: Bi, Chongke, et al.
Veröffentlicht: (2024)
von: Bi, Chongke, et al.
Veröffentlicht: (2024)
UniTalking: A Unified Audio-Video Framework for Talking Portrait Generation
von: Li, Hebeizi, et al.
Veröffentlicht: (2026)
von: Li, Hebeizi, et al.
Veröffentlicht: (2026)
MotionDreamer: One-to-Many Motion Synthesis with Localized Generative Masked Transformer
von: Wang, Yilin, et al.
Veröffentlicht: (2025)
von: Wang, Yilin, et al.
Veröffentlicht: (2025)
Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation
von: Kang, Fang, et al.
Veröffentlicht: (2025)
von: Kang, Fang, et al.
Veröffentlicht: (2025)
M2ORT: Many-To-One Regression Transformer for Spatial Transcriptomics Prediction from Histopathology Images
von: Wang, Hongyi, et al.
Veröffentlicht: (2024)
von: Wang, Hongyi, et al.
Veröffentlicht: (2024)
OneHOI: Unifying Human-Object Interaction Generation and Editing
von: Hoe, Jiun Tian, et al.
Veröffentlicht: (2026)
von: Hoe, Jiun Tian, et al.
Veröffentlicht: (2026)
High-fidelity and Lip-synced Talking Face Synthesis via Landmark-based Diffusion Model
von: Zhong, Weizhi, et al.
Veröffentlicht: (2024)
von: Zhong, Weizhi, et al.
Veröffentlicht: (2024)
QGFace: Quality-Guided Joint Training For Mixed-Quality Face Recognition
von: Song, Youzhe, et al.
Veröffentlicht: (2023)
von: Song, Youzhe, et al.
Veröffentlicht: (2023)
EditEmoTalk: Controllable Speech-Driven 3D Facial Animation with Continuous Expression Editing
von: Jiang, Diqiong, et al.
Veröffentlicht: (2026)
von: Jiang, Diqiong, et al.
Veröffentlicht: (2026)
One Size, Many Fits: Aligning Diverse Group-Wise Click Preferences in Large-Scale Advertising Image Generation
von: Lu, Shuo, et al.
Veröffentlicht: (2026)
von: Lu, Shuo, et al.
Veröffentlicht: (2026)
Talking Head Generation Driven by Speech-Related Facial Action Units and Audio- Based on Multimodal Representation Fusion
von: Chen, Sen, et al.
Veröffentlicht: (2022)
von: Chen, Sen, et al.
Veröffentlicht: (2022)
Memory-Guided View Refinement for Dynamic Human-in-the-loop EQA
von: Lu, Xin, et al.
Veröffentlicht: (2026)
von: Lu, Xin, et al.
Veröffentlicht: (2026)
GaussianTalker: Speaker-specific Talking Head Synthesis via 3D Gaussian Splatting
von: Yu, Hongyun, et al.
Veröffentlicht: (2024)
von: Yu, Hongyun, et al.
Veröffentlicht: (2024)
Dual Mutual Learning Network with Global-local Awareness for RGB-D Salient Object Detection
von: Yi, Kang, et al.
Veröffentlicht: (2025)
von: Yi, Kang, et al.
Veröffentlicht: (2025)
Generalized Face Forgery Detection via Adaptive Learning for Pre-trained Vision Transformer
von: Luo, Anwei, et al.
Veröffentlicht: (2023)
von: Luo, Anwei, et al.
Veröffentlicht: (2023)
HOP: Heterogeneous Topology-based Multimodal Entanglement for Co-Speech Gesture Generation
von: Cheng, Hongye, et al.
Veröffentlicht: (2025)
von: Cheng, Hongye, et al.
Veröffentlicht: (2025)
TOL: Textual Localization with OpenStreetMap
von: Liao, Youqi, et al.
Veröffentlicht: (2026)
von: Liao, Youqi, et al.
Veröffentlicht: (2026)
Divide-and-Conquer: Confluent Triple-Flow Network for RGB-T Salient Object Detection
von: Tang, Hao, et al.
Veröffentlicht: (2024)
von: Tang, Hao, et al.
Veröffentlicht: (2024)
Face Consistency Benchmark for GenAI Video
von: Podstawski, Michal, et al.
Veröffentlicht: (2025)
von: Podstawski, Michal, et al.
Veröffentlicht: (2025)
Reference-Guided Identity Preserving Face Restoration
von: Zhou, Mo, et al.
Veröffentlicht: (2025)
von: Zhou, Mo, et al.
Veröffentlicht: (2025)
M$^3$Face: A Unified Multi-Modal Multilingual Framework for Human Face Generation and Editing
von: Mofayezi, Mohammadreza, et al.
Veröffentlicht: (2024)
von: Mofayezi, Mohammadreza, et al.
Veröffentlicht: (2024)
HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning
von: Chen, Liyang, et al.
Veröffentlicht: (2025)
von: Chen, Liyang, et al.
Veröffentlicht: (2025)
EARTalking: End-to-end GPT-style Autoregressive Talking Head Synthesis with Frame-wise Control
von: Weng, Yuzhe, et al.
Veröffentlicht: (2026)
von: Weng, Yuzhe, et al.
Veröffentlicht: (2026)
StableDub: Taming Diffusion Prior for Generalized and Efficient Visual Dubbing
von: Chen, Liyang, et al.
Veröffentlicht: (2025)
von: Chen, Liyang, et al.
Veröffentlicht: (2025)
LAPIG: Language Guided Projector Image Generation with Surface Adaptation and Stylization
von: Deng, Yuchen, et al.
Veröffentlicht: (2025)
von: Deng, Yuchen, et al.
Veröffentlicht: (2025)
Kubrick: Multimodal Agent Collaborations for Synthetic Video Generation
von: He, Liu, et al.
Veröffentlicht: (2024)
von: He, Liu, et al.
Veröffentlicht: (2024)
Error Analyses of Auto-Regressive Video Diffusion Models: A Unified Framework
von: Wang, Jing, et al.
Veröffentlicht: (2025)
von: Wang, Jing, et al.
Veröffentlicht: (2025)
Single Image Dehazing Using Scene Depth Ordering
von: Ling, Pengyang, et al.
Veröffentlicht: (2024)
von: Ling, Pengyang, et al.
Veröffentlicht: (2024)
M3FAS: An Accurate and Robust MultiModal Mobile Face Anti-Spoofing System
von: Kong, Chenqi, et al.
Veröffentlicht: (2023)
von: Kong, Chenqi, et al.
Veröffentlicht: (2023)
Group-based Distinctive Image Captioning with Memory Difference Encoding and Attention
von: Wang, Jiuniu, et al.
Veröffentlicht: (2025)
von: Wang, Jiuniu, et al.
Veröffentlicht: (2025)
One Framework to Rule Them All: Unifying Multimodal Tasks with LLM Neural-Tuning
von: Sun, Hao, et al.
Veröffentlicht: (2024)
von: Sun, Hao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
DAE-Talker: High Fidelity Speech-Driven Talking Face Generation with Diffusion Autoencoder
von: Du, Chenpeng, et al.
Veröffentlicht: (2023) -
PoseTalk: Text-and-Audio-based Pose Control and Motion Refinement for One-Shot Talking Head Generation
von: Ling, Jun, et al.
Veröffentlicht: (2024) -
SegTalker: Segmentation-based Talking Face Generation with Mask-guided Local Editing
von: Xiong, Lingyu, et al.
Veröffentlicht: (2024) -
GAIA: Zero-shot Talking Avatar Generation
von: He, Tianyu, et al.
Veröffentlicht: (2023) -
OpFlowTalker: Realistic and Natural Talking Face Generation via Optical Flow Guidance
von: Ge, Shuheng, et al.
Veröffentlicht: (2024)