FacEDiT: Unified Talking Face Editing and Generation via Facial Motion Infilling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sung-Bin, Kim, Chang, Joohyun, Harwath, David, Oh, Tae-Hyun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MultiTalk: Enhancing 3D Talking Head Generation Across Languages with Multilingual Video Dataset
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2024)
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2024)
Perceptually Accurate 3D Talking Head Generation: New Definitions, Speech-Mesh Representation, and Evaluation Metrics
von: Chae-Yeon, Lee, et al.
Veröffentlicht: (2025)
von: Chae-Yeon, Lee, et al.
Veröffentlicht: (2025)
Controllable Talking Face Generation by Implicit Facial Keypoints Editing
von: Zhao, Dong, et al.
Veröffentlicht: (2024)
von: Zhao, Dong, et al.
Veröffentlicht: (2024)
UP-FacE: User-predictable Fine-grained Face Shape Editing
von: Strohm, Florian, et al.
Veröffentlicht: (2024)
von: Strohm, Florian, et al.
Veröffentlicht: (2024)
VoiceCraft-Dub: Automated Video Dubbing with Neural Codec Language Models
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2025)
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2025)
FacEnhance: Facial Expression Enhancing with Recurrent DDPMs
von: Bouzid, Hamza, et al.
Veröffentlicht: (2024)
von: Bouzid, Hamza, et al.
Veröffentlicht: (2024)
FaceEditTalker: Controllable Talking Head Generation with Facial Attribute Editing
von: Feng, Guanwen, et al.
Veröffentlicht: (2025)
von: Feng, Guanwen, et al.
Veröffentlicht: (2025)
Enhancing Speech-Driven 3D Facial Animation with Audio-Visual Guidance from Lip Reading Expert
von: EunGi, Han, et al.
Veröffentlicht: (2024)
von: EunGi, Han, et al.
Veröffentlicht: (2024)
IF-MDM: Implicit Face Motion Diffusion Model for High-Fidelity Realtime Talking Head Generation
von: Yang, Sejong, et al.
Veröffentlicht: (2024)
von: Yang, Sejong, et al.
Veröffentlicht: (2024)
Face2Parts: Exploring Coarse-to-Fine Inter-Regional Facial Dependencies for Generalized Deepfake Detection
von: Uddin, Kutub, et al.
Veröffentlicht: (2026)
von: Uddin, Kutub, et al.
Veröffentlicht: (2026)
SoundBrush: Sound as a Brush for Visual Scene Editing
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2024)
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2024)
Instruct-4DGS: Efficient Dynamic Scene Editing via 4D Gaussian-based Static-Dynamic Separation
von: Kwon, Joohyun, et al.
Veröffentlicht: (2025)
von: Kwon, Joohyun, et al.
Veröffentlicht: (2025)
Revisiting Learning-based Video Motion Magnification for Real-time Processing
von: Ha, Hyunwoo, et al.
Veröffentlicht: (2024)
von: Ha, Hyunwoo, et al.
Veröffentlicht: (2024)
Learning-based Axial Video Motion Magnification
von: Byung-Ki, Kwon, et al.
Veröffentlicht: (2023)
von: Byung-Ki, Kwon, et al.
Veröffentlicht: (2023)
AVHBench: A Cross-Modal Hallucination Benchmark for Audio-Visual Large Language Models
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2024)
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2024)
PC-Talk: Precise Facial Animation Control for Audio-Driven Talking Face Generation
von: Wang, Baiqin, et al.
Veröffentlicht: (2025)
von: Wang, Baiqin, et al.
Veröffentlicht: (2025)
DiffMagicFace: Identity Consistent Facial Editing of Real Videos
von: Yin, Huanghao, et al.
Veröffentlicht: (2026)
von: Yin, Huanghao, et al.
Veröffentlicht: (2026)
MotionVerse: A Unified Multimodal Framework for Motion Comprehension, Generation and Editing
von: Hou, Ruibing, et al.
Veröffentlicht: (2025)
von: Hou, Ruibing, et al.
Veröffentlicht: (2025)
Long-Term TalkingFace Generation via Motion-Prior Conditional Diffusion Model
von: Shen, Fei, et al.
Veröffentlicht: (2025)
von: Shen, Fei, et al.
Veröffentlicht: (2025)
MAGIC-Talk: Motion-aware Audio-Driven Talking Face Generation with Customizable Identity Control
von: Nazarieh, Fatemeh, et al.
Veröffentlicht: (2025)
von: Nazarieh, Fatemeh, et al.
Veröffentlicht: (2025)
Sound2Vision: Generating Diverse Visuals from Audio through Cross-Modal Latent Alignment
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2024)
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2024)
MemBench: Memorized Image Trigger Prompt Dataset for Diffusion Models
von: Hong, Chunsan, et al.
Veröffentlicht: (2024)
von: Hong, Chunsan, et al.
Veröffentlicht: (2024)
MotionLab: Unified Human Motion Generation and Editing via the Motion-Condition-Motion Paradigm
von: Guo, Ziyan, et al.
Veröffentlicht: (2025)
von: Guo, Ziyan, et al.
Veröffentlicht: (2025)
Cross-Modal Emotion Transfer for Emotion Editing in Talking Face Video
von: Choi, Chanhyuk, et al.
Veröffentlicht: (2026)
von: Choi, Chanhyuk, et al.
Veröffentlicht: (2026)
Generative Motion Infilling From Imprecisely Timed Keyframes
von: Goel, Purvi, et al.
Veröffentlicht: (2025)
von: Goel, Purvi, et al.
Veröffentlicht: (2025)
Noise Map Guidance: Inversion with Spatial Context for Real Image Editing
von: Cho, Hansam, et al.
Veröffentlicht: (2024)
von: Cho, Hansam, et al.
Veröffentlicht: (2024)
AniTalker: Animate Vivid and Diverse Talking Faces through Identity-Decoupled Facial Motion Encoding
von: Liu, Tao, et al.
Veröffentlicht: (2024)
von: Liu, Tao, et al.
Veröffentlicht: (2024)
AlignDiT: Multimodal Aligned Diffusion Transformer for Synchronized Speech Generation
von: Choi, Jeongsoo, et al.
Veröffentlicht: (2025)
von: Choi, Jeongsoo, et al.
Veröffentlicht: (2025)
A Generalist FaceX via Learning Unified Facial Representation
von: Han, Yue, et al.
Veröffentlicht: (2023)
von: Han, Yue, et al.
Veröffentlicht: (2023)
VQTalker: Towards Multilingual Talking Avatars through Facial Motion Tokenization
von: Liu, Tao, et al.
Veröffentlicht: (2024)
von: Liu, Tao, et al.
Veröffentlicht: (2024)
EditEmoTalk: Controllable Speech-Driven 3D Facial Animation with Continuous Expression Editing
von: Jiang, Diqiong, et al.
Veröffentlicht: (2026)
von: Jiang, Diqiong, et al.
Veröffentlicht: (2026)
SegTalker: Segmentation-based Talking Face Generation with Mask-guided Local Editing
von: Xiong, Lingyu, et al.
Veröffentlicht: (2024)
von: Xiong, Lingyu, et al.
Veröffentlicht: (2024)
FaceXFormer: A Unified Transformer for Facial Analysis
von: Narayan, Kartik, et al.
Veröffentlicht: (2024)
von: Narayan, Kartik, et al.
Veröffentlicht: (2024)
FaceGCD: Generalized Face Discovery via Dynamic Prefix Generation
von: Oh, Yunseok, et al.
Veröffentlicht: (2025)
von: Oh, Yunseok, et al.
Veröffentlicht: (2025)
FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis
von: Wang, Mengchao, et al.
Veröffentlicht: (2025)
von: Wang, Mengchao, et al.
Veröffentlicht: (2025)
InstaFace: Identity-Preserving Facial Editing with Single Image Inference
von: Khan, MD Wahiduzzaman, et al.
Veröffentlicht: (2025)
von: Khan, MD Wahiduzzaman, et al.
Veröffentlicht: (2025)
IP-FaceDiff: Identity-Preserving Facial Video Editing with Diffusion
von: Anand, Tharun, et al.
Veröffentlicht: (2025)
von: Anand, Tharun, et al.
Veröffentlicht: (2025)
IRASNet: Improved Feature-Level Clutter Reduction for Domain Generalized SAR-ATR
von: Jang, Oh-Tae, et al.
Veröffentlicht: (2024)
von: Jang, Oh-Tae, et al.
Veröffentlicht: (2024)
EDiT: Efficient Diffusion Transformers with Linear Compressed Attention
von: Becker, Philipp, et al.
Veröffentlicht: (2025)
von: Becker, Philipp, et al.
Veröffentlicht: (2025)
IMTalker: Efficient Audio-driven Talking Face Generation with Implicit Motion Transfer
von: Chen, Bo, et al.
Veröffentlicht: (2025)
von: Chen, Bo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MultiTalk: Enhancing 3D Talking Head Generation Across Languages with Multilingual Video Dataset
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2024) -
Perceptually Accurate 3D Talking Head Generation: New Definitions, Speech-Mesh Representation, and Evaluation Metrics
von: Chae-Yeon, Lee, et al.
Veröffentlicht: (2025) -
Controllable Talking Face Generation by Implicit Facial Keypoints Editing
von: Zhao, Dong, et al.
Veröffentlicht: (2024) -
UP-FacE: User-predictable Fine-grained Face Shape Editing
von: Strohm, Florian, et al.
Veröffentlicht: (2024) -
VoiceCraft-Dub: Automated Video Dubbing with Neural Codec Language Models
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2025)