Enhancing Speech-Driven 3D Facial Animation with Audio-Visual Guidance from Lip Reading Expert
Fuente:
arXiv
Saved in:
| Main Authors: | EunGi, Han, Hyun-Bin, Oh, Sung-Bin, Kim, Etcheberry, Corentin Nivelet, Nam, Suekyeong, Joo, Janghoon, Oh, Tae-Hyun |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Perceptually Accurate 3D Talking Head Generation: New Definitions, Speech-Mesh Representation, and Evaluation Metrics
by: Chae-Yeon, Lee, et al.
Published: (2025)
by: Chae-Yeon, Lee, et al.
Published: (2025)
MultiTalk: Enhancing 3D Talking Head Generation Across Languages with Multilingual Video Dataset
by: Sung-Bin, Kim, et al.
Published: (2024)
by: Sung-Bin, Kim, et al.
Published: (2024)
Audio Driven Real-Time Facial Animation for Social Telepresence
by: Lee, Jiye, et al.
Published: (2025)
by: Lee, Jiye, et al.
Published: (2025)
Content and Style Aware Audio-Driven Facial Animation
by: Liu, Qingju, et al.
Published: (2024)
by: Liu, Qingju, et al.
Published: (2024)
PESTalk: Speech-Driven 3D Facial Animation with Personalized Emotional Styles
by: Han, Tianshun, et al.
Published: (2025)
by: Han, Tianshun, et al.
Published: (2025)
LSF-Animation: Label-Free Speech-Driven Facial Animation via Implicit Feature Representation
by: Lu, Xin, et al.
Published: (2025)
by: Lu, Xin, et al.
Published: (2025)
FPGS: Feed-Forward Semantic-aware Photorealistic Style Transfer of Large-Scale Gaussian Splatting
by: Kim, GeonU, et al.
Published: (2025)
by: Kim, GeonU, et al.
Published: (2025)
FPRF: Feed-Forward Photorealistic Style Transfer of Large-Scale 3D Neural Radiance Fields
by: Kim, GeonU, et al.
Published: (2024)
by: Kim, GeonU, et al.
Published: (2024)
Paint-it: Text-to-Texture Synthesis via Deep Convolutional Texture Map Optimization and Physically-Based Rendering
by: Youwang, Kim, et al.
Published: (2023)
by: Youwang, Kim, et al.
Published: (2023)
Media2Face: Co-speech Facial Animation Generation With Multi-Modality Guidance
by: Zhao, Qingcheng, et al.
Published: (2024)
by: Zhao, Qingcheng, et al.
Published: (2024)
Model See Model Do: Speech-Driven Facial Animation with Style Control
by: Pan, Yifang, et al.
Published: (2025)
by: Pan, Yifang, et al.
Published: (2025)
DiffPoseTalk: Speech-Driven Stylistic 3D Facial Animation and Head Pose Generation via Diffusion Models
by: Sun, Zhiyao, et al.
Published: (2023)
by: Sun, Zhiyao, et al.
Published: (2023)
Democratizing the Creation of Animatable Facial Avatars
by: Zhu, Yilin, et al.
Published: (2024)
by: Zhu, Yilin, et al.
Published: (2024)
Lightweight Self-Driven Deformable Organ Animations
by: Kenwright, Benjamnin, et al.
Published: (2024)
by: Kenwright, Benjamnin, et al.
Published: (2024)
SAiD: Speech-driven Blendshape Facial Animation with Diffusion
by: Park, Inkyu, et al.
Published: (2023)
by: Park, Inkyu, et al.
Published: (2023)
KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation
by: Lyu, Tianle, et al.
Published: (2025)
by: Lyu, Tianle, et al.
Published: (2025)
Audio2Face-3D: Audio-driven Realistic Facial Animation For Digital Avatars
by: NVIDIA, et al.
Published: (2025)
by: NVIDIA, et al.
Published: (2025)
Animator-Centric Skeleton Generation on Objects with Fine-Grained Details
by: Sun, Mingze, et al.
Published: (2026)
by: Sun, Mingze, et al.
Published: (2026)
Interactive Holographic Visualization for 3D Facial Avatar
by: Nguyen, Tri Tung Nguyen, et al.
Published: (2025)
by: Nguyen, Tri Tung Nguyen, et al.
Published: (2025)
Near-realtime Facial Animation by Deep 3D Simulation Super-Resolution
by: Park, Hyojoon, et al.
Published: (2023)
by: Park, Hyojoon, et al.
Published: (2023)
Llanimation: Llama Driven Gesture Animation
by: J. Windle, et al.
Published: (2024)
by: J. Windle, et al.
Published: (2024)
AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation
by: Wei, Huawei, et al.
Published: (2024)
by: Wei, Huawei, et al.
Published: (2024)
Seam360GS: Seamless 360° Gaussian Splatting from Real-World Omnidirectional Images
by: Shin, Changha, et al.
Published: (2025)
by: Shin, Changha, et al.
Published: (2025)
FreeAvatar: Robust 3D Facial Animation Transfer by Learning an Expression Foundation Model
by: Qiu, Feng, et al.
Published: (2024)
by: Qiu, Feng, et al.
Published: (2024)
Personalizing Causal Audio-Driven Facial Motion via Dynamic Multi-modal Retrieval
by: Chu, Xuangeng, et al.
Published: (2026)
by: Chu, Xuangeng, et al.
Published: (2026)
PAVAS: Physics-Aware Video-to-Audio Synthesis
by: Hyun-Bin, Oh, et al.
Published: (2025)
by: Hyun-Bin, Oh, et al.
Published: (2025)
The Life and Legacy of Bui Tuong Phong
by: Oh, Yoehan, et al.
Published: (2024)
by: Oh, Yoehan, et al.
Published: (2024)
3DStyleGLIP: Part-Tailored Text-Guided 3D Neural Stylization
by: Chung, SeungJeh, et al.
Published: (2024)
by: Chung, SeungJeh, et al.
Published: (2024)
FacEDiT: Unified Talking Face Editing and Generation via Facial Motion Infilling
by: Sung-Bin, Kim, et al.
Published: (2025)
by: Sung-Bin, Kim, et al.
Published: (2025)
AVHBench: A Cross-Modal Hallucination Benchmark for Audio-Visual Large Language Models
by: Sung-Bin, Kim, et al.
Published: (2024)
by: Sung-Bin, Kim, et al.
Published: (2024)
Learning Disentangled Speech- and Expression-Driven Blendshapes for 3D Talking Face Animation
by: Mao, Yuxiang, et al.
Published: (2025)
by: Mao, Yuxiang, et al.
Published: (2025)
Real Time Animator: High-Quality Cartoon Style Transfer in 6 Animation Styles on Images and Videos
by: Yang, Liuxin, et al.
Published: (2025)
by: Yang, Liuxin, et al.
Published: (2025)
EmoFace: Audio-driven Emotional 3D Face Animation
by: Liu, Chang, et al.
Published: (2024)
by: Liu, Chang, et al.
Published: (2024)
PointTalk: Audio-Driven Dynamic Lip Point Cloud for 3D Gaussian-based Talking Head Synthesis
by: Xie, Yifan, et al.
Published: (2024)
by: Xie, Yifan, et al.
Published: (2024)
Sketch Animation: State-of-the-art Report
by: Rai, Gaurav, et al.
Published: (2025)
by: Rai, Gaurav, et al.
Published: (2025)
PASE: Phoneme-Aware Speech Encoder to Improve Lip Sync Accuracy for Talking Head Synthesis
by: Huang, Yihuan, et al.
Published: (2025)
by: Huang, Yihuan, et al.
Published: (2025)
Instant Facial Gaussians Translator for Relightable and Interactable Facial Rendering
by: Qin, Dafei, et al.
Published: (2024)
by: Qin, Dafei, et al.
Published: (2024)
Progressing Level-of-Detail Animation of Volumetric Elastodynamics
by: Zhang, Jiayi Eris, et al.
Published: (2025)
by: Zhang, Jiayi Eris, et al.
Published: (2025)
Cloth Animation with Time-dependent Persistent Wrinkles
by: Gong, Deshan, et al.
Published: (2025)
by: Gong, Deshan, et al.
Published: (2025)
3D Gaussian Model for Animation and Texturing
by: Wang, Xiangzhi Eric, et al.
Published: (2024)
by: Wang, Xiangzhi Eric, et al.
Published: (2024)
Similar Items
-
Perceptually Accurate 3D Talking Head Generation: New Definitions, Speech-Mesh Representation, and Evaluation Metrics
by: Chae-Yeon, Lee, et al.
Published: (2025) -
MultiTalk: Enhancing 3D Talking Head Generation Across Languages with Multilingual Video Dataset
by: Sung-Bin, Kim, et al.
Published: (2024) -
Audio Driven Real-Time Facial Animation for Social Telepresence
by: Lee, Jiye, et al.
Published: (2025) -
Content and Style Aware Audio-Driven Facial Animation
by: Liu, Qingju, et al.
Published: (2024) -
PESTalk: Speech-Driven 3D Facial Animation with Personalized Emotional Styles
by: Han, Tianshun, et al.
Published: (2025)