Learning Phonetic Context-Dependent Viseme for Enhancing Speech-Driven 3D Facial Animation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Hyung Kyu, Kim, Hak Gu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MemoryTalker: Personalized Speech-Driven 3D Facial Animation via Audio-Guided Stylization
von: Kim, Hyung Kyu, et al.
Veröffentlicht: (2025)
von: Kim, Hyung Kyu, et al.
Veröffentlicht: (2025)
DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer
von: Ma, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Ma, Zhiyuan, et al.
Veröffentlicht: (2024)
CSTalk: Correlation Supervised Speech-driven 3D Emotional Facial Animation Generation
von: Liang, Xiangyu, et al.
Veröffentlicht: (2024)
von: Liang, Xiangyu, et al.
Veröffentlicht: (2024)
ProbTalk3D: Non-Deterministic Emotion Controllable Speech-Driven 3D Facial Animation Synthesis Using VQ-VAE
von: Wu, Sichun, et al.
Veröffentlicht: (2024)
von: Wu, Sichun, et al.
Veröffentlicht: (2024)
Learning Disentangled Speech- and Expression-Driven Blendshapes for 3D Talking Face Animation
von: Mao, Yuxiang, et al.
Veröffentlicht: (2025)
von: Mao, Yuxiang, et al.
Veröffentlicht: (2025)
DIAMOND: An LLM-Driven Agent for Context-Aware Baseball Highlight Summarization
von: Kang, Jeonghun, et al.
Veröffentlicht: (2025)
von: Kang, Jeonghun, et al.
Veröffentlicht: (2025)
Diverse Code Query Learning for Speech-Driven Facial Animation
von: Gu, Chunzhi, et al.
Veröffentlicht: (2024)
von: Gu, Chunzhi, et al.
Veröffentlicht: (2024)
SAiD: Speech-driven Blendshape Facial Animation with Diffusion
von: Park, Inkyu, et al.
Veröffentlicht: (2023)
von: Park, Inkyu, et al.
Veröffentlicht: (2023)
3DiFACE: Synthesizing and Editing Holistic 3D Facial Animation
von: Thambiraja, Balamurugan, et al.
Veröffentlicht: (2025)
von: Thambiraja, Balamurugan, et al.
Veröffentlicht: (2025)
PC-Talk: Precise Facial Animation Control for Audio-Driven Talking Face Generation
von: Wang, Baiqin, et al.
Veröffentlicht: (2025)
von: Wang, Baiqin, et al.
Veröffentlicht: (2025)
Fast Registration of Photorealistic Avatars for VR Facial Animation
von: Patel, Chaitanya, et al.
Veröffentlicht: (2024)
von: Patel, Chaitanya, et al.
Veröffentlicht: (2024)
KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation
von: Lyu, Tianle, et al.
Veröffentlicht: (2025)
von: Lyu, Tianle, et al.
Veröffentlicht: (2025)
KeyFace: Expressive Audio-Driven Facial Animation for Long Sequences via KeyFrame Interpolation
von: Bigata, Antoni, et al.
Veröffentlicht: (2025)
von: Bigata, Antoni, et al.
Veröffentlicht: (2025)
Instruction-Driven 3D Facial Expression Generation and Transition
von: Vo, Anh H., et al.
Veröffentlicht: (2026)
von: Vo, Anh H., et al.
Veröffentlicht: (2026)
Controllable Expressive 3D Facial Animation via Diffusion in a Unified Multimodal Space
von: Liu, Kangwei, et al.
Veröffentlicht: (2025)
von: Liu, Kangwei, et al.
Veröffentlicht: (2025)
Polyglot: Multilingual Style Preserving Speech-Driven Facial Animation
von: Nocentini, Federico, et al.
Veröffentlicht: (2026)
von: Nocentini, Federico, et al.
Veröffentlicht: (2026)
Enhancing Speech-Driven 3D Facial Animation with Audio-Visual Guidance from Lip Reading Expert
von: EunGi, Han, et al.
Veröffentlicht: (2024)
von: EunGi, Han, et al.
Veröffentlicht: (2024)
Ex-Omni: Enabling 3D Facial Animation Generation for Omni-modal Large Language Models
von: Zhang, Haoyu, et al.
Veröffentlicht: (2026)
von: Zhang, Haoyu, et al.
Veröffentlicht: (2026)
TDMM-LM: Bridging Facial Understanding and Animation via Language Models
von: Song, Luchuan, et al.
Veröffentlicht: (2026)
von: Song, Luchuan, et al.
Veröffentlicht: (2026)
Text2Lip: Progressive Lip-Synced Talking Face Generation from Text via Viseme-Guided Rendering
von: Wang, Xu, et al.
Veröffentlicht: (2025)
von: Wang, Xu, et al.
Veröffentlicht: (2025)
KMTalk: Speech-Driven 3D Facial Animation with Key Motion Embedding
von: Xu, Zhihao, et al.
Veröffentlicht: (2024)
von: Xu, Zhihao, et al.
Veröffentlicht: (2024)
TCAN: Animating Human Images with Temporally Consistent Pose Guidance using Diffusion Models
von: Kim, Jeongho, et al.
Veröffentlicht: (2024)
von: Kim, Jeongho, et al.
Veröffentlicht: (2024)
ARBEx: Attentive Feature Extraction with Reliability Balancing for Robust Facial Expression Learning
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2023)
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2023)
DEEPTalk: Dynamic Emotion Embedding for Probabilistic Speech-Driven 3D Face Animation
von: Kim, Jisoo, et al.
Veröffentlicht: (2024)
von: Kim, Jisoo, et al.
Veröffentlicht: (2024)
SEAL-pose: Enhancing 3D Human Pose Estimation via a Learned Loss for Structural Consistency
von: Kim, Yeonsung, et al.
Veröffentlicht: (2026)
von: Kim, Yeonsung, et al.
Veröffentlicht: (2026)
3D Occupancy Prediction with Low-Resolution Queries via Prototype-aware View Transformation
von: Oh, Gyeongrok, et al.
Veröffentlicht: (2025)
von: Oh, Gyeongrok, et al.
Veröffentlicht: (2025)
InsertAnywhere: Bridging 4D Scene Geometry and Diffusion Models for Realistic Video Object Insertion
von: Jin, Hoiyeong, et al.
Veröffentlicht: (2025)
von: Jin, Hoiyeong, et al.
Veröffentlicht: (2025)
AniTalker: Animate Vivid and Diverse Talking Faces through Identity-Decoupled Facial Motion Encoding
von: Liu, Tao, et al.
Veröffentlicht: (2024)
von: Liu, Tao, et al.
Veröffentlicht: (2024)
DreamActor-M2: Universal Character Image Animation via Spatiotemporal In-Context Learning
von: Luo, Mingshuang, et al.
Veröffentlicht: (2026)
von: Luo, Mingshuang, et al.
Veröffentlicht: (2026)
Unexplored Faces of Robustness and Out-of-Distribution: Covariate Shifts in Environment and Sensor Domains
von: Baek, Eunsu, et al.
Veröffentlicht: (2024)
von: Baek, Eunsu, et al.
Veröffentlicht: (2024)
3DFacePolicy: Audio-Driven 3D Facial Animation Based on Action Control
von: Sha, Xuanmeng, et al.
Veröffentlicht: (2024)
von: Sha, Xuanmeng, et al.
Veröffentlicht: (2024)
GLDiTalker: Speech-Driven 3D Facial Animation with Graph Latent Diffusion Transformer
von: Lin, Yihong, et al.
Veröffentlicht: (2024)
von: Lin, Yihong, et al.
Veröffentlicht: (2024)
Task-Specific Adaptation of Segmentation Foundation Model via Prompt Learning
von: Kim, Hyung-Il, et al.
Veröffentlicht: (2024)
von: Kim, Hyung-Il, et al.
Veröffentlicht: (2024)
Contextual Gesture: Co-Speech Gesture Video Generation through Context-aware Gesture Representation
von: Liu, Pinxin, et al.
Veröffentlicht: (2025)
von: Liu, Pinxin, et al.
Veröffentlicht: (2025)
Task Prototype-Based Knowledge Retrieval for Multi-Task Learning from Partially Annotated Data
von: Oh, Youngmin, et al.
Veröffentlicht: (2026)
von: Oh, Youngmin, et al.
Veröffentlicht: (2026)
MoST: Motion Style Transformer between Diverse Action Contents
von: Kim, Boeun, et al.
Veröffentlicht: (2024)
von: Kim, Boeun, et al.
Veröffentlicht: (2024)
Collaborative Learning for 3D Hand-Object Reconstruction and Compositional Action Recognition from Egocentric RGB Videos Using Superquadrics
von: Tse, Tze Ho Elden, et al.
Veröffentlicht: (2025)
von: Tse, Tze Ho Elden, et al.
Veröffentlicht: (2025)
AdaMesh: Personalized Facial Expressions and Head Poses for Adaptive Speech-Driven 3D Facial Animation
von: Chen, Liyang, et al.
Veröffentlicht: (2023)
von: Chen, Liyang, et al.
Veröffentlicht: (2023)
Exploring Phonetic Context-Aware Lip-Sync For Talking Face Generation
von: Park, Se Jin, et al.
Veröffentlicht: (2023)
von: Park, Se Jin, et al.
Veröffentlicht: (2023)
LoopAnimate: Loopable Salient Object Animation
von: Wang, Fanyi, et al.
Veröffentlicht: (2024)
von: Wang, Fanyi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MemoryTalker: Personalized Speech-Driven 3D Facial Animation via Audio-Guided Stylization
von: Kim, Hyung Kyu, et al.
Veröffentlicht: (2025) -
DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer
von: Ma, Zhiyuan, et al.
Veröffentlicht: (2024) -
CSTalk: Correlation Supervised Speech-driven 3D Emotional Facial Animation Generation
von: Liang, Xiangyu, et al.
Veröffentlicht: (2024) -
ProbTalk3D: Non-Deterministic Emotion Controllable Speech-Driven 3D Facial Animation Synthesis Using VQ-VAE
von: Wu, Sichun, et al.
Veröffentlicht: (2024) -
Learning Disentangled Speech- and Expression-Driven Blendshapes for 3D Talking Face Animation
von: Mao, Yuxiang, et al.
Veröffentlicht: (2025)