Polyglot: Multilingual Style Preserving Speech-Driven Facial Animation
Fuente:
arXiv
Saved in:
| Main Authors: | Nocentini, Federico, Seo, Kwanggyoon, Liu, Qingju, Ferrari, Claudio, Berretti, Stefano, Ferman, David, Kim, Hyeongwoo, Garrido, Pablo, Caliskan, Akin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Content and Style Aware Audio-Driven Facial Animation
by: Liu, Qingju, et al.
Published: (2024)
by: Liu, Qingju, et al.
Published: (2024)
EmoVOCA: Speech-Driven Emotional 3D Talking Heads
by: Nocentini, Federico, et al.
Published: (2024)
by: Nocentini, Federico, et al.
Published: (2024)
FreeTalk: Emotional Topology-Free 3D Talking Heads
by: Nocentini, Federico, et al.
Published: (2026)
by: Nocentini, Federico, et al.
Published: (2026)
PAV: Personalized Head Avatar from Unstructured Video Collection
by: Caliskan, Akin, et al.
Published: (2024)
by: Caliskan, Akin, et al.
Published: (2024)
ProMode: A Speech Prosody Model Conditioned on Acoustic and Textual Inputs
by: Eren, Eray, et al.
Published: (2025)
by: Eren, Eray, et al.
Published: (2025)
SEDTalker: Emotion-Aware 3D Facial Animation Using Frame-Level Speech Emotion Diarization
by: Jafari, Farzaneh, et al.
Published: (2026)
by: Jafari, Farzaneh, et al.
Published: (2026)
3D Face Reconstruction Error Decomposed: A Modular Benchmark for Fair and Fast Method Evaluation
by: Sariyanidi, Evangelos, et al.
Published: (2025)
by: Sariyanidi, Evangelos, et al.
Published: (2025)
Beyond Fixed Topologies: Unregistered Training and Comprehensive Evaluation Metrics for 3D Talking Heads
by: Nocentini, Federico, et al.
Published: (2024)
by: Nocentini, Federico, et al.
Published: (2024)
ScanTalk: 3D Talking Heads from Unregistered Scans
by: Nocentini, Federico, et al.
Published: (2024)
by: Nocentini, Federico, et al.
Published: (2024)
FaceLift: Semi-supervised 3D Facial Landmark Localization
by: Ferman, David, et al.
Published: (2024)
by: Ferman, David, et al.
Published: (2024)
Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation
by: Huang, Nick Yiwen, et al.
Published: (2025)
by: Huang, Nick Yiwen, et al.
Published: (2025)
Model See Model Do: Speech-Driven Facial Animation with Style Control
by: Pan, Yifang, et al.
Published: (2025)
by: Pan, Yifang, et al.
Published: (2025)
PESTalk: Speech-Driven 3D Facial Animation with Personalized Emotional Styles
by: Han, Tianshun, et al.
Published: (2025)
by: Han, Tianshun, et al.
Published: (2025)
StyleCineGAN: Landscape Cinemagraph Generation using a Pre-trained StyleGAN
by: Choi, Jongwoo, et al.
Published: (2024)
by: Choi, Jongwoo, et al.
Published: (2024)
Neural Face Skinning for Mesh-agnostic Facial Expression Cloning
by: Cha, Sihun, et al.
Published: (2025)
by: Cha, Sihun, et al.
Published: (2025)
Neural Face Skinning for Mesh‐agnostic Facial Expression Cloning
by: Sihun Cha, et al.
Published: (2025)
by: Sihun Cha, et al.
Published: (2025)
StyleSpeaker: Audio-Enhanced Fine-Grained Style Modeling for Speech-Driven 3D Facial Animation
by: Yang, An, et al.
Published: (2025)
by: Yang, An, et al.
Published: (2025)
NeRFFaceSpeech: One-shot Audio-driven 3D Talking Head Synthesis via Generative Prior
by: Kim, Gihoon, et al.
Published: (2024)
by: Kim, Gihoon, et al.
Published: (2024)
Revisiting Emotions Representation for Recognition in the Wild
by: Neto, Joao Baptista Cardia, et al.
Published: (2026)
by: Neto, Joao Baptista Cardia, et al.
Published: (2026)
JambaTalk: Speech-Driven 3D Talking Head Generation Based on Hybrid Transformer-Mamba Model
by: Jafari, Farzaneh, et al.
Published: (2024)
by: Jafari, Farzaneh, et al.
Published: (2024)
Learning Phonetic Context-Dependent Viseme for Enhancing Speech-Driven 3D Facial Animation
by: Kim, Hyung Kyu, et al.
Published: (2025)
by: Kim, Hyung Kyu, et al.
Published: (2025)
Diverse Code Query Learning for Speech-Driven Facial Animation
by: Gu, Chunzhi, et al.
Published: (2024)
by: Gu, Chunzhi, et al.
Published: (2024)
LSF-Animation: Label-Free Speech-Driven Facial Animation via Implicit Feature Representation
by: Lu, Xin, et al.
Published: (2025)
by: Lu, Xin, et al.
Published: (2025)
Contextual Gesture: Co-Speech Gesture Video Generation through Context-aware Gesture Representation
by: Liu, Pinxin, et al.
Published: (2025)
by: Liu, Pinxin, et al.
Published: (2025)
MemoryTalker: Personalized Speech-Driven 3D Facial Animation via Audio-Guided Stylization
by: Kim, Hyung Kyu, et al.
Published: (2025)
by: Kim, Hyung Kyu, et al.
Published: (2025)
LeGO: Leveraging a Surface Deformation Network for Animatable Stylized Face Generation with One Example
by: Yoon, Soyeon, et al.
Published: (2024)
by: Yoon, Soyeon, et al.
Published: (2024)
KMTalk: Speech-Driven 3D Facial Animation with Key Motion Embedding
by: Xu, Zhihao, et al.
Published: (2024)
by: Xu, Zhihao, et al.
Published: (2024)
DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer
by: Ma, Zhiyuan, et al.
Published: (2024)
by: Ma, Zhiyuan, et al.
Published: (2024)
Generalized impulse response analysis: General or Extreme?
by: Hyeongwoo Kim
Published: (2013)
by: Hyeongwoo Kim
Published: (2013)
AdaMesh: Personalized Facial Expressions and Head Poses for Adaptive Speech-Driven 3D Facial Animation
by: Chen, Liyang, et al.
Published: (2023)
by: Chen, Liyang, et al.
Published: (2023)
Lookahead Anchoring: Preserving Character Identity in Audio-Driven Human Animation
by: Seo, Junyoung, et al.
Published: (2025)
by: Seo, Junyoung, et al.
Published: (2025)
AudioFace: Language-Assisted Speech-Driven Facial Animation with Multimodal Language Models
by: Zheng, Kai, et al.
Published: (2026)
by: Zheng, Kai, et al.
Published: (2026)
Representative Feature Extraction During Diffusion Process for Sketch Extraction with One Example
by: Yun, Kwan, et al.
Published: (2024)
by: Yun, Kwan, et al.
Published: (2024)
Investigating Multilingual Instruction-Tuning: Do Polyglot Models Demand for Multilingual Instructions?
by: Weber, Alexander Arno, et al.
Published: (2024)
by: Weber, Alexander Arno, et al.
Published: (2024)
Expressive Speech-driven Facial Animation with controllable emotions
by: Chen, Yutong, et al.
Published: (2023)
by: Chen, Yutong, et al.
Published: (2023)
KinMo: Kinematic-aware Human Motion Understanding and Generation
by: Zhang, Pengfei, et al.
Published: (2024)
by: Zhang, Pengfei, et al.
Published: (2024)
Enhancing Speech-Driven 3D Facial Animation with Audio-Visual Guidance from Lip Reading Expert
by: EunGi, Han, et al.
Published: (2024)
by: EunGi, Han, et al.
Published: (2024)
GLDiTalker: Speech-Driven 3D Facial Animation with Graph Latent Diffusion Transformer
by: Lin, Yihong, et al.
Published: (2024)
by: Lin, Yihong, et al.
Published: (2024)
Generation of Complex 3D Human Motion by Temporal and Spatial Composition of Diffusion Models
by: Mandelli, Lorenzo, et al.
Published: (2024)
by: Mandelli, Lorenzo, et al.
Published: (2024)
SAiD: Speech-driven Blendshape Facial Animation with Diffusion
by: Park, Inkyu, et al.
Published: (2023)
by: Park, Inkyu, et al.
Published: (2023)
Similar Items
-
Content and Style Aware Audio-Driven Facial Animation
by: Liu, Qingju, et al.
Published: (2024) -
EmoVOCA: Speech-Driven Emotional 3D Talking Heads
by: Nocentini, Federico, et al.
Published: (2024) -
FreeTalk: Emotional Topology-Free 3D Talking Heads
by: Nocentini, Federico, et al.
Published: (2026) -
PAV: Personalized Head Avatar from Unstructured Video Collection
by: Caliskan, Akin, et al.
Published: (2024) -
ProMode: A Speech Prosody Model Conditioned on Acoustic and Textual Inputs
by: Eren, Eray, et al.
Published: (2025)