Beyond Fixed Topologies: Unregistered Training and Comprehensive Evaluation Metrics for 3D Talking Heads
Fuente:
arXiv
Saved in:
| Main Authors: | Nocentini, Federico, Besnier, Thomas, Ferrari, Claudio, Arguillere, Sylvain, Daoudi, Mohamed, Berretti, Stefano |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ScanTalk: 3D Talking Heads from Unregistered Scans
by: Nocentini, Federico, et al.
Published: (2024)
by: Nocentini, Federico, et al.
Published: (2024)
FreeTalk: Emotional Topology-Free 3D Talking Heads
by: Nocentini, Federico, et al.
Published: (2026)
by: Nocentini, Federico, et al.
Published: (2026)
ScanMove: Motion Prediction and Transfer for Unregistered Body Meshes
by: Besnier, Thomas, et al.
Published: (2025)
by: Besnier, Thomas, et al.
Published: (2025)
EmoVOCA: Speech-Driven Emotional 3D Talking Heads
by: Nocentini, Federico, et al.
Published: (2024)
by: Nocentini, Federico, et al.
Published: (2024)
PaNDaS: Learnable Deformation Modeling with Localized Control
by: Besnier, Thomas, et al.
Published: (2024)
by: Besnier, Thomas, et al.
Published: (2024)
3D Face Reconstruction Error Decomposed: A Modular Benchmark for Fair and Fast Method Evaluation
by: Sariyanidi, Evangelos, et al.
Published: (2025)
by: Sariyanidi, Evangelos, et al.
Published: (2025)
JambaTalk: Speech-Driven 3D Talking Head Generation Based on Hybrid Transformer-Mamba Model
by: Jafari, Farzaneh, et al.
Published: (2024)
by: Jafari, Farzaneh, et al.
Published: (2024)
Polyglot: Multilingual Style Preserving Speech-Driven Facial Animation
by: Nocentini, Federico, et al.
Published: (2026)
by: Nocentini, Federico, et al.
Published: (2026)
Revisiting Emotions Representation for Recognition in the Wild
by: Neto, Joao Baptista Cardia, et al.
Published: (2026)
by: Neto, Joao Baptista Cardia, et al.
Published: (2026)
Generation of Complex 3D Human Motion by Temporal and Spatial Composition of Diffusion Models
by: Mandelli, Lorenzo, et al.
Published: (2024)
by: Mandelli, Lorenzo, et al.
Published: (2024)
SEDTalker: Emotion-Aware 3D Facial Animation Using Frame-Level Speech Emotion Diarization
by: Jafari, Farzaneh, et al.
Published: (2026)
by: Jafari, Farzaneh, et al.
Published: (2026)
Perceptually Accurate 3D Talking Head Generation: New Definitions, Speech-Mesh Representation, and Evaluation Metrics
by: Chae-Yeon, Lee, et al.
Published: (2025)
by: Chae-Yeon, Lee, et al.
Published: (2025)
Establishing a Unified Evaluation Framework for Human Motion Generation: A Comparative Analysis of Metrics
by: Ismail-Fawaz, Ali, et al.
Published: (2024)
by: Ismail-Fawaz, Ali, et al.
Published: (2024)
BIGFix: Bidirectional Image Generation with Token Fixing
by: Besnier, Victor, et al.
Published: (2025)
by: Besnier, Victor, et al.
Published: (2025)
FixTalk: Taming Identity Leakage for High-Quality Talking Head Generation in Extreme Cases
by: Tan, Shuai, et al.
Published: (2025)
by: Tan, Shuai, et al.
Published: (2025)
GaussianHeadTalk: Wobble-Free 3D Talking Heads with Audio Driven Gaussian Splatting
by: Agarwal, Madhav, et al.
Published: (2025)
by: Agarwal, Madhav, et al.
Published: (2025)
OT-Talk: Animating 3D Talking Head with Optimal Transportation
by: Wang, Xinmu, et al.
Published: (2025)
by: Wang, Xinmu, et al.
Published: (2025)
Advancing Talking Head Generation: A Comprehensive Survey of Multi-Modal Methodologies, Datasets, Evaluation Metrics, and Loss Functions
by: Rakesh, Vineet Kumar, et al.
Published: (2025)
by: Rakesh, Vineet Kumar, et al.
Published: (2025)
Measuring Anxiety Levels with Head Motion Patterns in Severe Depression Population
by: Boutaleb, Fouad, et al.
Published: (2025)
by: Boutaleb, Fouad, et al.
Published: (2025)
EmoTalk3D: High-Fidelity Free-View Synthesis of Emotional 3D Talking Head
by: He, Qianyun, et al.
Published: (2024)
by: He, Qianyun, et al.
Published: (2024)
DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations
by: Peng, Ziqiao, et al.
Published: (2025)
by: Peng, Ziqiao, et al.
Published: (2025)
EmoDiffTalk:Emotion-aware Diffusion for Editable 3D Gaussian Talking Head
by: Liu, Chang, et al.
Published: (2025)
by: Liu, Chang, et al.
Published: (2025)
TalkingGaussian: Structure-Persistent 3D Talking Head Synthesis via Gaussian Splatting
by: Li, Jiahe, et al.
Published: (2024)
by: Li, Jiahe, et al.
Published: (2024)
Unregistered Spectral Image Fusion: Unmixing, Adversarial Learning, and Recoverability
by: Song, Jiahui, et al.
Published: (2026)
by: Song, Jiahui, et al.
Published: (2026)
THEval. Evaluation Framework for Talking Head Video Generation
by: Quignon, Nabyl, et al.
Published: (2025)
by: Quignon, Nabyl, et al.
Published: (2025)
Localized Latent Editing for Dose-Response Modeling in Botulinum Toxin Injection Planning
by: Arnaud, Estèphe, et al.
Published: (2026)
by: Arnaud, Estèphe, et al.
Published: (2026)
A Comparative Study of Perceptual Quality Metrics for Audio-driven Talking Head Videos
by: Zhang, Weixia, et al.
Published: (2024)
by: Zhang, Weixia, et al.
Published: (2024)
MultiTalk: Enhancing 3D Talking Head Generation Across Languages with Multilingual Video Dataset
by: Sung-Bin, Kim, et al.
Published: (2024)
by: Sung-Bin, Kim, et al.
Published: (2024)
No MoCap Needed: Post-Training Motion Diffusion Models with Reinforcement Learning using Only Textual Prompts
by: Macaluso, Girolamo, et al.
Published: (2025)
by: Macaluso, Girolamo, et al.
Published: (2025)
Supervising 3D Talking Head Avatars with Analysis-by-Audio-Synthesis
by: Daněček, Radek, et al.
Published: (2025)
by: Daněček, Radek, et al.
Published: (2025)
SyncTalk: The Devil is in the Synchronization for Talking Head Synthesis
by: Peng, Ziqiao, et al.
Published: (2023)
by: Peng, Ziqiao, et al.
Published: (2023)
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion
by: Liu, Bin, et al.
Published: (2026)
by: Liu, Bin, et al.
Published: (2026)
Toward Fine-Grained Facial Control in 3D Talking Head Generation
by: Xie, Shaoyang, et al.
Published: (2026)
by: Xie, Shaoyang, et al.
Published: (2026)
R3DPA: Leveraging 3D Representation Alignment and RGB Pretrained Priors for LiDAR Scene Generation
by: Sereyjol-Garros, Nicolas, et al.
Published: (2026)
by: Sereyjol-Garros, Nicolas, et al.
Published: (2026)
DEGMC: Denoising Diffusion Models Based on Riemannian Equivariant Group Morphological Convolutions
by: Diop, El Hadji S., et al.
Published: (2026)
by: Diop, El Hadji S., et al.
Published: (2026)
Generative Deep Learning for Computational Destaining and Restaining of Unregistered Digital Pathology Images
by: Kulkarni, Aarushi, et al.
Published: (2026)
by: Kulkarni, Aarushi, et al.
Published: (2026)
Enhancing Unregistered Hyperspectral Image Super-Resolution via Unmixing-based Abundance Fusion Learning
by: Zhang, Yingkai, et al.
Published: (2026)
by: Zhang, Yingkai, et al.
Published: (2026)
CapTalk: Text-Guided Stylization and Speech-Driven 3D Head Animation
by: Chu, Xuangeng, et al.
Published: (2026)
by: Chu, Xuangeng, et al.
Published: (2026)
TalkingHeadBench: A Multi-Modal Benchmark & Analysis of Talking-Head DeepFake Detection
by: Xiong, Xinqi, et al.
Published: (2025)
by: Xiong, Xinqi, et al.
Published: (2025)
EmoTalkingGaussian: Continuous Emotion-conditioned Talking Head Synthesis
by: Cha, Junuk, et al.
Published: (2025)
by: Cha, Junuk, et al.
Published: (2025)
Similar Items
-
ScanTalk: 3D Talking Heads from Unregistered Scans
by: Nocentini, Federico, et al.
Published: (2024) -
FreeTalk: Emotional Topology-Free 3D Talking Heads
by: Nocentini, Federico, et al.
Published: (2026) -
ScanMove: Motion Prediction and Transfer for Unregistered Body Meshes
by: Besnier, Thomas, et al.
Published: (2025) -
EmoVOCA: Speech-Driven Emotional 3D Talking Heads
by: Nocentini, Federico, et al.
Published: (2024) -
PaNDaS: Learnable Deformation Modeling with Localized Control
by: Besnier, Thomas, et al.
Published: (2024)