Perceptually Accurate 3D Talking Head Generation: New Definitions, Speech-Mesh Representation, and Evaluation Metrics
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chae-Yeon, Lee, Hyun-Bin, Oh, EunGi, Han, Sung-Bin, Kim, Nam, Suekyeong, Oh, Tae-Hyun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Enhancing Speech-Driven 3D Facial Animation with Audio-Visual Guidance from Lip Reading Expert
von: EunGi, Han, et al.
Veröffentlicht: (2024)
von: EunGi, Han, et al.
Veröffentlicht: (2024)
MultiTalk: Enhancing 3D Talking Head Generation Across Languages with Multilingual Video Dataset
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2024)
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2024)
FPGS: Feed-Forward Semantic-aware Photorealistic Style Transfer of Large-Scale Gaussian Splatting
von: Kim, GeonU, et al.
Veröffentlicht: (2025)
von: Kim, GeonU, et al.
Veröffentlicht: (2025)
FPRF: Feed-Forward Photorealistic Style Transfer of Large-Scale 3D Neural Radiance Fields
von: Kim, GeonU, et al.
Veröffentlicht: (2024)
von: Kim, GeonU, et al.
Veröffentlicht: (2024)
Paint-it: Text-to-Texture Synthesis via Deep Convolutional Texture Map Optimization and Physically-Based Rendering
von: Youwang, Kim, et al.
Veröffentlicht: (2023)
von: Youwang, Kim, et al.
Veröffentlicht: (2023)
PAColorHolo: A Perceptually-Aware Color Management Framework for Holographic Displays
von: Chen, Chun, et al.
Veröffentlicht: (2026)
von: Chen, Chun, et al.
Veröffentlicht: (2026)
Perceptual Requirements for Low-Latency Head-Mounted Displays
von: Penner, Eric, et al.
Veröffentlicht: (2026)
von: Penner, Eric, et al.
Veröffentlicht: (2026)
PaMO: Parallel Mesh Optimization for Intersection-Free Low-Poly Modeling on the GPU
von: Oh, Seonghun, et al.
Veröffentlicht: (2025)
von: Oh, Seonghun, et al.
Veröffentlicht: (2025)
TetraSDF: Precise Mesh Extraction with Multi-resolution Tetrahedral Grid
von: Oh, Seonghun, et al.
Veröffentlicht: (2025)
von: Oh, Seonghun, et al.
Veröffentlicht: (2025)
Holographic Parallax Improves 3D Perceptual Realism
von: Kim, Dongyeon, et al.
Veröffentlicht: (2024)
von: Kim, Dongyeon, et al.
Veröffentlicht: (2024)
Piecewise Ruled Approximation for Freeform Mesh Surfaces
von: Pan, Yiling, et al.
Veröffentlicht: (2025)
von: Pan, Yiling, et al.
Veröffentlicht: (2025)
PASE: Phoneme-Aware Speech Encoder to Improve Lip Sync Accuracy for Talking Head Synthesis
von: Huang, Yihuan, et al.
Veröffentlicht: (2025)
von: Huang, Yihuan, et al.
Veröffentlicht: (2025)
Learning Correlation-aware Aleatoric Uncertainty for 3D Hand Pose Estimation
von: Chae-Yeon, Lee, et al.
Veröffentlicht: (2025)
von: Chae-Yeon, Lee, et al.
Veröffentlicht: (2025)
OT-Talk: Animating 3D Talking Head with Optimal Transportation
von: Wang, Xinmu, et al.
Veröffentlicht: (2025)
von: Wang, Xinmu, et al.
Veröffentlicht: (2025)
Perceptual Sensitivity to Stereo Geometry Errors in Head-Mounted Displays
von: Zhu, Raffles Xingqi, et al.
Veröffentlicht: (2025)
von: Zhu, Raffles Xingqi, et al.
Veröffentlicht: (2025)
DHPrep: Deep Hawkes Process based Dynamic Network Representation
von: Han, Ruixuan, et al.
Veröffentlicht: (2024)
von: Han, Ruixuan, et al.
Veröffentlicht: (2024)
Reality Check: How Avatar and Face Representation Affect the Perceptual Evaluation of Synthesized Gestures
von: Du, Haoyang, et al.
Veröffentlicht: (2026)
von: Du, Haoyang, et al.
Veröffentlicht: (2026)
FacEDiT: Unified Talking Face Editing and Generation via Facial Motion Infilling
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2025)
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2025)
The Life and Legacy of Bui Tuong Phong
von: Oh, Yoehan, et al.
Veröffentlicht: (2024)
von: Oh, Yoehan, et al.
Veröffentlicht: (2024)
Gaussian Fluids: A Grid-Free Fluid Solver based on Gaussian Spatial Representation
von: Xing, Jingrui, et al.
Veröffentlicht: (2024)
von: Xing, Jingrui, et al.
Veröffentlicht: (2024)
DiffPoseTalk: Speech-Driven Stylistic 3D Facial Animation and Head Pose Generation via Diffusion Models
von: Sun, Zhiyao, et al.
Veröffentlicht: (2023)
von: Sun, Zhiyao, et al.
Veröffentlicht: (2023)
Real-Time Cloth Simulation Using WebGPU: Evaluating Limits of High-Resolution
von: Sung, Nak-Jun, et al.
Veröffentlicht: (2025)
von: Sung, Nak-Jun, et al.
Veröffentlicht: (2025)
How Does a Virtual Agent Decide Where to Look? Symbolic Cognitive Reasoning for Embodied Head Rotation
von: Hwang, Juyeong, et al.
Veröffentlicht: (2025)
von: Hwang, Juyeong, et al.
Veröffentlicht: (2025)
MoNeRF: Deformable Neural Rendering for Talking Heads via Latent Motion Navigation
von: X. Li, et al.
Veröffentlicht: (2024)
von: X. Li, et al.
Veröffentlicht: (2024)
Mesh Processing Non-Meshes via Neural Displacement Fields
von: Noma, Yuta, et al.
Veröffentlicht: (2025)
von: Noma, Yuta, et al.
Veröffentlicht: (2025)
FreeMesh: Boosting Mesh Generation with Coordinates Merging
von: Liu, Jian, et al.
Veröffentlicht: (2025)
von: Liu, Jian, et al.
Veröffentlicht: (2025)
StyGazeTalk: Learning Stylized Generation of Gaze and Head Dynamics
von: Shi, Chengwei, et al.
Veröffentlicht: (2025)
von: Shi, Chengwei, et al.
Veröffentlicht: (2025)
MoDA: Multi-modal Diffusion Architecture for Talking Head Generation
von: Li, Xinyang, et al.
Veröffentlicht: (2025)
von: Li, Xinyang, et al.
Veröffentlicht: (2025)
Supervising 3D Talking Head Avatars with Analysis-by-Audio-Synthesis
von: Daněček, Radek, et al.
Veröffentlicht: (2025)
von: Daněček, Radek, et al.
Veröffentlicht: (2025)
Advancing Talking Head Generation: A Comprehensive Survey of Multi-Modal Methodologies, Datasets, Evaluation Metrics, and Loss Functions
von: Rakesh, Vineet Kumar, et al.
Veröffentlicht: (2025)
von: Rakesh, Vineet Kumar, et al.
Veröffentlicht: (2025)
Mesh Simplification For Unfolding
von: Bhargava, Manas, et al.
Veröffentlicht: (2024)
von: Bhargava, Manas, et al.
Veröffentlicht: (2024)
DMesh: A Differentiable Mesh Representation
von: Son, Sanghyun, et al.
Veröffentlicht: (2024)
von: Son, Sanghyun, et al.
Veröffentlicht: (2024)
Quadratic-Order Geodesics on Meshes
von: Ruan, Yue, et al.
Veröffentlicht: (2026)
von: Ruan, Yue, et al.
Veröffentlicht: (2026)
Learned Adaptive Mesh Generation
von: Zhang, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Zhiyuan, et al.
Veröffentlicht: (2025)
A Comprehensive Multi-scale Approach for Speech and Dynamics Synchrony in Talking Head Generation
von: Airale, Louis, et al.
Veröffentlicht: (2023)
von: Airale, Louis, et al.
Veröffentlicht: (2023)
MeshGen: Generating PBR Textured Mesh with Render-Enhanced Auto-Encoder and Generative Data Augmentation
von: Chen, Zilong, et al.
Veröffentlicht: (2025)
von: Chen, Zilong, et al.
Veröffentlicht: (2025)
PosterReward: Unlocking Accurate Evaluation for High-Quality Graphic Design Generation
von: Lai, Jianyu, et al.
Veröffentlicht: (2026)
von: Lai, Jianyu, et al.
Veröffentlicht: (2026)
A Robust Grid-Based Meshing Algorithm for Embedding Self-Intersecting Surfaces
von: Gagniere, Steven W., et al.
Veröffentlicht: (2022)
von: Gagniere, Steven W., et al.
Veröffentlicht: (2022)
From Far and Near: Perceptual Evaluation of Crowd Representations Across Levels of Detail
von: Sun, Xiaohan, et al.
Veröffentlicht: (2025)
von: Sun, Xiaohan, et al.
Veröffentlicht: (2025)
Projecting Radiance Fields to Mesh Surfaces
von: Lim, Adrian Xuan Wei, et al.
Veröffentlicht: (2024)
von: Lim, Adrian Xuan Wei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Enhancing Speech-Driven 3D Facial Animation with Audio-Visual Guidance from Lip Reading Expert
von: EunGi, Han, et al.
Veröffentlicht: (2024) -
MultiTalk: Enhancing 3D Talking Head Generation Across Languages with Multilingual Video Dataset
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2024) -
FPGS: Feed-Forward Semantic-aware Photorealistic Style Transfer of Large-Scale Gaussian Splatting
von: Kim, GeonU, et al.
Veröffentlicht: (2025) -
FPRF: Feed-Forward Photorealistic Style Transfer of Large-Scale 3D Neural Radiance Fields
von: Kim, GeonU, et al.
Veröffentlicht: (2024) -
Paint-it: Text-to-Texture Synthesis via Deep Convolutional Texture Map Optimization and Physically-Based Rendering
von: Youwang, Kim, et al.
Veröffentlicht: (2023)