OT-Talk: Animating 3D Talking Head with Optimal Transportation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Xinmu, Gao, Xiang, Song, Xiyun, Yu, Heather, Lin, Zongfang, Peng, Liang, Gu, Xianfeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Inverse Rendering for High-Genus Surface Meshes from Multi-View Images
von: Gao, Xiang, et al.
Veröffentlicht: (2025)
von: Gao, Xiang, et al.
Veröffentlicht: (2025)
ePBR: Extended PBR Materials in Image Synthesis
von: Guo, Yu, et al.
Veröffentlicht: (2025)
von: Guo, Yu, et al.
Veröffentlicht: (2025)
Neural Geometry Image-Based Representations with Optimal Transport (OT)
von: Gao, Xiang, et al.
Veröffentlicht: (2025)
von: Gao, Xiang, et al.
Veröffentlicht: (2025)
DiffPoseTalk: Speech-Driven Stylistic 3D Facial Animation and Head Pose Generation via Diffusion Models
von: Sun, Zhiyao, et al.
Veröffentlicht: (2023)
von: Sun, Zhiyao, et al.
Veröffentlicht: (2023)
MultiTalk: Enhancing 3D Talking Head Generation Across Languages with Multilingual Video Dataset
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2024)
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2024)
Supervising 3D Talking Head Avatars with Analysis-by-Audio-Synthesis
von: Daněček, Radek, et al.
Veröffentlicht: (2025)
von: Daněček, Radek, et al.
Veröffentlicht: (2025)
Learn2Talk: 3D Talking Face Learns from 2D Talking Face
von: Zhuang, Yixiang, et al.
Veröffentlicht: (2024)
von: Zhuang, Yixiang, et al.
Veröffentlicht: (2024)
MoDA: Multi-modal Diffusion Architecture for Talking Head Generation
von: Li, Xinyang, et al.
Veröffentlicht: (2025)
von: Li, Xinyang, et al.
Veröffentlicht: (2025)
Learning Disentangled Speech- and Expression-Driven Blendshapes for 3D Talking Face Animation
von: Mao, Yuxiang, et al.
Veröffentlicht: (2025)
von: Mao, Yuxiang, et al.
Veröffentlicht: (2025)
One Shot, One Talk: Whole-body Talking Avatar from a Single Image
von: Xiang, Jun, et al.
Veröffentlicht: (2024)
von: Xiang, Jun, et al.
Veröffentlicht: (2024)
Perceptually Accurate 3D Talking Head Generation: New Definitions, Speech-Mesh Representation, and Evaluation Metrics
von: Chae-Yeon, Lee, et al.
Veröffentlicht: (2025)
von: Chae-Yeon, Lee, et al.
Veröffentlicht: (2025)
StyGazeTalk: Learning Stylized Generation of Gaze and Head Dynamics
von: Shi, Chengwei, et al.
Veröffentlicht: (2025)
von: Shi, Chengwei, et al.
Veröffentlicht: (2025)
3D Gaussian Blendshapes for Head Avatar Animation
von: Ma, Shengjie, et al.
Veröffentlicht: (2024)
von: Ma, Shengjie, et al.
Veröffentlicht: (2024)
PASE: Phoneme-Aware Speech Encoder to Improve Lip Sync Accuracy for Talking Head Synthesis
von: Huang, Yihuan, et al.
Veröffentlicht: (2025)
von: Huang, Yihuan, et al.
Veröffentlicht: (2025)
PhysTalk: Language-driven Real-time Physics in 3D Gaussian Scenes
von: Collorone, Luca, et al.
Veröffentlicht: (2025)
von: Collorone, Luca, et al.
Veröffentlicht: (2025)
Make-It-Animatable: An Efficient Framework for Authoring Animation-Ready 3D Characters
von: Guo, Zhiyang, et al.
Veröffentlicht: (2024)
von: Guo, Zhiyang, et al.
Veröffentlicht: (2024)
Generalizable and Animatable Gaussian Head Avatar
von: Chu, Xuangeng, et al.
Veröffentlicht: (2024)
von: Chu, Xuangeng, et al.
Veröffentlicht: (2024)
Puppeteer: Rig and Animate Your 3D Models
von: Song, Chaoyue, et al.
Veröffentlicht: (2025)
von: Song, Chaoyue, et al.
Veröffentlicht: (2025)
SIRR-LMM: Single-image Reflection Removal via Large Multimodal Model
von: Guo, Yu, et al.
Veröffentlicht: (2026)
von: Guo, Yu, et al.
Veröffentlicht: (2026)
Animatable 3D Gaussian: Fast and High-Quality Reconstruction of Multiple Human Avatars
von: Liu, Yang, et al.
Veröffentlicht: (2023)
von: Liu, Yang, et al.
Veröffentlicht: (2023)
Lightning Fast Caching-based Parallel Denoising Prediction for Accelerating Talking Head Generation
von: Long, Jianzhi, et al.
Veröffentlicht: (2025)
von: Long, Jianzhi, et al.
Veröffentlicht: (2025)
ProgressiveAvatars: Progressive Animatable 3D Gaussian Avatars
von: Song, Kaiwen, et al.
Veröffentlicht: (2026)
von: Song, Kaiwen, et al.
Veröffentlicht: (2026)
PersonaTalk: Bring Attention to Your Persona in Visual Dubbing
von: Zhang, Longhao, et al.
Veröffentlicht: (2024)
von: Zhang, Longhao, et al.
Veröffentlicht: (2024)
Audio-Plane: Audio Factorization Plane Gaussian Splatting for Real-Time Talking Head Synthesis
von: Shen, Shuai, et al.
Veröffentlicht: (2025)
von: Shen, Shuai, et al.
Veröffentlicht: (2025)
READ: Real-time and Efficient Asynchronous Diffusion for Audio-driven Talking Head Generation
von: Wang, Haotian, et al.
Veröffentlicht: (2025)
von: Wang, Haotian, et al.
Veröffentlicht: (2025)
PSAvatar: A Point-based Shape Model for Real-Time Head Avatar Animation with 3D Gaussian Splatting
von: Zhao, Zhongyuan, et al.
Veröffentlicht: (2024)
von: Zhao, Zhongyuan, et al.
Veröffentlicht: (2024)
EditYourself: Audio-Driven Generation and Manipulation of Talking Head Videos with Diffusion Transformers
von: Flynn, John, et al.
Veröffentlicht: (2026)
von: Flynn, John, et al.
Veröffentlicht: (2026)
EmoFace: Audio-driven Emotional 3D Face Animation
von: Liu, Chang, et al.
Veröffentlicht: (2024)
von: Liu, Chang, et al.
Veröffentlicht: (2024)
Human Geometry Distribution for 3D Animation Generation
von: Tang, Xiangjun, et al.
Veröffentlicht: (2025)
von: Tang, Xiangjun, et al.
Veröffentlicht: (2025)
TOPOS: High-Fidelity and Efficient Industry-Grade 3D Head Generation
von: Xiong, Bojun, et al.
Veröffentlicht: (2026)
von: Xiong, Bojun, et al.
Veröffentlicht: (2026)
NeRF-3DTalker: Neural Radiance Field with 3D Prior Aided Audio Disentanglement for Talking Head Synthesis
von: Liu, Xiaoxing, et al.
Veröffentlicht: (2025)
von: Liu, Xiaoxing, et al.
Veröffentlicht: (2025)
Occlusion-robust Stylization for Drawing-based 3D Animation
von: Yoon, Sunjae, et al.
Veröffentlicht: (2025)
von: Yoon, Sunjae, et al.
Veröffentlicht: (2025)
GaussiAnimate: Reconstruct and Rig Animatable Categories with Level of Dynamics
von: Wang, Jiaxin, et al.
Veröffentlicht: (2026)
von: Wang, Jiaxin, et al.
Veröffentlicht: (2026)
Rigel3D: Rig-aware Latents for Animation-Ready 3D Asset Generation
von: Chatzis, Nikitas, et al.
Veröffentlicht: (2026)
von: Chatzis, Nikitas, et al.
Veröffentlicht: (2026)
AniGen: Unified $S^3$ Fields for Animatable 3D Asset Generation
von: Huang, Yi-Hua, et al.
Veröffentlicht: (2026)
von: Huang, Yi-Hua, et al.
Veröffentlicht: (2026)
Learning 3D Garment Animation from Trajectories of A Piece of Cloth
von: Shao, Yidi, et al.
Veröffentlicht: (2025)
von: Shao, Yidi, et al.
Veröffentlicht: (2025)
Sketch2Anim: Towards Transferring Sketch Storyboards into 3D Animation
von: Zhong, Lei, et al.
Veröffentlicht: (2025)
von: Zhong, Lei, et al.
Veröffentlicht: (2025)
FlashAvatar: High-fidelity Head Avatar with Efficient Gaussian Embedding
von: Xiang, Jun, et al.
Veröffentlicht: (2023)
von: Xiang, Jun, et al.
Veröffentlicht: (2023)
PhysAnimator: Physics-Guided Generative Cartoon Animation
von: Xie, Tianyi, et al.
Veröffentlicht: (2025)
von: Xie, Tianyi, et al.
Veröffentlicht: (2025)
FaceTalk: Audio-Driven Motion Diffusion for Neural Parametric Head Models
von: Aneja, Shivangi, et al.
Veröffentlicht: (2023)
von: Aneja, Shivangi, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Inverse Rendering for High-Genus Surface Meshes from Multi-View Images
von: Gao, Xiang, et al.
Veröffentlicht: (2025) -
ePBR: Extended PBR Materials in Image Synthesis
von: Guo, Yu, et al.
Veröffentlicht: (2025) -
Neural Geometry Image-Based Representations with Optimal Transport (OT)
von: Gao, Xiang, et al.
Veröffentlicht: (2025) -
DiffPoseTalk: Speech-Driven Stylistic 3D Facial Animation and Head Pose Generation via Diffusion Models
von: Sun, Zhiyao, et al.
Veröffentlicht: (2023) -
MultiTalk: Enhancing 3D Talking Head Generation Across Languages with Multilingual Video Dataset
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2024)