TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Haiyang, Yang, Xingchao, Akiyama, Tomoya, Huang, Yuantian, Li, Qiaoge, Kuriyama, Shigeru, Taketomi, Takafumi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FFHQ-Makeup: Paired Synthetic Makeup Dataset with Facial Consistency Across Multiple Styles
von: Yang, Xingchao, et al.
Veröffentlicht: (2025)
von: Yang, Xingchao, et al.
Veröffentlicht: (2025)
BeautyBank: Encoding Facial Makeup in Latent Space
von: Lu, Qianwen, et al.
Veröffentlicht: (2024)
von: Lu, Qianwen, et al.
Veröffentlicht: (2024)
FreeUV: Ground-Truth-Free Realistic Facial UV Texture Recovery via Cross-Assembly Inference Strategy
von: Yang, Xingchao, et al.
Veröffentlicht: (2025)
von: Yang, Xingchao, et al.
Veröffentlicht: (2025)
Makeup Prior Models for 3D Facial Makeup Estimation and Applications
von: Yang, Xingchao, et al.
Veröffentlicht: (2024)
von: Yang, Xingchao, et al.
Veröffentlicht: (2024)
Co-Speech Gesture Video Generation via Motion-Decoupled Diffusion Model
von: He, Xu, et al.
Veröffentlicht: (2024)
von: He, Xu, et al.
Veröffentlicht: (2024)
Cosh-DiT: Co-Speech Gesture Video Synthesis via Hybrid Audio-Visual Diffusion Transformers
von: Sun, Yasheng, et al.
Veröffentlicht: (2025)
von: Sun, Yasheng, et al.
Veröffentlicht: (2025)
Anchored Diffusion for Video Face Reenactment
von: Kligvasser, Idan, et al.
Veröffentlicht: (2024)
von: Kligvasser, Idan, et al.
Veröffentlicht: (2024)
Diverse Code Query Learning for Speech-Driven Facial Animation
von: Gu, Chunzhi, et al.
Veröffentlicht: (2024)
von: Gu, Chunzhi, et al.
Veröffentlicht: (2024)
EMAGE: Towards Unified Holistic Co-Speech Gesture Generation via Expressive Masked Audio Gesture Modeling
von: Liu, Haiyang, et al.
Veröffentlicht: (2023)
von: Liu, Haiyang, et al.
Veröffentlicht: (2023)
Neural Multi-View Self-Calibrated Photometric Stereo without Photometric Stereo Cues
von: Cao, Xu, et al.
Veröffentlicht: (2025)
von: Cao, Xu, et al.
Veröffentlicht: (2025)
GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling
von: Liu, Pinxin, et al.
Veröffentlicht: (2025)
von: Liu, Pinxin, et al.
Veröffentlicht: (2025)
Orientation-Aware Leg Movement Learning for Action-Driven Human Motion Prediction
von: Gu, Chunzhi, et al.
Veröffentlicht: (2023)
von: Gu, Chunzhi, et al.
Veröffentlicht: (2023)
PersonaGesture: Single-Reference Co-Speech Gesture Personalization for Unseen Speakers
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2026)
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2026)
Democratizing High-Fidelity Co-Speech Gesture Video Generation
von: Yang, Xu, et al.
Veröffentlicht: (2025)
von: Yang, Xu, et al.
Veröffentlicht: (2025)
AnyAct: Towards Human Reenactment of Character Motion From Video
von: Chen, Liuhan, et al.
Veröffentlicht: (2026)
von: Chen, Liuhan, et al.
Veröffentlicht: (2026)
PersonaGest: Personalized Co-Speech Gesture Generation with Semantic-Guided Hierarchical Motion Representation
von: Zhao, Junchuan, et al.
Veröffentlicht: (2026)
von: Zhao, Junchuan, et al.
Veröffentlicht: (2026)
Motion-aware Latent Diffusion Models for Video Frame Interpolation
von: Huang, Zhilin, et al.
Veröffentlicht: (2024)
von: Huang, Zhilin, et al.
Veröffentlicht: (2024)
MMGT: Motion Mask Guided Two-Stage Network for Co-Speech Gesture Video Generation
von: Wang, Siyuan, et al.
Veröffentlicht: (2025)
von: Wang, Siyuan, et al.
Veröffentlicht: (2025)
Reenact Anything: Semantic Video Motion Transfer Using Motion-Textual Inversion
von: Kansy, Manuel, et al.
Veröffentlicht: (2024)
von: Kansy, Manuel, et al.
Veröffentlicht: (2024)
Contextual Gesture: Co-Speech Gesture Video Generation through Context-aware Gesture Representation
von: Liu, Pinxin, et al.
Veröffentlicht: (2025)
von: Liu, Pinxin, et al.
Veröffentlicht: (2025)
EmotionGesture: Audio-Driven Diverse Emotional Co-Speech 3D Gesture Generation
von: Qi, Xingqun, et al.
Veröffentlicht: (2023)
von: Qi, Xingqun, et al.
Veröffentlicht: (2023)
Intentional Gesture: Deliver Your Intentions with Gestures for Speech
von: Liu, Pinxin, et al.
Veröffentlicht: (2025)
von: Liu, Pinxin, et al.
Veröffentlicht: (2025)
LiveGesture Streamable Co-Speech Gesture Generation Model
von: Saleem, Muhammad Usama, et al.
Veröffentlicht: (2026)
von: Saleem, Muhammad Usama, et al.
Veröffentlicht: (2026)
HoloGest: Decoupled Diffusion and Motion Priors for Generating Holisticly Expressive Co-speech Gestures
von: Cheng, Yongkang, et al.
Veröffentlicht: (2025)
von: Cheng, Yongkang, et al.
Veröffentlicht: (2025)
Moment-Reenacting: Inverse Motion Degradation with Cross-shutter Guidance
von: Ji, Xiang, et al.
Veröffentlicht: (2026)
von: Ji, Xiang, et al.
Veröffentlicht: (2026)
Navigating Large-Pose Challenge for High-Fidelity Face Reenactment with Video Diffusion Model
von: Guo, Mingtao, et al.
Veröffentlicht: (2025)
von: Guo, Mingtao, et al.
Veröffentlicht: (2025)
CoordSpeaker: Exploiting Gesture Captioning for Coordinated Caption-Empowered Co-Speech Gesture Generation
von: Fang, Fengyi, et al.
Veröffentlicht: (2025)
von: Fang, Fengyi, et al.
Veröffentlicht: (2025)
Re-HOLD: Video Hand Object Interaction Reenactment via adaptive Layout-instructed Diffusion Model
von: Fan, Yingying, et al.
Veröffentlicht: (2025)
von: Fan, Yingying, et al.
Veröffentlicht: (2025)
DiffTED: One-shot Audio-driven TED Talk Video Generation with Diffusion-based Co-speech Gestures
von: Hogue, Steven, et al.
Veröffentlicht: (2024)
von: Hogue, Steven, et al.
Veröffentlicht: (2024)
Recognizing Co-Speech Gestures in-the-Wild
von: Hegde, Sindhu B, et al.
Veröffentlicht: (2026)
von: Hegde, Sindhu B, et al.
Veröffentlicht: (2026)
Co-speech Gesture Video Generation via Motion-Based Graph Retrieval
von: Song, Yafei, et al.
Veröffentlicht: (2025)
von: Song, Yafei, et al.
Veröffentlicht: (2025)
ConvoFusion: Multi-Modal Conversational Diffusion for Co-Speech Gesture Synthesis
von: Mughal, Muhammad Hamza, et al.
Veröffentlicht: (2024)
von: Mughal, Muhammad Hamza, et al.
Veröffentlicht: (2024)
Adapting Image-to-Video Diffusion Models for Large-Motion Frame Interpolation
von: Jin, Luoxu, et al.
Veröffentlicht: (2024)
von: Jin, Luoxu, et al.
Veröffentlicht: (2024)
DiffusionAct: Controllable Diffusion Autoencoder for One-shot Face Reenactment
von: Bounareli, Stella, et al.
Veröffentlicht: (2024)
von: Bounareli, Stella, et al.
Veröffentlicht: (2024)
Hierarchical Flow Diffusion for Efficient Frame Interpolation
von: Hai, Yang, et al.
Veröffentlicht: (2025)
von: Hai, Yang, et al.
Veröffentlicht: (2025)
Streaming Generation of Co-Speech Gestures via Accelerated Rolling Diffusion
von: Vu, Evgeniia, et al.
Veröffentlicht: (2025)
von: Vu, Evgeniia, et al.
Veröffentlicht: (2025)
iDiT-HOI: Inpainting-based Hand Object Interaction Reenactment via Video Diffusion Transformer
von: Shen, Zhelun, et al.
Veröffentlicht: (2025)
von: Shen, Zhelun, et al.
Veröffentlicht: (2025)
Joint Co-Speech Gesture and Expressive Talking Face Generation using Diffusion with Adapters
von: Hogue, Steven, et al.
Veröffentlicht: (2024)
von: Hogue, Steven, et al.
Veröffentlicht: (2024)
Motion-Aware Video Frame Interpolation
von: Han, Pengfei, et al.
Veröffentlicht: (2024)
von: Han, Pengfei, et al.
Veröffentlicht: (2024)
Towards Unified Co-Speech Gesture Generation via Hierarchical Implicit Periodicity Learning
von: Guo, Xin, et al.
Veröffentlicht: (2025)
von: Guo, Xin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
FFHQ-Makeup: Paired Synthetic Makeup Dataset with Facial Consistency Across Multiple Styles
von: Yang, Xingchao, et al.
Veröffentlicht: (2025) -
BeautyBank: Encoding Facial Makeup in Latent Space
von: Lu, Qianwen, et al.
Veröffentlicht: (2024) -
FreeUV: Ground-Truth-Free Realistic Facial UV Texture Recovery via Cross-Assembly Inference Strategy
von: Yang, Xingchao, et al.
Veröffentlicht: (2025) -
Makeup Prior Models for 3D Facial Makeup Estimation and Applications
von: Yang, Xingchao, et al.
Veröffentlicht: (2024) -
Co-Speech Gesture Video Generation via Motion-Decoupled Diffusion Model
von: He, Xu, et al.
Veröffentlicht: (2024)