EMAGE: Towards Unified Holistic Co-Speech Gesture Generation via Expressive Masked Audio Gesture Modeling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Haiyang, Zhu, Zihao, Becherini, Giorgio, Peng, Yichen, Su, Mingyang, Zhou, You, Zhe, Xuefei, Iwamoto, Naoya, Zheng, Bo, Black, Michael J. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PersonaGesture: Single-Reference Co-Speech Gesture Personalization for Unseen Speakers
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2026)
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2026)
3DGesPolicy: Phoneme-Aware Holistic Co-Speech Gesture Generation Based on Action Control
von: Sha, Xuanmeng, et al.
Veröffentlicht: (2026)
von: Sha, Xuanmeng, et al.
Veröffentlicht: (2026)
Intentional Gesture: Deliver Your Intentions with Gestures for Speech
von: Liu, Pinxin, et al.
Veröffentlicht: (2025)
von: Liu, Pinxin, et al.
Veröffentlicht: (2025)
GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling
von: Liu, Pinxin, et al.
Veröffentlicht: (2025)
von: Liu, Pinxin, et al.
Veröffentlicht: (2025)
EmotionGesture: Audio-Driven Diverse Emotional Co-Speech 3D Gesture Generation
von: Qi, Xingqun, et al.
Veröffentlicht: (2023)
von: Qi, Xingqun, et al.
Veröffentlicht: (2023)
HoloGest: Decoupled Diffusion and Motion Priors for Generating Holisticly Expressive Co-speech Gestures
von: Cheng, Yongkang, et al.
Veröffentlicht: (2025)
von: Cheng, Yongkang, et al.
Veröffentlicht: (2025)
TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation
von: Liu, Haiyang, et al.
Veröffentlicht: (2024)
von: Liu, Haiyang, et al.
Veröffentlicht: (2024)
LiveGesture Streamable Co-Speech Gesture Generation Model
von: Saleem, Muhammad Usama, et al.
Veröffentlicht: (2026)
von: Saleem, Muhammad Usama, et al.
Veröffentlicht: (2026)
MIBURI: Towards Expressive Interactive Gesture Synthesis
von: Mughal, M. Hamza, et al.
Veröffentlicht: (2026)
von: Mughal, M. Hamza, et al.
Veröffentlicht: (2026)
HolisticSemGes: Semantic Grounding of Holistic Co-Speech Gesture Generation with Contrastive Flow-Matching
von: Liu, Lanmiao, et al.
Veröffentlicht: (2026)
von: Liu, Lanmiao, et al.
Veröffentlicht: (2026)
Contextual Gesture: Co-Speech Gesture Video Generation through Context-aware Gesture Representation
von: Liu, Pinxin, et al.
Veröffentlicht: (2025)
von: Liu, Pinxin, et al.
Veröffentlicht: (2025)
Joint Co-Speech Gesture and Expressive Talking Face Generation using Diffusion with Adapters
von: Hogue, Steven, et al.
Veröffentlicht: (2024)
von: Hogue, Steven, et al.
Veröffentlicht: (2024)
Towards Unified Co-Speech Gesture Generation via Hierarchical Implicit Periodicity Learning
von: Guo, Xin, et al.
Veröffentlicht: (2025)
von: Guo, Xin, et al.
Veröffentlicht: (2025)
CoCoGesture: Toward Coherent Co-speech 3D Gesture Generation in the Wild
von: Qi, Xingqun, et al.
Veröffentlicht: (2024)
von: Qi, Xingqun, et al.
Veröffentlicht: (2024)
Recognizing Co-Speech Gestures in-the-Wild
von: Hegde, Sindhu B, et al.
Veröffentlicht: (2026)
von: Hegde, Sindhu B, et al.
Veröffentlicht: (2026)
Conveying Meaning through Gestures: An Investigation into Semantic Co-Speech Gesture Generation
von: Voss, Hendric, et al.
Veröffentlicht: (2025)
von: Voss, Hendric, et al.
Veröffentlicht: (2025)
Speech2AffectiveGestures: Synthesizing Co-Speech Gestures with Generative Adversarial Affective Expression Learning
von: Bhattacharya, Uttaran, et al.
Veröffentlicht: (2021)
von: Bhattacharya, Uttaran, et al.
Veröffentlicht: (2021)
Gesture2Speech: How Far Can Hand Movements Shape Expressive Speech?
von: Kumar, Lokesh, et al.
Veröffentlicht: (2026)
von: Kumar, Lokesh, et al.
Veröffentlicht: (2026)
EmoDiffGes: Emotion‐Aware Co‐Speech Holistic Gesture Generation with Progressive Synergistic Diffusion
von: Xinru Li, et al.
Veröffentlicht: (2025)
von: Xinru Li, et al.
Veröffentlicht: (2025)
CoordSpeaker: Exploiting Gesture Captioning for Coordinated Caption-Empowered Co-Speech Gesture Generation
von: Fang, Fengyi, et al.
Veröffentlicht: (2025)
von: Fang, Fengyi, et al.
Veröffentlicht: (2025)
DuoGesture: Neuro-Inspired and Biomechanically Informed Dual-Stream Co-Speech Gesture Generation
von: Paar, Ferdinand, et al.
Veröffentlicht: (2026)
von: Paar, Ferdinand, et al.
Veröffentlicht: (2026)
MambaGesture: Enhancing Co-Speech Gesture Generation with Mamba and Disentangled Multi-Modality Fusion
von: Fu, Chencan, et al.
Veröffentlicht: (2024)
von: Fu, Chencan, et al.
Veröffentlicht: (2024)
ExGes: Expressive Human Motion Retrieval and Modulation for Audio-Driven Gesture Synthesis
von: Zhou, Xukun, et al.
Veröffentlicht: (2025)
von: Zhou, Xukun, et al.
Veröffentlicht: (2025)
MMGT: Motion Mask Guided Two-Stage Network for Co-Speech Gesture Video Generation
von: Wang, Siyuan, et al.
Veröffentlicht: (2025)
von: Wang, Siyuan, et al.
Veröffentlicht: (2025)
MDT-A2G: Exploring Masked Diffusion Transformers for Co-Speech Gesture Generation
von: Mao, Xiaofeng, et al.
Veröffentlicht: (2024)
von: Mao, Xiaofeng, et al.
Veröffentlicht: (2024)
Speech-Gesture GAN: Gesture Generation for Robots and Embodied Agents
von: Liu, Carson Yu, et al.
Veröffentlicht: (2023)
von: Liu, Carson Yu, et al.
Veröffentlicht: (2023)
A Unified Editing Method for Co-Speech Gesture Generation via Diffusion Inversion
von: Zhao, Zeyu, et al.
Veröffentlicht: (2024)
von: Zhao, Zeyu, et al.
Veröffentlicht: (2024)
DiM-Gesture: Co-Speech Gesture Generation with Adaptive Layer Normalization Mamba-2 framework
von: Zhang, Fan, et al.
Veröffentlicht: (2024)
von: Zhang, Fan, et al.
Veröffentlicht: (2024)
Co$^{3}$Gesture: Towards Coherent Concurrent Co-speech 3D Gesture Generation with Interactive Diffusion
von: Qi, Xingqun, et al.
Veröffentlicht: (2025)
von: Qi, Xingqun, et al.
Veröffentlicht: (2025)
Cosh-DiT: Co-Speech Gesture Video Synthesis via Hybrid Audio-Visual Diffusion Transformers
von: Sun, Yasheng, et al.
Veröffentlicht: (2025)
von: Sun, Yasheng, et al.
Veröffentlicht: (2025)
Democratizing High-Fidelity Co-Speech Gesture Video Generation
von: Yang, Xu, et al.
Veröffentlicht: (2025)
von: Yang, Xu, et al.
Veröffentlicht: (2025)
Learning Co-Speech Gesture for Multimodal Aphasia Type Detection
von: Lee, Daeun, et al.
Veröffentlicht: (2023)
von: Lee, Daeun, et al.
Veröffentlicht: (2023)
Semantic Gesticulator: Semantics-Aware Co-Speech Gesture Synthesis
von: Zhang, Zeyi, et al.
Veröffentlicht: (2024)
von: Zhang, Zeyi, et al.
Veröffentlicht: (2024)
Gelina: Unified Speech and Gesture Synthesis via Interleaved Token Prediction
von: Guichoux, Téo, et al.
Veröffentlicht: (2025)
von: Guichoux, Téo, et al.
Veröffentlicht: (2025)
EchoMask: Speech-Queried Attention-based Mask Modeling for Holistic Co-Speech Motion Generation
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2025)
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2025)
2D or not 2D: How Does the Dimensionality of Gesture Representation Affect 3D Co-Speech Gesture Generation?
von: Guichoux, Téo, et al.
Veröffentlicht: (2024)
von: Guichoux, Téo, et al.
Veröffentlicht: (2024)
Discovering Dynamical Laws for Speech Gestures
von: Sam Kirkham
Veröffentlicht: (2025)
von: Sam Kirkham
Veröffentlicht: (2025)
Semantic Motion Anchors: Bridging Motion and Meaning in Co-Speech Gestures
von: Suresh, Varsha, et al.
Veröffentlicht: (2026)
von: Suresh, Varsha, et al.
Veröffentlicht: (2026)
Co-Speech Gesture Detection through Multi-Phase Sequence Labeling
von: Ghaleb, Esam, et al.
Veröffentlicht: (2023)
von: Ghaleb, Esam, et al.
Veröffentlicht: (2023)
Streaming Generation of Co-Speech Gestures via Accelerated Rolling Diffusion
von: Vu, Evgeniia, et al.
Veröffentlicht: (2025)
von: Vu, Evgeniia, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PersonaGesture: Single-Reference Co-Speech Gesture Personalization for Unseen Speakers
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2026) -
3DGesPolicy: Phoneme-Aware Holistic Co-Speech Gesture Generation Based on Action Control
von: Sha, Xuanmeng, et al.
Veröffentlicht: (2026) -
Intentional Gesture: Deliver Your Intentions with Gestures for Speech
von: Liu, Pinxin, et al.
Veröffentlicht: (2025) -
GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling
von: Liu, Pinxin, et al.
Veröffentlicht: (2025) -
EmotionGesture: Audio-Driven Diverse Emotional Co-Speech 3D Gesture Generation
von: Qi, Xingqun, et al.
Veröffentlicht: (2023)