Towards Variable and Coordinated Holistic Co-Speech Motion Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Yifei, Cao, Qiong, Wen, Yandong, Jiang, Huaiguang, Ding, Changxing |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reconstruction-Anchored Diffusion Model for Text-to-Motion Generation
by: Liu, Yifei, et al.
Published: (2026)
by: Liu, Yifei, et al.
Published: (2026)
Coordinating Multiple Conditions for Trajectory-Controlled Human Motion Generation
by: Cai, Deli, et al.
Published: (2026)
by: Cai, Deli, et al.
Published: (2026)
Democratizing High-Fidelity Co-Speech Gesture Video Generation
by: Yang, Xu, et al.
Published: (2025)
by: Yang, Xu, et al.
Published: (2025)
HolisticSemGes: Semantic Grounding of Holistic Co-Speech Gesture Generation with Contrastive Flow-Matching
by: Liu, Lanmiao, et al.
Published: (2026)
by: Liu, Lanmiao, et al.
Published: (2026)
EchoMask: Speech-Queried Attention-based Mask Modeling for Holistic Co-Speech Motion Generation
by: Zhang, Xiangyue, et al.
Published: (2025)
by: Zhang, Xiangyue, et al.
Published: (2025)
EMAGE: Towards Unified Holistic Co-Speech Gesture Generation via Expressive Masked Audio Gesture Modeling
by: Liu, Haiyang, et al.
Published: (2023)
by: Liu, Haiyang, et al.
Published: (2023)
Guiding Human-Object Interactions with Rich Geometry and Relations
by: Xue, Mengqing, et al.
Published: (2025)
by: Xue, Mengqing, et al.
Published: (2025)
SpeechAct: Towards Generating Whole-body Motion from Speech
by: Zhang, Jinsong, et al.
Published: (2023)
by: Zhang, Jinsong, et al.
Published: (2023)
CAMs as Shapley Value-based Explainers
by: Cai, Huaiguang
Published: (2025)
by: Cai, Huaiguang
Published: (2025)
Beyond Talking -- Generating Holistic 3D Human Dyadic Motion for Communication
by: Sun, Mingze, et al.
Published: (2024)
by: Sun, Mingze, et al.
Published: (2024)
SemTalk: Holistic Co-speech Motion Generation with Frame-level Semantic Emphasis
by: Zhang, Xiangyue, et al.
Published: (2024)
by: Zhang, Xiangyue, et al.
Published: (2024)
Drifting Preference Optimization for One-Step Generative Models
by: Jiang, Zhou, et al.
Published: (2026)
by: Jiang, Zhou, et al.
Published: (2026)
Enabling Synergistic Full-Body Control in Prompt-Based Co-Speech Motion Generation
by: Chen, Bohong, et al.
Published: (2024)
by: Chen, Bohong, et al.
Published: (2024)
Rethinking Preference Alignment for Diffusion Models with Classifier-Free Guidance
by: Jiang, Zhou, et al.
Published: (2026)
by: Jiang, Zhou, et al.
Published: (2026)
HoloGest: Decoupled Diffusion and Motion Priors for Generating Holisticly Expressive Co-speech Gestures
by: Cheng, Yongkang, et al.
Published: (2025)
by: Cheng, Yongkang, et al.
Published: (2025)
Data-driven Head Motion Generation through Natural Gaze-Head Coordination
by: Liu, Xiaohan, et al.
Published: (2026)
by: Liu, Xiaohan, et al.
Published: (2026)
Coordinate-Based Dual-Constrained Autoregressive Motion Generation
by: Ding, Kang, et al.
Published: (2026)
by: Ding, Kang, et al.
Published: (2026)
Absolute Coordinates Make Motion Generation Easy
by: Meng, Zichong, et al.
Published: (2025)
by: Meng, Zichong, et al.
Published: (2025)
ParCo: Part-Coordinating Text-to-Motion Synthesis
by: Zou, Qiran, et al.
Published: (2024)
by: Zou, Qiran, et al.
Published: (2024)
MS-DETR: Towards Effective Video Moment Retrieval and Highlight Detection by Joint Motion-Semantic Learning
by: Ma, Hongxu, et al.
Published: (2025)
by: Ma, Hongxu, et al.
Published: (2025)
MMGT: Motion Mask Guided Two-Stage Network for Co-Speech Gesture Video Generation
by: Wang, Siyuan, et al.
Published: (2025)
by: Wang, Siyuan, et al.
Published: (2025)
Streamlined Open-Vocabulary Human-Object Interaction Detection
by: Sun, Chang, et al.
Published: (2026)
by: Sun, Chang, et al.
Published: (2026)
Holistic-Motion2D: Scalable Whole-body Human Motion Generation in 2D Space
by: Wang, Yuan, et al.
Published: (2024)
by: Wang, Yuan, et al.
Published: (2024)
THEMIS: Towards Holistic Evaluation of MLLMs for Scientific Paper Fraud Forensics
by: Ma, Tzu-Yen, et al.
Published: (2026)
by: Ma, Tzu-Yen, et al.
Published: (2026)
Lagrangian Motion Fields for Long-term Motion Generation
by: Yang, Yifei, et al.
Published: (2024)
by: Yang, Yifei, et al.
Published: (2024)
OCRVerse: Towards Holistic OCR in End-to-End Vision-Language Models
by: Zhong, Yufeng, et al.
Published: (2026)
by: Zhong, Yufeng, et al.
Published: (2026)
CoordSpeaker: Exploiting Gesture Captioning for Coordinated Caption-Empowered Co-Speech Gesture Generation
by: Fang, Fengyi, et al.
Published: (2025)
by: Fang, Fengyi, et al.
Published: (2025)
Contact-aware Human Motion Generation from Textual Descriptions
by: Ma, Sihan, et al.
Published: (2024)
by: Ma, Sihan, et al.
Published: (2024)
MotionWeaver: Holistic 4D-Anchored Framework for Multi-Humanoid Image Animation
by: Hu, Xirui, et al.
Published: (2026)
by: Hu, Xirui, et al.
Published: (2026)
Holistic Dynamic Frequency Transformer for Image Fusion and Exposure Correction
by: Shang, Xiaoke, et al.
Published: (2023)
by: Shang, Xiaoke, et al.
Published: (2023)
StoryMaker: Towards Holistic Consistent Characters in Text-to-image Generation
by: Zhou, Zhengguang, et al.
Published: (2024)
by: Zhou, Zhengguang, et al.
Published: (2024)
MotionTrack: Learning Motion Predictor for Multiple Object Tracking
by: Xiao, Changcheng, et al.
Published: (2023)
by: Xiao, Changcheng, et al.
Published: (2023)
Towards Holistic Surgical Scene Graph
by: Shin, Jongmin, et al.
Published: (2025)
by: Shin, Jongmin, et al.
Published: (2025)
Streaming Autoregressive Video Generation via Diagonal Distillation
by: Liu, Jinxiu, et al.
Published: (2026)
by: Liu, Jinxiu, et al.
Published: (2026)
Disentangled Pre-training for Human-Object Interaction Detection
by: Li, Zhuolong, et al.
Published: (2024)
by: Li, Zhuolong, et al.
Published: (2024)
ViHOI: Human-Object Interaction Synthesis with Visual Priors
by: Cai, Songjin, et al.
Published: (2026)
by: Cai, Songjin, et al.
Published: (2026)
Mitigating Error Accumulation in Co-Speech Motion Generation via Global Rotation Diffusion and Multi-Level Constraints
by: Zhang, Xiangyue, et al.
Published: (2025)
by: Zhang, Xiangyue, et al.
Published: (2025)
TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation
by: Liu, Haiyang, et al.
Published: (2024)
by: Liu, Haiyang, et al.
Published: (2024)
Effortless Active Labeling for Long-Term Test-Time Adaptation
by: Wang, Guowei, et al.
Published: (2025)
by: Wang, Guowei, et al.
Published: (2025)
Linking Modality Isolation in Heterogeneous Collaborative Perception
by: Liu, Changxing, et al.
Published: (2026)
by: Liu, Changxing, et al.
Published: (2026)
Similar Items
-
Reconstruction-Anchored Diffusion Model for Text-to-Motion Generation
by: Liu, Yifei, et al.
Published: (2026) -
Coordinating Multiple Conditions for Trajectory-Controlled Human Motion Generation
by: Cai, Deli, et al.
Published: (2026) -
Democratizing High-Fidelity Co-Speech Gesture Video Generation
by: Yang, Xu, et al.
Published: (2025) -
HolisticSemGes: Semantic Grounding of Holistic Co-Speech Gesture Generation with Contrastive Flow-Matching
by: Liu, Lanmiao, et al.
Published: (2026) -
EchoMask: Speech-Queried Attention-based Mask Modeling for Holistic Co-Speech Motion Generation
by: Zhang, Xiangyue, et al.
Published: (2025)