Enabling Synergistic Full-Body Control in Prompt-Based Co-Speech Motion Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Bohong, Li, Yumeng, Ding, Yao-Xiang, Shao, Tianjia, Zhou, Kun |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Motion-example-controlled Co-speech Gesture Generation Leveraging Large Language Models
by: Chen, Bohong, et al.
Published: (2025)
by: Chen, Bohong, et al.
Published: (2025)
Autonomous Imagination: Closed-Loop Decomposition of Visual-to-Textual Conversion in Visual Reasoning for Multimodal Large Language Models
by: Liu, Jingming, et al.
Published: (2024)
by: Liu, Jingming, et al.
Published: (2024)
When Gaussian Meets Surfel: Ultra-fast High-fidelity Radiance Field Rendering
by: Ye, Keyang, et al.
Published: (2025)
by: Ye, Keyang, et al.
Published: (2025)
Gaussian-Plus-SDF SLAM: High-fidelity 3D Reconstruction at 150+ fps
by: Peng, Zhexi, et al.
Published: (2025)
by: Peng, Zhexi, et al.
Published: (2025)
Slot-ID: Identity-Preserving Video Generation from Reference Videos via Slot-Based Temporal Identity Encoding
by: Lai, Yixuan, et al.
Published: (2026)
by: Lai, Yixuan, et al.
Published: (2026)
Sparse-to-Complete: From Sparse Image Captures to Complete 3D Scenes
by: Shen, Yiyang, et al.
Published: (2026)
by: Shen, Yiyang, et al.
Published: (2026)
High-fidelity 3D Object Generation from Single Image with RGBN-Volume Gaussian Reconstruction Model
by: Shen, Yiyang, et al.
Published: (2025)
by: Shen, Yiyang, et al.
Published: (2025)
TR-Gaussians: High-fidelity Real-time Rendering of Planar Transmission and Reflection with 3D Gaussian Splatting
by: Liu, Yong, et al.
Published: (2025)
by: Liu, Yong, et al.
Published: (2025)
3D Gaussian Blendshapes for Head Avatar Animation
by: Ma, Shengjie, et al.
Published: (2024)
by: Ma, Shengjie, et al.
Published: (2024)
Real-time High-fidelity Gaussian Human Avatars with Position-based Interpolation of Spatially Distributed MLPs
by: Zhan, Youyi, et al.
Published: (2025)
by: Zhan, Youyi, et al.
Published: (2025)
High-Fidelity Mobile Avatars with Pruned Local Blendshapes
by: Zhan, Youyi, et al.
Published: (2026)
by: Zhan, Youyi, et al.
Published: (2026)
Full-Body Motion Reconstruction with Sparse Sensing from Graph Perspective
by: Yao, Feiyu, et al.
Published: (2024)
by: Yao, Feiyu, et al.
Published: (2024)
Pattern Guided UV Recovery for Realistic Video Garment Texturing
by: Zhan, Youyi, et al.
Published: (2024)
by: Zhan, Youyi, et al.
Published: (2024)
Interactive Rendering of Relightable and Animatable Gaussian Avatars
by: Zhan, Youyi, et al.
Published: (2024)
by: Zhan, Youyi, et al.
Published: (2024)
Towards Variable and Coordinated Holistic Co-Speech Motion Generation
by: Liu, Yifei, et al.
Published: (2024)
by: Liu, Yifei, et al.
Published: (2024)
Motion Prompting: Controlling Video Generation with Motion Trajectories
by: Geng, Daniel, et al.
Published: (2024)
by: Geng, Daniel, et al.
Published: (2024)
SpeechAct: Towards Generating Whole-body Motion from Speech
by: Zhang, Jinsong, et al.
Published: (2023)
by: Zhang, Jinsong, et al.
Published: (2023)
GenesisTex2: Stable, Consistent and High-Quality Text-to-Texture Generation
by: Lu, Jiawei, et al.
Published: (2024)
by: Lu, Jiawei, et al.
Published: (2024)
DyStream: Streaming Dyadic Talking Heads Generation via Flow Matching-based Autoregressive Model
by: Chen, Bohong, et al.
Published: (2025)
by: Chen, Bohong, et al.
Published: (2025)
FUSION: Full-Body Unified Motion Prior for Body and Hands via Diffusion
by: Duran, Enes, et al.
Published: (2026)
by: Duran, Enes, et al.
Published: (2026)
Interactive Humanoid: Online Full-Body Motion Reaction Synthesis with Social Affordance Canonicalization and Forecasting
by: Liu, Yunze, et al.
Published: (2023)
by: Liu, Yunze, et al.
Published: (2023)
RTG-SLAM: Real-time 3D Reconstruction at Scale using Gaussian Splatting
by: Peng, Zhexi, et al.
Published: (2024)
by: Peng, Zhexi, et al.
Published: (2024)
InteracTalker: Prompt-Based Human-Object Interaction with Co-Speech Gesture Generation
by: Rajan, Sreehari, et al.
Published: (2025)
by: Rajan, Sreehari, et al.
Published: (2025)
MotionChain: Conversational Motion Controllers via Multimodal Prompts
by: Jiang, Biao, et al.
Published: (2024)
by: Jiang, Biao, et al.
Published: (2024)
HMPDM: A Diffusion Model for Driving Video Prediction with Historical Motion Priors
by: Li, Ke, et al.
Published: (2026)
by: Li, Ke, et al.
Published: (2026)
FreeFuse: Multi-Subject LoRA Fusion via Adaptive Token-Level Routing at Test Time
by: Liu, Yaoli, et al.
Published: (2025)
by: Liu, Yaoli, et al.
Published: (2025)
DEGAS: Detailed Expressions on Full-Body Gaussian Avatars
by: Shao, Zhijing, et al.
Published: (2024)
by: Shao, Zhijing, et al.
Published: (2024)
Autoregressive Image Generation with Vision Full-view Prompt
by: Cai, Miaomiao, et al.
Published: (2025)
by: Cai, Miaomiao, et al.
Published: (2025)
Efficient 3D Full-Body Motion Generation from Sparse Tracking Inputs with Temporal Windows
by: Angelis, Georgios Fotios, et al.
Published: (2025)
by: Angelis, Georgios Fotios, et al.
Published: (2025)
Controllable Text-to-Motion Generation via Modular Body-Part Phase Control
by: Dai, Minyue, et al.
Published: (2026)
by: Dai, Minyue, et al.
Published: (2026)
Generating Attribute-Aware Human Motions from Textual Prompt
by: Wang, Xinghan, et al.
Published: (2025)
by: Wang, Xinghan, et al.
Published: (2025)
Understanding the Vulnerability of Skeleton-based Human Activity Recognition via Black-box Attack
by: Diao, Yunfeng, et al.
Published: (2022)
by: Diao, Yunfeng, et al.
Published: (2022)
SymphoMotion: Joint Control of Camera Motion and Object Dynamics for Coherent Video Generation
by: Zhang, Guiyu, et al.
Published: (2026)
by: Zhang, Guiyu, et al.
Published: (2026)
OmniDrag: Enabling Motion Control for Omnidirectional Image-to-Video Generation
by: Li, Weiqi, et al.
Published: (2024)
by: Li, Weiqi, et al.
Published: (2024)
Beyond Full Labels: Energy-Double-Guided Single-Point Prompt for Infrared Small Target Label Generation
by: Yuan, Shuai, et al.
Published: (2024)
by: Yuan, Shuai, et al.
Published: (2024)
CoMo: Compositional Motion Customization for Text-to-Video Generation
by: Xu, Youcan, et al.
Published: (2025)
by: Xu, Youcan, et al.
Published: (2025)
EgoControl: Controllable Egocentric Video Generation via 3D Full-Body Poses
by: Pallotta, Enrico, et al.
Published: (2025)
by: Pallotta, Enrico, et al.
Published: (2025)
Gaussian Splashing: Unified Particles for Versatile Motion Synthesis and Rendering
by: Feng, Yutao, et al.
Published: (2024)
by: Feng, Yutao, et al.
Published: (2024)
HOSIG: Full-Body Human-Object-Scene Interaction Generation with Hierarchical Scene Perception
by: Yao, Wei, et al.
Published: (2025)
by: Yao, Wei, et al.
Published: (2025)
ElastoGen: 4D Generative Elastodynamics
by: Feng, Yutao, et al.
Published: (2024)
by: Feng, Yutao, et al.
Published: (2024)
Similar Items
-
Motion-example-controlled Co-speech Gesture Generation Leveraging Large Language Models
by: Chen, Bohong, et al.
Published: (2025) -
Autonomous Imagination: Closed-Loop Decomposition of Visual-to-Textual Conversion in Visual Reasoning for Multimodal Large Language Models
by: Liu, Jingming, et al.
Published: (2024) -
When Gaussian Meets Surfel: Ultra-fast High-fidelity Radiance Field Rendering
by: Ye, Keyang, et al.
Published: (2025) -
Gaussian-Plus-SDF SLAM: High-fidelity 3D Reconstruction at 150+ fps
by: Peng, Zhexi, et al.
Published: (2025) -
Slot-ID: Identity-Preserving Video Generation from Reference Videos via Slot-Based Temporal Identity Encoding
by: Lai, Yixuan, et al.
Published: (2026)