Beyond Talking -- Generating Holistic 3D Human Dyadic Motion for Communication
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Mingze, Xu, Chao, Jiang, Xinyu, Liu, Yang, Sun, Baigui, Huang, Ruqi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Combo: Co-speech holistic 3D human motion generation and efficient customizable adaptation in harmony
by: Xu, Chao, et al.
Published: (2024)
by: Xu, Chao, et al.
Published: (2024)
Topology-Agnostic Animal Motion Generation from Text Prompt
by: Chen, Keyi, et al.
Published: (2025)
by: Chen, Keyi, et al.
Published: (2025)
FaceChain-ImagineID: Freely Crafting High-Fidelity Diverse Talking Faces from Disentangled Audio
by: Xu, Chao, et al.
Published: (2024)
by: Xu, Chao, et al.
Published: (2024)
Think Before You Move: Latent Motion Reasoning for Text-to-Motion Generation
by: Qian, Yijie, et al.
Published: (2025)
by: Qian, Yijie, et al.
Published: (2025)
Dyadic Mamba: Long-term Dyadic Human Motion Synthesis
by: Tanke, Julian, et al.
Published: (2025)
by: Tanke, Julian, et al.
Published: (2025)
Holistic-Motion2D: Scalable Whole-body Human Motion Generation in 2D Space
by: Wang, Yuan, et al.
Published: (2024)
by: Wang, Yuan, et al.
Published: (2024)
SRIF: Semantic Shape Registration Empowered by Diffusion-based Image Morphing and Flow Estimation
by: Sun, Mingze, et al.
Published: (2024)
by: Sun, Mingze, et al.
Published: (2024)
DRiVE: Diffusion-based Rigging Empowers Generation of Versatile and Expressive Characters
by: Sun, Mingze, et al.
Published: (2024)
by: Sun, Mingze, et al.
Published: (2024)
Beyond Description: Cognitively Benchmarking Fine-Grained Action for Embodied Agents
by: Liu, Dayong, et al.
Published: (2025)
by: Liu, Dayong, et al.
Published: (2025)
SemTalk: Holistic Co-speech Motion Generation with Frame-level Semantic Emphasis
by: Zhang, Xiangyue, et al.
Published: (2024)
by: Zhang, Xiangyue, et al.
Published: (2024)
NFR: Neural Feature-Guided Non-Rigid Shape Registration
by: Chen, Zhangquan, et al.
Published: (2025)
by: Chen, Zhangquan, et al.
Published: (2025)
Facial Expression Generation Aligned with Human Preference for Natural Dyadic Interaction
by: Chen, Xu, et al.
Published: (2026)
by: Chen, Xu, et al.
Published: (2026)
FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis
by: Wang, Mengchao, et al.
Published: (2025)
by: Wang, Mengchao, et al.
Published: (2025)
Move-in-2D: 2D-Conditioned Human Motion Generation
by: Huang, Hsin-Ping, et al.
Published: (2024)
by: Huang, Hsin-Ping, et al.
Published: (2024)
Towards Variable and Coordinated Holistic Co-Speech Motion Generation
by: Liu, Yifei, et al.
Published: (2024)
by: Liu, Yifei, et al.
Published: (2024)
RobustMVS: Single Domain Generalized Deep Multi-view Stereo
by: Xu, Hongbin, et al.
Published: (2024)
by: Xu, Hongbin, et al.
Published: (2024)
DyStream: Streaming Dyadic Talking Heads Generation via Flow Matching-based Autoregressive Model
by: Chen, Bohong, et al.
Published: (2025)
by: Chen, Bohong, et al.
Published: (2025)
StyleDyRF: Zero-shot 4D Style Transfer for Dynamic Neural Radiance Fields
by: Xu, Hongbin, et al.
Published: (2024)
by: Xu, Hongbin, et al.
Published: (2024)
DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation
by: Chen, Junhao, et al.
Published: (2025)
by: Chen, Junhao, et al.
Published: (2025)
AVI-Talking: Learning Audio-Visual Instructions for Expressive 3D Talking Face Generation
by: Sun, Yasheng, et al.
Published: (2024)
by: Sun, Yasheng, et al.
Published: (2024)
FusionBERT: Multi-View Image-3D Retrieval via Cross-Attention Visual Fusion and Normal-Aware 3D Encoder
by: Li, Wei, et al.
Published: (2026)
by: Li, Wei, et al.
Published: (2026)
NEWTON: Agentic Planning for Physically Grounded Video Generation
by: Feng, Yuxiang, et al.
Published: (2026)
by: Feng, Yuxiang, et al.
Published: (2026)
TryOn-Adapter: Efficient Fine-Grained Clothing Identity Adaptation for High-Fidelity Virtual Try-On
by: Xing, Jiazheng, et al.
Published: (2024)
by: Xing, Jiazheng, et al.
Published: (2024)
MoDiT: Learning Highly Consistent 3D Motion Coefficients with Diffusion Transformer for Talking Head Generation
by: Wang, Yucheng, et al.
Published: (2025)
by: Wang, Yucheng, et al.
Published: (2025)
ARMO: Autoregressive Rigging for Multi-Category Objects
by: Sun, Mingze, et al.
Published: (2025)
by: Sun, Mingze, et al.
Published: (2025)
Holi-Spatial: Evolving Video Streams into Holistic 3D Spatial Intelligence
by: Gao, Yuanyuan, et al.
Published: (2026)
by: Gao, Yuanyuan, et al.
Published: (2026)
GenM$^3$: Generative Pretrained Multi-path Motion Model for Text Conditional Human Motion Generation
by: Shi, Junyu, et al.
Published: (2025)
by: Shi, Junyu, et al.
Published: (2025)
ControLRM: Fast and Controllable 3D Generation via Large Reconstruction Model
by: Xu, Hongbin, et al.
Published: (2024)
by: Xu, Hongbin, et al.
Published: (2024)
LaxMotion: Rethinking Supervision Granularity for 3D Human Motion Generation
by: Liu, Sheng, et al.
Published: (2025)
by: Liu, Sheng, et al.
Published: (2025)
OmniFit: Multi-modal 3D Body Fitting via Scale-agnostic Dense Landmark Prediction
by: Cai, Zeyu, et al.
Published: (2026)
by: Cai, Zeyu, et al.
Published: (2026)
Synergizing Motion and Appearance: Multi-Scale Compensatory Codebooks for Talking Head Video Generation
by: Zhao, Shuling, et al.
Published: (2024)
by: Zhao, Shuling, et al.
Published: (2024)
FaceChain-SuDe: Building Derived Class to Inherit Category Attributes for One-shot Subject-Driven Generation
by: Qiao, Pengchong, et al.
Published: (2024)
by: Qiao, Pengchong, et al.
Published: (2024)
LottieGPT: Tokenizing Vector Animation for Autoregressive Generation
by: Chen, Junhao, et al.
Published: (2026)
by: Chen, Junhao, et al.
Published: (2026)
ISExplore:Informative Segment Selection for Efficient Personalized 3D Talking Face Generation
by: Sun, Rui-Qing, et al.
Published: (2025)
by: Sun, Rui-Qing, et al.
Published: (2025)
SignAvatars: A Large-scale 3D Sign Language Holistic Motion Dataset and Benchmark
by: Yu, Zhengdi, et al.
Published: (2023)
by: Yu, Zhengdi, et al.
Published: (2023)
Seeing Beyond Views: Multi-View Driving Scene Video Generation with Holistic Attention
by: Lu, Hannan, et al.
Published: (2024)
by: Lu, Hannan, et al.
Published: (2024)
ControlGS: Consistent Structural Compression Control for Deployment-Aware Gaussian Splatting
by: Zhang, Fengdi, et al.
Published: (2025)
by: Zhang, Fengdi, et al.
Published: (2025)
Embracing Aleatoric Uncertainty: Generating Diverse 3D Human Motion
by: Qin, Zheng, et al.
Published: (2025)
by: Qin, Zheng, et al.
Published: (2025)
M2DAO-Talker: Harmonizing Multi-granular Motion Decoupling and Alternating Optimization for Talking-head Generation
by: Jiang, Kui, et al.
Published: (2025)
by: Jiang, Kui, et al.
Published: (2025)
MA-FSAR: Multimodal Adaptation of CLIP for Few-Shot Action Recognition
by: Xing, Jiazheng, et al.
Published: (2023)
by: Xing, Jiazheng, et al.
Published: (2023)
Similar Items
-
Combo: Co-speech holistic 3D human motion generation and efficient customizable adaptation in harmony
by: Xu, Chao, et al.
Published: (2024) -
Topology-Agnostic Animal Motion Generation from Text Prompt
by: Chen, Keyi, et al.
Published: (2025) -
FaceChain-ImagineID: Freely Crafting High-Fidelity Diverse Talking Faces from Disentangled Audio
by: Xu, Chao, et al.
Published: (2024) -
Think Before You Move: Latent Motion Reasoning for Text-to-Motion Generation
by: Qian, Yijie, et al.
Published: (2025) -
Dyadic Mamba: Long-term Dyadic Human Motion Synthesis
by: Tanke, Julian, et al.
Published: (2025)