Enabling Synergistic Full-Body Control in Prompt-Based Co-Speech Motion Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Bohong, Li, Yumeng, Ding, Yao-Xiang, Shao, Tianjia, Zhou, Kun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Motion-example-controlled Co-speech Gesture Generation Leveraging Large Language Models
di: Chen, Bohong, et al.
Pubblicazione: (2025)
di: Chen, Bohong, et al.
Pubblicazione: (2025)
Autonomous Imagination: Closed-Loop Decomposition of Visual-to-Textual Conversion in Visual Reasoning for Multimodal Large Language Models
di: Liu, Jingming, et al.
Pubblicazione: (2024)
di: Liu, Jingming, et al.
Pubblicazione: (2024)
When Gaussian Meets Surfel: Ultra-fast High-fidelity Radiance Field Rendering
di: Ye, Keyang, et al.
Pubblicazione: (2025)
di: Ye, Keyang, et al.
Pubblicazione: (2025)
Gaussian-Plus-SDF SLAM: High-fidelity 3D Reconstruction at 150+ fps
di: Peng, Zhexi, et al.
Pubblicazione: (2025)
di: Peng, Zhexi, et al.
Pubblicazione: (2025)
Slot-ID: Identity-Preserving Video Generation from Reference Videos via Slot-Based Temporal Identity Encoding
di: Lai, Yixuan, et al.
Pubblicazione: (2026)
di: Lai, Yixuan, et al.
Pubblicazione: (2026)
Sparse-to-Complete: From Sparse Image Captures to Complete 3D Scenes
di: Shen, Yiyang, et al.
Pubblicazione: (2026)
di: Shen, Yiyang, et al.
Pubblicazione: (2026)
High-fidelity 3D Object Generation from Single Image with RGBN-Volume Gaussian Reconstruction Model
di: Shen, Yiyang, et al.
Pubblicazione: (2025)
di: Shen, Yiyang, et al.
Pubblicazione: (2025)
TR-Gaussians: High-fidelity Real-time Rendering of Planar Transmission and Reflection with 3D Gaussian Splatting
di: Liu, Yong, et al.
Pubblicazione: (2025)
di: Liu, Yong, et al.
Pubblicazione: (2025)
3D Gaussian Blendshapes for Head Avatar Animation
di: Ma, Shengjie, et al.
Pubblicazione: (2024)
di: Ma, Shengjie, et al.
Pubblicazione: (2024)
Real-time High-fidelity Gaussian Human Avatars with Position-based Interpolation of Spatially Distributed MLPs
di: Zhan, Youyi, et al.
Pubblicazione: (2025)
di: Zhan, Youyi, et al.
Pubblicazione: (2025)
High-Fidelity Mobile Avatars with Pruned Local Blendshapes
di: Zhan, Youyi, et al.
Pubblicazione: (2026)
di: Zhan, Youyi, et al.
Pubblicazione: (2026)
Full-Body Motion Reconstruction with Sparse Sensing from Graph Perspective
di: Yao, Feiyu, et al.
Pubblicazione: (2024)
di: Yao, Feiyu, et al.
Pubblicazione: (2024)
Pattern Guided UV Recovery for Realistic Video Garment Texturing
di: Zhan, Youyi, et al.
Pubblicazione: (2024)
di: Zhan, Youyi, et al.
Pubblicazione: (2024)
Interactive Rendering of Relightable and Animatable Gaussian Avatars
di: Zhan, Youyi, et al.
Pubblicazione: (2024)
di: Zhan, Youyi, et al.
Pubblicazione: (2024)
Towards Variable and Coordinated Holistic Co-Speech Motion Generation
di: Liu, Yifei, et al.
Pubblicazione: (2024)
di: Liu, Yifei, et al.
Pubblicazione: (2024)
Motion Prompting: Controlling Video Generation with Motion Trajectories
di: Geng, Daniel, et al.
Pubblicazione: (2024)
di: Geng, Daniel, et al.
Pubblicazione: (2024)
SpeechAct: Towards Generating Whole-body Motion from Speech
di: Zhang, Jinsong, et al.
Pubblicazione: (2023)
di: Zhang, Jinsong, et al.
Pubblicazione: (2023)
GenesisTex2: Stable, Consistent and High-Quality Text-to-Texture Generation
di: Lu, Jiawei, et al.
Pubblicazione: (2024)
di: Lu, Jiawei, et al.
Pubblicazione: (2024)
DyStream: Streaming Dyadic Talking Heads Generation via Flow Matching-based Autoregressive Model
di: Chen, Bohong, et al.
Pubblicazione: (2025)
di: Chen, Bohong, et al.
Pubblicazione: (2025)
FUSION: Full-Body Unified Motion Prior for Body and Hands via Diffusion
di: Duran, Enes, et al.
Pubblicazione: (2026)
di: Duran, Enes, et al.
Pubblicazione: (2026)
Interactive Humanoid: Online Full-Body Motion Reaction Synthesis with Social Affordance Canonicalization and Forecasting
di: Liu, Yunze, et al.
Pubblicazione: (2023)
di: Liu, Yunze, et al.
Pubblicazione: (2023)
RTG-SLAM: Real-time 3D Reconstruction at Scale using Gaussian Splatting
di: Peng, Zhexi, et al.
Pubblicazione: (2024)
di: Peng, Zhexi, et al.
Pubblicazione: (2024)
InteracTalker: Prompt-Based Human-Object Interaction with Co-Speech Gesture Generation
di: Rajan, Sreehari, et al.
Pubblicazione: (2025)
di: Rajan, Sreehari, et al.
Pubblicazione: (2025)
MotionChain: Conversational Motion Controllers via Multimodal Prompts
di: Jiang, Biao, et al.
Pubblicazione: (2024)
di: Jiang, Biao, et al.
Pubblicazione: (2024)
HMPDM: A Diffusion Model for Driving Video Prediction with Historical Motion Priors
di: Li, Ke, et al.
Pubblicazione: (2026)
di: Li, Ke, et al.
Pubblicazione: (2026)
FreeFuse: Multi-Subject LoRA Fusion via Adaptive Token-Level Routing at Test Time
di: Liu, Yaoli, et al.
Pubblicazione: (2025)
di: Liu, Yaoli, et al.
Pubblicazione: (2025)
DEGAS: Detailed Expressions on Full-Body Gaussian Avatars
di: Shao, Zhijing, et al.
Pubblicazione: (2024)
di: Shao, Zhijing, et al.
Pubblicazione: (2024)
Autoregressive Image Generation with Vision Full-view Prompt
di: Cai, Miaomiao, et al.
Pubblicazione: (2025)
di: Cai, Miaomiao, et al.
Pubblicazione: (2025)
Efficient 3D Full-Body Motion Generation from Sparse Tracking Inputs with Temporal Windows
di: Angelis, Georgios Fotios, et al.
Pubblicazione: (2025)
di: Angelis, Georgios Fotios, et al.
Pubblicazione: (2025)
Controllable Text-to-Motion Generation via Modular Body-Part Phase Control
di: Dai, Minyue, et al.
Pubblicazione: (2026)
di: Dai, Minyue, et al.
Pubblicazione: (2026)
Generating Attribute-Aware Human Motions from Textual Prompt
di: Wang, Xinghan, et al.
Pubblicazione: (2025)
di: Wang, Xinghan, et al.
Pubblicazione: (2025)
Understanding the Vulnerability of Skeleton-based Human Activity Recognition via Black-box Attack
di: Diao, Yunfeng, et al.
Pubblicazione: (2022)
di: Diao, Yunfeng, et al.
Pubblicazione: (2022)
SymphoMotion: Joint Control of Camera Motion and Object Dynamics for Coherent Video Generation
di: Zhang, Guiyu, et al.
Pubblicazione: (2026)
di: Zhang, Guiyu, et al.
Pubblicazione: (2026)
OmniDrag: Enabling Motion Control for Omnidirectional Image-to-Video Generation
di: Li, Weiqi, et al.
Pubblicazione: (2024)
di: Li, Weiqi, et al.
Pubblicazione: (2024)
Beyond Full Labels: Energy-Double-Guided Single-Point Prompt for Infrared Small Target Label Generation
di: Yuan, Shuai, et al.
Pubblicazione: (2024)
di: Yuan, Shuai, et al.
Pubblicazione: (2024)
CoMo: Compositional Motion Customization for Text-to-Video Generation
di: Xu, Youcan, et al.
Pubblicazione: (2025)
di: Xu, Youcan, et al.
Pubblicazione: (2025)
EgoControl: Controllable Egocentric Video Generation via 3D Full-Body Poses
di: Pallotta, Enrico, et al.
Pubblicazione: (2025)
di: Pallotta, Enrico, et al.
Pubblicazione: (2025)
Gaussian Splashing: Unified Particles for Versatile Motion Synthesis and Rendering
di: Feng, Yutao, et al.
Pubblicazione: (2024)
di: Feng, Yutao, et al.
Pubblicazione: (2024)
HOSIG: Full-Body Human-Object-Scene Interaction Generation with Hierarchical Scene Perception
di: Yao, Wei, et al.
Pubblicazione: (2025)
di: Yao, Wei, et al.
Pubblicazione: (2025)
ElastoGen: 4D Generative Elastodynamics
di: Feng, Yutao, et al.
Pubblicazione: (2024)
di: Feng, Yutao, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Motion-example-controlled Co-speech Gesture Generation Leveraging Large Language Models
di: Chen, Bohong, et al.
Pubblicazione: (2025) -
Autonomous Imagination: Closed-Loop Decomposition of Visual-to-Textual Conversion in Visual Reasoning for Multimodal Large Language Models
di: Liu, Jingming, et al.
Pubblicazione: (2024) -
When Gaussian Meets Surfel: Ultra-fast High-fidelity Radiance Field Rendering
di: Ye, Keyang, et al.
Pubblicazione: (2025) -
Gaussian-Plus-SDF SLAM: High-fidelity 3D Reconstruction at 150+ fps
di: Peng, Zhexi, et al.
Pubblicazione: (2025) -
Slot-ID: Identity-Preserving Video Generation from Reference Videos via Slot-Based Temporal Identity Encoding
di: Lai, Yixuan, et al.
Pubblicazione: (2026)