Saved in:
| Main Authors: | Zhang, Xiangyue, Li, Jianfang, Zhang, Jiaxu, Dang, Ziqiang, Ren, Jianqiang, Bo, Liefeng, Tu, Zhigang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2412.16563 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
EchoMask: Speech-Queried Attention-based Mask Modeling for Holistic Co-Speech Motion Generation
by: Zhang, Xiangyue, et al.
Published: (2025)
by: Zhang, Xiangyue, et al.
Published: (2025)
Mitigating Error Accumulation in Co-Speech Motion Generation via Global Rotation Diffusion and Multi-Level Constraints
by: Zhang, Xiangyue, et al.
Published: (2025)
by: Zhang, Xiangyue, et al.
Published: (2025)
Generative Motion Stylization of Cross-structure Characters within Canonical Motion Space
by: Zhang, Jiaxu, et al.
Published: (2024)
by: Zhang, Jiaxu, et al.
Published: (2024)
Make-A-Character 2: Animatable 3D Character Generation From a Single Image
by: Liu, Lin, et al.
Published: (2025)
by: Liu, Lin, et al.
Published: (2025)
Not All Frames Are Equal: Complexity-Aware Masked Motion Generation via Motion Spectral Descriptors
by: Zhou, Pengfei, et al.
Published: (2026)
by: Zhou, Pengfei, et al.
Published: (2026)
Cascaded Dual Vision Transformer for Accurate Facial Landmark Detection
by: Dang, Ziqiang, et al.
Published: (2024)
by: Dang, Ziqiang, et al.
Published: (2024)
Symbiotic-MoE: Unlocking the Synergy between Generation and Understanding
by: Liu, Xiangyue, et al.
Published: (2026)
by: Liu, Xiangyue, et al.
Published: (2026)
HolisticSemGes: Semantic Grounding of Holistic Co-Speech Gesture Generation with Contrastive Flow-Matching
by: Liu, Lanmiao, et al.
Published: (2026)
by: Liu, Lanmiao, et al.
Published: (2026)
MikuDance: Animating Character Art with Mixed Motion Dynamics
by: Zhang, Jiaxu, et al.
Published: (2024)
by: Zhang, Jiaxu, et al.
Published: (2024)
Expressive Keypoints for Skeleton-based Action Recognition via Skeleton Transformation
by: Yang, Yijie, et al.
Published: (2024)
by: Yang, Yijie, et al.
Published: (2024)
HoloGest: Decoupled Diffusion and Motion Priors for Generating Holisticly Expressive Co-speech Gestures
by: Cheng, Yongkang, et al.
Published: (2025)
by: Cheng, Yongkang, et al.
Published: (2025)
MaskSem: Semantic-Guided Masking for Learning 3D Hybrid High-Order Motion Representation
by: Wei, Wei, et al.
Published: (2025)
by: Wei, Wei, et al.
Published: (2025)
Textoon: Generating Vivid 2D Cartoon Characters from Text Descriptions
by: He, Chao, et al.
Published: (2025)
by: He, Chao, et al.
Published: (2025)
Unison: Harmonizing Motion, Speech, and Sound for Human-Centric Audio-Video Generation
by: Cheng, Shihao, et al.
Published: (2026)
by: Cheng, Shihao, et al.
Published: (2026)
Exploring Timeline Control for Facial Motion Generation
by: Ma, Yifeng, et al.
Published: (2025)
by: Ma, Yifeng, et al.
Published: (2025)
ChatAnyone: Stylized Real-time Portrait Video Generation with Hierarchical Motion Diffusion Model
by: Qi, Jinwei, et al.
Published: (2025)
by: Qi, Jinwei, et al.
Published: (2025)
Beyond Talking -- Generating Holistic 3D Human Dyadic Motion for Communication
by: Sun, Mingze, et al.
Published: (2024)
by: Sun, Mingze, et al.
Published: (2024)
Co-speech Gesture Video Generation via Motion-Based Graph Retrieval
by: Song, Yafei, et al.
Published: (2025)
by: Song, Yafei, et al.
Published: (2025)
OmniTalker: One-shot Real-time Text-Driven Talking Audio-Video Generation With Multimodal Style Mimicking
by: Wang, Zhongjian, et al.
Published: (2025)
by: Wang, Zhongjian, et al.
Published: (2025)
PersonaGesture: Single-Reference Co-Speech Gesture Personalization for Unseen Speakers
by: Zhang, Xiangyue, et al.
Published: (2026)
by: Zhang, Xiangyue, et al.
Published: (2026)
Towards Variable and Coordinated Holistic Co-Speech Motion Generation
by: Liu, Yifei, et al.
Published: (2024)
by: Liu, Yifei, et al.
Published: (2024)
CoGenAV: Versatile Audio-Visual Representation Learning via Contrastive-Generative Synchronization
by: Bai, Detao, et al.
Published: (2025)
by: Bai, Detao, et al.
Published: (2025)
DiffTED: One-shot Audio-driven TED Talk Video Generation with Diffusion-based Co-speech Gestures
by: Hogue, Steven, et al.
Published: (2024)
by: Hogue, Steven, et al.
Published: (2024)
EMO: Emote Portrait Alive -- Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak Conditions
by: Tian, Linrui, et al.
Published: (2024)
by: Tian, Linrui, et al.
Published: (2024)
AnyText2: Visual Text Generation and Editing With Customizable Attributes
by: Tuo, Yuxiang, et al.
Published: (2024)
by: Tuo, Yuxiang, et al.
Published: (2024)
MoSAM: Motion-Guided Segment Anything Model with Spatial-Temporal Memory Selection
by: Yang, Qiushi, et al.
Published: (2025)
by: Yang, Qiushi, et al.
Published: (2025)
SemanticFormer: Holistic and Semantic Traffic Scene Representation for Trajectory Prediction using Knowledge Graphs
by: Sun, Zhigang, et al.
Published: (2024)
by: Sun, Zhigang, et al.
Published: (2024)
AdaptiveDrag: Semantic-Driven Dragging on Diffusion-Based Image Editing
by: Chen, DuoSheng, et al.
Published: (2024)
by: Chen, DuoSheng, et al.
Published: (2024)
Baton: Explicit Semantic Blueprints for Joint Video-Audio Generation
by: Tu, Shuyuan, et al.
Published: (2026)
by: Tu, Shuyuan, et al.
Published: (2026)
DiffuEraser: A Diffusion Model for Video Inpainting
by: Li, Xiaowen, et al.
Published: (2025)
by: Li, Xiaowen, et al.
Published: (2025)
SemGeoMo: Dynamic Contextual Human Motion Generation with Semantic and Geometric Guidance
by: Cong, Peishan, et al.
Published: (2025)
by: Cong, Peishan, et al.
Published: (2025)
FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis
by: Wang, Mengchao, et al.
Published: (2025)
by: Wang, Mengchao, et al.
Published: (2025)
Perception-as-Control: Fine-grained Controllable Image Animation with 3D-aware Motion Representation
by: Chen, Yingjie, et al.
Published: (2025)
by: Chen, Yingjie, et al.
Published: (2025)
EMO2: End-Effector Guided Audio-Driven Avatar Video Generation
by: Tian, Linrui, et al.
Published: (2025)
by: Tian, Linrui, et al.
Published: (2025)
Motion-Aware Generative Frame Interpolation
by: Zhang, Guozhen, et al.
Published: (2025)
by: Zhang, Guozhen, et al.
Published: (2025)
DreamDance: Animating Character Art via Inpainting Stable Gaussian Worlds
by: Zhang, Jiaxu, et al.
Published: (2025)
by: Zhang, Jiaxu, et al.
Published: (2025)
SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning
by: Liu, Lanmiao, et al.
Published: (2025)
by: Liu, Lanmiao, et al.
Published: (2025)
SemDP: Semantic-level Differential Privacy Protection for Face Datasets
by: Zhang, Xiaoting, et al.
Published: (2024)
by: Zhang, Xiaoting, et al.
Published: (2024)
SemGrasp: Semantic Grasp Generation via Language Aligned Discretization
by: Li, Kailin, et al.
Published: (2024)
by: Li, Kailin, et al.
Published: (2024)
MoGenTS: Motion Generation based on Spatial-Temporal Joint Modeling
by: Yuan, Weihao, et al.
Published: (2024)
by: Yuan, Weihao, et al.
Published: (2024)
Similar Items
-
EchoMask: Speech-Queried Attention-based Mask Modeling for Holistic Co-Speech Motion Generation
by: Zhang, Xiangyue, et al.
Published: (2025) -
Mitigating Error Accumulation in Co-Speech Motion Generation via Global Rotation Diffusion and Multi-Level Constraints
by: Zhang, Xiangyue, et al.
Published: (2025) -
Generative Motion Stylization of Cross-structure Characters within Canonical Motion Space
by: Zhang, Jiaxu, et al.
Published: (2024) -
Make-A-Character 2: Animatable 3D Character Generation From a Single Image
by: Liu, Lin, et al.
Published: (2025) -
Not All Frames Are Equal: Complexity-Aware Masked Motion Generation via Motion Spectral Descriptors
by: Zhou, Pengfei, et al.
Published: (2026)