MaskControl: Spatio-Temporal Control for Masked Motion Synthesis
Fuente:
arXiv
Saved in:
| Main Authors: | Pinyoanuntapong, Ekkasit, Saleem, Muhammad Usama, Karunratanakul, Korrawe, Wang, Pu, Xue, Hongfei, Chen, Chen, Guo, Chuan, Cao, Junli, Ren, Jian, Tulyakov, Sergey |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Walk Before You Dance: High-fidelity and Editable Dance Synthesis via Generative Masked Motion Prior
by: Shah, Foram N, et al.
Published: (2025)
by: Shah, Foram N, et al.
Published: (2025)
MaskHand: Generative Masked Modeling for Robust Hand Mesh Reconstruction in the Wild
by: Saleem, Muhammad Usama, et al.
Published: (2024)
by: Saleem, Muhammad Usama, et al.
Published: (2024)
MMM: Generative Masked Motion Model
by: Pinyoanuntapong, Ekkasit, et al.
Published: (2023)
by: Pinyoanuntapong, Ekkasit, et al.
Published: (2023)
GenHMR: Generative Human Mesh Recovery
by: Saleem, Muhammad Usama, et al.
Published: (2024)
by: Saleem, Muhammad Usama, et al.
Published: (2024)
BAMM: Bidirectional Autoregressive Motion Model
by: Pinyoanuntapong, Ekkasit, et al.
Published: (2024)
by: Pinyoanuntapong, Ekkasit, et al.
Published: (2024)
LiveGesture Streamable Co-Speech Gesture Generation Model
by: Saleem, Muhammad Usama, et al.
Published: (2026)
by: Saleem, Muhammad Usama, et al.
Published: (2026)
NaP-Control: Navigating Diffusion Prior for Versatile and Fast Character Control
by: Chen, Chia-Wen, et al.
Published: (2026)
by: Chen, Chia-Wen, et al.
Published: (2026)
UniPhys: Unified Planner and Controller with Diffusion for Flexible Physics-Based Character Control
by: Wu, Yan, et al.
Published: (2025)
by: Wu, Yan, et al.
Published: (2025)
Optimizing Diffusion Noise Can Serve As Universal Motion Priors
by: Karunratanakul, Korrawe, et al.
Published: (2023)
by: Karunratanakul, Korrawe, et al.
Published: (2023)
Dynamic Motion Synthesis: Masked Audio-Text Conditioned Spatio-Temporal Transformers
by: Anisetty, Sohan, et al.
Published: (2024)
by: Anisetty, Sohan, et al.
Published: (2024)
MaskedMimic: Unified Physics-Based Character Control Through Masked Motion Inpainting
by: Tessler, Chen, et al.
Published: (2024)
by: Tessler, Chen, et al.
Published: (2024)
BioPose: Biomechanically-accurate 3D Pose Estimation from Monocular Videos
by: Koleini, Farnoosh, et al.
Published: (2025)
by: Koleini, Farnoosh, et al.
Published: (2025)
MotionDreamer: One-to-Many Motion Synthesis with Localized Generative Masked Transformer
by: Wang, Yilin, et al.
Published: (2025)
by: Wang, Yilin, et al.
Published: (2025)
Monocular Models are Strong Learners for Multi-View Human Mesh Recovery
by: Xie, Haoyu, et al.
Published: (2026)
by: Xie, Haoyu, et al.
Published: (2026)
Towards Physical Understanding in Video Generation: A 3D Point Regularization Approach
by: Chen, Yunuo, et al.
Published: (2025)
by: Chen, Yunuo, et al.
Published: (2025)
AsCAN: Asymmetric Convolution-Attention Networks for Efficient Recognition and Generation
by: Kag, Anil, et al.
Published: (2024)
by: Kag, Anil, et al.
Published: (2024)
Beyond Boxes: Mask-Guided Spatio-Temporal Feature Aggregation for Video Object Detection
by: Hashmi, Khurram Azeem, et al.
Published: (2024)
by: Hashmi, Khurram Azeem, et al.
Published: (2024)
AlcheMinT: Fine-grained Temporal Control for Multi-Reference Consistent Video Generation
by: Girish, Sharath, et al.
Published: (2025)
by: Girish, Sharath, et al.
Published: (2025)
Spatio-Temporal Encoding of Brain Dynamics with Surface Masked Autoencoders
by: Dahan, Simon, et al.
Published: (2023)
by: Dahan, Simon, et al.
Published: (2023)
Controllable Text-to-Speech Synthesis with Masked-Autoencoded Style-Rich Representation
by: Wang, Yongqi, et al.
Published: (2025)
by: Wang, Yongqi, et al.
Published: (2025)
InterMask: 3D Human Interaction Generation via Collaborative Masked Modeling
by: Javed, Muhammad Gohar, et al.
Published: (2024)
by: Javed, Muhammad Gohar, et al.
Published: (2024)
Guiding Masked Representation Learning to Capture Spatio-Temporal Relationship of Electrocardiogram
by: Na, Yeongyeon, et al.
Published: (2024)
by: Na, Yeongyeon, et al.
Published: (2024)
Cluster-Wise Spatio-Temporal Masking for Efficient Video-Language Pretraining
by: Zhuang, Weijun, et al.
Published: (2026)
by: Zhuang, Weijun, et al.
Published: (2026)
Diffusion Priors for Dynamic View Synthesis from Monocular Videos
by: Wang, Chaoyang, et al.
Published: (2024)
by: Wang, Chaoyang, et al.
Published: (2024)
Accelerating Masked Image Generation by Learning Latent Controlled Dynamics
by: Zhu, Kaiwen, et al.
Published: (2026)
by: Zhu, Kaiwen, et al.
Published: (2026)
Send Less, Perceive More: Masked Quantized Point Cloud Communication for Loss-Tolerant Collaborative Perception
by: Xu, Sheng, et al.
Published: (2026)
by: Xu, Sheng, et al.
Published: (2026)
EndoControlMag: Robust Endoscopic Vascular Motion Magnification with Periodic Reference Resetting and Hierarchical Tissue-aware Dual-Mask Control
by: Wang, An, et al.
Published: (2025)
by: Wang, An, et al.
Published: (2025)
Text-driven Human Motion Generation with Motion Masked Diffusion Model
by: Chen, Xingyu
Published: (2024)
by: Chen, Xingyu
Published: (2024)
Towards Robust and Controllable Text-to-Motion via Masked Autoregressive Diffusion
by: Zhang, Zongye, et al.
Published: (2025)
by: Zhang, Zongye, et al.
Published: (2025)
Promptable Game Models: Text-Guided Game Simulation via Masked Diffusion Models
by: Menapace, Willi, et al.
Published: (2023)
by: Menapace, Willi, et al.
Published: (2023)
Mask What Matters: Controllable Text-Guided Masking for Self-Supervised Medical Image Analysis
by: Wang, Ruilang, et al.
Published: (2025)
by: Wang, Ruilang, et al.
Published: (2025)
Improving the O-GEHL Branch Prediction Accuracy Using Analytical Results
by: Ekkasit Tiamkaew
Published: (2007)
by: Ekkasit Tiamkaew
Published: (2007)
Plug-and-Play Controllable Generation for Discrete Masked Models
by: Guo, Wei, et al.
Published: (2024)
by: Guo, Wei, et al.
Published: (2024)
Resource-Efficient Motion Control for Video Generation via Dynamic Mask Guidance
by: Feng, Sicong, et al.
Published: (2025)
by: Feng, Sicong, et al.
Published: (2025)
ReMoMask: Retrieval-Augmented Masked Motion Generation
by: Li, Zhengdao, et al.
Published: (2025)
by: Li, Zhengdao, et al.
Published: (2025)
SpecTM: Spectral Targeted Masking for Trustworthy Foundation Models
by: Imtiaz, Syed Usama, et al.
Published: (2026)
by: Imtiaz, Syed Usama, et al.
Published: (2026)
Lightweight Predictive 3D Gaussian Splats
by: Cao, Junli, et al.
Published: (2024)
by: Cao, Junli, et al.
Published: (2024)
U-MASK: User-adaptive Spatio-Temporal Masking for Personalized Mobile AI Applications
by: Zhang, Shiyuan, et al.
Published: (2026)
by: Zhang, Shiyuan, et al.
Published: (2026)
GeoMAE: Masking Representation Learning for Spatio-Temporal Graph Forecasting with Missing Values
by: Ke, Songyu, et al.
Published: (2025)
by: Ke, Songyu, et al.
Published: (2025)
EchoMask: Speech-Queried Attention-based Mask Modeling for Holistic Co-Speech Motion Generation
by: Zhang, Xiangyue, et al.
Published: (2025)
by: Zhang, Xiangyue, et al.
Published: (2025)
Similar Items
-
Walk Before You Dance: High-fidelity and Editable Dance Synthesis via Generative Masked Motion Prior
by: Shah, Foram N, et al.
Published: (2025) -
MaskHand: Generative Masked Modeling for Robust Hand Mesh Reconstruction in the Wild
by: Saleem, Muhammad Usama, et al.
Published: (2024) -
MMM: Generative Masked Motion Model
by: Pinyoanuntapong, Ekkasit, et al.
Published: (2023) -
GenHMR: Generative Human Mesh Recovery
by: Saleem, Muhammad Usama, et al.
Published: (2024) -
BAMM: Bidirectional Autoregressive Motion Model
by: Pinyoanuntapong, Ekkasit, et al.
Published: (2024)