MixerMDM: Learnable Composition of Human Motion Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Ruiz-Ponce, Pablo, Barquero, German, Palmero, Cristina, Escalera, Sergio, García-Rodríguez, José |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Seamless Human Motion Composition with Blended Positional Encodings
by: Barquero, German, et al.
Published: (2024)
by: Barquero, German, et al.
Published: (2024)
in2IN: Leveraging individual Information to Generate Human INteractions
by: Ponce, Pablo Ruiz, et al.
Published: (2024)
by: Ponce, Pablo Ruiz, et al.
Published: (2024)
Interact2Ar: Full-Body Human-Human Interaction Generation via Autoregressive Diffusion Models
by: Ruiz-Ponce, Pablo, et al.
Published: (2025)
by: Ruiz-Ponce, Pablo, et al.
Published: (2025)
From Sparse Signal to Smooth Motion: Real-Time Motion Generation with Rolling Prediction Models
by: Barquero, German, et al.
Published: (2025)
by: Barquero, German, et al.
Published: (2025)
IF-MDM: Implicit Face Motion Diffusion Model for High-Fidelity Realtime Talking Head Generation
by: Yang, Sejong, et al.
Published: (2024)
by: Yang, Sejong, et al.
Published: (2024)
FG-MDM: Towards Zero-Shot Human Motion Generation via ChatGPT-Refined Descriptions
by: Shi, Xu, et al.
Published: (2023)
by: Shi, Xu, et al.
Published: (2023)
What Matters in Virtual Try-Off? Dual-UNet Diffusion Model For Garment Reconstruction
by: Truong, Loc-Phat, et al.
Published: (2026)
by: Truong, Loc-Phat, et al.
Published: (2026)
REACT 2024: the Second Multiple Appropriate Facial Reaction Generation Challenge
by: Song, Siyang, et al.
Published: (2024)
by: Song, Siyang, et al.
Published: (2024)
A Transformer Model for Boundary Detection in Continuous Sign Language
by: Rastgoo, Razieh, et al.
Published: (2024)
by: Rastgoo, Razieh, et al.
Published: (2024)
Enhancing Personality Recognition by Comparing the Predictive Power of Traits, Facets, and Nuances
by: Ansari, Amir, et al.
Published: (2026)
by: Ansari, Amir, et al.
Published: (2026)
DP-MDM: Detail-Preserving MR Reconstruction via Multiple Diffusion Models
by: Geng, Mengxiao, et al.
Published: (2024)
by: Geng, Mengxiao, et al.
Published: (2024)
REACT 2025: the Third Multiple Appropriate Facial Reaction Generation Challenge
by: Song, Siyang, et al.
Published: (2025)
by: Song, Siyang, et al.
Published: (2025)
EnergyMoGen: Compositional Human Motion Generation with Energy-Based Diffusion Model in Latent Space
by: Zhang, Jianrong, et al.
Published: (2024)
by: Zhang, Jianrong, et al.
Published: (2024)
Text-driven Human Motion Generation with Motion Masked Diffusion Model
by: Chen, Xingyu
Published: (2024)
by: Chen, Xingyu
Published: (2024)
ChebMixer: Efficient Graph Representation Learning with MLP Mixer
by: Kui, Xiaoyan, et al.
Published: (2024)
by: Kui, Xiaoyan, et al.
Published: (2024)
RDM: Recurrent Diffusion Model for Human Motion Generation
by: Mohamed, Mirgahney, et al.
Published: (2024)
by: Mohamed, Mirgahney, et al.
Published: (2024)
Realistic Human Motion Generation with Cross-Diffusion Models
by: Ren, Zeping, et al.
Published: (2023)
by: Ren, Zeping, et al.
Published: (2023)
FrankenMotion: Part-level Human Motion Generation and Composition
by: Li, Chuqiao, et al.
Published: (2026)
by: Li, Chuqiao, et al.
Published: (2026)
SOVABench: A Vehicle Surveillance Action Retrieval Benchmark for Multimodal Large Language Models
by: Rabasseda, Oriol, et al.
Published: (2026)
by: Rabasseda, Oriol, et al.
Published: (2026)
L-SWAG: Layer-Sample Wise Activation with Gradients information for Zero-Shot NAS on Vision Transformers
by: Casarin, Sofia, et al.
Published: (2025)
by: Casarin, Sofia, et al.
Published: (2025)
Learnability-Guided Diffusion for Dataset Distillation
by: Chan-Santiago, Jeffrey A., et al.
Published: (2026)
by: Chan-Santiago, Jeffrey A., et al.
Published: (2026)
The Learnability Gap in Medical Latent Diffusion
by: Dombrowski, Mischa, et al.
Published: (2026)
by: Dombrowski, Mischa, et al.
Published: (2026)
Point Cloud Resampling with Learnable Heat Diffusion
by: Xu, Wenqiang, et al.
Published: (2024)
by: Xu, Wenqiang, et al.
Published: (2024)
Shape Conditioned Human Motion Generation with Diffusion Model
by: Xue, Kebing, et al.
Published: (2024)
by: Xue, Kebing, et al.
Published: (2024)
Sparse-Dense Side-Tuner for efficient Video Temporal Grounding
by: Pujol-Perich, David, et al.
Published: (2025)
by: Pujol-Perich, David, et al.
Published: (2025)
SADA: Semantic adversarial unsupervised domain adaptation for Temporal Action Localization
by: Pujol-Perich, David, et al.
Published: (2023)
by: Pujol-Perich, David, et al.
Published: (2023)
STS-Mixer: Spatio-Temporal-Spectral Mixer for 4D Point Cloud Video Understanding
by: Li, Wenhao, et al.
Published: (2026)
by: Li, Wenhao, et al.
Published: (2026)
MixerFlow: MLP-Mixer meets Normalising Flows
by: English, Eshant, et al.
Published: (2023)
by: English, Eshant, et al.
Published: (2023)
Learnable Motion-Focused Tokenization for Effective and Efficient Video Unsupervised Domain Adaptation
by: Liu, Tzu Ling, et al.
Published: (2026)
by: Liu, Tzu Ling, et al.
Published: (2026)
TokenMotion: Motion-Guided Vision Transformer for Video Camouflaged Object Detection Via Learnable Token Selection
by: Yu, Zifan, et al.
Published: (2023)
by: Yu, Zifan, et al.
Published: (2023)
A Generative Multi-Resolution Pyramid and Normal-Conditioning 3D Cloth Draping
by: Laczkó, Hunor, et al.
Published: (2023)
by: Laczkó, Hunor, et al.
Published: (2023)
Neural Flow Diffusion Models: Learnable Forward Process for Improved Diffusion Modelling
by: Bartosh, Grigory, et al.
Published: (2024)
by: Bartosh, Grigory, et al.
Published: (2024)
GazeMoDiff: Gaze-guided Diffusion Model for Stochastic Human Motion Prediction
by: Yan, Haodong, et al.
Published: (2023)
by: Yan, Haodong, et al.
Published: (2023)
Causal Motion Diffusion Models for Autoregressive Motion Generation
by: Yu, Qing, et al.
Published: (2026)
by: Yu, Qing, et al.
Published: (2026)
Controlling Avatar Diffusion with Learnable Gaussian Embedding
by: Gao, Xuan, et al.
Published: (2025)
by: Gao, Xuan, et al.
Published: (2025)
LGTM: Local-to-Global Text-Driven Human Motion Diffusion Model
by: Sun, Haowen, et al.
Published: (2024)
by: Sun, Haowen, et al.
Published: (2024)
MotionPhysics: Learnable Motion Distillation for Text-Guided Simulation
by: Wang, Miaowei, et al.
Published: (2026)
by: Wang, Miaowei, et al.
Published: (2026)
Efficient Multi-scale Network with Learnable Discrete Wavelet Transform for Blind Motion Deblurring
by: Gao, Xin, et al.
Published: (2023)
by: Gao, Xin, et al.
Published: (2023)
CoMA: Compositional Human Motion Generation with Multi-modal Agents
by: Sun, Shanlin, et al.
Published: (2024)
by: Sun, Shanlin, et al.
Published: (2024)
RoHM: Robust Human Motion Reconstruction via Diffusion
by: Zhang, Siwei, et al.
Published: (2024)
by: Zhang, Siwei, et al.
Published: (2024)
Similar Items
-
Seamless Human Motion Composition with Blended Positional Encodings
by: Barquero, German, et al.
Published: (2024) -
in2IN: Leveraging individual Information to Generate Human INteractions
by: Ponce, Pablo Ruiz, et al.
Published: (2024) -
Interact2Ar: Full-Body Human-Human Interaction Generation via Autoregressive Diffusion Models
by: Ruiz-Ponce, Pablo, et al.
Published: (2025) -
From Sparse Signal to Smooth Motion: Real-Time Motion Generation with Rolling Prediction Models
by: Barquero, German, et al.
Published: (2025) -
IF-MDM: Implicit Face Motion Diffusion Model for High-Fidelity Realtime Talking Head Generation
by: Yang, Sejong, et al.
Published: (2024)