Motus: A Unified Latent Action World Model
Fuente:
arXiv
Saved in:
| Main Authors: | Bi, Hongzhe, Tan, Hengkai, Xie, Shenghao, Wang, Zeyuan, Huang, Shuhe, Liu, Haitian, Zhao, Ruowen, Feng, Yao, Xiang, Chendong, Rong, Yinze, Zhao, Hongyan, Liu, Hanyu, Su, Zhizhong, Ma, Lei, Su, Hang, Zhu, Jun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MotuBrain: An Advanced World Action Model for Robot Control
by: MotuBrain Team, et al.
Published: (2026)
by: MotuBrain Team, et al.
Published: (2026)
H-RDT: Human Manipulation Enhanced Bimanual Robotic Manipulation
by: Bi, Hongzhe, et al.
Published: (2025)
by: Bi, Hongzhe, et al.
Published: (2025)
Vidarc: Embodied Video Diffusion Model for Closed-loop Control
by: Feng, Yao, et al.
Published: (2025)
by: Feng, Yao, et al.
Published: (2025)
Vidar: Embodied Video Diffusion Model for Generalist Manipulation
by: Feng, Yao, et al.
Published: (2025)
by: Feng, Yao, et al.
Published: (2025)
AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation
by: Tan, Hengkai, et al.
Published: (2025)
by: Tan, Hengkai, et al.
Published: (2025)
ManiBox: Enhancing Embodied Spatial Generalization via Scalable Simulation Data Generations
by: Tan, Hengkai, et al.
Published: (2024)
by: Tan, Hengkai, et al.
Published: (2024)
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding
by: Ye, Junliang, et al.
Published: (2025)
by: Ye, Junliang, et al.
Published: (2025)
Fourier Controller Networks for Real-Time Decision-Making in Embodied Learning
by: Tan, Hengkai, et al.
Published: (2024)
by: Tan, Hengkai, et al.
Published: (2024)
Geometry-Aware Rotary Position Embedding for Consistent Video World Model
by: Xiang, Chendong, et al.
Published: (2026)
by: Xiang, Chendong, et al.
Published: (2026)
RDT2: Exploring the Scaling Limit of UMI Data Towards Zero-Shot Cross-Embodiment Generalization
by: Liu, Songming, et al.
Published: (2026)
by: Liu, Songming, et al.
Published: (2026)
Dynamic Nanostructure‐Based DNA Logic Gates for Cancer Diagnosis and Therapy
by: Shiyi Bi, et al.
Published: (2024)
by: Shiyi Bi, et al.
Published: (2024)
Post-Hoc Split-Point Self-Consistency Verification for Efficient, Unified Quantification of Aleatoric and Epistemic Uncertainty in Deep Learning
by: Zhao, Zhizhong, et al.
Published: (2025)
by: Zhao, Zhizhong, et al.
Published: (2025)
DriveWorld-VLA: Unified Latent-Space World Modeling with Vision-Language-Action for Autonomous Driving
by: jia, Feiyang, et al.
Published: (2026)
by: jia, Feiyang, et al.
Published: (2026)
Theoretical Analysis of Relative Errors in Gradient Computations for Adversarial Attacks with CE Loss
by: Yu, Yunrui, et al.
Published: (2025)
by: Yu, Yunrui, et al.
Published: (2025)
Fast-WAM: Do World Action Models Need Test-time Future Imagination?
by: Yuan, Tianyuan, et al.
Published: (2026)
by: Yuan, Tianyuan, et al.
Published: (2026)
Co-Evolving Latent Action World Models
by: Wang, Yucen, et al.
Published: (2025)
by: Wang, Yucen, et al.
Published: (2025)
RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation
by: Liu, Songming, et al.
Published: (2024)
by: Liu, Songming, et al.
Published: (2024)
Scaling Sim-to-Real Reinforcement Learning for Robot VLAs with Generative 3D Worlds
by: Choi, Andrew, et al.
Published: (2026)
by: Choi, Andrew, et al.
Published: (2026)
Causal Network Discovery from Interventional Count Data with Latent Linear DAGs
by: Zhang, Yijiao, et al.
Published: (2026)
by: Zhang, Yijiao, et al.
Published: (2026)
Self-Improving World Modelling with Latent Actions
by: Qiu, Yifu, et al.
Published: (2026)
by: Qiu, Yifu, et al.
Published: (2026)
ReCode: Unify Plan and Action for Universal Granularity Control
by: Yu, Zhaoyang, et al.
Published: (2025)
by: Yu, Zhaoyang, et al.
Published: (2025)
Towards Unifying Understanding and Generation in the Era of Vision Foundation Models: A Survey from the Autoregression Perspective
by: Xie, Shenghao, et al.
Published: (2024)
by: Xie, Shenghao, et al.
Published: (2024)
Latent-WAM: Latent World Action Modeling for End-to-End Autonomous Driving
by: Wang, Linbo, et al.
Published: (2026)
by: Wang, Linbo, et al.
Published: (2026)
Multiplicity and concentration of dual solutions for a Helmholtz system
by: Qiu, Ruowen, et al.
Published: (2026)
by: Qiu, Ruowen, et al.
Published: (2026)
Existence and multiplicity of $L^2$-Normalized solutions for the periodic Schrödinger system of Hamiltonian type
by: Qiu, Ruowen, et al.
Published: (2025)
by: Qiu, Ruowen, et al.
Published: (2025)
Motus: software para el análisis conductual de patrones de desplazamiento
by: Alejandro León
Published: (2020)
by: Alejandro León
Published: (2020)
Motus: A Framework for Human Motion Classification in a Notcontrolled Moving Environment
by: Joselyn Rodríguez-González
Published: (2020)
by: Joselyn Rodríguez-González
Published: (2020)
Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising
by: Guo, Jun, et al.
Published: (2026)
by: Guo, Jun, et al.
Published: (2026)
NANO3D: A Training-Free Approach for Efficient 3D Editing Without Masks
by: Ye, Junliang, et al.
Published: (2025)
by: Ye, Junliang, et al.
Published: (2025)
World Guidance: World Modeling in Condition Space for Action Generation
by: Su, Yue, et al.
Published: (2026)
by: Su, Yue, et al.
Published: (2026)
Multiphase Optimization of Fermentation Yield via Latent Variable Reinforcement Learning
by: Peng Su, et al.
Published: (2026)
by: Peng Su, et al.
Published: (2026)
EmbodiedGen: Towards a Generative 3D World Engine for Embodied Intelligence
by: Wang, Xinjie, et al.
Published: (2025)
by: Wang, Xinjie, et al.
Published: (2025)
Towards Unified and Lossless Latent Space for 3D Molecular Latent Diffusion Modeling
by: Luo, Yanchen, et al.
Published: (2025)
by: Luo, Yanchen, et al.
Published: (2025)
IRIS-SLAM: Unified Geo-Instance Representations for Robust Semantic Localization and Mapping
by: Xiao, Tingyang, et al.
Published: (2026)
by: Xiao, Tingyang, et al.
Published: (2026)
Crowdsourced bug report severity prediction based on text and image understanding via heterogeneous graph convolutional networks
by: Yifan Wu, et al.
Published: (2024)
by: Yifan Wu, et al.
Published: (2024)
Some $3$-designs invariant under $2.PΣL(2,49).$
by: Shi, Minjia, et al.
Published: (2024)
by: Shi, Minjia, et al.
Published: (2024)
AEMIM: Adversarial Examples Meet Masked Image Modeling
by: Xiang, Wenzhao, et al.
Published: (2024)
by: Xiang, Wenzhao, et al.
Published: (2024)
Molecular Techniques and Ecological Data for Taxonomically Difficult Groups: A Case Study of a Morphologically Variable New Species in the Genus (Coleoptera: Buprestidae).
by: Huang, Botao, et al.
Published: (2026)
by: Huang, Botao, et al.
Published: (2026)
The Fischer‐Lactonization‐Driven Mechanism for Ultra‐Efficient Recycling of Spent Lithium‐Ion Batteries
by: Miaomiao Zhou, et al.
Published: (2024)
by: Miaomiao Zhou, et al.
Published: (2024)
The Fischer‐Lactonization‐Driven Mechanism for Ultra‐Efficient Recycling of Spent Lithium‐Ion Batteries
by: Miaomiao Zhou, et al.
Published: (2024)
by: Miaomiao Zhou, et al.
Published: (2024)
Similar Items
-
MotuBrain: An Advanced World Action Model for Robot Control
by: MotuBrain Team, et al.
Published: (2026) -
H-RDT: Human Manipulation Enhanced Bimanual Robotic Manipulation
by: Bi, Hongzhe, et al.
Published: (2025) -
Vidarc: Embodied Video Diffusion Model for Closed-loop Control
by: Feng, Yao, et al.
Published: (2025) -
Vidar: Embodied Video Diffusion Model for Generalist Manipulation
by: Feng, Yao, et al.
Published: (2025) -
AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation
by: Tan, Hengkai, et al.
Published: (2025)