DexWorldModel: Causal Latent World Modeling towards Automated Learning of Embodied Tasks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Deng, Yueci, Liu, Guiliang, Jia, Kui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
World-Ego Modeling for Long-Horizon Evolution in Hybrid Embodied Tasks
von: Lin, Zuyao, et al.
Veröffentlicht: (2026)
von: Lin, Zuyao, et al.
Veröffentlicht: (2026)
You Only Teach Once: Learn One-Shot Bimanual Robotic Manipulation from Video Demonstrations
von: Zhou, Huayi, et al.
Veröffentlicht: (2025)
von: Zhou, Huayi, et al.
Veröffentlicht: (2025)
Rethinking Video Generation Model for the Embodied World
von: Deng, Yufan, et al.
Veröffentlicht: (2026)
von: Deng, Yufan, et al.
Veröffentlicht: (2026)
Learning Latent Action World Models In The Wild
von: Garrido, Quentin, et al.
Veröffentlicht: (2026)
von: Garrido, Quentin, et al.
Veröffentlicht: (2026)
ADAM: An Embodied Causal Agent in Open-World Environments
von: Yu, Shu, et al.
Veröffentlicht: (2024)
von: Yu, Shu, et al.
Veröffentlicht: (2024)
Latent Video Prediction Learns Better World Models
von: Alrasheed, Ali J, et al.
Veröffentlicht: (2026)
von: Alrasheed, Ali J, et al.
Veröffentlicht: (2026)
Embodied Tree of Thoughts: Deliberate Manipulation Planning with Embodied World Model
von: Xu, Wenjiang, et al.
Veröffentlicht: (2025)
von: Xu, Wenjiang, et al.
Veröffentlicht: (2025)
Chain of World: World Model Thinking in Latent Motion
von: Yang, Fuxiang, et al.
Veröffentlicht: (2026)
von: Yang, Fuxiang, et al.
Veröffentlicht: (2026)
EgoSim: Egocentric World Simulator for Embodied Interaction Generation
von: Hao, Jinkun, et al.
Veröffentlicht: (2026)
von: Hao, Jinkun, et al.
Veröffentlicht: (2026)
Other Vehicle Trajectories Are Also Needed: A Driving World Model Unifies Ego-Other Vehicle Trajectories in Video Latent Space
von: Zhu, Jian, et al.
Veröffentlicht: (2025)
von: Zhu, Jian, et al.
Veröffentlicht: (2025)
PAct: Part-Decomposed Single-View Articulated Object Generation
von: Liu, Qingming, et al.
Veröffentlicht: (2026)
von: Liu, Qingming, et al.
Veröffentlicht: (2026)
The Safety Challenge of World Models for Embodied AI Agents: A Review
von: Baraldi, Lorenzo, et al.
Veröffentlicht: (2025)
von: Baraldi, Lorenzo, et al.
Veröffentlicht: (2025)
UnrealZoo: Enriching Photo-realistic Virtual Worlds for Embodied AI
von: Zhong, Fangwei, et al.
Veröffentlicht: (2024)
von: Zhong, Fangwei, et al.
Veröffentlicht: (2024)
Lifting Embodied World Models for Planning and Control
von: Wang, Alex N., et al.
Veröffentlicht: (2026)
von: Wang, Alex N., et al.
Veröffentlicht: (2026)
AdaWorld: Learning Adaptable World Models with Latent Actions
von: Gao, Shenyuan, et al.
Veröffentlicht: (2025)
von: Gao, Shenyuan, et al.
Veröffentlicht: (2025)
EVA: Aligning Video World Models with Executable Robot Actions via Inverse Dynamics Rewards
von: Wang, Ruixiang, et al.
Veröffentlicht: (2026)
von: Wang, Ruixiang, et al.
Veröffentlicht: (2026)
ENACT: Evaluating Embodied Cognition with World Modeling of Egocentric Interaction
von: Wang, Qineng, et al.
Veröffentlicht: (2025)
von: Wang, Qineng, et al.
Veröffentlicht: (2025)
Olaf-World: Orienting Latent Actions for Video World Modeling
von: Jiang, Yuxin, et al.
Veröffentlicht: (2026)
von: Jiang, Yuxin, et al.
Veröffentlicht: (2026)
WoMAP: World Models For Embodied Open-Vocabulary Object Localization
von: Yin, Tenny, et al.
Veröffentlicht: (2025)
von: Yin, Tenny, et al.
Veröffentlicht: (2025)
Wow, wo, val! A Comprehensive Embodied World Model Evaluation Turing Test
von: Fan, Chun-Kai, et al.
Veröffentlicht: (2026)
von: Fan, Chun-Kai, et al.
Veröffentlicht: (2026)
DiLA: Disentangled Latent Action World Models
von: Zhang, Tianqiu, et al.
Veröffentlicht: (2026)
von: Zhang, Tianqiu, et al.
Veröffentlicht: (2026)
RynnEC: Bringing MLLMs into Embodied World
von: Dang, Ronghao, et al.
Veröffentlicht: (2025)
von: Dang, Ronghao, et al.
Veröffentlicht: (2025)
COMBO: Compositional World Models for Embodied Multi-Agent Cooperation
von: Zhang, Hongxin, et al.
Veröffentlicht: (2024)
von: Zhang, Hongxin, et al.
Veröffentlicht: (2024)
Beyond Pixels: Introducing Geometric-Semantic World Priors for Video-based Embodied Models via Spatio-temporal Alignment
von: Tang, Jinzhou, et al.
Veröffentlicht: (2025)
von: Tang, Jinzhou, et al.
Veröffentlicht: (2025)
Lattice Boltzmann Model for Learning Real-World Pixel Dynamicity
von: Zheng, Guangze, et al.
Veröffentlicht: (2025)
von: Zheng, Guangze, et al.
Veröffentlicht: (2025)
Render, Don't Decode: Weight-Space World Models with Latent Structural Disentanglement
von: Nzoyem, Roussel Desmond, et al.
Veröffentlicht: (2026)
von: Nzoyem, Roussel Desmond, et al.
Veröffentlicht: (2026)
EmbodieDreamer: Advancing Real2Sim2Real Transfer for Policy Training via Embodied World Modeling
von: Wang, Boyuan, et al.
Veröffentlicht: (2025)
von: Wang, Boyuan, et al.
Veröffentlicht: (2025)
WorldModelBench: Judging Video Generation Models As World Models
von: Li, Dacheng, et al.
Veröffentlicht: (2025)
von: Li, Dacheng, et al.
Veröffentlicht: (2025)
Rethinking Driving World Model as Synthetic Data Generator for Perception Tasks
von: Zeng, Kai, et al.
Veröffentlicht: (2025)
von: Zeng, Kai, et al.
Veröffentlicht: (2025)
Learning Additively Compositional Latent Actions for Embodied AI
von: Wei, Hangxing, et al.
Veröffentlicht: (2026)
von: Wei, Hangxing, et al.
Veröffentlicht: (2026)
Sekai: A Video Dataset towards World Exploration
von: Li, Zhen, et al.
Veröffentlicht: (2025)
von: Li, Zhen, et al.
Veröffentlicht: (2025)
Cloning Deterministic Worlds: The Critical Role of Latent Geometry in Long-Horizon World Models
von: Xia, Zaishuo, et al.
Veröffentlicht: (2025)
von: Xia, Zaishuo, et al.
Veröffentlicht: (2025)
RenderWorld: World Model with Self-Supervised 3D Label
von: Yan, Ziyang, et al.
Veröffentlicht: (2024)
von: Yan, Ziyang, et al.
Veröffentlicht: (2024)
Predictive but Not Plannable: RC-aux for Latent World Models
von: Li, Wenyuan, et al.
Veröffentlicht: (2026)
von: Li, Wenyuan, et al.
Veröffentlicht: (2026)
Accurate and Efficient World Modeling with Masked Latent Transformers
von: Burchi, Maxime, et al.
Veröffentlicht: (2025)
von: Burchi, Maxime, et al.
Veröffentlicht: (2025)
X-World: Controllable Ego-Centric Multi-Camera World Models for Scalable End-to-End Driving
von: Zheng, Chaoda, et al.
Veröffentlicht: (2026)
von: Zheng, Chaoda, et al.
Veröffentlicht: (2026)
An Embodied Generalist Agent in 3D World
von: Huang, Jiangyong, et al.
Veröffentlicht: (2023)
von: Huang, Jiangyong, et al.
Veröffentlicht: (2023)
GeoWorld-VLM: Geometry from World Models for Vision-Language Models
von: Gu, Renjie, et al.
Veröffentlicht: (2026)
von: Gu, Renjie, et al.
Veröffentlicht: (2026)
Recurrent Reasoning with Vision-Language Models for Estimating Long-Horizon Embodied Task Progress
von: Zhang, Yuelin, et al.
Veröffentlicht: (2026)
von: Zhang, Yuelin, et al.
Veröffentlicht: (2026)
Planning in 8 Tokens: A Compact Discrete Tokenizer for Latent World Model
von: Kim, Dongwon, et al.
Veröffentlicht: (2026)
von: Kim, Dongwon, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
World-Ego Modeling for Long-Horizon Evolution in Hybrid Embodied Tasks
von: Lin, Zuyao, et al.
Veröffentlicht: (2026) -
You Only Teach Once: Learn One-Shot Bimanual Robotic Manipulation from Video Demonstrations
von: Zhou, Huayi, et al.
Veröffentlicht: (2025) -
Rethinking Video Generation Model for the Embodied World
von: Deng, Yufan, et al.
Veröffentlicht: (2026) -
Learning Latent Action World Models In The Wild
von: Garrido, Quentin, et al.
Veröffentlicht: (2026) -
ADAM: An Embodied Causal Agent in Open-World Environments
von: Yu, Shu, et al.
Veröffentlicht: (2024)