DDP-WM: Disentangled Dynamics Prediction for Efficient World Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yin, Shicheng, Yin, Kaixuan, Chen, Weixing, Liu, Yang, Li, Guanbin, Lin, Liang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DART: Differentiable Dynamic Adaptive Region Tokenizer for Vision Foundation Models
von: Yin, Shicheng, et al.
Veröffentlicht: (2025)
von: Yin, Shicheng, et al.
Veröffentlicht: (2025)
VisionGRU: A Linear-Complexity RNN Model for Efficient Image Analysis
von: Yin, Shicheng, et al.
Veröffentlicht: (2024)
von: Yin, Shicheng, et al.
Veröffentlicht: (2024)
Aligning Cyber Space with Physical World: A Comprehensive Survey on Embodied AI
von: Liu, Yang, et al.
Veröffentlicht: (2024)
von: Liu, Yang, et al.
Veröffentlicht: (2024)
Driver-WM: A Driver-Centric Traffic-Conditioned Latent World Model for In-Cabin Dynamics Rollout
von: Chi, Haozhuang, et al.
Veröffentlicht: (2026)
von: Chi, Haozhuang, et al.
Veröffentlicht: (2026)
AutoLayout: Closed-Loop Layout Synthesis via Slow-Fast Collaborative Reasoning
von: Chen, Weixing, et al.
Veröffentlicht: (2025)
von: Chen, Weixing, et al.
Veröffentlicht: (2025)
Learning to See and Act: Task-Aware Virtual View Exploration for Robotic Manipulation
von: Bai, Yongjie, et al.
Veröffentlicht: (2025)
von: Bai, Yongjie, et al.
Veröffentlicht: (2025)
EA-WM: Event-Aware Generative World Model with Structured Kinematic-to-Visual Action Fields
von: Yang, Zhaoyang, et al.
Veröffentlicht: (2026)
von: Yang, Zhaoyang, et al.
Veröffentlicht: (2026)
ContactGaussian-WM: Learning Physics-Grounded World Model from Videos
von: Wang, Meizhong, et al.
Veröffentlicht: (2026)
von: Wang, Meizhong, et al.
Veröffentlicht: (2026)
MEIA: Multimodal Embodied Perception and Interaction in Unknown Environments
von: Liu, Yang, et al.
Veröffentlicht: (2024)
von: Liu, Yang, et al.
Veröffentlicht: (2024)
Beyond the Destination: A Novel Benchmark for Exploration-Aware Embodied Question Answering
von: Jiang, Kaixuan, et al.
Veröffentlicht: (2025)
von: Jiang, Kaixuan, et al.
Veröffentlicht: (2025)
3DAffordSplat: Efficient Affordance Reasoning with 3D Gaussians
von: Wei, Zeming, et al.
Veröffentlicht: (2025)
von: Wei, Zeming, et al.
Veröffentlicht: (2025)
Efficient Long-Horizon Vision-Language-Action Models via Static-Dynamic Disentanglement
von: Qiu, Weikang, et al.
Veröffentlicht: (2026)
von: Qiu, Weikang, et al.
Veröffentlicht: (2026)
Inference-Time Enhancement of Generative Robot Policies via Predictive World Modeling
von: Qi, Han, et al.
Veröffentlicht: (2025)
von: Qi, Han, et al.
Veröffentlicht: (2025)
LiDARCrafter: Dynamic 4D World Modeling from LiDAR Sequences
von: Liang, Ao, et al.
Veröffentlicht: (2025)
von: Liang, Ao, et al.
Veröffentlicht: (2025)
SKIP: Sparse Keyframe Interpolation Paradigm for Efficient Embodied World Models
von: He, Ziheng, et al.
Veröffentlicht: (2026)
von: He, Ziheng, et al.
Veröffentlicht: (2026)
Towards Long-Horizon Vision-Language Navigation: Platform, Benchmark and Method
von: Song, Xinshuai, et al.
Veröffentlicht: (2024)
von: Song, Xinshuai, et al.
Veröffentlicht: (2024)
Learning to Drive from a World Model
von: Goff, Mitchell, et al.
Veröffentlicht: (2025)
von: Goff, Mitchell, et al.
Veröffentlicht: (2025)
VectorWorld: Efficient Streaming World Model via Diffusion Flow on Vector Graphs
von: Jiang, Chaokang, et al.
Veröffentlicht: (2026)
von: Jiang, Chaokang, et al.
Veröffentlicht: (2026)
Is Your Driving World Model an All-Around Player?
von: Kong, Lingdong, et al.
Veröffentlicht: (2026)
von: Kong, Lingdong, et al.
Veröffentlicht: (2026)
3D and 4D World Modeling: A Survey
von: Kong, Lingdong, et al.
Veröffentlicht: (2025)
von: Kong, Lingdong, et al.
Veröffentlicht: (2025)
DynFlowDrive: Flow-Based Dynamic World Modeling for Autonomous Driving
von: Liu, Xiaolu, et al.
Veröffentlicht: (2026)
von: Liu, Xiaolu, et al.
Veröffentlicht: (2026)
KeyWorld: Key Frame Reasoning Enables Effective and Efficient World Models
von: Li, Sibo, et al.
Veröffentlicht: (2025)
von: Li, Sibo, et al.
Veröffentlicht: (2025)
Multi-modal Motion Prediction using Temporal Ensembling with Learning-based Aggregation
von: Hong, Kai-Yin, et al.
Veröffentlicht: (2024)
von: Hong, Kai-Yin, et al.
Veröffentlicht: (2024)
PhysFire-WM: A Physics-Informed World Model for Emulating Fire Spread Dynamics
von: Zhou, Nan, et al.
Veröffentlicht: (2025)
von: Zhou, Nan, et al.
Veröffentlicht: (2025)
OccRWKV: Rethinking Efficient 3D Semantic Occupancy Prediction with Linear Complexity
von: Wang, Junming, et al.
Veröffentlicht: (2024)
von: Wang, Junming, et al.
Veröffentlicht: (2024)
Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining
von: Zhang, Wenyao, et al.
Veröffentlicht: (2026)
von: Zhang, Wenyao, et al.
Veröffentlicht: (2026)
BridgeV2W: Bridging Video Generation Models to Embodied World Models via Embodiment Masks
von: Chen, Yixiang, et al.
Veröffentlicht: (2026)
von: Chen, Yixiang, et al.
Veröffentlicht: (2026)
DiLA: Disentangled Latent Action World Models
von: Zhang, Tianqiu, et al.
Veröffentlicht: (2026)
von: Zhang, Tianqiu, et al.
Veröffentlicht: (2026)
ReWorld: Multi-Dimensional Reward Modeling for Embodied World Models
von: Peng, Baorui, et al.
Veröffentlicht: (2026)
von: Peng, Baorui, et al.
Veröffentlicht: (2026)
ABot-PhysWorld: Interactive World Foundation Model for Robotic Manipulation with Physics Alignment
von: Chen, Yuzhi, et al.
Veröffentlicht: (2026)
von: Chen, Yuzhi, et al.
Veröffentlicht: (2026)
DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving
von: Zhang, Lingjun, et al.
Veröffentlicht: (2026)
von: Zhang, Lingjun, et al.
Veröffentlicht: (2026)
DriveWorld-VLA: Unified Latent-Space World Modeling with Vision-Language-Action for Autonomous Driving
von: jia, Feiyang, et al.
Veröffentlicht: (2026)
von: jia, Feiyang, et al.
Veröffentlicht: (2026)
Cross-Modal Causal Intervention for Medical Report Generation
von: Chen, Weixing, et al.
Veröffentlicht: (2023)
von: Chen, Weixing, et al.
Veröffentlicht: (2023)
DSPNet: Dual-vision Scene Perception for Robust 3D Question Answering
von: Luo, Jingzhou, et al.
Veröffentlicht: (2025)
von: Luo, Jingzhou, et al.
Veröffentlicht: (2025)
World Guidance: World Modeling in Condition Space for Action Generation
von: Su, Yue, et al.
Veröffentlicht: (2026)
von: Su, Yue, et al.
Veröffentlicht: (2026)
SANA-WM: Efficient Minute-Scale World Modeling with Hybrid Linear Diffusion Transformer
von: Zhu, Haoyi, et al.
Veröffentlicht: (2026)
von: Zhu, Haoyi, et al.
Veröffentlicht: (2026)
IntentionNav: A Benchmark for Intent-Driven Object Navigation from Implicit Human Instruction
von: Qian, Lin, et al.
Veröffentlicht: (2026)
von: Qian, Lin, et al.
Veröffentlicht: (2026)
U4D: Uncertainty-Aware 4D World Modeling from LiDAR Sequences
von: Xu, Xiang, et al.
Veröffentlicht: (2025)
von: Xu, Xiang, et al.
Veröffentlicht: (2025)
WorldMAP: Bootstrapping Vision-Language Navigation Trajectory Prediction with Generative World Models
von: Chen, Hongjin, et al.
Veröffentlicht: (2026)
von: Chen, Hongjin, et al.
Veröffentlicht: (2026)
Causal World Modeling for Robot Control
von: Li, Lin, et al.
Veröffentlicht: (2026)
von: Li, Lin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
DART: Differentiable Dynamic Adaptive Region Tokenizer for Vision Foundation Models
von: Yin, Shicheng, et al.
Veröffentlicht: (2025) -
VisionGRU: A Linear-Complexity RNN Model for Efficient Image Analysis
von: Yin, Shicheng, et al.
Veröffentlicht: (2024) -
Aligning Cyber Space with Physical World: A Comprehensive Survey on Embodied AI
von: Liu, Yang, et al.
Veröffentlicht: (2024) -
Driver-WM: A Driver-Centric Traffic-Conditioned Latent World Model for In-Cabin Dynamics Rollout
von: Chi, Haozhuang, et al.
Veröffentlicht: (2026) -
AutoLayout: Closed-Loop Layout Synthesis via Slow-Fast Collaborative Reasoning
von: Chen, Weixing, et al.
Veröffentlicht: (2025)