Chain of World: World Model Thinking in Latent Motion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Fuxiang, Di, Donglin, Tang, Lulu, Zhang, Xuancheng, Fan, Lei, Li, Hao, Wei, Chen, Su, Tonghua, Ma, Baorui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ReWorld: Multi-Dimensional Reward Modeling for Embodied World Models
von: Peng, Baorui, et al.
Veröffentlicht: (2026)
von: Peng, Baorui, et al.
Veröffentlicht: (2026)
Latent Chain-of-Thought World Modeling for End-to-End Driving
von: Tan, Shuhan, et al.
Veröffentlicht: (2025)
von: Tan, Shuhan, et al.
Veröffentlicht: (2025)
WorldMAP: Bootstrapping Vision-Language Navigation Trajectory Prediction with Generative World Models
von: Chen, Hongjin, et al.
Veröffentlicht: (2026)
von: Chen, Hongjin, et al.
Veröffentlicht: (2026)
EFDiT: Efficient Fine-grained Image Generation Using Diffusion Transformer Models
von: Wang, Kun, et al.
Veröffentlicht: (2025)
von: Wang, Kun, et al.
Veröffentlicht: (2025)
FOCA: Frequency-Oriented Cross-Domain Forgery Detection, Localization and Explanation via Multi-Modal Large Language Model
von: Liu, Zhou, et al.
Veröffentlicht: (2026)
von: Liu, Zhou, et al.
Veröffentlicht: (2026)
Global-Local Aware Scene Text Editing
von: Yang, Fuxiang, et al.
Veröffentlicht: (2025)
von: Yang, Fuxiang, et al.
Veröffentlicht: (2025)
Motus: A Unified Latent Action World Model
von: Bi, Hongzhe, et al.
Veröffentlicht: (2025)
von: Bi, Hongzhe, et al.
Veröffentlicht: (2025)
Latent-WAM: Latent World Action Modeling for End-to-End Autonomous Driving
von: Wang, Linbo, et al.
Veröffentlicht: (2026)
von: Wang, Linbo, et al.
Veröffentlicht: (2026)
DriveWorld-VLA: Unified Latent-Space World Modeling with Vision-Language-Action for Autonomous Driving
von: jia, Feiyang, et al.
Veröffentlicht: (2026)
von: jia, Feiyang, et al.
Veröffentlicht: (2026)
WorldRFT: Latent World Model Planning with Reinforcement Fine-Tuning for Autonomous Driving
von: Yang, Pengxuan, et al.
Veröffentlicht: (2025)
von: Yang, Pengxuan, et al.
Veröffentlicht: (2025)
Latent Action Pretraining Through World Modeling
von: Tharwat, Bahey, et al.
Veröffentlicht: (2025)
von: Tharwat, Bahey, et al.
Veröffentlicht: (2025)
WorldArena: A Unified Benchmark for Evaluating Perception and Functional Utility of Embodied World Models
von: Shang, Yu, et al.
Veröffentlicht: (2026)
von: Shang, Yu, et al.
Veröffentlicht: (2026)
CogPortrait: Fine-Grained Eye-Region Control in Portrait Animation via Hierarchical Agent Planning
von: Feng, He, et al.
Veröffentlicht: (2026)
von: Feng, He, et al.
Veröffentlicht: (2026)
RoboScape: Physics-informed Embodied World Model
von: Shang, Yu, et al.
Veröffentlicht: (2025)
von: Shang, Yu, et al.
Veröffentlicht: (2025)
WorldArena 2.0: Extending Embodied World Model Benchmarking on Modality, Functionality and Platform
von: Shang, Yu, et al.
Veröffentlicht: (2026)
von: Shang, Yu, et al.
Veröffentlicht: (2026)
FlowDreamer: A RGB-D World Model with Flow-based Motion Representations for Robot Manipulation
von: Guo, Jun, et al.
Veröffentlicht: (2025)
von: Guo, Jun, et al.
Veröffentlicht: (2025)
MWM: Mobile World Models for Action-Conditioned Consistent Prediction
von: Yan, Han, et al.
Veröffentlicht: (2026)
von: Yan, Han, et al.
Veröffentlicht: (2026)
Occupancy World Model for Robots
von: Zhang, Zhang, et al.
Veröffentlicht: (2025)
von: Zhang, Zhang, et al.
Veröffentlicht: (2025)
DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving
von: Zhang, Lingjun, et al.
Veröffentlicht: (2026)
von: Zhang, Lingjun, et al.
Veröffentlicht: (2026)
MetricAnything: Scaling Metric Depth Pretraining with Noisy Heterogeneous Sources
von: Ma, Baorui, et al.
Veröffentlicht: (2026)
von: Ma, Baorui, et al.
Veröffentlicht: (2026)
VLA-JEPA: Enhancing Vision-Language-Action Model with Latent World Model
von: Sun, Jingwen, et al.
Veröffentlicht: (2026)
von: Sun, Jingwen, et al.
Veröffentlicht: (2026)
Driver-WM: A Driver-Centric Traffic-Conditioned Latent World Model for In-Cabin Dynamics Rollout
von: Chi, Haozhuang, et al.
Veröffentlicht: (2026)
von: Chi, Haozhuang, et al.
Veröffentlicht: (2026)
World Guidance: World Modeling in Condition Space for Action Generation
von: Su, Yue, et al.
Veröffentlicht: (2026)
von: Su, Yue, et al.
Veröffentlicht: (2026)
KeyWorld: Key Frame Reasoning Enables Effective and Efficient World Models
von: Li, Sibo, et al.
Veröffentlicht: (2025)
von: Li, Sibo, et al.
Veröffentlicht: (2025)
ABot-PhysWorld: Interactive World Foundation Model for Robotic Manipulation with Physics Alignment
von: Chen, Yuzhi, et al.
Veröffentlicht: (2026)
von: Chen, Yuzhi, et al.
Veröffentlicht: (2026)
Learning Massively Multitask World Models for Continuous Control
von: Hansen, Nicklas, et al.
Veröffentlicht: (2025)
von: Hansen, Nicklas, et al.
Veröffentlicht: (2025)
DiTalker: A Unified DiT-based Framework for High-Quality and Speaking Styles Controllable Portrait Animation
von: Feng, He, et al.
Veröffentlicht: (2025)
von: Feng, He, et al.
Veröffentlicht: (2025)
Uni-World VLA: Interleaved World Modeling and Planning for Autonomous Driving
von: Liu, Qiqi, et al.
Veröffentlicht: (2026)
von: Liu, Qiqi, et al.
Veröffentlicht: (2026)
GeoWorld: Geometric World Models
von: Zhang, Zeyu, et al.
Veröffentlicht: (2026)
von: Zhang, Zeyu, et al.
Veröffentlicht: (2026)
UAV-Flow Colosseo: A Real-World Benchmark for Flying-on-a-Word UAV Imitation Learning
von: Wang, Xiangyu, et al.
Veröffentlicht: (2025)
von: Wang, Xiangyu, et al.
Veröffentlicht: (2025)
Learning 3D Persistent Embodied World Models
von: Zhou, Siyuan, et al.
Veröffentlicht: (2025)
von: Zhou, Siyuan, et al.
Veröffentlicht: (2025)
VLA-RFT: Vision-Language-Action Reinforcement Fine-tuning with Verified Rewards in World Simulators
von: Li, Hengtao, et al.
Veröffentlicht: (2025)
von: Li, Hengtao, et al.
Veröffentlicht: (2025)
DiLA: Disentangled Latent Action World Models
von: Zhang, Tianqiu, et al.
Veröffentlicht: (2026)
von: Zhang, Tianqiu, et al.
Veröffentlicht: (2026)
One-Shot Pose-Driving Face Animation Platform
von: Feng, He, et al.
Veröffentlicht: (2024)
von: Feng, He, et al.
Veröffentlicht: (2024)
Open-World Motion Forecasting
von: Schischka, Nicolas, et al.
Veröffentlicht: (2026)
von: Schischka, Nicolas, et al.
Veröffentlicht: (2026)
CoIRL-AD: Collaborative-Competitive Imitation-Reinforcement Learning in Latent World Models for Autonomous Driving
von: Zheng, Xiaoji, et al.
Veröffentlicht: (2025)
von: Zheng, Xiaoji, et al.
Veröffentlicht: (2025)
AdaWorld: Learning Adaptable World Models with Latent Actions
von: Gao, Shenyuan, et al.
Veröffentlicht: (2025)
von: Gao, Shenyuan, et al.
Veröffentlicht: (2025)
WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation
von: Zhao, Baining, et al.
Veröffentlicht: (2026)
von: Zhao, Baining, et al.
Veröffentlicht: (2026)
DeformMaster: An Interactive Physics-Neural World Model for Deformable Objects from Videos
von: Li, Can, et al.
Veröffentlicht: (2026)
von: Li, Can, et al.
Veröffentlicht: (2026)
A Survey of World Models for Autonomous Driving
von: Feng, Tuo, et al.
Veröffentlicht: (2025)
von: Feng, Tuo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ReWorld: Multi-Dimensional Reward Modeling for Embodied World Models
von: Peng, Baorui, et al.
Veröffentlicht: (2026) -
Latent Chain-of-Thought World Modeling for End-to-End Driving
von: Tan, Shuhan, et al.
Veröffentlicht: (2025) -
WorldMAP: Bootstrapping Vision-Language Navigation Trajectory Prediction with Generative World Models
von: Chen, Hongjin, et al.
Veröffentlicht: (2026) -
EFDiT: Efficient Fine-grained Image Generation Using Diffusion Transformer Models
von: Wang, Kun, et al.
Veröffentlicht: (2025) -
FOCA: Frequency-Oriented Cross-Domain Forgery Detection, Localization and Explanation via Multi-Modal Large Language Model
von: Liu, Zhou, et al.
Veröffentlicht: (2026)