VectorWorld: Efficient Streaming World Model via Diffusion Flow on Vector Graphs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jiang, Chaokang, Zhou, Desen, Liu, Jiuming, Sun, Kevin Li |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GMF-Drive: Gated Mamba Fusion with Spatial-Aware BEV Representation for End-to-End Autonomous Driving
von: Wang, Jian, et al.
Veröffentlicht: (2025)
von: Wang, Jian, et al.
Veröffentlicht: (2025)
Dual-Stream Diffusion for World-Model Augmented Vision-Language-Action Model
von: Won, John, et al.
Veröffentlicht: (2025)
von: Won, John, et al.
Veröffentlicht: (2025)
CapVector: Learning Transferable Capability Vectors in Parametric Space for Vision-Language-Action Models
von: Song, Wenxuan, et al.
Veröffentlicht: (2026)
von: Song, Wenxuan, et al.
Veröffentlicht: (2026)
DifFlow3D: Toward Robust Uncertainty-Aware Scene Flow Estimation with Diffusion Model
von: Liu, Jiuming, et al.
Veröffentlicht: (2023)
von: Liu, Jiuming, et al.
Veröffentlicht: (2023)
DynFlowDrive: Flow-Based Dynamic World Modeling for Autonomous Driving
von: Liu, Xiaolu, et al.
Veröffentlicht: (2026)
von: Liu, Xiaolu, et al.
Veröffentlicht: (2026)
KeyWorld: Key Frame Reasoning Enables Effective and Efficient World Models
von: Li, Sibo, et al.
Veröffentlicht: (2025)
von: Li, Sibo, et al.
Veröffentlicht: (2025)
Uni-World VLA: Interleaved World Modeling and Planning for Autonomous Driving
von: Liu, Qiqi, et al.
Veröffentlicht: (2026)
von: Liu, Qiqi, et al.
Veröffentlicht: (2026)
VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning
von: Jiang, Bo, et al.
Veröffentlicht: (2024)
von: Jiang, Bo, et al.
Veröffentlicht: (2024)
RealD$^2$iff: Bridging Real-World Gap in Robot Manipulation via Depth Diffusion
von: Liang, Xiujian, et al.
Veröffentlicht: (2025)
von: Liang, Xiujian, et al.
Veröffentlicht: (2025)
Representing Domain-Mixing Optical Degradation for Real-World Computational Aberration Correction via Vector Quantization
von: Jiang, Qi, et al.
Veröffentlicht: (2024)
von: Jiang, Qi, et al.
Veröffentlicht: (2024)
DDP-WM: Disentangled Dynamics Prediction for Efficient World Models
von: Yin, Shicheng, et al.
Veröffentlicht: (2026)
von: Yin, Shicheng, et al.
Veröffentlicht: (2026)
FlowDreamer: A RGB-D World Model with Flow-based Motion Representations for Robot Manipulation
von: Guo, Jun, et al.
Veröffentlicht: (2025)
von: Guo, Jun, et al.
Veröffentlicht: (2025)
GeoWorld: Geometric World Models
von: Zhang, Zeyu, et al.
Veröffentlicht: (2026)
von: Zhang, Zeyu, et al.
Veröffentlicht: (2026)
Residual Vector Quantization For Communication-Efficient Multi-Agent Perception
von: Shenkut, Dereje, et al.
Veröffentlicht: (2025)
von: Shenkut, Dereje, et al.
Veröffentlicht: (2025)
Scenario Dreamer: Vectorized Latent Diffusion for Generating Driving Simulation Environments
von: Rowe, Luke, et al.
Veröffentlicht: (2025)
von: Rowe, Luke, et al.
Veröffentlicht: (2025)
SKIP: Sparse Keyframe Interpolation Paradigm for Efficient Embodied World Models
von: He, Ziheng, et al.
Veröffentlicht: (2026)
von: He, Ziheng, et al.
Veröffentlicht: (2026)
World Guidance: World Modeling in Condition Space for Action Generation
von: Su, Yue, et al.
Veröffentlicht: (2026)
von: Su, Yue, et al.
Veröffentlicht: (2026)
TesserAct: Learning 4D Embodied World Models
von: Zhen, Haoyu, et al.
Veröffentlicht: (2025)
von: Zhen, Haoyu, et al.
Veröffentlicht: (2025)
WristWorld: Generating Wrist-Views via 4D World Models for Robotic Manipulation
von: Qian, Zezhong, et al.
Veröffentlicht: (2025)
von: Qian, Zezhong, et al.
Veröffentlicht: (2025)
VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers
von: Wang, Yating, et al.
Veröffentlicht: (2025)
von: Wang, Yating, et al.
Veröffentlicht: (2025)
OpenSGA: Efficient 3D Scene Graph Alignment in the Open World
von: Chen, Gang, et al.
Veröffentlicht: (2026)
von: Chen, Gang, et al.
Veröffentlicht: (2026)
ReWorld: Multi-Dimensional Reward Modeling for Embodied World Models
von: Peng, Baorui, et al.
Veröffentlicht: (2026)
von: Peng, Baorui, et al.
Veröffentlicht: (2026)
Is Your Driving World Model an All-Around Player?
von: Kong, Lingdong, et al.
Veröffentlicht: (2026)
von: Kong, Lingdong, et al.
Veröffentlicht: (2026)
Egocentric World Model for Photorealistic Hand-Object Interaction Synthesis
von: Li, Dayou, et al.
Veröffentlicht: (2026)
von: Li, Dayou, et al.
Veröffentlicht: (2026)
3DFlowAction: Learning Cross-Embodiment Manipulation from 3D Flow World Model
von: Zhi, Hongyan, et al.
Veröffentlicht: (2025)
von: Zhi, Hongyan, et al.
Veröffentlicht: (2025)
MAMBA4D: Efficient Long-Sequence Point Cloud Video Understanding with Disentangled Spatial-Temporal State Space Models
von: Liu, Jiuming, et al.
Veröffentlicht: (2024)
von: Liu, Jiuming, et al.
Veröffentlicht: (2024)
GigaWorld-0: World Models as Data Engine to Empower Embodied AI
von: GigaWorld Team, et al.
Veröffentlicht: (2025)
von: GigaWorld Team, et al.
Veröffentlicht: (2025)
MapTRv2: An End-to-End Framework for Online Vectorized HD Map Construction
von: Liao, Bencheng, et al.
Veröffentlicht: (2023)
von: Liao, Bencheng, et al.
Veröffentlicht: (2023)
Map-World: Masked Action planning and Path-Integral World Model for Autonomous Driving
von: Hu, Bin, et al.
Veröffentlicht: (2025)
von: Hu, Bin, et al.
Veröffentlicht: (2025)
Occupancy World Model for Robots
von: Zhang, Zhang, et al.
Veröffentlicht: (2025)
von: Zhang, Zhang, et al.
Veröffentlicht: (2025)
VLA-JEPA: Enhancing Vision-Language-Action Model with Latent World Model
von: Sun, Jingwen, et al.
Veröffentlicht: (2026)
von: Sun, Jingwen, et al.
Veröffentlicht: (2026)
Making Pose Representations More Expressive and Disentangled via Residual Vector Quantization
von: Jeong, Sukhyun, et al.
Veröffentlicht: (2025)
von: Jeong, Sukhyun, et al.
Veröffentlicht: (2025)
UniDWM: Towards a Unified Driving World Model via Multifaceted Representation Learning
von: Liu, Shuai, et al.
Veröffentlicht: (2026)
von: Liu, Shuai, et al.
Veröffentlicht: (2026)
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
von: Jiang, Anqing, et al.
Veröffentlicht: (2025)
von: Jiang, Anqing, et al.
Veröffentlicht: (2025)
WorldArena: A Unified Benchmark for Evaluating Perception and Functional Utility of Embodied World Models
von: Shang, Yu, et al.
Veröffentlicht: (2026)
von: Shang, Yu, et al.
Veröffentlicht: (2026)
WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation
von: Zhao, Baining, et al.
Veröffentlicht: (2026)
von: Zhao, Baining, et al.
Veröffentlicht: (2026)
Ego to World: Collaborative Spatial Reasoning in Embodied Systems via Reinforcement Learning
von: Zhou, Heng, et al.
Veröffentlicht: (2026)
von: Zhou, Heng, et al.
Veröffentlicht: (2026)
WorldArena 2.0: Extending Embodied World Model Benchmarking on Modality, Functionality and Platform
von: Shang, Yu, et al.
Veröffentlicht: (2026)
von: Shang, Yu, et al.
Veröffentlicht: (2026)
BridgeV2W: Bridging Video Generation Models to Embodied World Models via Embodiment Masks
von: Chen, Yixiang, et al.
Veröffentlicht: (2026)
von: Chen, Yixiang, et al.
Veröffentlicht: (2026)
ABot-PhysWorld: Interactive World Foundation Model for Robotic Manipulation with Physics Alignment
von: Chen, Yuzhi, et al.
Veröffentlicht: (2026)
von: Chen, Yuzhi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
GMF-Drive: Gated Mamba Fusion with Spatial-Aware BEV Representation for End-to-End Autonomous Driving
von: Wang, Jian, et al.
Veröffentlicht: (2025) -
Dual-Stream Diffusion for World-Model Augmented Vision-Language-Action Model
von: Won, John, et al.
Veröffentlicht: (2025) -
CapVector: Learning Transferable Capability Vectors in Parametric Space for Vision-Language-Action Models
von: Song, Wenxuan, et al.
Veröffentlicht: (2026) -
DifFlow3D: Toward Robust Uncertainty-Aware Scene Flow Estimation with Diffusion Model
von: Liu, Jiuming, et al.
Veröffentlicht: (2023) -
DynFlowDrive: Flow-Based Dynamic World Modeling for Autonomous Driving
von: Liu, Xiaolu, et al.
Veröffentlicht: (2026)