InstaDrive: Instance-Aware Driving World Models for Realistic and Consistent Video Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Zhuoran, Guo, Xi, Ding, Chenjing, Wang, Chiyu, Wu, Wei, Zhang, Yanyong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Physical Informed Driving World Model
by: Yang, Zhuoran, et al.
Published: (2024)
by: Yang, Zhuoran, et al.
Published: (2024)
ConsisDrive: Identity-Preserving Driving World Models for Video Generation by Instance Mask
by: Yang, Zhuoran, et al.
Published: (2026)
by: Yang, Zhuoran, et al.
Published: (2026)
DriveScape: Towards High-Resolution Controllable Multi-View Driving Video Generation
by: Wu, Wei, et al.
Published: (2024)
by: Wu, Wei, et al.
Published: (2024)
MyGo: Consistent and Controllable Multi-View Driving Video Generation with Camera Control
by: Yao, Yining, et al.
Published: (2024)
by: Yao, Yining, et al.
Published: (2024)
InfinityDrive: Breaking Time Limits in Driving World Models
by: Guo, Xi, et al.
Published: (2024)
by: Guo, Xi, et al.
Published: (2024)
SGC-VQGAN: Towards Complex Scene Representation via Semantic Guided Clustering Codebook
by: Ding, Chenjing, et al.
Published: (2024)
by: Ding, Chenjing, et al.
Published: (2024)
CVD-STORM: Cross-View Video Diffusion with Spatial-Temporal Reconstruction Model for Autonomous Driving
by: Zhang, Tianrui, et al.
Published: (2025)
by: Zhang, Tianrui, et al.
Published: (2025)
DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT
by: Hu, Xiaotao, et al.
Published: (2024)
by: Hu, Xiaotao, et al.
Published: (2024)
MARS: An Instance-aware, Modular and Realistic Simulator for Autonomous Driving
by: Wu, Zirui, et al.
Published: (2023)
by: Wu, Zirui, et al.
Published: (2023)
Drive&Gen: Co-Evaluating End-to-End Driving and Video Generation Models
by: Wang, Jiahao, et al.
Published: (2025)
by: Wang, Jiahao, et al.
Published: (2025)
GenieDrive: Towards Physics-Aware Driving World Model with 4D Occupancy Guided Video Generation
by: Yang, Zhenya, et al.
Published: (2025)
by: Yang, Zhenya, et al.
Published: (2025)
InstaVSR: Taming Diffusion for Efficient and Temporally Consistent Video Super-Resolution
by: Hu, Jintong, et al.
Published: (2026)
by: Hu, Jintong, et al.
Published: (2026)
InstaDA: Augmenting Instance Segmentation Data with Dual-Agent System
by: Hou, Xianbao, et al.
Published: (2025)
by: Hou, Xianbao, et al.
Published: (2025)
Toward Physically Consistent Driving Video World Models under Challenging Trajectories
by: Zhou, Jiawei, et al.
Published: (2026)
by: Zhou, Jiawei, et al.
Published: (2026)
FantasyWorld: Geometry-Consistent World Modeling via Unified Video and 3D Prediction
by: Dai, Yixiang, et al.
Published: (2025)
by: Dai, Yixiang, et al.
Published: (2025)
DriveWAM: Video Generative Priors Enable Scalable World-Action Modeling for Autonomous Driving
by: Shi, Chen, et al.
Published: (2026)
by: Shi, Chen, et al.
Published: (2026)
DriveFuture: Future-Aware Latent World Models for Autonomous Driving
by: Hong, Yufeng, et al.
Published: (2026)
by: Hong, Yufeng, et al.
Published: (2026)
DreamWorld: Unified World Modeling in Video Generation
by: Tan, Boming, et al.
Published: (2026)
by: Tan, Boming, et al.
Published: (2026)
DriveDreamer-2: LLM-Enhanced World Models for Diverse Driving Video Generation
by: Zhao, Guosheng, et al.
Published: (2024)
by: Zhao, Guosheng, et al.
Published: (2024)
InstDrive: Instance-Aware 3D Gaussian Splatting for Driving Scenes
by: Liu, Hongyuan, et al.
Published: (2025)
by: Liu, Hongyuan, et al.
Published: (2025)
Foundation Models for Amodal Video Instance Segmentation in Automated Driving
by: Breitenstein, Jasmin, et al.
Published: (2024)
by: Breitenstein, Jasmin, et al.
Published: (2024)
Realistic and Controllable 3D Gaussian-Guided Object Editing for Driving Video Generation
by: Li, Jiusi, et al.
Published: (2025)
by: Li, Jiusi, et al.
Published: (2025)
InstaGraM: Instance-level Graph Modeling for Vectorized HD Map Learning
by: Shin, Juyeb, et al.
Published: (2023)
by: Shin, Juyeb, et al.
Published: (2023)
MaskGWM: A Generalizable Driving World Model with Video Mask Reconstruction
by: Ni, Jingcheng, et al.
Published: (2025)
by: Ni, Jingcheng, et al.
Published: (2025)
DrivingGen: A Comprehensive Benchmark for Generative Video World Models in Autonomous Driving
by: Zhou, Yang, et al.
Published: (2026)
by: Zhou, Yang, et al.
Published: (2026)
HorizonDrive: Self-Corrective Autoregressive World Model for Long-horizon Driving Simulation
by: Zhang, Conglang, et al.
Published: (2026)
by: Zhang, Conglang, et al.
Published: (2026)
DriveLaW:Unifying Planning and Video Generation in a Latent Driving World
by: Xia, Tianze, et al.
Published: (2025)
by: Xia, Tianze, et al.
Published: (2025)
Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving
by: Wang, Shumin, et al.
Published: (2025)
by: Wang, Shumin, et al.
Published: (2025)
InstaScene: Towards Complete 3D Instance Decomposition and Reconstruction from Cluttered Scenes
by: Yang, Zesong, et al.
Published: (2025)
by: Yang, Zesong, et al.
Published: (2025)
DriveSceneGen: Generating Diverse and Realistic Driving Scenarios from Scratch
by: Sun, Shuo, et al.
Published: (2023)
by: Sun, Shuo, et al.
Published: (2023)
Comparison Drives Preference: Reference-Aware Modeling for AI-Generated Video Quality Assessment
by: Zou, Minghao, et al.
Published: (2026)
by: Zou, Minghao, et al.
Published: (2026)
DreamForge: Motion-Aware Autoregressive Video Generation for Multi-View Driving Scenes
by: Mei, Jianbiao, et al.
Published: (2024)
by: Mei, Jianbiao, et al.
Published: (2024)
Stag-1: Towards Realistic 4D Driving Simulation with Video Generation Model
by: Wang, Lening, et al.
Published: (2024)
by: Wang, Lening, et al.
Published: (2024)
DrivingGaussian++: Towards Realistic Reconstruction and Editable Simulation for Surrounding Dynamic Driving Scenes
by: Xiong, Yajiao, et al.
Published: (2025)
by: Xiong, Yajiao, et al.
Published: (2025)
Cosmos-Drive-Dreams: Scalable Synthetic Driving Data Generation with World Foundation Models
by: Ren, Xuanchi, et al.
Published: (2025)
by: Ren, Xuanchi, et al.
Published: (2025)
Instance-Aware Robust Consistency Regularization for Semi-Supervised Nuclei Instance Segmentation
by: Lin, Zenan, et al.
Published: (2025)
by: Lin, Zenan, et al.
Published: (2025)
Epona: Autoregressive Diffusion World Model for Autonomous Driving
by: Zhang, Kaiwen, et al.
Published: (2025)
by: Zhang, Kaiwen, et al.
Published: (2025)
Geometry-Aware Rotary Position Embedding for Consistent Video World Model
by: Xiang, Chendong, et al.
Published: (2026)
by: Xiang, Chendong, et al.
Published: (2026)
InstanceRSR: Real-World Super-Resolution via Instance-Aware Representation Alignment
by: Guo, Zixin, et al.
Published: (2026)
by: Guo, Zixin, et al.
Published: (2026)
UniDrive-WM: Unified Understanding, Planning and Generation World Model For Autonomous Driving
by: Xiong, Zhexiao, et al.
Published: (2026)
by: Xiong, Zhexiao, et al.
Published: (2026)
Similar Items
-
Physical Informed Driving World Model
by: Yang, Zhuoran, et al.
Published: (2024) -
ConsisDrive: Identity-Preserving Driving World Models for Video Generation by Instance Mask
by: Yang, Zhuoran, et al.
Published: (2026) -
DriveScape: Towards High-Resolution Controllable Multi-View Driving Video Generation
by: Wu, Wei, et al.
Published: (2024) -
MyGo: Consistent and Controllable Multi-View Driving Video Generation with Camera Control
by: Yao, Yining, et al.
Published: (2024) -
InfinityDrive: Breaking Time Limits in Driving World Models
by: Guo, Xi, et al.
Published: (2024)