Vista: A Generalizable Driving World Model with High Fidelity and Versatile Controllability
Fuente:
arXiv
Saved in:
| Main Authors: | Gao, Shenyuan, Yang, Jiazhi, Chen, Li, Chitta, Kashyap, Qiu, Yihang, Geiger, Andreas, Zhang, Jun, Li, Hongyang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ReSim: Reliable World Simulation for Autonomous Driving
by: Yang, Jiazhi, et al.
Published: (2025)
by: Yang, Jiazhi, et al.
Published: (2025)
GenAD: Generalized Predictive Model for Autonomous Driving
by: Yang, Jiazhi, et al.
Published: (2024)
by: Yang, Jiazhi, et al.
Published: (2024)
SLEDGE: Synthesizing Driving Environments with Generative Models and Rule-Based Traffic
by: Chitta, Kashyap, et al.
Published: (2024)
by: Chitta, Kashyap, et al.
Published: (2024)
End-to-end Autonomous Driving: Challenges and Frontiers
by: Chen, Li, et al.
Published: (2023)
by: Chen, Li, et al.
Published: (2023)
DriveLM: Driving with Graph Visual Question Answering
by: Sima, Chonghao, et al.
Published: (2023)
by: Sima, Chonghao, et al.
Published: (2023)
Hidden Biases of End-to-End Driving Datasets
by: Zimmerlin, Julian, et al.
Published: (2024)
by: Zimmerlin, Julian, et al.
Published: (2024)
Centaur: Robust End-to-End Autonomous Driving with Test-Time Training
by: Sima, Chonghao, et al.
Published: (2025)
by: Sima, Chonghao, et al.
Published: (2025)
SimScale: Learning to Drive via Real-World Simulation at Scale
by: Tian, Haochen, et al.
Published: (2025)
by: Tian, Haochen, et al.
Published: (2025)
LEAD: Minimizing Learner-Expert Asymmetry in End-to-End Driving
by: Nguyen, Long, et al.
Published: (2025)
by: Nguyen, Long, et al.
Published: (2025)
Pseudo-Simulation for Autonomous Driving
by: Cao, Wei, et al.
Published: (2025)
by: Cao, Wei, et al.
Published: (2025)
StereoCarla: A High-Fidelity Driving Dataset for Generalizable Stereo
by: Guo, Xianda, et al.
Published: (2025)
by: Guo, Xianda, et al.
Published: (2025)
Latent Chain-of-Thought World Modeling for End-to-End Driving
by: Tan, Shuhan, et al.
Published: (2025)
by: Tan, Shuhan, et al.
Published: (2025)
DriveX: Omni Scene Modeling for Learning Generalizable World Knowledge in Autonomous Driving
by: Shi, Chen, et al.
Published: (2025)
by: Shi, Chen, et al.
Published: (2025)
123D: Unifying Multi-Modal Autonomous Driving Data at Scale
by: Dauner, Daniel, et al.
Published: (2026)
by: Dauner, Daniel, et al.
Published: (2026)
Surface-Centric Modeling for High-Fidelity Generalizable Neural Surface Reconstruction
by: Peng, Rui, et al.
Published: (2024)
by: Peng, Rui, et al.
Published: (2024)
DOME: Taming Diffusion Model into High-Fidelity Controllable Occupancy World Model
by: Gu, Songen, et al.
Published: (2024)
by: Gu, Songen, et al.
Published: (2024)
AdaWorld: Learning Adaptable World Models with Latent Actions
by: Gao, Shenyuan, et al.
Published: (2025)
by: Gao, Shenyuan, et al.
Published: (2025)
REVISION: Rendering Tools Enable Spatial Fidelity in Vision-Language Models
by: Chatterjee, Agneet, et al.
Published: (2024)
by: Chatterjee, Agneet, et al.
Published: (2024)
Fail2Drive: Benchmarking Closed-Loop Driving Generalization
by: Gerstenecker, Simon, et al.
Published: (2026)
by: Gerstenecker, Simon, et al.
Published: (2026)
Optimization-Guided Diffusion for Interactive Scene Generation
by: Li, Shihao, et al.
Published: (2025)
by: Li, Shihao, et al.
Published: (2025)
NAVSIM: Data-Driven Non-Reactive Autonomous Vehicle Simulation and Benchmarking
by: Dauner, Daniel, et al.
Published: (2024)
by: Dauner, Daniel, et al.
Published: (2024)
High-Fidelity Document Stain Removal via A Large-Scale Real-World Dataset and A Memory-Augmented Transformer
by: Li, Mingxian, et al.
Published: (2024)
by: Li, Mingxian, et al.
Published: (2024)
SpaceDrive: Infusing Spatial Awareness into VLM-based Autonomous Driving
by: Li, Peizheng, et al.
Published: (2025)
by: Li, Peizheng, et al.
Published: (2025)
StaMo: Unsupervised Learning of Generalizable Robot Motion from Compact State Representation
by: Liu, Mingyu, et al.
Published: (2025)
by: Liu, Mingyu, et al.
Published: (2025)
Visual Point Cloud Forecasting enables Scalable Autonomous Driving
by: Yang, Zetong, et al.
Published: (2023)
by: Yang, Zetong, et al.
Published: (2023)
X-Scene: Large-Scale Driving Scene Generation with High Fidelity and Flexible Controllability
by: Yang, Yu, et al.
Published: (2025)
by: Yang, Yu, et al.
Published: (2025)
VistaGEN: Consistent Driving Video Generation with Fine-Grained Control Using Multiview Visual-Language Reasoning
by: Chen, Li-Heng, et al.
Published: (2026)
by: Chen, Li-Heng, et al.
Published: (2026)
MaskGWM: A Generalizable Driving World Model with Video Mask Reconstruction
by: Ni, Jingcheng, et al.
Published: (2025)
by: Ni, Jingcheng, et al.
Published: (2025)
UniSync: Towards Generalizable and High-Fidelity Lip Synchronization for Challenging Scenarios
by: Fan, Ruidi, et al.
Published: (2026)
by: Fan, Ruidi, et al.
Published: (2026)
VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
by: Li, Yunxin, et al.
Published: (2024)
by: Li, Yunxin, et al.
Published: (2024)
PlanT 2.0: Exposing Biases and Structural Flaws in Closed-Loop Driving
by: Gerstenecker, Simon, et al.
Published: (2025)
by: Gerstenecker, Simon, et al.
Published: (2025)
EvoMakeup: High-Fidelity and Controllable Makeup Editing with MakeupQuad
by: Wu, Huadong, et al.
Published: (2025)
by: Wu, Huadong, et al.
Published: (2025)
Risk-Aware World Model Predictive Control for Generalizable End-to-End Autonomous Driving
by: Sun, Jiangxin, et al.
Published: (2026)
by: Sun, Jiangxin, et al.
Published: (2026)
High-Fidelity and Generalizable Neural Surface Reconstruction with Sparse Feature Volumes
by: Fan, Aoxiang, et al.
Published: (2025)
by: Fan, Aoxiang, et al.
Published: (2025)
World4Drive: End-to-End Autonomous Driving via Intention-aware Physical Latent World Model
by: Zheng, Yupeng, et al.
Published: (2025)
by: Zheng, Yupeng, et al.
Published: (2025)
Decoupled Diffusion Sparks Adaptive Scene Generation
by: Zhou, Yunsong, et al.
Published: (2025)
by: Zhou, Yunsong, et al.
Published: (2025)
HybridWorldSim: A Scalable and Controllable High-fidelity Simulator for Autonomous Driving
by: Li, Qiang, et al.
Published: (2025)
by: Li, Qiang, et al.
Published: (2025)
MTGS: Multi-Traversal Gaussian Splatting
by: Li, Tianyu, et al.
Published: (2025)
by: Li, Tianyu, et al.
Published: (2025)
Depth-Centric Dehazing and Depth-Estimation from Real-World Hazy Driving Video
by: Fan, Junkai, et al.
Published: (2024)
by: Fan, Junkai, et al.
Published: (2024)
Embodied Understanding of Driving Scenarios
by: Zhou, Yunsong, et al.
Published: (2024)
by: Zhou, Yunsong, et al.
Published: (2024)
Similar Items
-
ReSim: Reliable World Simulation for Autonomous Driving
by: Yang, Jiazhi, et al.
Published: (2025) -
GenAD: Generalized Predictive Model for Autonomous Driving
by: Yang, Jiazhi, et al.
Published: (2024) -
SLEDGE: Synthesizing Driving Environments with Generative Models and Rule-Based Traffic
by: Chitta, Kashyap, et al.
Published: (2024) -
End-to-end Autonomous Driving: Challenges and Frontiers
by: Chen, Li, et al.
Published: (2023) -
DriveLM: Driving with Graph Visual Question Answering
by: Sima, Chonghao, et al.
Published: (2023)