HoloDrive: Holistic 2D-3D Multi-Modal Street Scene Generation for Autonomous Driving
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Zehuan, Ni, Jingcheng, Wang, Xiaodong, Guo, Yuxin, Chen, Rui, Lu, Lewei, Dai, Jifeng, Xiong, Yuwen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MaskGWM: A Generalizable Driving World Model with Video Mask Reconstruction
von: Ni, Jingcheng, et al.
Veröffentlicht: (2025)
von: Ni, Jingcheng, et al.
Veröffentlicht: (2025)
UniMLVG: Unified Framework for Multi-view Long Video Generation with Comprehensive Control Capabilities for Autonomous Driving
von: Chen, Rui, et al.
Veröffentlicht: (2024)
von: Chen, Rui, et al.
Veröffentlicht: (2024)
CVD-STORM: Cross-View Video Diffusion with Spatial-Temporal Reconstruction Model for Autonomous Driving
von: Zhang, Tianrui, et al.
Veröffentlicht: (2025)
von: Zhang, Tianrui, et al.
Veröffentlicht: (2025)
ReinDriveGen: Reinforcement Post-Training for Out-of-Distribution Driving Scene Generation
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
DriveMLM: Aligning Multi-Modal Large Language Models with Behavioral Planning States for Autonomous Driving
von: Cui, Erfei, et al.
Veröffentlicht: (2023)
von: Cui, Erfei, et al.
Veröffentlicht: (2023)
GA-Drive: Geometry-Appearance Decoupled Modeling for Free-viewpoint Driving Scene Generation
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
Multi-Modal Data-Efficient 3D Scene Understanding for Autonomous Driving
von: Kong, Lingdong, et al.
Veröffentlicht: (2024)
von: Kong, Lingdong, et al.
Veröffentlicht: (2024)
Copilot4D: Learning Unsupervised World Models for Autonomous Driving via Discrete Diffusion
von: Zhang, Lunjun, et al.
Veröffentlicht: (2023)
von: Zhang, Lunjun, et al.
Veröffentlicht: (2023)
MTDrive: Multi-turn Interactive Reinforcement Learning for Autonomous Driving
von: Li, Xidong, et al.
Veröffentlicht: (2026)
von: Li, Xidong, et al.
Veröffentlicht: (2026)
Seeing Beyond Views: Multi-View Driving Scene Video Generation with Holistic Attention
von: Lu, Hannan, et al.
Veröffentlicht: (2024)
von: Lu, Hannan, et al.
Veröffentlicht: (2024)
AnyScene: Towards Highly Controllable Driving Scene Generation at Anywhere and Beyond
von: Zhang, Haiming, et al.
Veröffentlicht: (2026)
von: Zhang, Haiming, et al.
Veröffentlicht: (2026)
DreamDrive: Generative 4D Scene Modeling from Street View Images
von: Mao, Jiageng, et al.
Veröffentlicht: (2024)
von: Mao, Jiageng, et al.
Veröffentlicht: (2024)
MagicDrive3D: Controllable 3D Generation for Any-View Rendering in Street Scenes
von: Gao, Ruiyuan, et al.
Veröffentlicht: (2024)
von: Gao, Ruiyuan, et al.
Veröffentlicht: (2024)
ReconDrive: Fast Feed-Forward 4D Gaussian Splatting for Autonomous Driving Scene Reconstruction
von: Yu, Haibao, et al.
Veröffentlicht: (2026)
von: Yu, Haibao, et al.
Veröffentlicht: (2026)
UniScene: Multi-Camera Unified Pre-training via 3D Scene Reconstruction for Autonomous Driving
von: Min, Chen, et al.
Veröffentlicht: (2023)
von: Min, Chen, et al.
Veröffentlicht: (2023)
DriveWorld: 4D Pre-trained Scene Understanding via World Models for Autonomous Driving
von: Min, Chen, et al.
Veröffentlicht: (2024)
von: Min, Chen, et al.
Veröffentlicht: (2024)
Holistic Autonomous Driving Understanding by Bird's-Eye-View Injected Multi-Modal Large Models
von: Ding, Xinpeng, et al.
Veröffentlicht: (2024)
von: Ding, Xinpeng, et al.
Veröffentlicht: (2024)
123D: Unifying Multi-Modal Autonomous Driving Data at Scale
von: Dauner, Daniel, et al.
Veröffentlicht: (2026)
von: Dauner, Daniel, et al.
Veröffentlicht: (2026)
InsightDrive: Insight Scene Representation for End-to-End Autonomous Driving
von: Song, Ruiqi, et al.
Veröffentlicht: (2025)
von: Song, Ruiqi, et al.
Veröffentlicht: (2025)
DriveLiDAR4D: Sequential and Controllable LiDAR Scene Generation for Autonomous Driving
von: Cai, Kaiwen, et al.
Veröffentlicht: (2025)
von: Cai, Kaiwen, et al.
Veröffentlicht: (2025)
OmniScene: Attention-Augmented Multimodal 4D Scene Understanding for Autonomous Driving
von: Liu, Pei, et al.
Veröffentlicht: (2025)
von: Liu, Pei, et al.
Veröffentlicht: (2025)
DriveTok: 3D Driving Scene Tokenization for Unified Multi-View Reconstruction and Understanding
von: Zhuo, Dong, et al.
Veröffentlicht: (2026)
von: Zhuo, Dong, et al.
Veröffentlicht: (2026)
Physics-Aware 3D Gaussian Editing for Driving Scene Generation
von: Zhou, Feng, et al.
Veröffentlicht: (2026)
von: Zhou, Feng, et al.
Veröffentlicht: (2026)
4D Driving Scene Generation With Stereo Forcing
von: Lu, Hao, et al.
Veröffentlicht: (2025)
von: Lu, Hao, et al.
Veröffentlicht: (2025)
HoloDreamer: Holistic 3D Panoramic World Generation from Text Descriptions
von: Zhou, Haiyang, et al.
Veröffentlicht: (2024)
von: Zhou, Haiyang, et al.
Veröffentlicht: (2024)
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency
von: Guo, Xiangyu, et al.
Veröffentlicht: (2025)
von: Guo, Xiangyu, et al.
Veröffentlicht: (2025)
RealDriveSim: A Realistic Multi-Modal Multi-Task Synthetic Dataset for Autonomous Driving
von: Jadon, Arpit, et al.
Veröffentlicht: (2025)
von: Jadon, Arpit, et al.
Veröffentlicht: (2025)
Classification Drives Geographic Bias in Street Scene Segmentation
von: Nair, Rahul, et al.
Veröffentlicht: (2024)
von: Nair, Rahul, et al.
Veröffentlicht: (2024)
DriveGen3D: Boosting Feed-Forward Driving Scene Generation with Efficient Video Diffusion
von: Wang, Weijie, et al.
Veröffentlicht: (2025)
von: Wang, Weijie, et al.
Veröffentlicht: (2025)
WorldSplat: Gaussian-Centric Feed-Forward 4D Scene Generation for Autonomous Driving
von: Zhu, Ziyue, et al.
Veröffentlicht: (2025)
von: Zhu, Ziyue, et al.
Veröffentlicht: (2025)
OG-Gaussian: Occupancy Based Street Gaussians for Autonomous Driving
von: Shen, Yedong, et al.
Veröffentlicht: (2025)
von: Shen, Yedong, et al.
Veröffentlicht: (2025)
ChatDyn: Language-Driven Multi-Actor Dynamics Generation in Street Scenes
von: Wei, Yuxi, et al.
Veröffentlicht: (2024)
von: Wei, Yuxi, et al.
Veröffentlicht: (2024)
GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation
von: Deng, Tianchen, et al.
Veröffentlicht: (2025)
von: Deng, Tianchen, et al.
Veröffentlicht: (2025)
EqDrive: Efficient Equivariant Motion Forecasting with Multi-Modality for Autonomous Driving
von: Wang, Yuping, et al.
Veröffentlicht: (2023)
von: Wang, Yuping, et al.
Veröffentlicht: (2023)
HoloCine: Holistic Generation of Cinematic Multi-Shot Long Video Narratives
von: Meng, Yihao, et al.
Veröffentlicht: (2025)
von: Meng, Yihao, et al.
Veröffentlicht: (2025)
$\textit{S}^3$Gaussian: Self-Supervised Street Gaussians for Autonomous Driving
von: Huang, Nan, et al.
Veröffentlicht: (2024)
von: Huang, Nan, et al.
Veröffentlicht: (2024)
OmniDrive: A Holistic Vision-Language Dataset for Autonomous Driving with Counterfactual Reasoning
von: Wang, Shihao, et al.
Veröffentlicht: (2025)
von: Wang, Shihao, et al.
Veröffentlicht: (2025)
OmniDrive: A Holistic Vision-Language Dataset for Autonomous Driving with Counterfactual Reasoning
von: Wang, Shihao, et al.
Veröffentlicht: (2024)
von: Wang, Shihao, et al.
Veröffentlicht: (2024)
HD$^2$-SSC: High-Dimension High-Density Semantic Scene Completion for Autonomous Driving
von: Yang, Zhiwen, et al.
Veröffentlicht: (2025)
von: Yang, Zhiwen, et al.
Veröffentlicht: (2025)
DME-Driver: Integrating Human Decision Logic and 3D Scene Perception in Autonomous Driving
von: Han, Wencheng, et al.
Veröffentlicht: (2024)
von: Han, Wencheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MaskGWM: A Generalizable Driving World Model with Video Mask Reconstruction
von: Ni, Jingcheng, et al.
Veröffentlicht: (2025) -
UniMLVG: Unified Framework for Multi-view Long Video Generation with Comprehensive Control Capabilities for Autonomous Driving
von: Chen, Rui, et al.
Veröffentlicht: (2024) -
CVD-STORM: Cross-View Video Diffusion with Spatial-Temporal Reconstruction Model for Autonomous Driving
von: Zhang, Tianrui, et al.
Veröffentlicht: (2025) -
ReinDriveGen: Reinforcement Post-Training for Out-of-Distribution Driving Scene Generation
von: Zhang, Hao, et al.
Veröffentlicht: (2026) -
DriveMLM: Aligning Multi-Modal Large Language Models with Behavioral Planning States for Autonomous Driving
von: Cui, Erfei, et al.
Veröffentlicht: (2023)