ReconDreamer++: Harmonizing Generative and Reconstructive Models for Driving Scene Representation

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Zhao, Guosheng, Wang, Xiaofeng, Ni, Chaojun, Zhu, Zheng, Qin, Wenkang, Huang, Guan, Wang, Xingang
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866912474063699968
author Zhao, Guosheng
Wang, Xiaofeng
Ni, Chaojun
Zhu, Zheng
Qin, Wenkang
Huang, Guan
Wang, Xingang
author_facet Zhao, Guosheng
Wang, Xiaofeng
Ni, Chaojun
Zhu, Zheng
Qin, Wenkang
Huang, Guan
Wang, Xingang
contents Combining reconstruction models with generative models has emerged as a promising paradigm for closed-loop simulation in autonomous driving. For example, ReconDreamer has demonstrated remarkable success in rendering large-scale maneuvers. However, a significant gap remains between the generated data and real-world sensor observations, particularly in terms of fidelity for structured elements, such as the ground surface. To address these challenges, we propose ReconDreamer++, an enhanced framework that significantly improves the overall rendering quality by mitigating the domain gap and refining the representation of the ground surface. Specifically, ReconDreamer++ introduces the Novel Trajectory Deformable Network (NTDNet), which leverages learnable spatial deformation mechanisms to bridge the domain gap between synthesized novel views and original sensor observations. Moreover, for structured elements such as the ground surface, we preserve geometric prior knowledge in 3D Gaussians, and the optimization process focuses on refining appearance attributes while preserving the underlying geometric structure. Experimental evaluations conducted on multiple datasets (Waymo, nuScenes, PandaSet, and EUVS) confirm the superior performance of ReconDreamer++. Specifically, on Waymo, ReconDreamer++ achieves performance comparable to Street Gaussians for the original trajectory while significantly outperforming ReconDreamer on novel trajectories. In particular, it achieves substantial improvements, including a 6.1% increase in NTA-IoU, a 23. 0% improvement in FID, and a remarkable 4.5% gain in the ground surface metric NTL-IoU, highlighting its effectiveness in accurately reconstructing structured elements such as the road surface.
format Preprint
id arxiv_https___arxiv_org_abs_2503_18438
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle ReconDreamer++: Harmonizing Generative and Reconstructive Models for Driving Scene Representation
Zhao, Guosheng
Wang, Xiaofeng
Ni, Chaojun
Zhu, Zheng
Qin, Wenkang
Huang, Guan
Wang, Xingang
Computer Vision and Pattern Recognition
Combining reconstruction models with generative models has emerged as a promising paradigm for closed-loop simulation in autonomous driving. For example, ReconDreamer has demonstrated remarkable success in rendering large-scale maneuvers. However, a significant gap remains between the generated data and real-world sensor observations, particularly in terms of fidelity for structured elements, such as the ground surface. To address these challenges, we propose ReconDreamer++, an enhanced framework that significantly improves the overall rendering quality by mitigating the domain gap and refining the representation of the ground surface. Specifically, ReconDreamer++ introduces the Novel Trajectory Deformable Network (NTDNet), which leverages learnable spatial deformation mechanisms to bridge the domain gap between synthesized novel views and original sensor observations. Moreover, for structured elements such as the ground surface, we preserve geometric prior knowledge in 3D Gaussians, and the optimization process focuses on refining appearance attributes while preserving the underlying geometric structure. Experimental evaluations conducted on multiple datasets (Waymo, nuScenes, PandaSet, and EUVS) confirm the superior performance of ReconDreamer++. Specifically, on Waymo, ReconDreamer++ achieves performance comparable to Street Gaussians for the original trajectory while significantly outperforming ReconDreamer on novel trajectories. In particular, it achieves substantial improvements, including a 6.1% increase in NTA-IoU, a 23. 0% improvement in FID, and a remarkable 4.5% gain in the ground surface metric NTL-IoU, highlighting its effectiveness in accurately reconstructing structured elements such as the road surface.
title ReconDreamer++: Harmonizing Generative and Reconstructive Models for Driving Scene Representation
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2503.18438