AdaOcc: Adaptive-Resolution Occupancy Prediction

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Chen, Chao, Wang, Ruoyu, Guo, Yuliang, Zhao, Cheng, Huang, Xinyu, Feng, Chen, Ren, Liu
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866914922339762176
author Chen, Chao
Wang, Ruoyu
Guo, Yuliang
Zhao, Cheng
Huang, Xinyu
Feng, Chen
Ren, Liu
author_facet Chen, Chao
Wang, Ruoyu
Guo, Yuliang
Zhao, Cheng
Huang, Xinyu
Feng, Chen
Ren, Liu
contents Autonomous driving in complex urban scenarios requires 3D perception to be both comprehensive and precise. Traditional 3D perception methods focus on object detection, resulting in sparse representations that lack environmental detail. Recent approaches estimate 3D occupancy around vehicles for a more comprehensive scene representation. However, dense 3D occupancy prediction increases computational demands, challenging the balance between efficiency and resolution. High-resolution occupancy grids offer accuracy but demand substantial computational resources, while low-resolution grids are efficient but lack detail. To address this dilemma, we introduce AdaOcc, a novel adaptive-resolution, multi-modal prediction approach. Our method integrates object-centric 3D reconstruction and holistic occupancy prediction within a single framework, performing highly detailed and precise 3D reconstruction only in regions of interest (ROIs). These high-detailed 3D surfaces are represented in point clouds, thus their precision is not constrained by the predefined grid resolution of the occupancy map. We conducted comprehensive experiments on the nuScenes dataset, demonstrating significant improvements over existing methods. In close-range scenarios, we surpass previous baselines by over 13% in IOU, and over 40% in Hausdorff distance. In summary, AdaOcc offers a more versatile and effective framework for delivering accurate 3D semantic occupancy prediction across diverse driving scenarios.
format Preprint
id arxiv_https___arxiv_org_abs_2408_13454
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle AdaOcc: Adaptive-Resolution Occupancy Prediction
Chen, Chao
Wang, Ruoyu
Guo, Yuliang
Zhao, Cheng
Huang, Xinyu
Feng, Chen
Ren, Liu
Computer Vision and Pattern Recognition
Autonomous driving in complex urban scenarios requires 3D perception to be both comprehensive and precise. Traditional 3D perception methods focus on object detection, resulting in sparse representations that lack environmental detail. Recent approaches estimate 3D occupancy around vehicles for a more comprehensive scene representation. However, dense 3D occupancy prediction increases computational demands, challenging the balance between efficiency and resolution. High-resolution occupancy grids offer accuracy but demand substantial computational resources, while low-resolution grids are efficient but lack detail. To address this dilemma, we introduce AdaOcc, a novel adaptive-resolution, multi-modal prediction approach. Our method integrates object-centric 3D reconstruction and holistic occupancy prediction within a single framework, performing highly detailed and precise 3D reconstruction only in regions of interest (ROIs). These high-detailed 3D surfaces are represented in point clouds, thus their precision is not constrained by the predefined grid resolution of the occupancy map. We conducted comprehensive experiments on the nuScenes dataset, demonstrating significant improvements over existing methods. In close-range scenarios, we surpass previous baselines by over 13% in IOU, and over 40% in Hausdorff distance. In summary, AdaOcc offers a more versatile and effective framework for delivering accurate 3D semantic occupancy prediction across diverse driving scenarios.
title AdaOcc: Adaptive-Resolution Occupancy Prediction
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2408.13454