Not All Voxels Are Equal: Hardness-Aware Semantic Scene Completion with Self-Distillation

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Wang, Song, Yu, Jiawei, Li, Wentong, Liu, Wenyu, Liu, Xiaolu, Chen, Junbo, Zhu, Jianke
Natura: Preprint
Pubblicazione: 2024
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866909173940224000
author Wang, Song
Yu, Jiawei
Li, Wentong
Liu, Wenyu
Liu, Xiaolu
Chen, Junbo
Zhu, Jianke
author_facet Wang, Song
Yu, Jiawei
Li, Wentong
Liu, Wenyu
Liu, Xiaolu
Chen, Junbo
Zhu, Jianke
contents Semantic scene completion, also known as semantic occupancy prediction, can provide dense geometric and semantic information for autonomous vehicles, which attracts the increasing attention of both academia and industry. Unfortunately, existing methods usually formulate this task as a voxel-wise classification problem and treat each voxel equally in 3D space during training. As the hard voxels have not been paid enough attention, the performance in some challenging regions is limited. The 3D dense space typically contains a large number of empty voxels, which are easy to learn but require amounts of computation due to handling all the voxels uniformly for the existing models. Furthermore, the voxels in the boundary region are more challenging to differentiate than those in the interior. In this paper, we propose HASSC approach to train the semantic scene completion model with hardness-aware design. The global hardness from the network optimization process is defined for dynamical hard voxel selection. Then, the local hardness with geometric anisotropy is adopted for voxel-wise refinement. Besides, self-distillation strategy is introduced to make training process stable and consistent. Extensive experiments show that our HASSC scheme can effectively promote the accuracy of the baseline model without incurring the extra inference cost. Source code is available at: https://github.com/songw-zju/HASSC.
format Preprint
id arxiv_https___arxiv_org_abs_2404_11958
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Not All Voxels Are Equal: Hardness-Aware Semantic Scene Completion with Self-Distillation
Wang, Song
Yu, Jiawei
Li, Wentong
Liu, Wenyu
Liu, Xiaolu
Chen, Junbo
Zhu, Jianke
Computer Vision and Pattern Recognition
Robotics
Semantic scene completion, also known as semantic occupancy prediction, can provide dense geometric and semantic information for autonomous vehicles, which attracts the increasing attention of both academia and industry. Unfortunately, existing methods usually formulate this task as a voxel-wise classification problem and treat each voxel equally in 3D space during training. As the hard voxels have not been paid enough attention, the performance in some challenging regions is limited. The 3D dense space typically contains a large number of empty voxels, which are easy to learn but require amounts of computation due to handling all the voxels uniformly for the existing models. Furthermore, the voxels in the boundary region are more challenging to differentiate than those in the interior. In this paper, we propose HASSC approach to train the semantic scene completion model with hardness-aware design. The global hardness from the network optimization process is defined for dynamical hard voxel selection. Then, the local hardness with geometric anisotropy is adopted for voxel-wise refinement. Besides, self-distillation strategy is introduced to make training process stable and consistent. Extensive experiments show that our HASSC scheme can effectively promote the accuracy of the baseline model without incurring the extra inference cost. Source code is available at: https://github.com/songw-zju/HASSC.
title Not All Voxels Are Equal: Hardness-Aware Semantic Scene Completion with Self-Distillation
topic Computer Vision and Pattern Recognition
Robotics
url https://arxiv.org/abs/2404.11958