BFANet: Revisiting 3D Semantic Segmentation with Boundary Feature Analysis

Fuente: arXiv
Guardado en:
Detalles Bibliográficos
Autores principales: Zhao, Weiguang, Zhang, Rui, Wang, Qiufeng, Cheng, Guangliang, Huang, Kaizhu
Formato: Preprint
Publicado: 2025
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866929761801994240
author Zhao, Weiguang
Zhang, Rui
Wang, Qiufeng
Cheng, Guangliang
Huang, Kaizhu
author_facet Zhao, Weiguang
Zhang, Rui
Wang, Qiufeng
Cheng, Guangliang
Huang, Kaizhu
contents 3D semantic segmentation plays a fundamental and crucial role to understand 3D scenes. While contemporary state-of-the-art techniques predominantly concentrate on elevating the overall performance of 3D semantic segmentation based on general metrics (e.g. mIoU, mAcc, and oAcc), they unfortunately leave the exploration of challenging regions for segmentation mostly neglected. In this paper, we revisit 3D semantic segmentation through a more granular lens, shedding light on subtle complexities that are typically overshadowed by broader performance metrics. Concretely, we have delineated 3D semantic segmentation errors into four comprehensive categories as well as corresponding evaluation metrics tailored to each. Building upon this categorical framework, we introduce an innovative 3D semantic segmentation network called BFANet that incorporates detailed analysis of semantic boundary features. First, we design the boundary-semantic module to decouple point cloud features into semantic and boundary features, and fuse their query queue to enhance semantic features with attention. Second, we introduce a more concise and accelerated boundary pseudo-label calculation algorithm, which is 3.9 times faster than the state-of-the-art, offering compatibility with data augmentation and enabling efficient computation in training. Extensive experiments on benchmark data indicate the superiority of our BFANet model, confirming the significance of emphasizing the four uniquely designed metrics. Code is available at https://github.com/weiguangzhao/BFANet.
format Preprint
id arxiv_https___arxiv_org_abs_2503_12539
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle BFANet: Revisiting 3D Semantic Segmentation with Boundary Feature Analysis
Zhao, Weiguang
Zhang, Rui
Wang, Qiufeng
Cheng, Guangliang
Huang, Kaizhu
Computer Vision and Pattern Recognition
3D semantic segmentation plays a fundamental and crucial role to understand 3D scenes. While contemporary state-of-the-art techniques predominantly concentrate on elevating the overall performance of 3D semantic segmentation based on general metrics (e.g. mIoU, mAcc, and oAcc), they unfortunately leave the exploration of challenging regions for segmentation mostly neglected. In this paper, we revisit 3D semantic segmentation through a more granular lens, shedding light on subtle complexities that are typically overshadowed by broader performance metrics. Concretely, we have delineated 3D semantic segmentation errors into four comprehensive categories as well as corresponding evaluation metrics tailored to each. Building upon this categorical framework, we introduce an innovative 3D semantic segmentation network called BFANet that incorporates detailed analysis of semantic boundary features. First, we design the boundary-semantic module to decouple point cloud features into semantic and boundary features, and fuse their query queue to enhance semantic features with attention. Second, we introduce a more concise and accelerated boundary pseudo-label calculation algorithm, which is 3.9 times faster than the state-of-the-art, offering compatibility with data augmentation and enabling efficient computation in training. Extensive experiments on benchmark data indicate the superiority of our BFANet model, confirming the significance of emphasizing the four uniquely designed metrics. Code is available at https://github.com/weiguangzhao/BFANet.
title BFANet: Revisiting 3D Semantic Segmentation with Boundary Feature Analysis
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2503.12539