TSceneJAL: Joint Active Learning of Traffic Scenes for 3D Object Detection

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Lei, Chenyang, Peng, Weiyuan, Zhou, Guang, Zhang, Meiying, Hao, Qi, Ji, Chunlin, Xu, Chengzhong
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866910914806022144
author Lei, Chenyang
Peng, Weiyuan
Zhou, Guang
Zhang, Meiying
Hao, Qi
Ji, Chunlin
Xu, Chengzhong
author_facet Lei, Chenyang
Peng, Weiyuan
Zhou, Guang
Zhang, Meiying
Hao, Qi
Ji, Chunlin
Xu, Chengzhong
contents Most autonomous driving (AD) datasets incur substantial costs for collection and labeling, inevitably yielding a plethora of low-quality and redundant data instances, thereby compromising performance and efficiency. Many applications in AD systems necessitate high-quality training datasets using both existing datasets and newly collected data. In this paper, we propose a traffic scene joint active learning (TSceneJAL) framework that can efficiently sample the balanced, diverse, and complex traffic scenes from both labeled and unlabeled data. The novelty of this framework is threefold: 1) a scene sampling scheme based on a category entropy, to identify scenes containing multiple object classes, thus mitigating class imbalance for the active learner; 2) a similarity sampling scheme, estimated through the directed graph representation and a marginalize kernel algorithm, to pick sparse and diverse scenes; 3) an uncertainty sampling scheme, predicted by a mixture density network, to select instances with the most unclear or complex regression outcomes for the learner. Finally, the integration of these three schemes in a joint selection strategy yields an optimal and valuable subdataset. Experiments on the KITTI, Lyft, nuScenes and SUScape datasets demonstrate that our approach outperforms existing state-of-the-art methods on 3D object detection tasks with up to 12% improvements.
format Preprint
id arxiv_https___arxiv_org_abs_2412_18870
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle TSceneJAL: Joint Active Learning of Traffic Scenes for 3D Object Detection
Lei, Chenyang
Peng, Weiyuan
Zhou, Guang
Zhang, Meiying
Hao, Qi
Ji, Chunlin
Xu, Chengzhong
Computer Vision and Pattern Recognition
Most autonomous driving (AD) datasets incur substantial costs for collection and labeling, inevitably yielding a plethora of low-quality and redundant data instances, thereby compromising performance and efficiency. Many applications in AD systems necessitate high-quality training datasets using both existing datasets and newly collected data. In this paper, we propose a traffic scene joint active learning (TSceneJAL) framework that can efficiently sample the balanced, diverse, and complex traffic scenes from both labeled and unlabeled data. The novelty of this framework is threefold: 1) a scene sampling scheme based on a category entropy, to identify scenes containing multiple object classes, thus mitigating class imbalance for the active learner; 2) a similarity sampling scheme, estimated through the directed graph representation and a marginalize kernel algorithm, to pick sparse and diverse scenes; 3) an uncertainty sampling scheme, predicted by a mixture density network, to select instances with the most unclear or complex regression outcomes for the learner. Finally, the integration of these three schemes in a joint selection strategy yields an optimal and valuable subdataset. Experiments on the KITTI, Lyft, nuScenes and SUScape datasets demonstrate that our approach outperforms existing state-of-the-art methods on 3D object detection tasks with up to 12% improvements.
title TSceneJAL: Joint Active Learning of Traffic Scenes for 3D Object Detection
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2412.18870