StreamMOTP: Streaming and Unified Framework for Joint 3D Multi-Object Tracking and Trajectory Prediction

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Zhuang, Jiaheng, Wang, Guoan, Zhang, Siyu, Wang, Xiyang, Zhou, Hangning, Xu, Ziyao, Zhang, Chi, Li, Zhiheng
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866910506162323456
author Zhuang, Jiaheng
Wang, Guoan
Zhang, Siyu
Wang, Xiyang
Zhou, Hangning
Xu, Ziyao
Zhang, Chi
Li, Zhiheng
author_facet Zhuang, Jiaheng
Wang, Guoan
Zhang, Siyu
Wang, Xiyang
Zhou, Hangning
Xu, Ziyao
Zhang, Chi
Li, Zhiheng
contents 3D multi-object tracking and trajectory prediction are two crucial modules in autonomous driving systems. Generally, the two tasks are handled separately in traditional paradigms and a few methods have started to explore modeling these two tasks in a joint manner recently. However, these approaches suffer from the limitations of single-frame training and inconsistent coordinate representations between tracking and prediction tasks. In this paper, we propose a streaming and unified framework for joint 3D Multi-Object Tracking and trajectory Prediction (StreamMOTP) to address the above challenges. Firstly, we construct the model in a streaming manner and exploit a memory bank to preserve and leverage the long-term latent features for tracked objects more effectively. Secondly, a relative spatio-temporal positional encoding strategy is introduced to bridge the gap of coordinate representations between the two tasks and maintain the pose-invariance for trajectory prediction. Thirdly, we further improve the quality and consistency of predicted trajectories with a dual-stream predictor. We conduct extensive experiments on popular nuSences dataset and the experimental results demonstrate the effectiveness and superiority of StreamMOTP, which outperforms previous methods significantly on both tasks. Furthermore, we also prove that the proposed framework has great potential and advantages in actual applications of autonomous driving.
format Preprint
id arxiv_https___arxiv_org_abs_2406_19844
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle StreamMOTP: Streaming and Unified Framework for Joint 3D Multi-Object Tracking and Trajectory Prediction
Zhuang, Jiaheng
Wang, Guoan
Zhang, Siyu
Wang, Xiyang
Zhou, Hangning
Xu, Ziyao
Zhang, Chi
Li, Zhiheng
Computer Vision and Pattern Recognition
Robotics
3D multi-object tracking and trajectory prediction are two crucial modules in autonomous driving systems. Generally, the two tasks are handled separately in traditional paradigms and a few methods have started to explore modeling these two tasks in a joint manner recently. However, these approaches suffer from the limitations of single-frame training and inconsistent coordinate representations between tracking and prediction tasks. In this paper, we propose a streaming and unified framework for joint 3D Multi-Object Tracking and trajectory Prediction (StreamMOTP) to address the above challenges. Firstly, we construct the model in a streaming manner and exploit a memory bank to preserve and leverage the long-term latent features for tracked objects more effectively. Secondly, a relative spatio-temporal positional encoding strategy is introduced to bridge the gap of coordinate representations between the two tasks and maintain the pose-invariance for trajectory prediction. Thirdly, we further improve the quality and consistency of predicted trajectories with a dual-stream predictor. We conduct extensive experiments on popular nuSences dataset and the experimental results demonstrate the effectiveness and superiority of StreamMOTP, which outperforms previous methods significantly on both tasks. Furthermore, we also prove that the proposed framework has great potential and advantages in actual applications of autonomous driving.
title StreamMOTP: Streaming and Unified Framework for Joint 3D Multi-Object Tracking and Trajectory Prediction
topic Computer Vision and Pattern Recognition
Robotics
url https://arxiv.org/abs/2406.19844