VIPeR: Visual Incremental Place Recognition with Adaptive Mining and Continual Learning

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Ming, Yuhang, Xu, Minyang, Yang, Xingrui, Ye, Weicai, Wang, Weihan, Peng, Yong, Dai, Weichen, Kong, Wanzeng
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866915147391434752
author Ming, Yuhang
Xu, Minyang
Yang, Xingrui
Ye, Weicai
Wang, Weihan
Peng, Yong
Dai, Weichen
Kong, Wanzeng
author_facet Ming, Yuhang
Xu, Minyang
Yang, Xingrui
Ye, Weicai
Wang, Weihan
Peng, Yong
Dai, Weichen
Kong, Wanzeng
contents Visual place recognition (VPR) is an essential component of many autonomous and augmented/virtual reality systems. It enables the systems to robustly localize themselves in large-scale environments. Existing VPR methods demonstrate attractive performance at the cost of heavy pre-training and limited generalizability. When deployed in unseen environments, these methods exhibit significant performance drops. Targeting this issue, we present VIPeR, a novel approach for visual incremental place recognition with the ability to adapt to new environments while retaining the performance of previous environments. We first introduce an adaptive mining strategy that balances the performance within a single environment and the generalizability across multiple environments. Then, to prevent catastrophic forgetting in lifelong learning, we draw inspiration from human memory systems and design a novel memory bank for our VIPeR. Our memory bank contains a sensory memory, a working memory and a long-term memory, with the first two focusing on the current environment and the last one for all previously visited environments. Additionally, we propose a probabilistic knowledge distillation to explicitly safeguard the previously learned knowledge. We evaluate our proposed VIPeR on three large-scale datasets, namely Oxford Robotcar, Nordland, and TartanAir. For comparison, we first set a baseline performance with naive finetuning. Then, several more recent lifelong learning methods are compared. Our VIPeR achieves better performance in almost all aspects with the biggest improvement of 13.65% in average performance.
format Preprint
id arxiv_https___arxiv_org_abs_2407_21416
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle VIPeR: Visual Incremental Place Recognition with Adaptive Mining and Continual Learning
Ming, Yuhang
Xu, Minyang
Yang, Xingrui
Ye, Weicai
Wang, Weihan
Peng, Yong
Dai, Weichen
Kong, Wanzeng
Computer Vision and Pattern Recognition
Robotics
Visual place recognition (VPR) is an essential component of many autonomous and augmented/virtual reality systems. It enables the systems to robustly localize themselves in large-scale environments. Existing VPR methods demonstrate attractive performance at the cost of heavy pre-training and limited generalizability. When deployed in unseen environments, these methods exhibit significant performance drops. Targeting this issue, we present VIPeR, a novel approach for visual incremental place recognition with the ability to adapt to new environments while retaining the performance of previous environments. We first introduce an adaptive mining strategy that balances the performance within a single environment and the generalizability across multiple environments. Then, to prevent catastrophic forgetting in lifelong learning, we draw inspiration from human memory systems and design a novel memory bank for our VIPeR. Our memory bank contains a sensory memory, a working memory and a long-term memory, with the first two focusing on the current environment and the last one for all previously visited environments. Additionally, we propose a probabilistic knowledge distillation to explicitly safeguard the previously learned knowledge. We evaluate our proposed VIPeR on three large-scale datasets, namely Oxford Robotcar, Nordland, and TartanAir. For comparison, we first set a baseline performance with naive finetuning. Then, several more recent lifelong learning methods are compared. Our VIPeR achieves better performance in almost all aspects with the biggest improvement of 13.65% in average performance.
title VIPeR: Visual Incremental Place Recognition with Adaptive Mining and Continual Learning
topic Computer Vision and Pattern Recognition
Robotics
url https://arxiv.org/abs/2407.21416