The Solution for Single Object Tracking Task of Perception Test Challenge 2024

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Zhong, Zhiqiang, Yang, Yang, Wan, Fengqiang, Wei, Henglu, Ji, Xiangyang
Format: Preprint
Veröffentlicht: 2024
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866916448612384768
author Zhong, Zhiqiang
Yang, Yang
Wan, Fengqiang
Wei, Henglu
Ji, Xiangyang
author_facet Zhong, Zhiqiang
Yang, Yang
Wan, Fengqiang
Wei, Henglu
Ji, Xiangyang
contents This report presents our method for Single Object Tracking (SOT), which aims to track a specified object throughout a video sequence. We employ the LoRAT method. The essence of the work lies in adapting LoRA, a technique that fine-tunes a small subset of model parameters without adding inference latency, to the domain of visual tracking. We train our model using the extensive LaSOT and GOT-10k datasets, which provide a solid foundation for robust performance. Additionally, we implement the alpha-refine technique for post-processing the bounding box outputs. Although the alpha-refine method does not yield the anticipated results, our overall approach achieves a score of 0.813, securing first place in the competition.
format Preprint
id arxiv_https___arxiv_org_abs_2410_16329
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle The Solution for Single Object Tracking Task of Perception Test Challenge 2024
Zhong, Zhiqiang
Yang, Yang
Wan, Fengqiang
Wei, Henglu
Ji, Xiangyang
Computer Vision and Pattern Recognition
This report presents our method for Single Object Tracking (SOT), which aims to track a specified object throughout a video sequence. We employ the LoRAT method. The essence of the work lies in adapting LoRA, a technique that fine-tunes a small subset of model parameters without adding inference latency, to the domain of visual tracking. We train our model using the extensive LaSOT and GOT-10k datasets, which provide a solid foundation for robust performance. Additionally, we implement the alpha-refine technique for post-processing the bounding box outputs. Although the alpha-refine method does not yield the anticipated results, our overall approach achieves a score of 0.813, securing first place in the competition.
title The Solution for Single Object Tracking Task of Perception Test Challenge 2024
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2410.16329