Computer Science > Computer Vision and Pattern Recognition

arXiv:2208.09787 (cs)

[Submitted on 21 Aug 2022 (v1), last revised 30 Dec 2022 (this version, v3)]

Title:RGBD1K: A Large-scale Dataset and Benchmark for RGB-D Object Tracking

Authors:Xue-Feng Zhu, Tianyang Xu, Zhangyong Tang, Zucheng Wu, Haodong Liu, Xiao Yang, Xiao-Jun Wu, Josef Kittler

View PDF

Abstract:RGB-D object tracking has attracted considerable attention recently, achieving promising performance thanks to the symbiosis between visual and depth channels. However, given a limited amount of annotated RGB-D tracking data, most state-of-the-art RGB-D trackers are simple extensions of high-performance RGB-only trackers, without fully exploiting the underlying potential of the depth channel in the offline training stage. To address the dataset deficiency issue, a new RGB-D dataset named RGBD1K is released in this paper. The RGBD1K contains 1,050 sequences with about 2.5M frames in total. To demonstrate the benefits of training on a larger RGB-D data set in general, and RGBD1K in particular, we develop a transformer-based RGB-D tracker, named SPT, as a baseline for future visual object tracking studies using the new dataset. The results, of extensive experiments using the SPT tracker emonstrate the potential of the RGBD1K dataset to improve the performance of RGB-D tracking, inspiring future developments of effective tracker designs. The dataset and codes will be available on the project homepage: this https URL.

Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2208.09787 [cs.CV]
	(or arXiv:2208.09787v3 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2208.09787

Submission history

From: Xue-Feng Zhu [view email]
[v1] Sun, 21 Aug 2022 03:07:36 UTC (6,715 KB)
[v2] Tue, 13 Dec 2022 10:30:06 UTC (6,715 KB)
[v3] Fri, 30 Dec 2022 23:23:37 UTC (6,716 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:RGBD1K: A Large-scale Dataset and Benchmark for RGB-D Object Tracking

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:RGBD1K: A Large-scale Dataset and Benchmark for RGB-D Object Tracking

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators