Home › Datasets › task › Visual Object Tracking

Visual Object Tracking datasets

archive 2025-07-28

29 datasets carry the task tag "Visual Object Tracking" (the task itself: Visual Object Tracking), ordered by the archive's paper count. Page 1 of 1: 29 shown of 29. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 51 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets

Visual Object Tracking datasets 1–29 of 29

Object Tracking Benchmark (OTB) is a visual tracking benchmark that is widely used to evaluate the performance of a visual tracking algorithm.
416 papers · 1 benchmark
LaSOT (Large-scale Single Object Tracking)
LaSOT is a high-quality benchmark for Large-scale Single Object Tracking.
275 papers · 3 benchmarks
GOT-10k (Generic Object Tracking Benchmark)
The GOT-10k dataset contains more than 10,000 video segments of real-world moving objects and over 1.5 million manually labelled bounding boxes.
239 papers · 2 benchmarks
TrackingNet is a large-scale tracking dataset consisting of videos in the wild.
210 papers · 2 benchmarks
YouTube-VOS 2018 (Youtube Video Object Segmentation)
Youtube-VOS is a Video Object Segmentation dataset that contains 4,453 videos - 3,471 for training, 474 for validation, and 508 for testing.
203 papers · 10 benchmarks
OTB-2015, also referred as Visual Tracker Benchmark, is a visual tracking dataset.
182 papers · 1 benchmark
VOT2018 is a dataset for visual object tracking.
129 papers · 1 benchmark
VOT2016 is a video dataset for visual object tracking.
113 papers · 1 benchmark
OTB2013 is the previous version of the current OTB2015 Visual Tracker Benchmark.
110 papers · 2 benchmarks
TNL2K (Tracking by natural language)
Tracking by Natural Language (TNL2K) is constructed for the evaluation of tracking by natural language specification.
62 papers · 2 benchmarks
VOT2017 (Visual Object Tracking Challenge)
VOT2017 is a Visual Object Tracking dataset for different tasks that contains 60 short sequences annotated with 6 different attributes.
56 papers · 2 benchmarks
VOTChallenge (Visual Object Tracking)
The Visual Object Tracking (VOT) dataset is a collection of video sequences used for evaluating and benchmarking visual object tracking algorithms.
36 papers · 0 benchmarks
20 papers · 1 benchmark
CDTB (Color-and-Depth Tracking)
Source: https://www.vicos.si/Projects/CDTB 4.2 State-of-the-art Comparison A TH CTB (color-and-depth visual object tracking) dataset is recorded by several passive and active RGB-D setups and contains indoor as well as outdoor sequences…
17 papers · 0 benchmarks
TLP (Track Long and Prosper)
A new long video dataset and benchmark for single object tracking.
14 papers · 0 benchmarks
VOT2014 (Visual Object Tracking Challenge 2014)
The dataset comprises 25 short sequences showing various objects in challenging backgrounds.
12 papers · 1 benchmark
DiDi (Distractor Distilled Dataset)
DiDi is a distractor-distilled tracking dataset created to address the limitation of low distractor presence in current visual object tracking benchmarks.
10 papers · 1 benchmark
VOT2020 is a Visual Object Tracking benchmark for short-term tracking in RGB.
9 papers · 1 benchmark
AVisT (A Benchmark for Visual Object Tracking in Adverse Visibility)
One of the key factors behind the recent success in visual tracking is the availability of dedicated benchmarks.
7 papers · 1 benchmark
TREK-150 is a benchmark dataset for object tracking in First Person Vision (FPV) videos composed of 150 densely annotated video sequences.
7 papers · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
7 papers · 1 benchmark
VideoCube is a high-quality and large-scale benchmark to create a challenging real-world experimental environment for Global Instance Tracking (GIT).
6 papers · 1 benchmark
RF100 (Roboflow 100)
The evaluation of object detection models is usually performed by optimizing a single metric, e.g.
5 papers · 1 benchmark
VOT2019 is a Visual Object Tracking benchmark for short-term tracking in RGB.
4 papers · 1 benchmark
The AU-AIR is a multi-modal aerial dataset captured by a UAV.
2 papers · 0 benchmarks
ITB (Informative Tracking Benchmark)
Informative Tracking Benchmark (ITB) is a small and informative tracking benchmark with 7% out of 1.2 M frames of existing and newly collected datasets, which enables efficient evaluation while ensuring effectiveness.
2 papers · 1 benchmark
RGBD1K (A Large-scale Dataset and Benchmark for RGB-D Object Tracking)
RGBD1K is a benchmark for RGB-D Object Tracking which contains 1050 sequences with about 2.5M frames in total.
2 papers · 0 benchmarks
BioDrone is the first bionic drone-based single object tracking benchmark, it features videos captured from a flapping-wing UAV system with a major camera shake due to its aerodynamics.
1 paper · 0 benchmarks
SOTVerse is a user-defined task space of single object tracking.
1 paper · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.