REVIEW 4 cited by
GOT-10k: A Large High-Diversity Benchmark for Generic Object Tracking in the Wild
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
We introduce here a large tracking database that offers an unprecedentedly wide coverage of common moving objects in the wild, called GOT-10k. Specifically, GOT-10k is built upon the backbone of WordNet structure and it populates the majority of over 560 classes of moving objects and 87 motion patterns, magnitudes wider than the most recent similar-scale counterparts. The contributions of this paper are summarized in the following: (1) GOT-10k offers over 10,000 video segments with more than 1.5 million manually labeled bounding boxes, enabling unified training and stable evaluation of deep trackers. (2) GOT-10k is by far the first video trajectory dataset that uses the semantic hierarchy of WordNet to guide class population. (3) For the first time, GOT-10k introduces the one-shot protocol for tracker evaluation, where the training and test classes are zero-overlapped. The protocol avoids biased evaluation results towards familiar objects and it promotes generalization in tracker development. (4) We conduct extensive tracking experiments with 39 typical tracking algorithms on GOT-10k and analyze their results in this paper. (5) Finally, we develop a comprehensive platform for the tracking community that offers full-featured evaluation toolkits, an online evaluation server, and a responsive leaderboard. The annotations of GOT-10k's test data are kept private to avoid tuning parameters on it. The database, toolkits, evaluation server and baseline results are available at http://got-10k.aitestunion.com.
Forward citations
Cited by 4 Pith papers
-
Multi-Modal Fusion for End-to-End RGB-T Tracking
A feature-level fusion of RGB and thermal features, trained end-to-end on synthetic paired data, beats the DiMP baseline by 6.4% EAO and sets new state-of-the-art results on VOT-RGBT2019 and RGBT210.
-
Effects of Blur and Deblurring to Visual Object Tracking
The paper builds a synthetic motion-blur tracking benchmark, finds light blur helps and heavy blur hurts most trackers, and proposes fine-tuning a DeblurGAN discriminator to selectively deblur frames, improving six of...
-
RBCN: Rectified Binary Convolutional Networks for Enhancing the Performance of 1-bit DCNNs
A GAN-guided rectified training scheme narrows the accuracy gap between binary and full-precision convolutional networks on classification and tracking.
-
DomainSiam: Domain-Aware Siamese Network for Visual Object Tracking
DomainSiam couples a Siamese tracker with channel selection and a variant of Barron's robust loss, but its main SOTA claim is undercut by its own VOT2017 numbers.
Discussion (0). Continue with ORCID to comment.