REVIEW 3 cited by
Self-EMD: Self-Supervised Object Detection without ImageNet
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
In this paper, we propose a novel self-supervised representation learning method, Self-EMD, for object detection. Our method directly trained on unlabeled non-iconic image dataset like COCO, instead of commonly used iconic-object image dataset like ImageNet. We keep the convolutional feature maps as the image embedding to preserve spatial structures and adopt Earth Mover's Distance (EMD) to compute the similarity between two embeddings. Our Faster R-CNN (ResNet50-FPN) baseline achieves 39.8% mAP on COCO, which is on par with the state of the art self-supervised methods pre-trained on ImageNet. More importantly, it can be further improved to 40.4% mAP with more unlabeled images, showing its great potential for leveraging more easily obtained unlabeled data. Code will be made available.
Forward citations
Cited by 3 Pith papers
-
Multiple Object Stitching for Unsupervised Representation Learning
Multiple Object Stitching improves self-supervised representations by training on stitched multi-object images and reports gains on ImageNet, CIFAR, and COCO.
-
Unlocking the Potential of Weakly Labeled Data: A Co-Evolutionary Learning Framework for Abnormality Detection and Report Generation
A co-evolutionary framework that alternates between chest X-ray abnormality detection and radiology report generation, using each task to refine the other's pseudo-labels, achieves state-of-the-art results on MS-CXR a...
-
Long-Tailed Object Detection Pre-training: Dynamic Rebalancing Contrastive Learning with Dual Reconstruction
A detection pre-training framework that dynamically rebalances rare classes and adds dual reconstruction improves tail-class AP on COCO and LVIS by about 0.5 to 1.5 points in most configurations.
Discussion (0). Continue with ORCID to comment.