REVIEW 4 cited by
Know Your Self-supervised Learning: A Survey on Image-based Generative and Discriminative Training
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
abstract
Although supervised learning has been highly successful in improving the state-of-the-art in the domain of image-based computer vision in the past, the margin of improvement has diminished significantly in recent years, indicating that a plateau is in sight. Meanwhile, the use of self-supervised learning (SSL) for the purpose of natural language processing (NLP) has seen tremendous successes during the past couple of years, with this new learning paradigm yielding powerful language models. Inspired by the excellent results obtained in the field of NLP, self-supervised methods that rely on clustering, contrastive learning, distillation, and information-maximization, which all fall under the banner of discriminative SSL, have experienced a swift uptake in the area of computer vision. Shortly afterwards, generative SSL frameworks that are mostly based on masked image modeling, complemented and surpassed the results obtained with discriminative SSL. Consequently, within a span of three years, over $100$ unique general-purpose frameworks for generative and discriminative SSL, with a focus on imaging, were proposed. In this survey, we review a plethora of research efforts conducted on image-oriented SSL, providing a historic view and paying attention to best practices as well as useful software packages. While doing so, we discuss pretext tasks for image-based SSL, as well as techniques that are commonly used in image-based SSL. Lastly, to aid researchers who aim at contributing to image-focused SSL, we outline a number of promising research directions.
Forward citations
Cited by 4 Pith papers
-
Color Flow Imaging Microscopy Improves Identification of Stress Sources of Protein Aggregates in Biopharmaceuticals
Deep learning trained on color flow imaging microscopy images identifies stress sources of protein aggregates more accurately than on grayscale images.
-
An efficient unsupervised classification model for galaxy morphology: Voting clustering based on coding from ConvNeXt large model
An unsupervised pipeline using ConvNeXt encoding, PCA, and multi-model voting classifies about 53% of COSMOS galaxies into 20 clusters, later merged into five morphology types.
-
Identifying Critical Tokens for Accurate Predictions in Transformer-based Medical Imaging Models
Token Insight iteratively removes the image token that most reduces a vision transformer's polyp-class confidence until the prediction flips, exposing the patches that drive the decision.
-
Handwritten Text Recognition: A Survey
A survey of handwritten text recognition that categorizes methods by reading-order complexity and compares reported performance on the IAM database.
Discussion (0). Continue with ORCID to comment.