Pith. sign in

REVIEW 2 cited by

Towards the Generalization of Contrastive Self-Supervised Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2111.00743 v4 pith:6B6RKBZP submitted 2021-11-01 cs.LG cs.AIcs.CVstat.ML

classification cs.LGcs.AIcs.CVstat.ML
keywords datalearningcontrastiveself-supervisedfactorsgeneralizationabilityaugmentation
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

Recently, self-supervised learning has attracted great attention, since it only requires unlabeled data for model training. Contrastive learning is one popular method for self-supervised learning and has achieved promising empirical performance. However, the theoretical understanding of its generalization ability is still limited. To this end, we define a kind of $(\sigma,\delta)$-measure to mathematically quantify the data augmentation, and then provide an upper bound of the downstream classification error rate based on the measure. It reveals that the generalization ability of contrastive self-supervised learning is related to three key factors: alignment of positive samples, divergence of class centers, and concentration of augmented data. The first two factors are properties of learned representations, while the third one is determined by pre-defined data augmentation. We further investigate two canonical contrastive losses, InfoNCE and cross-correlation, to show how they provably achieve the first two factors. Moreover, we conduct experiments to study the third factor, and observe a strong correlation between downstream performance and the concentration of augmented data.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. OpenAlex reports about 21 citations worldwide. Full citation record

  1. What Has Been Overlooked in Contrastive Source-Free Domain Adaptation: Leveraging Source-Informed Latent Augmentation within Neighborhood Context

    cs.CV 2024-12 conditional novelty 6.0 of 10

    SiLAN augments target-neighborhood centroids with Gaussian noise whose variance comes from the frozen source model's neighbor dispersion, improving contrastive SFDA accuracy on three benchmarks.

  2. Where Did Your Model Learn That? Label-free Influence for Self-supervised Learning

    cs.LG 2024-12 conditional novelty 5.0 of 10

    Influence-SSL defines influence as the sensitivity of a sample's representation to augmentation, and shows the resulting scores can identify duplicates, outliers, and fairness-relevant examples in SSL models.

Pith tools