Pith. sign in

REVIEW 1 cited by

Exploring the Limits of Deep Image Clustering using Pretrained Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2303.17896 v2 pith:JCCXV46L submitted 2023-03-31 cs.CV cs.AI

classification cs.CVcs.AI
keywords pretrainedclusteringfeatureaccuracyimageimagenetlearnsmodels
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

We present a general methodology that learns to classify images without labels by leveraging pretrained feature extractors. Our approach involves self-distillation training of clustering heads based on the fact that nearest neighbours in the pretrained feature space are likely to share the same label. We propose a novel objective that learns associations between image features by introducing a variant of pointwise mutual information together with instance weighting. We demonstrate that the proposed objective is able to attenuate the effect of false positive pairs while efficiently exploiting the structure in the pretrained feature space. As a result, we improve the clustering accuracy over $k$-means on $17$ different pretrained models by $6.1$\% and $12.2$\% on ImageNet and CIFAR100, respectively. Finally, using self-supervised vision transformers, we achieve a clustering accuracy of $61.6$\% on ImageNet. The code is available at https://github.com/HHU-MMBS/TEMI-official-BMVC2023.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Dimensionally Reduced Open-World Clustering: DROWCULA

    cs.CV 2025-09 conditional novelty 3.0 of 10

    A no-labels pipeline built from known parts (DINOv2, normalization, UMAP/t-SNE, K-means) reports strong clustering and novel-class accuracies on four benchmarks.

Pith tools