Pith. sign in

REVIEW 1 cited by

Revisiting Active Learning in the Era of Vision Foundation Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2401.14555 v2 pith:QQ3JFVC6 submitted 2024-01-25 cs.CV cs.LG

classification cs.CVcs.LG
keywords foundationmodelsactivelearningdiverseimagesincludingnatural
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Foundation vision or vision-language models are trained on large unlabeled or noisy data and learn robust representations that can achieve impressive zero- or few-shot performance on diverse tasks. Given these properties, they are a natural fit for active learning (AL), which aims to maximize labeling efficiency. However, the full potential of foundation models has not been explored in the context of AL, specifically in the low-budget regime. In this work, we evaluate how foundation models influence three critical components of effective AL, namely, 1) initial labeled pool selection, 2) ensuring diverse sampling, and 3) the trade-off between representative and uncertainty sampling. We systematically study how the robust representations of foundation models (DINOv2, OpenCLIP) challenge existing findings in active learning. Our observations inform the principled construction of a new simple and elegant AL strategy that balances uncertainty estimated via dropout with sample diversity. We extensively test our strategy on many challenging image classification benchmarks, including natural images as well as out-of-domain biomedical images that are relatively understudied in the AL literature. We also provide a highly performant and efficient implementation of modern AL strategies (including our method) at https://github.com/sanketx/AL-foundation-models.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. FunnelAL: Retrieve-then-Rank Active Learning for Single-Class Discovery

    cs.CV 2026-07 conditional novelty 6.0 of 10

    A precision-triggered retrieve-then-rank active-learning pipeline achieves better label efficiency and F1 than recent single-class discovery baselines on three image benchmarks.

Pith tools