Pith. sign in

REVIEW 1 cited by

All models are wrong, some are useful: Model Selection with Limited Labels

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2410.13609 v2 pith:CYB35DIJ submitted 2024-10-17 cs.LG

classification cs.LG
keywords modelselectorbestcostlabelingpretrainedreducesselection
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We introduce MODEL SELECTOR, a framework for label-efficient selection of pretrained classifiers. Given a pool of unlabeled target data, MODEL SELECTOR samples a small subset of highly informative examples for labeling, in order to efficiently identify the best pretrained model for deployment on this target dataset. Through extensive experiments, we demonstrate that MODEL SELECTOR drastically reduces the need for labeled data while consistently picking the best or near-best performing model. Across 18 model collections on 16 different datasets, comprising over 1,500 pretrained models, MODEL SELECTOR reduces the labeling cost by up to 94.15% to identify the best model compared to the cost of the strongest baseline. Our results further highlight the robustness of MODEL SELECTOR in model selection, as it reduces the labeling cost by up to 72.41% when selecting a near-best model, whose accuracy is only within 1% of the best model.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Consensus-Driven Active Model Selection

    cs.LG 2025-07 conditional novelty 7.0 of 10

    CODA uses consensus-based priors and Bayesian updating to select the best candidate model with far fewer labels than prior active model selection methods, beating them on 18 of 26 benchmark tasks.

Pith tools