Pith. sign in

REVIEW 1 cited by

Predicting Human Similarity Judgments Using Large Language Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2202.04728 v1 pith:S6Y6TFOO submitted 2022-02-09 cs.LG cs.CL

classification cs.LGcs.CL
keywords similarityjudgmentsnumberdescriptionsmodelspredictingstimulidatasets
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Similarity judgments provide a well-established method for accessing mental representations, with applications in psychology, neuroscience and machine learning. However, collecting similarity judgments can be prohibitively expensive for naturalistic datasets as the number of comparisons grows quadratically in the number of stimuli. One way to tackle this problem is to construct approximation procedures that rely on more accessible proxies for predicting similarity. Here we leverage recent advances in language models and online recruitment, proposing an efficient domain-general procedure for predicting human similarity judgments based on text descriptions. Intuitively, similar stimuli are likely to evoke similar descriptions, allowing us to use description similarity to predict pairwise similarity judgments. Crucially, the number of descriptions required grows only linearly with the number of stimuli, drastically reducing the amount of data required. We test this procedure on six datasets of naturalistic images and show that our models outperform previous approaches based on visual information.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. The time course of visuo-semantic representations in the human brain is captured by combining vision and language models

    q-bio.NC 2025-06 conditional novelty 7.0 of 10

    A fusion of vision DNN and LLM representations predicts human EEG responses to images better than either model alone, with distinct temporal and spectral signatures.

Pith tools