Pith. sign in

Paper Citation Record · LEDGER

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks

As of 22 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 1 inbound Pith citation observation for arXiv:2412.20682.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.20682 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T23:19:02.251395Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T00:32:20.408599Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-12T00:32:21.103085Z

Reference resolution

60 of 60 outbound references displayed

  • verified exact0
  • verified fuzzy47
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c6080f1f-968c-4258-abc4-92ad33c29ca7 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Learning transferable visual models from natural language supervision,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.853203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.062335Z digest=sha256:2d8867c694e9300c20dd0162f88d9972a665f036c4e80d440e0db57f94a7936c

Observation 24a9c0f8-f2fe-4376-a37e-59c0a3822ab3 · outbound

This paper cites Scaling up visual and vision-language representation learning with noisy text supervision,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Scaling up visual and vision-language representation learning with noisy text supervision,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.844441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.067085Z digest=sha256:7573248d7d6daabaa42e25fc0f895ee70e7cee79224a9ec49b7af8c905e71e73

Observation 1496bf16-a508-4a7d-9e0c-b9a0d8b127c3 · outbound

This paper cites Sigmoid loss for language image pre-training,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Sigmoid loss for language image pre-training,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.836193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.070476Z digest=sha256:a1d9d0b21281d740aa6632168676b314acb62aa950369e6bccb9eeef2aa4663f

Observation 1d7dc89c-8759-4ea7-bd9f-7f6e10e24a54 · outbound

This paper cites Sgva-clip: Semantic- guided visual adapting of vision-language models for few-shot image classification,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Sgva-clip: Semantic- guided visual adapting of vision-language models for few-shot image classification,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.827397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.074512Z digest=sha256:2d038205f83e16688f679f305b007a6fb5ac4555d93cb90fdb201abfcbab510f

Observation f125e2cf-d345-4ff8-8bf6-1d65bf841a07 · outbound

This paper cites Clip-vg: Self-paced curriculum adapting of clip for visual grounding,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Clip-vg: Self-paced curriculum adapting of clip for visual grounding,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.818810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.078024Z digest=sha256:aa419a0f8227da859ec347d5da5e13eeb4fce4f3595457e55a5c8c2758035081

Observation 4f0ff8ac-de93-4d10-bae8-4f677414e50c · outbound

This paper cites Effective end-to-end vision language pre- training with semantic visual loss,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Effective end-to-end vision language pre- training with semantic visual loss,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.809768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.081342Z digest=sha256:81d82872385796184d8f7d17fdf535cf7951b9972674611f58aa4829a7a77a9d

Observation 39e2d1f9-77ae-45d4-8827-e40e13faa139 · outbound

This paper cites Neural logic vision language explainer,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Neural logic vision language explainer,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.799422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.085792Z digest=sha256:38849e2cb56cb2b4935ba1b793ea958b683f6ee25e9ccc92bc1d3fdbcd66603e

Observation 45f4ac16-7638-4079-96fe-a911d77bd1f4 · outbound

This paper cites Lovm: Language- only vision model selection,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Lovm: Language- only vision model selection,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.788305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.088657Z digest=sha256:91d64c0148d8a778660c4141d567447437c9e25012e783cddf678f74b60de28f

Observation ed98bd14-cf8f-4e55-a8c6-7b2f33da4b48 · outbound

This paper cites Bridge the Modality and Capability Gaps in Vision-Language Model Selection.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Bridge the Modality and Capability Gaps in Vision-Language Model Selection

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.092268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.092268Z digest=sha256:63c9ffb83b23a6accc0e17c7135a74727952c42332bedef6133e70f989d217e7

Observation a9663b70-939f-42e3-b691-b47a749e5d47 · outbound

This paper cites Imagenet large scale visual recognition challenge,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Imagenet large scale visual recognition challenge,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.096576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.096576Z digest=sha256:7854b0530c8966489deb48836bb4812bc2ba9a962876d081655dee1a207d8469

Observation 59c80df0-e92e-4a65-bf70-fb67a9d8e0d8 · outbound

This paper cites GPT-4 Technical Report.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks GPT-4 Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.099108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.099108Z digest=sha256:742a816e98146e2aecd34949bd111d370f4deb03d008ab530ee742a9a54bdaf1

Observation a1291799-f858-46f5-b114-950b4b26af82 · outbound

This paper cites Leveraging unlabeled data to predict out-of-distribution performance,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Leveraging unlabeled data to predict out-of-distribution performance,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.776644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.102861Z digest=sha256:39edbd115bfeadfb96cd064cb46cb6ce3f2cc10b0312c3667a0301e6a467f259

Observation a13c2413-be53-44db-b299-acae43f47066 · outbound

This paper cites Are labels always necessary for classifier accuracy evaluation?.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Are labels always necessary for classifier accuracy evaluation?

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.769117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.106942Z digest=sha256:fa986e748364dfc7c6c2832290eff4ecb74796472c6237f7eb6101941a2c1b16

Observation 327ab1b6-d597-49f8-b18e-7b59a2ddeee1 · outbound

This paper cites Predicting out-of- distribution error with the projection norm,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Predicting out-of- distribution error with the projection norm,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.760868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.109667Z digest=sha256:22da991a2c1a30526ad5535f229654442d8a3be9a6730ab3bf119834807c2443

Observation c3384c91-811e-458a-8cee-ce8ae80128a6 · outbound

This paper cites Data determines distributional robustness in contrastive language image pre-training (clip),.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Data determines distributional robustness in contrastive language image pre-training (clip),

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.752385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.112361Z digest=sha256:7c509fb896883bb29a856ec8e87c2651d645e082e2891a8100b362d00d1ca553

Observation 18584f6a-cff5-4044-b263-60498db4077c · outbound

This paper cites Does clip’s generalization performance mainly stem from high train- test similarity?.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Does clip’s generalization performance mainly stem from high train- test similarity?

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.744442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.115218Z digest=sha256:e49536f3c54a351b3fa71039c8f524da766e647ccface3b3b579053bb4c1194d

Observation 3bde1082-2188-4a4f-a509-69890d4088ec · outbound

This paper cites A Survey on Evaluation of Out-of-Distribution Generalization.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks A Survey on Evaluation of Out-of-Distribution Generalization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.117741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.117741Z digest=sha256:d03c01db0c73d2b299810a8281ed1f39202e7dd2620216a695efd8f23f3dab8e

Observation 23247191-a828-4b72-9d61-913dc3761483 · outbound

This paper cites Which Model to Transfer? A Survey on Transferability Estimation.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Which Model to Transfer? A Survey on Transferability Estimation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.120616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.120616Z digest=sha256:dc2c4c2f9a3237c5b88fb91ed4074b0e69660d0079813f9f301842be697284d3

Observation 67e21d47-42b9-47c7-841a-68ede3983428 · outbound

This paper cites Rankme: Assessing the downstream performance of pretrained self-supervised representations by their rank,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Rankme: Assessing the downstream performance of pretrained self-supervised representations by their rank,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.736740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.123697Z digest=sha256:e3dbcab04497c6dd58ef32d44c565f6a2c7ebb40a240b4edd24e1c8ae9fb8aa5

Observation 1a5b9b43-afd2-4f53-83e0-2523ef975313 · outbound

This paper cites Identifying useful learnwares for heterogeneous label spaces,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Identifying useful learnwares for heterogeneous label spaces,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.727669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.126635Z digest=sha256:26cf5aa6815fa46067204c8a17bc49cb43de32c97bd01a4ecfd4b9a9c164c363

Observation bac8c173-40be-4e03-930c-158b51cd1e34 · outbound

This paper cites Etran: Energy-based transferability estimation,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Etran: Energy-based transferability estimation,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.720403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.129072Z digest=sha256:2c6e047b7a134cc4b8b354c7556b842eeb0d53770d7c1be549449459933417e6

Observation 27d233fe-7d78-4b5d-a545-499a6c8ed392 · outbound

This paper cites Predicting out-of-distribution error with confidence optimal transport,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Predicting out-of-distribution error with confidence optimal transport,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.713537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.132874Z digest=sha256:ac55419d0a2791c1042328290baef0b97260cb23efef0ac7c36b9121bd712dd5

Observation e3096a9c-fbdd-48cf-9bb4-7fa431d6cd33 · outbound

This paper cites Data analysis and regression. a second course in statistics,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Data analysis and regression. a second course in statistics,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.706480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.136115Z digest=sha256:d4a7bb33294b5097686191c4cb4d62402b36ceca22deed09e74a73beeafd59b9

Observation 07e1f4b7-5bf3-4985-9975-d8648163fea9 · outbound

This paper cites Tune it the right way: Unsupervised validation of domain adaptation via soft neighborhood density,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Tune it the right way: Unsupervised validation of domain adaptation via soft neighborhood density,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.699122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.138546Z digest=sha256:8201d946e003a3e554ae8ad6f903c55293fa237382037b9c14ec19d2e6838502

Observation 50eff583-92e6-4574-bb0a-40ebbff8ca86 · outbound

This paper cites Covariate shift adap- tation by importance weighted cross validation.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Covariate shift adap- tation by importance weighted cross validation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.689740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.142516Z digest=sha256:5378c6b949ee15bfd07c4beedd8170215ec9aceeba1e72f94073c5a40e30bbc4

Observation 591eb3d8-46f2-42c8-9f10-df2da74637e7 · outbound

This paper cites Towards accurate model selection in deep unsupervised domain adaptation,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Towards accurate model selection in deep unsupervised domain adaptation,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.680320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.146091Z digest=sha256:efd00360d4ff6bc4e36f93208ed624950b27300ba2d5232a0f23dbdf456e79ee

Observation 7b35aab9-69dc-4412-a0f5-3156d5e7e5c2 · outbound

This paper cites Stochastic gradient methods for dis- tributionally robust optimization with f-divergences,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Stochastic gradient methods for dis- tributionally robust optimization with f-divergences,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.672181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.148593Z digest=sha256:1879b7fbbf3b73192279b64528f6015fd6978378f6c9a8ae4f7786be52953088

Observation add8cb75-5990-468f-ae31-0288f3206bb9 · outbound

This paper cites Invariant Risk Minimization.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Invariant Risk Minimization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.151663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.151663Z digest=sha256:1ce7eee097ac92a2b315c9e52e94e441333e4b3c5907ee7334fef4a55c16e61e

Observation 4c547453-dc72-408a-92b8-62b718f6e057 · outbound

This paper cites Stable learning via sample reweighting,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Stable learning via sample reweighting,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.664007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.155701Z digest=sha256:b3904161c178758daed2b8473c72da26d1b6bb379863e44d26ff3a106c6e7fc0

Observation a61de187-c254-4af9-8795-bcda9ebcb386 · outbound

This paper cites A baseline for detecting misclassified and out-of-distribution examples in neural networks,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks A baseline for detecting misclassified and out-of-distribution examples in neural networks,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.655597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.159008Z digest=sha256:6cfef35c4b1df2a8c698983280703f8ca86d89cfd4500e8ea8266b7dc41caca4

Observation b578e508-0b05-421a-8fdb-39e58819caf8 · outbound

This paper cites What does rotation prediction tell us about classifier accuracy under varying testing environments?.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks What does rotation prediction tell us about classifier accuracy under varying testing environments?

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.647830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.161587Z digest=sha256:de914294f43dc41e613cf37df7406303dd886e8b168c9a2b998c1636b19bb966

Observation 42141123-3547-4e66-940d-9e732916013f · outbound

This paper cites Agreement-on-the- line: Predicting the performance of neural networks under distribution shift,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Agreement-on-the- line: Predicting the performance of neural networks under distribution shift,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.637822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.164953Z digest=sha256:fbcd7165f67e88acbee04d29895268f1268d3ac4c5e11b1ea7d847b924f58297

Observation 69d6d889-6245-47ff-84ed-d09684bf3972 · outbound

This paper cites Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.630112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.168153Z digest=sha256:0be63ff779fc8776b0851da3298aa2cd08d9e0ac12608b02010c042c902c25c0

Observation d0258e10-b13b-47bd-aa46-36d5a80fa574 · outbound

This paper cites I. mathematical contributions to the theory of evolu- tion.—vii. on the correlation of characters not quantitatively measur- able,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks I. mathematical contributions to the theory of evolu- tion.—vii. on the correlation of characters not quantitatively measur- able,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.621254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.171155Z digest=sha256:208c99890396bbb02e09af6fdd10aaeb2308ffd38405c985b725ad2ad22c21d8

Observation 3a956975-6cb9-4cde-a6d6-06c8787a739b · outbound

This paper cites Learning multiple layers of features from tiny images,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Learning multiple layers of features from tiny images,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.173527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.173527Z digest=sha256:f7782ab5d1575426b2b77f080365150e6c527cc10707cb16dc97512b4819e6e5

Observation 7101e1b9-9972-4dd6-9fd8-8c9d9db1b391 · outbound

This paper cites Cats and dogs,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Cats and dogs,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.610138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.176231Z digest=sha256:9ef1d3559f6b9a6bb0c20d1195fa611a684897f3b86f7cb2ce60767cc369d931

Observation 2c700529-10b4-4c1e-8269-d91c1adf3467 · outbound

This paper cites Automated flower classification over a large number of classes,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Automated flower classification over a large number of classes,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.603200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.178912Z digest=sha256:4acd9f9409b9be6a428d41a1ea43b846b9e103663c9b2134bbab4b713b10951f

Observation ba33f10b-0d1f-43da-aa42-52ceab64c3e8 · outbound

This paper cites Reading digits in natural images with unsupervised feature learning,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Reading digits in natural images with unsupervised feature learning,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.595212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.181346Z digest=sha256:9bdbcb95bf272fc2fc592969b8a3ccdfdcaefc83cc6192ff16d744ebe8260a02

Observation fc7e523e-c7f9-45ea-ac45-d9a27606758f · outbound

This paper cites Detection of traffic signs in real-world images: The german traffic sign detection benchmark,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Detection of traffic signs in real-world images: The german traffic sign detection benchmark,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.585413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.185130Z digest=sha256:4fe77e3aedf6908893f4b08846b6269288b9694bece5bb19290edeb2a75ecb98

Observation 7ca0654a-f47a-4828-81f5-6b5bfc88c71a · outbound

This paper cites Describing textures in the wild,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Describing textures in the wild,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.576201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.189062Z digest=sha256:cb518985a801610e991c7c6c25a78336494c844f4a71661273cbfdb55ff08e98

Observation e08b47f2-ef10-484a-b07f-6fc1400b7c45 · outbound

This paper cites Yfcc100m: The new data in multimedia research,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Yfcc100m: The new data in multimedia research,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.566206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.192584Z digest=sha256:1a07f6d43a334ebdd9282b6de431acd8d72937b572d16098a06573ddd888584c

Observation 0c4880aa-d471-4885-a240-66142a8a7251 · outbound

This paper cites Sun database: Large-scale scene recognition from abbey to zoo,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Sun database: Large-scale scene recognition from abbey to zoo,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.554864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.198741Z digest=sha256:5b7a8193b2137f3bbfc73210b0bfae2b2d31340bdef5a9bb031d10ae5a1da5db

Observation 2ad95db0-05d0-42b8-b66c-0ea03d373dca · outbound

This paper cites Gradient-based learning applied to document recognition,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Gradient-based learning applied to document recognition,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.544234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.201847Z digest=sha256:9ee2422251b0dd89cbfbf19d32f45d9359c04e4e07e5371b9c591e898abeb9a6

Observation a387ba64-2dae-4eb9-a1ef-9d21e78f5e04 · outbound

This paper cites Challenges in representation learning: Facial expression recognition challenge,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Challenges in representation learning: Facial expression recognition challenge,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.535101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.205503Z digest=sha256:cce8b0ab1b6a4392720667d481e751effb8fe0e01a7208825819f3a38a002fc9

Observation 45258e42-4956-4e3a-aa1a-87627e74055d · outbound

This paper cites On the importance of feature separability in predicting out-of-distribution error,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks On the importance of feature separability in predicting out-of-distribution error,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.527167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.209061Z digest=sha256:4887d1b105a0e05ae9e3978c3cfb02f3f000a3b00b11303ab0cab86f03b03746

Observation 315a85bc-873d-463f-97c0-b37bf5325664 · outbound

This paper cites Unsupervised representation learning by predicting image rotations,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Unsupervised representation learning by predicting image rotations,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.517108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.212042Z digest=sha256:81e8507232d3caf66d6c60d506005076939bbfbac8bf7cfc36fa6416c91b3397

Observation a88f3d46-786e-4e23-acdb-27219d5b6f06 · outbound

This paper cites The use of multiple measurements in taxonomic prob- lems,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks The use of multiple measurements in taxonomic prob- lems,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.508560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.214634Z digest=sha256:acd360e391d5cdba5be22e2b0575aebda013027602736e235c01c034d79438d2

Observation 218e96b0-d58b-4d22-9b7a-d5f41067cfa0 · outbound

This paper cites Silhouettes: a graphical aid to the interpretation and validation of cluster analysis,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Silhouettes: a graphical aid to the interpretation and validation of cluster analysis,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.217025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.217025Z digest=sha256:3ea670d2840debdc249019482286af305e07236d58d711fbc329ed194c008ff9

Observation c39edd1e-2beb-494c-b54b-d9f797f5ae56 · outbound

This paper cites Deep residual learning for image recognition,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Deep residual learning for image recognition,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.219537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.219537Z digest=sha256:4bd20930178de1f0e3e1ced8a14af3da7ba7a9c7b7942035ad55723424b08298

Observation 7995dc37-9bf7-441f-b9fc-c302ac82e98e · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks An image is worth 16x16 words: Transformers for image recognition at scale,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.422639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.222327Z digest=sha256:5bc26fcffe550e51f280d664e262f952f1ce222fa7fe77727ca0895a453ecbd0

Observation 119a2684-2c0c-439d-9f05-922bad6cc5cc · outbound

This paper cites A convnet for the 2020s,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks A convnet for the 2020s,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.382172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.225312Z digest=sha256:acdcaaafe8e28d7fa7af2414a072fec1d2c328a572077caa8498e4f664b1ca01

Observation 675996f6-57fd-4a99-ae23-b26f0b48b90a · outbound

This paper cites Laion-400m: Open dataset of clip-filtered 400 million image-text pairs,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Laion-400m: Open dataset of clip-filtered 400 million image-text pairs,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.372675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.227883Z digest=sha256:678c87e23f2e7304f224bd39a25127a27765d1f9a05a46a71844979481dd865a

Observation 82cc4b91-9e7e-4ad3-8f3c-c26e6c5d8102 · outbound

This paper cites AltCLIP: Altering the language encoder in CLIP for extended language capabilities,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks AltCLIP: Altering the language encoder in CLIP for extended language capabilities,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.362356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.230309Z digest=sha256:9cc02a23328e6a5e04197062566b778673007873aa81c27d363632ca7ce27c3c

Observation d7237b80-b802-4761-859c-448fcacc02af · outbound

This paper cites Groupvit: Semantic segmentation emerges from text supervision,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Groupvit: Semantic segmentation emerges from text supervision,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.350740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.232910Z digest=sha256:bc2117d56125d83851f685bdd96b8251e4842de59c001fa5d6f86188b0da24b4

Observation a3b48329-45f1-4ab2-856a-bec9f04e4f0e · outbound

This paper cites Learning Generalized Zero-Shot Learners for Open-Domain Image Geolocalization.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Learning Generalized Zero-Shot Learners for Open-Domain Image Geolocalization

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.235754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.235754Z digest=sha256:b85a4515bbb0e3459895352d8d1cc135008ad804dfcf695e51742359520ade94

Observation b63e8fcc-ea48-4330-8d0f-23b2ae3a5fef · outbound

This paper cites Demystifying CLIP Data.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Demystifying CLIP Data

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.239160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.239160Z digest=sha256:8be417d7b50b4d9d1f370e34d8b10d489659ed8e304d3dc3ab00b5e17720ed25

Observation 8829e10d-c0a3-4c87-a026-2ede33fedede · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.242350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.242350Z digest=sha256:b9e96390dce32ba2b399708e7f7afa61ed967bb76ec6c8c674681c85215920cf

Observation 60cb4d52-1c1d-4b93-8581-ff0c846425f1 · outbound

This paper cites Quilt-1M: One Million Image-Text Pairs for Histopathology.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Quilt-1M: One Million Image-Text Pairs for Histopathology

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T23:19:02.245892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:19:02.245892Z digest=sha256:a309434b0eeaeab319df02b03a195cf005b5c866c7b87c4fb2ddbd5b90468a5f

Observation 8e2650a0-8233-468f-8cc3-09fc29526389 · outbound

This paper cites BioCLIP: A vision foundation model for the tree of life,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks BioCLIP: A vision foundation model for the tree of life,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.341422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.248744Z digest=sha256:de15a254c9a4b1cd6e582a1e9a3e355d5f31d6317e38733f901ac74b90065f17

Observation cb36a3e1-a99f-4dcc-8527-2511724e5102 · outbound

This paper cites Gpt-4: Generative pre-trained transformer,.

Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks Gpt-4: Generative pre-trained transformer,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T23:19:02.331857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-10T23:19:02.251395Z digest=sha256:5580f641917d81eabea0d8d1eb5d3c1557f9bc274bd939462f1f3c1602832ece

Pith citing papers

Observation 12f4859a-834e-4817-a2cf-bfb3385c5b4e · inbound

HugSelect: An Explainable Multi-Criteria Decision-Support Framework for foundation-model selection cites this paper.

HugSelect: An Explainable Multi-Criteria Decision-Support Framework for foundation-model selection Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-08-12T00:32:21.110763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T00:32:20.408599Z digest=sha256:8054517fd9b87cb55dfa4ae3f5be7b89e0c78d6dcd88bf94ee6372b203796e17