Pith. sign in

Paper Citation Record · LEDGER

Adapting Vision-Language Models Without Labels: A Comprehensive Survey

As of 18 August 2026, this Paper Citation Record lists 100 of 299 outbound references and 4 inbound Pith citation observations for arXiv:2508.05547.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.05547 v1

Coverage vector

measured 100 of 299 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T23:17:06.729087Z

measured 104 of 104 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-11T23:10:08.033785Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T12:54:40.377437Z

Reference resolution

100 of 299 outbound references displayed

  • verified exact5
  • verified fuzzy0
  • unresolved95
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c7b6bc77-500a-472d-865f-2f872bda4b73 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Learning transferable visual models from natural language supervision,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:03.411463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:03.411463Z digest=sha256:baf660ea919407ded2f5c6fa885498ab257a16cccd9dd3a36b5cc710854bad16

Observation b31b2cf6-d9fc-4828-b62b-048077f931a8 · outbound

This paper cites Scaling up visual and vision-language representation learning with noisy text supervision,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Scaling up visual and vision-language representation learning with noisy text supervision,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:03.440245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:03.440245Z digest=sha256:20c7461ef5cfee5d182681cf00340cfb3835607bc96b746867fbb82a82ef373a

Observation c25732c9-0c30-4640-bae1-4ef40b5f49f0 · outbound

This paper cites Flamingo: a visual language model for few-shot learning,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Flamingo: a visual language model for few-shot learning,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:03.506315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:03.506315Z digest=sha256:0d008eff5301b86a66c951a19879019266792f600d5cc7a1f70adc696522d699

Observation a6d4c706-f83b-4876-9841-aa92642aeb6c · outbound

This paper cites Visual instruction tuning,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Visual instruction tuning,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:03.573874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:03.573874Z digest=sha256:3e21ae1c6a1565b52d7a90345a1ce06e9bed7a11b185932a8ad96bfc94dd2b99

Observation 871f2069-6a30-41ea-8e04-b01f027af56a · outbound

This paper cites Laion-5b: An open large-scale dataset for training next generation image-text models,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Laion-5b: An open large-scale dataset for training next generation image-text models,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:03.616036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:03.616036Z digest=sha256:309437e503dcbec27829b28348d52239a8c0081014c99597ef0dd8caae656c43

Observation f656fd03-df30-4533-8938-90be9d1ffa37 · outbound

This paper cites Clip2scene: Towards label-efficient 3d scene understanding by clip,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Clip2scene: Towards label-efficient 3d scene understanding by clip,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:03.674684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:03.674684Z digest=sha256:e9f76376123a56ac18be52ce4dd24baf320f50e92738acc85ce3ede4aefbba6d

Observation 3beeec1f-bdac-4ce2-877c-6bee9c981e19 · outbound

This paper cites CLIP-Fields: Weakly Supervised Semantic Fields for Robotic Memory.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey CLIP-Fields: Weakly Supervised Semantic Fields for Robotic Memory

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:03.779796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:03.779796Z digest=sha256:528cf0301bf37eebd71a65ae7090befa2a71eb838149fc5ab0e971b0ccba1f7e

Observation a2af32f6-8260-4f30-8822-38d51e270b87 · outbound

This paper cites Unseen visual anomaly generation,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Unseen visual anomaly generation,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:03.831119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:03.831119Z digest=sha256:4b769a63909cd55ff524f40cdec29bc621bb7b07722f7672e04067dac571256b

Observation fd09b1c1-6566-4185-b79b-c82440621154 · outbound

This paper cites Probabilistic embeddings for cross-modal retrieval,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Probabilistic embeddings for cross-modal retrieval,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:03.910821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:03.910821Z digest=sha256:cf3fcbf3186e151aa1b5d7e526611b4c11d2ad547a90d52a43ab9e28b9dc782a

Observation ba6eecd3-805d-46a9-87f9-170a1bbb4d0b · outbound

This paper cites Learning to prompt for vision-language models,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Learning to prompt for vision-language models,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:03.951281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:03.951281Z digest=sha256:d9bbe1aa4f89e3c728108653da771be073a4311897988691e71959c08911e5c3

Observation a204632c-839e-4b13-b0f0-d6ad72687abb · outbound

This paper cites Conditional prompt learning for vision-language models,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Conditional prompt learning for vision-language models,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:04.008476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:04.008476Z digest=sha256:de55f03a1ff680f4c83a8f3940baf7d01a57af947a1e33d48ddc0e8281bc1ce7

Observation dc17dc4d-4f79-4b3b-be30-068c5e9da5a6 · outbound

This paper cites MaPLe: Multi-modal Prompt Learning.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey MaPLe: Multi-modal Prompt Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:04.069590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:04.069590Z digest=sha256:e2ae4dec2c8e68c23055a06f8a31ce60fa6999337bbf01a694361f43121cdfa7

Observation 37b296ab-d061-46b0-af81-05101b73d4aa · outbound

This paper cites A hard-to- beat baseline for training-free clip-based adaptation,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey A hard-to- beat baseline for training-free clip-based adaptation,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:04.125513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:04.125513Z digest=sha256:fb7cc5a283c0608f022df4b363aa9296d758218fd8e979321baf67b771344c2b

Observation c7d5e9f4-b8b5-48a1-b6f1-03b734620888 · outbound

This paper cites Adapting visual category models to new domains,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Adapting visual category models to new domains,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:04.180514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:04.180514Z digest=sha256:3e1e0315db3e58048be2a70e1a96fc938fcdeffb4f826de2849acf70b523cc04

Observation ae56fbcf-d5fb-4397-b0b8-c5b88ba57c63 · outbound

This paper cites Visual classification via description from large language models,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Visual classification via description from large language models,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:04.240023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:04.240023Z digest=sha256:d6793280d107e873502c406b58173ade7d49e4da54f3fa13b48217ba7e7848d5

Observation 43569026-8929-4db7-a916-50f67bb0c716 · outbound

This paper cites Extract free dense labels from clip,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Extract free dense labels from clip,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:04.287319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:04.287319Z digest=sha256:242e45f2263007b1a2960b62eca20895f11611981f8171273aeedbd5c5e52ea1

Observation c3942aff-a4f6-4c99-945b-09a255845dd3 · outbound

This paper cites Unsupervised Prompt Learning for Vision-Language Models.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Unsupervised Prompt Learning for Vision-Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:04.343681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:04.343681Z digest=sha256:40e6f4d4c3522f5c68a9da61c2ac80a2893755693ab2813e73340fe7af7667ed

Observation 68b5ba6b-c2c9-496f-b81f-0d203c35a7ca · outbound

This paper cites Test-time prompt tuning for zero-shot generalization in vision-language models,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Test-time prompt tuning for zero-shot generalization in vision-language models,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:04.396702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:04.396702Z digest=sha256:738b89f6add27594abf9b2975414998320cc3e4016205ee0c0a5cddb037ae2da

Observation 788dd14e-fd13-4fbc-ad9d-ba36c89c061c · outbound

This paper cites Swapprompt: Test-time prompt adaptation for vision-language models,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Swapprompt: Test-time prompt adaptation for vision-language models,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:04.466211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:04.466211Z digest=sha256:282cc27338dbd17b7acb18a408ccba7748d6b719ed279face40308e1e26f8037

Observation c3c3c35f-78cd-46df-b823-32edca0210c0 · outbound

This paper cites The illusion of progress? a critical look at test-time adaptation for vision-language models,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey The illusion of progress? a critical look at test-time adaptation for vision-language models,

Reference 20

Resolution
verified exact
raw_fallback, observed 2026-08-05T23:17:15.006412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T23:17:04.506791Z digest=sha256:8c4f427bcd074d651e2237891d8eeba4340ab2e5459b021487da93a9eb1a70be

Observation e124257a-b3a7-41f6-8f27-cb4a0b3476b3 · outbound

This paper cites What does a platypus look like? generating customized prompts for zero-shot image classi- fication,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey What does a platypus look like? generating customized prompts for zero-shot image classi- fication,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:04.542044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:04.542044Z digest=sha256:378a6b8763b4e7895aba15770e915abc7c202730a81c2c6b35388407f978737d

Observation 7fca3fa5-e0cd-4f0b-a89b-0fdc766993b2 · outbound

This paper cites Improving Zero-Shot Models with Label Distribution Priors.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Improving Zero-Shot Models with Label Distribution Priors

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-05T23:17:14.928712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T23:17:04.589997Z digest=sha256:742b404c39fd6740f031a040425ad64a906378d5b79791955e2c1ae9711dc556

Observation 447f5e3e-7f2c-4d92-aa47-da94c1dec0c3 · outbound

This paper cites Online zero-shot classification with clip,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Online zero-shot classification with clip,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:04.649384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:04.649384Z digest=sha256:9f96e02156b176f476f6f6ae1d59f18b279a01a83946ad4c37a415593c4ee39b

Observation 0d32d8bf-997f-4b63-b1ed-f92733283f9a · outbound

This paper cites Align your prompts: Test- time prompting with distribution alignment for zero-shot generaliza- tion,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Align your prompts: Test- time prompting with distribution alignment for zero-shot generaliza- tion,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:04.682395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:04.682395Z digest=sha256:3226aff997719a9c83111402615cf8a747711bfd112a2806dbe48afe814a470b

Observation 40ceada1-b4e4-47a8-9dcc-76ad2c947cd3 · outbound

This paper cites Efficient test-time adaptation of vision-language models,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Efficient test-time adaptation of vision-language models,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:04.745293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:04.745293Z digest=sha256:a00286ad96146f80be7aa449ff17287ad1866a1f2fde36643bb8815d0873fc12

Observation d26b76ca-92e6-460b-a75d-33fff16c2494 · outbound

This paper cites Masked unsupervised self-training for label-free image classification,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Masked unsupervised self-training for label-free image classification,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:04.829357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:04.829357Z digest=sha256:b6dd40708fbb126dac6429802c8a25b6c0ef68fc05eadbcd1beeea7d1d31918f

Observation 2588556e-e9e0-440e-b2f1-903b9e85aba7 · outbound

This paper cites Realistic unsupervised clip fine-tuning with universal entropy optimization,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Realistic unsupervised clip fine-tuning with universal entropy optimization,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:04.889601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:04.889601Z digest=sha256:a264644ddbdde421c9c7688d1a70f43547a9aa93a160756216e1865536f369a4

Observation e42f5036-091d-4e86-9a67-2fb5ac166d2d · outbound

This paper cites Reco: Retrieve and co-segment for zero-shot transfer,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Reco: Retrieve and co-segment for zero-shot transfer,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:04.931657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:04.931657Z digest=sha256:5b4b2ee9b8afc51baeff08bde393b778b3ee529efd2da11c9c230a37c505994b

Observation ea382039-b3dc-40c6-96ab-0c3079b25544 · outbound

This paper cites Pay attention to your neighbours: Training-free open-vocabulary semantic segmentation,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Pay attention to your neighbours: Training-free open-vocabulary semantic segmentation,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:04.977067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:04.977067Z digest=sha256:ef47e3e69ae813e8f7196383758ba90227abb8eeff51973cd5ce533bd7682f46

Observation 9386ebcb-8a65-4608-9c31-ec0d63031710 · outbound

This paper cites Can language-guided unsupervised adaptation improve medical image classification using unpaired images and texts?.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Can language-guided unsupervised adaptation improve medical image classification using unpaired images and texts?

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:05.046092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:05.046092Z digest=sha256:8a115ffa56582ddd8fbe19c9fc77f3cad59a1d66008eca654f99a18a1eade870

Observation 0b96e069-5ce9-46fb-ba23-618196242cbd · outbound

This paper cites A chatgpt aided explainable framework for zero-shot medical image diagnosis,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey A chatgpt aided explainable framework for zero-shot medical image diagnosis,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:05.070471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:05.070471Z digest=sha256:05c8d78c37fcda6a2529f4ccc396b8506f42bbf460c96e41e70ca2c50b8cadac

Observation e432cc70-4a70-4a78-834c-ee7a7e2fb61c · outbound

This paper cites Text-enhanced zero-shot action recognition: A training-free approach,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Text-enhanced zero-shot action recognition: A training-free approach,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:05.138131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:05.138131Z digest=sha256:1a4a34a353da09eec67f5309223cc0ff236f8f390f4072f1f7b6d0867fd8bf72

Observation efb7bc96-b047-4ae8-ab73-df044ce405da · outbound

This paper cites Dts-tpt: dual temporal-sync test-time prompt tuning for zero-shot activity recogni- tion,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Dts-tpt: dual temporal-sync test-time prompt tuning for zero-shot activity recogni- tion,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:05.189370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:05.189370Z digest=sha256:63d5c61023474bbf167c104d867746ecbf04bfd33e572d0f7a312ac4955211ca

Observation 2d5e7e2c-9e54-4fd9-9876-a1e3ec263850 · outbound

This paper cites Pouf: Prompt- oriented unsupervised fine-tuning for large pre-trained models,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Pouf: Prompt- oriented unsupervised fine-tuning for large pre-trained models,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:05.253496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:05.253496Z digest=sha256:2e075b3ff6b14e3e767e8fc1f28ed4894d11952379f8a825e5464305593b8b9c

Observation a4d550c2-8289-4794-a8c9-5bea88691405 · outbound

This paper cites Label propagation for zero-shot classification with vision-language models,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Label propagation for zero-shot classification with vision-language models,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:05.308737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:05.308737Z digest=sha256:ea12e4e3cf51cce53bed9650d6d9f096a81c024dfc2f98e752698a0070eec147

Observation 2bc59056-7b73-454a-8dd2-bed00cc3b9be · outbound

This paper cites On the test-time zero-shot generalization of vision-language models: Do we really need prompt learning?.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey On the test-time zero-shot generalization of vision-language models: Do we really need prompt learning?

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:05.371056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:05.371056Z digest=sha256:29b31ee368eeb237d5cb0224a682b36fafb1e62f382c14d8c68f54c6640e2665

Observation ea13eed2-730c-4441-91de-ef40aadca1f7 · outbound

This paper cites Vision-language models for vision tasks: A survey,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Vision-language models for vision tasks: A survey,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:05.461471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:05.461471Z digest=sha256:d0584f293d8ac6e37e80a03cd7536adc2b28913edd021ff79fae1892b36ff146

Observation e5790b02-7f2d-4dad-a3f2-33261cc94dfd · outbound

This paper cites Advances in multimodal adaptation and generalization: From traditional approaches to foundation models,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Advances in multimodal adaptation and generalization: From traditional approaches to foundation models,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:05.538275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:05.538275Z digest=sha256:4f135d4f18c315b03ef5cf585ca562d7da884d7d3eecab0ae6f9a0f49a739aed

Observation f8531e7b-6bdf-4510-b747-0e5a4a7f87ed · outbound

This paper cites Generalizing vision-language models to novel domains: A comprehensive survey.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Generalizing vision-language models to novel domains: A comprehensive survey

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:05.585412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:05.585412Z digest=sha256:b03a3c7ce58078ff6a81b94616ef658ab5668a3c79f48075cf39d87900baa0ae

Observation bf69ec72-773f-48b1-8cca-2da2df8b104c · outbound

This paper cites A comprehensive survey on test-time adaptation under distribution shifts,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey A comprehensive survey on test-time adaptation under distribution shifts,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:05.671452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:05.671452Z digest=sha256:2149b606d917c101942db8f035413e1d8a4210b9d5fa305cb24be357c74c4140

Observation 842cf4bc-c21d-4531-8f59-f062c69b0885 · outbound

This paper cites In search of lost online test-time adaptation: A survey,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey In search of lost online test-time adaptation: A survey,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:05.713046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:05.713046Z digest=sha256:98d0b460dd6b02f04f043c1c85a34178e49c791087fa5db2589989e81bb04397

Observation d0d3fa61-1766-4e2b-9cae-d689705ac2fb · outbound

This paper cites Beyond Model Adaptation at Test Time: A Survey.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Beyond Model Adaptation at Test Time: A Survey

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:05.771665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:05.771665Z digest=sha256:3db7ac644e27288763c5252b300acdb424adab98cf9da6d017e9ac63145449fd

Observation e716f882-24dc-4ea4-b2f4-4fc56b7832f5 · outbound

This paper cites FILIP: Fine-grained Interactive Language-Image Pre-Training.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey FILIP: Fine-grained Interactive Language-Image Pre-Training

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:05.859329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:05.859329Z digest=sha256:2db25bf7eef29e5bf96c00c087c32d70a7e2f06d528579225f57efbc4b9789f0

Observation c30dc607-dab7-4d22-bae1-125ce99ae1eb · outbound

This paper cites Attention is all you need,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Attention is all you need,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:05.910520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:05.910520Z digest=sha256:f9aafb18d626f1d8b666e5abcb14f5f93c17808f204250a403a0b9dacc6c1d2a

Observation 200d380b-0fc0-4d23-91a2-4c94b932b669 · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey An image is worth 16x16 words: Transformers for image recognition at scale,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:05.951097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:05.951097Z digest=sha256:a9bfafbe4017f27fc6d4ad2f99ebc24218f1f8893078752c069a6267248426f1

Observation aaf5043a-1208-4a05-9244-460064c6cb40 · outbound

This paper cites Microsoft COCO Captions: Data Collection and Evaluation Server.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Microsoft COCO Captions: Data Collection and Evaluation Server

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.035332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.035332Z digest=sha256:7fa44875e3d35e0a01ea2baafae1fd2fbbc9273064a440e0b8d9c1390e51f9e7

Observation 57e4bcf7-a40f-4541-86cc-45e50f583cc9 · outbound

This paper cites Image captioning with semantic attention,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Image captioning with semantic attention,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.095913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.095913Z digest=sha256:4ef004df71656d64626b31ae11d76cf02be2c82acc16c1f88fe1cb335ce9760b

Observation ee405704-3202-4be4-a1ab-e3073d8f9d58 · outbound

This paper cites Visual question answering: A survey of methods and datasets,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Visual question answering: A survey of methods and datasets,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.198250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.198250Z digest=sha256:720161399b4278d072d16ae9ec0ec70e31489a1278563e5ed418cf158b63780b

Observation b11b603b-f8e3-497e-be26-5e4aa4056122 · outbound

This paper cites High- resolution image synthesis with latent diffusion models,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey High- resolution image synthesis with latent diffusion models,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.276256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.276256Z digest=sha256:d0283dd9b9c676803059c9895957bdfaf8274d480436f73f287c7a092e8260c7

Observation 02359158-71fa-4a62-9dd2-f003c7892612 · outbound

This paper cites Deep supervised cross-modal retrieval,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Deep supervised cross-modal retrieval,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.317978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.317978Z digest=sha256:aea25d6a815eceadece31e844685d05a757b65d456c722e3fb6f5119d6185b36

Observation 83f1a346-24ac-4557-93a1-f3d22b65a286 · outbound

This paper cites A comprehensive survey on pretrained foundation models: A history from bert to chatgpt,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey A comprehensive survey on pretrained foundation models: A history from bert to chatgpt,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.384111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.384111Z digest=sha256:da70032000e93947c2261a5420aaa6791fe387712a21e0517a90563c7c8fce52

Observation 61ef6ba4-54e1-479b-b3c1-e5a66aef93b7 · outbound

This paper cites Learning to detect unseen object classes by between-class attribute transfer,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Learning to detect unseen object classes by between-class attribute transfer,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.445225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.445225Z digest=sha256:e23803fb898ebc6d0ba7a570294c11178c49fa55cb5fd4b1cf21ec1860099abf

Observation c38ba418-de6d-4f35-9acc-7332b1c1c721 · outbound

This paper cites Describing objects by their attributes,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Describing objects by their attributes,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.511582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.511582Z digest=sha256:5aba9954620c4e73569e0cfbb573faa55e3faabcfed3ddc2feb851965f3be738

Observation 7df03e6e-1b56-43ad-8a7d-7a52637ed111 · outbound

This paper cites Devise: A deep visual-semantic embedding model,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Devise: A deep visual-semantic embedding model,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.519380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.519380Z digest=sha256:de95aeac2827846e2929da66c319a551446aff8cb71dc7a27e04a2be445ee6d1

Observation 3536aa89-4f5e-4bed-9844-69467245b3d1 · outbound

This paper cites An embarrassingly simple approach to zero-shot learning,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey An embarrassingly simple approach to zero-shot learning,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.524000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.524000Z digest=sha256:26438f1cd505aae1184de5f7ae4153011baf0b2cf1bdd7a6f5d698c0358929e3

Observation bd00ec66-9ca8-4ca5-bea0-5d1ad47cabce · outbound

This paper cites Generalized zero-shot learning via synthesized examples,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Generalized zero-shot learning via synthesized examples,

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.528338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.528338Z digest=sha256:f48f180075ad96e22184ce2584ccc51271a103d60cfacb00df900a09d8333a61

Observation c293621c-f38d-4d18-af71-7364ada6ef61 · outbound

This paper cites Multi-modal cycle-consistent generalized zero-shot learning,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Multi-modal cycle-consistent generalized zero-shot learning,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.532571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.532571Z digest=sha256:525ef08df1a82c1953728c85e5c3ffab749d98cd7ed8aa4f10d06578bf91d1b8

Observation fd4c9846-139c-474a-beec-29711a63aee4 · outbound

This paper cites f-vaegan-d2: A feature generating framework for any-shot learning,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey f-vaegan-d2: A feature generating framework for any-shot learning,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.537074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.537074Z digest=sha256:8176e1cb2d99f172f6056ef46c8261829ab3781738d33ecd50f5e96e684a4543

Observation c72e388f-fee6-4710-90e6-e6bbd9715452 · outbound

This paper cites An empirical study and analysis of generalized zero-shot learning for object recognition in the wild,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey An empirical study and analysis of generalized zero-shot learning for object recognition in the wild,

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.541736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.541736Z digest=sha256:5f0c928eb0998f853fce09f50ea12333e18e4650a120f12add7a8f57916c58bf

Observation e9ef69b8-6ff8-43b3-8a42-d537f92582ce · outbound

This paper cites Zero-shot learning-the good, the bad and the ugly,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Zero-shot learning-the good, the bad and the ugly,

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.546148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.546148Z digest=sha256:fba88f6137451cbf833959520fe3eb4ec49aa1600ac06fd02572cdd92e5de322

Observation 0e9509dd-9864-47a8-aba1-19cbb38e52f9 · outbound

This paper cites A review of generalized zero-shot learning meth- ods,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey A review of generalized zero-shot learning meth- ods,

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.550354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.550354Z digest=sha256:680c1b951b2210f16693c2034992dd517163f56589e823c8a04e7943a8c892e0

Observation 9f0d1a5f-f402-4753-9031-61e150c492d7 · outbound

This paper cites A survey of zero-shot learning: Settings, methods, and applications,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey A survey of zero-shot learning: Settings, methods, and applications,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.554621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.554621Z digest=sha256:ad4ada6ecc4dca8e9bd3bc6f6dd63089f8e8763ca31fa44fee77f48bc110942f

Observation cedb5ca5-4f96-49a2-800f-3bbf385f578d · outbound

This paper cites Understanding and Mitigating Overfitting in Prompt Tuning for Vision-Language Models.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Understanding and Mitigating Overfitting in Prompt Tuning for Vision-Language Models

Reference 63

Resolution
verified exact
local_arxiv, observed 2026-08-05T23:17:14.764669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T23:17:06.559119Z digest=sha256:f37f25e2f00af08eebbc1473ec394b1c35887467def7655c589be3abc603afea

Observation 295c011c-078c-4ccc-a1bf-3ebd20648b9e · outbound

This paper cites Clip-adapter: Better vision-language models with feature adapters,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Clip-adapter: Better vision-language models with feature adapters,

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.563682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.563682Z digest=sha256:bf5f10f9286470d42b7ac0875e35fae5e633f2c6a000f1c9c9b4d03fc536aaee

Observation 6bdd2d81-edc0-4840-8603-264b06a8fbad · outbound

This paper cites Tip-Adapter: Training-free CLIP-Adapter for Better Vision-Language Modeling.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Tip-Adapter: Training-free CLIP-Adapter for Better Vision-Language Modeling

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.568562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.568562Z digest=sha256:357007f2d241c11868633340ddeaaa8cd1050f9ab5b1d949bd06826a4e1a4652

Observation 508ca37d-ec6c-41f9-9955-d8a8c376ed91 · outbound

This paper cites Low-rank few-shot adaptation of vision- language models,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Low-rank few-shot adaptation of vision- language models,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.572790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.572790Z digest=sha256:9a1bd4d558f15cd180b2a9906b81c4a4af8c08d1d77485f7ca056989e4ad31ad

Observation be2961b0-4a51-4b47-995f-fc2953daa493 · outbound

This paper cites Visual Classification via Description from Large Language Models.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Visual Classification via Description from Large Language Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.577174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.577174Z digest=sha256:d1906669ee0ce79e8cf0f3e142eba5132a9f8f82ad54bc01c146580ff6978306

Observation 647103fe-dd31-47ff-ae31-feed2951415b · outbound

This paper cites Denseclip: Language-guided dense prediction with context- aware prompting,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Denseclip: Language-guided dense prediction with context- aware prompting,

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.581490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.581490Z digest=sha256:1a7ba2ed15287f1bc8bd2877df33ad17c2fe68b2e00bedc975f484f6d808ba1e

Observation a0d3e291-6774-4c4b-b33c-5f3ddc703e08 · outbound

This paper cites A Systematic Survey of Prompt Engineering on Vision-Language Foundation Models.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey A Systematic Survey of Prompt Engineering on Vision-Language Foundation Models

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.585528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.585528Z digest=sha256:d9073fe3088954b0b44f13885dc035c5470347ca8b2f09dc7d1191636c10c279

Observation fb8af6a8-7d5e-4f5e-a6b1-9d43331e4572 · outbound

This paper cites Recall and Refine: A Simple but Effective Source-free Open-set Domain Adaptation Framework.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Recall and Refine: A Simple but Effective Source-free Open-set Domain Adaptation Framework

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-08-05T23:17:14.695358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T23:17:06.589956Z digest=sha256:57bf7cb3d9e14efe6f50956f59eed2a20f2690be42279ad10b3823c38e0f1439

Observation 19322e0b-ce33-4da2-b424-77e88498a2e5 · outbound

This paper cites Do we really need to access the source data? source hypothesis transfer for unsupervised domain adaptation,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Do we really need to access the source data? source hypothesis transfer for unsupervised domain adaptation,

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.594580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.594580Z digest=sha256:bf86dc6d1328b0c3b152cbf09f8b388b0c32daa68c65f71757a76dea82361cb5

Observation 8e7bf716-5549-4aad-9d98-634e953e60d6 · outbound

This paper cites Exploiting local feature patterns for unsupervised domain adaptation,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Exploiting local feature patterns for unsupervised domain adaptation,

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.598787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.598787Z digest=sha256:fdbe100117997173c9f6d90ec42967ea7be3b42322e201d411a6ef0f29f67c12

Observation ce03de49-3ecc-4add-90dd-634f432ef179 · outbound

This paper cites Contrastive test-time adaptation,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Contrastive test-time adaptation,

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.602801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.602801Z digest=sha256:36595ed1ff0343fcfcee7ae13bbbf13c2c31aae0791c95e5e65fa5c4aac11040

Observation d8d98dbb-c695-44bf-afd4-c9a410ea73b1 · outbound

This paper cites Universal source-free domain adaptation,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Universal source-free domain adaptation,

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.607476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.607476Z digest=sha256:9b616e7546d75b09f77bb7d42e5ae3e1f24894c7c4d0741bb5291c2b8f4546a6

Observation 338b0a1c-c833-4f03-aea6-808e47f1a14f · outbound

This paper cites Model adaptation: Historical contrastive learning for unsupervised domain adaptation without source data,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Model adaptation: Historical contrastive learning for unsupervised domain adaptation without source data,

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.611750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.611750Z digest=sha256:31d092abaa79106c2ac2c11686d388b9f9e9952524b24befbafe67863fc7904e

Observation 1861f6fd-2daa-4cf2-9940-09ceecbc5bc9 · outbound

This paper cites Domain adaptation without source data,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Domain adaptation without source data,

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.615871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.615871Z digest=sha256:e179e056388b0da72a27a5765e6856a4eaa9921c1771a90cf2aceec89f67e2c8

Observation 7fa2327c-4e39-4d95-85d2-6db255df0849 · outbound

This paper cites A comprehensive survey on source-free domain adaptation,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey A comprehensive survey on source-free domain adaptation,

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.620242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.620242Z digest=sha256:3225afa881062326afb813c5314785420f229509533fe487675f9f94865df9c4

Observation f4ce1167-0449-48b6-9528-9dba3aff35d9 · outbound

This paper cites Source-free unsu- pervised domain adaptation: A survey,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Source-free unsu- pervised domain adaptation: A survey,

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.624666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.624666Z digest=sha256:f84aab31eaccb7fcc261f40f75714dab24b5e5f3aab47d553e51fb27f7d48aee

Observation 5b0de065-0440-4bfc-9d68-f2fc31c4619d · outbound

This paper cites Tent: Fully test-time adaptation by entropy minimization,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Tent: Fully test-time adaptation by entropy minimization,

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.628996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.628996Z digest=sha256:d2d77ebab5503a179d6c52ad83608d920746b27db1e7a18283f0a272ced5fc3b

Observation df703f37-0672-45f7-9e34-c28f0e8d0c7a · outbound

This paper cites Towards robust multimodal open-set test-time adaptation via adaptive entropy-aware optimization,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Towards robust multimodal open-set test-time adaptation via adaptive entropy-aware optimization,

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.633652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.633652Z digest=sha256:093c5a22df26bd3599afb498a42ab01602a90cf0db3e7c49d8a1df383d8d694c

Observation 76497343-3cc5-4a9a-be0f-0ba739fcdaae · outbound

This paper cites Efficient test-time model adaptation without forgetting,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Efficient test-time model adaptation without forgetting,

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.638132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.638132Z digest=sha256:6e17378c714c81599894f854b74e82873a26c5133665b587027f8e69e74e1b3a

Observation dc774672-c694-45f1-a379-0a41fb77b784 · outbound

This paper cites Sotta: Robust test-time adaptation on noisy data streams,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Sotta: Robust test-time adaptation on noisy data streams,

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.642480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.642480Z digest=sha256:9339c3d8dca11902c6a7a710c049ffbba4a4c2f3e9f93c436853ff074f218fd8

Observation 62cbaaf8-0c64-44be-8665-0a97165148f2 · outbound

This paper cites Continual test-time domain adaptation,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Continual test-time domain adaptation,

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.648686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.648686Z digest=sha256:7f118c0bddff62a2d310462d9fc2f4d6ceff06efa5719538fb9fa6251f687bd1

Observation dac2b887-aefc-4861-b6b5-6dda7765fbf6 · outbound

This paper cites Ecotta: Memory- efficient continual test-time adaptation via self-distilled regularization,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Ecotta: Memory- efficient continual test-time adaptation via self-distilled regularization,

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.653156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.653156Z digest=sha256:542a00767635290684ae7fe5f630db8f393e99fdf129fe386c7bc2b7bbd341ce

Observation e4377699-4fc5-4b85-8662-e509ee2fe5ae · outbound

This paper cites Diverse data augmentation with diffusions for effective test-time prompt tuning,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Diverse data augmentation with diffusions for effective test-time prompt tuning,

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.657503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.657503Z digest=sha256:ea0890f2f71ed64d4964f5649a3b089e7820a79ec7355a410eea621ca09d1f66

Observation 4b059712-fd22-4477-b17a-1ad35fee2618 · outbound

This paper cites EVA-CLIP: Improved Training Techniques for CLIP at Scale.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey EVA-CLIP: Improved Training Techniques for CLIP at Scale

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.662323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.662323Z digest=sha256:bd935f90fc8b9d7a7f878d2456a702b759af03236ac01b8d9c64a9377a6ba5e8

Observation fde89f92-ca02-403b-80ba-a5ea5b6555bb · outbound

This paper cites Sigmoid loss for language image pre-training,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Sigmoid loss for language image pre-training,

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.667461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.667461Z digest=sha256:1af73d0415128262b11252f6eddadab5847675f20dc6b02b21f5a0481d0e5728

Observation 1763b2da-8de3-4af7-83ce-b084a89c1661 · outbound

This paper cites Chils: Zero-shot image classification with hierarchical label sets,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Chils: Zero-shot image classification with hierarchical label sets,

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.671557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.671557Z digest=sha256:78295506a32d6745a1ee422c1edab157bac12b39eacad12ccddba359daf009f4

Observation fdef4187-3409-4253-91db-20dcf330a566 · outbound

This paper cites Sus-x: Training-free name- only transfer of vision-language models,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Sus-x: Training-free name- only transfer of vision-language models,

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.680502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.680502Z digest=sha256:c204d9cc392210a3cf9bf84612bb470dd47ee753cf8ec6a6f6c98555a43496d7

Observation fe24334b-0686-42d4-b91d-db55f481f4fc · outbound

This paper cites Neural priming for sample- efficient adaptation,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Neural priming for sample- efficient adaptation,

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.685106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.685106Z digest=sha256:297f90d685a95d9a44647e6744e02ad66ad2a1fc9c67e09a8e6548796ac5d5bc

Observation d0e8afd2-32b2-41d1-9647-5c5193546abc · outbound

This paper cites Just say the name: Online continual learning with category names only via data generation,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Just say the name: Online continual learning with category names only via data generation,

Reference 92

Resolution
verified exact
raw_fallback, observed 2026-08-05T23:17:14.657902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T23:17:06.689789Z digest=sha256:30d285f37f7ebf92319bc66ca47ddc47fc5ecb68e5013fb6d5766e68d1cca76c

Observation 794fb8eb-5dba-4078-a90a-b01a70774e3d · outbound

This paper cites Calip: Zero-shot enhancement of clip with parameter-free attention,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Calip: Zero-shot enhancement of clip with parameter-free attention,

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.693916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.693916Z digest=sha256:8d5a33ed55e086238893e05a408ff705b73dfe3b0085200b82bb7f83e027970a

Observation 520db2c9-ed3c-4e35-baec-e294df52c37f · outbound

This paper cites Sclip: Rethinking self-attention for dense vision-language inference,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Sclip: Rethinking self-attention for dense vision-language inference,

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.698630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.698630Z digest=sha256:879c24c38fb41a9e15e35ca69be8749dd18de220091f2d8197075afdf87d7ce2

Observation dd531c91-a9a6-473a-8f67-3bc93551ea42 · outbound

This paper cites Proxyclip: Proxy attention improves clip for open-vocabulary segmentation,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Proxyclip: Proxy attention improves clip for open-vocabulary segmentation,

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.702964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.702964Z digest=sha256:e5be0baa2b2a63dbb2077e747e9744848c13948365a13fb8580246b0441e099d

Observation 6909462e-7405-4d0d-a95f-a8a0c478f0fd · outbound

This paper cites Language models are few-shot learners,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Language models are few-shot learners,

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.707303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.707303Z digest=sha256:a2d00f3abbe6b7a1993a91bd3f8200599deb8fbc45192ccb3c80efb5951cebf1

Observation 58118cb6-2a68-4a9c-985c-e96e0734f571 · outbound

This paper cites Meta-prompting for automating zero- shot visual recognition with llms,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Meta-prompting for automating zero- shot visual recognition with llms,

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.711448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.711448Z digest=sha256:e7b8aded458ea0ff00beb6c757daf79f9eabfcaaad3a99e37b96e58b794f934d

Observation 0f584590-04b1-4ba6-841b-cf602dbaa2ac · outbound

This paper cites The neglected tails in vision-language models,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey The neglected tails in vision-language models,

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.715573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.715573Z digest=sha256:6bb0594795a60a25abbacf9a8561462c6a7557d6cbc69f84b9a0bbefb9fe1099

Observation 0a162a8a-de35-43ad-9d0c-ab0017399326 · outbound

This paper cites Introducing chatgpt,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Introducing chatgpt,

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.720336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.720336Z digest=sha256:b29574a6007316c4814b43b35fe3cd43714d9fcfca10d3e8a9d6c0db6711167e

Observation 3dacc8ed-c96d-498c-8816-f21d7e5c368f · outbound

This paper cites Prompting scientific names for zero-shot species recognition,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Prompting scientific names for zero-shot species recognition,

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.724722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.724722Z digest=sha256:d26e98bc4e77381c8aa8baedde1dea5b64b00ba0dc61cb4486147c544788d8aa

Observation 1470bece-a105-4470-b068-35d2358f5e17 · outbound

This paper cites Waffling around for performance: Visual classification with random words and broad concepts,.

Adapting Vision-Language Models Without Labels: A Comprehensive Survey Waffling around for performance: Visual classification with random words and broad concepts,

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:06.729087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:06.729087Z digest=sha256:61cb92b2c796f05ad39b7148faacce2bb846daf7862b0a4aad3a33a2d34f5a48

Pith citing papers

Observation a924611e-c64a-4e38-af94-4b9e9e490014 · inbound

VIP: Visual-guided Prompt Evolution for Efficient Dense Vision-Language Inference cites this paper.

VIP: Visual-guided Prompt Evolution for Efficient Dense Vision-Language Inference Adapting Vision-Language Models Without Labels: A Comprehensive Survey

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:27:29.164716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T07:27:02.544831Z digest=sha256:842e1e4dd2ef0136fc345b378371afc3f1a2eeac17a46eb4c342a7802191c9e1

Observation 63954e70-ee0f-45f8-a0c9-8e86a17425f9 · inbound

VIP: Visual-guided Prompt Evolution for Efficient Dense Vision-Language Inference cites this paper.

VIP: Visual-guided Prompt Evolution for Efficient Dense Vision-Language Inference Adapting Vision-Language Models Without Labels: A Comprehensive Survey

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:43:00.188003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-14T21:42:36.088959Z digest=sha256:743af229e5513e3410bdead0f6a8192eb2cb4a9a3bbeaf47ac3e785638e87598

Observation 21374d3e-ec90-4c91-86c3-b6e740d7c3d3 · inbound

ComMem: Complementary Memory Systems for Test-Time Adaptation of Vision-Language Models cites this paper.

ComMem: Complementary Memory Systems for Test-Time Adaptation of Vision-Language Models Adapting Vision-Language Models Without Labels: A Comprehensive Survey

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-30T12:54:40.378855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-30T09:57:02.377866Z digest=sha256:6d7c56b5bd0341b8d97983db7292fad0a3c34de7386ac8aa3ce9451f3d56a1b2

Observation 7646928a-0959-4ed0-9ea9-14e3b74238db · inbound

USE: A Unified Self-Ensembling Framework for Test-Time Prompt Tuning cites this paper.

USE: A Unified Self-Ensembling Framework for Test-Time Prompt Tuning Adapting Vision-Language Models Without Labels: A Comprehensive Survey

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-11T23:10:08.033785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T23:10:08.033785Z digest=sha256:64058494a0c06cbe82da272e59e1a331f2584bcdb221e3f121d4599c4143dcbe