Pith. sign in

Paper Citation Record · LEDGER

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model

As of 23 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 1 inbound Pith citation observation for arXiv:2506.23822.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.23822 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:34:42.355397Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-29T12:35:51.321835Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T12:43:25.937018Z

Reference resolution

57 of 57 outbound references displayed

  • verified exact0
  • verified fuzzy55
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 40b413b2-a1a8-4381-a53c-ad456714f883 · outbound

This paper cites Label-embedding for image classification.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Label-embedding for image classification

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.278673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:38.119515Z digest=sha256:9b9b02f10bf5b5a1e4273f2be9c1e679db77c0e800fe0c117af7ba852e59b0cf

Observation f1f2df2e-7d76-4288-9f15-2fbd7e3a5a4f · outbound

This paper cites Wasserstein generative adversarial networks.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Wasserstein generative adversarial networks

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.271135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:38.187652Z digest=sha256:67df5ff8a69356cf8336d13a4fc942a8169b08ff65491ab9bbfe758d9d1a547f

Observation 97750993-608f-40eb-a06d-f943ac381e8f · outbound

This paper cites Food-101 - mining discriminative components with random forests.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Food-101 - mining discriminative components with random forests

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.263407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:38.247061Z digest=sha256:3f99165c2697a0f04091f0c3617d11d2c38724be3f39832b82b853c3b8b2b193

Observation 576058e6-045c-434d-9872-8b9873384dd6 · outbound

This paper cites PLOT: prompt learning with optimal transport for vision-language models.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model PLOT: prompt learning with optimal transport for vision-language models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.254621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:38.317911Z digest=sha256:bfba5db12ee6aa8f8c633ad3b2833f89705901d1adea8b55dc1f3e52ed267762

Observation 0d7687e3-e8b2-4e8e-9802-43069ae99510 · outbound

This paper cites Hsva: Hi- erarchical semantic-visual adaptation for zero-shot learning.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Hsva: Hi- erarchical semantic-visual adaptation for zero-shot learning

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.247518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:38.405260Z digest=sha256:85283dc79089757c54f675f3a5cf58d2d3cebb86c951d6cfb82160b9e79889c6

Observation 29e7a758-0192-40e3-939b-71eb25c0a6f1 · outbound

This paper cites MSDN: mutually semantic distillation network for zero-shot learn- ing.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model MSDN: mutually semantic distillation network for zero-shot learn- ing

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.238807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:38.491003Z digest=sha256:d7ff46af60dc0c674df49e4c5e6788e2e4a7e5695c07b5372175eef1e8537509

Observation 8383f5d3-e207-49fe-9539-d36a81351640 · outbound

This paper cites Transzero++: Cross attribute-guided transformer for zero-shot learning.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Transzero++: Cross attribute-guided transformer for zero-shot learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.229905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:38.575798Z digest=sha256:c5a218486dd11f9e837b6465b8247bdcb5a9ac04729b43534090d928c0a7befd

Observation bff84e56-9018-4421-a534-2207a90e60bd · outbound

This paper cites Evolving semantic prototype improves generative zero-shot learning.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Evolving semantic prototype improves generative zero-shot learning

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.219668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:38.661869Z digest=sha256:a3d56753c7dbadbbf887c39329b5049d24929d8fce9d91a81db557be670406fe

Observation c7705c4c-b248-4996-b224-f92f00091485 · outbound

This paper cites Khan, and Fa- had Shahbaz Khan.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Khan, and Fa- had Shahbaz Khan

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.210594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:38.741203Z digest=sha256:08a79a6fcd3a05ff521af946ab2969788ddd6795457930eb242a149b6870b90a

Observation 077d3ffe-47c1-4773-8831-a7c972902570 · outbound

This paper cites Khan, and Fa- had Shahbaz Khan.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Khan, and Fa- had Shahbaz Khan

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.202557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:38.823129Z digest=sha256:f7aac91091c630db3a9e01cab3be7800acc51e5a2db58f9416e55e0cad113cfb

Observation eae02ef5-81fb-4e95-9277-d48b3f25a511 · outbound

This paper cites Semantics-conditioned generative zero-shot learning via fea- ture refinement.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Semantics-conditioned generative zero-shot learning via fea- ture refinement

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.195108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:38.883354Z digest=sha256:446c4e562a5ef6fc7af1c6e3f115ed69ac1079e6c7048c1d964d4ddcfc4bb655

Observation 27ef409b-7c26-42f4-9ecd-c46a31848d96 · outbound

This paper cites Evolving interpretable visual classifiers with large language models.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Evolving interpretable visual classifiers with large language models

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.188074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:38.966148Z digest=sha256:c111ec2560b58d71344aa78e1106e58fdc4e7e83f0e1f85c6edc32abaa084396

Observation 81d3dcfe-0038-4188-8425-43755b51664f · outbound

This paper cites Sinkhorn distances: Lightspeed computation of optimal transport.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Sinkhorn distances: Lightspeed computation of optimal transport

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.181638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:39.058199Z digest=sha256:5ce41f122f0e87faeac85e3ab26ddbece5277b8d02596f22aff987f716bfa416

Observation aa7d2f61-7e30-4a7e-8c14-c6951c5f4b90 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Imagenet: A large-scale hierarchical image database

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T21:34:39.128557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:34:39.128557Z digest=sha256:6e202a9b35f607710e1d1187f0d1f904430d64b7b9a1c44d2d3849303f3e902b

Observation 46337fe5-7c4f-4deb-98b9-28ab3d0b1892 · outbound

This paper cites Image2sentence based asymmetrical zero-shot composed image retrieval.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Image2sentence based asymmetrical zero-shot composed image retrieval

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.170000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:39.193778Z digest=sha256:0c68fb161ac86669d98f9861be6c545f8f2d6b964deb060d6a2702175ffdb805

Observation 7f06d496-7393-4c0e-9191-4984a3db0337 · outbound

This paper cites Improving clip training with language rewrites.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Improving clip training with language rewrites

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.162692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:39.261986Z digest=sha256:d49566de44b78628b45ec78def8971322b16a25496c83d897379d79733559f2f

Observation 50ff2b36-6f07-4d71-9e31-5881476ebcc6 · outbound

This paper cites Diverse data augmentation with diffusions for effective test-time prompt tuning.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Diverse data augmentation with diffusions for effective test-time prompt tuning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.156722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:39.334243Z digest=sha256:1269650105959b97bf417fc79e8401d2e2a5c85d743914fe01aa355eea726726

Observation 8809b69a-0799-40fe-8cbb-979f50296397 · outbound

This paper cites Clip-adapter: Better vision-language models with feature adapters.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Clip-adapter: Better vision-language models with feature adapters

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.150383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:39.412790Z digest=sha256:0d31138f39636680c5afe9afb3501ea55fbbae71950eac53d2cf2bb276ae9b99

Observation 9ab5ffc0-fd63-4977-854f-19d5ad9495a8 · outbound

This paper cites The many faces of robustness: A criti- cal analysis of out-of-distribution generalization.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model The many faces of robustness: A criti- cal analysis of out-of-distribution generalization

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.144076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:39.507419Z digest=sha256:13ae1cd4a77c8ab2b72d01d6d1603f18b0f64071cd4dc6f0b97ace8c8df35354

Observation 38767e60-f574-4482-a15b-6f37d7592907 · outbound

This paper cites Natural adversarial examples.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Natural adversarial examples

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.136938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:39.555060Z digest=sha256:ef5036b33443cdcc56b92724b13b48d7dd4554bde13df0af284bbfc1cd5895c4

Observation ddb8f569-5c5b-46e6-9d7c-070499bb4da4 · outbound

This paper cites Fine-grained generalized zero-shot learning via dense attribute-based attention.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Fine-grained generalized zero-shot learning via dense attribute-based attention

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.131033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:39.642698Z digest=sha256:83a8432254204a5de3875aca9b0a8fd78a84ba26922138d4f3eee5571d952ee9

Observation 13a7fd2e-058c-4df0-bf02-afc81298e31b · outbound

This paper cites Le, Yun-Hsuan Sung, Zhen Li, and Tom Duerig.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Le, Yun-Hsuan Sung, Zhen Li, and Tom Duerig

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.122621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:39.707291Z digest=sha256:39a962ad60f087cc68a615ba86b81b456ac1e0b126d594d515abe47b76934286

Observation 92bce210-3093-4185-9b86-4de4fdcf0aac · outbound

This paper cites Multi- modal classifiers for open-vocabulary object detection.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Multi- modal classifiers for open-vocabulary object detection

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.114989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:39.797286Z digest=sha256:1878b4369274c3a850c76b3a5acbdf01915b976aa29c9f8eee54c313b30fe554

Observation d0d90183-cbbe-4849-97cb-536e4ce67108 · outbound

This paper cites Khan, and Fahad Shahbaz Khan.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Khan, and Fahad Shahbaz Khan

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.105812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:39.863413Z digest=sha256:d06080bdb879fabd9a38fcb78a1ce660faabfa9f22ebebc80af2c610adb4025e

Observation ab8f1ad2-bcb2-4b5b-b43e-38fa875ee780 · outbound

This paper cites Kolkin, Jason Salavon, and Gregory Shakhnarovich.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Kolkin, Jason Salavon, and Gregory Shakhnarovich

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.099868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:39.938287Z digest=sha256:1b20b127fb0dda72648be07876dc6fadee0d4cd33d63b3b26323e4ec80eae86d

Observation 9bb93967-9739-474b-a061-785bb67415b3 · outbound

This paper cites Co-clustering through optimal transport.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Co-clustering through optimal transport

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.093614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:40.037549Z digest=sha256:324a7dcb35f351f1f63607e3a12d03d555b39e314f24b652b0b97fd5b319e947

Observation ef2909ca-1805-4c29-aabf-46c241faaed0 · outbound

This paper cites Visual-text cross alignment: Refining the similarity score in vision-language models.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Visual-text cross alignment: Refining the similarity score in vision-language models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.086316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:40.126420Z digest=sha256:0fcff7c3afcb439ffb881e11687d0ad49ead93393b366ed3a978c8057d8613d5

Observation c06d7e7e-fb8d-4072-a48f-c264e7da6b5e · outbound

This paper cites Patchct: Align- ing patch set and label set with conditional transport for multi-label image classification.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Patchct: Align- ing patch set and label set with conditional transport for multi-label image classification

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.076976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:40.220034Z digest=sha256:94772e47238d84f807b389875d7279f1742e2f036ac12f9f14165ece9b6f5a0c

Observation f20368ac-095b-46c4-acc0-b067581ecdf3 · outbound

This paper cites Progressive semantic-visual mutual adaption for generalized zero-shot learning.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Progressive semantic-visual mutual adaption for generalized zero-shot learning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.069234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:40.310639Z digest=sha256:3ff30e1a233a4faf0831bf54b97bf1e3ffe5e7e99c0176c863c08d1b0e7213e0

Observation b82a2603-7b41-4931-bbc9-e4a4baf1b5b4 · outbound

This paper cites Visual classification via description from large language models.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Visual classification via description from large language models

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.060118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:40.396768Z digest=sha256:c664ceecbb063014e1a7cd610a0fbb15c9506dbedf77af686dd5aac67aba9891

Observation 56e367c0-2645-40ce-8f97-c362dcdfa37b · outbound

This paper cites Robust calibration of large vision- language adapters.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Robust calibration of large vision- language adapters

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.051149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:40.490179Z digest=sha256:3c581c0494f72c5a50fc01c0ea56710a002a48e1b73598df6b09fabb70d35aff

Observation fa8fd233-7fde-43f0-8449-f86d097fce3d · outbound

This paper cites Robust calibration of large vision- language adapters.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Robust calibration of large vision- language adapters

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.043443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:40.558778Z digest=sha256:e8709288c5e6b0079a26d8f60f3f19916b77c3529d550c0e9c5fa3d9c2272839

Observation 8d17f56b-fb2f-4ca6-a73e-680b06728d7d · outbound

This paper cites I2dformer+: Learning image to document summary attention for zero-shot image classifi- cation.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model I2dformer+: Learning image to document summary attention for zero-shot image classifi- cation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.035914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:40.637193Z digest=sha256:56c1fbc4326acc5686a7696480167e119c3274b07d7b20f3abc79628edb6b5f3

Observation d003a31f-4eb8-4d97-a58e-14dda1893eee · outbound

This paper cites Pomerleau, Geoffrey E.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Pomerleau, Geoffrey E

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.028313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:40.705915Z digest=sha256:b5d15012ea883b610d39ebf3df825a8cebf85c9639c6731084e5bacc2b34dd59

Observation 3195ebc4-ff21-4ccd-b9c5-a254b9ce0073 · outbound

This paper cites Parkhi, Andrea Vedaldi, Andrew Zisserman, and C.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Parkhi, Andrea Vedaldi, Andrew Zisserman, and C

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T21:34:40.802004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:34:40.802004Z digest=sha256:584cf9777d32f638d10461247251f7e7680fc3e7989445c410dc67777807225d

Observation c1f54dfa-500b-4d0f-ac6f-ddad495d6bb5 · outbound

This paper cites Computational optimal transport.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Computational optimal transport

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.016913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:40.874839Z digest=sha256:9bae5391a9fd60a150d76867df22540f5754688494e1f3a64f29739e8c4b58b2

Observation d1dac71e-820d-4434-9b06-e8cfc6c5e08c · outbound

This paper cites What does a platypus look like? generating customized prompts for zero- shot image classification.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model What does a platypus look like? generating customized prompts for zero- shot image classification

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.008948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:40.951345Z digest=sha256:f36d391c00986c13ea232fc43ed2f0930cff0f961bc9120b31aa5313c6b4cb62

Observation f60793b9-96d0-4e32-b84c-f2101a735e61 · outbound

This paper cites Pratt, Ian Covert, Rosanne Liu, and Ali Farhadi.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Pratt, Ian Covert, Rosanne Liu, and Ali Farhadi

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:45.001796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:41.030911Z digest=sha256:ee06cb979a1b1897b8b3cec57fab39654671fbe9c462d98f06a5ef6a5c80ea51

Observation fd36e31a-c486-459d-9cb9-207a2557eadb · outbound

This paper cites Proapo: Pro- gressively automatic prompt optimization for visual classifi- cation.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Proapo: Pro- gressively automatic prompt optimization for visual classifi- cation

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:44.995071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:41.106333Z digest=sha256:02d2e6e0b15f586b46475c5981b57420ab1ba69f1214aa2b4597269030f30181

Observation e54bfc82-2d2c-4346-be19-9de2a842179a · outbound

This paper cites Learning transferable visual models from natural language supervision.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Learning transferable visual models from natural language supervision

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:44.838610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:41.181037Z digest=sha256:d7198bf6af9dcdcc58d11fe9eb22712901603db97eb1c8e3891ea76f38f8312f

Observation 64c9ff44-6c22-4dc3-a3eb-c742f18e01f0 · outbound

This paper cites Do imagenet classifiers generalize to im- agenet? In ICML, pages 5389–5400, 2019.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Do imagenet classifiers generalize to im- agenet? In ICML, pages 5389–5400, 2019

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:44.700283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:41.234392Z digest=sha256:b13e38f7cfcd591768f55787e719c1053ecb7d6f0aa6eaee7fed631159563a5e

Observation a62323ec-9ffd-4390-b42a-8b970de10065 · outbound

This paper cites Sophia Koepke, Oriol Vinyals, Cordelia Schmid, and Zeynep Akata.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Sophia Koepke, Oriol Vinyals, Cordelia Schmid, and Zeynep Akata

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:44.539090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:41.319904Z digest=sha256:f0cec04e27fd1a6a96b60778e37ac411340a1208bd011cbef84a93f62c58a9a7

Observation 294ff7c5-fa3f-48b9-877d-b075229d6aa9 · outbound

This paper cites Generalized zero- and few- shot learning via aligned variational autoencoders.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Generalized zero- and few- shot learning via aligned variational autoencoders

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:44.478190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:41.409051Z digest=sha256:37b4da1b02faa3ac9cc21229921b98e8cc88e34a093e941b6bc1b2a2d64cd75d

Observation 84644f33-3bed-47f6-bb0d-5eac2285653a · outbound

This paper cites Test- time prompt tuning for zero-shot generalization in vision- language models.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Test- time prompt tuning for zero-shot generalization in vision- language models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:44.406486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:41.455920Z digest=sha256:e5a8c61203a8e8beb004a310056163818fbf992df5c4be786cc0a3c142c4b00a

Observation 1d6d6c11-3c33-47ce-a64f-16ac96086f80 · outbound

This paper cites Hunting attributes: Con- text prototype-aware learning for weakly supervised seman- tic segmentation.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Hunting attributes: Con- text prototype-aware learning for weakly supervised seman- tic segmentation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:44.185725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:41.521708Z digest=sha256:eceedbb2f19376ad33ea032577c1eecae7853b4cb3d1c3f7e799c2ebc48b6602

Observation 30ee8ef2-1c23-438c-84f7-946459cbd1ef · outbound

This paper cites Argue: Attribute-guided prompt tuning for vision-language models.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Argue: Attribute-guided prompt tuning for vision-language models

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:44.058181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:41.582088Z digest=sha256:5596f82c647bf26ff02aa57739b947de93cf6b8e709d75dad261170dc5931aa8

Observation fc5846fe-897e-482c-b385-8c54c2b36623 · outbound

This paper cites Tuning multi-mode token- level prompt alignment across modalities.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Tuning multi-mode token- level prompt alignment across modalities

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:43.960764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:41.646705Z digest=sha256:ee4198fdca89f84834af6d2d707d04e1b52def90cd27cf5b71bd2ebef509ed1f

Observation bf6e56df-5db3-401f-bc4e-cfcb3777d3a7 · outbound

This paper cites Lipton, and Eric P.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Lipton, and Eric P

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:43.948255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:41.728611Z digest=sha256:3e2aa993682c573d74de5084babdd214c527b7a817aaea33e57f64ec4c1cab55

Observation 65327196-1cb1-458f-922b-861af002a7c9 · outbound

This paper cites Zero-shot visual recognition via bidirectional latent embedding.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Zero-shot visual recognition via bidirectional latent embedding

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:43.858773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:41.814733Z digest=sha256:f12f440bd13932c230373752c573eb5181c88a36c14df1ab08b978441a575083

Observation f8ff87f1-fec1-4ec5-a183-362e6082e662 · outbound

This paper cites Welinder, S.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Welinder, S

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:43.640489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:41.883185Z digest=sha256:ba981943fba806a2933ae4ac1f791a123179b111bf0556e19962331974009f8a

Observation 028b4ce2-dfe8-41a6-929b-f19f062ff9da · outbound

This paper cites Schiele, and Zeynep Akata.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Schiele, and Zeynep Akata

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:43.462090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:41.964125Z digest=sha256:464168c973ab59ef1bb1daf6cd49b88cdd5eef6f1632523a69af17db57b095fc

Observation 5128ea2a-be29-476c-b3ac-1f4fd8f29558 · outbound

This paper cites Lorenz, B.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Lorenz, B

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:43.236837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:42.010840Z digest=sha256:2e09b45e69d5d8b834f62e06ccf5e3e1abae5ef7b46781ae6f9c014f6711822f

Observation de7ada3a-90ad-4bf3-9832-02e5e8230e77 · outbound

This paper cites Khan, Zhiqiang Shen, Muzammal Naseer, Guangyi Chen, and Fahad Shahbaz Khan.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Khan, Zhiqiang Shen, Muzammal Naseer, Guangyi Chen, and Fahad Shahbaz Khan

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:43.015907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:42.098251Z digest=sha256:4e766180f841373071ff94d3e78f10f4914bff142ce881b14a22f2c9378ba1fd

Observation 7c075e03-1d4c-4391-a993-57253318aabe · outbound

This paper cites Places: A 10 million image database for scene recognition.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Places: A 10 million image database for scene recognition

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:42.831883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:42.165157Z digest=sha256:9c6281419445c8ee4416c6fdda1b339ba02f4f0ca7257f1501e7e60a8a7f8249

Observation 55fef51c-db4c-490a-907b-361581c2074d · outbound

This paper cites Conditional prompt learning for vision-language models.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Conditional prompt learning for vision-language models

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:42.697098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:42.232511Z digest=sha256:e1dd7e5c7dbe89bfde91dcb28b2feaf3b6e7b35e49bc0009e7bbfd0e4933a0a6

Observation 30a63017-5af8-4e64-ac96-eb058b7cc3cb · outbound

This paper cites Learning to prompt for vision-language models.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Learning to prompt for vision-language models

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:42.577356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:42.300495Z digest=sha256:4b0c5a1af8e5df2a906679ed2b747fdfbca0c3cd96c6b4559d2db52aaee62f04

Observation 3530d133-5caa-4089-924a-4d0abafedcb5 · outbound

This paper cites Prompt-aligned gradient for prompt tuning.

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model Prompt-aligned gradient for prompt tuning

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:34:42.478926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T21:34:42.355397Z digest=sha256:37e7e97295ebf08f41f2b1ddf0c67e989233ec012b5917fcd965693e9f0d61b7

Pith citing papers

Observation 603626a4-5abb-4478-a7a8-f87782c6921b · inbound

FedMPT: Federated Multi-label Prompt Tuning of Vision-Language Models cites this paper.

FedMPT: Federated Multi-label Prompt Tuning of Vision-Language Models Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:43:25.938440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-29T12:35:51.321835Z digest=sha256:545d9f50bf36dc753d5f500d681e6376734d18d5b54ab6697a28a279f96d213c