Pith. sign in

Paper Citation Record · LEDGER

Variational Adapter for Cross-modal Similarity Representation

As of 9 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 0 inbound Pith citation observations for arXiv:2605.30968.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.30968 v1

Coverage vector

measured 16 of 16 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-28T23:04:14.351521Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

16 of 16 outbound references displayed

  • verified exact8
  • verified fuzzy0
  • unresolved7
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2b563439-ea02-4ae3-898b-1aec1e37d9fe · outbound

This paper cites Microsoft COCO Captions: Data Collection and Evaluation Server.

Variational Adapter for Cross-modal Similarity Representation Microsoft COCO Captions: Data Collection and Evaluation Server

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:16:00.174405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:3c50ad75466334df58704837f5d6dc41b7e4efc633a15e50e0a1e8d55410e577

Observation fbc31ecb-526e-434e-8303-1d2d37adcf6a · outbound

This paper cites VSE++: Improving Visual-Semantic Embeddings with Hard Negatives.

Variational Adapter for Cross-modal Similarity Representation VSE++: Improving Visual-Semantic Embeddings with Hard Negatives

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:16:00.157696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:819a4b49c225d85f210b3eafc16e5ce7f936744b810ae1bbcf0ceb559ba09028

Observation 0e81e320-1d72-430d-9fba-7448eabe7cdf · outbound

This paper cites Learning generative visual models from few training examples: An incremen- tal bayesian approach tested on 101 object categories.

Variational Adapter for Cross-modal Similarity Representation Learning generative visual models from few training examples: An incremen- tal bayesian approach tested on 101 object categories

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-28T23:04:14.351521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:9b3a79ef2e8a592deada2255c37ee5179bd65790032ba1d5c83bb62da83a1711

Observation bdb8f07f-9cad-4f81-bdd2-46f9bd55d584 · outbound

This paper cites Auto-Encoding Variational Bayes.

Variational Adapter for Cross-modal Similarity Representation Auto-Encoding Variational Bayes

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:16:00.170432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:42d1dea080c56343e1b478cf7f15b00e85f2d6eb71a8e2a34484c9e64478b7f6

Observation 10bebb61-460d-4435-82e2-c612871c838d · outbound

This paper cites Fine-Grained Visual Classification of Aircraft.

Variational Adapter for Cross-modal Similarity Representation Fine-Grained Visual Classification of Aircraft

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:16:00.171923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:a8bd33e599b495a2abedd09e8a4454e151259d79a196f6afe3a5e32e16c58b57

Observation c973190d-b69f-430a-9a31-e1af1bca4a65 · outbound

This paper cites Crisscrossed Captions: Extended Intramodal and Intermodal Semantic Similarity Judgments for MS-COCO.

Variational Adapter for Cross-modal Similarity Representation Crisscrossed Captions: Extended Intramodal and Intermodal Semantic Similarity Judgments for MS-COCO

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:16:00.168002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:6a1024bca073157d1fd1e012b216ac05b5170c50de76b0dd5cefa875ca749232

Observation 38b61483-6ae9-48e5-b3c7-05dc67544cc1 · outbound

This paper cites Consistency-guided Prompt Learning for Vision-Language Models.

Variational Adapter for Cross-modal Similarity Representation Consistency-guided Prompt Learning for Vision-Language Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:16:00.162367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:151d76be0a2eab4cf23b39b1b70e8895b1f8b4754cd00284ff612b561ab8adc3

Observation 2e8fe5c2-f876-46cf-a546-4b863c1a7096 · outbound

This paper cites UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild.

Variational Adapter for Cross-modal Similarity Representation UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:16:00.159049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:d4f6c4a920983c0a8f40ba1919c42d963bbffde30f73b1c8b872d656a16734ab

Observation b8a1df95-1f29-4a19-9706-7db4a60409b8 · outbound

This paper cites Probvlm: Probabilistic adapter for frozen vison-language models.

Variational Adapter for Cross-modal Similarity Representation Probvlm: Probabilistic adapter for frozen vison-language models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-28T23:04:14.351521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:55573af03ea2c04aaa7e5bb656136254c6e02fc84667ba33cfcdfe5505a70364

Observation 85604e52-8341-4db0-863f-673c1eb56d2c · outbound

This paper cites Towards Instance-wise Personalized Federated Learning via Semi-Implicit Bayesian Prompt Tuning.

Variational Adapter for Cross-modal Similarity Representation Towards Instance-wise Personalized Federated Learning via Semi-Implicit Bayesian Prompt Tuning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:16:00.175678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:5577a1f57753bd8803574d4379924b662be39215e5a7b0637a171d56abe31af7

Observation 2e869e0b-8e4b-4480-923e-6131f33a2db7 · outbound

This paper cites C., and Liu, Z.

Variational Adapter for Cross-modal Similarity Representation C., and Liu, Z

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-28T23:04:14.351521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:b29662ae37213c1198482453a83c5ca84f94b41a38c6b66b0683745fe18e13be

Observation 8cc7b580-0128-408f-bf6a-9a7b7cd755d7 · outbound

This paper cites Upadhyay et al.(Upadhyay et al.,.

Variational Adapter for Cross-modal Similarity Representation Upadhyay et al.(Upadhyay et al.,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-28T23:04:14.351521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:4b0770e43a2b2dba393d6b06553ac6eac3a0c0b90ab5eb7bbc4040e100e97a3b

Observation c8176680-ae13-416e-aa11-3f3aba5675ad · outbound

This paper cites Collectively, these approaches expand the set of potential results, constructing a richer semantic retrieval space.

Variational Adapter for Cross-modal Similarity Representation Collectively, these approaches expand the set of potential results, constructing a richer semantic retrieval space

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-28T23:04:14.351521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:f1f1af4037c90b8f772c5385546eb57ef45987eb06e85e231d144b94ba1b54b1

Observation ed0e5583-2841-40a2-9fc2-87fbf0dc992a · outbound

This paper cites We employ a 16-shot setting and use the template ”a photo of a <category>” for the word embeddings.

Variational Adapter for Cross-modal Similarity Representation We employ a 16-shot setting and use the template ”a photo of a <category>” for the word embeddings

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-28T23:04:14.351521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:2ac3305f4c96b5796949b100d71359506409b4eaae103b9975289e8da1a48cf2

Observation 4c5c5e61-3753-4fce-a9cb-435a77b34b69 · outbound

This paper cites an unresolved cited work.

Variational Adapter for Cross-modal Similarity Representation Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-28T23:04:14.351521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:455101c3a4172d5ff778ea93d8316b99f61818e0a04f11e6ac711fdf4bdd9c54

Observation 11ceda50-bd12-426f-8f46-7c4d4f6081d3 · outbound

This paper cites the sample may not actually be a negative sample.

Variational Adapter for Cross-modal Similarity Representation the sample may not actually be a negative sample

Reference 16

Resolution
malformed identifier
arxiv_id, observed 2026-07-01T19:16:00.173027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T23:04:14.351521Z digest=sha256:58bc9efa7436906bce4a704940f5377585f5f274443525790bec6e5ce5c6530a

Pith citing papers

No inbound Pith citation observations are available.