Pith. sign in

Paper Citation Record · LEDGER

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning

As of 13 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 0 inbound Pith citation observations for arXiv:2412.19289.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.19289 v3

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T00:47:52.532503Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

50 of 50 outbound references displayed

  • verified exact6
  • verified fuzzy3
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 226f241b-a6e8-4021-897a-0b4095f4c58e · outbound

This paper cites , " * write output.state after.block = add.period write newline.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.337518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.337518Z digest=sha256:3ac26b41896edbcae1152281b9421aa79af5ebedc033284c83a998831e6d1d2a

Observation 0dae187d-1325-49b4-b159-77c04d02c43b · outbound

This paper cites write newline.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.342475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.342475Z digest=sha256:aa0552b8606837b04dfb048c7091ca46c3ea7da8fb36802e2eed3f38f076d00b

Observation 7a818e42-d0de-4553-b90a-7c042df8278c · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.347000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.347000Z digest=sha256:4e4d8f1881beb24dcec7358a45d8d006e176898393db519a0b907ad1a7e78bea

Observation cff45548-ce89-437c-9505-aa5874459687 · outbound

This paper cites SPICE: Semantic Propositional Image Caption Evaluation.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning SPICE: Semantic Propositional Image Caption Evaluation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.350966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.350966Z digest=sha256:52c295969f43641f0c28f9db71529e3c34eb82d38ea1392ddc0188311fdd55f3

Observation d878a7b6-5845-47db-b99b-6ce4702cdc54 · outbound

This paper cites Exploring Visual Prompts for Adapting Large-Scale Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Exploring Visual Prompts for Adapting Large-Scale Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.355230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.355230Z digest=sha256:1592b478bbb22f9f2268a568be15982eac224900ea7ccde39f834792329cad9d

Observation 28adb247-1126-4345-9a50-40070c3b7b43 · outbound

This paper cites CaMEL: Mean Teacher Learning for Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning CaMEL: Mean Teacher Learning for Image Captioning

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-11T00:47:52.932804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.359461Z digest=sha256:fbd429e745a6565a0039b5cf2d52011c10c3ae5c095c808e0793c1dc94d3b99f

Observation a8898f81-acc2-45bb-a3c5-fa492e59170d · outbound

This paper cites PaLI: A Jointly-Scaled Multilingual Language-Image Model.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning PaLI: A Jointly-Scaled Multilingual Language-Image Model

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.363799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.363799Z digest=sha256:fab600e868d925c7164fe43722d9d02bc55f3fe9ec3af9b47e9f5c2265653efb

Observation e671ddb3-7ad5-420c-9562-2c709005c4c4 · outbound

This paper cites E.; Stoica, I.; and Xing, E.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning E.; Stoica, I.; and Xing, E

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.367994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.367994Z digest=sha256:7ecdf88b984db8a4ee0c65b4431a48fa17c64795b7c4b5205c7e1722cfd2bcc6

Observation 8aa8e136-3c00-4fcf-9234-2aebaf9832a4 · outbound

This paper cites J.; and Lavie, A.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning J.; and Lavie, A

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:47:53.053822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.372139Z digest=sha256:eafd1e9ec9f8660f7def5dd428d55cd1772e5f9e913aaf103610d775578da633

Observation 023456da-f5ef-4eab-86e7-e8132b51b89e · outbound

This paper cites Transferable Decoding with Visual Entities for Zero-Shot Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Transferable Decoding with Visual Entities for Zero-Shot Image Captioning

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-11T00:47:52.907331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.377156Z digest=sha256:f78695f86cc25af04792dbc4c9eef129fecb5886082b59a52f2dfb034386ae95

Observation b9500f88-5c56-4d37-8c59-36dc115bd496 · outbound

This paper cites Making Pre-trained Language Models Better Few-shot Learners.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Making Pre-trained Language Models Better Few-shot Learners

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.381700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.381700Z digest=sha256:b724b65f068486295516abfa441fd03b7d3fd34f639e689bf812af76cfcc44cb

Observation 49d42f1d-8228-409b-9826-da1b7e270051 · outbound

This paper cites Language-only Efficient Training of Zero-shot Composed Image Retrieval.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Language-only Efficient Training of Zero-shot Composed Image Retrieval

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.386106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.386106Z digest=sha256:5200c0e80be566de5dfa23a98ca4c482cfad7b7c52f7cc969ae5e830092723bc

Observation f8c9a401-df39-491f-bc5d-a55c981f36aa · outbound

This paper cites Scaling Up Vision-Language Pre-training for Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Scaling Up Vision-Language Pre-training for Image Captioning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.390168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.390168Z digest=sha256:3898b1a400d903090ec59f5b4226593e90933ea8af703e89a571ee21e81ee55c

Observation 0d0790bb-0b6f-4354-a632-5587ddb2eefb · outbound

This paper cites REVEAL: Retrieval-Augmented Visual-Language Pre-Training with Multi-Source Multimodal Knowledge Memory.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning REVEAL: Retrieval-Augmented Visual-Language Pre-Training with Multi-Source Multimodal Knowledge Memory

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.394111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.394111Z digest=sha256:e8cbb0a73558b1415acdab00ad7c6267cba9b5bff8111d01279fdde6ed00798f

Observation cdcd0b62-1ac9-42fd-beca-718e10dd8200 · outbound

This paper cites Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.398127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.398127Z digest=sha256:3cbc24c1a540e68a8de0a23ddac745f848378bba56ac8e5dd5199f9a2e08d83b

Observation 68cfa15d-5001-412c-8014-6bba038e47b0 · outbound

This paper cites Visual Prompt Tuning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Visual Prompt Tuning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.402265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.402265Z digest=sha256:9882323df8bd0d67e1ee0c860e0218d0756654b2b557e094780b62a521f3e44b

Observation f6844831-ca97-454d-923c-9795e2a6a256 · outbound

This paper cites Billion-scale similarity search with GPUs.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Billion-scale similarity search with GPUs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.406187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.406187Z digest=sha256:990484abbcb50c5b1cc06bbe50a081773fc00c034a4379580f6f0e6bf743f0ac

Observation fbf6f1ea-74aa-4474-9dad-76ac4b4f5126 · outbound

This paper cites Deep Visual-Semantic Alignments for Generating Image Descriptions.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Deep Visual-Semantic Alignments for Generating Image Descriptions

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.410230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.410230Z digest=sha256:ca8834b165c8a6b6eac0ae8f9c8acf04f029e9c5576873b0ad5f92fed77b3375

Observation 2e8b41ff-a450-4e9c-af0d-c5631d105fa0 · outbound

This paper cites Auto-Encoding Variational Bayes.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Auto-Encoding Variational Bayes

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.414573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.414573Z digest=sha256:f0d073399fcf050ba8456e390052ed16c08d26a3c6deccbd7be8e9f6fc278224

Observation 778da2bf-ee88-4ffc-a9d5-3a556e13dc02 · outbound

This paper cites The Power of Scale for Parameter-Efficient Prompt Tuning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning The Power of Scale for Parameter-Efficient Prompt Tuning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.418671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.418671Z digest=sha256:8ff4614a986f4c9128f95a39670081658b3302d3ceaa3d8e488c438bc1a8b1cb

Observation 59ca8558-0ed1-43ae-bf88-aa1a3ad6f4a5 · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.422712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.422712Z digest=sha256:7c015671da1171befdbb08dff5d20c4214b361d43e0477af0005fb8872ed543e

Observation 230e9ed6-070d-49f1-8d6d-6f975a8b0aa3 · outbound

This paper cites EVCap: Retrieval-Augmented Image Captioning with External Visual-Name Memory for Open-World Comprehension.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning EVCap: Retrieval-Augmented Image Captioning with External Visual-Name Memory for Open-World Comprehension

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.426766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.426766Z digest=sha256:ed17bd6b8b9fba9a7c0f30e81ad95cfa4cbdfd76f7ca2c2c1ad32cad4afa97d2

Observation a9f4eac5-2562-4f10-b109-2dda3ee40556 · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:47:53.041629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.430620Z digest=sha256:69fd0def1bab05c3bc4297c902292940358d3901a6df881c5d019c588eb3a196

Observation fd13970e-17f6-4ac2-becd-daad3b68fff5 · outbound

This paper cites Oscar: Object-Semantics Aligned Pre-training for Vision-Language Tasks.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Oscar: Object-Semantics Aligned Pre-training for Vision-Language Tasks

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.434115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.434115Z digest=sha256:2de82a6a5cb2ff82cdcbfc5aed54fd544f2ce4f54052ccc82337e78738f98a3f

Observation 0395bd76-e967-4b2f-b5fd-f7a6b66007e1 · outbound

This paper cites Prefix-Tuning: Optimizing Continuous Prompts for Generation.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Prefix-Tuning: Optimizing Continuous Prompts for Generation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.437835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.437835Z digest=sha256:8c41c53c5d73474a3e4b3269c00437ca1d6dee9f6b2df9fdd8169cdded632a3c

Observation 434cd9d7-14b2-4a18-83bf-708115380476 · outbound

This paper cites Microsoft COCO: Common Objects in Context.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Microsoft COCO: Common Objects in Context

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.441953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.441953Z digest=sha256:744be8d97fbcb6d024c21370dc830704c0db2ff2464a2c8bb67dc02b20f40c9f

Observation 2f665086-7745-455f-ac61-ae254c66e6c2 · outbound

This paper cites Few-shot Learning with Multilingual Language Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Few-shot Learning with Multilingual Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.446185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.446185Z digest=sha256:c43391bb177878b1566790bc579cb04267ff77d9f4694a6d22e226b4a392749b

Observation 93255380-34c1-443d-88a2-11ec4e3a7bee · outbound

This paper cites Visual Instruction Tuning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Visual Instruction Tuning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.450290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.450290Z digest=sha256:c1f216f2d2e4190f9ec8ce2d39fe637892cd07bf6e6de440ede6da2168cc3825

Observation 226c086b-add0-4e3f-839c-d969a4be0fc6 · outbound

This paper cites I-Tuning: Tuning Frozen Language Models with Image for Lightweight Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning I-Tuning: Tuning Frozen Language Models with Image for Lightweight Image Captioning

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-11T00:47:52.727828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.454181Z digest=sha256:42df83173a3d5b1903ed91d606738b17680c6b234dc51f762592f74c27d75b06

Observation 230a2993-3479-4c04-8728-416ccb573566 · outbound

This paper cites MAPL: Parameter-Efficient Adaptation of Unimodal Pre-Trained Models for Vision-Language Few-Shot Prompting.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning MAPL: Parameter-Efficient Adaptation of Unimodal Pre-Trained Models for Vision-Language Few-Shot Prompting

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.458002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.458002Z digest=sha256:2ea9722b3d0812f74023c62eb33bc18c750f550a0ef9636a95b6856b881b50b6

Observation 6f40e8af-c06c-4d64-87db-5b4eb3febfb7 · outbound

This paper cites ClipCap: CLIP Prefix for Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning ClipCap: CLIP Prefix for Image Captioning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.461697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.461697Z digest=sha256:0f31a6379f1e8cbe346f82c9ea82f4c3e33dff7f867a5e03851d4bd1931fc89f

Observation e739e8bb-f0fe-4bc1-b724-fd797411b8b6 · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:47:53.030299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.465468Z digest=sha256:3df3f75d183d335a5ed22beef6c3de8ae52ca30c2365465e8c7142b8769ffd14

Observation e56161bb-1edb-48aa-b691-edaf369245f4 · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:47:53.018872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.468715Z digest=sha256:ab31e182277c338c16943d17f490fdd58babcb7694f663bd3c0edaadaa9e359d

Observation 5841e810-0639-45f3-bec8-1f34f573c01e · outbound

This paper cites Flickr30k Entities: Collecting Region-to-Phrase Correspondences for Richer Image-to-Sentence Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Flickr30k Entities: Collecting Region-to-Phrase Correspondences for Richer Image-to-Sentence Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.472320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.472320Z digest=sha256:2ccdc099a62bd50a901692de4271b052705e5a014e539a78e4aafbf301bbd9cb

Observation afa7388d-2396-49cc-b606-d7566f664077 · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Learning Transferable Visual Models From Natural Language Supervision

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.476285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.476285Z digest=sha256:28159c5879693e205ae8081caacc298a191dd5c0341b2905464b00533a70453b

Observation 982a31ea-aaee-44bb-8600-1f34d2273480 · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:47:53.007004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.480296Z digest=sha256:6eb02c38669d3bd1fcd28d8a275aa86f2c4e34866353508b34cf128949344adc

Observation 3ea24812-6756-427b-9c9a-6a89fb70bec9 · outbound

This paper cites LMCap: Few-shot Multilingual Image Captioning by Retrieval Augmented Language Model Prompting.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning LMCap: Few-shot Multilingual Image Captioning by Retrieval Augmented Language Model Prompting

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.483794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.483794Z digest=sha256:cd85c75f5e5907ec65b87b53fc0d2d97a16d93711ceb437b7a793d4e08ed3463

Observation a5fbff30-4ddb-4018-9f30-1e344aa2a986 · outbound

This paper cites SmallCap: Lightweight Image Captioning Prompted with Retrieval Augmentation.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning SmallCap: Lightweight Image Captioning Prompted with Retrieval Augmentation

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-11T00:47:52.660913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.487763Z digest=sha256:1e73a009ee1141d163ab6745568b59f01157fd7196a8974d0e41bfd481c8afc8

Observation 658f9a71-d444-4e33-b749-0caa9763f830 · outbound

This paper cites P.; Elliott, D.; and Martins, B.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning P.; Elliott, D.; and Martins, B

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:47:52.996295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.491626Z digest=sha256:245e6c817e11dc853a18e5c1b7410f05504cf268178804412340f69b6421a739

Observation 54b2ca25-4afb-4bff-91d4-045daf8eb7ad · outbound

This paper cites EVA-CLIP: Improved Training Techniques for CLIP at Scale.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning EVA-CLIP: Improved Training Techniques for CLIP at Scale

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.495100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.495100Z digest=sha256:36e4c0bed3e231658d34f667ceb0d8ba4568e8f226e9f4023458e8b257225508

Observation 99d81053-2031-459e-a814-9b90d8503bfb · outbound

This paper cites L.; and Parikh, D.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning L.; and Parikh, D

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:47:52.985218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.498807Z digest=sha256:fb71c7298091d8c84071698344ad10fb1356861634827a0077ded1a4ce0c17a4

Observation 049e45b1-06fd-41c7-9935-b80f1758c45c · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:47:52.974325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.502607Z digest=sha256:86995dd634585c775b9433fd7e37389ec369c39b41800b36c2ae7485a82261f3

Observation cda9c693-6e3a-4b49-84b7-35ea6863a103 · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:47:52.963583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.506044Z digest=sha256:d0dcea5c4074fa6261b4fad9fed29c804cc5688ac10ceb1c939de66fc72411e3

Observation e4514247-2ebd-4ed2-bc74-e7053b154cfd · outbound

This paper cites CogVLM: Visual Expert for Pretrained Language Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning CogVLM: Visual Expert for Pretrained Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.509718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.509718Z digest=sha256:e548708b0ab3eb703a63b3a78c16ce70fcedf9d217f6d63ee25b51922fd57950

Observation d03c0505-7d79-4c8e-a6d6-0fb2112b0a24 · outbound

This paper cites SimVLM: Simple Visual Language Model Pretraining with Weak Supervision.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning SimVLM: Simple Visual Language Model Pretraining with Weak Supervision

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.513342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.513342Z digest=sha256:6353af099ae8f82e8a830500dd1f2074b8e528485dedbfd5c8fe43510a5ca1f8

Observation 9268e79b-dd81-410f-803d-edb3aa4c880f · outbound

This paper cites DualPrompt: Complementary Prompting for Rehearsal-free Continual Learning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning DualPrompt: Complementary Prompting for Rehearsal-free Continual Learning

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-11T00:47:52.615222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.517018Z digest=sha256:be934453b7bbd3b1f0dc27f4e570464a3ad70a8b00bccbdc08900692b0f5f42a

Observation 91256ac0-9dd6-41e1-9938-660fafd9122c · outbound

This paper cites Re-ViLM: Retrieval-Augmented Visual Language Model for Zero and Few-Shot Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Re-ViLM: Retrieval-Augmented Visual Language Model for Zero and Few-Shot Image Captioning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.520772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.520772Z digest=sha256:58a050f7731a05c2f59ec0b58afa07261e81560f117dbcd3acd564b9f4ae1702

Observation c8eb9d5e-d19c-4e4b-bd0c-15a77cdf443e · outbound

This paper cites MeaCap: Memory-Augmented Zero-shot Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning MeaCap: Memory-Augmented Zero-shot Image Captioning

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-11T00:47:52.590167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.524701Z digest=sha256:b52971945900c03c181227c01c4222f00a3ef887c47a58d315af73b736af7770

Observation a85aaa13-e464-4ee1-90cf-551ff68e6e21 · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning OPT: Open Pre-trained Transformer Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.528543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.528543Z digest=sha256:1ff4c295b7aa63634849bc6cfff19ee45a8ce3147fb802491baa3fb08407ff5e

Observation a0fcaac9-22ea-4392-b6f1-fb1b67373a17 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.532503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.532503Z digest=sha256:8e32813d1f3e10f64c0e0d68e7d02c30814bd982afd2ab8e1b515661b7e0480e

Pith citing papers

No inbound Pith citation observations are available.