Pith. sign in

Paper Citation Record · LEDGER

PromptCap: Prompt-Guided Task-Aware Image Captioning

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2211.09699.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2211.09699 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T15:53:24.202171Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T00:35:10.154544Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 9ecfa34d-365a-404f-b7ee-4bcbe5998cb7 · inbound

REPLUG: Retrieval-Augmented Black-Box Language Models cites this paper.

REPLUG: Retrieval-Augmented Black-Box Language Models PromptCap: Prompt-Guided Task-Aware Image Captioning

Reference 73

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T12:41:54.142073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-17T12:41:53.833754Z digest=sha256:69d515dde36d0f21d97168e7f0789fc95910eac3b5fb36c5356bb0760bb9e16b

Observation f2a92179-2fc7-47e5-9ef9-d475416e8044 · inbound

ViperGPT: Visual Inference via Python Execution for Reasoning cites this paper.

ViperGPT: Visual Inference via Python Execution for Reasoning PromptCap: Prompt-Guided Task-Aware Image Captioning

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T18:15:14.634893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T18:15:14.382011Z digest=sha256:bc0546c4708de204533d90e48c896255518dc8b11c3ce87c188cf26ac1d61176

Observation fd4f8231-d625-484e-aed9-a7820f8d7232 · inbound

MM-REACT: Prompting ChatGPT for Multimodal Reasoning and Action cites this paper.

MM-REACT: Prompting ChatGPT for Multimodal Reasoning and Action PromptCap: Prompt-Guided Task-Aware Image Captioning

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T01:17:58.759014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T01:17:58.678036Z digest=sha256:274c828df7376e395b3a0654bcb12c9abed40c8f03432fc6a3aae4df5cb9bafc

Observation 9a16ecd1-667b-4023-b391-602484d52b8b · inbound

BLINK: Multimodal Large Language Models Can See but Not Perceive cites this paper.

BLINK: Multimodal Large Language Models Can See but Not Perceive PromptCap: Prompt-Guided Task-Aware Image Captioning

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T20:18:15.672292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T20:18:15.439163Z digest=sha256:ea86984d0a2c080c7285a16b1c9bac3b614fd8cfe6ae49d74bd3527ff14973f1

Observation b7089b11-485d-499a-95c0-5986a3eb4d3e · inbound

Prompt-Driven Continual Graph Learning cites this paper.

Prompt-Driven Continual Graph Learning PromptCap: Prompt-Guided Task-Aware Image Captioning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T15:53:24.202171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:53:24.202171Z digest=sha256:49a40e3c339d8cd0ea7ef74a092370bedf216e919d20657c23e3a58f4accdc7f

Observation fd3227d4-e8d7-4861-9073-fe3ae9c5e2cf · inbound

GC-KBVQA: A New Four-Stage Framework for Enhancing Knowledge Based Visual Question Answering Performance cites this paper.

GC-KBVQA: A New Four-Stage Framework for Enhancing Knowledge Based Visual Question Answering Performance PromptCap: Prompt-Guided Task-Aware Image Captioning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:43.951510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:21:43.951510Z digest=sha256:bae06b3421de010a807d1cb991900769053b78eb58a3ad252d4c7b8fc4d6ee5c

Observation d55fa72d-bf2f-4a12-959e-5e0f50f827f7 · inbound

MMIG-Bench: Towards Comprehensive and Explainable Evaluation of Multi-Modal Image Generation Models cites this paper.

MMIG-Bench: Towards Comprehensive and Explainable Evaluation of Multi-Modal Image Generation Models PromptCap: Prompt-Guided Task-Aware Image Captioning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:17:33.369845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:17:33.369845Z digest=sha256:67bb1106bd29537b950cb32075a367d5ea29d0d2ffd0a3bd620c8c6fa9d7861c

Observation fc7e8496-8281-43bf-bcc8-769778750263 · inbound

TableVision: A Large-Scale Benchmark for Spatially Grounded Reasoning over Complex Hierarchical Tables cites this paper.

TableVision: A Large-Scale Benchmark for Spatially Grounded Reasoning over Complex Hierarchical Tables PromptCap: Prompt-Guided Task-Aware Image Captioning

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T17:23:02.401153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T17:20:28.036531Z digest=sha256:e52aee9d23aa2f7fef0f7d8607c0dc7d0ba814541a1b352989a4325a229caf74

Observation dbbb60ad-08af-4e87-9db6-b671c70de804 · inbound

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning cites this paper.

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning PromptCap: Prompt-Guided Task-Aware Image Captioning

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T08:11:04.452618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T16:46:36.010169Z digest=sha256:411c9b7a5b899d11c8343842639c28601b5f3d7c706002a42a8d308dfd986f78

Observation 045090c8-b373-418c-87c2-d776aab5cb79 · inbound

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning cites this paper.

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning PromptCap: Prompt-Guided Task-Aware Image Captioning

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:01:26.873090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T04:38:06.774877Z digest=sha256:aa3c16c314b4c0f8e914ffbc50f4d1bd88d56cad8d3fef66b8013693d16fde76

Observation 44b1e129-d23c-4b4a-adaf-399b9e26d85a · inbound

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning cites this paper.

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning PromptCap: Prompt-Guided Task-Aware Image Captioning

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:27:29.401352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T07:26:59.917150Z digest=sha256:7e3c3968083531329d31714013693b7ddfc35a6d80854835a04e43057f5b5d46

Observation 89982f76-0757-4fac-b735-22112e3c952e · inbound

GEASS: Gated Evidence-Adaptive Selective Caption Trust for Vision-Language Models cites this paper.

GEASS: Gated Evidence-Adaptive Selective Caption Trust for Vision-Language Models PromptCap: Prompt-Guided Task-Aware Image Captioning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-21T00:43:53.579809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T00:41:58.198403Z digest=sha256:6011fe644eae6dbb16040f6abd2347d70c440a9c0c287a86489b4fc7131f6ec7

Observation b6103e5e-9a70-4702-bbf2-0ef7b5bd78bd · inbound

GEASS: Gated Evidence-Adaptive Selective Caption Trust for Vision-Language Models cites this paper.

GEASS: Gated Evidence-Adaptive Selective Caption Trust for Vision-Language Models PromptCap: Prompt-Guided Task-Aware Image Captioning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-01T00:35:10.156265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-01T00:32:49.666202Z digest=sha256:45ea6ff9df898c0c2fc5c5956bd2c4ef5f50ef3c2bd34eab56b8146d66089e2e

Observation 4d739551-ebd8-4a51-b91b-f911deee577d · inbound

Salient Knowledge Pathways: Sparse Cross-Modal Routing for Efficient Knowledge-Intensive Multimodal Question Answering cites this paper.

Salient Knowledge Pathways: Sparse Cross-Modal Routing for Efficient Knowledge-Intensive Multimodal Question Answering PromptCap: Prompt-Guided Task-Aware Image Captioning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T02:33:36.014715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:33:36.014715Z digest=sha256:415a5b22e466cedff20dd0b206d10c7845271c472a5cb802bdf1a598e6aa0d29

Observation 98f682b3-5e4a-4352-978e-dfc17d2b0022 · inbound

Salient Knowledge Pathways: Sparse Cross-Modal Routing for Efficient Knowledge-Intensive Multimodal Question Answering cites this paper.

Salient Knowledge Pathways: Sparse Cross-Modal Routing for Efficient Knowledge-Intensive Multimodal Question Answering PromptCap: Prompt-Guided Task-Aware Image Captioning

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-01T02:33:36.018539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:33:36.018539Z digest=sha256:30a26fe281e6810a415384df2a9213cb35bae0e42eb1468d053cc6c4f8b92ba3

Observation 1f83d1d8-d99a-4b81-a0da-95d0766194d8 · inbound

VideoArgus: Agentic Rubric-Grounded Unified Evaluation for Video Generation and Editing cites this paper.

VideoArgus: Agentic Rubric-Grounded Unified Evaluation for Video Generation and Editing PromptCap: Prompt-Guided Task-Aware Image Captioning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T12:16:44.671113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:16:44.671113Z digest=sha256:b9ed24accbb7e28e52f336602f97e71040dee0d53c88ab902daf7eef1713a545