Pith. sign in

Paper Citation Record · LEDGER

Visual In-Context Learning for Large Vision-Language Models

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2402.11574.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.11574 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T21:36:09.484285Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T08:13:15.073925Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 570f10ac-3dca-4739-a70c-b20c8b8f04e8 · inbound

MageBench: Bridging Large Multimodal Models to Agents cites this paper.

MageBench: Bridging Large Multimodal Models to Agents Visual In-Context Learning for Large Vision-Language Models

Reference 113

Resolution
unresolved
no resolver link, observed 2026-08-11T21:36:09.484285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:36:09.484285Z digest=sha256:6d13ac11c7d6cd42f6118adb30fe07cd98a4bdff2525a4845f1808dca7a0b337

Observation da0910bb-19b9-4471-9ef9-71335f49ecbd · inbound

Draw ALL Your Imagine: A Holistic Benchmark and Agent Framework for Complex Instruction-based Image Generation cites this paper.

Draw ALL Your Imagine: A Holistic Benchmark and Agent Framework for Complex Instruction-based Image Generation Visual In-Context Learning for Large Vision-Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:18:34.987500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:18:34.987500Z digest=sha256:1733fbfd9c3e3603803c6f65e3a9dbccd31ff45fbae993c8bd6fcbeb29b17c75

Observation 11bfecb5-6557-496a-ba71-7a345bf79b2f · inbound

ComplexBench-Edit: Benchmarking Complex Instruction-Driven Image Editing via Compositional Dependencies cites this paper.

ComplexBench-Edit: Benchmarking Complex Instruction-Driven Image Editing via Compositional Dependencies Visual In-Context Learning for Large Vision-Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T00:40:22.398794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:40:22.398794Z digest=sha256:159072fb6dd87f8fcef6778d291c4a8ded0b28eaba0096e2d684390681899b83

Observation 02829d29-b733-492b-a1f6-e66d5a2936d7 · inbound

Online In-Context Distillation for Low-Resource Vision Language Models cites this paper.

Online In-Context Distillation for Low-Resource Vision Language Models Visual In-Context Learning for Large Vision-Language Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-18T05:40:55.977866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T05:36:38.735914Z digest=sha256:4cbe7134c1f93437e582604c3be6837dfef845a0fd8351dc15128f7a3fd8baba

Observation 4624d63d-39c9-47ec-86ac-47d5c4ceab53 · inbound

Learning to Select Visual In-Context Demonstrations cites this paper.

Learning to Select Visual In-Context Demonstrations Visual In-Context Learning for Large Vision-Language Models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-04T05:46:09.450965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:46:09.450965Z digest=sha256:33fb9c43520701952adf344814d578eb3774aad5eaaa11a3fbadb6e512c830e7

Observation b401ebef-7684-4fbe-8e19-e56b5142d7b8 · inbound

DocRetriever: A Plug-and-Play Framework for Multimodal Document Retrieval with Comprehensive Benchmark cites this paper.

DocRetriever: A Plug-and-Play Framework for Multimodal Document Retrieval with Comprehensive Benchmark Visual In-Context Learning for Large Vision-Language Models

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:13:15.075880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:09:41.068000Z digest=sha256:af8274a28ec208f09f4f93eb1b69becf5ba2b47e52ecb3ce7d2eba0ef4a823d8