Pith. sign in

Paper Citation Record · LEDGER

ViP-LLaVA: Making Large Multimodal Models Understand Arbitrary Visual Prompts

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 3 inbound Pith citation observations for arXiv:2312.00784.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.00784 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 3 of 3 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:45:53.573721Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T18:00:43.120575Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation cf14ef7f-11e3-4a74-9ed1-e4f5ea5cb9e4 · inbound

LPOI: Listwise Preference Optimization for Vision Language Models cites this paper.

LPOI: Listwise Preference Optimization for Vision Language Models ViP-LLaVA: Making Large Multimodal Models Understand Arbitrary Visual Prompts

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:45:53.573721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:45:53.573721Z digest=sha256:5de0e16da97f7319cace4b5af12ea87e8fe6572cee39d0abf46b0c8d12b31469

Observation f2508823-1c46-4071-bbf4-ea4e41a2ca42 · inbound

Alignment Quality Index (AQI) : Beyond Refusals: AQI as an Intrinsic Alignment Diagnostic via Latent Geometry, Cluster Divergence, and Layer wise Pooled Representations cites this paper.

Alignment Quality Index (AQI) : Beyond Refusals: AQI as an Intrinsic Alignment Diagnostic via Latent Geometry, Cluster Divergence, and Layer wise Pooled Representations ViP-LLaVA: Making Large Multimodal Models Understand Arbitrary Visual Prompts

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:32.709825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:32.709825Z digest=sha256:788301b4457f69af1f76c9fd730a7f4bfe5f06197a680bf8b6ac063d791bc0ce

Observation f98d3dc1-595c-4b73-a7b4-724dd6aabe57 · inbound

DesignCLIP: Multimodal Learning with CLIP for Design Patent Understanding cites this paper.

DesignCLIP: Multimodal Learning with CLIP for Design Patent Understanding ViP-LLaVA: Making Large Multimodal Models Understand Arbitrary Visual Prompts

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T18:00:43.188358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:00:33.177933Z digest=sha256:2ff79eead01301632d4de9fd2c8e2810f6c3477e129011699fe46132b05e9517