Pith. sign in

Paper Citation Record · LEDGER

ViP-LLaVA: Making Large Multimodal Models Understand Arbitrary Visual Prompts

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 3 inbound Pith citation observations for arXiv:2312.00784.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.00784 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 3 of 3 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:45:53.573721Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T18:00:43.120575Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation cf14ef7f-11e3-4a74-9ed1-e4f5ea5cb9e4 · inbound

LPOI: Listwise Preference Optimization for Vision Language Models cites this paper.

LPOI: Listwise Preference Optimization for Vision Language Models ViP-LLaVA: Making Large Multimodal Models Understand Arbitrary Visual Prompts

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:45:53.573721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:45:53.573721Z digest=sha256:b5da8330edc65dec17bf943ed3a8fc3d0ac75897d69a8589952f2e1ac466955a

Observation f2508823-1c46-4071-bbf4-ea4e41a2ca42 · inbound

Alignment Quality Index (AQI) : Beyond Refusals: AQI as an Intrinsic Alignment Diagnostic via Latent Geometry, Cluster Divergence, and Layer wise Pooled Representations cites this paper.

Alignment Quality Index (AQI) : Beyond Refusals: AQI as an Intrinsic Alignment Diagnostic via Latent Geometry, Cluster Divergence, and Layer wise Pooled Representations ViP-LLaVA: Making Large Multimodal Models Understand Arbitrary Visual Prompts

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:32.709825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:32.709825Z digest=sha256:eed3a40c8e2f4cd70bb330250b32c740fc994d71474f2cb58eff6790e055744f

Observation f98d3dc1-595c-4b73-a7b4-724dd6aabe57 · inbound

DesignCLIP: Multimodal Learning with CLIP for Design Patent Understanding cites this paper.

DesignCLIP: Multimodal Learning with CLIP for Design Patent Understanding ViP-LLaVA: Making Large Multimodal Models Understand Arbitrary Visual Prompts

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T18:00:43.188358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T18:00:33.177933Z digest=sha256:d757a459d5c48c46549a6aeaf23a637a744c676893b74081dd4175b1032ba3b0