Pith. sign in

Paper Citation Record · LEDGER

StableLLaVA: Enhanced Visual Instruction Tuning with Synthesized Image-Dialogue Data

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2308.10253.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.10253 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T19:51:43.490848Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-17T10:16:43.917557Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation fc66c79f-d895-4146-9bb9-ab2d894100b1 · inbound

HallusionBench: An Advanced Diagnostic Suite for Entangled Language Hallucination and Visual Illusion in Large Vision-Language Models cites this paper.

HallusionBench: An Advanced Diagnostic Suite for Entangled Language Hallucination and Visual Illusion in Large Vision-Language Models StableLLaVA: Enhanced Visual Instruction Tuning with Synthesized Image-Dialogue Data

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-17T01:22:04.276991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-17T01:22:04.035994Z digest=sha256:e1df8a223bf4b895a5ed6662f34a1f37f81f2630187e28fa86c182a4a351b88a

Observation 86880e42-d145-4175-86a2-1cfbdf6bacdb · inbound

AppAgent: Multimodal Agents as Smartphone Users cites this paper.

AppAgent: Multimodal Agents as Smartphone Users StableLLaVA: Enhanced Visual Instruction Tuning with Synthesized Image-Dialogue Data

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-17T10:16:43.920193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-17T10:16:43.364787Z digest=sha256:33896786394652fa8b15f1fe01c6ace14e9ec75dcb4cc0ac708974fae1ef2b0d

Observation 195e6fd4-f7b0-4573-967a-ec81e199c2de · inbound

LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations cites this paper.

LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations StableLLaVA: Enhanced Visual Instruction Tuning with Synthesized Image-Dialogue Data

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T19:51:43.490848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:51:43.490848Z digest=sha256:201997828fb312921cb9ad03aaa2a1a998a69e658e1559ea83d8cca41788cacc

Observation f47bf0db-c32c-4eb2-904f-1aedc998b5f0 · inbound

Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey cites this paper.

Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey StableLLaVA: Enhanced Visual Instruction Tuning with Synthesized Image-Dialogue Data

Reference 250

Resolution
unresolved
no resolver link, observed 2026-08-11T14:59:02.236974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:59:02.236974Z digest=sha256:6daa54c2f2e3c56779ab9453f25131ee7258299271e5abf68c1c1dd49f965f27

Observation a304fd03-3513-4f8d-b377-20c6df1a35f8 · inbound

Visual Large Language Models for Generalized and Specialized Applications cites this paper.

Visual Large Language Models for Generalized and Specialized Applications StableLLaVA: Enhanced Visual Instruction Tuning with Synthesized Image-Dialogue Data

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T22:08:09.080258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:08:09.080258Z digest=sha256:e54090d82afa404e8557e7fc80ad8d23ef96c6f3272e8dfcda28ed0de2d3217f

Observation f0a8bc6c-093a-47cc-b13c-4e056eb34284 · inbound

MTPChat: A Multimodal Time-Aware Persona Dataset for Conversational Agents cites this paper.

MTPChat: A Multimodal Time-Aware Persona Dataset for Conversational Agents StableLLaVA: Enhanced Visual Instruction Tuning with Synthesized Image-Dialogue Data

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T17:37:38.092080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:37:38.092080Z digest=sha256:1fcbd93fc5752fe66db548cdf8ba453c1779136c691265f330381bbe05485884

Observation e92f5ebe-1327-4337-bc68-0d2a152b7e54 · inbound

Advancing Multimodal LLMs by Large-Scale 3D Visual Instruction Dataset Generation cites this paper.

Advancing Multimodal LLMs by Large-Scale 3D Visual Instruction Dataset Generation StableLLaVA: Enhanced Visual Instruction Tuning with Synthesized Image-Dialogue Data

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T18:24:52.156204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:24:52.156204Z digest=sha256:65246e4d14438d55330cda8295018fad09c4f8c14d54dd8fc5f90791e6525efc