Pith. sign in

Paper Citation Record · LEDGER

Quilt-LLaVA: Visual Instruction Tuning by Extracting Localized Narratives from Open-Source Histopathology Videos

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2312.04746.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.04746 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T14:21:46.598304Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T09:55:40.926803Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 88167289-04e2-459e-a19f-024c2704ddfb · inbound

PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering cites this paper.

PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering Quilt-LLaVA: Visual Instruction Tuning by Extracting Localized Narratives from Open-Source Histopathology Videos

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-15T23:08:21.779850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-15T23:08:21.673251Z digest=sha256:f0032420489e47ba59a6c9828d56597a8a8f11a21ac8ddb31e4cd1dea343c5e6

Observation 0a345d8a-f109-4f4d-a1e3-96854d0a3ebf · inbound

CPath-Omni: A Unified Multimodal Foundation Model for Patch and Whole Slide Image Analysis in Computational Pathology cites this paper.

CPath-Omni: A Unified Multimodal Foundation Model for Patch and Whole Slide Image Analysis in Computational Pathology Quilt-LLaVA: Visual Instruction Tuning by Extracting Localized Narratives from Open-Source Histopathology Videos

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T14:21:46.598304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:21:46.598304Z digest=sha256:0a3b9fe43f32bfbd18c727dda14a8c6a9ed4a8af940f83f7ede6b19f590fbf9d

Observation de8d5563-737d-4acf-b3df-d12cecef4cdd · inbound

Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning cites this paper.

Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning Quilt-LLaVA: Visual Instruction Tuning by Extracting Localized Narratives from Open-Source Histopathology Videos

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T23:32:35.294994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:32:35.294994Z digest=sha256:95d2744495b88018fac29c5f6ead8de9c46946ce77a78325a7cb820775dbebc1

Observation 74dc3aa4-1ade-49da-87e3-c200c9605f5a · inbound

MEDSYN: Benchmarking Multi-EviDence SYNthesis in Complex Clinical Cases for Multimodal Large Language Models cites this paper.

MEDSYN: Benchmarking Multi-EviDence SYNthesis in Complex Clinical Cases for Multimodal Large Language Models Quilt-LLaVA: Visual Instruction Tuning by Extracting Localized Narratives from Open-Source Histopathology Videos

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-15T19:36:32.801465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-15T19:34:42.244909Z digest=sha256:a30308ccd288065e45dc9d9888d52c036ed0565e704acdadb31342ed8fde2c3a

Observation bef98d89-4b53-4285-9591-c3ef4a30bfed · inbound

Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning cites this paper.

Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning Quilt-LLaVA: Visual Instruction Tuning by Extracting Localized Narratives from Open-Source Histopathology Videos

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:55:40.928780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-01T06:03:12.182207Z digest=sha256:a463a28f7cc2cd2a4ecba0278ec8ad3f646d19d99351e5f3068daa390436497d