Pith. sign in

Paper Citation Record · LEDGER

Filter Images First, Generate Instructions Later: Pre-Instruction Data Selection for Visual Instruction Tuning

As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2503.07591.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.07591 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T19:40:42.676920Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T00:29:15.781730Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a8049146-2ddc-4b79-b76d-c852fad4b72a · inbound

MLAN: Language-Based Instruction Tuning Preserves and Transfers Knowledge in Multimodal Language Models cites this paper.

MLAN: Language-Based Instruction Tuning Preserves and Transfers Knowledge in Multimodal Language Models Filter Images First, Generate Instructions Later: Pre-Instruction Data Selection for Visual Instruction Tuning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-12T19:40:42.676920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T19:40:42.676920Z digest=sha256:76248ef560e37a1af4b49003f0b952268a8fdb804d57d9350ff436295274cbde

Observation ccab11fa-0912-4f2e-bc9f-632d3cd2331f · inbound

Certainty and Uncertainty Guided Active Domain Adaptation cites this paper.

Certainty and Uncertainty Guided Active Domain Adaptation Filter Images First, Generate Instructions Later: Pre-Instruction Data Selection for Visual Instruction Tuning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:17:47.315703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:17:47.315703Z digest=sha256:215b9356faca0e1059aa865dafa7f9c2d53f8411ae882345cc13941030747330

Observation f97533ce-a39f-442b-9c12-cfa9776916e3 · inbound

VisNec: Measuring and Leveraging Visual Necessity for Multimodal Instruction Tuning cites this paper.

VisNec: Measuring and Leveraging Visual Necessity for Multimodal Instruction Tuning Filter Images First, Generate Instructions Later: Pre-Instruction Data Selection for Visual Instruction Tuning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T19:46:20.848412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:46:20.848412Z digest=sha256:98d7d76ef4e562665cf0550ba22ff2db1a977716d915dee61ead33e4ac5a856a

Observation 2f95f597-39a6-4c94-b6b6-c44dd66c3cc9 · inbound

Once-For-All: A Train-Once and Select-Anytime Framework for Multimodal Instruction Tuning cites this paper.

Once-For-All: A Train-Once and Select-Anytime Framework for Multimodal Instruction Tuning Filter Images First, Generate Instructions Later: Pre-Instruction Data Selection for Visual Instruction Tuning

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T18:33:50.129745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T18:32:53.558370Z digest=sha256:cde0f2eb028a1210aa54c83169b97084730f349d87473971b8030f20c5121c40

Observation 096bac68-9c4c-46ce-a63d-9aba7e39b33c · inbound

Seeing Before Reasoning: Decoupling Perception and Reasoning for Shortcut-Resilient Multimodal On-Policy Self-Distillation cites this paper.

Seeing Before Reasoning: Decoupling Perception and Reasoning for Shortcut-Resilient Multimodal On-Policy Self-Distillation Filter Images First, Generate Instructions Later: Pre-Instruction Data Selection for Visual Instruction Tuning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:29:15.785087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-26T21:12:26.064012Z digest=sha256:06b25bcdab657ff1a515ae855299c72e16fe6edb1184434ef7ddf5bc583c69c4