Pith. sign in

Paper Citation Record · LEDGER

Scaling Text-Rich Image Understanding via Code-Guided Synthetic Multimodal Data Generation

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2502.14846.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.14846 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:49:28.118433Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-05T12:30:59.857969Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 290d4613-fbd5-4354-b505-ccf4a1af24e1 · inbound

Prioritizing Image-Related Tokens Enhances Vision-Language Pre-Training cites this paper.

Prioritizing Image-Related Tokens Enhances Vision-Language Pre-Training Scaling Text-Rich Image Understanding via Code-Guided Synthetic Multimodal Data Generation

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-15T21:49:28.118433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:49:28.118433Z digest=sha256:2bc743cbdd34c631f0ffc326500c65ea65064e155e63d80b5bf224de9e5f6338

Observation 6a0ab228-bffd-4b7a-b5f4-ca9c246f834b · inbound

GoVector: An I/O-Efficient Caching Strategy for High-Dimensional Vector Nearest Neighbor Search cites this paper.

GoVector: An I/O-Efficient Caching Strategy for High-Dimensional Vector Nearest Neighbor Search Scaling Text-Rich Image Understanding via Code-Guided Synthetic Multimodal Data Generation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T17:45:21.174003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:45:21.174003Z digest=sha256:bbb743e5b38adce56ec32781bb4476ec904983683228d37f32b2f4394ebf35cc

Observation 9007a525-5709-4b49-9068-3908ac91f327 · inbound

ChartVerse: Scaling Chart Reasoning via Reliable Programmatic Synthesis from Scratch cites this paper.

ChartVerse: Scaling Chart Reasoning via Reliable Programmatic Synthesis from Scratch Scaling Text-Rich Image Understanding via Code-Guided Synthetic Multimodal Data Generation

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-16T13:12:54.846538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T13:12:01.889341Z digest=sha256:9f4d41b45f75d289c6a5b9785d351bbc81886115eae9af01b26fb27e67ac2bc3

Observation 72c80788-53bd-4f1f-991f-976f80a30ae4 · inbound

Multilingual Training and Evaluation Resources for Vision-Language Models cites this paper.

Multilingual Training and Evaluation Resources for Vision-Language Models Scaling Text-Rich Image Understanding via Code-Guided Synthetic Multimodal Data Generation

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T09:38:43.169078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T05:11:00.461167Z digest=sha256:2fab7a4eb9938bfcb9d0928f8e11d6747976a684cda43cd28bcbff3683e4e051

Observation 143edbfc-d1ba-4bfd-be0c-1568758d717c · inbound

Multilingual Training and Evaluation Resources for Vision-Language Models cites this paper.

Multilingual Training and Evaluation Resources for Vision-Language Models Scaling Text-Rich Image Understanding via Code-Guided Synthetic Multimodal Data Generation

Reference 42

Resolution
metadata mismatch
local_arxiv, observed 2026-07-05T12:30:59.860761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-05T12:25:00.027624Z digest=sha256:d78edf95afafab127e859d3904fbe3d1d7c064c8c0211bce88c8f38e4002de3b

Observation d3427634-5aa2-4cd0-b22f-5cd624a4adf1 · inbound

Erase Persona, Forget Lore: Benchmarking Multimodal Copyright Unlearning in Large Vision Language Models cites this paper.

Erase Persona, Forget Lore: Benchmarking Multimodal Copyright Unlearning in Large Vision Language Models Scaling Text-Rich Image Understanding via Code-Guided Synthetic Multimodal Data Generation

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:11:18.039930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-07T17:58:17.879345Z digest=sha256:4fa37cb3c4326af152b4e20992e9c1fc8888c74c39287f71c053e501dd38de2b

Observation e0470cf6-f6bd-4403-9e6b-377d17a1e59c · inbound

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone cites this paper.

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone Scaling Text-Rich Image Understanding via Code-Guided Synthetic Multimodal Data Generation

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:57:09.498413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-13T02:52:43.674969Z digest=sha256:5fd25ee166cd48af5c6b1961928d3645e221a647ea1b8ba32c8f88f4ecfb7935

Observation 8c94b694-d13a-4be5-b797-71308ea3926f · inbound

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone cites this paper.

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone Scaling Text-Rich Image Understanding via Code-Guided Synthetic Multimodal Data Generation

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:29:28.675611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-14T21:28:37.680681Z digest=sha256:5376caef97891428da4c21d86b1c60f437a5f06e22a361e0edcf794b85db6074

Observation 3e2a6aa2-ea7c-4d47-887a-04be621d19e7 · inbound

Zone of Proximal Policy Optimization: Teacher in Prompts, Not Gradients cites this paper.

Zone of Proximal Policy Optimization: Teacher in Prompts, Not Gradients Scaling Text-Rich Image Understanding via Code-Guided Synthetic Multimodal Data Generation

Reference 102

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:48:56.137363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T01:08:52.981296Z digest=sha256:a5490743126cc6cee5926679c8315564281901865770f766d9c245f551585b96

Observation a5c2662f-3ec8-487d-93c5-f8aeb59384c0 · inbound

Program-as-Weights: A Programming Paradigm for Fuzzy Functions cites this paper.

Program-as-Weights: A Programming Paradigm for Fuzzy Functions Scaling Text-Rich Image Understanding via Code-Guided Synthetic Multimodal Data Generation

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T16:18:37.234379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-07-03T16:17:59.522449Z digest=sha256:dd4a369b97e45b8eba01b54901c49749b4b75ff8dd13880b98e3ab8a98e8e70f