Pith. sign in

Paper Citation Record · LEDGER

VisText: A Benchmark for Semantically Rich Chart Captioning

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2307.05356.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.05356 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T11:47:45.080370Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T17:53:46.746830Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7a7cd7fd-8a48-4bd7-9b16-11f97db346c7 · inbound

Otter: A Multi-Modal Model with In-Context Instruction Tuning cites this paper.

Otter: A Multi-Modal Model with In-Context Instruction Tuning VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 80

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:43:47.859889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T02:43:47.775691Z digest=sha256:3f8fadfe9c71f7edc37dc72de4c5f207726414b531387af1f4547eddf0cce949

Observation 410ddfd0-f1b3-4ba7-a7d1-86983ebf87a1 · inbound

Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning cites this paper.

Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T17:34:56.974820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T17:34:56.836034Z digest=sha256:66afb17daff765ca9cfeabf38b27e19c7d1d32b6ace9ba9ace0b938731585849

Observation 2fd067ba-9792-4da9-8622-ec2eb13d51dc · inbound

Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs cites this paper.

Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 125

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T00:05:03.720092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T00:05:03.547664Z digest=sha256:bf8d414e269ea5d4f6aeeeba8e49734db51880ec0ef8518db3cecba42b7fdac6

Observation 02c29023-905d-4a25-ac3c-be39f194f06a · inbound

Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling cites this paper.

Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 227

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T13:23:58.032587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T13:23:57.588851Z digest=sha256:d104e0547741a3b43e6073a7557f85491ed96d9615864c9a01e00f4c4148a63a

Observation 8516a13f-ee44-4b43-9ecb-9822d129ab35 · inbound

MetaMorph: Multimodal Understanding and Generation via Instruction Tuning cites this paper.

MetaMorph: Multimodal Understanding and Generation via Instruction Tuning VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 139

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T07:51:13.239657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-17T07:51:12.953777Z digest=sha256:5c71e4bc38bed17d999a266fd6d837de40ec1939a05f37f28f969dc1aec40e25

Observation c72e89a8-60a6-4c19-b4c0-9cade7122bff · inbound

Pluto: Authoring Semantically Aligned Text and Charts for Data-Driven Communication cites this paper.

Pluto: Authoring Semantically Aligned Text and Charts for Data-Driven Communication VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-08T11:47:45.080370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:47:45.080370Z digest=sha256:12162792fc7170dbde8d00603a5e76755269d579a1704bb1a3e10fe92628789f

Observation b15b11c2-bddf-4362-8445-2ccf572f18de · inbound

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding cites this paper.

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:52:01.842766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T19:49:00.961388Z digest=sha256:eb90040caa756784d6e5ea46779c9ae930018556345d4f30ac09b80dcd46cdc0

Observation fee02811-32cc-4cbf-8c58-30ea7bd237dd · inbound

Chart-to-Experience: Benchmarking Multimodal LLMs for Predicting Experiential Impact of Charts cites this paper.

Chart-to-Experience: Benchmarking Multimodal LLMs for Predicting Experiential Impact of Charts VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:53:09.700402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:53:09.700402Z digest=sha256:2932780b92391c9c4cce0941030804ada985e24b4090a9678b283a9f3274dfd6

Observation bf171219-4f17-444c-862b-8a273c2bd491 · inbound

ChartLens: Fine-grained Visual Attribution in Charts cites this paper.

ChartLens: Fine-grained Visual Attribution in Charts VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:19:24.218373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:19:24.218373Z digest=sha256:38a177c8bb0d2772c64466b259e97a20766784a44641716dbd936bdfcc94a81f

Observation 40e0ab61-176a-48d7-ad58-facc1c2cd120 · inbound

Boosting Document Parsing Efficiency and Performance with Coarse-to-Fine Visual Processing cites this paper.

Boosting Document Parsing Efficiency and Performance with Coarse-to-Fine Visual Processing VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 48

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T00:28:23.540411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T00:25:19.782732Z digest=sha256:b2519b5100cc2210eff649debe35b9a5a35a62d1e47814717be2af9c12f6039e

Observation dfc42b26-6ab1-4da9-90d6-ff1ccf5656b0 · inbound

LocateAnything: Fast and High-Quality Vision-Language Grounding with Parallel Box Decoding cites this paper.

LocateAnything: Fast and High-Quality Vision-Language Grounding with Parallel Box Decoding VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T17:53:46.748360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T17:52:59.346637Z digest=sha256:23fe339fea767428d75fc19590933f2a73a05399cfec4f12557adfecf5d1b556

Observation 5f77c327-dc9a-4794-a710-ae630298d438 · inbound

ABot-OCR Technical Report cites this paper.

ABot-OCR Technical Report VisText: A Benchmark for Semantically Rich Chart Captioning

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:33:28.135686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T13:29:17.221676Z digest=sha256:394c2c474fbbeb6edd889bb35739eec49b5035d54bb3908a7cbaaf579fa79644