Pith. sign in

Paper Citation Record · LEDGER

Good at captioning, bad at counting: Benchmarking GPT-4V on Earth observation data

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2401.17600.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.17600 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:35:41.868792Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-15T21:20:59.417576Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 58606be8-348a-477c-ae5c-a3004d93affe · inbound

Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models cites this paper.

Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models Good at captioning, bad at counting: Benchmarking GPT-4V on Earth observation data

Reference 184

Resolution
verified exact
arxiv_id, observed 2026-05-15T21:20:59.418991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T21:20:59.128986Z digest=sha256:e06bf84f6c21af043400535dec2821efa9d86d09305be207fc364ee6bf1880a1

Observation 910cafb6-7591-4ee6-b69e-bc74298e466a · inbound

BYO-Eval: Build Your Own Dataset for Fine-Grained Visual Assessment of Multimodal Language Models cites this paper.

BYO-Eval: Build Your Own Dataset for Fine-Grained Visual Assessment of Multimodal Language Models Good at captioning, bad at counting: Benchmarking GPT-4V on Earth observation data

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:41.868792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:35:41.868792Z digest=sha256:0aab55e20096ef1ef76f9b4a28dc46de8eb9738dabea4f3c78fad4a39375bf55

Observation 6d95f1cf-d097-4862-953f-7717116ca134 · inbound

Domain Specific Benchmarks for Evaluating Multimodal Large Language Models cites this paper.

Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Good at captioning, bad at counting: Benchmarking GPT-4V on Earth observation data

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T00:39:41.737914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:39:41.737914Z digest=sha256:8643fc28bddd81cbe3e1b364a801d387ff85cbe34342d67c7428387f27c7d180

Observation 699aaa9e-581e-4a0c-afd9-1752e25544cb · inbound

Pedestrian Intention Prediction via Vision-Language Foundation Models cites this paper.

Pedestrian Intention Prediction via Vision-Language Foundation Models Good at captioning, bad at counting: Benchmarking GPT-4V on Earth observation data

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T19:59:56.092349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:59:56.092349Z digest=sha256:9f40b89275ed4ad449a7c8683866d0af5e5ae88f8c8ba64b91d2bcdf244bdcb2

Observation 55ef12ed-f951-4cde-b56a-4d92b8b5c6d6 · inbound

Unlocking Multi-Spectral Data for Multi-Modal Models with Guided Inputs and Chain-of-Thought Reasoning cites this paper.

Unlocking Multi-Spectral Data for Multi-Modal Models with Guided Inputs and Chain-of-Thought Reasoning Good at captioning, bad at counting: Benchmarking GPT-4V on Earth observation data

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:24:47.547071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T00:05:15.384332Z digest=sha256:7aee1a2fb998a018434ec1b35d453da69bce8275b599c1a31133d1bb380b3637