Pith. sign in

Paper Citation Record · LEDGER

NPHardEval4V: Dynamic Evaluation of Large Vision-Language Models with Effects of Vision

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2403.01777.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.01777 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T00:36:22.162543Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T23:06:20.250022Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 559a3bdd-c09f-4c2a-8d9b-c345957d927e · inbound

AgentReview: Exploring Peer Review Dynamics with LLM Agents cites this paper.

AgentReview: Exploring Peer Review Dynamics with LLM Agents NPHardEval4V: Dynamic Evaluation of Large Vision-Language Models with Effects of Vision

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-23T23:38:37.063344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-23T23:38:28.005028Z digest=sha256:91417f4248ae6d903984c6071728f953968996f14a1ad33a12cc6f8de2cdbefb

Observation 97d54cd4-ff90-40b7-93e9-00ce9a5c2f64 · inbound

Evaluating Visual Mathematics in Multimodal LLMs: A Multilingual Benchmark Based on the Kangaroo Tests cites this paper.

Evaluating Visual Mathematics in Multimodal LLMs: A Multilingual Benchmark Based on the Kangaroo Tests NPHardEval4V: Dynamic Evaluation of Large Vision-Language Models with Effects of Vision

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:38:58.647891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:38:58.647891Z digest=sha256:09abfd03cccbfe56a695da0556265b364b35649388e955545feb5be49ab2a26a

Observation 686843be-bdbf-4b96-8ac9-ce58a65f5568 · inbound

From Pixels to Places: A Systematic Benchmark for Evaluating Image Geolocalization Ability in Large Language Models cites this paper.

From Pixels to Places: A Systematic Benchmark for Evaluating Image Geolocalization Ability in Large Language Models NPHardEval4V: Dynamic Evaluation of Large Vision-Language Models with Effects of Vision

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-21T23:50:47.252533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-21T23:49:59.795575Z digest=sha256:9297d572a1ea5b2c44601cab4e4da8ffcd124180797f6913a8d18d54f9c8a499

Observation c0742a6e-d65f-4615-8edc-5615664f1d9d · inbound

FATHOMS-RAG: A Framework for the Assessment of Thinking and Observation in Multimodal Systems that use Retrieval Augmented Generation cites this paper.

FATHOMS-RAG: A Framework for the Assessment of Thinking and Observation in Multimodal Systems that use Retrieval Augmented Generation NPHardEval4V: Dynamic Evaluation of Large Vision-Language Models with Effects of Vision

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-25T08:15:33.902797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-25T08:13:05.328746Z digest=sha256:6a8d59d350c1598daccf84f2b95b77a97fc30054bb353f848eff5dbd51b261d9

Observation a6b714fc-882f-4c84-a98d-4e7180312c6c · inbound

Do VLMs See What Sensors Feel? A Scalable Expert-Guided Design for Wheelchair Accessibility Assessment from Street View cites this paper.

Do VLMs See What Sensors Feel? A Scalable Expert-Guided Design for Wheelchair Accessibility Assessment from Street View NPHardEval4V: Dynamic Evaluation of Large Vision-Language Models with Effects of Vision

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T23:06:20.252581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-28T14:42:47.815192Z digest=sha256:f1dbf82f7eb21107f7e85ddb05848dbb58882fc39a1c8aa6c436371d827e9f05

Observation 7c5160d0-5f51-4c7f-994d-825831750f7c · inbound

ComplexityWorld: Benchmarking Vision-Language Models on Verifiable Visual Decision Making cites this paper.

ComplexityWorld: Benchmarking Vision-Language Models on Verifiable Visual Decision Making NPHardEval4V: Dynamic Evaluation of Large Vision-Language Models with Effects of Vision

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T00:36:22.162543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:36:22.162543Z digest=sha256:af5e85b302cc07e090323f0a3d4a13ffddb0d72c638688de6a66c3b42e1d0323