Pith. sign in

Paper Citation Record · LEDGER

NPHardEval4V: Dynamic Evaluation of Large Vision-Language Models with Effects of Vision

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2403.01777.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.01777 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:38:58.647891Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T23:06:20.250022Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 559a3bdd-c09f-4c2a-8d9b-c345957d927e · inbound

AgentReview: Exploring Peer Review Dynamics with LLM Agents cites this paper.

AgentReview: Exploring Peer Review Dynamics with LLM Agents NPHardEval4V: Dynamic Evaluation of Large Vision-Language Models with Effects of Vision

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-23T23:38:37.063344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T23:38:28.005028Z digest=sha256:b7defbe8b2b23536207b28b1ba346732b513ade6fc87696cbba1ae9f9c4a4eeb

Observation 97d54cd4-ff90-40b7-93e9-00ce9a5c2f64 · inbound

Evaluating Visual Mathematics in Multimodal LLMs: A Multilingual Benchmark Based on the Kangaroo Tests cites this paper.

Evaluating Visual Mathematics in Multimodal LLMs: A Multilingual Benchmark Based on the Kangaroo Tests NPHardEval4V: Dynamic Evaluation of Large Vision-Language Models with Effects of Vision

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:38:58.647891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:38:58.647891Z digest=sha256:9dc80e8870a2773594115341701cb23a7ff66629fdbc54d96090f6246a6ed706

Observation 686843be-bdbf-4b96-8ac9-ce58a65f5568 · inbound

From Pixels to Places: A Systematic Benchmark for Evaluating Image Geolocalization Ability in Large Language Models cites this paper.

From Pixels to Places: A Systematic Benchmark for Evaluating Image Geolocalization Ability in Large Language Models NPHardEval4V: Dynamic Evaluation of Large Vision-Language Models with Effects of Vision

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-21T23:50:47.252533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T23:49:59.795575Z digest=sha256:8b91a2ed949497af299c7c1cd2f2eb039b9858daadfc17ba3cc215d92f60b95c

Observation c0742a6e-d65f-4615-8edc-5615664f1d9d · inbound

FATHOMS-RAG: A Framework for the Assessment of Thinking and Observation in Multimodal Systems that use Retrieval Augmented Generation cites this paper.

FATHOMS-RAG: A Framework for the Assessment of Thinking and Observation in Multimodal Systems that use Retrieval Augmented Generation NPHardEval4V: Dynamic Evaluation of Large Vision-Language Models with Effects of Vision

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-25T08:15:33.902797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T08:13:05.328746Z digest=sha256:2147be854d74b18b0408c26ea0955f7a456f290bc05f454eed0e850505eea1a0

Observation a6b714fc-882f-4c84-a98d-4e7180312c6c · inbound

Do VLMs See What Sensors Feel? A Scalable Expert-Guided Design for Wheelchair Accessibility Assessment from Street View cites this paper.

Do VLMs See What Sensors Feel? A Scalable Expert-Guided Design for Wheelchair Accessibility Assessment from Street View NPHardEval4V: Dynamic Evaluation of Large Vision-Language Models with Effects of Vision

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T23:06:20.252581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T14:42:47.815192Z digest=sha256:80bb8fbb3e0c897b5a6e1a2546e1a1a8f8eb7fb7143b2c026d4f817647af9b1c