Pith. sign in

Paper Citation Record · LEDGER

VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2503.23064.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.23064 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:50:32.535373Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ffbe223e-d9c3-4ae9-9117-c5fa30c22be1 · inbound

Jigsaw-Puzzles: From Seeing to Understanding to Reasoning in Vision-Language Models cites this paper.

Jigsaw-Puzzles: From Seeing to Understanding to Reasoning in Vision-Language Models VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T13:50:32.535373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:50:32.535373Z digest=sha256:4d99b2ddba9f25ee2ac0729121184170d60366c9529c3bb68032ad3081c5a8df

Observation e694965d-e68e-4cd1-8aed-ce8277157796 · inbound

VisualSphinx: Large-Scale Synthetic Vision Logic Puzzles for RL cites this paper.

VisualSphinx: Large-Scale Synthetic Vision Logic Puzzles for RL VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:18.617428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:18.617428Z digest=sha256:f6c071f586f07759c475eafd6f74a60f784d3ca8f966b4703c7eb7f5e5746305

Observation 8d876b5e-022d-40bd-b173-4d0e16cc8068 · inbound

TextAtari: 100K Frames Game Playing with Language Agents cites this paper.

TextAtari: 100K Frames Game Playing with Language Agents VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T10:51:58.020886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:51:58.020886Z digest=sha256:b9b491790a890ae97374e4cb851b2975fda72ca0401a90deead51d7a3938df7f

Observation 820be703-276d-4e5e-a29d-d32f4f8c9e90 · inbound

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey cites this paper.

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:19.516932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:19.516932Z digest=sha256:41b7a138d677132773828a4d36a8a95d812a24b52a4e8fb105cabd012a38c1b8

Observation c4745894-f539-49e8-b423-da4d3c2e6961 · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:16:34.576642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T20:12:46.385646Z digest=sha256:721ccdb9ed2156d3df51d6c5b8ba20711753b570788c517dbde8b0cc2cd22ae6

Observation 5c555827-bff2-497e-a72c-51c0a2fe2ba9 · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:31:25.363121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T10:30:06.829915Z digest=sha256:e690e9b5e90386acdcb058983ad521decf07d4455d427ad99a7bfd8ad5584edf

Observation 7235af6b-ec94-4326-80c7-96d88d53c2e4 · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T22:00:08.370312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:00:08.370312Z digest=sha256:ef4f9b964f1770e06e7a8e62146a1d878a79b063714ecdafb5f2d990167f18cc

Observation f8ba29c2-58f6-4551-ba96-2858f788498f · inbound

Mind's Eye: A Benchmark of Visual Abstraction, Transformation and Composition for Multimodal LLMs cites this paper.

Mind's Eye: A Benchmark of Visual Abstraction, Transformation and Composition for Multimodal LLMs VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:23:37.242098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T09:18:33.234354Z digest=sha256:ce5cd7ba527a904ab85889d0dc43f196fdf5c75c3fd4febfd46ce782d94c5f75

Observation b3fe7372-8644-4f0a-857f-b28b22be425c · inbound

The Cartesian Shortcut: Re-evaluate Vision Reasoning in Polar Coordinate Space cites this paper.

The Cartesian Shortcut: Re-evaluate Vision Reasoning in Polar Coordinate Space VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:11:25.099997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T04:30:54.053958Z digest=sha256:6e70886b596551857cd975fef1d2954dde32f965fe269d7555fa8b8ea1281527

Observation a15b774b-cb81-4ad8-a409-51396aa81da1 · inbound

The Cartesian Shortcut: Re-evaluate Vision Reasoning in Polar Coordinate Space cites this paper.

The Cartesian Shortcut: Re-evaluate Vision Reasoning in Polar Coordinate Space VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:35:46.591007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T22:58:04.574536Z digest=sha256:32f62081c2096093a2e2f591d8b82cdd6fa470c5cb507ffe6a90f12e1c78d3be

Observation 4e369fa1-6a0e-4b11-a4a8-4a508d15ca4e · inbound

Do Vision-Language-Models show human-like logical problem-solving capability in point and click puzzle games? cites this paper.

Do Vision-Language-Models show human-like logical problem-solving capability in point and click puzzle games? VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T02:02:05.512354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-13T01:58:39.476408Z digest=sha256:ead9f962dfa5c7d0f4e58ebaf3055c3c918c165f7bee39a621ac027c321b30c7

Observation 98026f46-0f5a-4fa3-9851-1b2ecc01cdc5 · inbound

Do Vision-Language-Models show human-like logical problem-solving capability in point and click puzzle games? cites this paper.

Do Vision-Language-Models show human-like logical problem-solving capability in point and click puzzle games? VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:07:08.781223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-13T01:58:39.476408Z digest=sha256:ff76cf05dba85ffadac6305bfc35774c654e61502520396735d7c442867e0e33

Observation 05b8ec7d-02ef-49b8-bc20-4a3b338ace98 · inbound

Sci-Rho: A Multilingual Visually-Grounded Symbolic Benchmark for STEM Problems cites this paper.

Sci-Rho: A Multilingual Visually-Grounded Symbolic Benchmark for STEM Problems VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:22.848985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T20:11:02.445626Z digest=sha256:947871f4b6c187e9f51205600f4db8823ea678cd01c40b0094a30f2a22aaa9b9

Observation 31e4795b-8560-4d8a-9b08-2ad1f7908213 · inbound

Beyond the Current Observation: Evaluating Multimodal Large Language Models in Controllable Non-Markov Games cites this paper.

Beyond the Current Observation: Evaluating Multimodal Large Language Models in Controllable Non-Markov Games VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:19:13.644799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T21:17:02.332687Z digest=sha256:b1e252e2abb6db2a9e33a245143b1d6dbc6858f761803e2b5e10214164cd1a73

Observation 0114579a-b248-47e7-a0ab-997e05354579 · inbound

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models cites this paper.

ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T02:59:26.128499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T18:37:27.171215Z digest=sha256:590abc8e74bfa009da72cb6a22a84e8b1d91df5c779c57e7e7608bf703188240

Observation 54c48fe6-78a2-42b3-ac40-38996f921129 · inbound

JigShape: Evaluating Visual-Geometric Reasoning in VLMs through Jigsaw Puzzles cites this paper.

JigShape: Evaluating Visual-Geometric Reasoning in VLMs through Jigsaw Puzzles VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T03:39:41.733830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:39:41.733830Z digest=sha256:bac2866def294f113a32187271c816cf7feea7ee32ba7c6c19e136db3deca65f

Observation 96159599-af9e-4b36-88b6-34d8ad3f2012 · inbound

JigShape: Evaluating Visual-Geometric Reasoning in VLMs through Jigsaw Puzzles cites this paper.

JigShape: Evaluating Visual-Geometric Reasoning in VLMs through Jigsaw Puzzles VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T04:25:21.511122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T04:25:21.511122Z digest=sha256:87df6f8d813e5e57880f87036a7e1cc7dd41757d0821e40d667ac035fa970de9