Pith. sign in

Paper Citation Record · LEDGER

VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2504.10342.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.10342 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:50:32.927998Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:09:43.184299Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation dc85e25c-5131-46e0-a594-a8fb0bacfe77 · inbound

Jigsaw-Puzzles: From Seeing to Understanding to Reasoning in Vision-Language Models cites this paper.

Jigsaw-Puzzles: From Seeing to Understanding to Reasoning in Vision-Language Models VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:50:32.927998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:50:32.927998Z digest=sha256:c480cdd9b9a84c7cac6f84dd4d41b88418b0b2a385a2fb528c2ef08574e325ca

Observation e8aab8db-8a7b-4c70-acd6-5180a716456b · inbound

MME-Reasoning: A Comprehensive Benchmark for Logical Reasoning in MLLMs cites this paper.

MME-Reasoning: A Comprehensive Benchmark for Logical Reasoning in MLLMs VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T13:35:59.800576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:35:59.800576Z digest=sha256:0c491e82926472bffd46b4938ea28a425db7420f54ad359a842ae877c202be75

Observation 0299d2e3-07dc-4d17-91fe-4aac498c87bb · inbound

VisualSphinx: Large-Scale Synthetic Vision Logic Puzzles for RL cites this paper.

VisualSphinx: Large-Scale Synthetic Vision Logic Puzzles for RL VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:18.894158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:18.894158Z digest=sha256:de1620a1bc4a52814ad837939f5c70e44a95ec12806c3fd0c5237a3ad79a5c41

Observation 1bd79913-9728-404c-a379-a26deb8e6c40 · inbound

PyVision: Agentic Vision with Dynamic Tooling cites this paper.

PyVision: Agentic Vision with Dynamic Tooling VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T18:31:31.285831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:31:31.285831Z digest=sha256:6bf8a53c624d54c947a64c096fde510473c1295667cceeebfea36264695e5e84

Observation 157d773d-61dd-4579-a0a3-75a727390367 · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:16:34.725580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T20:12:46.385646Z digest=sha256:1fc493e4b945a089fe6c7f5d0d2563a5f6b5acc13daa1478bdc71bc97433049d

Observation 14012aca-e9b3-4b09-ba82-9391edd17e32 · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:31:25.285893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T10:30:06.829915Z digest=sha256:a288f23c581582072b198f7a069afed1b4da245fef10aa7731e297c85e60bce0

Observation 36dc0cea-f4ce-4c15-8b78-2988d4cbaf90 · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-02T22:00:09.028700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:00:09.028700Z digest=sha256:f6161a359371e9f8a44b0e33bba0bd3841fedd6fd14d0306f9e1f5be1248d919

Observation 28dea71a-aa68-40be-9c65-92537819bd3b · inbound

Limits of Spatial Imagery Reasoning in Frontier LLM Models cites this paper.

Limits of Spatial Imagery Reasoning in Frontier LLM Models VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-13T19:19:52.557913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T19:19:52.557913Z digest=sha256:342bfa639edd3f86941e9f9965195205587d11f9d6c130383e650ab94eb5f712

Observation 8c9eab83-c05a-49e5-9dea-746e01f64a12 · inbound

SALLIE: Safeguarding Against Latent Language & Image Exploits cites this paper.

SALLIE: Safeguarding Against Latent Language & Image Exploits VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:30:51.295449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T19:04:46.426969Z digest=sha256:9c68f9dc95629d8520cb9a25948d4ef987e3b226b5c93845fb16a7cbcd156031

Observation 32710ea2-6bc0-4cea-a032-ea0f33c017a4 · inbound

TraversalBench: Challenging Paths to Follow for Vision Language Models cites this paper.

TraversalBench: Challenging Paths to Follow for Vision Language Models VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:36:02.685510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T15:25:38.215013Z digest=sha256:96b3b050e4f1448e258a7a1b21678f63418011152194b3d1da41d2922e9a6039

Observation 7a59ac05-a125-454e-ae97-d8497ee5de87 · inbound

Reinforcing Multimodal Reasoning Against Visual Degradation cites this paper.

Reinforcing Multimodal Reasoning Against Visual Degradation VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:06:25.204696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T04:37:21.451146Z digest=sha256:3d175d28eac8f4da859e88d5c5df1502f4627bdf26b83934b6c587569b45f9ff

Observation 17dca504-fc9a-489b-bffe-4dcc82decf55 · inbound

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model cites this paper.

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:17:29.199028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:14:48.918959Z digest=sha256:abf79bbbd32b3d56694c2e1347c13f224f578476bb687155f51eccd694eb2c4a

Observation 048ff051-86d8-4ffa-9f85-f94b1affb024 · inbound

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model cites this paper.

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:48:00.572342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T21:47:50.595481Z digest=sha256:3f5d86b47f0e0afc505fcaed0f980ce20fe9f8266860609745d6e9f6c9fafba8

Observation 4cd7d506-de3d-4848-95b4-e600246450b2 · inbound

Semantic-Enriched Latent Visual Reasoning cites this paper.

Semantic-Enriched Latent Visual Reasoning VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T06:43:05.875829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T06:40:55.537488Z digest=sha256:0ccb4e51f1d355537639bfaa67684759f072a60ca8c95e37286c2494f9d631e2

Observation db4ee2f6-c4f7-4e38-97bb-966aba36fff3 · inbound

Semantic-Enriched Latent Visual Reasoning cites this paper.

Semantic-Enriched Latent Visual Reasoning VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T18:55:00.779176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T18:48:04.370230Z digest=sha256:a15129e8492ea18d2a56156614f72b4cf9c4ef89441616da251d84f063834c05

Observation 6b9996d1-738b-4331-a89d-f3355faf8ac1 · inbound

StemBind: When MLLMs Get Lost Between Rules and Instances in Abstract Visual Reasoning cites this paper.

StemBind: When MLLMs Get Lost Between Rules and Instances in Abstract Visual Reasoning VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-06-29T00:12:50.317499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T23:15:56.598968Z digest=sha256:f2bb462580040242620cbe779bcb15fdd693ae295737499f929a16db9b64052b

Observation 0d4afd7a-2761-4209-b41e-ba36b85a72e4 · inbound

Zone of Proximal Policy Optimization: Teacher in Prompts, Not Gradients cites this paper.

Zone of Proximal Policy Optimization: Teacher in Prompts, Not Gradients VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 141

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:48:56.116501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T01:08:52.981296Z digest=sha256:911b31f331ccec1da2af77591820c11d849276165623952094150f70a1f4ff7f

Observation 91627fbf-52d5-4376-b522-4392b0c57f4a · inbound

Look Light, Think Heavy: What Multimodal Chain-of-Thought Reasoning Can and Cannot Do cites this paper.

Look Light, Think Heavy: What Multimodal Chain-of-Thought Reasoning Can and Cannot Do VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:09:43.185822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T10:27:43.396559Z digest=sha256:62bf6589af2d0287d38d207fb54b333080f2f784cc1404b4436ac75c035de4b6

Observation 843e4628-03f0-49c5-80f1-819965140d95 · inbound

PACE: A Proxy for Agentic Capability Evaluation cites this paper.

PACE: A Proxy for Agentic Capability Evaluation VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 58

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T13:38:18.468017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-03T13:34:39.350893Z digest=sha256:a87f55053e6b7331a5f8b9e95727ec168e2a4f400c8936e7e0b756d365342526

Observation ef855c3d-5503-47e4-8adb-f9d5f82bc823 · inbound

PACE: A Proxy for Agentic Capability Evaluation cites this paper.

PACE: A Proxy for Agentic Capability Evaluation VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-12T08:29:58.561496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T08:29:58.561496Z digest=sha256:05ca6bb0765c25d077168c0d6bf2f0cb960270c3edc138934baf5eb0644be9a5

Observation ca6f6a0e-2630-43d9-9ccf-325d95d6ef81 · inbound

Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning cites this paper.

Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T11:47:26.293957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:47:26.293957Z digest=sha256:7e7522f9fb2ef4a2c02fb8b21a07339b5c7c4f50ab0ea677a63bb3f806cbae78

Observation 6b89ba06-8336-4c64-8b10-f79a53b39f0d · inbound

Beacon: Knowing When and How to Perform Agentic Visual Reasoning cites this paper.

Beacon: Knowing When and How to Perform Agentic Visual Reasoning VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-31T02:45:28.647722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T02:45:28.647722Z digest=sha256:8471a3b075fb212a1fab5673d5249274b5b250d7d900c9a02f00b352c47dc20a

Observation 911b0bbe-095d-4888-802d-e7e9880745fa · inbound

LUT: Latent Utility Training for Visual Reasoning cites this paper.

LUT: Latent Utility Training for Visual Reasoning VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T00:28:16.463177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:28:16.463177Z digest=sha256:0e89f006949a170b1c0c1bf05646b9005e0212581099b8242b223a331dacc281