Pith. sign in

Paper Citation Record · LEDGER

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence

As of 19 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 1 inbound Pith citation observation for arXiv:2505.10604.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.10604 v2

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:14:42.975394Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-15T13:03:40.459461Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

34 of 34 outbound references displayed

  • verified exact0
  • verified fuzzy12
  • unresolved21
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 87e88345-65e0-419a-b17c-5659c3503a55 · outbound

This paper cites Flamingo: a visual language model for few-shot learning, 2022.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Flamingo: a visual language model for few-shot learning, 2022

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:14:43.428372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:14:42.834213Z digest=sha256:d1dbb3f6dcec925b9f6b6099a7a0f47d2ef14fe09b3436b6abfe2cd06e31fe82

Observation 07686218-6953-49cc-840b-de759a06fcc6 · outbound

This paper cites Qwen-vl: A versatile vision-language model for understanding, localization, text reading, and beyond, 2023.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Qwen-vl: A versatile vision-language model for understanding, localization, text reading, and beyond, 2023

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T21:14:42.840002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:14:42.840002Z digest=sha256:7acfdc4c8c0979be73ece17fa6dc5cf5e562dcd40d90f12936da92e26e25fa03

Observation 6ad4ae12-e25b-4d13-a6c5-086cad8308df · outbound

This paper cites Spatialbot: Precise spatial understanding with vision language models, 2025.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Spatialbot: Precise spatial understanding with vision language models, 2025

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T21:14:42.843997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:14:42.843997Z digest=sha256:8da6d3cd901f5b23dfd954764c17442e4adbb2481c6d878333a7e1c72d9df44f

Observation 550d2283-d28e-4a7d-b7c5-40d966de3523 · outbound

This paper cites How far are we to gpt-4v? closing the gap to commercial multimodal models with open-source suites, 2024.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence How far are we to gpt-4v? closing the gap to commercial multimodal models with open-source suites, 2024

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T21:14:42.848911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:14:42.848911Z digest=sha256:d3edb58f786cee5a612ffaab59324c0f05ebdf0955d96bf841bcc90486d52665

Observation 996e1359-3b3e-4eb9-8385-dd6c5f6d13a4 · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks, 2024.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks, 2024

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T21:14:42.853472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:14:42.853472Z digest=sha256:12302b7e7f4f415fe89bb9c404b3d11e1098e6bb9f3caf43c429f0671c79f468

Observation be14955a-a184-4eaa-b759-de79b3014db8 · outbound

This paper cites Scaling egocentric vision: The epic-kitchens dataset.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Scaling egocentric vision: The epic-kitchens dataset

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T21:14:42.857675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:14:42.857675Z digest=sha256:6cf382d646dba575abe3c0f9ef194fd049493055a4c4d9a8bf83f4330aa81631

Observation c046bf85-b54c-416a-ba39-7949a06a74bd · outbound

This paper cites Geobench-vlm: Benchmarking vision-language models for geospatial tasks, 2025.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Geobench-vlm: Benchmarking vision-language models for geospatial tasks, 2025

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:14:43.362150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:14:42.862165Z digest=sha256:57d171f2898d1f1ae14f5616c7b87047ed5d8f6cb187fde97ba2bf971ff478ac

Observation 17266c76-1d65-4af0-98af-028779f3b487 · outbound

This paper cites Mm-spatial: Exploring 3d spatial understanding in multimodal llms, 2025.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Mm-spatial: Exploring 3d spatial understanding in multimodal llms, 2025

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T21:14:42.866609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:14:42.866609Z digest=sha256:0071ff5f9a3950463cdb0848ba74fd0f8791fc2b3f46565f24e888a94a44c5ae

Observation 42e69c59-eac7-4258-8f7d-b19164ff9dde · outbound

This paper cites Patch n’ pack: Navit, a vision transformer for any aspect ratio and resolution, 2023.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Patch n’ pack: Navit, a vision transformer for any aspect ratio and resolution, 2023

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:14:43.341176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:14:42.870189Z digest=sha256:72b7bf78184af29f6e1cfb29e833f968788ed6034c87a00efb36e9713cec1394

Observation e0e633db-67bb-4eca-bc9a-d3e91b5be9c2 · outbound

This paper cites Smith, Hannaneh Hajishirzi, Ross Girshick, Ali Farhadi, and Aniruddha Kembhavi.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Smith, Hannaneh Hajishirzi, Ross Girshick, Ali Farhadi, and Aniruddha Kembhavi

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T21:14:42.873980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:14:42.873980Z digest=sha256:bfcec5c0281315b2a0c01014d885e48ce43eccbf71bab8f939696f927b2dedca

Observation 31bfc3c0-47aa-4d7c-ba77-88b5c6288cfd · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale, 2021.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence An image is worth 16x16 words: Transformers for image recognition at scale, 2021

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T21:14:42.878132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:14:42.878132Z digest=sha256:ccf0e8cf8fbdb7f9c8481acb363774f9900d4156820f7e2687cc52798d1d0105

Observation e8723a1f-e289-433b-bbe3-ce8d23587007 · outbound

This paper cites Minicpm: Unveiling the potential of small language models with scalable training strategies, 2024.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Minicpm: Unveiling the potential of small language models with scalable training strategies, 2024

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T21:14:42.882592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:14:42.882592Z digest=sha256:1f25453d2dbedc991028cb53d3653354a423288878fbf279202c572872b0bdaf

Observation ea6c64e2-2557-452d-b62e-c97c500880d9 · outbound

This paper cites Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation, 2022.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation, 2022

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T21:14:42.887409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:14:42.887409Z digest=sha256:368e0f72dc615d671b0c0e9f5584e2ec009157944f711f14bf56197cef1050fe

Observation bf4a7934-a9fb-45f8-b5ae-e428bc15bf20 · outbound

This paper cites Sti- bench: Are mllms ready for precise spatial-temporal world understanding?, 2025.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Sti- bench: Are mllms ready for precise spatial-temporal world understanding?, 2025

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:14:43.297126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:14:42.892170Z digest=sha256:ed7b54516fc5eae7d29a4ae554b354162595da79f6ef20ee6a3c403c8d406d70

Observation a5127df8-e170-47cd-8905-165b1a3d83cc · outbound

This paper cites Visual instruction tuning, 2023.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Visual instruction tuning, 2023

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T21:14:42.896925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:14:42.896925Z digest=sha256:0f30eab612eaa2bfcffcaf5aa758bb77993765ddd25d69266fa9d591347f4d1d

Observation 1499ebe9-a050-4dd8-819a-34749f8621e0 · outbound

This paper cites ivispar – an interactive visual-spatial reasoning benchmark for vlms, 2025.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence ivispar – an interactive visual-spatial reasoning benchmark for vlms, 2025

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:14:43.277779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:14:42.901535Z digest=sha256:cdae3d5035209fa26ac9a706e0d629afdd2f3b208b66b985590beaf9a9b77cc2

Observation 6a684c88-8b16-42f6-8533-24566499da66 · outbound

This paper cites Learning transferable visual models from natural language supervision, 2021.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Learning transferable visual models from natural language supervision, 2021

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T21:14:42.906309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:14:42.906309Z digest=sha256:d516dc0ef7d5bfcad4445cf5ad6dec1bf5ea5b26c621183315f02c84f2415f05

Observation d21812cd-ad75-485a-8782-3a566f5efc54 · outbound

This paper cites Gsr-bench: A benchmark for grounded spatial reasoning evaluation via multimodal llms, 2024.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Gsr-bench: A benchmark for grounded spatial reasoning evaluation via multimodal llms, 2024

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:14:43.245849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:14:42.909907Z digest=sha256:6fbecc90d31cbf38cbdf060c8194befc34dabfb0fcc92005ea816b4b90fc65ce

Observation d5289e28-15f5-4bcf-abf0-e7dd2e3a4887 · outbound

This paper cites an unresolved cited work.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:14:43.224809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:14:42.913857Z digest=sha256:203083dc85ab44838032fc07c9332cf391122378dfa685e06de00a222b75c2f9

Observation 2f9c5f1d-a126-44e2-a7a3-a61d24c260c1 · outbound

This paper cites Qwen2-vl: Enhancing vision- language model’s perception of the world at any resolution, 2024.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Qwen2-vl: Enhancing vision- language model’s perception of the world at any resolution, 2024

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T21:14:42.919220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:14:42.919220Z digest=sha256:9c319a7ecb09787bb03e298e4efb586a0402fdccd849509d8418efcb3ee351f6

Observation 96bb1d28-a914-49dc-8036-b391f3124010 · outbound

This paper cites ST-Think: How Multimodal Large Language Models Reason About 4D Worlds from Ego-Centric Videos.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence ST-Think: How Multimodal Large Language Models Reason About 4D Worlds from Ego-Centric Videos

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T21:14:42.923386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:14:42.923386Z digest=sha256:6caee840a2b76ab727d1b15b88ce278fc224dde5aab61e6190284c5f430f75b7

Observation 13f3de2b-29b2-439c-b235-d8ccd82f46b4 · outbound

This paper cites Lvlm-ehub: A comprehensive evaluation benchmark for large vision-language models, 2023.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Lvlm-ehub: A comprehensive evaluation benchmark for large vision-language models, 2023

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:14:43.202941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:14:42.927704Z digest=sha256:7e449f19c11e9ab142d38b0f66abf08be4e12a609b460f044c2f86e33a5a8c43

Observation 508ec8fc-6d1a-449d-811f-9db2594959f3 · outbound

This paper cites Gupta, Rilyn Han, Li Fei-Fei, and Saining Xie.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Gupta, Rilyn Han, Li Fei-Fei, and Saining Xie

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T21:14:42.931630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:14:42.931630Z digest=sha256:47e71c4972fe8284643f211cf9acc4e2aa5e27508767238b855307a477d5509b

Observation 56015db9-22e8-4f78-868c-05ac3d0a976f · outbound

This paper cites Minicpm-v: A gpt-4v level mllm on your phone, 2024.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Minicpm-v: A gpt-4v level mllm on your phone, 2024

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T21:14:42.935144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:14:42.935144Z digest=sha256:674a66e1e7b0cc3ff935536b3962844fb0feb373bf570d9b0d8f59b5be6439f1

Observation 7a48a703-c9b7-47da-ac1f-d364b701f184 · outbound

This paper cites Good at captioning, bad at counting: Benchmarking gpt-4v on earth observation data, 2024.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Good at captioning, bad at counting: Benchmarking gpt-4v on earth observation data, 2024

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:14:43.168321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:14:42.938659Z digest=sha256:b5195f2e7fc1ee8cfef203cd99ff95c5eced88fdcb117ce5a13fcbfc2c39a641

Observation 7869ba2d-d9e5-4933-a60e-25f151893c0c · outbound

This paper cites image_caption.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence image_caption

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:14:43.153469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:14:42.942260Z digest=sha256:aaf93eccac47dae1fc6ef92b34931ba1b59fece6506127e402fdc235b6b366b7

Observation de0e1514-0314-414e-9cd3-c2ddbb478ab8 · outbound

This paper cites an unresolved cited work.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:14:43.138770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:14:42.946256Z digest=sha256:783944102a7ddeea0fa6dcfba89092f009a6c17f674ab8c60dce2f2100d2c15d

Observation 6cf4dfac-5101-4e18-940b-1c504da11ffc · outbound

This paper cites an unresolved cited work.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:14:43.124542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:14:42.950263Z digest=sha256:95306a0880286b71df5bb2fa5105862416a6b2c01f9e8a26f0e0b369a69b17f4

Observation 10f6f34c-a698-4d3c-b115-630d2ef5a125 · outbound

This paper cites an unresolved cited work.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:14:43.109455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:14:42.954616Z digest=sha256:fe21a3a8a481ba1a617599724211eb707d1b0e4057b73b221d3ab022bd7afc1e

Observation d18fda69-9fef-406e-9282-80c02c983538 · outbound

This paper cites an unresolved cited work.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:14:43.091845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:14:42.958473Z digest=sha256:2e3cca67334fc952027aa132f70aa46418fa43a087b06c611c5fec32be5458a6

Observation 602e3c84-08a4-4704-a2ad-0a28f95d1a03 · outbound

This paper cites } Stage 2: Task-Specific Questions a. Spatial Relation Task RELATION BASE PROMPT You should output a json string with format {.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence } Stage 2: Task-Specific Questions a. Spatial Relation Task RELATION BASE PROMPT You should output a json string with format {

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:14:43.079210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:14:42.962107Z digest=sha256:da053e3ef0caba1209da73d01153aa00e59f908697717d27e8187308d4cb9f69

Observation 663d2553-90c8-4201-a73d-d98f83ccf89f · outbound

This paper cites 3Implemented with PIL.Image.transpose.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence 3Implemented with PIL.Image.transpose

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:14:43.059612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:14:42.966737Z digest=sha256:7ca19a0d4b44eeb25e521c6c3bc2fe0dd91f5abe2e7fbd62511cc91fb602518b

Observation 64ddf566-fd70-4027-8179-40acbeb1c2d9 · outbound

This paper cites Limitations.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Limitations

Reference 33

Resolution
malformed identifier
raw_fallback, observed 2026-08-15T21:14:43.045458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:14:42.970742Z digest=sha256:a890a68c9dc1931337795732ae840ca8603b690ca737e54f9a9fe4449573499d

Observation 7c749a1e-b4b1-4bbb-a9ce-a288f0d04e9e · outbound

This paper cites Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects.

MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:14:43.031455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:14:42.975394Z digest=sha256:08c8d09356379f4b874cc32ce88986e336d7a62c9039c2e6d87c585a08cf9d12

Pith citing papers

Observation a09fba45-765b-4fce-a3cf-5210b421166d · inbound

It's Time to Get It Right: Improving Analog Clock Reading and Clock-Hand Spatial Reasoning in Vision-Language Models cites this paper.

It's Time to Get It Right: Improving Analog Clock Reading and Clock-Hand Spatial Reasoning in Vision-Language Models MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-15T13:03:40.459461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T13:03:40.459461Z digest=sha256:4b48e787feb305d62e8ac62b8ae482097242fe647dc8ef824cccd6c54713e2f9