Pith. sign in

Paper Citation Record · LEDGER

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models

As of 7 August 2026, this Paper Citation Record lists 70 of 70 outbound references and 0 inbound Pith citation observations for arXiv:2508.06009.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.06009 v1

Coverage vector

measured 70 of 70 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T23:03:10.178730Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

70 of 70 outbound references displayed

  • verified exact2
  • verified fuzzy3
  • unresolved65
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a3f53cd3-5ee1-46be-8659-9225c646233f · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.136528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:05.479730Z digest=sha256:6a7fa422b46202b1c7dfc44d3c5a1b726912fc4e33957c0b919b561c888f0313

Observation b38bca70-330c-4537-bf98-e86dfc74ca3c · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.125949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:05.548855Z digest=sha256:940972bc0e11862afd7f588bbdfcff9766da8c8d693f37bbb25109c85812f7d4

Observation 7e50cb86-3a88-40d2-bf01-b7e465fa5f7b · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.116421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:05.630453Z digest=sha256:957773d0c00c59648bbc0a66a76fa06ee874e00eea8d11742ee31c3ec4af8bca

Observation 661b3ec9-63de-46cb-bba7-6d56bd1afe86 · outbound

This paper cites D.; and Ammanabrolu, P.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models D.; and Ammanabrolu, P

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:03:11.107598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:05.698730Z digest=sha256:548c42dc555650fdce03e91acf86729e1f2cecff307df0c6f3bb3d2b7f7cc6ec

Observation 019de136-a5fb-49e7-b721-ea8225ea87f3 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.098134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:05.782415Z digest=sha256:e370f1b213410f2d4b5b84ad7263956b55f9230ae689d67592b8f8749bc647c6

Observation 67e0bbaa-98a7-4d03-be3f-2c61303d8811 · outbound

This paper cites S.; Bahaj, S.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models S.; Bahaj, S

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:03:11.090252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:05.902707Z digest=sha256:64579f0e38bff9dafe18003d316706a8974fbd45ca2318ce398e4b7f4096355b

Observation 67ae9123-8417-4077-898e-4bc46c7a5b94 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:06.024471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:06.024471Z digest=sha256:a97d646999799593df67b595847ae6c5ddcd11a877836d7a9acadb9a4c60cc55

Observation a0f2249e-6e3a-4c6a-9816-39f714cb71d7 · outbound

This paper cites Qwen2.5-VL Technical Report.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Qwen2.5-VL Technical Report

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:06.129793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:06.129793Z digest=sha256:7db82aa9171e6681bda30ceafa5506981f2e98fe5c9e6e91b6e30a1ed8a65d2d

Observation e144e3fe-8be4-40c8-a66c-c85a2f0d0488 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.080301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:06.189698Z digest=sha256:ed1e80a41450b59e30bf8ab8b8ddb088aa04d4405db0d1d8c09228adfcc4a3e5

Observation 7256ba6d-cb3d-432f-b1ec-11ae92db001e · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.069911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:06.274113Z digest=sha256:482e2b8c4ddf6472fc304666552c948aeba1e6673b78038e682dd34a67abb923

Observation a97923c0-6d6c-47cb-a13c-450e89675f53 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.054461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:06.405881Z digest=sha256:843ea71674e1e588b35e3f6aa187fde4970f6385c3c4774161bb71d72979a7f5

Observation 2f187234-bd6e-4fc0-917d-39b20e354fcd · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.044836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:06.513232Z digest=sha256:a26913277b94019987aec702cbc88f33c44a645f9ffce4b67240edf45a88769b

Observation 510143d2-1596-4e11-a7df-93bb2fa1d1c5 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.033015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:06.534871Z digest=sha256:0546fda8371f5c634ff4321b3f7f43951f7f31856991205e7f64b6448b3f3233

Observation 52ab568b-7125-4c5f-8753-0f7549bbe07d · outbound

This paper cites SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:06.609642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:06.609642Z digest=sha256:d1f5ddf76510e8f7f7a5340e18f0c9b9d4e821aac1bd184c4f22c6a6cc3faf11

Observation 7463ee3b-7a8b-4b4c-8b17-39c278af18b2 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:06.826107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:06.826107Z digest=sha256:a2f8b4865cef7575c26e09eaf9d2b909324597b47cdca99c6621b3b2a7056868

Observation f661969b-fbb4-4be4-87a8-f55ab6236a67 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Training Verifiers to Solve Math Word Problems

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:06.901216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:06.901216Z digest=sha256:28a0c49bcbfe49a9510d523392c59f66f6709ab053d1e9b33d136f19405346c7

Observation 4bf6a102-a30a-416f-b249-14055def1dc4 · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:07.061721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:07.061721Z digest=sha256:95d4cc1a2d1e05ae294f426cc76e6822a69e710ed56e7ea429235e7c08813c48

Observation bb1a9588-635b-4ae0-9df9-a79c863ab9a3 · outbound

This paper cites OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:07.169525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:07.169525Z digest=sha256:b4fd02728c3818fb663ffb579521d66c3e9142d06ed5ea11d4b97ffe06cd55d5

Observation e26bc2f0-7d3a-407e-8a53-524173d99e44 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:07.244354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:07.244354Z digest=sha256:d4cb5a3de33443da5b40474f730e96e5676d9a0882908210f79f0857a1d248e5

Observation cb3702ed-2f72-4b02-b978-4fae2d648b4f · outbound

This paper cites OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:07.345505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:07.345505Z digest=sha256:c00c232ed0beb7be71b30f8550613871665a59f15141783ee0b405f942732486

Observation 0c7f0ac6-c693-4f66-b267-b8e515ff00ee · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:07.437202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:07.437202Z digest=sha256:16ff83a2784f7051331fbc9474f1a4cae84b067096eba8ab59679bd9b65da370

Observation e0211619-0a7c-4cb0-a0f0-f8ff398862f6 · outbound

This paper cites OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:07.545378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:07.545378Z digest=sha256:6bc6c8fb0b39b80ca27446674aead97caf86c327dee84ccd03af900132246ae1

Observation b1f05515-02b9-4a4c-9d86-6a0c6faa08c2 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Measuring Mathematical Problem Solving With the MATH Dataset

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:07.632051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:07.632051Z digest=sha256:70f34c2a938fee177d6e2329b125578e2578679e6bdae90e5187d47fc72afa10

Observation 899150c5-4673-40c3-b6a9-e19a38c1dc7a · outbound

This paper cites Benchmarking LLMs' Mathematical Reasoning with Unseen Random Variables Questions.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Benchmarking LLMs' Mathematical Reasoning with Unseen Random Variables Questions

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:07.734635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:07.734635Z digest=sha256:68f66a508e0c92856db52f5a7fb75cbbfd6c2b5b12fab47478adaca54d6b8992

Observation 429db48e-c95c-4a12-bdd3-388ffca76dd4 · outbound

This paper cites OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:07.854784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:07.854784Z digest=sha256:e174bd9010b5e311c7f86461fd0cb5549f6daa8ab7111d60d8c2d2d71775a93d

Observation 0fdc2059-8eac-43f0-8bff-108a0c209620 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.023470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:07.949013Z digest=sha256:8fb7056f52be9027edb2091759c4d0c8b548c1791ee533da510c3055eb5ae95c

Observation f4afe626-f042-4eb8-948c-6478f66f7fef · outbound

This paper cites VisOnlyQA: Large Vision Language Models Still Struggle with Visual Perception of Geometric Information.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models VisOnlyQA: Large Vision Language Models Still Struggle with Visual Perception of Geometric Information

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:08.012180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:08.012180Z digest=sha256:e5a29bc22485af8ea1d045ce4f533a628f4c21fa5a106f81168211125d79ca72

Observation 5f44f766-ee04-4484-9033-28be5d49da4d · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.014370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:08.096139Z digest=sha256:ab77e25da275b18933bc436697bbe1a6797d001e642f2790c788b11760e277fe

Observation e446c60f-acc8-4f68-9a7d-af9aa250ad29 · outbound

This paper cites SEED-Bench-2-Plus: Benchmarking Multimodal Large Language Models with Text-Rich Visual Comprehension.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models SEED-Bench-2-Plus: Benchmarking Multimodal Large Language Models with Text-Rich Visual Comprehension

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:08.173192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:08.173192Z digest=sha256:113a73c2417dfc7cb680bbbf846a0ce7e1b099023e86896f70eac4bdb340e71e

Observation 7458803d-b9af-4f7d-9139-0725b7e7c1df · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 30

Resolution
verified exact
raw_fallback, observed 2026-08-05T23:03:10.609040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:08.264837Z digest=sha256:00529062e63a02c79d3f1ab0482a092d2344981df0f17e2ec75bdb59df6edc75

Observation 3e9f344b-1c98-4941-af2c-1bf31b044908 · outbound

This paper cites CMMaTH: A Chinese Multi-modal Math Skill Evaluation Benchmark for Foundation Models.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models CMMaTH: A Chinese Multi-modal Math Skill Evaluation Benchmark for Foundation Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:08.364386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:08.364386Z digest=sha256:987b73726995d2396db76a8c80f08e7aecbb771e4918b0bccd217c2bf2f7085d

Observation 65025616-6cc4-4899-8aaf-823dfe965e17 · outbound

This paper cites DeepSeek-V3 Technical Report.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models DeepSeek-V3 Technical Report

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:08.435992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:08.435992Z digest=sha256:296fc7aaff057c47a82d3be4ec74128795fcb84b2e6f0d850f40063baffad0fd

Observation 8066f43e-ca67-4b67-8ca5-d290acc7a358 · outbound

This paper cites Focus Anywhere for Fine-grained Multi-page Document Understanding.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Focus Anywhere for Fine-grained Multi-page Document Understanding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:08.515997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:08.515997Z digest=sha256:563a3ee68a1381acd8438b48e42027b5549951cc10e9fbda89c33835c33ae916

Observation 380ac013-5ac1-43fe-b9ce-d221e9b4ed3c · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.004621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:08.606167Z digest=sha256:6edef135965813c8dcd318d5cc077306148e17796f88b3c55181cbee3cc122dc

Observation 5e9ad217-3b3c-4ffe-bc04-8c7437c425c5 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.992808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:08.696090Z digest=sha256:a0fc56827ea00ac8bc227afbaebf79b4f1d8e9efb518410d773d9838653d909a

Observation accb2ee4-d4c1-4e0a-8015-82579f104c99 · outbound

This paper cites ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:08.779737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:08.779737Z digest=sha256:d1d1a2947a8e642648e162757f05f7c962541157f156927cf2b3039a9af3cdda

Observation 45c6c99f-4dc9-45fe-87ad-1a2aef8d2741 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:08.881627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:08.881627Z digest=sha256:ae33d29cc433328885a49c3dfd3b4d35dbc478efdd5e95113f6504f744ecbab5

Observation e2016163-bd89-441e-a1da-aa3a6e5fb986 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:09.004368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:09.004368Z digest=sha256:edbb8a33761d10e9599a0ee5ca0e8b6e0e5e969f62c26952a9773de2d3bbc09b

Observation 0f940b45-b68a-4cf9-acee-4d2492a24848 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:09.168731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:09.168731Z digest=sha256:c5a8677a00ce6674dee57f229d708aec44ce517a5ca08e3d48a7ef0e3baf6e15

Observation c1d456e7-75c6-4dac-8b03-f18346f8f951 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.968289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:09.300212Z digest=sha256:e44306975f0b55674a6a7d60205d9ef4c096188579dec09a5f12612e3f2253c5

Observation 907d46af-5286-4823-b493-aba88da5abf9 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.960560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:09.418539Z digest=sha256:dd373e5d596c484010ad7e423c05f80ce5dfd7def5c2bde54e6ddee702668f98

Observation 999c36b2-a78d-47d1-b25d-8364c7a3cbe8 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.952770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:09.585949Z digest=sha256:0e9d6fd3f4008c6a4a5a4f7f86904861c9ef0fbc3b905c8e711433a950e7cdbf

Observation cbb68054-14f4-434d-bc7c-2dd937822bb7 · outbound

This paper cites Skywork-R1V3 Technical Report.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Skywork-R1V3 Technical Report

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:09.720243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:09.720243Z digest=sha256:aefdb204d9a82822b0f7193706be8d1cb69a1dc45bcb8c3b48aa31241a8c956d

Observation f226de3c-ad66-4d4e-854a-0c91a587c763 · outbound

This paper cites W.; Tay, Y.; Ruder, S.; Zhou, D.; et al.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models W.; Tay, Y.; Ruder, S.; Zhou, D.; et al

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:03:10.944633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:09.892367Z digest=sha256:b140567dc7b86a77d834e389a1a430adfb96879024423977a14bc89c8a884670

Observation 13955807-f937-4d3e-adea-6aa6c22de0b5 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.937022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:09.990297Z digest=sha256:4faf4f8ddcb1749cbde518d92bd58068da1b1e7a84ab6e0ea2b36f26278ca0d1

Observation 7d52edc9-2494-4423-99e0-35a83f14b4f1 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.928498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:10.101605Z digest=sha256:2df8084bc521e7eaefdb2456693a82036319a43858dae7bc12ef0d82367811fd

Observation 71aee071-de03-4200-b863-2ceaaa125bea · outbound

This paper cites MiMo-VL Technical Report.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models MiMo-VL Technical Report

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.109377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.109377Z digest=sha256:6b096e87c1f6ff86f14a3bb6b56b35080c4983c0608bc247f625a08847344818

Observation 2ea5c41e-9f20-4505-9e8e-01755079951f · outbound

This paper cites Gemma 3 Technical Report.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Gemma 3 Technical Report

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.112701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.112701Z digest=sha256:5198d301263658ffe0cea6c88e91b1eb78c1177649e729b12aa1abd0acf43e27

Observation 5e485052-805b-4905-8a77-1f30ab78ac14 · outbound

This paper cites Kimi-VL Technical Report.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Kimi-VL Technical Report

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.115570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.115570Z digest=sha256:7212e02b24176dd64f2df4246dd3a6e9dbf0124da60630cf52d81a8fa44166ea

Observation 1e0d2528-1958-4e3f-b77f-1efc4f76eaf6 · outbound

This paper cites Kwai Keye-VL Technical Report.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Kwai Keye-VL Technical Report

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.118939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.118939Z digest=sha256:5bbf1dbd2b0eaaf0b95891b3562846e88a98378c2dafdfa511da0e705ffd97c1

Observation 618dd905-d1d5-49ec-99d7-26c90ea45130 · outbound

This paper cites VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.122029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.122029Z digest=sha256:b6381e1f4896af6eed2d796920a4fe2846ffb31665cce77d4d2fce3b6ef6bd7f

Observation 260dc3eb-3791-42c0-8cfe-437c4713d715 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.920534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:10.125020Z digest=sha256:385c136dcb81ff909e18db00247a838fcf8672fca483461993e3803fbb7a29d7

Observation 3871fa7b-7502-439a-94b1-7fcaea51847c · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.911319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:10.128111Z digest=sha256:a5cbcb40cc37ed80aa4cf291858dd531b110b0ebaf249c70b34ccbdf0179b055

Observation 2a694aaf-8b1a-4551-b2ec-17d8ecf284fb · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.902563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:10.131052Z digest=sha256:650d26ff534d3eb2a385bd85aafb52f459ed1a94f552b8c44e942caba8918baf

Observation f055fe9b-4553-40eb-bc57-d58d410b375a · outbound

This paper cites SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.133767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.133767Z digest=sha256:3a6ef7f94ef08215e69378b62c66577cb93cb0c96432c649ebd7de31379c954b

Observation 41cbcac5-d37f-48cf-91ae-e0388f1b2d00 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.137008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.137008Z digest=sha256:b44ee3a97ef386650da313075b4ace2d135bdc0e9047d1acec13afc6ae858b3e

Observation 1bd96295-23af-4c3e-a5db-9dc5ee03bd66 · outbound

This paper cites Benchmarking Multimodal Mathematical Reasoning with Explicit Visual Dependency.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Benchmarking Multimodal Mathematical Reasoning with Explicit Visual Dependency

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-08-05T23:03:10.382590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:10.139666Z digest=sha256:0d13b04b63f9a88954629695451c9af48fa5fbf114d30721ab0f23714fc42379

Observation 4fb2e29a-2ccb-4639-9c21-9808a76826c5 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.142900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.142900Z digest=sha256:1590649cf951896d08cd4092d87d2a9e278b4b11f346cc47695bb4d6cf130de6

Observation 9ce07fde-0d99-462f-9dcc-29e2e4c88fe9 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.893404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:10.146186Z digest=sha256:4b9ce58a5ca57d69e54182b0bd4bad6a18700fca3fce3e11e6d54348709aae73

Observation a7d2f13b-9698-4f37-8ec8-bb2072bd816f · outbound

This paper cites LogicVista: Multimodal LLM Logical Reasoning Benchmark in Visual Contexts.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models LogicVista: Multimodal LLM Logical Reasoning Benchmark in Visual Contexts

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.149147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.149147Z digest=sha256:55a634984543dbb4198397da2b72bc65f38a661053ab5f609c639fcccdb646a0

Observation d42b273b-e04a-4d0d-b3fa-ccdea184d038 · outbound

This paper cites SuperCLUE-Math6: Graded Multi-Step Math Reasoning Benchmark for LLMs in Chinese.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models SuperCLUE-Math6: Graded Multi-Step Math Reasoning Benchmark for LLMs in Chinese

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.152722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.152722Z digest=sha256:2c54253dac825f706a0806e6ca137b10c8fa9eb8eb84c1f9cc74a3fc0794162b

Observation ac6f4c5d-f2f3-4688-ac28-1377088bed02 · outbound

This paper cites Qwen3 Technical Report.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Qwen3 Technical Report

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.155851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.155851Z digest=sha256:820f5608658c7f6d81fdf3c69d405bd5907647ba2d0907186fcba3902db33423

Observation 81f9e82d-4567-4976-9de0-b0ba9a4a3bb7 · outbound

This paper cites WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.158919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.158919Z digest=sha256:0a927fc68c50c863b747d3706ad9c269015ed31ef7a52493826cc51ac34101b9

Observation 0b2729ac-0c6d-42a0-a2a7-401f20bcb0e0 · outbound

This paper cites CC-OCR: A Comprehensive and Challenging OCR Benchmark for Evaluating Large Multimodal Models in Literacy.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models CC-OCR: A Comprehensive and Challenging OCR Benchmark for Evaluating Large Multimodal Models in Literacy

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.162023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.162023Z digest=sha256:5e6985b6d9213c184c5079100a55a9b1ba1d1a439d5f12047ce3f4217bf0f8d7

Observation ab2b6813-c8b3-4b9b-b62b-dc614589eade · outbound

This paper cites Benchmarking Reasoning Robustness in Large Language Models.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Benchmarking Reasoning Robustness in Large Language Models

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.164969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.164969Z digest=sha256:2916d26269d6d577674b87855f45df360b920a4ee89743bbfe41e9b5f9d43b6d

Observation 91144bb5-3f36-47df-adb1-01aa8ccf59ce · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.885047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:10.167996Z digest=sha256:4214dc090d1a25f0fe40db7bfdbb003c63512142bcdb9eab05eac861b138fd57

Observation 3547a3a2-b23f-4526-acf2-6c9d0fb60d39 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.875416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:03:10.170967Z digest=sha256:2b6d0c4fd4747e83ab4f0bce5bd5fb647aa9765a9d1c2949f23bfa4ea7307a34

Observation 1cf70cc8-b296-4921-858d-26877ae6012a · outbound

This paper cites Multimodal Table Understanding.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Multimodal Table Understanding

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.173220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.173220Z digest=sha256:1beb0f2a6d583b51577938f79b97db3dad9bccde3156be4711dcdc292d9744c3

Observation c7cda48d-4955-4a87-b877-5afcf3d4c5f3 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.175885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.175885Z digest=sha256:0c53d3bc4307f44c3ee81ff4c8a1ce400d45fefb9e7541c8755e69994b0a6d09

Observation 803706d1-06d3-4b48-8f87-64fdea209998 · outbound

This paper cites DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.178730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.178730Z digest=sha256:0cce66d5a746c36c289dc7b197320bec4223f8cf9cce7dc8f87cc0fc38bedef7

Pith citing papers

No inbound Pith citation observations are available.