Pith. sign in

Paper Citation Record · LEDGER

DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 41 inbound Pith citation observations for arXiv:2411.00836.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.00836 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 41 of 41 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T15:26:07.809728Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:48:56.041994Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4794ee11-da2e-46c1-84dc-2d3eaaa2a716 · inbound

MATH-Perturb: Benchmarking LLMs' Math Reasoning Abilities against Hard Perturbations cites this paper.

MATH-Perturb: Benchmarking LLMs' Math Reasoning Abilities against Hard Perturbations DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T15:26:07.809728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:26:07.809728Z digest=sha256:1ff357882d5ff29025e14e53837abb18d55747b80ef1506cc0ee5cec35cc7b41

Observation c2dd2c26-2ff9-4d2e-9eea-9e2b62982ada · inbound

R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization cites this paper.

R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-16T00:19:20.565097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T00:19:20.462455Z digest=sha256:d4a20d87e4b7d9a169e1ae5748c6644aaa03fd031eba901b4d8b3e2c650b918d

Observation 54d46fca-135d-447d-8e2c-1e722cc02ab6 · inbound

R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization cites this paper.

R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-16T15:04:22.722652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T15:04:22.690503Z digest=sha256:160c5ab6e723ab0cfb5edd3b9f18c74bcbcb1e7fe47b1cceeb69b1afdaa45e4f

Observation cae3aea9-046d-48a6-b246-c7d7fac5a7dd · inbound

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models cites this paper.

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 156

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:41:08.033152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T13:41:07.991012Z digest=sha256:bddb4c36d5f6075e4b8a7525c48a2500bf2496414dcb7691cf3a014abea0e2bb

Observation c1496f25-d005-4f09-80d0-1a64808b027e · inbound

DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning cites this paper.

DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:42:56.825358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T14:42:56.565621Z digest=sha256:70b0680a070f52faa82c3b87c7bee492fde7ba6215a35de9feab67bef02fea96

Observation 5ae7bb7b-10ed-462e-908c-fd816a7b30b4 · inbound

TimeCausality: Evaluating the Causal Ability in Time Dimension for Vision Language Models cites this paper.

TimeCausality: Evaluating the Causal Ability in Time Dimension for Vision Language Models DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:08.559862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:21:08.559862Z digest=sha256:1b30be911260d5992ab63bd5a4d5aea33922aaa505684704ed74dae81fe16bd4

Observation 1b61006d-4166-4d49-ad64-4c018a689bad · inbound

MiMo-VL Technical Report cites this paper.

MiMo-VL Technical Report DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T11:04:14.304184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:04:14.304184Z digest=sha256:84f4a0b27107dbb0ac615935dd34ae30fe711512f08059b5dc743feeac884083

Observation a5bc9403-a0d6-4efd-b92f-bc0e7a252ec1 · inbound

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning cites this paper.

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:05.225253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:05.225253Z digest=sha256:68d9ff370b786d27cae7a55d3f65ca6c82c29b801e24f8a51e6c9663d8f70bff

Observation 68b418f4-f233-4f31-b187-9a27584609bf · inbound

MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI cites this paper.

MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T21:41:41.103071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:41:41.103071Z digest=sha256:e3639ea3e45ddf2cf9bd462917fea010c15e4f3ec9f93060aff371172d635f26

Observation b7f21d7b-1304-4d8e-aa23-099a7a46970d · inbound

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning cites this paper.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.265524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:46677c43e787d23b704f024e3dd1077b512dc2bf984e6ba1a72f0dfddb15b031

Observation 9da45d34-702a-41cc-9fa9-a0241d60580d · inbound

Kwai Keye-VL Technical Report cites this paper.

Kwai Keye-VL Technical Report DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T20:45:10.546613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:45:10.546613Z digest=sha256:89f5a115f140128ffe546155aa80b27e275690ec09558d17498971cca1f6f788

Observation e509b6e9-ec44-41bd-865e-d1e0d170bfc1 · inbound

BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset cites this paper.

BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T20:15:52.701699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:15:52.701699Z digest=sha256:a20e28a0458bae2249b6cd69e1fd0e755d18fb9a8a184ff93f6d773de1510e08

Observation 78c0a6e1-1605-47cb-86ed-90f98db30e62 · inbound

Position: Reasoning After Perception Means Reasoning Without Vision cites this paper.

Position: Reasoning After Perception Means Reasoning Without Vision DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T15:26:02.359760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:26:02.359760Z digest=sha256:c8d4bada12c5761a4217397ddad7a044f5273ef1118d7134576460e71dc8637e

Observation 25648867-0a60-4698-8cd4-62272875f645 · inbound

GM-PRM: A Generative Multimodal Process Reward Model for Multimodal Mathematical Reasoning cites this paper.

GM-PRM: A Generative Multimodal Process Reward Model for Multimodal Mathematical Reasoning DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T00:59:54.936904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:59:54.936904Z digest=sha256:9fcb12bdaa20418b9f85cecf37cfdb955456a4235279b69998839dc583c8d823

Observation 803706d1-06d3-4b48-8f87-64fdea209998 · inbound

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models cites this paper.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.178730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.178730Z digest=sha256:0cce66d5a746c36c289dc7b197320bec4223f8cf9cce7dc8f87cc0fc38bedef7

Observation 6b2461dd-f50c-4b0b-b157-80895269b884 · inbound

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency cites this paper.

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 190

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:58:59.011849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T11:58:58.660564Z digest=sha256:f56df941eba9eccb50f4327b65aa3e672b9b3cc139cda345cbe1b52d5da76662

Observation 9e06aefb-62f4-4a2c-a405-f24a37aebcb5 · inbound

R-4B: Incentivizing General-Purpose Auto-Thinking Capability in MLLMs via Bi-Mode Annealing and Reinforce Learning cites this paper.

R-4B: Incentivizing General-Purpose Auto-Thinking Capability in MLLMs via Bi-Mode Annealing and Reinforce Learning DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:21.125850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:21.125850Z digest=sha256:375fa30f9e39d31925e9d7cc91b090027b61b66c208f35c90090b69f4445a17a

Observation 2a76cfdb-d8cb-4be2-b25b-a98ce65010b4 · inbound

Kwai Keye-VL 1.5 Technical Report cites this paper.

Kwai Keye-VL 1.5 Technical Report DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:31.683925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:31.683925Z digest=sha256:8ced412e2c6308ac4f57e81df5de3e8ca3142edb54dd338ff2a262cd1d88b97d

Observation 1ff0b8d2-eeb1-4d9b-8ca6-3d5e54c15a04 · inbound

Can Vision-Language Models Solve Visual Math Equations? cites this paper.

Can Vision-Language Models Solve Visual Math Equations? DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T19:53:10.732656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T19:53:10.732656Z digest=sha256:0c687000327239a28c4b02c90da0ec8f2cee1a735b7bb8a82f5d49fa7cc91ba9

Observation bb89482a-18be-46e9-b761-04106fad767e · inbound

Boosting Reasoning in Large Multimodal Models via Activation Replay cites this paper.

Boosting Reasoning in Large Multimodal Models via Activation Replay DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:09:04.048673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T05:05:48.682057Z digest=sha256:10e78c7f7e049f7b17d68b0fadf824d5b3dcd60f89c97ad6a8be6bf0386f7210

Observation 3abf123a-2591-4894-920f-4ff320b2353d · inbound

Addressing Overthinking in Large Vision-Language Models via Gated Perception-Reasoning Optimization cites this paper.

Addressing Overthinking in Large Vision-Language Models via Gated Perception-Reasoning Optimization DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T16:08:04.410716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T16:06:52.052810Z digest=sha256:203a9187c1fc96e4015b74f7f6ea1ba39bec4854d36f4f1e2f99a29f62fab7e6

Observation 6ff6ff25-35b0-498f-9c44-152740a151c8 · inbound

ChartVerse: Scaling Chart Reasoning via Reliable Programmatic Synthesis from Scratch cites this paper.

ChartVerse: Scaling Chart Reasoning via Reliable Programmatic Synthesis from Scratch DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T13:12:54.723766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T13:12:01.889341Z digest=sha256:4bfa8c80dd2ce33e6589cd5f82733a7370e78640df794f0d74690aa1be910f03

Observation d81264c5-2706-49e1-bb91-813a388323b3 · inbound

CodePercept: Code-Grounded Visual STEM Perception for MLLMs cites this paper.

CodePercept: Code-Grounded Visual STEM Perception for MLLMs DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 64

Resolution
unresolved
no resolver link, observed 2026-07-14T23:22:13.847876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T23:22:13.847876Z digest=sha256:3ab9ef3ea0cd2df3af70233caac140120c479a4020870b9fcc60ac8698a0971c

Observation 583be613-923c-472f-96ff-0a2bc5223d62 · inbound

Seed1.8 Model Card: Towards Generalized Real-World Agency cites this paper.

Seed1.8 Model Card: Towards Generalized Real-World Agency DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:45:14.496421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T07:44:02.827006Z digest=sha256:d146add56ffab772b8bea7bf25924975cbeed198c17f5cf3d4292c927fed6ddb

Observation 367a7780-a059-4402-9a3c-a76e46315220 · inbound

Act Wisely: Cultivating Meta-Cognitive Tool Use in Agentic Multimodal Models cites this paper.

Act Wisely: Cultivating Meta-Cognitive Tool Use in Agentic Multimodal Models DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:20:53.738571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:35:21.514502Z digest=sha256:4be6ed0ec40fe54107e740d477bc8507200c1c7b857b66a37f39430aa0355c2c

Observation 7b7096a6-de08-413d-9bd5-22762d1f2f95 · inbound

Why and When Visual Token Pruning Fails? A Study on Relevant Visual Information Shift in MLLMs Decoding cites this paper.

Why and When Visual Token Pruning Fails? A Study on Relevant Visual Information Shift in MLLMs Decoding DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:31:01.518013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T14:50:37.022338Z digest=sha256:e69b2e3d13dbe91e19b3f9294664d7a747b6c70139d61e7d36339c7f7d63ea80

Observation a5a5f8eb-ce0c-4d76-8b63-65250ae29914 · inbound

ReaLB: Real-Time Load Balancing for Multimodal MoE Inference cites this paper.

ReaLB: Real-Time Load Balancing for Multimodal MoE Inference DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:01:29.079542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T01:23:42.529129Z digest=sha256:fab47f7e44eea163587679cd3f5937607dfe61b1afe681ba8523bc01a0ef2cfd

Observation 574323eb-0efb-4a20-8e7b-e5dc35346274 · inbound

Segment-Aligned Policy Optimization for Multi-Modal Reasoning cites this paper.

Segment-Aligned Policy Optimization for Multi-Modal Reasoning DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:51:07.917445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-09T14:44:31.160543Z digest=sha256:5b69226b2f6fc9db208569dd4e63b6deb98eb106b22b6577c916aeb3b7f1c891

Observation e8a2003e-ca14-4039-aa51-13d48769aa30 · inbound

Uni-OPD: Unifying On-Policy Distillation with a Dual-Perspective Recipe cites this paper.

Uni-OPD: Unifying On-Policy Distillation with a Dual-Perspective Recipe DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:26:16.356487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T17:00:49.448352Z digest=sha256:e424acc807cb7337da2441bc1a22c9da750ca93e91714c0fc0cb78223069d0a0

Observation 94f731f4-8d62-4817-aa77-2f3d071c6fc5 · inbound

CEPO: RLVR Self-Distillation using Contrastive Evidence Policy Optimization cites this paper.

CEPO: RLVR Self-Distillation using Contrastive Evidence Policy Optimization DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:28:06.725569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T07:25:34.637803Z digest=sha256:d72c75ff8a6186e1f1feac30481bca6a217c862aedcd5f7548c1ddb501139e39

Observation 77cdde0a-33d9-4c61-8880-9092fa519084 · inbound

Are Tools Always Beneficial? Learning to Invoke Tools Adaptively for Dual-Mode Multimodal LLM Reasoning cites this paper.

Are Tools Always Beneficial? Learning to Invoke Tools Adaptively for Dual-Mode Multimodal LLM Reasoning DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 55

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T06:18:05.215516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T06:16:47.650748Z digest=sha256:24baf400a73a6ffa80e3f2fd646e0199d2a0d38d894898db66a8d3a111b4f77d

Observation 62ba18d2-50e4-444a-b080-abbb816179e0 · inbound

TRON: Targeted Rule-Verifiable Online Environments for Visual Reasoning RL cites this paper.

TRON: Targeted Rule-Verifiable Online Environments for Visual Reasoning RL DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 50

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:56:19.945174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T14:55:07.045625Z digest=sha256:bbd7985caef2a69a124c1e1972ae4c97073b10d77484858adce1bbed70f6a155

Observation bcdf8a07-284a-487e-8e7d-6cbf472d189b · inbound

Sci-Rho: A Multilingual Visually-Grounded Symbolic Benchmark for STEM Problems cites this paper.

Sci-Rho: A Multilingual Visually-Grounded Symbolic Benchmark for STEM Problems DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:22.863324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T20:11:02.445626Z digest=sha256:798f0c67d64539db4ef3a2fec9753bdc26b952e10caed235abaa8134213e3906

Observation 7e73f050-7a90-42da-aa3d-e2d49f48d2ae · inbound

Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning cites this paper.

Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 120

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:37:25.272238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T19:36:57.231932Z digest=sha256:7470da3c06c2b68baceb0a5dc2c7f45def15c184e23ccc69dfe1342573b4804b

Observation 2eac8e13-ee06-4d26-acba-29fdfa65db91 · inbound

Kwai Keye-VL-2.0 Technical Report cites this paper.

Kwai Keye-VL-2.0 Technical Report DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 113

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T04:27:37.126894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T13:53:10.352603Z digest=sha256:c187996872c20c0683266a76e99e595286dad386a65309776f57820a9bf38d98

Observation 735cf272-874d-4908-b4d9-26409cdabd90 · inbound

Zone of Proximal Policy Optimization: Teacher in Prompts, Not Gradients cites this paper.

Zone of Proximal Policy Optimization: Teacher in Prompts, Not Gradients DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 133

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:48:56.043493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T01:08:52.981296Z digest=sha256:97112383c9dd6730079fada5e9555c773165df386279d8cb1203c303fee21200

Observation 203aa784-5ef5-4555-9b26-f50f1b9fd1b7 · inbound

TuringViT: Making SOTA Vision Transformers Accessible to All cites this paper.

TuringViT: Making SOTA Vision Transformers Accessible to All DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-06-29T15:03:32.172756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T05:32:26.746776Z digest=sha256:f5de753e5664f66d97e996f418e1f9f9f60397ad85606904b0b3a7ae2fbc294f

Observation 0d3444e8-d223-4509-9899-5de7a94a8763 · inbound

Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity cites this paper.

Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 150

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:07:17.456810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-02T18:57:46.841456Z digest=sha256:ea269b4788716f108e8ff0c5e06045a3935ff21f8c9ab0d722b05dfcf6128062

Observation ee54c62c-eef4-4ccd-ac6c-f89ced35fad5 · inbound

Model Guides You How to Draw: Adaptive Visual Gating for Unified Multimodal Reasoning cites this paper.

Model Guides You How to Draw: Adaptive Visual Gating for Unified Multimodal Reasoning DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-14T01:07:06.861298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T01:07:06.861298Z digest=sha256:caffcda1755042ea845504442f931490c0f6363815244a09a358622ac7d5c3b9

Observation a6346352-c820-4d72-88fc-7d33e3b9970a · inbound

What Transfers from Text to Vision? Capability Scaling Laws and Transfer Dynamics for VLMs cites this paper.

What Transfers from Text to Vision? Capability Scaling Laws and Transfer Dynamics for VLMs DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T02:46:49.739428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:46:49.739428Z digest=sha256:8f445405eadc7c11bead92470ccc16720c0f17e91a7fa08b29ad56ef297d879e

Observation d9ba0465-73f0-47d9-a34b-4f562c052503 · inbound

MIDAL: A Dataset of Math Image Descriptions for Accessible Learning cites this paper.

MIDAL: A Dataset of Math Image Descriptions for Accessible Learning DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T00:49:46.500785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:49:46.500785Z digest=sha256:180e9c7752a3f0a1c5ab7ac5e2b883e97c25107482132f14cc09636972253238