Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T14:29:05.241019Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 15 inbound Pith citation observations for arXiv:2412.11936.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T14:29:05.241019Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:24:33.123081Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-23T04:32:32.800150Z
49 of 49 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation dba59bef-4069-472a-b74c-f344e7c58290 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 8ecd0505-1d87-42a7-9340-5b828ef930b8 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 97be51ba-aebc-406b-bc6c-3bed9092f080 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 39a73c8e-ddb0-44f6-bde8-197070c9fd57 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges RoMath: A Mathematical Reasoning Benchmark in Romanian
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99e823c4-d10a-44b7-8267-03c14e08053b · outbound
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 8f3609ab-92a0-4350-b45f-acce94fb949a · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges MathWriting: A Dataset For Handwritten Mathematical Expression Recognition
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a26adb0-27c2-4099-bdec-56177def482e · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges OpenAI o1 System Card
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ba5b79c-1a85-47ad-8470-fae1c7eee7eb · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges ATHENA: Mathematical Reasoning with Thought Expansion
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3c05723f-389a-4926-80e6-a6452dd96748 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Token-Supervised Value Models for Enhancing Mathematical Problem-Solving Capabilities of Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e137c8f-3e85-4ebb-8143-dd8c7d61a794 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74039854-4833-48c0-b298-fe04cc271d58 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Large Language Models Meet Symbolic Provers for Logical Reasoning Evaluation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d3b8f3c-48c8-4bfe-af81-9f0bc5cdc11f · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Wu's Method can Boost Symbolic AI to Rival Silver Medalists and AlphaGeometry to Outperform Gold Medalists at IMO Geometry
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae4675dc-c168-4270-b17a-f8975df4f775 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges How to Bridge the Gap between Modalities: Survey on Multimodal Large Language Model
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5205e4a4-ebd8-4836-9aad-4842212751ab · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Proving Olympiad Algebraic Inequalities without Human Demonstrations
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6c1acb48-a2b3-44e6-80f5-9c98c69cdd80 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges DeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic Data
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e62a1cc0-3c9f-4ca8-8ad3-b97b642f8995 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges MAmmoTH: Building Math Generalist Models through Hybrid Instruction Tuning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf14c607-42d8-4e3e-a264-24b39842d7f6 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges SEGO: Sequential Subgoal Optimization for Mathematical Problem-Solving
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f126ae01-c912-468f-9899-8d718e7c2bc2 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d1ede7c3-80f4-49d0-a788-44747cc9c376 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f516fd9a-df6e-45eb-853a-b9fc42bb8151 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges ❷ Bottleneck in Data Diversity:
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7c0344c6-3a91-4975-9fb3-68349f7e9834 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7fbdf47a-d43f-4d72-8a29-8f07dccf4b64 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges ❸ Bottleneck in Data Scale:
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 42ae1ead-5673-4829-8ea2-84d0008536a4 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f1fa7cb8-3a00-4e12-84cb-de91dca62478 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges ❹ Based on recent trends in the latest works, we further propose the following actionable sugges- tions to address these dataset bottlenecks:
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0509f27b-61dd-4520-a39b-4d72047310c3 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a6ea1e75-6233-4dea-8ead-24338f149242 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 34f89647-9b89-4ecc-a502-643d4901d3ba · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d2387a89-a605-4f65-8a6e-98ead558ba33 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges spatial reason- ing
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 87b450de-ba20-4746-b819-e6677eb04a31 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Use domain-specific few-shot examples (e.g., providing figure-text associations in geometry) to guide the model in switching reasoning modes
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation dc4caeb2-0385-4b9c-b915-3f51b631eedd · outbound
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 578fe127-015b-4dfb-86e4-4ba4bc36f342 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges The limitation is that labeling error types is costly and it’s difficult to cover all long-tail errors
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 72c164b1-2f2e-4abe-aa08-1a4a876ea596 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges computation-logic-conclusion
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 617e69fa-b12c-459d-8fc7-86a75370afc4 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges The limitation is that tool invoca- tion delays affect real-time performance, and some errors require manually defined detec- tion rules
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 17331e94-06c2-48d6-ba1c-c777f22b8354 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Unresolved cited work
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b3132461-8613-4f86-adea-aed2e29164c6 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 77173011-fb98-4b4d-95be-5e327b63df6b · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges If the training data coverage is insufficient, test-time strategies may not be able to compensate (Ke et al., 2025; Chen et al., 2025c)
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 61ec9488-f54b-4770-b841-6691ff24c0d7 · outbound
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3433ae2a-0a2f-4284-8b6b-2ebc63c08367 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 385697d0-fe96-42a0-86e6-ea305380a1ea · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges ❸ Error Feedback Limitations:
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation cf36e3f8-1cf0-47f0-99b0-e5df23e2bebd · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Unresolved cited work
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation dc4f8666-a545-44b0-94fa-83dfa2de3e96 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Long-tail errors such as rare symbol confusions may be overlooked
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2e35f305-c370-4cec-9ace-74211b0a6271 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Large Language Models for Mathematical Reasoning: Progresses and Challenges
Reference 147
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d6b954e-ddac-4748-a237-4c8a0eb053d6 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges MathQA: Towards Interpretable Math Word Problem Solving with Operation-Based Formalisms
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 100240ca-376d-4f4f-8463-db2c396533e6 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges MWP-BERT: Numeracy-Augmented Pre-training for Math Word Problem Solving
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b82d1d28-879f-4225-8e89-516e33005d76 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges MindStar: Enhancing Math Reasoning in Pre-trained LLMs at Inference Time
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17ade6d4-05a1-499d-bb98-db823689930b · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Advances in Neural Information Processing Systems, 36:5539–5568
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation efd5c1c7-8f9c-4680-b841-34e91896f3d2 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges Evaluating Large Vision-and-Language Models on Children's Mathematical Olympiads
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9aea032e-f4f4-49d0-8fe4-e25da0fb9f47 · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges arXiv preprint arXiv:2501.04686
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2f9304d-de2f-4d5a-9e88-fc8542099eea · outbound
A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges BloombergGPT: A Large Language Model for Finance
Reference 2256
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82193361-4bc8-42a1-bc94-2b4db7320522 · inbound
Reasoning Language Models: A Blueprint A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges
Reference 179
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fe2dede-2be1-4b50-9830-bc703c188f0f · inbound
Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges
Reference 223
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 71045df9-91b9-452a-a982-4d4a36911c30 · inbound
A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6df75775-2aec-45fd-b1fb-a298f94f6a80 · inbound
Are Multimodal Large Language Models Ready for Omnidirectional Spatial Reasoning? A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85ca2948-0a06-4aa2-a7c7-f3df5c4d17da · inbound
CAFES: A Collaborative Multi-Agent Framework for Multi-Granular Multimodal Essay Scoring A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5c85ee0-ca3d-4a0d-ba8d-3ebe4f30f42f · inbound
Towards Omnidirectional Reasoning with 360-R1: A Dataset, Benchmark, and GRPO-based Method A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b5a7fe5-dd8f-4268-bb38-eaeb8f7e9bd6 · inbound
Plane Geometry Problem Solving with Multi-modal Reasoning: A Survey A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8aa1ca2e-288e-453a-b26f-315745af4d4e · inbound
Pierce the Mists, Greet the Sky: Decipher Knowledge Overshadowing via Knowledge Circuit Analysis A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d5b4e13-f700-482d-b896-2f0c99511bb8 · inbound
Unveiling Instruction-Specific Neurons & Experts: An Analytical Framework for LLM's Instruction-Following Capabilities A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c2ee46f-164e-4739-9ac5-5f2c1d463460 · inbound
Understanding Financial Reasoning in AI: A Multimodal Benchmark and Error Learning Approach A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96b8d507-a2f3-417b-807c-1d4d8df6c155 · inbound
A Survey on Large Language Models for Mathematical Reasoning A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eaa1e2b9-109e-4539-9e06-d948c0c055f4 · inbound
Are Large Language Models Capable of Deep Relational Reasoning? Insights from DeepSeek-R1 and Benchmark Comparisons A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37c2ba42-2db1-4040-a1bc-a13d2cc8e47c · inbound
GM-PRM: A Generative Multimodal Process Reward Model for Multimodal Mathematical Reasoning A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73c582c6-f868-4444-b2fe-27fc3443fa46 · inbound
GeoLaux: A Benchmark for Evaluating MLLMs' Geometry Performance on Long-Step Problems Requiring Auxiliary Lines A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f34545ab-8bf3-4f76-a439-9cd855772702 · inbound
WAR: Workload-Aware Rollouts for Synchronous Agentic Reinforcement Learning A Survey of Mathematical Reasoning in the Era of Multimodal Large Language Model: Benchmark, Method & Challenges
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.