Pith. sign in

Paper Citation Record · LEDGER

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems?

As of 21 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 3 inbound Pith citation observations for arXiv:2506.06034.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.06034 v1

Coverage vector

measured 67 of 67 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T06:08:28.246141Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:59:38.760845Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T18:48:53.330917Z

Reference resolution

67 of 67 outbound references displayed

  • verified exact1
  • verified fuzzy22
  • unresolved42
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation acb883b3-5c29-4aef-bbf7-34413ea631c2 · outbound

This paper cites MathQA: Towards Interpretable Math Word Problem Solving with Operation-Based Formalisms.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? MathQA: Towards Interpretable Math Word Problem Solving with Operation-Based Formalisms

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:21.855106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:21.855106Z digest=sha256:ff6cd5253405ce9c3dade797c6fa41353e0690dbe113ee275b766527df6f5a6c

Observation 3c532b0d-ebde-4dbd-9019-67174cb88e00 · outbound

This paper cites Claude Sonnet.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Claude Sonnet

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:34.362771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:21.905044Z digest=sha256:46e863a63ca562325526b97b722ceebcacd88b0d0ba45c21b9b243606e23fd30

Observation 167569e3-65a8-40a0-b5f9-7fca7b9d3718 · outbound

This paper cites ProofNet: Autoformalizing and Formally Proving Undergraduate-Level Mathematics.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? ProofNet: Autoformalizing and Formally Proving Undergraduate-Level Mathematics

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:21.997740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:21.997740Z digest=sha256:f18977eafaabca8e6a8aa8e8095e82ee3ca3645863b463d9b0fcfb60f5c940ed

Observation 7d4958a9-4096-4eac-b70a-c377c00380c6 · outbound

This paper cites Llemma: An Open Language Model For Mathematics.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Llemma: An Open Language Model For Mathematics

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:22.093809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:22.093809Z digest=sha256:9a9ed8c893abc7315117b1a029e07d215266d08b472e289e575d0f71fc916da4

Observation 51611b4a-13b8-4d83-bfff-1fbe85b86676 · outbound

This paper cites Springer Science & Business Me- dia, 2013.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Springer Science & Business Me- dia, 2013

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:34.231119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:22.216322Z digest=sha256:9cd656480c33b08a07fbefc4ac2605979baf1f888e8f8c980e98a64977732252

Observation 06fefbf5-c199-4210-801e-6c73f27cb789 · outbound

This paper cites An augmented benchmark dataset for geometric question answering through dual parallel text encoding.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? An augmented benchmark dataset for geometric question answering through dual parallel text encoding

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:34.089762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:22.366269Z digest=sha256:a7242a76dafed1c203928307e52c29788134e8d33ca35790b05dd0a8cc272450

Observation 1179a2cd-654d-4699-b767-6589a11cde83 · outbound

This paper cites UniGeo: Unifying Geometry Logical Reasoning via Reformulating Mathematical Expression.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? UniGeo: Unifying Geometry Logical Reasoning via Reformulating Mathematical Expression

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:22.470480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:22.470480Z digest=sha256:f2ae047397263b9fa3c8009cf1d6721af17cb0388a0479d064aa5828a6026eb7

Observation 0702550d-4da1-4819-bed9-5efa66ede162 · outbound

This paper cites Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:22.574013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:22.574013Z digest=sha256:97a338a099616f7653c238e4f95bd60723ee56cef8c2e6d03664ac4b0b75b53e

Observation 70a54f9c-515b-4817-9044-6e1832e5a516 · outbound

This paper cites MIT Press, 2013.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? MIT Press, 2013

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:33.843539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:22.670481Z digest=sha256:5236765d949daab85d76ceb8c02806c4b9fbca146b0ac801464043322051717e

Observation 5570880d-b59e-43e4-ae83-25a7980259e4 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Training Verifiers to Solve Math Word Problems

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:22.834147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:22.834147Z digest=sha256:60b5a7449107111ba4482fd836be5934ffcb29f4fa3803dc57f16d51fdead1e2

Observation e22ca28c-fe33-42ba-842b-ff10eb2eef9c · outbound

This paper cites Compfiles.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Compfiles

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:33.499806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:22.950530Z digest=sha256:58a2e1791a8bfd96f4f055e3b758b03475976c1c4ec4e1182975c9a5205b9e67

Observation 75df8126-e94b-4bda-a1c8-59953f60c123 · outbound

This paper cites Start building with gemini 2.5 flash.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Start building with gemini 2.5 flash

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:33.307670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:23.014503Z digest=sha256:33238e9634cdc08eb6f2346d3f8f68b3cc4be4a2b416d1d68b52ee2ba16f3268

Observation 1664df8e-3d06-4dba-9890-c813c0175c08 · outbound

This paper cites Mathematical capabilities of chatgpt.Advances in neural information processing systems, 36:27699–27744,.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Mathematical capabilities of chatgpt.Advances in neural information processing systems, 36:27699–27744,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:33.104964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:23.173500Z digest=sha256:c5ee81402be2870bddefff36851e11262c781fca197c3fe6fab96120bb7f005e

Observation 76cda449-1bc7-42e0-8076-fd4cf77adfcf · outbound

This paper cites Measuring Massive Multitask Language Understanding.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Measuring Massive Multitask Language Understanding

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:23.581488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:23.581488Z digest=sha256:379b80dd56bbc523e334a3c9aa1ea9ad155de5934d78f1e8d8b22cbd116366da

Observation 3e870678-2d33-4427-a766-40a00f7f57be · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Measuring Mathematical Problem Solving With the MATH Dataset

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:23.697375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:23.697375Z digest=sha256:02b0f26bf85b41fb9b3e4f60763227d985a82b688ae93df15f07e2fba0e6ccc8

Observation 77bd8e6a-ba53-422d-8865-9b17c424115f · outbound

This paper cites Visual Sketchpad: Sketching as a Visual Chain of Thought for Multimodal Language Models.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Visual Sketchpad: Sketching as a Visual Chain of Thought for Multimodal Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:23.824053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:23.824053Z digest=sha256:aff41e6398ee0e6af1e60d2ae7cf2a699da9f660d310cce3927afa961e7093d7

Observation 0824c7d0-488f-4018-87c1-dc4aa55f425e · outbound

This paper cites Draft, Sketch, and Prove: Guiding Formal Theorem Provers with Informal Proofs.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Draft, Sketch, and Prove: Guiding Formal Theorem Provers with Informal Proofs

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:23.894970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:23.894970Z digest=sha256:1ad55065e0334b7f05ad5af2e5ad62351b25f62f5c16e07334c159d322d212c8

Observation fa850464-a228-45b4-9e60-772ec49e8481 · outbound

This paper cites Lisa: Language models of isabelle proofs.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Lisa: Language models of isabelle proofs

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:32.660889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:23.998470Z digest=sha256:56e6becd3de309bd822cff00c03289b7c78d1dff3182bbf1352f844d5407b2ee

Observation 54a6364c-4c60-452d-9fc2-8d3704295318 · outbound

This paper cites Hypertree proof search for neural theorem proving.Advances in neural information processing systems, 35:26337– 26349, 2022.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Hypertree proof search for neural theorem proving.Advances in neural information processing systems, 35:26337– 26349, 2022

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:32.457411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:24.130906Z digest=sha256:7ecae6ce0bbcd00f0914bc0b8ad9c6f169ed8cfea078929b1f0ba6d4ba7915a4

Observation 1330ca35-564b-4e96-b1b6-6053bf94d57a · outbound

This paper cites Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:24.242469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:24.242469Z digest=sha256:926a6635d7642bdf7c34da196c9abd4aff955b47cfa5c76a97090fb1172bc0f0

Observation 654347db-a614-4318-a427-e98efc433bac · outbound

This paper cites Goedel-Prover: A Frontier Model for Open-Source Automated Theorem Proving.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Goedel-Prover: A Frontier Model for Open-Source Automated Theorem Proving

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:24.408162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:24.408162Z digest=sha256:db85d6c06e60bdfd68b5479a24532caeeacc014c6da91c2188261a7a2f8e8df7

Observation 7feea3de-c201-430f-bc5e-3b4f8599f66f · outbound

This paper cites Program Induction by Rationale Generation : Learning to Solve and Explain Algebraic Word Problems.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Program Induction by Rationale Generation : Learning to Solve and Explain Algebraic Word Problems

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:24.478551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:24.478551Z digest=sha256:8f6a2a63b63cd7f8f5e5479266f8df7f6d87cd4698e6ba9bb54cc2c7107c10cb

Observation fdba6ed3-a7f0-48d2-b337-eca28ea169e1 · outbound

This paper cites FIMO: A Challenge Formal Dataset for Automated Theorem Proving.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? FIMO: A Challenge Formal Dataset for Automated Theorem Proving

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:24.562226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:24.562226Z digest=sha256:1023cab237156124bd910a9189efd24c425a8e82d6695eb431b4cbc668c279ca

Observation 4ecda9a4-ecc5-4ac3-975b-12072c97a093 · outbound

This paper cites Improved baselines with visual instruction tuning, 2023.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Improved baselines with visual instruction tuning, 2023

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:32.154902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:24.650929Z digest=sha256:1018f3cd9f47a3e6530dbe5502802428304eb37dbd3138c4e412e13423381aea

Observation 08f2d713-e59b-4ce1-88ec-09763e6149fe · outbound

This paper cites CMM-Math: A Chinese Multimodal Math Dataset To Evaluate and Enhance the Mathematics Reasoning of Large Multimodal Models.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? CMM-Math: A Chinese Multimodal Math Dataset To Evaluate and Enhance the Mathematics Reasoning of Large Multimodal Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:24.733133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:24.733133Z digest=sha256:19f2b7e0658ecf7caba7fd7bca53f2baea5a41f96ebb041782a7a73cc627e6f2

Observation a2bc0fef-0689-471b-9728-ded97078fcc0 · outbound

This paper cites Inter-GPS: Interpretable Geometry Problem Solving with Formal Language and Symbolic Reasoning.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Inter-GPS: Interpretable Geometry Problem Solving with Formal Language and Symbolic Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:24.805851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:24.805851Z digest=sha256:c7979071d6cb8f75900a0a5ce4b75913c2c2ad65d2de528175151239d9e104ea

Observation 8fde812d-926a-46bb-8ab2-30f19f339af1 · outbound

This paper cites MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:24.896278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:24.896278Z digest=sha256:623545ffabcdbfd160ba381955f884cf7b08d9d6bef799177ba64ad7f8ae016b

Observation cf742dc9-84b6-44f5-94d2-768785c543df · outbound

This paper cites ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:25.001892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:25.001892Z digest=sha256:e5ea3b2d9a694f0b8cd26ed818ced48022a8dca64c58b0568b9acbc27c91ff42

Observation c396e12a-dcb1-4eeb-affc-457d8cff22d4 · outbound

This paper cites Lila: A Unified Benchmark for Mathematical Reasoning.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Lila: A Unified Benchmark for Mathematical Reasoning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:25.116856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:25.116856Z digest=sha256:457cd43196ab0503a85d257ff96f97e497150492455b8b15db957b9f3ca4e330

Observation af94ed3b-2234-40df-b93b-aa149fee2ba2 · outbound

This paper cites The lean 4 theorem prover and programming language.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? The lean 4 theorem prover and programming language

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:25.221848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:25.221848Z digest=sha256:5d66770382e0b3b848aa7bc043e02f6b11f0e7692060c00ad43a43981e55e3ad

Observation d4a8acd1-bb00-4abc-a119-28f0c9f6ef44 · outbound

This paper cites Autoformalizing Euclidean Geometry.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Autoformalizing Euclidean Geometry

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:25.341953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:25.341953Z digest=sha256:ec246980c204c87ddb51814f4332011bdf4d580602d8228e73c9e65a1f2a9f35

Observation 895eebe7-1527-42c7-9204-608ea94e3d63 · outbound

This paper cites GPT-4.1.https://openai.com/index/gpt-4-1/, 2024.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? GPT-4.1.https://openai.com/index/gpt-4-1/, 2024

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:31.885301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:25.405965Z digest=sha256:e22a78447d15eca49626951939f7154fd90bc994f1f42ed4569477995baea336

Observation 2ebd40f4-579d-43d8-8a08-c0584261b028 · outbound

This paper cites Introducing o3 and o4-mini.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Introducing o3 and o4-mini

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:31.653004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:25.460899Z digest=sha256:ec05096a3b739ac14a2c864c5582f2ba67612540c99495428c9f02a02b1acc3f

Observation 95ab73d0-ed12-4361-90c3-1cd5ce07448d · outbound

This paper cites Generative Language Modeling for Automated Theorem Proving.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Generative Language Modeling for Automated Theorem Proving

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:25.530207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:25.530207Z digest=sha256:a6f6f18e5cb5819879dc3c588c0b49705e6fe4eb2b925a9864b1295b59e945f7

Observation 88a9890a-34c1-4455-b04b-7add4ca06e26 · outbound

This paper cites Artificial intelligence mathematical olympiad (aimo) prize, 2023.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Artificial intelligence mathematical olympiad (aimo) prize, 2023

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:31.411389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:25.727399Z digest=sha256:4a595239a9d4086d0564bacd6d6f9e8a43edfd4062097fbde87adc927cd09e4a

Observation 71412192-617b-4eba-a436-b2c54db3699b · outbound

This paper cites A survey of efficient reasoning for large reasoning models: Language, multimodality, and beyond.arXiv preprint arXiv:2503.21614, 2025.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? A survey of efficient reasoning for large reasoning models: Language, multimodality, and beyond.arXiv preprint arXiv:2503.21614, 2025

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:25.795135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:25.795135Z digest=sha256:c01535afc847198c80b3f16afdaa7106e1b6501ebface206979c02d25648985f

Observation 043df018-ce33-4d94-b359-2086772ec1a4 · outbound

This paper cites DeepSeek-Prover-V2: Advancing Formal Mathematical Reasoning via Reinforcement Learning for Subgoal Decomposition.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? DeepSeek-Prover-V2: Advancing Formal Mathematical Reasoning via Reinforcement Learning for Subgoal Decomposition

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:25.882735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:25.882735Z digest=sha256:c8dad50c74d2e08bab3af06a203df19d903d8d23829ed8991d35195a4a4b0070

Observation 59268741-d604-42b1-93f7-8fd56a470ae8 · outbound

This paper cites Elsevier,.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Elsevier,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:31.187743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:25.957670Z digest=sha256:d6ca7e3b56f5e897fc4708332ff624df31555a9c983003daa42f14b2ac53f2b0

Observation 2a222a7c-ea52-456e-88b8-673346dc9dc7 · outbound

This paper cites Imo grand challenge.URL https://imo-grand-challenge.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Imo grand challenge.URL https://imo-grand-challenge

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:30.768042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:26.111225Z digest=sha256:45dc1d8d01f4904c007b21daf4175dd57e89aa8ffc924c1af1b2470ddde11599

Observation b11cb3a4-a417-42f0-8723-1fb1c0e4722c · outbound

This paper cites Solving geometry problems: Combining text and diagram interpretation.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Solving geometry problems: Combining text and diagram interpretation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:30.633467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:26.188441Z digest=sha256:a411d421b07ac034c59dbcb99a9074f519a5cc7d3155eff96561be4b9d5b66f6

Observation 0c23ffa2-fbcb-4e59-894f-f65eea910526 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:26.270766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:26.270766Z digest=sha256:adeee329093b95526d33f5a2e8630ee37efd15d87b74665b12577bb667ca57b0

Observation 2bfd9f50-6f4c-412d-a024-af64874f16a4 · outbound

This paper cites Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:26.344518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:26.344518Z digest=sha256:2d41137473f5580e779c65b26fdd0bcf7959fe71f62013154c1d1d1e81482a23

Observation f79ce605-4d23-4551-8430-bcc1d18c50c0 · outbound

This paper cites What does clip know about a red circle? visual prompt engineering for vlms.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? What does clip know about a red circle? visual prompt engineering for vlms

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:30.455017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:26.412194Z digest=sha256:b8e3002bdae1ee527208f33a70027166c5f5ff787564ae1ce75c5e3f490836c6

Observation d2c77063-030d-4d18-b78f-a08451c9678e · outbound

This paper cites Qwen2.5-vl, January 2025.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Qwen2.5-vl, January 2025

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:26.482320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:26.482320Z digest=sha256:7d00fb1bfdb5e901050639db0d2084d40f2b63afb5cb6201a197f60f17bca126

Observation b874b35f-e245-4960-842f-c504424062fe · outbound

This paper cites An In-Context Learning Agent for Formal Theorem-Proving.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? An In-Context Learning Agent for Formal Theorem-Proving

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:26.563979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:26.563979Z digest=sha256:fbeb297cc558b61f72e486daa78af9b978d6e7c5742e290389ff30bb78558dac

Observation 03dc5019-849f-4099-864a-cab55bc96b42 · outbound

This paper cites Solving olympiad geometry without human demonstrations.Nature, 625(7995):476–482, 2024.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Solving olympiad geometry without human demonstrations.Nature, 625(7995):476–482, 2024

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:30.288689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:26.633889Z digest=sha256:8d4ff7b938cff0a02effe702a3991ba53819788abe2830699ebbe6a1f9e578d3

Observation d0cab9c9-272d-4b39-9ffc-587b1563e01b · outbound

This paper cites PutnamBench: Evaluating Neural Theorem-Provers on the Putnam Mathematical Competition.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? PutnamBench: Evaluating Neural Theorem-Provers on the Putnam Mathematical Competition

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:26.694429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:26.694429Z digest=sha256:987f5f8561d5e0cd1a7232f6334b8e25e43d97ab0bc3a5edec53d257202d9e28

Observation fcd33a01-a62e-4fd9-9be8-14d4b9b53fac · outbound

This paper cites Proving Theorems Recursively.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Proving Theorems Recursively

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:26.801891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:26.801891Z digest=sha256:8a17544e34dba2145b1fb055c59a87f525f92c93845340ca80cc0137a98902d3

Observation 15779d59-3c4f-42c2-9cd0-896e99a74108 · outbound

This paper cites Measuring multimodal mathematical reasoning with math- vision dataset.Advances in Neural Information Processing Systems, 37:95095–95169,.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Measuring multimodal mathematical reasoning with math- vision dataset.Advances in Neural Information Processing Systems, 37:95095–95169,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:30.128153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:26.868017Z digest=sha256:3b27701baaee2182c1186a7f1f303f5ba919e478c92352be3d44ccea2a8ec183

Observation bc114644-a2fc-4d2d-afaa-9d4a5207644e · outbound

This paper cites MV-MATH: Evaluating Multimodal Math Reasoning in Multi-Visual Contexts.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? MV-MATH: Evaluating Multimodal Math Reasoning in Multi-Visual Contexts

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:27.028489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:27.028489Z digest=sha256:27a3033b9e07ba40953ac110c61643fbef876498bd87ab186d1ade159bf74c5f

Observation 0a9a7827-cbb4-4c0c-87e6-d355e50d3b1a · outbound

This paper cites TheoremLlama: Transforming General-Purpose LLMs into Lean4 Experts.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? TheoremLlama: Transforming General-Purpose LLMs into Lean4 Experts

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:27.144650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:27.144650Z digest=sha256:c5cd803b2c7f52b067da8cd38ada491cec771645a965421030d9087659155bba

Observation 7f4882c0-5c08-4ebb-a227-a3678941626b · outbound

This paper cites Fung, and Tong Zhang.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Fung, and Tong Zhang

Reference 54

Resolution
verified exact
raw_fallback, observed 2026-08-07T06:08:28.602806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:27.227184Z digest=sha256:0ebee66af94bd4f235e254ef33112fdea07a9510a71990c7acc749fab0f734a1

Observation ff394c0c-ea07-48e3-86c5-7240c56add5f · outbound

This paper cites The isabelle framework.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? The isabelle framework

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:29.763857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:27.313857Z digest=sha256:5ee781a56a3525e79edb2fb1b8a64752defff314155dc8bb65f3bd84e0d234a8

Observation cdeb2da0-f271-4d9b-9b11-2df79c5e0364 · outbound

This paper cites Leandojo: Theorem proving with retrieval-augmented language models.Advances in Neural Information Processing Systems, 36:21573–21612, 2023.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Leandojo: Theorem proving with retrieval-augmented language models.Advances in Neural Information Processing Systems, 36:21573–21612, 2023

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:27.815011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:27.815011Z digest=sha256:0966fa369d60a12f34c23049b3f206027fd70a2f5fdcb5bb05e2f06da19f5a86

Observation 094cce33-edbc-4774-b7e6-68681c8359da · outbound

This paper cites Lean Workbook: A large-scale Lean problem set formalized from natural language math problems.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Lean Workbook: A large-scale Lean problem set formalized from natural language math problems

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:27.915796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:27.915796Z digest=sha256:cda268701804f4f778b3142e6fbaa62067265ee85583d7bb1e68e44e309502c4

Observation 0ae23a6e-440f-4b82-a9ca-6af557a91687 · outbound

This paper cites Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:29.601101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:28.013211Z digest=sha256:ed6c86b45ca2cbee536528e00d9ff449c04501b908a59e9559bf6d0c54c387a3

Observation 6bbee102-0225-4836-9239-3276f462bdd6 · outbound

This paper cites Mathverse: Does your multi-modal llm truly see the diagrams in visual math problems? InEuropean Conference on Computer Vision, pages 169–186.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Mathverse: Does your multi-modal llm truly see the diagrams in visual math problems? InEuropean Conference on Computer Vision, pages 169–186

Reference 61

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T06:08:29.380731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:28.099080Z digest=sha256:42369ecd08a703e30c01bbf975c9dc75c370bb2f123e0fee25bb5633d4158276

Observation 7d941f0c-0830-4b49-8308-69e3209a41f7 · outbound

This paper cites MiniF2F: a cross-system benchmark for formal Olympiad-level mathematics.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? MiniF2F: a cross-system benchmark for formal Olympiad-level mathematics

Reference 62

Resolution
malformed identifier
no resolver link, observed 2026-08-07T06:08:28.160898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:28.160898Z digest=sha256:2443dbe236f31c56bc95dc2f0bda442bf19d592acdfba1f23d7f1a78b95c5ad9

Observation 31f0da6c-8af6-4be9-8cf0-71392c1708dc · outbound

This paper cites DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:27.508982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:27.508982Z digest=sha256:0a270ce0c27eb10505987e2a1ad26c469e26fd02057b677ad406c55075fcd571

Observation aa45ff06-b898-4bc1-b725-c9796301a215 · outbound

This paper cites Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:27.708050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:27.708050Z digest=sha256:524d1f5452dab8a693e221052074ac30b8c6be6f01a272833f795b3f9d04b9d1

Observation e5580767-efeb-4fd3-b020-f941661aaa9b · outbound

This paper cites Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:08:29.190739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:28.246141Z digest=sha256:4b495ee073cb504d347d798f612a22718f1b4c0f25483c5a759e1330e3bd8d11

Observation 092f1825-15f8-44eb-884c-dda23176ad6a · outbound

This paper cites an unresolved cited work.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Unresolved cited work

Reference 2001

Resolution
unresolved
raw_fallback, observed 2026-08-07T06:08:30.971403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:26.033854Z digest=sha256:9cd69017bdc7240ef3ee20773f6af4e3f9d6dbb2c92e38cc3f8a93f344773e5c

Observation 34ed8203-6c5e-46ce-809e-84c419f43418 · outbound

This paper cites an unresolved cited work.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Unresolved cited work

Reference 2008

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:27.379237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:27.379237Z digest=sha256:7cfd07509056214721701e6719df8b7074ccf169c55f1d7f0e8bcdd280c727c6

Observation d582042d-f4dd-4256-aa62-48b20b6fda73 · outbound

This paper cites Formal Mathematics Statement Curriculum Learning.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Formal Mathematics Statement Curriculum Learning

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:25.636471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:25.636471Z digest=sha256:6685c69a2a0dbe956d03cef96c5ef55fccaeebffe6e1546f939438369bb60136

Observation 931d50d9-794f-491f-a820-75208c7e1060 · outbound

This paper cites an unresolved cited work.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Unresolved cited work

Reference 2023

Resolution
unresolved
raw_fallback, observed 2026-08-07T06:08:32.898395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:23.249411Z digest=sha256:797250053eeea6d7d55c7a5cce3dce84c3db2fd8117c7f69df5852e843d7f7ba

Observation 71a7e03a-04b1-436d-b31f-50a16ec99525 · outbound

This paper cites an unresolved cited work.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Unresolved cited work

Reference 2024

Resolution
unresolved
raw_fallback, observed 2026-08-07T06:08:29.941786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T06:08:26.951158Z digest=sha256:0c97b93f367be9017e4145692ae2232dc4652f4cf308c6f0b4c9d02ffc119996

Observation c1aa5b9c-eddd-4e06-b4cc-64f5e96b9d71 · outbound

This paper cites an unresolved cited work.

MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems? Unresolved cited work

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T06:08:23.445192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:08:23.445192Z digest=sha256:4225b8c1c713d2f8826419899cc33388431a0e15f4b6834e6368436d785926c9

Pith citing papers

Observation 1724996d-a25d-4c98-8a72-ba5310689128 · inbound

MAC-Tuning: LLM Multi-Compositional Problem Reasoning with Enhanced Knowledge Boundary Awareness cites this paper.

MAC-Tuning: LLM Multi-Compositional Problem Reasoning with Enhanced Knowledge Boundary Awareness MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems?

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T04:59:38.760845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:59:38.760845Z digest=sha256:3d3aa4ef23784fabf2fc3b5b4e757c376a3b4a7b00c3b9b4fdebeb60e00931a4

Observation 589d9889-1a60-42a7-8ffc-cf3a2b2f9b8a · inbound

Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models cites this paper.

Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems?

Reference 158

Resolution
unresolved
no resolver link, observed 2026-08-15T23:21:12.632396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:21:12.632396Z digest=sha256:b0adc193787a2fcc3e9bd418462d831c1c35a5c5f85477b166c99193197bd8a6

Observation 5d4f9b23-04e4-4cda-83df-d3f0e84c26d8 · inbound

PAGER: Bridging the Semantic-Execution Gap in Point-Precise Geometric GUI Control cites this paper.

PAGER: Bridging the Semantic-Execution Gap in Point-Precise Geometric GUI Control MATP-BENCH: Can MLLM Be a Good Automated Theorem Prover for Multimodal Problems?

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-20T18:48:53.332391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-20T18:46:20.417039Z digest=sha256:39233718d13fcd2eb010a761276c5426b68d3adc9a0703bf6db3327053757e0a