Pith. sign in

Paper Citation Record · LEDGER

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models

As of 9 August 2026, this Paper Citation Record lists 70 of 70 outbound references and 11 inbound Pith citation observations for arXiv:2502.00698.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.00698 v2

Coverage vector

measured 70 of 70 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T18:06:04.220507Z

measured 81 of 81 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:31:19.801487Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T16:58:42.493790Z

Reference resolution

70 of 70 outbound references displayed

  • verified exact0
  • verified fuzzy25
  • unresolved45
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fc6d63ee-f459-4e6f-972f-e1365e6caf8c · outbound

This paper cites GPT-4 Technical Report.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:03.886634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:03.886634Z digest=sha256:7252cf8febf22d1418209ad2c7f04b3b7b841149f39eb22c6e94fddb4ff4235a

Observation bda3e211-9671-43e7-b55b-1a42373c7a73 · outbound

This paper cites The Claude 3 Model Family: Opus, Sonnet, Haiku.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models The Claude 3 Model Family: Opus, Sonnet, Haiku

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:05.288313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:03.892117Z digest=sha256:4f407eb165e6cdbf86d5e986f1d3fb71152186c38025f81778ef89528fb5f7a3

Observation dffe1d6e-a602-4e4d-887c-f1540693059a · outbound

This paper cites Qwen2.5-VL Technical Report.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Qwen2.5-VL Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:03.897113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:03.897113Z digest=sha256:8ec8eb6f8c9dac6175c1c677e452bc969faddacdcc05c31f5fe626470de4b8dd

Observation 6838ee51-a742-48c7-a6c8-f2936e1f5ba0 · outbound

This paper cites DeepSeek LLM: Scaling Open-Source Language Models with Longtermism.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models DeepSeek LLM: Scaling Open-Source Language Models with Longtermism

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:03.902859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:03.902859Z digest=sha256:d6c1d37da3bd3fff62ce7a623b0e77cb595f4ce0a57ed78071874110bf0f9a74

Observation 8a11ddc1-4f72-4e5d-b363-110f19e12ab5 · outbound

This paper cites SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:03.908096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:03.908096Z digest=sha256:a951a2e22b4c0623676e48c5d54356416a3888f8c55a5ed51367f7f9d5673b5c

Observation 358cd1c6-2971-48ec-b42d-c0c5e8ef1240 · outbound

This paper cites On the Measure of Intelligence.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models On the Measure of Intelligence

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:03.913307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:03.913307Z digest=sha256:f77aee0c5d7d379b547a75f57116babe07a16514d0df502cb68d993445ef0226

Observation 214146f6-a64a-48f4-b495-23b53bc898ea · outbound

This paper cites Comparing machines and humans on a visual categorization test.Proceedings of the National Academy of Sciences, 108(43):17621–17625, 2011.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Comparing machines and humans on a visual categorization test.Proceedings of the National Academy of Sciences, 108(43):17621–17625, 2011

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:05.272176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:03.919052Z digest=sha256:f96aded96f01933e56ac07f283b8f2afd2ebd14d34c72ddd39cfb123568b22c7

Observation 2b16b301-c08c-48e7-8233-95fb54efcc23 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:03.923444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:03.923444Z digest=sha256:eaebf1191bc2dec8715181351821073c0850515036e1f77bc58bcec3a7d35b58

Observation 74444b0b-af22-4982-a8af-78c2b047bdf3 · outbound

This paper cites Learning to Make Analogies by Contrasting Abstract Relational Structure.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Learning to Make Analogies by Contrasting Abstract Relational Structure

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:03.928372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:03.928372Z digest=sha256:88a026ce4ade93eb7bde988456434b8555783571bce8969552f943acd42f035c

Observation 6fddcb76-3ab2-4f5c-9b99-fe8aa811847c · outbound

This paper cites Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:03.933538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:03.933538Z digest=sha256:59fbc6d080e7b9c349fab185bd6d37e1e10ceb09728315dc2187f9c2c035b693

Observation 4deb3423-5cbe-4ebb-8166-64d2c7db421c · outbound

This paper cites OpenAI o1 System Card.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models OpenAI o1 System Card

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:03.938657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:03.938657Z digest=sha256:ce3b92093cb10b1703ac5dc10b06fa8d740782c8824ee690704594861cb7e25a

Observation 3ef6e551-24ea-4a67-8fcf-528de41938d9 · outbound

This paper cites MARVEL: Multidimensional Abstraction and Reasoning through Visual Evaluation and Learning.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models MARVEL: Multidimensional Abstraction and Reasoning through Visual Evaluation and Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:03.943931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:03.943931Z digest=sha256:6dd023eff0bf038d328d4b43d8afa84241ae60c8811fbf48a4287b1fde84f0f6

Observation 38664aa0-e1c5-4112-b5ee-8646e409633a · outbound

This paper cites Llava-next: Improved reasoning, ocr, and world knowledge, January 2024.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Llava-next: Improved reasoning, ocr, and world knowledge, January 2024

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:03.949106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:03.949106Z digest=sha256:4fad44fdec0a41a55cb3813aa1e9f201fd399483d3106392c77e76df9bf5f12c

Observation 86c2aa38-5944-4b0d-8e6b-4603f4b50d5d · outbound

This paper cites LogiQA: A Challenge Dataset for Machine Reading Comprehension with Logical Reasoning.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models LogiQA: A Challenge Dataset for Machine Reading Comprehension with Logical Reasoning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:03.953539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:03.953539Z digest=sha256:dead3318c73d91f18ea074ca88add382940d15763d62a0154e9fe3deb816dbc4

Observation 67b9f4cd-4f4b-4bd7-a732-443654d26866 · outbound

This paper cites Mmbench: Is your multi-modal model an all-around player? InEuropean conference on computer vision, pages 216–233.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Mmbench: Is your multi-modal model an all-around player? InEuropean conference on computer vision, pages 216–233

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:03.958420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:03.958420Z digest=sha256:43bb8afba64ee6def7bf1d06132642a36f3dbd19f3ba535a92baff1ddb5d41b2

Observation a15d39fc-d06e-4b6d-9242-c26b4e0c844b · outbound

This paper cites OCRBench: On the Hidden Mystery of OCR in Large Multimodal Models.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models OCRBench: On the Hidden Mystery of OCR in Large Multimodal Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:03.962984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:03.962984Z digest=sha256:1c590bf13c3fae69ed3ac33dd6cc2e8519b8df8c02b69526a68133d5fe2d98f8

Observation cab5201e-5ff8-449e-860c-17500c72fe40 · outbound

This paper cites MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:03.967371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:03.967371Z digest=sha256:79931700b65a1e7c238569ae4cf30bf968a5480a1898d4d3be7c0781018de552

Observation 3943d110-a7db-467c-a159-9b4d43e73e40 · outbound

This paper cites Deep Learning Methods for Abstract Visual Reasoning: A Survey on Raven's Progressive Matrices.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Deep Learning Methods for Abstract Visual Reasoning: A Survey on Raven's Progressive Matrices

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:03.972475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:03.972475Z digest=sha256:180a431923e791065b3856f8989f5e8b39d3206704775b7f48a30792aa57d47d

Observation 4d10e9e2-5d7d-4ff6-aeef-81b57408ccdc · outbound

This paper cites A review of emerging research directions in abstract visual reasoning.Information Fusion, 91:713–736, 2023.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models A review of emerging research directions in abstract visual reasoning.Information Fusion, 91:713–736, 2023

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:05.236885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:03.977070Z digest=sha256:5bdd836d49160467d34317867579b0c68b24a003805a51592b8544bd1edcf1ce

Observation 0c2f7318-fc21-4c79-84d7-ae9db69c6ec6 · outbound

This paper cites Deepiq: A human-inspired ai system for solving iq test problems.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Deepiq: A human-inspired ai system for solving iq test problems

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:05.221324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:03.981759Z digest=sha256:193d7954c516e18fe4c2c2df3b3783ddac7b371420a7c5be00b8e29169afe96c

Observation e4b5421e-9679-48ab-a2f8-be58931aba29 · outbound

This paper cites Bongard- logo: A new benchmark for human-level concept learning and reasoning.Advances in Neural Information Processing Systems, 33:16468–16480, 2020.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Bongard- logo: A new benchmark for human-level concept learning and reasoning.Advances in Neural Information Processing Systems, 33:16468–16480, 2020

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:05.206273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:03.986259Z digest=sha256:61ec7c361aba29f7299a54ce93880a5940a0803667a54a240c0818761e047f9f

Observation 396a489b-d404-4af2-9e39-0144ed3e722a · outbound

This paper cites Learning transferable visual models from natural language supervision.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Learning transferable visual models from natural language supervision

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:03.990607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:03.990607Z digest=sha256:c43d8e93302d1fd4c751502c065ab2a8729f476d0ca2adaf368af4301264f554

Observation 26433668-27e8-49cd-a00f-8233b7ac134c · outbound

This paper cites Raven progressive matrices.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Raven progressive matrices

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:05.180552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:03.995165Z digest=sha256:d63a40ad58b40033e729cf592ea179b3729eaab6421df3f0608b5ed862770d02

Observation ed51ad7c-6072-42c0-9b06-fd3432a68182 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:03.999938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:03.999938Z digest=sha256:a680716edc89468b84680fac359bd229b95fbac0278e9d83899a042a3f7e9b8e

Observation 24999903-3709-44ec-ba3d-0da7622c8660 · outbound

This paper cites The topography of ability and learning correlations.Advances in the psychology of human intelligence/Erlbaum, 1984.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models The topography of ability and learning correlations.Advances in the psychology of human intelligence/Erlbaum, 1984

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:05.165436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.004904Z digest=sha256:66a75b569ecf2870ac617b30d749018ad3373003a8b77bb246fd55ca60f18433

Observation 71676622-1760-43da-a4df-7d0f17b61eaf · outbound

This paper cites Link-context learning for multimodal llms.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Link-context learning for multimodal llms

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:05.149595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.009645Z digest=sha256:7084a6dc53964afb489cee8925919ccbc57bb8128b6795730e22d0e29a84dce3

Observation 260dae39-3c78-4443-b21c-d056ea0a4d2c · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Gemini: A Family of Highly Capable Multimodal Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:04.014290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:04.014290Z digest=sha256:f159f8ee8280960dfc441db36e3a2a31ef7ef6c988f12eb427d1e3b60981d04d

Observation d585e0d6-6e48-4864-ac66-920e0e235b9d · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:04.019501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:04.019501Z digest=sha256:f5eaf775a70e855d7e31195a906367bf6456cbbfdc674650b51d6cf08f5e3cb7

Observation cca9cd4b-9da9-4032-94bf-49c63db57838 · outbound

This paper cites Qvq: To see the world with wisdom, December 2024.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Qvq: To see the world with wisdom, December 2024

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:04.024353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:04.024353Z digest=sha256:c107615e2d4528225cddd0d8a6432e83939cf84ea344722d4f3d1e8635ac8581

Observation 2bf3c440-69e0-4c2f-964f-01ec255c6499 · outbound

This paper cites How much intelligence is there in artificial intelligence? a 2020 update.Intelligence, 87:101548, 2021.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models How much intelligence is there in artificial intelligence? a 2020 update.Intelligence, 87:101548, 2021

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:05.124995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.028818Z digest=sha256:036cf3171064577720f26e92f23cc4a8d583a8d55c2c170f9fcdee50bb994ef2

Observation cffd74f7-de87-4028-9f86-499fb68d939a · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:04.033388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:04.033388Z digest=sha256:d3c2a8612cf97bd9350378ca4cc285fbfc7eb8389431ef87a59ca17957363e7d

Observation ab7c8d52-68ca-4410-817b-e2a9772b1dd9 · outbound

This paper cites Learning representations that support extrapolation.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Learning representations that support extrapolation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:05.109716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.038369Z digest=sha256:c1f6cc2bfd19d5e4ca44404af76b386218ad09b3a4346f49591a83e7e86b44d3

Observation a64fa8b0-7a57-4dd3-9583-863d39926c8f · outbound

This paper cites A Survey on Multimodal Large Language Models.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models A Survey on Multimodal Large Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:04.042938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:04.042938Z digest=sha256:7f7698674deaad1b2d39475139097d7b28c5a3e176eddd58c79441cc01cd57d7

Observation 031b8d0c-73c5-46ea-8651-a9691123dc72 · outbound

This paper cites Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:04.047529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:04.047529Z digest=sha256:e5dc6d154fe14a3b510090a8c5a057b610c452c390c27c51d76f9e5e03ad76d9

Observation 195ca3f1-d2c1-48cd-bb28-255d2332bfaa · outbound

This paper cites Raven: A dataset for relational and analogical visual reasoning.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Raven: A dataset for relational and analogical visual reasoning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:05.084588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.052333Z digest=sha256:fd4b8d7e65bdb109e938d81caf71a13d40aa8e54825302a5078d131a366941ed

Observation 95d1fb53-049e-4b5d-b539-15b734a0a2cb · outbound

This paper cites CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T18:06:04.057379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:06:04.057379Z digest=sha256:09248a51d262f7d496778613be757e0d4b2d48d64b9d378f64498b48abb5afbf

Observation bba131e5-a45e-4281-ab74-fb86208af389 · outbound

This paper cites Machine number sense: A dataset of visual arithmetic problems for abstract and relational reasoning.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Machine number sense: A dataset of visual arithmetic problems for abstract and relational reasoning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:05.069608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.062557Z digest=sha256:74aa265a71c0174025f21e7d91d89b2308ae698bdec05473fdb45b047aed29e8

Observation 72cc5841-c2ff-466b-837f-a15c5bf151ca · outbound

This paper cites Easyr1: An efficient, scalable, multi-modality rl training framework.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Easyr1: An efficient, scalable, multi-modality rl training framework

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:05.055031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.067238Z digest=sha256:131a3cd2fc0aecb8fe5d82c1ca146e0b58b20780082301121bf33cc8ccec76e1

Observation 1f91ed44-2fe0-4c90-bd2b-ff72933dc929 · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:05.039700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.072463Z digest=sha256:1c74f520cc3107f0a8623d87819d154e4b0e5cf97275877aeb9aee23f783dcb9

Observation 04c2621d-b498-47e5-9c89-fa3948b14f11 · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:05.025006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.077441Z digest=sha256:2505e5aacaa7a2cb356ce8b9d40300bd05973ceb92fffa085111e88351f6a077

Observation e28739c9-a3c2-444f-be06-e618f855b23d · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:05.010495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.082238Z digest=sha256:6586c11bb176abf4840f2830f6b445df0f069fb3ccc2a98fa6ab7a07d98c2f74

Observation 49e4fd93-5e76-4ecc-945e-20c4963d3aad · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:04.995615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.086995Z digest=sha256:c27d9dd9209ba1af9d9477abc54b624fd694f2ed7c162732f3f855fb43c19199

Observation ed686d0c-c8e7-47a2-98d4-476723a63c69 · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:04.980965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.091982Z digest=sha256:dc15990e796189462ce4f60efa7caf00ae801c751b6ffe685ed2e5b05589062a

Observation d0a2b464-e376-431d-9961-f87c70482cfe · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:04.966469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.097359Z digest=sha256:e1cfa1d988d799fab595674356dc292fa4bb226035eb223c89f49294ae6b0e17

Observation 8fcd1032-48d2-47db-9292-c51fde1ca50f · outbound

This paper cites The pattern seems to be categorization based on function or use.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models The pattern seems to be categorization based on function or use

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:04.951879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.101879Z digest=sha256:af39e875f77a5669caa3aba5aab9eb210425757f07b21c6d065569cccd02a534

Observation 70c0bb7d-9284-4dae-9818-7f111de08f72 · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:04.937339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.106626Z digest=sha256:a86f587b32c7613e63fac31939071147a02e4a39749a0d2577a4cfe797eaa7c8

Observation aa4eed75-4a5f-4939-abed-b1f0d64cbb45 · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:04.923086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.111388Z digest=sha256:5aadcb02f134302ce13fa3396830c591815af32318c9ebad53e93626526e015c

Observation 88c8f5b5-4ea9-4428-ba56-82e3597810e8 · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:04.909056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.115909Z digest=sha256:783fb36403bd2e36d47c35fced8687985495b03a44725960a2890df9fa32dae9

Observation 80d2efe4-6d31-4b6d-b4f3-7cee78bec5ac · outbound

This paper cites The pattern seems to be increasing the number of points by one each time, while also increasing the complexity of the connections between these points.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models The pattern seems to be increasing the number of points by one each time, while also increasing the complexity of the connections between these points

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:04.894331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.120681Z digest=sha256:91eaefb51608a6b384d8538617d2bc45e91ca74cbaeab80e6bd5289009617b70

Observation bcb80351-fbc7-4646-be45-2b46b734593a · outbound

This paper cites They seem to rotate or shift.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models They seem to rotate or shift

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:04.878976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.125528Z digest=sha256:3b9f02f9bec64cc9c6cbd0cebc974254a35661563970507c07ab7ec73c71cf8c

Observation 148f5b3b-092a-47c1-beee-41d9e66c12eb · outbound

This paper cites There’s a sort of flipping or swapping happening.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models There’s a sort of flipping or swapping happening

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:04.863995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.130293Z digest=sha256:824fac03a36c894853d0b5095d04c7b41e9630d79ac96a7a2160fe093b155a6b

Observation 7adc02e0-94b6-4ce5-b3c8-a572a16bd324 · outbound

This paper cites If you follow the observed changes in connections and circle colors, option C best fits the next step in the sequence.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models If you follow the observed changes in connections and circle colors, option C best fits the next step in the sequence

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:04.848200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.135674Z digest=sha256:b700296e6661896ad60a5682d0e22bd77b0c2cad7cf229d87a754b20a1537c73

Observation b7c2639c-8f6c-4ef7-b581-3257db5fa642 · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:04.832798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.140755Z digest=sha256:c1f407ed86b942d9718e38d03d6701684c8a4a036c1a5dbadf9b416e165b1c0e

Observation b9a7217d-1e9b-4699-99d5-f24a4f3b883f · outbound

This paper cites The missing square needs a mixed pair of circles and a triple line base to satisfy both rules within its row and column.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models The missing square needs a mixed pair of circles and a triple line base to satisfy both rules within its row and column

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:04.818614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.145419Z digest=sha256:c42fc5d17cdf1124319322480eccb553b9543c706fc613fa70067634acd430db

Observation 1c4fc662-ed96-45d9-a27b-a83197439836 · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:04.803921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.150544Z digest=sha256:c432975be24361c09beaf3edb9884386cb726a072d50ee0d02f7c42abd8be153

Observation f8353fd6-2f21-48c2-888f-597013efa831 · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:04.789868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.155361Z digest=sha256:933bd97e4adbdac61e57f6ad34a5c893ad9a9b9fbfa350dedfdd38cddb0be5fb

Observation 18fc108f-fb67-46b0-914b-e857c42a21ea · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:04.775890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.159972Z digest=sha256:628190788220ee084afa0446f4d9722569192ca183bdb1c89af59d75452a47f5

Observation 5a2b98ac-677f-48cd-977e-c16421488f68 · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:04.761354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.164532Z digest=sha256:91c6d5f1f7ecab14fee69435090a7899056808e996020519417240955597ceb1

Observation f059f19c-71cd-4347-82ec-c90b5d8b042e · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:04.746513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.169285Z digest=sha256:9be4959979049d6bd590486b615118cd4a3b2362157813dd78c1fc1bbe38c062

Observation 236d2c45-d6a5-4e9e-90ca-54aff8f02a41 · outbound

This paper cites Now, let’s categorize them: - Category 1: Figures with shapes pointing in a specific direction.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Now, let’s categorize them: - Category 1: Figures with shapes pointing in a specific direction

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:04.731935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.174242Z digest=sha256:1d81a798dd480640cadcc73513fc4a2c1f74ec54f342155d72402af543bea5dc

Observation e021564c-90c6-499c-a5ae-3f688d62b211 · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:04.716964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.178738Z digest=sha256:d4c5a21a8a8cbe826b6865d7092ccc6e69b7e46e48a8ed16220976565f25c67b

Observation 4e4815ff-b3e4-45b6-a7a0-035acb71aeeb · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:04.702286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.183679Z digest=sha256:10cc7678ad3eb339f60e22d9db476b01ab669860c9d4d7d1a116d391cc2c2f8e

Observation 1e0a9b70-0821-44c6-b365-0421305417c2 · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:04.687134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.188175Z digest=sha256:dd0af76a9f0e6f3b1b151e7a8876984203b19af992b2e073dc30262ff0d5c97c

Observation 753689fd-775c-4485-97dc-336d6b2faf06 · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:04.672023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.192497Z digest=sha256:44ce5fd07a6725ea73c3ff1e1b36c8f52c0fb4585ff8eca150f8fced6fc1725a

Observation dbc16bfe-3120-454c-85b8-706fcb29261e · outbound

This paper cites an unresolved cited work.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:06:04.654326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.197043Z digest=sha256:97c11c9ec01a4912f25d2964e912dcebc446dbcd8ba05be6e1097e70f833b429

Observation 78cede7f-745c-4727-b1e5-93c9291830fb · outbound

This paper cites Now, let’s group them: - Group 1: Figures that are symmetrical or have a clear pattern.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models Now, let’s group them: - Group 1: Figures that are symmetrical or have a clear pattern

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:04.637270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.201568Z digest=sha256:f6fb7dc7ff802981cb1fc788d582a662c60c2ce1b5cba49e255a2c668cc058d3

Observation 2bc40f9a-e8a6-491e-9946-171821119ee7 · outbound

This paper cites The middle number (4) is obtained by adding the numbers on opposite sides: 3 + 5 = 8, 6 + 2 = 8, and 4 + 4 = 8.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models The middle number (4) is obtained by adding the numbers on opposite sides: 3 + 5 = 8, 6 + 2 = 8, and 4 + 4 = 8

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:04.621946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.206578Z digest=sha256:8e839f1955a1474be5ed9c1b3ebf63486e78491c99c1615567c41929a76a5477

Observation f5339b44-af68-499e-a8b3-b6642ac9aaa8 · outbound

This paper cites The middle number (6) is obtained by adding the products of the numbers on opposite sides: 15 * 5 = 75, 12 * 4 = 48, and 6 * 6 = 36.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models The middle number (6) is obtained by adding the products of the numbers on opposite sides: 15 * 5 = 75, 12 * 4 = 48, and 6 * 6 = 36

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:04.605817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.211342Z digest=sha256:fc38175ae8f2b1a9983401db2345c9e25f99f37871bd571bb50fe97a9fc4f5a2

Observation 90933992-c2b7-4714-8277-b5bdbcda446f · outbound

This paper cites The middle number (7) is obtained by adding the differences of the numbers on opposite sides: 24 - 14 = 10, 6 - 5 = 1, and 7 - 7 = 0.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models The middle number (7) is obtained by adding the differences of the numbers on opposite sides: 24 - 14 = 10, 6 - 5 = 1, and 7 - 7 = 0

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:04.590248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.215882Z digest=sha256:5e9303c2d58edaf3649da8b8e0a722dc356bf0b44cce66add72e7e776ac1771d

Observation a74069cb-9fe7-420b-bd69-8714d5423489 · outbound

This paper cites The middle number is obtained by adding the quotients of the numbers on opposite sides: 1 / 4 = 0.25, 5 / 12 = 0.4167, and 12 / 5 = 2.4.

MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models The middle number is obtained by adding the quotients of the numbers on opposite sides: 1 / 4 = 0.25, 5 / 12 = 0.4167, and 12 / 5 = 2.4

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:06:04.573802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T18:06:04.220507Z digest=sha256:45647b96fbfe3357e42fa55e1fc3668049f7ea66ffc45dfa6bf509502a048ecc

Pith citing papers

Observation c71a46a2-925a-48c9-80c1-159d570db7ad · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models

Reference 128

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:19.801487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:19.801487Z digest=sha256:2e226b7f2cb63b385ef2de37779bf630ee850fe1e590b0e38ff42fc682b731c3

Observation 60d8943b-51d0-4c67-8ab1-fbec4f246770 · inbound

MME-Reasoning: A Comprehensive Benchmark for Logical Reasoning in MLLMs cites this paper.

MME-Reasoning: A Comprehensive Benchmark for Logical Reasoning in MLLMs MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:35:55.549583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:35:55.549583Z digest=sha256:45831e3cb0229e828a0166a05cf2213c75a1c8d62f908bcc0fdbeb7c5fee628c

Observation 0ac1d8d1-47ac-46b8-a4fa-ba1fb0af3714 · inbound

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM cites this paper.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.268894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.268894Z digest=sha256:408201f012544077eb6d3fbf49e5ae05bb930f5cc07e18f15c5cf23d0fa67392

Observation 9c436822-d943-4e1f-a1c2-924f170c08be · inbound

Boosting Vulnerability Detection of LLMs via Curriculum Preference Optimization with Synthetic Reasoning Data cites this paper.

Boosting Vulnerability Detection of LLMs via Curriculum Preference Optimization with Synthetic Reasoning Data MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models

Reference 4455

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:54.389863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:54.389863Z digest=sha256:c6bfd6c86565e4ee54ff2941dfb5eed03a74a3d8b458312da95330cf069f17d7

Observation 21186ae1-501d-4c0f-aca8-09281ecef484 · inbound

VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning cites this paper.

VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T11:35:10.875843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:35:10.875843Z digest=sha256:0e7a4885ef1275cb08524cebbfdeef1b21f3cff9cb8dfb1d476beb2b73a21426

Observation 82d4b4d0-877f-47f0-855d-f7272efa9c0f · inbound

Physical Plausibility Reasoning via HCM-GRPO: Empowering Compact Model for Superior Performance cites this paper.

Physical Plausibility Reasoning via HCM-GRPO: Empowering Compact Model for Superior Performance MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T22:36:09.932091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:36:09.932091Z digest=sha256:b5d1b5aa11290c9010ee286f6bf57e5e338c4203f247bf33165e44f57ccb9d03

Observation 29de4c11-ee20-4c29-b729-08c2b809fffa · inbound

Enhancing Multimodal In-Context Learning via Inductive-Deductive Reasoning cites this paper.

Enhancing Multimodal In-Context Learning via Inductive-Deductive Reasoning MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-09T05:55:31.190380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T19:21:30.235583Z digest=sha256:80065303f4e7309c0cc791195519f3d1a4a5487b69241bfb03772394572b5191

Observation 226454fd-a7e1-45c9-9446-a7d9a9157029 · inbound

StemBind: When MLLMs Get Lost Between Rules and Instances in Abstract Visual Reasoning cites this paper.

StemBind: When MLLMs Get Lost Between Rules and Instances in Abstract Visual Reasoning MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-29T00:12:50.337644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T23:15:56.598968Z digest=sha256:536a05e4ce7a194c57b886ed766dde2f6a716859bb4d9087f9702598938536e8

Observation 59705aeb-c08c-4a07-b580-a35532272c24 · inbound

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes cites this paper.

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:57:41.642430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T12:59:51.091008Z digest=sha256:18fbcff68a7b2d5ce82e2a90041d450cc32c028a81f25a85ce5ad794c9e42188

Observation 0f024524-d78f-4914-bc05-efeb83432add · inbound

LASER: A Corrective Lens for LVLMs via Visual Attention Preservation and Sink Suppression cites this paper.

LASER: A Corrective Lens for LVLMs via Visual Attention Preservation and Sink Suppression MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T16:58:42.495183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-03T16:53:43.427625Z digest=sha256:4ef4620c7ad291308db42e1bcda8b93c46d0028d194c89b1c2b01f5e0eef88ec

Observation cefddff5-b290-4d9e-8837-d75c549dc05b · inbound

Boogu-Image-0.1: Boosting Open Agentic Multimodal Generation via Understanding under a Minimal Budget cites this paper.

Boogu-Image-0.1: Boosting Open Agentic Multimodal Generation via Understanding under a Minimal Budget MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T06:13:50.663556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:13:50.663556Z digest=sha256:a684b3fcdd8dc8cb0d34c36e3a99ddec26e97ea6799b812b74ccbb9c66cb7cc7