Pith. sign in

Paper Citation Record · LEDGER

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models

As of 9 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 4 inbound Pith citation observations for arXiv:2505.23091.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.23091 v3

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:59:23.141014Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-03T16:53:43.427625Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:39:46.064161Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved30
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 44d759aa-6212-40d0-bb87-aa29433b17ce · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:19.315667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:19.315667Z digest=sha256:159768e97035f93af8948d2b39c8644fac636799f46c34d4241aa0f57fd94b30

Observation 65a29c40-e7d4-4c9a-b8de-0221cfd36bf8 · outbound

This paper cites Qwen2 Technical Report.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Qwen2 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:19.431681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:19.431681Z digest=sha256:29eb4b3c897f54314e26a470cd33410fae3753afb4c34439e64b291377a32568

Observation 9cb0006b-35d7-4660-ba07-8f10ad333518 · outbound

This paper cites OpenAI o1 System Card.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models OpenAI o1 System Card

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:19.523089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:19.523089Z digest=sha256:01e89e79c1e6f2f70a62ddfba84a864bcd379334d4cfbbb8693ce735c4378d72

Observation c01072d5-3cee-4766-bfb9-26bbf4b826a7 · outbound

This paper cites Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:19.624700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:19.624700Z digest=sha256:7a45a36df2c8c85fb7cc545f412a512e9a8ebd4314f6cffddd1c2b968b6f8d36

Observation 0c95002a-6c4f-4b32-9bdb-c7906314ef78 · outbound

This paper cites Kosaraju, Y.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Kosaraju, Y

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:25.681023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:59:19.764183Z digest=sha256:50f6a5a01bcf1853eec6fd98b72074e97b7c01c4c3159ad9c37034c9a3a64f43

Observation 19d83683-c19d-4022-ad09-68fcb74e33fb · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:25.531225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:59:19.906443Z digest=sha256:c44e907e8dc5f5121976889d3e5f9cb228c0a5cc84535cef166ba228b5f7e3cc

Observation 7711fb21-8c83-49c6-8593-de6873b052a7 · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:25.393970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:59:20.054294Z digest=sha256:eba8a7186b956f853e2d708657b8d957e3e88fecdd6155e3e3c8bc296eb47f6d

Observation 042b01ec-2aec-4b7e-b9d0-fc7cabbb8249 · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:20.185816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:20.185816Z digest=sha256:382a635faf4b5a55587fbc067c64202f4c35a9e5999f9f481680eb86c6a8e99b

Observation 39c6486a-1a54-499d-a35c-aff0c2dd67a6 · outbound

This paper cites Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:20.280238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:20.280238Z digest=sha256:1ab0b822c8a7c6539312b3da5da8d841a793a0922cc79ff90d42debee191a641

Observation dc87b2d7-493c-4505-8140-ad931d9fce0b · outbound

This paper cites Donahue, P.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Donahue, P

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:25.277908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:59:20.410033Z digest=sha256:0f168316a04b24f9cca3ce9c75653079ac6f3a8ac890936498ec85d897334a1c

Observation 654ae170-2358-4d99-874b-ea1b6f4f9f42 · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:25.183689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:59:20.492336Z digest=sha256:94668348ec6e2aace78062672d95989ab689150c2ec6bb2ccb4969c4cf7baf96

Observation bbc362b5-e8fe-4c70-8155-8d80da3ff91d · outbound

This paper cites Improve Vision Language Model Chain-of-thought Reasoning.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Improve Vision Language Model Chain-of-thought Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:20.586668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:20.586668Z digest=sha256:9e65834c5617149ba8dd837e613fa08f6d18085ae1d4028b1e3829764964540d

Observation 55887e46-2bf9-4415-9ab0-450044554533 · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:25.067268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:59:20.666995Z digest=sha256:a8c6d82c3af4cf7af98f753ac01f0984f40e9be607ceb99fcefb297bfc33f12a

Observation 31096d2e-7c27-41ac-a0ea-5011229d128f · outbound

This paper cites R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:20.808202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:20.808202Z digest=sha256:fc2a10898ccc879bd2f5f605f229035adf4498988f34412726c9b6e5e4e260d4

Observation 1f49ad59-bbbc-40c1-8dd9-2b7c2e7e9b40 · outbound

This paper cites X-Reasoner: Towards Generalizable Reasoning Across Modalities and Domains.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models X-Reasoner: Towards Generalizable Reasoning Across Modalities and Domains

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:20.896118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:20.896118Z digest=sha256:8d1b8a6a019a710a2c88af5c2269134856dde36235ced8bdab24d181302783f6

Observation e4014433-8b02-47a4-84e6-9bccf0045322 · outbound

This paper cites LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:20.994541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:20.994541Z digest=sha256:f67b7e70f33742b8d529930e1146bed31895d8e8369ba09bd0ce3a1e1a629be7

Observation 5415c9ba-dfb6-4ca7-98f3-f57147b36ef6 · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:24.879439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:59:21.107805Z digest=sha256:330b780d6e5ae55ccc08e58a97e141de95e2553864fbb78c8de3271ad552135f

Observation be86a894-3054-417b-b73c-fc6fb900a8d4 · outbound

This paper cites Competence-based Curriculum Learning for Neural Machine Translation.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Competence-based Curriculum Learning for Neural Machine Translation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.225659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.225659Z digest=sha256:cf8e54066eb8ce54ca64540ab8a1e28f3444f08688e6b7560cb590a8c27f7c1f

Observation 975416c1-4126-4765-a05e-f68e8aa722fc · outbound

This paper cites Proximal Policy Optimization Algorithms.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Proximal Policy Optimization Algorithms

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.343445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.343445Z digest=sha256:0f0f548c0f124b86e4aa1ae3a7cb7a661bf06241771f7e3064d80afb3821de23

Observation dfb51c60-7593-4a3c-a539-9c0f8e795978 · outbound

This paper cites OmniCaptioner: One Captioner to Rule Them All.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models OmniCaptioner: One Captioner to Rule Them All

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.482694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.482694Z digest=sha256:7a7ce2d9a61bfaa5132a5147f5f08b71589e5fa48c6f83ba6deb7538c5cb3559

Observation e0a63174-7431-42af-8c92-dd95f1880b22 · outbound

This paper cites Qwen2.5-VL Technical Report.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Qwen2.5-VL Technical Report

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.585019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.585019Z digest=sha256:828c9a3b9b0eaad2b7a1355f1c7824ad61f445c68b80a5bf7614afee6496edb3

Observation 699b119a-0813-4c66-8920-4c58e0225a34 · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:24.706564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:59:21.742628Z digest=sha256:94ca8c5468f0efc3fdb022fa66c08b143163a34f089a6cd1681fbbde6338be3d

Observation c93cfb6a-aad6-46df-8592-84352ab045ac · outbound

This paper cites VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.841015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.841015Z digest=sha256:53715fd956bdeda3f0bef1871da1ee431b07c7c358b856ecb3726ea51b97118c

Observation 40d1d414-ca7a-4e9a-8d5a-2ee63473299f · outbound

This paper cites Light-R1: Curriculum SFT, DPO and RL for Long COT from Scratch and Beyond.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Light-R1: Curriculum SFT, DPO and RL for Long COT from Scratch and Beyond

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.956041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.956041Z digest=sha256:5fec4044c268b50e2b1b97ca60735073b1fed2373db0df03db57b76e35f307de

Observation deb10754-20f1-4f67-b05d-d636b3530df1 · outbound

This paper cites Zhang, W.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Zhang, W

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:24.519630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:59:22.076981Z digest=sha256:a71a2f731488b27a20983ca57750c838cc95cfa81a2f4ba2b48b1d309b01708a

Observation 97be1d3a-03fc-41f3-a6e9-d4326a340a24 · outbound

This paper cites Jiang, Y.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Jiang, Y

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:24.272710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:59:22.208111Z digest=sha256:3d6f889ea918d2909f4084ad0eee1bac485e3c97e20cd666c08755be8a78fbc1

Observation 83d0c1a8-6092-48b4-858e-ac0f90b51c9e · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:22.296769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:22.296769Z digest=sha256:1b3e4610dc7ab2bd19fe315039987744bbe77cf33177fe50e0bdb34b2c8ea998

Observation 8f977acf-2f75-4a43-af1b-84457d8b858c · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:24.038461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:59:22.365820Z digest=sha256:105e3376163ddab853f8d72880efbd95346df84bc4ca8880a617b1ac5e44df95

Observation 795d9b8f-73d6-49f7-b812-0e8a96d55842 · outbound

This paper cites Bansal, T.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Bansal, T

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:23.844769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:59:22.459073Z digest=sha256:55d305217f8c31bc111984f460cb09b761cd1afce64680f03c750c72af8e2f86

Observation 1ca04bda-1f1b-44e7-b44d-3c3c52e73e67 · outbound

This paper cites GPT-4o System Card.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models GPT-4o System Card

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:22.596946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:22.596946Z digest=sha256:0d4980bc7d9eb46f548b7ccd4bfae61cc60e7d6143f34ebafe3cfc14b204e2a4

Observation 7da94750-5079-4bce-b70c-4ae33764ce60 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:22.692295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:22.692295Z digest=sha256:216c44923fb55fd642a3aecfa06318547ae5f54ae7e0b8ba345061f7bc727f2a

Observation 3e0959b3-67cf-4dbc-b435-144b138447c1 · outbound

This paper cites Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:22.779575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:22.779575Z digest=sha256:e7a7a8bad7e66fcc52831e1502524bfe94a93874a6be256a4e5e76cd99f3c7de

Observation 064d151c-f6d6-48f8-abe0-6684ee452d39 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:22.902200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:22.902200Z digest=sha256:8174a1e93d7b6694c97597616938289e3d8e575dca244880bfaa176e3256894f

Observation 93489210-a930-4f3e-b293-6f091b59bba1 · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:23.018390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:23.018390Z digest=sha256:f3f866624b9640efba2412031d980522157a712709d2c7c0748b1fc616555b35

Observation 8f16e8fb-f784-4672-a0c3-8e0d2dff52f7 · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:23.141014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:23.141014Z digest=sha256:32426e6e1c218fa5e39cb4f36e82cf5165d7b3955495aab341ebaf9b907df5fd

Pith citing papers

Observation 12bd2d1f-7a14-4920-a638-28601499118e · inbound

Latent Visual Reasoning cites this paper.

Latent Visual Reasoning Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T18:41:30.420970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T18:41:30.307521Z digest=sha256:87231bd6c818f468b383f8c103cb07699ae0f9594e715267b0d8797fa2575859

Observation 600bdcfc-05fa-415f-b1e7-da58c3897ad4 · inbound

D$^2$Evo: Dual Difficulty-Aware Self-Evolution for Data-Efficient Reinforcement Learning cites this paper.

D$^2$Evo: Dual Difficulty-Aware Self-Evolution for Data-Efficient Reinforcement Learning Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:22:45.278394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-19T20:21:48.657926Z digest=sha256:2b9ad93fd1240b435137cbddf3092e3ae34e452beb0d56766ebc0c4cb33e8d1c

Observation 3bfa07d7-4e8b-4106-8fa1-53a296c7ad16 · inbound

VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct cites this paper.

VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:39:46.065840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T08:34:27.719022Z digest=sha256:a28dcded8a320b9f43b5eea5eb05e018555ce80fd006577683d0113af56de44a

Observation 8c55e748-8cfe-4f10-84db-cdea0c04c937 · inbound

LASER: A Corrective Lens for LVLMs via Visual Attention Preservation and Sink Suppression cites this paper.

LASER: A Corrective Lens for LVLMs via Visual Attention Preservation and Sink Suppression Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-03T16:58:42.492709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-03T16:53:43.427625Z digest=sha256:6f57b91ba515db8d45cb7c11075d9e9e58daed282a087abed921e10f53b0b745