Pith. sign in

Paper Citation Record · LEDGER

VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 55 inbound Pith citation observations for arXiv:2503.10291.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.10291 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 55 of 55 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:34:55.820224Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0357deb5-f945-4f28-954b-f79a9557f837 · inbound

T2I-FactualBench: Benchmarking the Factuality of Text-to-Image Models with Knowledge-Intensive Concepts cites this paper.

T2I-FactualBench: Benchmarking the Factuality of Text-to-Image Models with Knowledge-Intensive Concepts VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:02:43.301816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-23T08:00:12.781392Z digest=sha256:47da59c892a63f8a9d830ccfc565c2889d2f39560c48823caed057d83748c6a1

Observation 0c8b6bab-40fd-4ef5-a69b-4f51776f7c37 · inbound

From System 1 to System 2: A Survey of Reasoning Large Language Models cites this paper.

From System 1 to System 2: A Survey of Reasoning Large Language Models VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 177

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:36:24.230946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T01:36:23.845366Z digest=sha256:ae5451fcb9bf025112a26b7b4712d5810d0f0380170f2ed936d3f5273b72197f

Observation 93397a2e-c44b-4da0-b9ff-85e0fa4b7d70 · inbound

OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles cites this paper.

OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-05-19T06:59:03.295558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T06:59:03.112252Z digest=sha256:f63626a8c576618ba8daebd1d6ce54a8b5d2eab74b44a38966535d8e0ccbd8ea

Observation 4416468e-560f-42f8-b27d-2bfabe7a4572 · inbound

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models cites this paper.

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 126

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:41:08.123645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T13:41:07.991012Z digest=sha256:959724749aeb361145ef23b71131f486a6183576bc892ea839d9d8dd1b197f3a

Observation 8231b32a-9cff-4b97-b07b-433a37fa33fb · inbound

Reinforcement Learning from Human Feedback cites this paper.

Reinforcement Learning from Human Feedback VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 99

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:32:01.103457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T19:27:40.991325Z digest=sha256:38ccd0c3e5eb1b959df086c433edf8293c8f69ed207f58fb61116e63632fb1ea

Observation e1b25499-ef54-4575-90e3-e45f4e117118 · inbound

UniGen: Enhanced Training & Test-Time Strategies for Unified Multimodal Understanding and Generation cites this paper.

UniGen: Enhanced Training & Test-Time Strategies for Unified Multimodal Understanding and Generation VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T15:34:55.820224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:34:55.820224Z digest=sha256:8e4e33b855c70f44c04577b61ceac285b01c181daa2b9e6966db37077cdf4b75

Observation 96c7fdab-7bb1-4dbd-b817-ff0c0b83b648 · inbound

OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning cites this paper.

OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:20.837510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:57:20.837510Z digest=sha256:d51b63cfccd0696a1f3846e11c48358486fef6732bc1b75f703edbbc643f521f

Observation 91ff927c-8585-4733-abf8-33827210c667 · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:13.239675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:13.239675Z digest=sha256:10aa0b67e3ff88097f7d7e8e189f4e34ff89a1e7f922f127d40538ed4eb041b0

Observation 1c994559-1c4f-4a4e-a084-4f9d5dee187f · inbound

Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning cites this paper.

Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:42.507040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:24:42.507040Z digest=sha256:054444ce7bff6f42d4ea72be63bb1c820300e9eac4c920822a8f733667611bda

Observation 1aa3ba3e-bc83-4e45-923a-36e8737cec09 · inbound

UI-Genie: A Self-Improving Approach for Iteratively Boosting MLLM-based Mobile GUI Agents cites this paper.

UI-Genie: A Self-Improving Approach for Iteratively Boosting MLLM-based Mobile GUI Agents VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:34:15.560978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:34:15.560978Z digest=sha256:f426be7baefaaf65651617a8c4e414afc02c60e5d8268b404f7819d7dc91b6ed

Observation 976e89c3-815f-41cc-8fb3-508f1184dac0 · inbound

More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models cites this paper.

More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:35.324780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:35.324780Z digest=sha256:4611a53e087db9ad78a437a0d1c295ba6c0a5cc70ee4cadf8ac5d985d2f0a2f6

Observation 98bcf2ab-b1af-405d-b0dd-0ffed01820e9 · inbound

PixelThink: Towards Efficient Chain-of-Pixel Reasoning cites this paper.

PixelThink: Towards Efficient Chain-of-Pixel Reasoning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:45.679803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:45.679803Z digest=sha256:2e13f5001318eea36d6bed8c33166fc3ab7434471d65042517b6d1dc843fbf3b

Observation a17f981c-2931-4cd4-a188-00780aeb2972 · inbound

GThinker: Towards General Multimodal Reasoning via Cue-Guided Rethinking cites this paper.

GThinker: Towards General Multimodal Reasoning via Cue-Guided Rethinking VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T11:55:13.976791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:55:13.976791Z digest=sha256:eac2bdc97a3d8826ab5c382f07a8f5b79644922abb8588a360de317fec69a32a

Observation 35195a0a-124e-4c7d-a219-b93373d467c6 · inbound

K12Vista: Exploring the Boundaries of MLLMs in K-12 Education cites this paper.

K12Vista: Exploring the Boundaries of MLLMs in K-12 Education VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:42:31.819593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:42:31.819593Z digest=sha256:5871048704cd47ba9975a2fbd819445574ade3004b9891f394a1f3b69d96493a

Observation ede0b217-4f58-4d77-bde2-4636fe296336 · inbound

RewardBench 2: Advancing Reward Model Evaluation cites this paper.

RewardBench 2: Advancing Reward Model Evaluation VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:22:16.709903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T11:18:03.965711Z digest=sha256:2c82c090ad36ce21c0e073aea2be6eefa44a8317d55f4a32a57c658283ea1f1c

Observation 6290e526-21b7-4b54-ad8e-4f63cb734c17 · inbound

Evaluating MLLMs with Multimodal Multi-image Reasoning Benchmark cites this paper.

Evaluating MLLMs with Multimodal Multi-image Reasoning Benchmark VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:05:19.613480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:05:19.613480Z digest=sha256:03e149ffbd55ee1b480d0b20a91feac383d5b1372b53d7febcabe318c2672f11

Observation b9926470-386d-4ba4-9230-85aa4286e314 · inbound

MMRefine: Unveiling the Obstacles to Robust Refinement in Multimodal Large Language Models cites this paper.

MMRefine: Unveiling the Obstacles to Robust Refinement in Multimodal Large Language Models VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T10:39:59.406354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:39:59.406354Z digest=sha256:f96861d990aed9146e620ff06d7ca28909e4e333dbc61c8dfcb8e6ba236871ba

Observation b65f2783-cb81-4d0b-81c6-d4f0b5b9c89e · inbound

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning cites this paper.

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:04.949338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:04.949338Z digest=sha256:fa1f6b9ec3898dd53342700d16aa69f932828e9e53a87f87f4d3ecdee7709f5f

Observation e58ef32d-61b3-4a67-bba3-05a0c7d124eb · inbound

FinLMM-R1: Enhancing Financial Reasoning in LMM through Scalable Data and Reward Design cites this paper.

FinLMM-R1: Enhancing Financial Reasoning in LMM through Scalable Data and Reward Design VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:40:40.281992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:40:40.281992Z digest=sha256:433cae0dc1fa52be89ae18b63df79ca333262b729a8a89de635c16dac3b6f327

Observation e6bbca02-a527-4cb8-a7e2-05fd2c2b8596 · inbound

BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset cites this paper.

BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T20:15:52.646946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:15:52.646946Z digest=sha256:78663d98daa1ff755afa6e1cb111a079ca4a88e24df6cf2db76d40600a6e366e

Observation 397007ca-f815-49d8-bfee-c49396f092d9 · inbound

EduFlow: Advancing MLLMs' Problem-Solving Proficiency through Multi-Stage, Multi-Perspective Critique cites this paper.

EduFlow: Advancing MLLMs' Problem-Solving Proficiency through Multi-Stage, Multi-Perspective Critique VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:04:54.825440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:04:54.825440Z digest=sha256:f03868c313e3d15b7eeab3b2f90a9a0ce83e6acd6c501f56461c1e0f45bfbdcf

Observation af3cbdf7-50fb-4977-a426-7995841d4cec · inbound

VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning cites this paper.

VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T11:35:16.939733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:35:16.939733Z digest=sha256:a698e2da572a11f6d1ab11e87155b4f3b3fc5e4dbee1a1906611849cc2351f21

Observation c67da1dd-31f8-4225-9259-3c6597d91784 · inbound

VRPRM: Process Reward Modeling via Visual Reasoning cites this paper.

VRPRM: Process Reward Modeling via Visual Reasoning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T12:21:30.970706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T12:20:17.430881Z digest=sha256:2db41079d81ebbd845c4c7ca0075908d357e4ac2acee399e9ae96f4c525560bd

Observation c4e37aac-a1b6-47b5-8351-96dc2335b822 · inbound

VRPRM: Process Reward Modeling via Visual Reasoning cites this paper.

VRPRM: Process Reward Modeling via Visual Reasoning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T04:27:05.367965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:27:05.367965Z digest=sha256:611973bdc9734ecbbcd93d2959fe64560a82d6c9e2265ceacb37f362dbf31457

Observation c733a776-8f7c-4f95-a8c5-460a68bb050a · inbound

GM-PRM: A Generative Multimodal Process Reward Model for Multimodal Mathematical Reasoning cites this paper.

GM-PRM: A Generative Multimodal Process Reward Model for Multimodal Mathematical Reasoning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T00:59:53.654101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:59:53.654101Z digest=sha256:489afa19a1e8c145f9e31b93cdf94012e3db17731e44ca4eb1d0a6ebf001087c

Observation 09a9147e-039b-4d1b-9b85-d77cc603ab32 · inbound

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency cites this paper.

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 144

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:58:58.855293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T11:58:58.660564Z digest=sha256:7b9fb36ac148e71e57c8d8b9258c82ff5bf9cfa1f001aa60a297fe5986a7ca62

Observation 5e30e8a2-f9a7-480f-ac96-142f40427ac7 · inbound

LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model cites this paper.

LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T13:24:39.845687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:24:39.845687Z digest=sha256:bfc2f6c01e98fb3b206e44ade3028044d8f4eac36a6695f17af597fce8ae2ca5

Observation 9563bb18-9800-47b5-8b4d-38580fbf2181 · inbound

Reinforced Visual Perception with Tools cites this paper.

Reinforced Visual Perception with Tools VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T12:27:05.053667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:27:05.053667Z digest=sha256:6a0f59d858201f60194ce2e842d3ca0d3044db5422b5a7c84937d0dc4447ef31

Observation d8ac7af6-dc72-4af8-b27a-13e3231807a1 · inbound

Physical Plausibility Reasoning via HCM-GRPO: Empowering Compact Model for Superior Performance cites this paper.

Physical Plausibility Reasoning via HCM-GRPO: Empowering Compact Model for Superior Performance VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T22:36:10.373032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:36:10.373032Z digest=sha256:a6f0906bc042a44d59c9d81500be440b7d091fe1677d777a19915308696b29f5

Observation fe9e57f4-5e9e-4184-984c-23255ecbc06e · inbound

EvoLMM: Self-Evolving Large Multimodal Models with Continuous Rewards cites this paper.

EvoLMM: Self-Evolving Large Multimodal Models with Continuous Rewards VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T21:09:25.222181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:09:25.222181Z digest=sha256:f26d0b3d498a66e9874724bc9906aefc0b53a65688aa27b19bb6da038624974e

Observation b47ab510-e6f2-4a28-9631-23b49267ba75 · inbound

PaLMR: Towards Faithful Visual Reasoning via Multimodal Process Alignment cites this paper.

PaLMR: Towards Faithful Visual Reasoning via Multimodal Process Alignment VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T19:57:33.167864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:57:33.167864Z digest=sha256:f9b984a6bcdedf387efe098da7738ace1413f80016f1b4719f52342eb6ae0ced

Observation ab5f3789-1d32-4599-bbbe-d830fb9ca578 · inbound

CoVR-R:Reason-Aware Composed Video Retrieval cites this paper.

CoVR-R:Reason-Aware Composed Video Retrieval VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-13T21:37:55.887477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T21:37:55.887477Z digest=sha256:0fcb1b37ff0640e4919d718057d8c6b4ac824d8ed3a5f9b1bdcc1024ba1dccbf

Observation cff69bee-e71b-48b7-82bc-7adcb4899906 · inbound

Process Reward Models Meet Planning: Generating Precise and Scalable Datasets for Step-Level Rewards cites this paper.

Process Reward Models Meet Planning: Generating Precise and Scalable Datasets for Step-Level Rewards VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-10T04:45:21.123931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T04:40:52.854907Z digest=sha256:2ea31470704d575f289c43ccf1c343d88bdb7e2c03c76650a43ce46cd1eae70d

Observation d2b4ed60-afb0-46a3-a858-67eabcabdc6e · inbound

DT2IT-MRM: Debiased Preference Construction and Iterative Training for Multimodal Reward Modeling cites this paper.

DT2IT-MRM: Debiased Preference Construction and Iterative Training for Multimodal Reward Modeling VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:26:04.091377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T01:45:30.001398Z digest=sha256:ee3b6fea4210aa3c4ba3b9220632a979ac6fdb1cc21635a31a55b10499ebcf2d

Observation c97eab3c-2eb9-442d-ad84-ff0870ecc7c8 · inbound

V-tableR1: Process-Supervised Multimodal Table Reasoning with Critic-Guided Policy Optimization cites this paper.

V-tableR1: Process-Supervised Multimodal Table Reasoning with Critic-Guided Policy Optimization VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:01:04.215642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-09T23:48:32.613988Z digest=sha256:86f5775805b3c9838768c141440d42f7703190508c5389cdb30e7dacb5416cf8

Observation ba2121d0-d693-4a90-be7d-127fd5e02659 · inbound

Chart-FR1: Visual Focus-Driven Fine-Grained Reasoning on Dense Charts cites this paper.

Chart-FR1: Visual Focus-Driven Fine-Grained Reasoning on Dense Charts VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:16:02.907342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T16:08:15.393476Z digest=sha256:ba7cbe74f8038cbd1f4564293b54aa4bd265916fc776e68f942d944c8f04067f

Observation e51d29f8-5501-46de-a428-82aec7ed1639 · inbound

Verification Mirage: Mapping the Reliability Boundary of Self-Verification in Medical VQA cites this paper.

Verification Mirage: Mapping the Reliability Boundary of Self-Verification in Medical VQA VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:56:25.498891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T04:48:09.071812Z digest=sha256:c238fb25625486317053386f21aef4e7ba1d609304aebd1d32542f9e2803a5ab

Observation 45c67538-77e6-4ca5-b22c-fa3be027f5ca · inbound

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model cites this paper.

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:17:29.175391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T07:14:48.918959Z digest=sha256:654b326c4a3a7d2052e9646e5349b6068f5b7277bea73dc5fe3d801da8006b7f

Observation 4a07c751-3688-4944-89f9-33c75bf95245 · inbound

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model cites this paper.

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:48:00.564018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T21:47:50.595481Z digest=sha256:7fa64a4c61326f0599b82aed5ccef46b85f965b3118a15d363cc52525c281cec

Observation c8f15d7c-bc4d-4bc7-9d8e-958a90137a98 · inbound

PDCR: Perception-Decomposed Confidence Reward for Vision-Language Reasoning cites this paper.

PDCR: Perception-Decomposed Confidence Reward for Vision-Language Reasoning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:22:50.527662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T19:20:32.435135Z digest=sha256:812928efa5b56012eda07d35d87861320082701e161962b67ade8729f727028e

Observation 2a218807-cbe8-47ed-901d-fca5d1fefff5 · inbound

LLMs Know When They Know, but Do Not Act on It: A Metacognitive Harness for Test-time Scaling cites this paper.

LLMs Know When They Know, but Do Not Act on It: A Metacognitive Harness for Test-time Scaling VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:49:43.921154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T04:49:20.837352Z digest=sha256:b081e5e8bb25df5f7fb85b11a5dc2b4cc0ee249ce7eef27755f6c5b75bbbeac5

Observation 47c4b1de-10bd-4aa9-ab26-e8c771386952 · inbound

From Failure to Feedback: Group Revision Unlocks Hard Cases in Object-Level Grounding cites this paper.

From Failure to Feedback: Group Revision Unlocks Hard Cases in Object-Level Grounding VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-20T18:43:38.786474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T18:39:11.904941Z digest=sha256:85cb25422a8d3e8b60d22548f85a3ba038a85329f26f336a39c4e6db023946d8

Observation 8e3a314e-a995-4df0-a4dc-2b262804e1c8 · inbound

PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization cites this paper.

PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:43:12.481573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T10:41:25.205368Z digest=sha256:8dcb87cbb1ebcaabd2a3d09f35bed46e642a00c86bcc1a95d493684650e37cb1

Observation 99005f9f-9931-4576-98e8-0ad675cdd3f2 · inbound

PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization cites this paper.

PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T02:23:31.506677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:23:31.506677Z digest=sha256:e99e8c7037971a1f98dd59fb52c32231714dc43715cc8e3841f1a18897b7301a

Observation edb38eaa-85f3-4dfe-bd80-7f9e3c602e8b · inbound

CaptchaMind: Training CAPTCHA Solvers via Reinforcement Learning with Explicit Reasoning Supervision cites this paper.

CaptchaMind: Training CAPTCHA Solvers via Reinforcement Learning with Explicit Reasoning Supervision VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T06:18:05.601029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T06:14:43.288597Z digest=sha256:2f52886076139f2758ff28816f911dfc17ff51a6ec292bf64c45c7b8c0d8c4c1

Observation 79388d26-6c86-48c3-809a-70c48ed6a507 · inbound

PRO-CUA: Process-Reward Optimization for Computer Use Agents cites this paper.

PRO-CUA: Process-Reward Optimization for Computer Use Agents VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T11:53:24.213227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T11:46:38.569528Z digest=sha256:61345d6a83d291aedc15f4335ec38c8e8bbc79b533dfb6104b7415143154961c

Observation 4b3b86c1-8001-40c5-b9d7-3665d5975513 · inbound

StemBind: When MLLMs Get Lost Between Rules and Instances in Abstract Visual Reasoning cites this paper.

StemBind: When MLLMs Get Lost Between Rules and Instances in Abstract Visual Reasoning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-06-29T00:02:50.470385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T23:15:56.598968Z digest=sha256:4eb09a99189f8670d1c82f0de1edb267612a9c20081596d76f19ed9016d4d1a3

Observation e3088156-84dc-4d24-b31a-df4e8c9fe142 · inbound

ATLAS: Agentic Test-time Learning-to-Allocate Scaling cites this paper.

ATLAS: Agentic Test-time Learning-to-Allocate Scaling VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:26:17.018851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T15:27:28.290178Z digest=sha256:b18454e46949265304e22b43dbc7c9afeffb313c9d1400f642193cac8ce653a8

Observation c311efaa-8915-439f-9605-8e9a7ce9317b · inbound

SCI-PRM: A Tool Aware Process Reward Model for Scientific Reasoning Verification cites this paper.

SCI-PRM: A Tool Aware Process Reward Model for Scientific Reasoning Verification VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:46:46.511202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T06:38:37.969788Z digest=sha256:923a297a37dfd83fab5b181a670e1a4559885146473a43880b811639ca52f823

Observation ed2bd02a-a54d-46d2-bded-3debe6d62791 · inbound

Improving Multimodal Reasoning via Worst Dimension Optimization cites this paper.

Improving Multimodal Reasoning via Worst Dimension Optimization VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:37:14.806679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T21:55:57.691185Z digest=sha256:5f3b9bdacd23e53b4d2b53d5e9ff90f0f8fc6663a9243e55c3a93c33f1272cf0

Observation 8172c0e5-dec6-45cc-be10-b308f5d40ae3 · inbound

Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning cites this paper.

Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:37:25.318226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T19:36:57.231932Z digest=sha256:317a562c58ebe0391dff71a59bb1049716e3ce9c4be5b2d7a38bc050ce23b1ff

Observation 88dd4483-2fe0-4c59-aaed-a601380ca739 · inbound

Reasoning as Intersection: Consensus-Frame Alignment for Visual Focus in Video-MLLMs cites this paper.

Reasoning as Intersection: Consensus-Frame Alignment for Visual Focus in Video-MLLMs VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:58:57.733804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T01:02:37.632461Z digest=sha256:1e9342ca4a6fb504ca8a88bb9a8f885e33cb1d0f852fd3145f2f1187b75aefa7

Observation 58d63079-3feb-46bd-9be9-c55760f13cbe · inbound

Test-Time Scaling for Small VLMs on Multilingual Visual MCQ cites this paper.

Test-Time Scaling for Small VLMs on Multilingual Visual MCQ VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-13T03:00:51.318412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:00:51.318412Z digest=sha256:d47a28aa7ecfd939df0dd49e543e5c6528f0a7acc47adbef680291620761ae5b

Observation 27a4cdd8-8675-4686-940c-1dea5a4259e8 · inbound

Correcting What You Cannot See: Credit Assignment for Perception Distillation in Multimodal Reasoners cites this paper.

Correcting What You Cannot See: Credit Assignment for Perception Distillation in Multimodal Reasoners VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-31T10:52:57.434420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T10:52:57.434420Z digest=sha256:ff711ac1184c7ae127eb6ee56ea02971d040cb90c4e97be2508939ba0d0f55d7

Observation f9026cc3-10b4-4545-9445-6bb030723ba1 · inbound

Correcting What You Cannot See: Credit Assignment for Perception Distillation in Multimodal Reasoners cites this paper.

Correcting What You Cannot See: Credit Assignment for Perception Distillation in Multimodal Reasoners VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T01:24:27.043856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:24:27.043856Z digest=sha256:bfc69276106d17623daa788b5a58f50740d7e8ae6597cca1886ed6e55643bda5