Pith. sign in

Paper Citation Record · LEDGER

VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 56 inbound Pith citation observations for arXiv:2503.10291.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.10291 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 56 of 56 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T19:56:01.947345Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0357deb5-f945-4f28-954b-f79a9557f837 · inbound

T2I-FactualBench: Benchmarking the Factuality of Text-to-Image Models with Knowledge-Intensive Concepts cites this paper.

T2I-FactualBench: Benchmarking the Factuality of Text-to-Image Models with Knowledge-Intensive Concepts VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:02:43.301816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-23T08:00:12.781392Z digest=sha256:b13dc657fefd829ac912e5c5693e6d194bd64acf59582fed5fdc0d1c8e00aeb2

Observation 0c8b6bab-40fd-4ef5-a69b-4f51776f7c37 · inbound

From System 1 to System 2: A Survey of Reasoning Large Language Models cites this paper.

From System 1 to System 2: A Survey of Reasoning Large Language Models VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 177

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:36:24.230946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-13T01:36:23.845366Z digest=sha256:47fa778798e409f149aef5a1c822def18647d02d47bd5dbb837d49e0f348d636

Observation 93397a2e-c44b-4da0-b9ff-85e0fa4b7d70 · inbound

OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles cites this paper.

OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-05-19T06:59:03.295558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-19T06:59:03.112252Z digest=sha256:320f5d9e041495b467c9af95d1bdf05f8f5fe38b9b9f27118508e42a8263e072

Observation 4416468e-560f-42f8-b27d-2bfabe7a4572 · inbound

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models cites this paper.

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 126

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:41:08.123645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T13:41:07.991012Z digest=sha256:b71911fc11f52914bc0c65cec20dc38f4113105e7ebd93110900e35c7c66d73e

Observation 8231b32a-9cff-4b97-b07b-433a37fa33fb · inbound

Reinforcement Learning from Human Feedback cites this paper.

Reinforcement Learning from Human Feedback VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 99

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:32:01.103457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-22T19:27:40.991325Z digest=sha256:73a258f9bb5c47f79c817f9e9fdbab56f4a98701ab8bf50ea5786ba2c3445584

Observation e1b25499-ef54-4575-90e3-e45f4e117118 · inbound

UniGen: Enhanced Training & Test-Time Strategies for Unified Multimodal Understanding and Generation cites this paper.

UniGen: Enhanced Training & Test-Time Strategies for Unified Multimodal Understanding and Generation VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T15:34:55.820224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:34:55.820224Z digest=sha256:48aeef826e26d113768c9e07abb1b5a1a606f2dfd37b992819f917f0ba73b3cd

Observation 96c7fdab-7bb1-4dbd-b817-ff0c0b83b648 · inbound

OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning cites this paper.

OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:20.837510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:57:20.837510Z digest=sha256:6fb708fc393a1297da42c09c2af074c68bf610cdd4dfaac54e8d8471461a2521

Observation 91ff927c-8585-4733-abf8-33827210c667 · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:13.239675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:13.239675Z digest=sha256:68f1e98e33eac5458cb28bb7f42dd9cd5df65ec96980c31361d96b880649d008

Observation 1c994559-1c4f-4a4e-a084-4f9d5dee187f · inbound

Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning cites this paper.

Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:42.507040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:24:42.507040Z digest=sha256:e1ae597f38d70df7d25bc437deaf394702e05a1135daa6d2f97dc66cb59af650

Observation 1aa3ba3e-bc83-4e45-923a-36e8737cec09 · inbound

UI-Genie: A Self-Improving Approach for Iteratively Boosting MLLM-based Mobile GUI Agents cites this paper.

UI-Genie: A Self-Improving Approach for Iteratively Boosting MLLM-based Mobile GUI Agents VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:34:15.560978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:34:15.560978Z digest=sha256:4c68d9b3852c7dadda3b1dba0e38519d313da65fbb1b3ad5831abe986d4b97ac

Observation 976e89c3-815f-41cc-8fb3-508f1184dac0 · inbound

More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models cites this paper.

More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:35.324780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:35.324780Z digest=sha256:a61fd70c3e25c55253da51b4eeaa6bc157805b75a4e306346c4841d65854d143

Observation 98bcf2ab-b1af-405d-b0dd-0ffed01820e9 · inbound

PixelThink: Towards Efficient Chain-of-Pixel Reasoning cites this paper.

PixelThink: Towards Efficient Chain-of-Pixel Reasoning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:45.679803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:45.679803Z digest=sha256:83b89f39d54979874853868c3361b1f50e60ba9c7da69a8c0f91f94b3b3ae492

Observation a17f981c-2931-4cd4-a188-00780aeb2972 · inbound

GThinker: Towards General Multimodal Reasoning via Cue-Guided Rethinking cites this paper.

GThinker: Towards General Multimodal Reasoning via Cue-Guided Rethinking VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T11:55:13.976791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:55:13.976791Z digest=sha256:6921549a91e07028ddaeb73a0a890e20d24ea9d80cd5e92c3bbab3ca310313af

Observation 35195a0a-124e-4c7d-a219-b93373d467c6 · inbound

K12Vista: Exploring the Boundaries of MLLMs in K-12 Education cites this paper.

K12Vista: Exploring the Boundaries of MLLMs in K-12 Education VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:42:31.819593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:42:31.819593Z digest=sha256:dca4ca6ade986696b71f31df48c887a50d1960793d80b92635b6eaa4ea188a89

Observation ede0b217-4f58-4d77-bde2-4636fe296336 · inbound

RewardBench 2: Advancing Reward Model Evaluation cites this paper.

RewardBench 2: Advancing Reward Model Evaluation VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:22:16.709903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-19T11:18:03.965711Z digest=sha256:b43de82f779898671b035cbc7ef348113db718eb0055aa8cbc0f55ef316562b1

Observation 6290e526-21b7-4b54-ad8e-4f63cb734c17 · inbound

Evaluating MLLMs with Multimodal Multi-image Reasoning Benchmark cites this paper.

Evaluating MLLMs with Multimodal Multi-image Reasoning Benchmark VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:05:19.613480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:05:19.613480Z digest=sha256:7703f358889cc3a84fcc66195c92269a34ef47a88fd10d9602e4ad47a2c3862d

Observation b9926470-386d-4ba4-9230-85aa4286e314 · inbound

MMRefine: Unveiling the Obstacles to Robust Refinement in Multimodal Large Language Models cites this paper.

MMRefine: Unveiling the Obstacles to Robust Refinement in Multimodal Large Language Models VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T10:39:59.406354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:39:59.406354Z digest=sha256:4e634b6dd3e1e422e69b573c6eeaee0a1980165e21a32bf180f90586fe1be929

Observation b65f2783-cb81-4d0b-81c6-d4f0b5b9c89e · inbound

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning cites this paper.

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:04.949338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:04.949338Z digest=sha256:020639a33b2d72e4ba96394295769fe5159dcfbe3a4ecdf261f93a099b7c185f

Observation e58ef32d-61b3-4a67-bba3-05a0c7d124eb · inbound

FinLMM-R1: Enhancing Financial Reasoning in LMM through Scalable Data and Reward Design cites this paper.

FinLMM-R1: Enhancing Financial Reasoning in LMM through Scalable Data and Reward Design VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:40:40.281992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:40:40.281992Z digest=sha256:7cb4c00971810eb6ad6f869afb5821ba47be86b8000066e41a60eda7cbe35dd1

Observation e6bbca02-a527-4cb8-a7e2-05fd2c2b8596 · inbound

BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset cites this paper.

BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T20:15:52.646946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:15:52.646946Z digest=sha256:f6e8c5c013de9dc6170ef6a424e1006a53e595d201a919b55a9add7cfd91513f

Observation 397007ca-f815-49d8-bfee-c49396f092d9 · inbound

EduFlow: Advancing MLLMs' Problem-Solving Proficiency through Multi-Stage, Multi-Perspective Critique cites this paper.

EduFlow: Advancing MLLMs' Problem-Solving Proficiency through Multi-Stage, Multi-Perspective Critique VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:04:54.825440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:04:54.825440Z digest=sha256:a8125e44dae7f4cc56cf8dd33c39254d2d13dbe8228a7512013d233691d1d618

Observation af3cbdf7-50fb-4977-a426-7995841d4cec · inbound

VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning cites this paper.

VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T11:35:16.939733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:35:16.939733Z digest=sha256:e66510b86d919214b3d4d3f4a56e50935f4ce0e2cdf7ca3097d6820bd3fa1d65

Observation c67da1dd-31f8-4225-9259-3c6597d91784 · inbound

VRPRM: Process Reward Modeling via Visual Reasoning cites this paper.

VRPRM: Process Reward Modeling via Visual Reasoning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T12:21:30.970706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-22T12:20:17.430881Z digest=sha256:e7264fab00209802326903286caf35daa66f02618c48e3e4916206679ea562b4

Observation c4e37aac-a1b6-47b5-8351-96dc2335b822 · inbound

VRPRM: Process Reward Modeling via Visual Reasoning cites this paper.

VRPRM: Process Reward Modeling via Visual Reasoning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T04:27:05.367965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:27:05.367965Z digest=sha256:56e6bc18ce0c332eda04d190422e1016bbb133ca6460b587d7aeeb7531d235c1

Observation c733a776-8f7c-4f95-a8c5-460a68bb050a · inbound

GM-PRM: A Generative Multimodal Process Reward Model for Multimodal Mathematical Reasoning cites this paper.

GM-PRM: A Generative Multimodal Process Reward Model for Multimodal Mathematical Reasoning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T00:59:53.654101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:59:53.654101Z digest=sha256:451f6d96c22a369fb8e8445feefb4cf6249e42b4c10409ef5bd0c85be2681969

Observation 09a9147e-039b-4d1b-9b85-d77cc603ab32 · inbound

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency cites this paper.

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 144

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:58:58.855293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T11:58:58.660564Z digest=sha256:7e2ab17cfa5f1ecf3a1cebda3d13ce119613a5f90b1276f64c573e89561caf52

Observation 5e30e8a2-f9a7-480f-ac96-142f40427ac7 · inbound

LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model cites this paper.

LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T13:24:39.845687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:24:39.845687Z digest=sha256:080fba97a9f9f455d1c0012da9c3e27fbbd2b16430ca7c8b75a63921dfef4a1b

Observation 9563bb18-9800-47b5-8b4d-38580fbf2181 · inbound

Reinforced Visual Perception with Tools cites this paper.

Reinforced Visual Perception with Tools VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T12:27:05.053667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:27:05.053667Z digest=sha256:8630d824242a572efa5ad87d842bd7a0f238a39eface2d71e3a39b738f906444

Observation d8ac7af6-dc72-4af8-b27a-13e3231807a1 · inbound

Physical Plausibility Reasoning via HCM-GRPO: Empowering Compact Model for Superior Performance cites this paper.

Physical Plausibility Reasoning via HCM-GRPO: Empowering Compact Model for Superior Performance VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T22:36:10.373032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:36:10.373032Z digest=sha256:315e280383e91c09bdee159e47533d1d2c72c3672aa06779cb2fe562147f1262

Observation fe9e57f4-5e9e-4184-984c-23255ecbc06e · inbound

EvoLMM: Self-Evolving Large Multimodal Models with Continuous Rewards cites this paper.

EvoLMM: Self-Evolving Large Multimodal Models with Continuous Rewards VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T21:09:25.222181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:09:25.222181Z digest=sha256:766d960184de617518a0d15e340e7da76e95dc50533015bc94f4a2a56f7c55d6

Observation b47ab510-e6f2-4a28-9631-23b49267ba75 · inbound

PaLMR: Towards Faithful Visual Reasoning via Multimodal Process Alignment cites this paper.

PaLMR: Towards Faithful Visual Reasoning via Multimodal Process Alignment VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T19:57:33.167864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:57:33.167864Z digest=sha256:b155c7d5ce065d7572a3a49ff385b4df9c2e731f1563a07ad2c60f463061865e

Observation ab5f3789-1d32-4599-bbbe-d830fb9ca578 · inbound

CoVR-R:Reason-Aware Composed Video Retrieval cites this paper.

CoVR-R:Reason-Aware Composed Video Retrieval VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-13T21:37:55.887477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T21:37:55.887477Z digest=sha256:9fa97201b95d1fe92f40c3dc911c2841572629f9e9ef09cabf7f29a812a5e295

Observation cff69bee-e71b-48b7-82bc-7adcb4899906 · inbound

Process Reward Models Meet Planning: Generating Precise and Scalable Datasets for Step-Level Rewards cites this paper.

Process Reward Models Meet Planning: Generating Precise and Scalable Datasets for Step-Level Rewards VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-10T04:45:21.123931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-10T04:40:52.854907Z digest=sha256:392a7d9beb331591fd04f44f036ba74be9b8958f56df68c9b3ea10fcaaa4d7d3

Observation d2b4ed60-afb0-46a3-a858-67eabcabdc6e · inbound

DT2IT-MRM: Debiased Preference Construction and Iterative Training for Multimodal Reward Modeling cites this paper.

DT2IT-MRM: Debiased Preference Construction and Iterative Training for Multimodal Reward Modeling VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:26:04.091377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T01:45:30.001398Z digest=sha256:cfa4214a7a365ae352bb01d5c263e7166a5a6b8dd51c03ce8cee1e224bde9ec7

Observation c97eab3c-2eb9-442d-ad84-ff0870ecc7c8 · inbound

V-tableR1: Process-Supervised Multimodal Table Reasoning with Critic-Guided Policy Optimization cites this paper.

V-tableR1: Process-Supervised Multimodal Table Reasoning with Critic-Guided Policy Optimization VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:01:04.215642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-09T23:48:32.613988Z digest=sha256:2505f3ffd23876d29a5136151db8b62ea7e5cc8f2dbeca7f9d5246859cd1220e

Observation ba2121d0-d693-4a90-be7d-127fd5e02659 · inbound

Chart-FR1: Visual Focus-Driven Fine-Grained Reasoning on Dense Charts cites this paper.

Chart-FR1: Visual Focus-Driven Fine-Grained Reasoning on Dense Charts VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:16:02.907342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T16:08:15.393476Z digest=sha256:9f5c1642109838d53a12b7f62f6de00558e683f4ce42e5e72e480e550e96022d

Observation e51d29f8-5501-46de-a428-82aec7ed1639 · inbound

Verification Mirage: Mapping the Reliability Boundary of Self-Verification in Medical VQA cites this paper.

Verification Mirage: Mapping the Reliability Boundary of Self-Verification in Medical VQA VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:56:25.498891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-12T04:48:09.071812Z digest=sha256:cd299d97ecf539499c1d59271cd2ea9065ffdf872afabe41d110c36ade9d0aca

Observation 45c67538-77e6-4ca5-b22c-fa3be027f5ca · inbound

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model cites this paper.

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:17:29.175391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-13T07:14:48.918959Z digest=sha256:954cbab83e5e44c38277e7ce98369a63144946640fdd7487855045575ec4bf0e

Observation 4a07c751-3688-4944-89f9-33c75bf95245 · inbound

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model cites this paper.

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:48:00.564018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-14T21:47:50.595481Z digest=sha256:a238d6d61851b54bdf703882870f3d606063851f7dc3f1fcc10a0c98a9e139f3

Observation c8f15d7c-bc4d-4bc7-9d8e-958a90137a98 · inbound

PDCR: Perception-Decomposed Confidence Reward for Vision-Language Reasoning cites this paper.

PDCR: Perception-Decomposed Confidence Reward for Vision-Language Reasoning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:22:50.527662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-14T19:20:32.435135Z digest=sha256:95ee36d8100dbeb1ae43bc721d3c1600f29b9a8cb228ab377d10ade546dc973f

Observation 2a218807-cbe8-47ed-901d-fca5d1fefff5 · inbound

LLMs Know When They Know, but Do Not Act on It: A Metacognitive Harness for Test-time Scaling cites this paper.

LLMs Know When They Know, but Do Not Act on It: A Metacognitive Harness for Test-time Scaling VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:49:43.921154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-15T04:49:20.837352Z digest=sha256:3d0059f9dd9e9b226de6927bafde2eb72db1cff0f70e50d0162aed3205fcf749

Observation 47c4b1de-10bd-4aa9-ab26-e8c771386952 · inbound

From Failure to Feedback: Group Revision Unlocks Hard Cases in Object-Level Grounding cites this paper.

From Failure to Feedback: Group Revision Unlocks Hard Cases in Object-Level Grounding VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-20T18:43:38.786474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-20T18:39:11.904941Z digest=sha256:9e7cc52c669d9c59bed88a5e4061ccc5447c2c3570200f363a480f09e44e2ef6

Observation 8e3a314e-a995-4df0-a4dc-2b262804e1c8 · inbound

PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization cites this paper.

PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:43:12.481573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-20T10:41:25.205368Z digest=sha256:0907cf2463c4bd981eefe56d73aa2f648b01a3e71623f94da3fb489619e222bb

Observation 99005f9f-9931-4576-98e8-0ad675cdd3f2 · inbound

PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization cites this paper.

PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T02:23:31.506677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:23:31.506677Z digest=sha256:9282bd12bed441f9b4b791fde35bb8633af10f084782f7ef6bf04ed6302b6354

Observation edb38eaa-85f3-4dfe-bd80-7f9e3c602e8b · inbound

CaptchaMind: Training CAPTCHA Solvers via Reinforcement Learning with Explicit Reasoning Supervision cites this paper.

CaptchaMind: Training CAPTCHA Solvers via Reinforcement Learning with Explicit Reasoning Supervision VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T06:18:05.601029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T06:14:43.288597Z digest=sha256:0a9ce1852a4370a9151a40a88947c449ce1f44a61314bb29dde310ab83c574fe

Observation 79388d26-6c86-48c3-809a-70c48ed6a507 · inbound

PRO-CUA: Process-Reward Optimization for Computer Use Agents cites this paper.

PRO-CUA: Process-Reward Optimization for Computer Use Agents VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T11:53:24.213227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-29T11:46:38.569528Z digest=sha256:8a9232f355a1bb636329644ab908102c4a3de92a7b81da64d175b0871dd1bbdb

Observation 4b3b86c1-8001-40c5-b9d7-3665d5975513 · inbound

StemBind: When MLLMs Get Lost Between Rules and Instances in Abstract Visual Reasoning cites this paper.

StemBind: When MLLMs Get Lost Between Rules and Instances in Abstract Visual Reasoning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-06-29T00:02:50.470385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-28T23:15:56.598968Z digest=sha256:c82620bd793af9e3bdfa7c4a2c8413a8dcc938275e16af420e347168c38a115b

Observation e3088156-84dc-4d24-b31a-df4e8c9fe142 · inbound

ATLAS: Agentic Test-time Learning-to-Allocate Scaling cites this paper.

ATLAS: Agentic Test-time Learning-to-Allocate Scaling VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:26:17.018851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-28T15:27:28.290178Z digest=sha256:7a56ff266a36ae571623f58a5e2d0b1ed8097b987ea1ccb75fdd8e194799c0ff

Observation c311efaa-8915-439f-9605-8e9a7ce9317b · inbound

SCI-PRM: A Tool Aware Process Reward Model for Scientific Reasoning Verification cites this paper.

SCI-PRM: A Tool Aware Process Reward Model for Scientific Reasoning Verification VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:46:46.511202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-28T06:38:37.969788Z digest=sha256:b4f69d266a60da09443f1cd32321ff7b68d71559aeb457cb5f9634406daa3a4a

Observation ed2bd02a-a54d-46d2-bded-3debe6d62791 · inbound

Improving Multimodal Reasoning via Worst Dimension Optimization cites this paper.

Improving Multimodal Reasoning via Worst Dimension Optimization VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:37:14.806679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-27T21:55:57.691185Z digest=sha256:0b74a0380f0247befa8209ed542b8d84e4271779f4b4e100abadf3e152221170

Observation 8172c0e5-dec6-45cc-be10-b308f5d40ae3 · inbound

Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning cites this paper.

Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:37:25.318226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-27T19:36:57.231932Z digest=sha256:90be14bafca17bdc4c791bad34dfc986ef64459fac4a864ecb940e078560ab8b

Observation 88dd4483-2fe0-4c59-aaed-a601380ca739 · inbound

Reasoning as Intersection: Consensus-Frame Alignment for Visual Focus in Video-MLLMs cites this paper.

Reasoning as Intersection: Consensus-Frame Alignment for Visual Focus in Video-MLLMs VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:58:57.733804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-27T01:02:37.632461Z digest=sha256:dc89b57040d3f30b300fe3ca66b41bac713f953b8ae75869938235ee3920c1c0

Observation 58d63079-3feb-46bd-9be9-c55760f13cbe · inbound

Test-Time Scaling for Small VLMs on Multilingual Visual MCQ cites this paper.

Test-Time Scaling for Small VLMs on Multilingual Visual MCQ VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-13T03:00:51.318412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:00:51.318412Z digest=sha256:7dbe24e3a7b1fc0aa80d8a657f303e633bda7caab774cfb3ab1266b732d3d3a9

Observation 27a4cdd8-8675-4686-940c-1dea5a4259e8 · inbound

Correcting What You Cannot See: Credit Assignment for Perception Distillation in Multimodal Reasoners cites this paper.

Correcting What You Cannot See: Credit Assignment for Perception Distillation in Multimodal Reasoners VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-31T10:52:57.434420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T10:52:57.434420Z digest=sha256:0b762a1316e75bc21ae49536af3df77c1e53fb5b15cc0bf29e044e9a328c2061

Observation f9026cc3-10b4-4545-9445-6bb030723ba1 · inbound

Correcting What You Cannot See: Credit Assignment for Perception Distillation in Multimodal Reasoners cites this paper.

Correcting What You Cannot See: Credit Assignment for Perception Distillation in Multimodal Reasoners VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T01:24:27.043856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:24:27.043856Z digest=sha256:5bfc1d91e856315f2facacb57e2b660c6a52b47c5f23d91c24c0de748971b742

Observation 4c6225d7-f8a3-4908-81aa-9d65d80833e6 · inbound

VERDICT: Training-Free Step-Wise Verification of Multimodal Reasoning via Disagreement-Aware Consensus cites this paper.

VERDICT: Training-Free Step-Wise Verification of Multimodal Reasoning via Disagreement-Aware Consensus VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T19:56:01.947345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:56:01.947345Z digest=sha256:e9908eecfe529a6928fb25e70d5d66942f8fd0473e8dd2b8e785d7d4ba82a6f7