Pith. sign in

Paper Citation Record · LEDGER

Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2505.07263.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.07263 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:31:22.244781Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T08:25:33.331848Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 29c752e8-4d25-4393-9110-51be7b734697 · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 148

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:22.244781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:22.244781Z digest=sha256:2a1338866a122bfe1eca085baec3d86a0b79932cf2adade29258300cbd18d48d

Observation daea2ec7-3631-4e7d-a0ea-a91d883a4ca9 · inbound

Skywork-R1V3 Technical Report cites this paper.

Skywork-R1V3 Technical Report Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T19:14:05.742695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:14:05.742695Z digest=sha256:a99c40d075a8a211ecf0f558423ff1f695b41793175d876f0b6fde010e05d8fe

Observation afe50853-a83f-4c34-9d08-5af78b86a8a9 · inbound

TAR: Temporal Anchor-Constrained Reasoning for Video Temporal Grounding cites this paper.

TAR: Temporal Anchor-Constrained Reasoning for Video Temporal Grounding Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T22:01:25.998214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:01:25.998214Z digest=sha256:a9b4e9526b31cea2f5109379d67066cd373c1a04281f7ae15980f7e5ca957ff6

Observation 061fc6a0-f652-4dc6-9d5b-c112bd9a4b5b · inbound

Visual Preference Optimization with Rubric Rewards cites this paper.

Visual Preference Optimization with Rubric Rewards Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:56:00.954463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T15:45:52.980881Z digest=sha256:20399ecc7fe3b4d3665a9bd2ba1bdc5dcfd18e612681284b1f7c0ce2ccdf0b14

Observation 55746557-662c-4734-8315-55f47fd05c61 · inbound

DT2IT-MRM: Debiased Preference Construction and Iterative Training for Multimodal Reward Modeling cites this paper.

DT2IT-MRM: Debiased Preference Construction and Iterative Training for Multimodal Reward Modeling Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:26:03.881700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T01:45:30.001398Z digest=sha256:f75a7a3ada3286e54bb7b7b447b3082b49c8d59246d2672318bbc4c94f07ba0c

Observation 8da7362c-25cc-4882-92e1-9c33fa08d5be · inbound

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance cites this paper.

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:46:26.660162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T09:24:42.020782Z digest=sha256:9d76de3142f48a83f27fb81f39a93e17f92d75519e58c33d065aaef92281977f

Observation 072a0e18-81d2-4d77-b783-e6f9fc6f7f6a · inbound

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance cites this paper.

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-21T09:29:57.206393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T09:25:19.439423Z digest=sha256:faa4c4bf8acd5da066d4a16b62c38b4014c9df2e5f8bdbb68fe5a1bae9aa30dd

Observation 68e5921d-2ec6-4578-8812-980a02904756 · inbound

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance cites this paper.

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-07-01T08:25:33.334411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-01T08:19:37.044714Z digest=sha256:f4e5f27c8061ad0268adaca18aac5e726fb4e4b5c617cf6104680f3e8c144d05

Observation e45cb040-cb95-4b16-8068-23c9f4e0fb6a · inbound

Video Understanding Reward Modeling: A Robust Benchmark and Performant Reward Models cites this paper.

Video Understanding Reward Modeling: A Robust Benchmark and Performant Reward Models Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:45:58.370830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T02:18:20.880231Z digest=sha256:7e6bd501f99917c202ab5c72be47536384d7abca7b3af59865d8328b2492de13