Pith. sign in

Paper Citation Record · LEDGER

Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2505.07263.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.07263 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:31:22.244781Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T08:25:33.331848Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 29c752e8-4d25-4393-9110-51be7b734697 · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 148

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:22.244781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:22.244781Z digest=sha256:2a1338866a122bfe1eca085baec3d86a0b79932cf2adade29258300cbd18d48d

Observation daea2ec7-3631-4e7d-a0ea-a91d883a4ca9 · inbound

Skywork-R1V3 Technical Report cites this paper.

Skywork-R1V3 Technical Report Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T19:14:05.742695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:14:05.742695Z digest=sha256:a99c40d075a8a211ecf0f558423ff1f695b41793175d876f0b6fde010e05d8fe

Observation afe50853-a83f-4c34-9d08-5af78b86a8a9 · inbound

TAR: Temporal Anchor-Constrained Reasoning for Video Temporal Grounding cites this paper.

TAR: Temporal Anchor-Constrained Reasoning for Video Temporal Grounding Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T22:01:25.998214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:01:25.998214Z digest=sha256:a9b4e9526b31cea2f5109379d67066cd373c1a04281f7ae15980f7e5ca957ff6

Observation 061fc6a0-f652-4dc6-9d5b-c112bd9a4b5b · inbound

Visual Preference Optimization with Rubric Rewards cites this paper.

Visual Preference Optimization with Rubric Rewards Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:56:00.954463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T15:45:52.980881Z digest=sha256:ec946595021600aef17c2a9df8561f954b3dc1ab5eace315157cf440e4d5ce5d

Observation 55746557-662c-4734-8315-55f47fd05c61 · inbound

DT2IT-MRM: Debiased Preference Construction and Iterative Training for Multimodal Reward Modeling cites this paper.

DT2IT-MRM: Debiased Preference Construction and Iterative Training for Multimodal Reward Modeling Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:26:03.881700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T01:45:30.001398Z digest=sha256:c773ac571d2c9b5bd04548deeaf014b10739151e88d7bce73134de74dc275c9b

Observation 8da7362c-25cc-4882-92e1-9c33fa08d5be · inbound

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance cites this paper.

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:46:26.660162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T09:24:42.020782Z digest=sha256:372f1497b84ab4c5e45606e8e0a6038b9d14788f5623bdd990a5a8f881f8bef2

Observation 072a0e18-81d2-4d77-b783-e6f9fc6f7f6a · inbound

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance cites this paper.

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-21T09:29:57.206393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T09:25:19.439423Z digest=sha256:0efed443c14042f04777e7d70246706a524db659f810935418a281460ad6c720

Observation 68e5921d-2ec6-4578-8812-980a02904756 · inbound

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance cites this paper.

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-07-01T08:25:33.334411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-01T08:19:37.044714Z digest=sha256:29b141bb788ad6cfdb7a4cee203b21432902a2e418204f103afb25920482d5b5

Observation e45cb040-cb95-4b16-8068-23c9f4e0fb6a · inbound

Video Understanding Reward Modeling: A Robust Benchmark and Performant Reward Models cites this paper.

Video Understanding Reward Modeling: A Robust Benchmark and Performant Reward Models Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:45:58.370830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T02:18:20.880231Z digest=sha256:6586095a0d676c6f5ca9c684c19cceeeddc3329762f06bdcdd1931c38515fc6e