Pith. sign in

Paper Citation Record · LEDGER

Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2504.16656.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.16656 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:01:37.530069Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T20:10:08.032034Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 43041165-440b-400b-96f9-d1d2cafa7e9f · inbound

R1-ShareVL: Incentivizing Reasoning Capability of Multimodal Large Language Models via Share-GRPO cites this paper.

R1-ShareVL: Incentivizing Reasoning Capability of Multimodal Large Language Models via Share-GRPO Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T15:01:37.530069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:01:37.530069Z digest=sha256:3ff5246f8d32988bb853c4dc134625bc77599180cd4d9e7f4b43b3f273f96dd0

Observation 00eac3aa-c4fa-40f9-ac6b-2a9297bbe228 · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:15.321786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:15.321786Z digest=sha256:fc0d653f677ccab50937a678f149274217c3e15cc3b85ba5dbe1e25fd8ebaa62

Observation c52f5b08-33f5-4921-a7eb-fdd7fb77eefd · inbound

Seeing is Believing, but How Much? A Comprehensive Analysis of Verbalized Calibration in Vision-Language Models cites this paper.

Seeing is Believing, but How Much? A Comprehensive Analysis of Verbalized Calibration in Vision-Language Models Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:01:22.483974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:01:22.483974Z digest=sha256:3ebb496e242fffd521aa541f10b96a7ff9281528e3a33ea5b21e2aeb80d6a9c1

Observation 7a5d3b8e-0064-4738-9167-8b751905a84a · inbound

MedBookVQA: A Systematic and Comprehensive Medical Benchmark Derived from Open-Access Book cites this paper.

MedBookVQA: A Systematic and Comprehensive Medical Benchmark Derived from Open-Access Book Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:03.055536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:03.055536Z digest=sha256:7fb213000b1e0c76a9cc3e51c7138a7bdd886d50e3112d29d1c3f741b320acfe

Observation b797d668-c259-444f-899a-0e6aa87ddfc8 · inbound

Rex-Thinker: Grounded Object Referring via Chain-of-Thought Reasoning cites this paper.

Rex-Thinker: Grounded Object Referring via Chain-of-Thought Reasoning Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T10:55:14.826127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:55:14.826127Z digest=sha256:10fc4a7636a940d525900e18bac517bb043f637fc64e0c78d0dc11ee94bb91d0

Observation c9527092-abfe-40c2-90fc-066cb510a9a0 · inbound

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning cites this paper.

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:04.960499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:04.960499Z digest=sha256:2fbfa48032d1b82fcbb4b5961cb6bea6f73d3d1ff64d3731ed1ffe60c06bc4a3

Observation 822e85af-b74d-4cbe-8011-983260eafac0 · inbound

FinLMM-R1: Enhancing Financial Reasoning in LMM through Scalable Data and Reward Design cites this paper.

FinLMM-R1: Enhancing Financial Reasoning in LMM through Scalable Data and Reward Design Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:40:40.201848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:40:40.201848Z digest=sha256:3de671bf4f4905aa1d06d2994e9fd4bc1c87c8a0e5ce83e869302c6e0404a88f

Observation 37c69bd4-167e-4749-ae2e-08166d75dcbc · inbound

Skywork-R1V3 Technical Report cites this paper.

Skywork-R1V3 Technical Report Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T19:14:05.468540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:14:05.468540Z digest=sha256:1834388e7032c3ed27472f349a79d17d9d59d2408e1d7706007161082f692ab6

Observation f68c5350-190b-4d9d-8bb9-663afbc9c051 · inbound

Perception-Aware Policy Optimization for Multimodal Reasoning cites this paper.

Perception-Aware Policy Optimization for Multimodal Reasoning Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-19T05:12:04.894974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-19T05:11:54.685897Z digest=sha256:32354fa5d46d30acfd2121249bc0f001e64afc6f33481b297d843649ff7dc872

Observation 1c3f7eaa-e072-4070-a191-14f807f18d2d · inbound

The Synergy Dilemma of Long-CoT SFT and RL: Investigating Post-Training Techniques for Reasoning VLMs cites this paper.

The Synergy Dilemma of Long-CoT SFT and RL: Investigating Post-Training Techniques for Reasoning VLMs Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T18:43:13.461715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:43:13.461715Z digest=sha256:48e9b2daaf8b99f41ea9fc983558d3aa44d2e69f7af2c578c3704957901bebd4

Observation 7557b985-2c27-4fa3-a62c-bb6b5b8b145b · inbound

VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning cites this paper.

VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T11:35:16.929602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:35:16.929602Z digest=sha256:d1a7c3a7194455e3df93cea0ac73fb3ac97478a9786037fbb4e323a69cb24a5d

Observation 4c08be35-c821-4880-8015-d56a3704094c · inbound

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency cites this paper.

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 138

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:58:58.831501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T11:58:58.660564Z digest=sha256:31599e974915b56792b99631c405da30247ed9fd57e3b2552d9bffee6588905f

Observation b2df9d6a-be8e-4b2f-ba4f-e856480b8e1c · inbound

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey cites this paper.

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-18T19:21:48.436079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T19:19:36.427337Z digest=sha256:3bad871f9e2a0da7c12c47d52dffd5cceee41b85b6356dde137933d25e78dd3f

Observation d23576fd-95c2-4705-993c-d1c6faa8a2bc · inbound

Boosting Reasoning in Large Multimodal Models via Activation Replay cites this paper.

Boosting Reasoning in Large Multimodal Models via Activation Replay Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:09:03.913660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T05:05:48.682057Z digest=sha256:6ec14c1c1c493bb812ffe0c3c0ce9c59cf1d4a39f4fe9668de3d13b271388239

Observation b7ae3f4e-507d-498b-8fa1-8a22ed821df6 · inbound

MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings cites this paper.

MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-10T03:29:21.666566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T03:27:50.144706Z digest=sha256:e9d5a7999c11b96f9521b2628937d62d3bb394d96539e1152b04d0b806414b3c

Observation 58fdb66b-61b1-44cb-b81d-fe39515eeb57 · inbound

Structured Role-Aware Policy Optimization for Multimodal Reasoning cites this paper.

Structured Role-Aware Policy Optimization for Multimodal Reasoning Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:25:58.316033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-11T01:24:20.286120Z digest=sha256:2aca60960b3dee208b1480e233866fc69403b4664afa0fb24b40aa6f11011777

Observation d3a16a42-9743-40dd-bb18-c9f6530313e1 · inbound

VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct cites this paper.

VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:39:46.028994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T08:34:27.719022Z digest=sha256:1ebe2792eaa21dfabc787012113fb0074253b39debc455ecf510c68a5b384a6f

Observation 19cae604-2808-4b8b-bb41-6ef22e2c502e · inbound

OracleAnalyser: Analysing Implicit Semantics of Oracle Bone Scripts through MLLMs with Post-training cites this paper.

OracleAnalyser: Analysing Implicit Semantics of Oracle Bone Scripts through MLLMs with Post-training Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-04T20:10:08.034694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-25T20:36:16.927195Z digest=sha256:6b8e533c1e4061904ed473c43d686e7de4c99d77c16fe0e78992c98495090ebe