Pith. sign in

Paper Citation Record · LEDGER

Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2504.16656.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.16656 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:01:37.530069Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T20:10:08.032034Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 43041165-440b-400b-96f9-d1d2cafa7e9f · inbound

R1-ShareVL: Incentivizing Reasoning Capability of Multimodal Large Language Models via Share-GRPO cites this paper.

R1-ShareVL: Incentivizing Reasoning Capability of Multimodal Large Language Models via Share-GRPO Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T15:01:37.530069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:01:37.530069Z digest=sha256:3ff5246f8d32988bb853c4dc134625bc77599180cd4d9e7f4b43b3f273f96dd0

Observation 00eac3aa-c4fa-40f9-ac6b-2a9297bbe228 · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:15.321786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:15.321786Z digest=sha256:fc0d653f677ccab50937a678f149274217c3e15cc3b85ba5dbe1e25fd8ebaa62

Observation c52f5b08-33f5-4921-a7eb-fdd7fb77eefd · inbound

Seeing is Believing, but How Much? A Comprehensive Analysis of Verbalized Calibration in Vision-Language Models cites this paper.

Seeing is Believing, but How Much? A Comprehensive Analysis of Verbalized Calibration in Vision-Language Models Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:01:22.483974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:01:22.483974Z digest=sha256:1871087a640b5172862be172213e7b56208d854878b634d0ed27b14f7602421a

Observation 7a5d3b8e-0064-4738-9167-8b751905a84a · inbound

MedBookVQA: A Systematic and Comprehensive Medical Benchmark Derived from Open-Access Book cites this paper.

MedBookVQA: A Systematic and Comprehensive Medical Benchmark Derived from Open-Access Book Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:03.055536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:03.055536Z digest=sha256:783c42170694ae2b7e2b98043d39d7e3d943468c5ab9059a3024f98886170006

Observation b797d668-c259-444f-899a-0e6aa87ddfc8 · inbound

Rex-Thinker: Grounded Object Referring via Chain-of-Thought Reasoning cites this paper.

Rex-Thinker: Grounded Object Referring via Chain-of-Thought Reasoning Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T10:55:14.826127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:55:14.826127Z digest=sha256:10fc4a7636a940d525900e18bac517bb043f637fc64e0c78d0dc11ee94bb91d0

Observation c9527092-abfe-40c2-90fc-066cb510a9a0 · inbound

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning cites this paper.

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:04.960499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:04.960499Z digest=sha256:2fbfa48032d1b82fcbb4b5961cb6bea6f73d3d1ff64d3731ed1ffe60c06bc4a3

Observation 822e85af-b74d-4cbe-8011-983260eafac0 · inbound

FinLMM-R1: Enhancing Financial Reasoning in LMM through Scalable Data and Reward Design cites this paper.

FinLMM-R1: Enhancing Financial Reasoning in LMM through Scalable Data and Reward Design Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:40:40.201848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:40:40.201848Z digest=sha256:3de671bf4f4905aa1d06d2994e9fd4bc1c87c8a0e5ce83e869302c6e0404a88f

Observation 37c69bd4-167e-4749-ae2e-08166d75dcbc · inbound

Skywork-R1V3 Technical Report cites this paper.

Skywork-R1V3 Technical Report Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T19:14:05.468540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:14:05.468540Z digest=sha256:1834388e7032c3ed27472f349a79d17d9d59d2408e1d7706007161082f692ab6

Observation f68c5350-190b-4d9d-8bb9-663afbc9c051 · inbound

Perception-Aware Policy Optimization for Multimodal Reasoning cites this paper.

Perception-Aware Policy Optimization for Multimodal Reasoning Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-19T05:12:04.894974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-19T05:11:54.685897Z digest=sha256:fb26d72e841fd7f42324cea07dc05b6d7fdbf1dc5ff5581eea7d000e98a8dde9

Observation 1c3f7eaa-e072-4070-a191-14f807f18d2d · inbound

The Synergy Dilemma of Long-CoT SFT and RL: Investigating Post-Training Techniques for Reasoning VLMs cites this paper.

The Synergy Dilemma of Long-CoT SFT and RL: Investigating Post-Training Techniques for Reasoning VLMs Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T18:43:13.461715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:43:13.461715Z digest=sha256:48e9b2daaf8b99f41ea9fc983558d3aa44d2e69f7af2c578c3704957901bebd4

Observation 7557b985-2c27-4fa3-a62c-bb6b5b8b145b · inbound

VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning cites this paper.

VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T11:35:16.929602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:35:16.929602Z digest=sha256:f770fe947aac0c54e7dead730862f88254f39098a78b32eedebcc276b3d6a3ec

Observation 4c08be35-c821-4880-8015-d56a3704094c · inbound

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency cites this paper.

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 138

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:58:58.831501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T11:58:58.660564Z digest=sha256:94dfc7fb1d4805a9adf939934db3342747f0b1bd617c2871092338915fd882e6

Observation b2df9d6a-be8e-4b2f-ba4f-e856480b8e1c · inbound

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey cites this paper.

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-18T19:21:48.436079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T19:19:36.427337Z digest=sha256:af4f498cbd59998772a7895bdbca5eec9d0408ed919f2018819851a129b24f7f

Observation d23576fd-95c2-4705-993c-d1c6faa8a2bc · inbound

Boosting Reasoning in Large Multimodal Models via Activation Replay cites this paper.

Boosting Reasoning in Large Multimodal Models via Activation Replay Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:09:03.913660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T05:05:48.682057Z digest=sha256:24ce78e300c45695c11f5e09b3d91a9b30b25bea3203e2ba025deb321806468d

Observation b7ae3f4e-507d-498b-8fa1-8a22ed821df6 · inbound

MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings cites this paper.

MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-10T03:29:21.666566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T03:27:50.144706Z digest=sha256:70ce4529aa710074b35bf015345272c3e023314ec4fc91549c062c7e86e1b826

Observation 58fdb66b-61b1-44cb-b81d-fe39515eeb57 · inbound

Structured Role-Aware Policy Optimization for Multimodal Reasoning cites this paper.

Structured Role-Aware Policy Optimization for Multimodal Reasoning Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:25:58.316033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-11T01:24:20.286120Z digest=sha256:d9715802e184630c0aadd7f3378ad1ba2c19816fdde5d09dec6086c7f5099a6f

Observation d3a16a42-9743-40dd-bb18-c9f6530313e1 · inbound

VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct cites this paper.

VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:39:46.028994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T08:34:27.719022Z digest=sha256:14ddba412a668abfcc91742134d4db93cbcedeb0d643ea6774ba4bf00c8fde5b

Observation 19cae604-2808-4b8b-bb41-6ef22e2c502e · inbound

OracleAnalyser: Analysing Implicit Semantics of Oracle Bone Scripts through MLLMs with Post-training cites this paper.

OracleAnalyser: Analysing Implicit Semantics of Oracle Bone Scripts through MLLMs with Post-training Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-04T20:10:08.034694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-25T20:36:16.927195Z digest=sha256:db8576ef3aab6382af0aa953380cf4f4924e0eded28f710e2a2f5747180034c3