Pith. sign in

Paper Citation Record · LEDGER

Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 31 inbound Pith citation observations for arXiv:2503.07065.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.07065 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 31 of 31 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:12:18.162015Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T15:09:54.970421Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c68ae557-5d86-4d4b-a033-12e69baae016 · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 245

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.556644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:251c8108274320e7e695fc3bb5f90240edd33bcd7dccde3a53ab453a9cd7f739

Observation 2c866938-5ec6-4e64-b313-5723a5276481 · inbound

VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning cites this paper.

VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:56:07.673376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T20:56:07.247122Z digest=sha256:32feb19788db8d40e61c97b90264674932dccc657ffdcec391d711b5dfbaf71d

Observation 87cfbd50-de8b-49ec-b945-027c72db7606 · inbound

VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model cites this paper.

VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T01:13:57.417184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-13T01:13:57.368874Z digest=sha256:bed56bff716290e4eee93af82c807fa9887470ef917a3e033e7719075db7b108

Observation 447b65fd-a8ba-4c2a-8a9d-28d9a619c78c · inbound

Reinforced MLLM: A Survey on RL-Based Reasoning in Multimodal Large Language Models cites this paper.

Reinforced MLLM: A Survey on RL-Based Reasoning in Multimodal Large Language Models Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T05:12:18.162015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:12:18.162015Z digest=sha256:3c9dd834b1ce6f1800f65e3fae673d5b4a71a379f5c555882fa6330327c37009

Observation d0ff539c-5fd0-4e91-aaa5-997b077ec6c6 · inbound

EchoInk-R1: Exploring Audio-Visual Reasoning in Multimodal LLMs via Reinforcement Learning cites this paper.

EchoInk-R1: Exploring Audio-Visual Reasoning in Multimodal LLMs via Reinforcement Learning Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T23:27:27.670138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:27:27.670138Z digest=sha256:2b33c805cd73fe47da4d2b1ffe911a37fa251dc0f35e6a0c4e3a67ac045f1e6d

Observation 69ec2b47-087f-4ba7-b52a-c4cc0a292bf4 · inbound

Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models cites this paper.

Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-15T23:21:12.317779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:21:12.317779Z digest=sha256:5110d7fa6d60da2571cc7a44abd4e204c110dd8edd6d50e775d872f0d38320b4

Observation ef240fc9-f479-47ad-bb22-f648706214c8 · inbound

UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning cites this paper.

UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:50.648770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:50.648770Z digest=sha256:96d7b9f583ecf17c5dd2e2c5a6a975c5c2db9e141da6c901f80db87e61a1cd8b

Observation 49a03a3e-854f-4b06-978e-666e8a225aa5 · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 121

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:19.391570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:19.391570Z digest=sha256:176ad8017a90be9a2bdd0d7700df926f5c66883f6a7cf0f2d04a6a638a3529e1

Observation c15369b9-d485-4683-a15b-39a08f08b63a · inbound

Align and Surpass Human Camouflaged Perception: Visual Refocus Reinforcement Fine-Tuning cites this paper.

Align and Surpass Human Camouflaged Perception: Visual Refocus Reinforcement Fine-Tuning Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:52.196940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:52.196940Z digest=sha256:75f69122bbec3b271b15da4eead427cd06322a379a1b8c7853a7add7813ccdf5

Observation 5fd999b5-d794-44e5-8d07-b494ae53ca2f · inbound

ZeroGUI: Automating Online GUI Learning at Zero Human Cost cites this paper.

ZeroGUI: Automating Online GUI Learning at Zero Human Cost Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:41:21.033702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:41:21.033702Z digest=sha256:fc0533a54cac5b532d8f77859ffd8b1351cbb2db3abec53a5d432d3405195df5

Observation 6348801c-8c22-4cb9-8943-fd8dbb1ab974 · inbound

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning cites this paper.

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:05.136689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:05.136689Z digest=sha256:69ef2fbf43461536474d93cc1e4765cc3ad736b469cdfd32675c1edd1fbeb069

Observation 8a7a9780-6cb9-4c16-a4e5-72318cf4db33 · inbound

ChartReasoner: Code-Driven Modality Bridging for Long-Chain Reasoning in Chart Question Answering cites this paper.

ChartReasoner: Code-Driven Modality Bridging for Long-Chain Reasoning in Chart Question Answering Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T04:38:32.144672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:38:32.144672Z digest=sha256:4dc319c5d4fcd7aff9498802459bfbfe7f1b5ebd8ee19047422cc363ea2e3259

Observation 6acd7ba0-3709-48da-8847-e86b78f7681a · inbound

MM-R5: MultiModal Reasoning-Enhanced ReRanker via Reinforcement Learning for Document Retrieval cites this paper.

MM-R5: MultiModal Reasoning-Enhanced ReRanker via Reinforcement Learning for Document Retrieval Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:57:01.730564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:57:01.730564Z digest=sha256:4028b7a5514c7720d9806195fabb5b86c019470f3770b309df197328c6e02210

Observation 8a3c534e-f5b7-401f-8b45-8355bd13bd18 · inbound

CaughtCheating: Is Your MLLM a Good Cheating Detective? Exploring the Boundary of Visual Perception and Reasoning cites this paper.

CaughtCheating: Is Your MLLM a Good Cheating Detective? Exploring the Boundary of Visual Perception and Reasoning Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T18:39:18.633525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:39:18.633525Z digest=sha256:fca38adc99a266a3e5d33062938f24fd09ae335a52c4f9b39c6b72db55fbb75d

Observation 0c20257e-3e3b-4776-aef8-ab908f7d19ba · inbound

Improving the Reasoning of Multi-Image Grounding in MLLMs via Reinforcement Learning cites this paper.

Improving the Reasoning of Multi-Image Grounding in MLLMs via Reinforcement Learning Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T06:52:07.977161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T06:50:02.607136Z digest=sha256:159a95bf1a2e439e2143ec2bcdf0164f2617903d6d493e5f067fa206612a59ee

Observation 06c3a419-31f2-4e3d-8846-29766c49ee89 · inbound

SyncLoop: A Multimodal Dual-Loop Framework for Self-Improving Mathematical Reasoning cites this paper.

SyncLoop: A Multimodal Dual-Loop Framework for Self-Improving Mathematical Reasoning Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:12:01.785071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:12:01.785071Z digest=sha256:55bb89b782bd494a855d0c43b34e583d6999f771248579528a178b6c2073aa2d

Observation f3a197cc-5b96-4ff5-ae41-bd22d2977949 · inbound

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey cites this paper.

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 141

Resolution
unresolved
no resolver link, observed 2026-08-05T20:28:55.647921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:28:55.647921Z digest=sha256:e92f830eb0717f3e49b495648a8f6a83590f9e42b39d7c660da51042651adf6e

Observation 81da2cfc-2953-4d37-9707-0803e844cfca · inbound

Aligning Large Vision-Language Models by Deep Reinforcement Learning and Direct Preference Optimization cites this paper.

Aligning Large Vision-Language Models by Deep Reinforcement Learning and Direct Preference Optimization Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T23:11:33.680238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:11:33.680238Z digest=sha256:8a8a0d7938c36debb98a9c4470f260cbdab93bc0da3b0bc487501d42f8fe21a8

Observation 994cb6c9-e3dc-4725-8a93-657b246f0fd7 · inbound

Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle cites this paper.

Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T16:07:28.711477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:07:28.711477Z digest=sha256:aa09e0b6d78bf327c82a0c1a6f2586ec6e33c4607c0a4fcbafc33f217e58265d

Observation 7fbcfb2d-1706-4683-aed3-2eb00a406eee · inbound

Rethinking Reward Signals in Video GRPO: When Scores Become Targets cites this paper.

Rethinking Reward Signals in Video GRPO: When Scores Become Targets Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T20:36:05.524886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:36:05.524886Z digest=sha256:a24a231f2ca96b14dbdf69870842c23292848d1ceef466a09f399de81f397e9c

Observation be8d7fde-2bbf-49a6-83d9-053626d88997 · inbound

RL-RIG: A Generative Spatial Reasoner via Intrinsic Reflection cites this paper.

RL-RIG: A Generative Spatial Reasoner via Intrinsic Reflection Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:36:35.334320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T20:33:09.627731Z digest=sha256:344c7219a9b41b770bb115bacba1aafc148104998a398a0bb038f2514698674a

Observation b4429a10-cffb-411e-9217-1b9d4f2b86c7 · inbound

Curr-RLCER:Curriculum Reinforcement Learning For Coherence Explainable Recommendation cites this paper.

Curr-RLCER:Curriculum Reinforcement Learning For Coherence Explainable Recommendation Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:35:49.933305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T19:43:11.384787Z digest=sha256:5f466ab0f920383c7a9118878fb5d69d0291f154d5ffd1bc529db923320cb5ab

Observation 4c59afd4-f3ee-41e1-bbb5-a5d8f8f33cc3 · inbound

S-GRPO: Unified Post-Training for Large Vision-Language Models cites this paper.

S-GRPO: Unified Post-Training for Large Vision-Language Models Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:22:37.495818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T08:18:59.432486Z digest=sha256:2776d86828cf8acccbe43e9a7b917c845800b91a1cabf0dc517b0cd7c59e4594

Observation 80d34535-592e-4bde-815e-ce842d8991ec · inbound

S-GRPO: Unified Post-Training for Large Vision-Language Models cites this paper.

S-GRPO: Unified Post-Training for Large Vision-Language Models Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T16:10:54.025142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:10:54.025142Z digest=sha256:f048f51405bf893f95597609c441e3e802be7df6c38bc9ebd8659ff0e59748b4

Observation d1b1a740-f640-4f1f-9599-3afc69605ee2 · inbound

DARE: Difficulty-Adaptive Reinforcement Learning with Co-Evolved Difficulty Estimation cites this paper.

DARE: Difficulty-Adaptive Reinforcement Learning with Co-Evolved Difficulty Estimation Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T03:16:19.353516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T03:12:49.428954Z digest=sha256:425c929e622da411529520644161d8fc54a815b05c811a3c859553366057194f

Observation e1d6424c-6e84-42cb-abdf-bcddad31fd48 · inbound

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology cites this paper.

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:56:21.641871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T03:55:55.359488Z digest=sha256:4c185544c98e861e8b2377e803e3c8ae47bfe64c97c698ad8ccb3d3fd3ac6617

Observation 22522478-2499-4884-aed2-626270644382 · inbound

D$^2$Evo: Dual Difficulty-Aware Self-Evolution for Data-Efficient Reinforcement Learning cites this paper.

D$^2$Evo: Dual Difficulty-Aware Self-Evolution for Data-Efficient Reinforcement Learning Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:22:45.283893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-19T20:21:48.657926Z digest=sha256:f002a6201de377d6045887586233e9b3f89c1d277279a4c0e11b8214cae52648

Observation 7cb9933f-e0ff-4e72-9b76-03f0233c8ed6 · inbound

Towards One-to-Many Temporal Grounding cites this paper.

Towards One-to-Many Temporal Grounding Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:16:57.688274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-28T02:11:48.455492Z digest=sha256:825fa4feda1922b27f38867747bfeec26af10b39795daf706bbba9d007f6c119

Observation e7c2e331-f3c0-4d96-ab6b-e6e9a403e464 · inbound

Stage-1 Controls the Entropy Regime, Not the Outcome cites this paper.

Stage-1 Controls the Entropy Regime, Not the Outcome Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:17:29.710973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-27T17:16:39.814082Z digest=sha256:3343b9b1aeb7325a36717cecaa80a6588740ab3e9498fdf8627b9902b229b3c6

Observation efe7a114-88d6-46da-9f60-3ac4fa777060 · inbound

Dense Reward for Multi-View 3D Reasoning with Global Maps and Local Views cites this paper.

Dense Reward for Multi-View 3D Reasoning with Global Maps and Local Views Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:39:45.368196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T08:38:46.044079Z digest=sha256:46dd2cd689dcf3cc7a238255fd779cf7f32e1a8165b522e7574a9d56c2c76df3

Observation f5baae0b-b052-45ef-9647-37a64218d44a · inbound

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models cites this paper.

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 176

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:09:54.971976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T01:50:54.242508Z digest=sha256:dca4ad38ffee7515279b716aca35a067a2fd3ee060710b9830c8e47f4481feec