Pith. sign in

Paper Citation Record · LEDGER

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis

As of 8 August 2026, this Paper Citation Record lists 76 of 76 outbound references and 4 inbound Pith citation observations for arXiv:2506.02096.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02096 v1

Coverage vector

measured 76 of 76 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:35:48.656605Z

measured 80 of 80 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T11:47:27.883967Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T16:28:05.708078Z

Reference resolution

76 of 76 outbound references displayed

  • verified exact1
  • verified fuzzy11
  • unresolved64
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c6187cca-91c3-42de-8472-71c141a21060 · outbound

This paper cites @esa (Ref.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis @esa (Ref

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:41.936149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:41.936149Z digest=sha256:839d98de8954aca4b4bd8f5d4b2edabbf88126d0ede3097e8bec293104f6ff01

Observation 38def998-febb-43b9-8077-7ad487d07064 · outbound

This paper cites an unresolved cited work.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:41.977593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:41.977593Z digest=sha256:542c22f2bfc0237785a63135ea8c4ceac25c30e25424aee3273b9071cc55d0a0

Observation 06d238d5-bf6a-4608-a501-72ae996bdf8b · outbound

This paper cites an unresolved cited work.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:35:51.317857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:35:42.069904Z digest=sha256:e8a238c4a3a512791a2fda49c3ef0d5ecf3993e32e8cd1ff851fc79c3e30b06c

Observation 7833eefd-422b-4e3d-87df-b4bf2f8ceafd · outbound

This paper cites Phi-4-reasoning Technical Report.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Phi-4-reasoning Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:42.180779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:42.180779Z digest=sha256:96a241ba0715ce5b269a42bacec61b6e4e95eab48cded28e4bad0568d2b46ad5

Observation 7ec446e8-42d1-46ff-8d8c-e3df9114472a · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Flamingo: a visual language model for few-shot learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:42.252690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:42.252690Z digest=sha256:e408a87936ae50152f98c9fa9d7db6daba586d7b8f605bda1b6b6a6384e65868

Observation a53a029f-0c75-451b-9f4b-32788710af0b · outbound

This paper cites Claude 3.7 sonnet.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Claude 3.7 sonnet

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:51.299837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:35:42.377165Z digest=sha256:ec50d035f33a40aa95f28b3de1662744c4a77ac686971bf0e871abb42b57ccf4

Observation 24131374-3806-4923-a548-3d4b089c74d4 · outbound

This paper cites Qwen2.5-VL Technical Report.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Qwen2.5-VL Technical Report

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:42.453668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:42.453668Z digest=sha256:80625c2075e45b7571e09e0f84803b19d79eed2899b33a5efbced2c23347da2f

Observation 2ee0dfea-5b0c-4812-9145-a8dba7689ad3 · outbound

This paper cites A Survey of Multimodal Large Language Model from A Data-centric Perspective.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis A Survey of Multimodal Large Language Model from A Data-centric Perspective

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:42.563075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:42.563075Z digest=sha256:11fe56dbf18d75d72fb640dd67b8db0a8b8ca213ade593ce766ef424ecb34bd2

Observation c22a30bb-9997-482b-914d-9d91e59d0a67 · outbound

This paper cites Rank analysis of incomplete block designs: I.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Rank analysis of incomplete block designs: I

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:42.642750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:42.642750Z digest=sha256:43076f92034f96c53b430875da6727b2db54fb1fddeaf497b11d499262a12bab

Observation 4bc9233c-8ceb-46ba-8cda-9ae023448025 · outbound

This paper cites Sft or rl? an early investigation into training r1-like reasoning large vision-language models.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Sft or rl? an early investigation into training r1-like reasoning large vision-language models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:51.279167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:35:42.716759Z digest=sha256:2d26148a64bc4145db9611754fb81a75afa0b10762c85413822460d8e97473b6

Observation ae23a085-3fb9-4407-85a0-a934e545035e · outbound

This paper cites R1-v: Reinforcing super generalization ability in vision-language models with less than \ 3.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis R1-v: Reinforcing super generalization ability in vision-language models with less than \ 3

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:51.265716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:35:42.780766Z digest=sha256:22f582e669e671e595b286f0a02e9ccde226cf4ffa1b2bdc4647a2c7a0dfcc02

Observation 34d88ea0-756b-4e34-ae93-b1fdaa4ed3b3 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:42.887829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:42.887829Z digest=sha256:48ee9456846090216fbc55426719b64f48118028e204aa1d3ef7fbe94f68cdb6

Observation 6e2a18e6-d565-4ba6-8acc-37d8924be19d · outbound

This paper cites Chatbot arena: An open platform for evaluating llms by human preference.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Chatbot arena: An open platform for evaluating llms by human preference

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:42.993498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:42.993498Z digest=sha256:9900160d5f03f616cdafd1b6d974966b3f479a484078bdc72ab9f81272d094f3

Observation f831c92e-44e1-4380-a5fa-01ac675de575 · outbound

This paper cites Biomedical Visual Instruction Tuning with Clinician Preference Alignment.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Biomedical Visual Instruction Tuning with Clinician Preference Alignment

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:43.091029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:43.091029Z digest=sha256:8ae0313a0159fdccafe85c0249fc5adb871b67b649253e2247bb44488cfb0559

Observation 79b86d50-af30-48d0-9a86-9aad26524e5f · outbound

This paper cites Theorem-Validated Reverse Chain-of-Thought Problem Generation for Geometric Reasoning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Theorem-Validated Reverse Chain-of-Thought Problem Generation for Geometric Reasoning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:43.221451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:43.221451Z digest=sha256:11bae027a6645229d2f6dd550a3a3fa73197dc869b43a3c28aa459d7d0be98a3

Observation ed87c766-a6dc-47cb-b5cb-d88aa5d72c1b · outbound

This paper cites OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:43.330907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:43.330907Z digest=sha256:184d698203286b4b72ba5e7e5ee28b52f2401c356f0dbf7c9cf107f3d2ab4c29

Observation 0fe375c1-c7b5-484e-a161-31bcf0593054 · outbound

This paper cites How Abilities in Large Language Models are Affected by Supervised Fine-tuning Data Composition.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis How Abilities in Large Language Models are Affected by Supervised Fine-tuning Data Composition

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:43.435240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:43.435240Z digest=sha256:3c81e9e3dc0b62c0cc801fb91e98f36e3b5340fc33ce5f04611ed6b927691c72

Observation 4a671aa5-87d9-4126-adb4-878b76fe86be · outbound

This paper cites What Makes for Good Visual Instructions? Synthesizing Complex Visual Reasoning Instructions for Visual Instruction Tuning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis What Makes for Good Visual Instructions? Synthesizing Complex Visual Reasoning Instructions for Visual Instruction Tuning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:43.517803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:43.517803Z digest=sha256:2171c3d4877ef46b460b5ba78da72878ec06d9912d5b51150bbf9362a43aba42

Observation cf0d2612-911f-4d4c-8209-2f33268bd7b4 · outbound

This paper cites Solution of a ranking problem from binary comparisons.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Solution of a ranking problem from binary comparisons

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:51.222509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:35:43.625141Z digest=sha256:cf0c402e034ad2ea3fb08e9c07dee8e410913a2db3bef06ea3a941b82d3ce455

Observation 650db308-b638-43f2-8d5d-8f3faaed4d97 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Gemini: A Family of Highly Capable Multimodal Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:43.700132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:43.700132Z digest=sha256:f32c3508f2048156bcf98ec90073bea76628b181a64f0e609cdd262ce897393c

Observation ade4cd4c-d9ac-44d1-b66e-6ac3a711b9fc · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:43.799581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:43.799581Z digest=sha256:2100d3a5738fe14393bdd23bb604e3791cdf1219d4e0464422f31f9066d17a2d

Observation cabece22-cff2-4f1c-a96d-23c6efec9290 · outbound

This paper cites Bias in Large Language Models: Origin, Evaluation, and Mitigation.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Bias in Large Language Models: Origin, Evaluation, and Mitigation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:43.887636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:43.887636Z digest=sha256:b1fff213ba5837c70798e7b93a2723c578b342e7a8175c9350e762f14d2a66dc

Observation 2db529f2-7589-4f6d-8e6e-4ba01d9e4c8e · outbound

This paper cites Minimax-optimal inference from partial rankings.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Minimax-optimal inference from partial rankings

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:51.184027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:35:43.964365Z digest=sha256:dff87708b99795d2ef6fb48024cccb86eb09f85c80065ffb13b12e0b5904ce14

Observation d12f446e-6c97-4192-adc2-204a1ced061a · outbound

This paper cites Multi-modal Synthetic Data Training and Model Collapse: Insights from VLMs and Diffusion Models.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Multi-modal Synthetic Data Training and Model Collapse: Insights from VLMs and Diffusion Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:44.065746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:44.065746Z digest=sha256:33f02b2d3813bc9d00aec532d64144278dc772fb5e8d9903134cd8c44fd83891

Observation bf28e19e-dbe7-4195-b91d-6aed987e7192 · outbound

This paper cites Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:44.142868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:44.142868Z digest=sha256:ad735c2cdeb5c91f8429728eeb8999e6adb890921817737a436797f857a57b83

Observation 9ced2ab8-255d-4e7b-b565-a4a820322838 · outbound

This paper cites GPT-4o System Card.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis GPT-4o System Card

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:44.242073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:44.242073Z digest=sha256:bebdd000a2e205f4ca3f79bf860b5ca29e15f70dace7a82c7e8a0333acc9e574

Observation 747b2510-aa0e-4a01-90fb-270351244842 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:44.334513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:44.334513Z digest=sha256:00c46d9fdbbabdbfcde5db84ef5c290bb5d890d23a472b2c00f8653f22392144

Observation 3de466ae-685b-4712-88f4-c2efc8f103b9 · outbound

This paper cites Kimi-VL Technical Report.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Kimi-VL Technical Report

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:44.484635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:44.484635Z digest=sha256:7ff4b373ca367fb8a62e7f9d57af0f2565699532ce472af27a3950dc0d4def3e

Observation 2fe1a62f-9281-4da4-a788-23b525eb8245 · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Gonzalez, Hao Zhang, and Ion Stoica

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:44.582583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:44.582583Z digest=sha256:91260bb2953b63d29456ae6e9a34bb6b7e2a8b5521e955038483697c82537358

Observation 31790328-208c-4bb1-bc96-3a1ef76af91e · outbound

This paper cites Llava-next: Stronger llms supercharge multimodal capabilities in the wild, May 2024 a.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Llava-next: Stronger llms supercharge multimodal capabilities in the wild, May 2024 a

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:51.110422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:35:44.677095Z digest=sha256:f4915efe451d965b2dad441d9f983fe5a11d41e6c7d96c9f06fbc724f1f9e392

Observation 2904867b-9b6f-4ca3-bad0-3d10a0cf5cb8 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis LLaVA-OneVision: Easy Visual Task Transfer

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:44.784577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:44.784577Z digest=sha256:7fd2f1d524b76554df074e168c57f062a654c6888a58416b513557b03db15d41

Observation de55569d-7d7e-4a88-ad9a-826430d87c03 · outbound

This paper cites LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:44.888994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:44.888994Z digest=sha256:fd5df447ceb90ecb94bac5d01a9c1a41a88cd304ec27a65ab4f98b1da012d997

Observation c716f602-1221-444c-be7a-4a8e7e71a635 · outbound

This paper cites TextBind: Multi-turn Interleaved Multimodal Instruction-following in the Wild.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis TextBind: Multi-turn Interleaved Multimodal Instruction-following in the Wild

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:35:49.344555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:35:44.982415Z digest=sha256:dbc6a5bc49bb74d6e386ef701ac4fd7f59a04d7658f9d6c149dc5f7b147ef9ca

Observation d630d306-b179-4a60-87f6-a214bf57fa22 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:50.850222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:35:45.070368Z digest=sha256:6a4f51e591e06476c40fddec23a4512de1d80c25fb1218833dbc80274b034f26

Observation 4d4f2cae-024a-4301-a03e-21f1e91b980c · outbound

This paper cites VLFeedback: A Large-Scale AI Feedback Dataset for Large Vision-Language Models Alignment.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis VLFeedback: A Large-Scale AI Feedback Dataset for Large Vision-Language Models Alignment

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.165859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.165859Z digest=sha256:04909cd978baf3174352c0fe127d083f2dce7e9d0c52a2aeba8c5ae30a912915

Observation 510385c4-0869-43fe-8cbe-16199966b57d · outbound

This paper cites Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.240736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.240736Z digest=sha256:7f7ed5d4552d2347f4e4cfa58823de2bfc9a56b9e36955bd281a29c100564a8f

Observation 07818c08-b6e3-4a43-a672-edcb1c042834 · outbound

This paper cites LIMR: Less is More for RL Scaling.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis LIMR: Less is More for RL Scaling

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.327280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.327280Z digest=sha256:da03a7e4889e13d9630f689f169b00afc7cc403abf96e9dd0e55f9e0e0bcdc6f

Observation 3df0d95d-4c97-4b3c-8f89-3abd196b1efb · outbound

This paper cites Visual instruction tuning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Visual instruction tuning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.411544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.411544Z digest=sha256:06427ae0401249cca68cb53c4dd2680147687c51457b711c2089b7c51bdd6f6c

Observation 7d6f53ad-e226-4e89-b634-75ebfba23a18 · outbound

This paper cites Improved baselines with visual instruction tuning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Improved baselines with visual instruction tuning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.521488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.521488Z digest=sha256:9fb4f11f526adaa0b0533c64eb36acd173b191e0063b4a27e6c6a0d8ca628978

Observation 3b0fda6a-c7cb-4060-b6c8-0f1741d7eeed · outbound

This paper cites What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction Tuning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction Tuning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.608563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.608563Z digest=sha256:1a7675e7757cf1bb6b7dcf1147eefb9e1e6b1625f6dbdc179ea99fc261260dfe

Observation ba55c519-66f4-4e96-97b8-e4c7746baaf7 · outbound

This paper cites Noisyrollout: Reinforcing visual reasoning with data augmentation, 2025 a.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Noisyrollout: Reinforcing visual reasoning with data augmentation, 2025 a

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.685696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.685696Z digest=sha256:987f2c1266b9587d1eaaca766740fc5cac54b436b0ffa003c29c964991a425f6

Observation e4243140-b616-4e12-a060-0fa3aa769e70 · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Understanding R1-Zero-Like Training: A Critical Perspective

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.759914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.759914Z digest=sha256:a22675b6dbbb8f43352b6b797a76436b5d7dcbbc74b0697b7b0efe723ecc4c00

Observation 4a7f63d8-7d4d-4f82-a545-29279563f384 · outbound

This paper cites MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.833316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.833316Z digest=sha256:28f06619e4fae6a2fcd2d6c71d64abf1a253984c0c6a0ef0f574077ae723a78f

Observation a0e4118f-8b8e-4283-ba7e-455aca61d857 · outbound

This paper cites Ursa: Understanding and verifying chain-of-thought reasoning in multimodal mathematics.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Ursa: Understanding and verifying chain-of-thought reasoning in multimodal mathematics

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.903561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.903561Z digest=sha256:c8509e7fee2510a5296e16f4dcf8ea8e5bf235b480a354f34da827087e56116c

Observation 0a3c5ca4-c5d8-4559-ac15-4050a424d602 · outbound

This paper cites MMEvol: Empowering Multimodal Large Language Models with Evol-Instruct.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis MMEvol: Empowering Multimodal Large Language Models with Evol-Instruct

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.979553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.979553Z digest=sha256:0a658fafc32ac8f49d273631fb36f76bad707d0ba53d4144f3d0c1e354f13368

Observation a098ea00-8092-40cc-8023-b1a7c52b535b · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.053056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.053056Z digest=sha256:2259f3658749753e24605e5a02f7f29e077e0c2401ca0a6bd8b4caaaa66c4864

Observation 12809237-5dff-433c-aa41-d28f3afc8c38 · outbound

This paper cites Iterative ranking from pair-wise comparisons.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Iterative ranking from pair-wise comparisons

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.131832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.131832Z digest=sha256:ab8b3bade5da22ebad313b69ddc88da7def95a97d35823be324f9548e847c874

Observation 16ee4cb3-d35d-43bc-bbca-0e8a180b1b8f · outbound

This paper cites LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.225972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.225972Z digest=sha256:5201c7ad97ded4e369f7aba08b7eb28d265690fc7f97f0ca35fefcedaf966ee1

Observation b474b799-dfb0-4dc5-b822-dc123530c909 · outbound

This paper cites We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.295497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.295497Z digest=sha256:9cd8d6dc9388bf3bec403f4f6bd4a7a47b0e244472443b597874d49b82ff0116

Observation 02404c31-7369-43d5-8fcc-1bd93236094a · outbound

This paper cites Proximal Policy Optimization Algorithms.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Proximal Policy Optimization Algorithms

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.364671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.364671Z digest=sha256:3e8d0dba593c63a682ff70ec0023fd50bcd65678ff464545e3e4edabd5bb1189

Observation 1a776c3c-4134-4707-bc44-50b7809548a0 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.455575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.455575Z digest=sha256:ffeefdd0311d5c0a1952ddfcfd5a85dedbf2b11393c3bb3d9b954c71b87c2a92

Observation 3607d390-bcba-4434-aa0f-3a24e1420cde · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis HybridFlow: A Flexible and Efficient RLHF Framework

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.547506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.547506Z digest=sha256:d581aec2f47f1f065d926677a47b5760408501b1b285f8e2bcc4e161b6078331

Observation 4a3f3a05-f5fc-4a3f-a88d-7f220ec9a4a6 · outbound

This paper cites Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.632301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.632301Z digest=sha256:a8175f905862d9d882d88399a044c7e138328a7de962a790338baaf0bf408b66

Observation 814b1d37-28b1-4257-8d38-56b0aa65d387 · outbound

This paper cites Some rank order tests which are most powerful against specific parametric alternatives.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Some rank order tests which are most powerful against specific parametric alternatives

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:50.605631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:35:46.730566Z digest=sha256:a37edfd989b884da383319d352cc5a0041e537b74fffd5dd3b1fdbfdb49abc23

Observation c7cd2c15-bf9d-4f7e-a6e5-441590bccb17 · outbound

This paper cites Dart-math: Difficulty-aware rejection tuning for mathematical problem-solving.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Dart-math: Difficulty-aware rejection tuning for mathematical problem-solving

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.817678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.817678Z digest=sha256:9c755826ebb78ce77427511d4cf49530eb09b5129b681f546cd09e01418433dd

Observation a903f51d-b1df-4c2a-b767-3e18cd230981 · outbound

This paper cites Measuring multimodal mathematical reasoning with math-vision dataset.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Measuring multimodal mathematical reasoning with math-vision dataset

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:50.271280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:35:46.908919Z digest=sha256:b068b2fad7057a2ff85184684d80a9f58b4878495cf45d3078d3f3330677c574

Observation db2b616b-3222-4f49-9bcc-45395e56de21 · outbound

This paper cites Agri-LLaVA: Knowledge-Infused Large Multimodal Assistant on Agricultural Pests and Diseases.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Agri-LLaVA: Knowledge-Infused Large Multimodal Assistant on Agricultural Pests and Diseases

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.989091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.989091Z digest=sha256:5fd390822bafe07c252b40faaea1c33cc14b2534f085ae81ea59b95de455c9a8

Observation 04625a6a-c12b-4741-a4d4-b573b244768e · outbound

This paper cites SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.087104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.087104Z digest=sha256:0d8cfed0d74108f741aca7d9ab7c351ca3b4d73566c05eddd7ae1b72cc09a707

Observation 1debfe81-31f3-4173-9bf5-783109667a66 · outbound

This paper cites QuRating: Selecting High-Quality Data for Training Language Models.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis QuRating: Selecting High-Quality Data for Training Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.179506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.179506Z digest=sha256:9a1ad052ae176f3e601a4b2b72b541c3fa48159b8193268761eca1cb3042c8e9

Observation 1fd56d91-44f8-4727-9c8f-17c273d6c01f · outbound

This paper cites LESS: Selecting Influential Data for Targeted Instruction Tuning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.274238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.274238Z digest=sha256:9cbeb32813dc648124e112e0405f1aaa9aba95fa517c3c3f7e8a21bca5a32c65

Observation e8e5cbd9-5d77-44bd-99a1-e8527c71921b · outbound

This paper cites WizardLM: Empowering large pre-trained language models to follow complex instructions.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis WizardLM: Empowering large pre-trained language models to follow complex instructions

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.349754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.349754Z digest=sha256:c66bfbe1569fcc6c4cbd111a265b90736582c5217ba61835143ebea664e8d8fb

Observation 48ce8203-2fb1-4340-8e8a-e8328f62ec26 · outbound

This paper cites Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.442353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.442353Z digest=sha256:71643cf2bb693d924f2a43ce161df71609d95e42c0cb53d1e89c146f2834fb73

Observation 9945eb6f-2e2a-44af-a382-649b997bffdb · outbound

This paper cites R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.538238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.538238Z digest=sha256:67b6b932b7f331000ecc9381cf52c5c47dd74472eded9a3bd624091f878ea04e

Observation e11d0494-ee93-48e0-b805-ec5b346c0859 · outbound

This paper cites Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.633710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.633710Z digest=sha256:e64b7e342945486f5d0b2d69fe0a14b402fbea228dd0c5459f34fd06bfa4592f

Observation 4d6ed220-c3b6-406f-a283-3abb183608da · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.704809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.704809Z digest=sha256:88a4adbee2cf46c38e9d69c412f9b66f7f0f2d7aa75b118c46a45cb823a890b6

Observation 5816efab-176e-4812-b4af-4645f840aa09 · outbound

This paper cites VAPO: Efficient and Reliable Reinforcement Learning for Advanced Reasoning Tasks.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis VAPO: Efficient and Reliable Reinforcement Learning for Advanced Reasoning Tasks

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.789603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.789603Z digest=sha256:02a930f880a9af85eb1d07ffeba9f1d312acc1290edcf81f62147c702973838a

Observation 02528de9-975f-4329-9a32-8d8299e7940b · outbound

This paper cites SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.858390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.858390Z digest=sha256:7ccd9ce3f50b4f159565bac1e09d2b9be9f41a35911711e4913a0e9bed861308

Observation 3cdf3911-6cfc-49e3-9a6e-db665d108f64 · outbound

This paper cites R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.951687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.951687Z digest=sha256:465f1a0927b05ea97058000b8fc1e67a7d871770e3e3d986fd066aac5e2abc5a

Observation 2b492e3b-9148-4658-aeac-c5b03ebb118c · outbound

This paper cites Mathverse: Does your multi-modal llm truly see the diagrams in visual math problems? In European Conference on Computer Vision, pp.\ 169--186.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Mathverse: Does your multi-modal llm truly see the diagrams in visual math problems? In European Conference on Computer Vision, pp.\ 169--186

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:48.064298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:48.064298Z digest=sha256:d225154b135a7ff2c097da41925115bae30f5cb0867ff524ea31704a872fcfd9

Observation 85b0c8d5-e330-4b54-b0ee-0a09458f3881 · outbound

This paper cites Mavis: Mathematical visual instruction tuning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Mavis: Mathematical visual instruction tuning

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:49.943274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:35:48.156999Z digest=sha256:fc57891730ef8cb174b40f103b211235dfac7c27ece917a32164932bddb8bd23

Observation 0b217a9b-0c96-425e-be90-cf67935278db · outbound

This paper cites Easyr1: An efficient, scalable, multi-modality rl training framework.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Easyr1: An efficient, scalable, multi-modality rl training framework

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:48.225424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:48.225424Z digest=sha256:6c10041565ad3b9e067519ef5d7c4e10fa47f373f91b619b89f512f4dd1483c3

Observation 6efe3762-95b1-408a-bc3b-ca878f1a1378 · outbound

This paper cites Lima: Less is more for alignment.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Lima: Less is more for alignment

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:48.312203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:48.312203Z digest=sha256:3c7e8fe47750a7411d5edfeeed9a140bd9c85e8c24c508e120857d05a386f2d9

Observation 9f7571fe-8f07-48d5-a41e-884b3b99e191 · outbound

This paper cites NavGPT-2: Unleashing Navigational Reasoning Capability for Large Vision-Language Models.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis NavGPT-2: Unleashing Navigational Reasoning Capability for Large Vision-Language Models

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:48.404458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:48.404458Z digest=sha256:3e0754c73fbe64f9de90971254de17b678301ebf4b9dc84b4e810d7877a6df32

Observation a78461e9-b9f9-4af6-bfdf-124de97dc7f6 · outbound

This paper cites Anyprefer: An Agentic Framework for Preference Data Synthesis.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Anyprefer: An Agentic Framework for Preference Data Synthesis

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:48.503722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:48.503722Z digest=sha256:ea2b3753b8ea40d9dbde2883ae727ec1dc92dfdbae9770b324efccf881b9e135

Observation 78b6cf58-0845-45a2-95e5-16bb60f350e7 · outbound

This paper cites Dynamath: A dynamic visual benchmark for evaluating mathematical reasoning robustness of vision language models, 2024.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Dynamath: A dynamic visual benchmark for evaluating mathematical reasoning robustness of vision language models, 2024

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:49.749976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:35:48.604808Z digest=sha256:ec384bf1270fbbbfa835b622b0ae46c5358d0cc107308bf27ece95cf365e88b0

Observation f1303788-f8f7-4a4d-a01a-64db254eba24 · outbound

This paper cites write newline.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis write newline

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:48.656605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:48.656605Z digest=sha256:85c0a02f423e9bafc55b9fcb513ba960a8cb83b4bd920e84d05b129d274e5140

Pith citing papers

Observation 3e6ecd59-cd6a-4989-995d-28ea935f3b0c · inbound

SCALER:Synthetic Scalable Adaptive Learning Environment for Reasoning cites this paper.

SCALER:Synthetic Scalable Adaptive Learning Environment for Reasoning SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:28:05.710085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T16:24:04.132572Z digest=sha256:43d0cbdf7b6b048ae77b59d590c1e5c3cc63a65f3fdab185d143e339ad56274b

Observation a65cee66-e7eb-462a-a1eb-7559aa9d25da · inbound

GaMMA: Towards Joint Global-Temporal Music Understanding in Large Multimodal Models cites this paper.

GaMMA: Towards Joint Global-Temporal Music Understanding in Large Multimodal Models SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:47:17.195171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-09T19:12:28.290483Z digest=sha256:c15a358d8340dfa19a62998f776709bda1654018b62a569172e779847c728e8f

Observation 6f0c6a90-80f6-4654-a8b5-00947dc73b4e · inbound

Generate, Filter, Control, Replay: A Comprehensive Survey of Rollout Strategies for LLM Reinforcement Learning cites this paper.

Generate, Filter, Control, Replay: A Comprehensive Survey of Rollout Strategies for LLM Reinforcement Learning SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis

Reference 142

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:15:49.390205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T19:15:27.406778Z digest=sha256:34a505be8e354c127221c1ca2c8402e964aba387a855e53217cf72564f34217c

Observation d658f3bf-0e1d-456a-b0ec-cbf42d7578a0 · inbound

Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning cites this paper.

Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T11:47:27.883967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:47:27.883967Z digest=sha256:73dee4cdb2d0622c4992857882959012b3e571391777e5eea375f3b67479a0cc