Pith. sign in

Paper Citation Record · LEDGER

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis

As of 16 August 2026, this Paper Citation Record lists 76 of 76 outbound references and 4 inbound Pith citation observations for arXiv:2506.02096.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02096 v1

Coverage vector

measured 76 of 76 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:35:48.656605Z

measured 80 of 80 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T11:47:27.883967Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T16:28:05.708078Z

Reference resolution

76 of 76 outbound references displayed

  • verified exact1
  • verified fuzzy11
  • unresolved64
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c6187cca-91c3-42de-8472-71c141a21060 · outbound

This paper cites @esa (Ref.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis @esa (Ref

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:41.936149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:41.936149Z digest=sha256:d7f638bd1554c98f53fe275397e8ab543116b93f3e2241ae66bb089a34c89ba6

Observation 38def998-febb-43b9-8077-7ad487d07064 · outbound

This paper cites an unresolved cited work.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:41.977593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:41.977593Z digest=sha256:9e4ae9498a11ccbc2b50412671ec404d89aa31afa6deb0d6032ccee5b7812122

Observation 06d238d5-bf6a-4608-a501-72ae996bdf8b · outbound

This paper cites an unresolved cited work.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:35:51.317857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T11:35:42.069904Z digest=sha256:f8b2bdd1ec577e8ff6d7b6c4b2278d460a7fd49bafcdff8b798ded1cee303ba9

Observation 7833eefd-422b-4e3d-87df-b4bf2f8ceafd · outbound

This paper cites Phi-4-reasoning Technical Report.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Phi-4-reasoning Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:42.180779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:42.180779Z digest=sha256:86e8e224d721a3adfd419f2db180fd9722c96f73a79bcdf95568f3521fbfece1

Observation 7ec446e8-42d1-46ff-8d8c-e3df9114472a · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Flamingo: a visual language model for few-shot learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:42.252690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:42.252690Z digest=sha256:ae68f3d670503367c1bc5ca57285e811520056649f9d27aef0cd85c2f3d34c5c

Observation a53a029f-0c75-451b-9f4b-32788710af0b · outbound

This paper cites Claude 3.7 sonnet.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Claude 3.7 sonnet

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:51.299837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T11:35:42.377165Z digest=sha256:ca74f36d0ec4eb972148c00c1a2988501cea81d240ac3c9dc83411fe11fff81d

Observation 24131374-3806-4923-a548-3d4b089c74d4 · outbound

This paper cites Qwen2.5-VL Technical Report.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Qwen2.5-VL Technical Report

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:42.453668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:42.453668Z digest=sha256:c58f6fe0fd45bb9917b09450226dca8bf4cbfb01ff44febac23b518b3c0d789c

Observation 2ee0dfea-5b0c-4812-9145-a8dba7689ad3 · outbound

This paper cites A Survey of Multimodal Large Language Model from A Data-centric Perspective.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis A Survey of Multimodal Large Language Model from A Data-centric Perspective

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:42.563075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:42.563075Z digest=sha256:6bae93a11de3072057cf7aa32d344cbd1eaf46055e98bbfbaf3bc9b4ab8f7516

Observation c22a30bb-9997-482b-914d-9d91e59d0a67 · outbound

This paper cites Rank analysis of incomplete block designs: I.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Rank analysis of incomplete block designs: I

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:42.642750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:42.642750Z digest=sha256:7bee03bc9e16b16a014437e935f25c8ec45a396631895e727bbc1e6026095cc3

Observation 4bc9233c-8ceb-46ba-8cda-9ae023448025 · outbound

This paper cites Sft or rl? an early investigation into training r1-like reasoning large vision-language models.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Sft or rl? an early investigation into training r1-like reasoning large vision-language models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:51.279167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T11:35:42.716759Z digest=sha256:c1590bb24799dbff9e86b097dc66954e101ab3ba411d5e4ecd282478dea3c0f9

Observation ae23a085-3fb9-4407-85a0-a934e545035e · outbound

This paper cites R1-v: Reinforcing super generalization ability in vision-language models with less than \ 3.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis R1-v: Reinforcing super generalization ability in vision-language models with less than \ 3

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:51.265716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T11:35:42.780766Z digest=sha256:9e207968f40f6f7e128ebc685dd29efb3c29183c87ce7d1416f227bec0a3e931

Observation 34d88ea0-756b-4e34-ae93-b1fdaa4ed3b3 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:42.887829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:42.887829Z digest=sha256:5a35b3cd4d7d9db3c1ae79ca6a62cebed79102e8c004dc684a7b9cd52ddc0e15

Observation 6e2a18e6-d565-4ba6-8acc-37d8924be19d · outbound

This paper cites Chatbot arena: An open platform for evaluating llms by human preference.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Chatbot arena: An open platform for evaluating llms by human preference

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:42.993498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:42.993498Z digest=sha256:067d57efa35023d0e7689f78e34267a14bc0f2a81ae9d4679b1e72998de9dae8

Observation f831c92e-44e1-4380-a5fa-01ac675de575 · outbound

This paper cites Biomedical Visual Instruction Tuning with Clinician Preference Alignment.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Biomedical Visual Instruction Tuning with Clinician Preference Alignment

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:43.091029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:43.091029Z digest=sha256:c5cc11779637bb9b3bbf9b99c9e1d4af75c89a3a06dd03c62fae1a2d48e16ec2

Observation 79b86d50-af30-48d0-9a86-9aad26524e5f · outbound

This paper cites Theorem-Validated Reverse Chain-of-Thought Problem Generation for Geometric Reasoning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Theorem-Validated Reverse Chain-of-Thought Problem Generation for Geometric Reasoning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:43.221451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:43.221451Z digest=sha256:4485347d5686810ba7d5fbd2f760484a58f391ecd67e9557e6606fc78e83046f

Observation ed87c766-a6dc-47cb-b5cb-d88aa5d72c1b · outbound

This paper cites OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:43.330907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:43.330907Z digest=sha256:68d191d9141117bba658458cbfe6aa38b075e42eff301462b4958d15b3a3f828

Observation 0fe375c1-c7b5-484e-a161-31bcf0593054 · outbound

This paper cites How Abilities in Large Language Models are Affected by Supervised Fine-tuning Data Composition.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis How Abilities in Large Language Models are Affected by Supervised Fine-tuning Data Composition

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:43.435240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:43.435240Z digest=sha256:a4f09c75bb522644b5e194e247fe2d1e3441237f45ec63ccf04ea498a41b26d5

Observation 4a671aa5-87d9-4126-adb4-878b76fe86be · outbound

This paper cites What Makes for Good Visual Instructions? Synthesizing Complex Visual Reasoning Instructions for Visual Instruction Tuning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis What Makes for Good Visual Instructions? Synthesizing Complex Visual Reasoning Instructions for Visual Instruction Tuning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:43.517803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:43.517803Z digest=sha256:2ba12b086b4d51df64fb31fb9539a2cba5922f70d8ac70e77d66b2894c20ba83

Observation cf0d2612-911f-4d4c-8209-2f33268bd7b4 · outbound

This paper cites Solution of a ranking problem from binary comparisons.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Solution of a ranking problem from binary comparisons

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:51.222509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T11:35:43.625141Z digest=sha256:15cea4bf909237df51d3934adc208e6bbe88115a9d60800f67bb2ff152bfb5dc

Observation 650db308-b638-43f2-8d5d-8f3faaed4d97 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Gemini: A Family of Highly Capable Multimodal Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:43.700132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:43.700132Z digest=sha256:60730e171d82e20c2138edcd1b23d0e090ff04f78e3d940795437aed50856a28

Observation ade4cd4c-d9ac-44d1-b66e-6ac3a711b9fc · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:43.799581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:43.799581Z digest=sha256:d429921d7a764f46aac9620e63ae8eab7458bdbebf1d93849cb637e0b0deec9f

Observation cabece22-cff2-4f1c-a96d-23c6efec9290 · outbound

This paper cites Bias in Large Language Models: Origin, Evaluation, and Mitigation.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Bias in Large Language Models: Origin, Evaluation, and Mitigation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:43.887636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:43.887636Z digest=sha256:4be996f2bdbee73546beefc4c4e3e2807009fe9e0c079b4da3b8612e7d4dc474

Observation 2db529f2-7589-4f6d-8e6e-4ba01d9e4c8e · outbound

This paper cites Minimax-optimal inference from partial rankings.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Minimax-optimal inference from partial rankings

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:51.184027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T11:35:43.964365Z digest=sha256:89ee7d7e491d85bddff0f286e9c1ac3f06ecd5e1e2eda2ef8697f06fa881b8e4

Observation d12f446e-6c97-4192-adc2-204a1ced061a · outbound

This paper cites Multi-modal Synthetic Data Training and Model Collapse: Insights from VLMs and Diffusion Models.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Multi-modal Synthetic Data Training and Model Collapse: Insights from VLMs and Diffusion Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:44.065746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:44.065746Z digest=sha256:08b05afe91660261363833b2b21ec36744f44ce6bda7a6725eed6cc52a3719f8

Observation bf28e19e-dbe7-4195-b91d-6aed987e7192 · outbound

This paper cites Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:44.142868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:44.142868Z digest=sha256:75fbd3d46fb561dfadb3952b8c0a023099c2af0824fb537df5c73f0362d20d95

Observation 9ced2ab8-255d-4e7b-b565-a4a820322838 · outbound

This paper cites GPT-4o System Card.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis GPT-4o System Card

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:44.242073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:44.242073Z digest=sha256:99945850b97d7cdc9025bc8591b62b8e29da19491493a0e9865218248f1743fa

Observation 747b2510-aa0e-4a01-90fb-270351244842 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:44.334513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:44.334513Z digest=sha256:cf329489c7984528b8303a9fca1e39cde1bdc4fcb38eed4f8703d25135b2ef5e

Observation 3de466ae-685b-4712-88f4-c2efc8f103b9 · outbound

This paper cites Kimi-VL Technical Report.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Kimi-VL Technical Report

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:44.484635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:44.484635Z digest=sha256:a3d1fc32e79ee1636f26f204048cb6701e9a9db5cd83f0e1701cfeac2ccbdfb9

Observation 2fe1a62f-9281-4da4-a788-23b525eb8245 · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Gonzalez, Hao Zhang, and Ion Stoica

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:44.582583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:44.582583Z digest=sha256:bbcb2145fdfdd9de092f23f17bd291c27485d5e3c76c7def626efb378acbe367

Observation 31790328-208c-4bb1-bc96-3a1ef76af91e · outbound

This paper cites Llava-next: Stronger llms supercharge multimodal capabilities in the wild, May 2024 a.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Llava-next: Stronger llms supercharge multimodal capabilities in the wild, May 2024 a

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:51.110422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T11:35:44.677095Z digest=sha256:424900fd889aab5544f06e08b44a964487a4c0bbfde14e77af263706c8e930fb

Observation 2904867b-9b6f-4ca3-bad0-3d10a0cf5cb8 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis LLaVA-OneVision: Easy Visual Task Transfer

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:44.784577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:44.784577Z digest=sha256:90aa2d4d163986f9cfa72291a8f2ea3dbd4145045d2cd73639c47506129be5dd

Observation de55569d-7d7e-4a88-ad9a-826430d87c03 · outbound

This paper cites LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:44.888994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:44.888994Z digest=sha256:08b349eb509dd46f18c342e5205dbb753cd0fbacb54b2af79d8141e22fca5d45

Observation c716f602-1221-444c-be7a-4a8e7e71a635 · outbound

This paper cites TextBind: Multi-turn Interleaved Multimodal Instruction-following in the Wild.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis TextBind: Multi-turn Interleaved Multimodal Instruction-following in the Wild

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:35:49.344555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T11:35:44.982415Z digest=sha256:9e920b5587b10b5b3baa145a9509286e5f0e7e9384566f3cfbd59fa025e08133

Observation d630d306-b179-4a60-87f6-a214bf57fa22 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:50.850222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T11:35:45.070368Z digest=sha256:90b2e44c28c1e3d03e040d93ddd454abdc7d5a44c6627bcb89caa2dfa3d218d3

Observation 4d4f2cae-024a-4301-a03e-21f1e91b980c · outbound

This paper cites VLFeedback: A Large-Scale AI Feedback Dataset for Large Vision-Language Models Alignment.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis VLFeedback: A Large-Scale AI Feedback Dataset for Large Vision-Language Models Alignment

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.165859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.165859Z digest=sha256:8b3bb488094ee609089ba2c917c9b56711885fd84c927093d6dd76e4947615e4

Observation 510385c4-0869-43fe-8cbe-16199966b57d · outbound

This paper cites Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.240736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.240736Z digest=sha256:6c676b1745b860503df02cc14e7a13eeea4c00cc60652597febe48a7b5ddfa88

Observation 07818c08-b6e3-4a43-a672-edcb1c042834 · outbound

This paper cites LIMR: Less is More for RL Scaling.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis LIMR: Less is More for RL Scaling

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.327280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.327280Z digest=sha256:3771f728794c919f11d28101cc6656743f93e8ffd66c71ee27d76d23b03baeef

Observation 3df0d95d-4c97-4b3c-8f89-3abd196b1efb · outbound

This paper cites Visual instruction tuning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Visual instruction tuning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.411544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.411544Z digest=sha256:c5967fd4738c88ad14719b037835a8b388f47526f0a4b96105268fa2a8b4254e

Observation 7d6f53ad-e226-4e89-b634-75ebfba23a18 · outbound

This paper cites Improved baselines with visual instruction tuning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Improved baselines with visual instruction tuning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.521488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.521488Z digest=sha256:f535526feedddfc981f655744fa8539cab361528f088b593ee580ab917993a94

Observation 3b0fda6a-c7cb-4060-b6c8-0f1741d7eeed · outbound

This paper cites What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction Tuning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction Tuning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.608563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.608563Z digest=sha256:f2bfcf965662e6134ae5d4dd9fe08b674b77b1ae549f3c0ccfc15a0b16a02230

Observation ba55c519-66f4-4e96-97b8-e4c7746baaf7 · outbound

This paper cites Noisyrollout: Reinforcing visual reasoning with data augmentation, 2025 a.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Noisyrollout: Reinforcing visual reasoning with data augmentation, 2025 a

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.685696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.685696Z digest=sha256:40e1b913c4db3ab33b3967efd2cfb79f4e9d9cb541ab9414dfd614b978342cd3

Observation e4243140-b616-4e12-a060-0fa3aa769e70 · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Understanding R1-Zero-Like Training: A Critical Perspective

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.759914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.759914Z digest=sha256:d82f952979b02ffc2f4d37f827429d0d5558f394600bfd258b8e9f5e8704b592

Observation 4a7f63d8-7d4d-4f82-a545-29279563f384 · outbound

This paper cites MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.833316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.833316Z digest=sha256:d0e7827d8c98e779c12ff735c431bad3b609f634f148ec5f19f627defbd71eb1

Observation a0e4118f-8b8e-4283-ba7e-455aca61d857 · outbound

This paper cites Ursa: Understanding and verifying chain-of-thought reasoning in multimodal mathematics.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Ursa: Understanding and verifying chain-of-thought reasoning in multimodal mathematics

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.903561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.903561Z digest=sha256:787a81400381f3e86e3dccc093b7134a326b3fb5c2f16ab0acb1af6f9b58868d

Observation 0a3c5ca4-c5d8-4559-ac15-4050a424d602 · outbound

This paper cites MMEvol: Empowering Multimodal Large Language Models with Evol-Instruct.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis MMEvol: Empowering Multimodal Large Language Models with Evol-Instruct

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:45.979553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:45.979553Z digest=sha256:28e2102a113ec7a8a5eb16baa553c150c1757b6ac30bbc4b95854e720d9121ca

Observation a098ea00-8092-40cc-8023-b1a7c52b535b · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.053056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.053056Z digest=sha256:9639b7e07778a575c42114b3f54fb37dbbfa162e984a7983cf005bbba2ad9514

Observation 12809237-5dff-433c-aa41-d28f3afc8c38 · outbound

This paper cites Iterative ranking from pair-wise comparisons.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Iterative ranking from pair-wise comparisons

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.131832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.131832Z digest=sha256:d942c1cb3258bacbea1bfc93198ce67ee6fc9731e71dde0075789fba8bf73b30

Observation 16ee4cb3-d35d-43bc-bbca-0e8a180b1b8f · outbound

This paper cites LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.225972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.225972Z digest=sha256:02095e3d86eca129217be4e9e433997cab2201b86603fad4afe0c504c3ae3215

Observation b474b799-dfb0-4dc5-b822-dc123530c909 · outbound

This paper cites We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.295497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.295497Z digest=sha256:dd0a34c6bc8ee80cd46013ee5da43e21e2db98abb9c3fc015334463451606749

Observation 02404c31-7369-43d5-8fcc-1bd93236094a · outbound

This paper cites Proximal Policy Optimization Algorithms.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Proximal Policy Optimization Algorithms

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.364671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.364671Z digest=sha256:1977210177a4522837eeb2a8eb4e2f9662b5083baefd2dc374859fd5a0397a12

Observation 1a776c3c-4134-4707-bc44-50b7809548a0 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.455575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.455575Z digest=sha256:cb717e3e79df1728cc236fab4c3aa9762787701b979b7ab1f71a0247c7f09517

Observation 3607d390-bcba-4434-aa0f-3a24e1420cde · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis HybridFlow: A Flexible and Efficient RLHF Framework

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.547506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.547506Z digest=sha256:ad14673c8ecef4727fec98c13288a3f865ff24e06d3be77fabfa1d16fc9c0d2a

Observation 4a3f3a05-f5fc-4a3f-a88d-7f220ec9a4a6 · outbound

This paper cites Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.632301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.632301Z digest=sha256:43831936fda8bb0430a8d27e6a5030eb50df3323f95103e8eeba9727be1cac10

Observation 814b1d37-28b1-4257-8d38-56b0aa65d387 · outbound

This paper cites Some rank order tests which are most powerful against specific parametric alternatives.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Some rank order tests which are most powerful against specific parametric alternatives

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:50.605631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T11:35:46.730566Z digest=sha256:95e1d78ea8e684e47f505733aa6aab9c90341e23442fb144dde9a2b0633ca422

Observation c7cd2c15-bf9d-4f7e-a6e5-441590bccb17 · outbound

This paper cites Dart-math: Difficulty-aware rejection tuning for mathematical problem-solving.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Dart-math: Difficulty-aware rejection tuning for mathematical problem-solving

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.817678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.817678Z digest=sha256:cedc943ce2cf45d3e3af362982c1ff78bfaf9a88e5cf7214d341fd0b98819493

Observation a903f51d-b1df-4c2a-b767-3e18cd230981 · outbound

This paper cites Measuring multimodal mathematical reasoning with math-vision dataset.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Measuring multimodal mathematical reasoning with math-vision dataset

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:50.271280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T11:35:46.908919Z digest=sha256:226fd39c512d219dc9f4965c39cb7bb0367ca2eaad9ba3fa36517052b3ac2a97

Observation db2b616b-3222-4f49-9bcc-45395e56de21 · outbound

This paper cites Agri-LLaVA: Knowledge-Infused Large Multimodal Assistant on Agricultural Pests and Diseases.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Agri-LLaVA: Knowledge-Infused Large Multimodal Assistant on Agricultural Pests and Diseases

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:46.989091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:46.989091Z digest=sha256:04f025d4c9d5e0c0e1f7ef64a7a04b59e567927e55e21f061ce7516faac192dd

Observation 04625a6a-c12b-4741-a4d4-b573b244768e · outbound

This paper cites SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.087104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.087104Z digest=sha256:3f34a141c9076533b82dab2cd3bdeb8f56025e3df8063f618047e40000febc67

Observation 1debfe81-31f3-4173-9bf5-783109667a66 · outbound

This paper cites QuRating: Selecting High-Quality Data for Training Language Models.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis QuRating: Selecting High-Quality Data for Training Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.179506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.179506Z digest=sha256:87374a310379a7e6c53c6c18b0526d52d1fd39a6fe123b0cf6897095b08d5551

Observation 1fd56d91-44f8-4727-9c8f-17c273d6c01f · outbound

This paper cites LESS: Selecting Influential Data for Targeted Instruction Tuning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.274238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.274238Z digest=sha256:27d0ea1ca0c586ee9456187bce7b12fa180ffb36c3b9d80188d12ebc976564bf

Observation e8e5cbd9-5d77-44bd-99a1-e8527c71921b · outbound

This paper cites WizardLM: Empowering large pre-trained language models to follow complex instructions.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis WizardLM: Empowering large pre-trained language models to follow complex instructions

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.349754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.349754Z digest=sha256:e5129e98b90c01e61b142fa351e1b8610d8a9202743d43708fe79eb8747fd172

Observation 48ce8203-2fb1-4340-8e8a-e8328f62ec26 · outbound

This paper cites Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.442353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.442353Z digest=sha256:86ee33c16071e5d297afc776d04a0c2f30cd3ccea1fee49785f3a2fe538e4315

Observation 9945eb6f-2e2a-44af-a382-649b997bffdb · outbound

This paper cites R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.538238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.538238Z digest=sha256:ed945e20c26d23f7e3013978c11e469f61df3793c448c78adddd29598d5dde80

Observation e11d0494-ee93-48e0-b805-ec5b346c0859 · outbound

This paper cites Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.633710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.633710Z digest=sha256:8c832c4a798d3dc02e8b7f6f36c46f9d62c397ea610e12944f5b1543ef86ab4b

Observation 4d6ed220-c3b6-406f-a283-3abb183608da · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.704809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.704809Z digest=sha256:49ab67d3b31ecc845077dafe87c0221a8e8f5055e8c996099f0be8c3551191dc

Observation 5816efab-176e-4812-b4af-4645f840aa09 · outbound

This paper cites VAPO: Efficient and Reliable Reinforcement Learning for Advanced Reasoning Tasks.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis VAPO: Efficient and Reliable Reinforcement Learning for Advanced Reasoning Tasks

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.789603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.789603Z digest=sha256:b08e6443391b76de8179adcfbb2fd704eedd473aabbfd1fcb99016e2e1406227

Observation 02528de9-975f-4329-9a32-8d8299e7940b · outbound

This paper cites SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.858390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.858390Z digest=sha256:9d063913390eeaab543d03367845514f67030b9f01bf1197c57804f5c57ef91c

Observation 3cdf3911-6cfc-49e3-9a6e-db665d108f64 · outbound

This paper cites R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.951687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.951687Z digest=sha256:1dd1b68f08e8079e43a0f174461570c66220b0d804d4fa9e351d02f6d7827a32

Observation 2b492e3b-9148-4658-aeac-c5b03ebb118c · outbound

This paper cites Mathverse: Does your multi-modal llm truly see the diagrams in visual math problems? In European Conference on Computer Vision, pp.\ 169--186.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Mathverse: Does your multi-modal llm truly see the diagrams in visual math problems? In European Conference on Computer Vision, pp.\ 169--186

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:48.064298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:48.064298Z digest=sha256:d68d510e1217c1b0f2f465593b9f7eaf270ee51146c47f369f1d10ffef57258e

Observation 85b0c8d5-e330-4b54-b0ee-0a09458f3881 · outbound

This paper cites Mavis: Mathematical visual instruction tuning.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Mavis: Mathematical visual instruction tuning

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:49.943274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T11:35:48.156999Z digest=sha256:209e664b142014427bf4980459e4b4871391c7b95cd2906f43392b20f53996f0

Observation 0b217a9b-0c96-425e-be90-cf67935278db · outbound

This paper cites Easyr1: An efficient, scalable, multi-modality rl training framework.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Easyr1: An efficient, scalable, multi-modality rl training framework

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:48.225424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:48.225424Z digest=sha256:9432428308fd116a84549e10622cfb65453b79dd82a8cdda70153e9ca4129e5b

Observation 6efe3762-95b1-408a-bc3b-ca878f1a1378 · outbound

This paper cites Lima: Less is more for alignment.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Lima: Less is more for alignment

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:48.312203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:48.312203Z digest=sha256:4c326778b4a96a09801f66ec17beff8c3b1cad2b1778edf8cde6874b3faf6216

Observation 9f7571fe-8f07-48d5-a41e-884b3b99e191 · outbound

This paper cites NavGPT-2: Unleashing Navigational Reasoning Capability for Large Vision-Language Models.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis NavGPT-2: Unleashing Navigational Reasoning Capability for Large Vision-Language Models

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:48.404458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:48.404458Z digest=sha256:1302632394f90b8623773f497a31dfb4ca39dd7de732a8b8a9663f920d1c6b19

Observation a78461e9-b9f9-4af6-bfdf-124de97dc7f6 · outbound

This paper cites Anyprefer: An Agentic Framework for Preference Data Synthesis.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Anyprefer: An Agentic Framework for Preference Data Synthesis

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:48.503722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:48.503722Z digest=sha256:c7df81822cb251b4c58c5c7d20599f71c3b1ba26ce2d27543e24515ab728d6e3

Observation 78b6cf58-0845-45a2-95e5-16bb60f350e7 · outbound

This paper cites Dynamath: A dynamic visual benchmark for evaluating mathematical reasoning robustness of vision language models, 2024.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Dynamath: A dynamic visual benchmark for evaluating mathematical reasoning robustness of vision language models, 2024

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:35:49.749976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T11:35:48.604808Z digest=sha256:52d6675662159e10cf30a67b58958130f47946fd0a7204e7ae0631b5551a5814

Observation f1303788-f8f7-4a4d-a01a-64db254eba24 · outbound

This paper cites write newline.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis write newline

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:48.656605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:48.656605Z digest=sha256:82eca94a4b0c676d493a765be4121675243540c4ac2858cc5eb73c42a3e91709

Pith citing papers

Observation 3e6ecd59-cd6a-4989-995d-28ea935f3b0c · inbound

SCALER:Synthetic Scalable Adaptive Learning Environment for Reasoning cites this paper.

SCALER:Synthetic Scalable Adaptive Learning Environment for Reasoning SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:28:05.710085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-16T16:24:04.132572Z digest=sha256:24a5e36181595ddd154852b8ba2dd73b7384134223314b6cbc37e16ed72165f8

Observation a65cee66-e7eb-462a-a1eb-7559aa9d25da · inbound

GaMMA: Towards Joint Global-Temporal Music Understanding in Large Multimodal Models cites this paper.

GaMMA: Towards Joint Global-Temporal Music Understanding in Large Multimodal Models SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:47:17.195171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-09T19:12:28.290483Z digest=sha256:b93fb6e73e738b9b794cfbee3597595c1e02e5cb9d4fbcfba4b5a937798f0fa1

Observation 6f0c6a90-80f6-4654-a8b5-00947dc73b4e · inbound

Generate, Filter, Control, Replay: A Comprehensive Survey of Rollout Strategies for LLM Reinforcement Learning cites this paper.

Generate, Filter, Control, Replay: A Comprehensive Survey of Rollout Strategies for LLM Reinforcement Learning SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis

Reference 142

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:15:49.390205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-05-10T19:15:27.406778Z digest=sha256:657ee895804f613e79d8b383d05d36cd30395c89046a2ee2407bb0630bc8b152

Observation d658f3bf-0e1d-456a-b0ec-cbf42d7578a0 · inbound

Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning cites this paper.

Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T11:47:27.883967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:47:27.883967Z digest=sha256:9f471769fcbf697d706db62ae101b339e65d875c2b4034aa39c7f69e764a16ea