Pith. sign in

Paper Citation Record · LEDGER

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization

As of 15 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 1 inbound Pith citation observation for arXiv:2505.23331.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.23331 v2

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:52:44.525510Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-15T12:50:13.764159Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-15T12:50:37.089876Z

Reference resolution

28 of 28 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 73815b03-b814-48bb-a670-c0c9faef498e · outbound

This paper cites Training Diffusion Models with Reinforcement Learning.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Training Diffusion Models with Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:40.638024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:40.638024Z digest=sha256:7aa902a3b8717006027ee479406e7f254af9f1b054308ec35e36cc41037044aa

Observation 334b404d-992e-4452-8b8e-260ab5c5120c · outbound

This paper cites Language Models are Few-Shot Learners.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Language Models are Few-Shot Learners

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:40.747777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:40.747777Z digest=sha256:10bde3c9708d4f331fad10ececd6bf7c68903953d77c5dcc1613a6f891227903

Observation 52d7ffff-1436-4f80-bd33-e7b061866528 · outbound

This paper cites Large scale visual recognition.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Large scale visual recognition

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:46.101052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T12:52:40.913201Z digest=sha256:ece2e11ac26817d0dabcc4f42576657bcff6711423033b148d0e29d2041423a7

Observation 5991cb67-d15b-46e1-80bf-7168df972419 · outbound

This paper cites Scaling rectified flow transformers for high-resolution image synthesis.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Scaling rectified flow transformers for high-resolution image synthesis

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:45.859940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T12:52:41.164927Z digest=sha256:394c958957b1d7f3bb09579c3d83e4d68eb8c171148d82de5e2a9ee271d22f93

Observation 2ed005df-04d0-48d3-a3bf-ad74248e098f · outbound

This paper cites Generative Adversarial Networks.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Generative Adversarial Networks

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:41.243685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:41.243685Z digest=sha256:9d411c4ca5f16eec90520de7233891e71874ec9b10d5448760212429f77db31d

Observation ab1063cd-2ba3-4608-b879-3c1a9e9ef0c5 · outbound

This paper cites Infinity: Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Infinity: Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:41.349047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:41.349047Z digest=sha256:ab1bd7b14b7469a61de823f4878e844b23eff12a77675630f3d1e4eba11c5d0c

Observation 5477b8c5-cc75-40cb-a2a0-fb6f6172dcbd · outbound

This paper cites Classifier-Free Diffusion Guidance.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Classifier-Free Diffusion Guidance

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:41.478481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:41.478481Z digest=sha256:aeb225a21a804f96532656fdad8e08b512aa4feac3d4e17ff5678589ecccebc2

Observation bd5c164b-df26-4f3e-916c-8be800341f7a · outbound

This paper cites Auto-Encoding Variational Bayes.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Auto-Encoding Variational Bayes

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:41.621301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:41.621301Z digest=sha256:c51e86312aaf0e44608b2f730e6e84d2731f6b798c5b0da4a4e15fe58617dae8

Observation 51859940-21cf-42c4-8067-b34c9cafe27f · outbound

This paper cites Aligning Text-to-Image Models using Human Feedback.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Aligning Text-to-Image Models using Human Feedback

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:41.738949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:41.738949Z digest=sha256:d52c21f14f21030be449c77750f26d823184be3fa08610fc6873597e83643293

Observation 703e1f93-8a83-4f89-9b77-7105a1b9e159 · outbound

This paper cites an unresolved cited work.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:41.864074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:41.864074Z digest=sha256:67a35b8352d711189bd6bef9d6f401ab3e9c18650e84ee67eaa89c9c956e83fc

Observation c5992993-0260-40c2-9dc9-4be5b02b0dc3 · outbound

This paper cites Flow straight and fast: Learning to generate and transfer data with rectified flow.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Flow straight and fast: Learning to generate and transfer data with rectified flow

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:45.627084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T12:52:41.936495Z digest=sha256:986af0b969a83a7988a8264e6c2765a10c05cd1c339ff6011a2ac30b85046576

Observation e3998470-5cf0-4671-800e-3d41743b7aad · outbound

This paper cites Training language models to follow instructions with human feedback.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Training language models to follow instructions with human feedback

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:42.078302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:42.078302Z digest=sha256:53d26edf824581c0e16f41502f08666a8c057a4ae6c3fe837d88e5b3906db38d

Observation 046d19a2-5ff6-4aba-834b-9913e021f5f3 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:42.261911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:42.261911Z digest=sha256:a3de53f69f16818942c9937a7c1bec955726c08cabebc3cac146bb6629ccbb8f

Observation a630f947-8778-49b6-b006-c4032dd31a2f · outbound

This paper cites W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:42.415543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:42.415543Z digest=sha256:906fdfbebf653e27d0ec4ff7e75d6375564008eaed06f793a52b5a9cf3b14278

Observation 0b42df90-d4f4-4cfd-8661-84c51b32b6e1 · outbound

This paper cites Zero-Shot Text-to-Image Generation.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Zero-Shot Text-to-Image Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:42.565243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:42.565243Z digest=sha256:5edfe0a641acbbd9c75c7ab7d14bebcd9dd1f0a8f64ca5dc9d7f5882a146cf36

Observation 780b7e3e-a93a-4b50-aa7c-edc000dce459 · outbound

This paper cites High-Resolution Image Synthesis with Latent Diffusion Models.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization High-Resolution Image Synthesis with Latent Diffusion Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:42.747054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:42.747054Z digest=sha256:e933e74469566ea0cf91a298c7aa6787baae9939a21f277083d89f3033b5cbc8

Observation 44166280-3365-4342-b26b-8b78cd705ea1 · outbound

This paper cites Laion-5b: An open large-scale dataset for training next generation image-text models.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Laion-5b: An open large-scale dataset for training next generation image-text models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:42.903491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:42.903491Z digest=sha256:597b7f667e1c0823390ba968f6e8c9a5789c3e1b00fcf4c38bb7171d508ae9d2

Observation b271deda-3b76-42c6-b890-2e06f44c2710 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Proximal Policy Optimization Algorithms

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:43.081066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:43.081066Z digest=sha256:e86257668b9bc42d9a6237c807ea2892ab7a790cbde8b645bd9065ebc81021ae

Observation 38dd85b0-368d-4f72-a15c-01c90780abd0 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:43.292943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:43.292943Z digest=sha256:580a1de87423c7c5fd30b9ca7e8e92f10f64e3624a17f6e4c4f3dc608a4509b3

Observation 579793ff-e0a3-43cc-b4c9-3f9d5d71bf7b · outbound

This paper cites Deep Unsupervised Learning using Nonequilibrium Thermodynamics.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Deep Unsupervised Learning using Nonequilibrium Thermodynamics

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:43.377350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:43.377350Z digest=sha256:e52acb1a38496657887868e980ef1cf5147cd5d7251ebbe91eaa78b8dbbb00b7

Observation 9203ac51-8d84-457a-a6b8-823261cb1e37 · outbound

This paper cites Visual autoregressive modeling: Scalable image generation via next-scale prediction.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Visual autoregressive modeling: Scalable image generation via next-scale prediction

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:45.322555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T12:52:43.504353Z digest=sha256:2f14c8bd211f1485a6dec7c145340ea3f7b407cf927d9ef6de1dcbe4ad73816c

Observation f39b1de4-0e02-450d-9b88-391ac128758d · outbound

This paper cites Neural Discrete Representation Learning.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Neural Discrete Representation Learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:43.667477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:43.667477Z digest=sha256:3710dce5b1ded3e0f4fc5ec4f1377231444aa30bdf6af7c5662280a2bf9b87a2

Observation d0c26fd1-8aa3-41b9-a381-e86f887c6072 · outbound

This paper cites Switti: Designing Scale-Wise Transformers for Text-to-Image Synthesis.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Switti: Designing Scale-Wise Transformers for Text-to-Image Synthesis

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:43.800620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:43.800620Z digest=sha256:2a162d80c9109fe90581352a4a06bea980b9dd6eb6a4b1ba22d2a96b1cbf72f2

Observation f2e27a8e-fd7b-4ad4-9e4b-39b97adddefe · outbound

This paper cites How to train state-of-the-art models using torchvision’s latest primitives.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization How to train state-of-the-art models using torchvision’s latest primitives

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:45.092645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T12:52:43.943917Z digest=sha256:19aade81d4ad25f68c1ab8a5703ac2dd4be8c4f6babeab628f38477b5602f079

Observation 8468dbaf-0c2b-4632-95a6-54c98420e1a6 · outbound

This paper cites SimpleAR: Pushing the Frontier of Autoregressive Visual Generation through Pretraining, SFT, and RL.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization SimpleAR: Pushing the Frontier of Autoregressive Visual Generation through Pretraining, SFT, and RL

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:44.091056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:44.091056Z digest=sha256:afb636355e6aff64ceb4471ef19a7061b879c0fc4bab962346304d149c8e047f

Observation 3f74f70c-104d-4e19-a9e4-4b4bd4b81cac · outbound

This paper cites DanceGRPO: Unleashing GRPO on Visual Generation.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization DanceGRPO: Unleashing GRPO on Visual Generation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:44.218213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:44.218213Z digest=sha256:c92fab42d366f78da58d87ea5d56f7faf7c5fcf638f816da25a1b397f9fb8c0d

Observation 0791ead8-c531-4512-9ea2-917af645cbec · outbound

This paper cites Using Human Feedback to Fine-tune Diffusion Models without Any Reward Model.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Using Human Feedback to Fine-tune Diffusion Models without Any Reward Model

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:44.361394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:44.361394Z digest=sha256:fc03b703cec8f79e36cbc2a6961282a027cf242535efe0b1a9ebeff952a1985f

Observation cb5f053f-4694-4d01-aa53-332260fdf51d · outbound

This paper cites write newline.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization write newline

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:44.525510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:44.525510Z digest=sha256:e7c48d56c75edbee18c26414bb113f8f9582c12a80bc7c3ec3b2974f64085b42

Pith citing papers

Observation 7d5451d7-35be-4d72-8e06-480e1e4701e8 · inbound

From Broad Exploration to Stable Synthesis: Entropy-Guided Optimization for Autoregressive Image Generation cites this paper.

From Broad Exploration to Stable Synthesis: Entropy-Guided Optimization for Autoregressive Image Generation Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:50:37.096138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-15T12:50:13.764159Z digest=sha256:f4e2297bdc577271f966c7a78d9ab5a3c5e070322e5abf8e6ad52d00deae65f8