Pith. sign in

Paper Citation Record · LEDGER

RewardDance: Reward Scaling in Visual Generation

As of 18 August 2026, this Paper Citation Record lists 72 of 72 outbound references and 30 inbound Pith citation observations for arXiv:2509.08826.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.08826 v1

Coverage vector

measured 72 of 72 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T20:09:01.822336Z

measured 102 of 102 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 30 of 30 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T21:47:20.112331Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:39:45.924030Z

Reference resolution

72 of 72 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved72
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 86905ea5-c022-43bb-8b50-c37819f87f6f · outbound

This paper cites Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models.

RewardDance: Reward Scaling in Visual Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:55.396895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:55.396895Z digest=sha256:4e568bca6a6bf050cd691ab7ff2bba72a110527f71885627720b23a59c78d28e

Observation ea855739-cf41-40b2-aa1e-1d10a6592d1f · outbound

This paper cites Improving image generation with better captions.Computer Science.

RewardDance: Reward Scaling in Visual Generation Improving image generation with better captions.Computer Science

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:55.529478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:55.529478Z digest=sha256:81a46f4527b2b5b215520b1182cd4cf7bc4c5db7014a9da2b05b26a4ad6ae29d

Observation 132c6830-d12c-43f7-9156-dddaf4ed4d40 · outbound

This paper cites Training Diffusion Models with Reinforcement Learning.

RewardDance: Reward Scaling in Visual Generation Training Diffusion Models with Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:55.650307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:55.650307Z digest=sha256:08d8be5ee1ea682cf796140a79fc55e5ceccb68d43e88a80a1b7de09e9c48e35

Observation 7efa567e-ef00-4989-a064-45e848d36bbf · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

RewardDance: Reward Scaling in Visual Generation Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:55.746553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:55.746553Z digest=sha256:40ba9a236acde9dc1e24b6e8e5424b31e1330554bd181f9bb8a2747d6ff1edf9

Observation 3abdd8b0-0899-4016-a961-6a5bcadc57df · outbound

This paper cites Rank analysis of incomplete block designs: I.

RewardDance: Reward Scaling in Visual Generation Rank analysis of incomplete block designs: I

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:55.865292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:55.865292Z digest=sha256:28ef9242b6cf8b7a02e51995a6e892c851844940395e92f5ac70153c50bc501a

Observation c3a37bb2-8a9b-4395-8d20-11ff4747b2f4 · outbound

This paper cites Video generation models as world simulators.OpenAI Blog, 1(8):1, 2024.

RewardDance: Reward Scaling in Visual Generation Video generation models as world simulators.OpenAI Blog, 1(8):1, 2024

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:55.944987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:55.944987Z digest=sha256:d3d3964f35fc341e025c22854c151ba744744dbaeab501ed1c4b5953f8d038c8

Observation 0f4bcd2d-bb19-4f13-b4b8-18379f110cfc · outbound

This paper cites PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis.

RewardDance: Reward Scaling in Visual Generation PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.030031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.030031Z digest=sha256:960eab273ffe6d564c4061a5bb54c7d289dc9dc3a52941256b614b24a9ea472c

Observation e91579b9-d6e5-4c0c-9e9d-818bfe04fc4c · outbound

This paper cites Control-a-video: Controllable text-to-video generation with diffusion models.CoRR, 2023.

RewardDance: Reward Scaling in Visual Generation Control-a-video: Controllable text-to-video generation with diffusion models.CoRR, 2023

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.108706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.108706Z digest=sha256:588441805fc46f19b42b3ef2670539ecb81d712d45c8d289d0cabb3c827c8581

Observation b272c883-8db7-4937-86eb-367ad3b86f30 · outbound

This paper cites The Accuracy Paradox in RLHF: When Better Reward Models Don't Yield Better Language Models.

RewardDance: Reward Scaling in Visual Generation The Accuracy Paradox in RLHF: When Better Reward Models Don't Yield Better Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.202728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.202728Z digest=sha256:f30f1dcb0ef83fb096619ba443cf173394c7e2362d80fb9488b0bff85a98ed6a

Observation 486d1163-e00a-4ea2-9d8a-5cf8cd809e93 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

RewardDance: Reward Scaling in Visual Generation Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.309210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.309210Z digest=sha256:677fcf86eb73b8b5d6c310b917de3578bb2921bf7f674615a572c3b70d0fee9c

Observation e2b42bf0-44f1-4691-ad5a-fc91b0429d1a · outbound

This paper cites RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment.

RewardDance: Reward Scaling in Visual Generation RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.415726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.415726Z digest=sha256:b38d2596e4ce771552df5b63d9384b2489e79207e9844a8ba71027d0d7491727

Observation e8c40eb9-35e7-4a24-aa42-90d44b3ff2ab · outbound

This paper cites Scaling rectified flow transformers for high-resolution image synthesis.

RewardDance: Reward Scaling in Visual Generation Scaling rectified flow transformers for high-resolution image synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.497198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.497198Z digest=sha256:d01b0fd7c14b8bd3b2be8a1b34654f181c309bb5c9a1c2a41f067df8f8336f73

Observation 1382ada2-b0c5-4221-8d9c-7225d5b6876a · outbound

This paper cites Reinforcement learning for fine-tuning text-to-image diffusion mod- els.

RewardDance: Reward Scaling in Visual Generation Reinforcement learning for fine-tuning text-to-image diffusion mod- els

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.572535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.572535Z digest=sha256:1c8e75b735370a1fdbfc74d48a11455c74018f318a20f4d482f6b020a83446ab

Observation c4bbc6c6-5620-430b-94b2-27eedc40a862 · outbound

This paper cites Seedream 3.0 Technical Report.

RewardDance: Reward Scaling in Visual Generation Seedream 3.0 Technical Report

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.678650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.678650Z digest=sha256:093003a4adc838171d38ccfa6bd5533eef82c38b4f779a6133f43ac2280b760b

Observation e1ed9b66-e8ca-4b34-9f27-5433927eb701 · outbound

This paper cites Seedance 1.0: Exploring the Boundaries of Video Generation Models.

RewardDance: Reward Scaling in Visual Generation Seedance 1.0: Exploring the Boundaries of Video Generation Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.802057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.802057Z digest=sha256:fafb686a99c06efd151bd7c608cda4f9dcbadeff520a7f43e58296873d45f1bf

Observation da4d94ab-2adb-47a2-8c1f-a6dcf2fc3c85 · outbound

This paper cites Seedream 2.0: A Native Chinese-English Bilingual Image Generation Foundation Model.

RewardDance: Reward Scaling in Visual Generation Seedream 2.0: A Native Chinese-English Bilingual Image Generation Foundation Model

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.884737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.884737Z digest=sha256:97046a5cfbac7942c75dd648a6b844567ea647047e817ad925a33b3e12cb34c4

Observation d6e63974-c7c7-459b-a90b-de923f36774b · outbound

This paper cites Veo.https://deepmind.google/models/veo/, 2025.

RewardDance: Reward Scaling in Visual Generation Veo.https://deepmind.google/models/veo/, 2025

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.953258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.953258Z digest=sha256:1b088ca86302ca647411474bce53e5c96e845103ec4d5bbaf44799bdc7aa127f

Observation 301b6b30-f1d6-4a81-8021-787bedd2734f · outbound

This paper cites Multi-Reward as Condition for Instruction-based Image Editing.

RewardDance: Reward Scaling in Visual Generation Multi-Reward as Condition for Instruction-based Image Editing

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.041134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.041134Z digest=sha256:45e5efd0f029ac564deea1f925c3540e5378a510e56236d55c05a2039da1cfeb

Observation 9093d8ed-5bfb-4fb1-b9d5-4663fcf14436 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

RewardDance: Reward Scaling in Visual Generation AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.137515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.137515Z digest=sha256:fd1d8476dd8532f6589663cf40734ca287b9bea9d21bc4573cd34f9af23735b1

Observation 32d7c0f9-754f-47c4-b243-f3d6527613a3 · outbound

This paper cites A simple and effective reinforcement learning method for text-to-image diffusion fine-tuning.

RewardDance: Reward Scaling in Visual Generation A simple and effective reinforcement learning method for text-to-image diffusion fine-tuning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.228352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.228352Z digest=sha256:2f6bb04385e02b9d2d02f4ca7d086e703104bf0c95235326d86ec671d276987c

Observation 31caed26-a32f-4c1d-a531-0594ffda0736 · outbound

This paper cites Denoising diffusion probabilistic models.

RewardDance: Reward Scaling in Visual Generation Denoising diffusion probabilistic models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.297593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.297593Z digest=sha256:1854363bd0bea437a900b4aa8ccf446f5cb44af92ab5c205ca976d3f2efba023

Observation 79c2c946-32a1-47a5-a3c1-8ee6c77af671 · outbound

This paper cites Ideogram.https://about.ideogram.ai/1.0., 2024.

RewardDance: Reward Scaling in Visual Generation Ideogram.https://about.ideogram.ai/1.0., 2024

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.397897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.397897Z digest=sha256:9c2d93a633d7d3b62cd01e85d607428517704a98a005abb7e7d68e523a1f4ac3

Observation b88cd1fa-c9b6-41d9-aa71-bf3d89cbaa92 · outbound

This paper cites Pick-a-pic: An open dataset of user preferences for text-to-image generation.Advancesin neural information processing systems, 36:36652–36663, 2023.

RewardDance: Reward Scaling in Visual Generation Pick-a-pic: An open dataset of user preferences for text-to-image generation.Advancesin neural information processing systems, 36:36652–36663, 2023

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.483267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.483267Z digest=sha256:c48bb5b0ee8f2169fac4cc5a8aabc3c7808f9cf193e4483ec70723ab82484d5c

Observation 200e654a-6de4-4478-86be-48c758073b93 · outbound

This paper cites klingai.https://app.klingai.com/cn/, 2025.

RewardDance: Reward Scaling in Visual Generation klingai.https://app.klingai.com/cn/, 2025

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.585056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.585056Z digest=sha256:9f16ae68d9704b343d55866e1b29bca189e9ae2b57ccb8848733c97dc773fb41

Observation 836d2260-1b27-41f8-82c9-82b7380589d3 · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

RewardDance: Reward Scaling in Visual Generation HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.666163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.666163Z digest=sha256:375c7fe0fe62b1b2cc8da603a9bdcc3075bcdfce99e65bddfaa8f56e2c6a0589

Observation af1ebe2a-2734-4c43-a81c-231af56b104d · outbound

This paper cites Flux: Official inference repository for flux.1 models, 2024.

RewardDance: Reward Scaling in Visual Generation Flux: Official inference repository for flux.1 models, 2024

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.744104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.744104Z digest=sha256:0e9dc66023397a5b583d107297ceb9983f4580a0829c0e5655548b86aa6826f4

Observation 94a153bb-e4c9-46f1-bbac-6a8c7f108c32 · outbound

This paper cites Flux.https://github.com/black-forest-labs/flux, 2024.

RewardDance: Reward Scaling in Visual Generation Flux.https://github.com/black-forest-labs/flux, 2024

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.838549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.838549Z digest=sha256:3addf0373ef57fee4bd6c7745c021c51114f768d4a9067dd7eaba867db7be641

Observation 8e4c83b1-f2ed-44cd-91a9-df9a6b582506 · outbound

This paper cites Controlnet++: Improving conditional controls with efficient consistency feedback: Project page: liming-ai.

RewardDance: Reward Scaling in Visual Generation Controlnet++: Improving conditional controls with efficient consistency feedback: Project page: liming-ai

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.926797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.926797Z digest=sha256:2ce1d1b1dc54f54f3e122b680387b9cad7b17b1fdf1200285c7e7a24e0119ffd

Observation fba2b18b-871c-4917-8670-bbc8e802f8ee · outbound

This paper cites SuperEdit: Rectifying and Facilitating Supervision for Instruction-Based Image Editing.

RewardDance: Reward Scaling in Visual Generation SuperEdit: Rectifying and Facilitating Supervision for Instruction-Based Image Editing

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.003125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.003125Z digest=sha256:833c2831e53742f66ca3107b96a4d590fea1ebb93904c58a80c2a8dd17e6bf26

Observation 980525a9-64b4-4034-bc9c-1b258ed6ce7a · outbound

This paper cites Exploring How Generative MLLMs Perceive More Than CLIP with the Same Vision Encoder.

RewardDance: Reward Scaling in Visual Generation Exploring How Generative MLLMs Perceive More Than CLIP with the Same Vision Encoder

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.110097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.110097Z digest=sha256:8f7ed92acd4732afb90fac92cef6fd595f452041035a962a38216a8a64a16c6f

Observation 40a2ebf7-fa29-4820-92fe-524f2545a508 · outbound

This paper cites An inverse scaling law for clip training.Advancesin Neural Information Processing Systems, 36:49068–49087, 2023.

RewardDance: Reward Scaling in Visual Generation An inverse scaling law for clip training.Advancesin Neural Information Processing Systems, 36:49068–49087, 2023

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.185503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.185503Z digest=sha256:da6edd410907adec508484291899e41e9724d6c56030b1247bc93e1137aa9ee4

Observation 9e56985d-a38c-4a4a-bc8a-387d8247ca8e · outbound

This paper cites Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation.

RewardDance: Reward Scaling in Visual Generation Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.270191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.270191Z digest=sha256:8bf9fe2631d0fefb432080112829e6b813f130e7dcbeb4d1b2e3a209c37244ed

Observation dbb474ed-47dd-4165-879f-85a61565a7b8 · outbound

This paper cites Flow Matching for Generative Modeling.

RewardDance: Reward Scaling in Visual Generation Flow Matching for Generative Modeling

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.373978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.373978Z digest=sha256:f0c02d51a0d4119f2a83ed65bd6961f8a3425c050a8723547414502c784d9954

Observation 8fa26720-2d27-44f9-ab9f-4fc28f386f7a · outbound

This paper cites Flow-GRPO: Training Flow Matching Models via Online RL.

RewardDance: Reward Scaling in Visual Generation Flow-GRPO: Training Flow Matching Models via Online RL

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.469893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.469893Z digest=sha256:f8b76526d85504dd59efdba2c3c46a3b38cc1731f12d80292564733ce9fb1c0a

Observation dcf4dd50-8905-4504-afa9-b61b94c65d46 · outbound

This paper cites Improving Video Generation with Human Feedback.

RewardDance: Reward Scaling in Visual Generation Improving Video Generation with Human Feedback

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.572097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.572097Z digest=sha256:cff5aadea14cee3c09ebd4d7a30e4e2d807b241c4940ca46194deaa8bacd2531

Observation b90777e3-0f54-4c9f-b360-3e758cd57151 · outbound

This paper cites Inference-time scaling for generalist reward modeling.arXiv preprint arXiv:2504.02495, 2025.

RewardDance: Reward Scaling in Visual Generation Inference-time scaling for generalist reward modeling.arXiv preprint arXiv:2504.02495, 2025

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.678868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.678868Z digest=sha256:d7013c7f075477de4bd758e8c9a668d3fc9dfc5bf1a758465a47d421cabd1b63

Observation d26e027a-12f0-4535-b7df-455007611c6f · outbound

This paper cites lumalabs.https://lumalabs.ai/, 2024.

RewardDance: Reward Scaling in Visual Generation lumalabs.https://lumalabs.ai/, 2024

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.744034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.744034Z digest=sha256:dcf74fb32090a647c073a360add363d59d36ef57409e28ef3c78e3992fbf6140

Observation e5445f7e-d9e2-48ea-89bf-662cbd1b64c2 · outbound

This paper cites Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model.

RewardDance: Reward Scaling in Visual Generation Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.832391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.832391Z digest=sha256:428eed64e8a401461149b5f5de79383c7badffc21e6bf8383ff71ae538512a5b

Observation 873f0fa2-ce1c-463e-81ea-5695252c4ed2 · outbound

This paper cites Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps.

RewardDance: Reward Scaling in Visual Generation Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.890930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.890930Z digest=sha256:8ddbd140d97fed41f3a512dc2eb4f54bf86426d68865ba2fbf104ec050669cfc

Observation 3930fdb3-dd8b-4d9a-abf8-1030b39d043a · outbound

This paper cites HPSv3: Towards Wide-Spectrum Human Preference Score.

RewardDance: Reward Scaling in Visual Generation HPSv3: Towards Wide-Spectrum Human Preference Score

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.995776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.995776Z digest=sha256:83d5d09d0bdf6fa0ce1ca006932a7b10b9be91034f089eb4321f650124cd641a

Observation 69192038-94c5-445a-bf15-8916ef183f2f · outbound

This paper cites midjourney.https://www.midjourney.com/home, 2024.

RewardDance: Reward Scaling in Visual Generation midjourney.https://www.midjourney.com/home, 2024

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.094853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.094853Z digest=sha256:3d3d5508e446e8c02215d5315aacf148609b0bf4ce73d9d1a28e5438d4d4b61a

Observation d0664f12-d885-48ec-a8de-318928b50deb · outbound

This paper cites Inference-time text-to-video alignment with diffusion latent beam search.arXiv preprint arXiv:2501.19252, 2025.

RewardDance: Reward Scaling in Visual Generation Inference-time text-to-video alignment with diffusion latent beam search.arXiv preprint arXiv:2501.19252, 2025

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.185710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.185710Z digest=sha256:9658c9d0126ca790ebb8464f12a67dd1983885c2bf2c71ca7a0c0285141507d8

Observation 668c5b5a-e40c-4444-9aaf-14c4c7247767 · outbound

This paper cites Training language models to follow instructions with human feedback.

RewardDance: Reward Scaling in Visual Generation Training language models to follow instructions with human feedback

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.294363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.294363Z digest=sha256:c099149e3315834e7b633e3d241d02f31e2e7002a2a183539cf22c8f0fde6647

Observation 866901a6-8d1e-4f11-9f59-ecec90bfcced · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

RewardDance: Reward Scaling in Visual Generation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.377336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.377336Z digest=sha256:93dac3d17b716db9ca0f23aa8ede3c4dfd687b80d27d6ffaa5d236d0bad3252b

Observation 977bd8af-fed2-4656-ba94-2b087858437a · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

RewardDance: Reward Scaling in Visual Generation Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.470265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.470265Z digest=sha256:27282049a5da34f601ec5d5e6cdbf8d16b11be8ecdd3c8f2644991ad7d6eec64

Observation efc721e4-fdb0-4f99-a5ea-fab919a1130e · outbound

This paper cites What makes a reward model a good teacher? an optimization perspective.arXiv preprint arXiv:2503.15477, 2025.

RewardDance: Reward Scaling in Visual Generation What makes a reward model a good teacher? an optimization perspective.arXiv preprint arXiv:2503.15477, 2025

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.549832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.549832Z digest=sha256:bf42b5c88df6f87549cc88507473c5347be3c94c31c7a4c923c3b70c85d7273f

Observation e3b317e8-9a99-42d8-87b3-6274f9ad808e · outbound

This paper cites recraft.https://www.recraft.ai/, 2024.

RewardDance: Reward Scaling in Visual Generation recraft.https://www.recraft.ai/, 2024

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.607627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.607627Z digest=sha256:3ccd29e71f341a493332db052f9815acd3789d76c9311973decc6e095d58b12f

Observation 81fb23ff-d3e2-4f00-99b9-c2f63a0f3a80 · outbound

This paper cites Byteedit: Boost, comply and accelerate generative image editing.

RewardDance: Reward Scaling in Visual Generation Byteedit: Boost, comply and accelerate generative image editing

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.703606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.703606Z digest=sha256:6c90ba1c6799dd5d43cf6596bf691da9b17dd95dc340ab1fa94856fb8f51fd12

Observation 5da2af02-3f3f-409f-baf2-707e228638a6 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

RewardDance: Reward Scaling in Visual Generation High-resolution image synthesis with latent diffusion models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.800075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.800075Z digest=sha256:a0a0ea54ca9b6095564b673625595bee2cc661dd523fd7e7f2a935cadbc39221

Observation f14a6ad6-6c8e-4c6e-bb67-41c035ef8863 · outbound

This paper cites Runway.https://runwayml.com/research/introducing-runway-gen-4, 2025.

RewardDance: Reward Scaling in Visual Generation Runway.https://runwayml.com/research/introducing-runway-gen-4, 2025

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.895610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.895610Z digest=sha256:c5ea2016a51d3bff185218e75be44f038e71db9fca34f90b295896d1fb675588

Observation 064fc98f-5828-48a3-ba5a-292bb297fa5d · outbound

This paper cites Seaweed-7B: Cost-Effective Training of Video Generation Foundation Model.

RewardDance: Reward Scaling in Visual Generation Seaweed-7B: Cost-Effective Training of Video Generation Foundation Model

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.968870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.968870Z digest=sha256:49bf349d1f005879e2cfe2d939f1b34232db4c34f72591b3083e3cc191ea41ca

Observation d208184c-7791-4c8a-a987-fc85114cded9 · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

RewardDance: Reward Scaling in Visual Generation Deep unsupervised learning using nonequilibrium thermodynamics

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.052186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.052186Z digest=sha256:77eec35bc58529d5d8a5c5da832c48ad7bbc43d37478dc3e4d6adbf5f193f830

Observation 83058bb5-07f5-4989-a8b5-0a2f1f9925b6 · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

RewardDance: Reward Scaling in Visual Generation Score-Based Generative Modeling through Stochastic Differential Equations

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.112624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.112624Z digest=sha256:42a44921b472396c5254486becca1ab6dd76ba9ed2531774a1b331cb69b6d6ae

Observation a6df593d-ba6d-4eeb-a27b-ee3828f84c4b · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

RewardDance: Reward Scaling in Visual Generation Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.245755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.245755Z digest=sha256:5b00e47b4dc0d360d81ae82c28be38c0e1411d924a6e6627f5d7b63724e143ab

Observation 4b61a4f2-c478-4e5b-bc2e-49d7706edeef · outbound

This paper cites Diffusion model alignment using direct preference optimization.

RewardDance: Reward Scaling in Visual Generation Diffusion model alignment using direct preference optimization

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.340485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.340485Z digest=sha256:c150d33183598d02fe074891c62cc4956ed5358966c414087a36204a29833603

Observation 02bedef1-7f2a-4ce3-bb51-59c74b8c2387 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

RewardDance: Reward Scaling in Visual Generation Wan: Open and Advanced Large-Scale Video Generative Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.409906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.409906Z digest=sha256:24c551753c635db96ce20196c05ccb3bbda6d37ff11a5bb39019f780889e0899

Observation c6887f61-aa82-4a10-9f0e-16a07b78309b · outbound

This paper cites WorldPM: Scaling Human Preference Modeling.

RewardDance: Reward Scaling in Visual Generation WorldPM: Scaling Human Preference Modeling

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.504202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.504202Z digest=sha256:424adee035efa8be36acd16726b60f22cd1c427a7b08cbed69554784407998fb

Observation de72a69d-f92c-4349-9115-e5f9f698a074 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

RewardDance: Reward Scaling in Visual Generation Emu3: Next-Token Prediction is All You Need

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.588199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.588199Z digest=sha256:6ded1062440af6f833f330fb1ed7fe9b1c0ff0829018462d1764b001332b2cc4

Observation 6efd3f03-2a1b-4d8e-b84f-c6ee7772ab80 · outbound

This paper cites Unified multimodal chain-of-thought reward model through reinforcement fine-tuning.arXiv preprint arXiv:2505.03318, 2025.

RewardDance: Reward Scaling in Visual Generation Unified multimodal chain-of-thought reward model through reinforcement fine-tuning.arXiv preprint arXiv:2505.03318, 2025

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.663870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.663870Z digest=sha256:53bbc2cf4723acac42186d272d8376937b3cf7bd96006b009cadb7ef9a4804b8

Observation efae6b3f-4095-4f72-8914-968501479d83 · outbound

This paper cites Rethinking Reward Model Evaluation: Are We Barking up the Wrong Tree?.

RewardDance: Reward Scaling in Visual Generation Rethinking Reward Model Evaluation: Are We Barking up the Wrong Tree?

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.786327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.786327Z digest=sha256:fbacd29b6932c346efb3967ca3031a9e4ee8612b87fbbbd0062a17c699801ee3

Observation eb924bd3-99c4-4e38-a602-3e7f9fda43e7 · outbound

This paper cites Qwen-Image Technical Report.

RewardDance: Reward Scaling in Visual Generation Qwen-Image Technical Report

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.883913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.883913Z digest=sha256:cdf13ccf859cfb71d382028848c7f8dfb5846b749482e83b4bba78a209ddf12e

Observation 46e3a5f9-eca2-47eb-a390-548ed9974217 · outbound

This paper cites Human Preference Score: Better Aligning Text-to-Image Models with Human Preference.

RewardDance: Reward Scaling in Visual Generation Human Preference Score: Better Aligning Text-to-Image Models with Human Preference

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.976315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.976315Z digest=sha256:f52bfe0b8387b545b071204e1fc5e19884498ec3bdfc403f288ff842b642d39d

Observation 2d92b21d-ec74-43a2-b9b2-9b58ff1aa849 · outbound

This paper cites Human preference score: Better aligning text-to-image models with human preference.

RewardDance: Reward Scaling in Visual Generation Human preference score: Better aligning text-to-image models with human preference

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.048080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.048080Z digest=sha256:4ab794bbb770e04a5d3da6884d7b0e4a613258c1d5218076aa41d6faae86a62d

Observation aa80b625-a1c6-4678-a710-15e30fb59008 · outbound

This paper cites Imagereward: Learning and evaluating human preferences for text-to-image generation.

RewardDance: Reward Scaling in Visual Generation Imagereward: Learning and evaluating human preferences for text-to-image generation

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.130666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.130666Z digest=sha256:7c253f211aee37cc70a22fda5a4b12b7d28711aceeaad4db60d5ad26ba1e65c6

Observation 5101a8db-ce9f-4fd1-bb50-799474662b82 · outbound

This paper cites VisionReward: Fine-Grained Multi-Dimensional Human Preference Learning for Image and Video Generation.

RewardDance: Reward Scaling in Visual Generation VisionReward: Fine-Grained Multi-Dimensional Human Preference Learning for Image and Video Generation

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.179272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.179272Z digest=sha256:cb9a45028ed0d6961f8ed33f2861386fb78d55974b689b3f09d6825aa0f01156

Observation 96ff8a98-aaa1-4426-84c9-7d72a24e8711 · outbound

This paper cites A Unified Pairwise Framework for RLHF: Bridging Generative Reward Modeling and Policy Optimization.

RewardDance: Reward Scaling in Visual Generation A Unified Pairwise Framework for RLHF: Bridging Generative Reward Modeling and Policy Optimization

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.240055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.240055Z digest=sha256:15540a9e93552adffff07538ffe75526422a0ffc94c55bcea73bf612906bb185

Observation 04c91260-2b23-467c-90c4-d1a0b87aea58 · outbound

This paper cites DanceGRPO: Unleashing GRPO on Visual Generation.

RewardDance: Reward Scaling in Visual Generation DanceGRPO: Unleashing GRPO on Visual Generation

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.340171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.340171Z digest=sha256:3394112f5dee295833ae0f0489795e77c3b7b404560035733f727c404e2a81ec

Observation b0c408f1-1a71-4588-a01e-4be385e58372 · outbound

This paper cites Schedule on the fly: Diffusion time prediction for faster and better image generation.

RewardDance: Reward Scaling in Visual Generation Schedule on the fly: Diffusion time prediction for faster and better image generation

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.438578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.438578Z digest=sha256:7ab6738ed32a520dd194dec43950f61ac005a7bdfd9131731a000366e6a02d91

Observation a019f125-4279-4470-a980-f33dec0ff0f8 · outbound

This paper cites Make pixels dance: High-dynamic video generation.

RewardDance: Reward Scaling in Visual Generation Make pixels dance: High-dynamic video generation

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.521145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.521145Z digest=sha256:79a0a23e26ae2633a0e28cc88495257ec781a71874a4f71aa971137900844460

Observation 14d30f9e-c529-4d55-86ff-ac7fe409c8c4 · outbound

This paper cites Onlinevpo: Align video diffusion model with online video-centric preference optimization.arXiv preprint arXiv:2412.15159, 2024.

RewardDance: Reward Scaling in Visual Generation Onlinevpo: Align video diffusion model with online video-centric preference optimization.arXiv preprint arXiv:2412.15159, 2024

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.617649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.617649Z digest=sha256:feca5d7f2dac56e929e18d7e612df4560f974ee6341e4147feb88d85cd77658a

Observation 844e159e-4617-4f71-98e5-b22ea819c96f · outbound

This paper cites Unifl: Improve latent diffusion model via unified feedback learning.Advances in Neural Information Processing Systems, 37:67355–67382, 2024.

RewardDance: Reward Scaling in Visual Generation Unifl: Improve latent diffusion model via unified feedback learning.Advances in Neural Information Processing Systems, 37:67355–67382, 2024

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.735214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.735214Z digest=sha256:fff343d1baa91ffd3abe94d5fb8a124cf0613e20ea2e39093324f7f21c69832a

Observation a1fc5e6d-cd38-4165-878f-3d41f14e96dd · outbound

This paper cites Fine-Tuning Language Models from Human Preferences.

RewardDance: Reward Scaling in Visual Generation Fine-Tuning Language Models from Human Preferences

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.822336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.822336Z digest=sha256:9bab80dbc9c50c6bf4e3d216cc8dce1c11687f8d3bbac0ab3b0f6f6b6b87c290

Pith citing papers

Observation aa41ecd4-f187-471c-8c5d-4737894c39c5 · inbound

MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE cites this paper.

MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE RewardDance: Reward Scaling in Visual Generation

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-13T13:27:50.078729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T13:27:50.031781Z digest=sha256:2a36be8f6c5b3cfcf24b89f58825370ae8c1349e170d209fbddad3a455079074

Observation 0f9c6109-5bb3-4599-9e82-6fd95e2f7a29 · inbound

Seedream 4.0: Toward Next-generation Multimodal Image Generation cites this paper.

Seedream 4.0: Toward Next-generation Multimodal Image Generation RewardDance: Reward Scaling in Visual Generation

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-12T16:39:00.896443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-12T16:39:00.828442Z digest=sha256:b31f0dcb28289c74ffb922d9ce5b6dc8bc4f7af568a653fe64c2ed6d490a406e

Observation 88b1eb00-efe6-485c-8474-5262e6456e6c · inbound

Uniworld-V2: Reinforce Image Editing with Diffusion Negative-aware Finetuning and MLLM Implicit Feedback cites this paper.

Uniworld-V2: Reinforce Image Editing with Diffusion Negative-aware Finetuning and MLLM Implicit Feedback RewardDance: Reward Scaling in Visual Generation

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:01:19.893657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-21T18:01:19.748677Z digest=sha256:0ae9f187faa12438be1659ae5925b4df733448ad5d5ddf03bac34226ba26fa11

Observation 5bc90bef-5c59-4c41-accd-9ced908fd7a5 · inbound

Distribution Matching Distillation Meets Reinforcement Learning cites this paper.

Distribution Matching Distillation Meets Reinforcement Learning RewardDance: Reward Scaling in Visual Generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-03T21:47:20.112331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:47:20.112331Z digest=sha256:6f54737089f5bcbd28a9fec0ae8f77ceccb4ef150d21935dd1867a44b38648bc

Observation 53be49cc-2695-4ae0-a27a-2e225b9edff9 · inbound

Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation cites this paper.

Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation RewardDance: Reward Scaling in Visual Generation

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-05-16T18:17:55.054498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T18:17:54.943863Z digest=sha256:6cf9de3e31f01e9039b35718fc7002c4c5bd968f9704df92b7ec9a4f9fe01994

Observation 8e67035b-89db-48f8-9eda-19f252874dad · inbound

Seedance 1.5 pro: A Native Audio-Visual Joint Generation Foundation Model cites this paper.

Seedance 1.5 pro: A Native Audio-Visual Joint Generation Foundation Model RewardDance: Reward Scaling in Visual Generation

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-16T01:35:37.891530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T01:35:37.817082Z digest=sha256:3082a047cebf8e374764c366fd7bceb8a508263f73ee57055341303ba8f6b528

Observation 38ee40f4-d9f5-422d-b436-c1dd5598246c · inbound

Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling cites this paper.

Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling RewardDance: Reward Scaling in Visual Generation

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:45:26.073726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-25T06:41:33.493927Z digest=sha256:1f8c6d8229bf766b8c28a542fe87f686eb0b3216d96cc24a9fdea8a3c35559f4

Observation ee3ef21c-0e70-4519-894f-780fdc8f522d · inbound

Seedance 2.0: Advancing Video Generation for World Complexity cites this paper.

Seedance 2.0: Advancing Video Generation for World Complexity RewardDance: Reward Scaling in Visual Generation

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:35:26.276205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T13:34:36.248186Z digest=sha256:f9d71b4fbfe0c0b147ec57953a76027b9e27ab87259fb36a419e60efaf1f0741

Observation e7328192-0136-46d3-8da4-955bba962ac0 · inbound

A Systematic Post-Train Framework for Video Generation cites this paper.

A Systematic Post-Train Framework for Video Generation RewardDance: Reward Scaling in Visual Generation

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:26:20.184788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-07T16:58:33.014401Z digest=sha256:2bc10a1103c11c1a50d20e81fdd00bf6b08baa2f6b9cd9a5222d335a1bbea298

Observation 055871ed-053e-4653-a29e-8ef0bf1af01f · inbound

Leveraging Verifier-Based Reinforcement Learning in Image Editing cites this paper.

Leveraging Verifier-Based Reinforcement Learning in Image Editing RewardDance: Reward Scaling in Visual Generation

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:06:27.520658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-07T08:00:33.307429Z digest=sha256:d0ea7f0a449c7624bc004c92629c3658e03a52f9dfc13862b83e348a107d89fe

Observation 5eaba5f4-7dea-4564-b3bc-8bb8268a148c · inbound

Leveraging Verifier-Based Reinforcement Learning in Image Editing cites this paper.

Leveraging Verifier-Based Reinforcement Learning in Image Editing RewardDance: Reward Scaling in Visual Generation

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-21T09:14:05.973733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-21T09:11:02.183133Z digest=sha256:ec145f5753636e5884a0edf6b77596969eb65772c585679a168b557463d545d0

Observation b1b1e6f1-ac7f-48aa-bf61-77fe883a442c · inbound

Think, then Score: Decoupled Reasoning and Scoring for Video Reward Modeling cites this paper.

Think, then Score: Decoupled Reasoning and Scoring for Video Reward Modeling RewardDance: Reward Scaling in Visual Generation

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:42:30.786573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T07:37:52.346280Z digest=sha256:2064b6e9360d979e34acf3d3de56555fa121e567707e6046940365cd93c68bab

Observation 18e9cbc2-15da-49eb-b302-ae263df069f0 · inbound

TMPO: Trajectory Matching Policy Optimization for Diverse and Efficient Diffusion Alignment cites this paper.

TMPO: Trajectory Matching Policy Optimization for Diverse and Efficient Diffusion Alignment RewardDance: Reward Scaling in Visual Generation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.181018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T07:36:16.811765Z digest=sha256:831d072d7d645b7da7b04dc24d6ba910ea8102fac4f1658dcb99f122442d7de4

Observation 89a90e18-c6af-43be-ab91-cb40c0d2ff94 · inbound

TMPO: Trajectory Matching Policy Optimization for Diverse and Efficient Diffusion Alignment cites this paper.

TMPO: Trajectory Matching Policy Optimization for Diverse and Efficient Diffusion Alignment RewardDance: Reward Scaling in Visual Generation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:03:02.863836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-14T22:01:21.695270Z digest=sha256:db10c9820240e71292e22e46d5048f248ae25394aa20ae8ce8fc8857988ddc29

Observation a3dbd52d-cab0-45e8-b875-8a9872158d8c · inbound

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating cites this paper.

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating RewardDance: Reward Scaling in Visual Generation

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:02:22.029828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T06:00:31.582714Z digest=sha256:e7bf1eeada3d6e5fbfa61e025eca576564b2976bce5d6214c52f3646cdf0d821

Observation 68a52997-103e-422e-81c8-d98091d48fc1 · inbound

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating cites this paper.

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating RewardDance: Reward Scaling in Visual Generation

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:55:45.300168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-30T22:38:43.102769Z digest=sha256:1922441c68cff4441df81bdf66424c056bd1baa5d187f87c4151ad5a0c535aaa

Observation dadec1d0-7167-44be-bfb7-631db441c849 · inbound

DiffusionOPD: A Unified Perspective of On-Policy Distillation in Diffusion Models cites this paper.

DiffusionOPD: A Unified Perspective of On-Policy Distillation in Diffusion Models RewardDance: Reward Scaling in Visual Generation

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-01T14:25:46.853546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-30T21:20:23.420424Z digest=sha256:1f816a079493bf971c77f286ffef026cb6c8d9f38be26a8296e05a7d04a234ed

Observation cfaf8c48-414d-495c-b4fb-a7cb715e6ad4 · inbound

Edit-GRPO: A Locality-Preserving Policy Optimization Framework for Image Editing cites this paper.

Edit-GRPO: A Locality-Preserving Policy Optimization Framework for Image Editing RewardDance: Reward Scaling in Visual Generation

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:47:45.836625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-19T20:44:54.356658Z digest=sha256:9d505f1302dbd350d97eeb44135f820093385e3b4c023088c5a01ad4dd33c8d5

Observation bc19bdb5-f87f-4e3f-8b2b-b1779a76eed8 · inbound

StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement cites this paper.

StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement RewardDance: Reward Scaling in Visual Generation

Reference 118

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:25:59.831028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T22:45:58.629263Z digest=sha256:586f9140bb527cc251907ee8d769d5702262c698d1bf5a779043c06a46b3d1e5

Observation 485db53f-e8ef-480a-bb21-91d596b0bdc5 · inbound

Improving Visual Representation Alignment Generation with GRPO cites this paper.

Improving Visual Representation Alignment Generation with GRPO RewardDance: Reward Scaling in Visual Generation

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T19:32:35.129428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T19:24:09.597127Z digest=sha256:f8ab42e6d7bfd6a88a5cbd453fccccf2e67ce40aab1e50f63bc969c2e8de57d4

Observation 876aee65-1061-4fbc-b3d5-b22c4f1ad8aa · inbound

Are we really tilting? The mechanics of reward guidance in flow and diffusion models cites this paper.

Are we really tilting? The mechanics of reward guidance in flow and diffusion models RewardDance: Reward Scaling in Visual Generation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:26:18.153780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T15:20:52.980868Z digest=sha256:3888a297bf5cf7a6e7f59e1f7e753e613b5761838de9f32454e657904e6e629d

Observation 1c45d939-a4dd-476c-9e57-1990a9cbc875 · inbound

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions cites this paper.

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions RewardDance: Reward Scaling in Visual Generation

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:07:27.944548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T17:30:57.001021Z digest=sha256:e0a66392595e0547830e44a9433692b92918bcc461e4264727a189d15fc80c49

Observation 6e797adc-88a9-454b-89af-7a97da480d64 · inbound

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions cites this paper.

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions RewardDance: Reward Scaling in Visual Generation

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-15T10:53:37.186361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T10:53:37.186361Z digest=sha256:b51c457fd760e016929ad12e916251a6ae56887e93c941573ca71a853498ee62

Observation 494cc52e-0a45-4538-8d3d-0487d70aa0d5 · inbound

Each Judge Its Own Yardstick: Discovering Per-VLM Taxonomies for Physical Video Evaluation cites this paper.

Each Judge Its Own Yardstick: Discovering Per-VLM Taxonomies for Physical Video Evaluation RewardDance: Reward Scaling in Visual Generation

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:09:45.413050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-26T09:04:23.965554Z digest=sha256:24f5ba2ced6bbaf3eaf9b5f114b1ced1c852b04a78b9df40d4b9d83f7714e2a1

Observation 0b5713ab-4f56-45d7-9d42-b547b47d54e2 · inbound

DiT-Reward: Generative Representations for Text-to-Image Reward Modeling cites this paper.

DiT-Reward: Generative Representations for Text-to-Image Reward Modeling RewardDance: Reward Scaling in Visual Generation

Reference 118

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:39:45.925822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-26T08:35:04.038960Z digest=sha256:936c100ed9a5c37ddcbe3fe4092b340e01ea1786ed4c350f094c7d73847411d3

Observation 47831f1c-db81-48f8-b061-18d80119d225 · inbound

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning cites this paper.

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning RewardDance: Reward Scaling in Visual Generation

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:03:51.757351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T05:00:04.037802Z digest=sha256:6be96163aa5e8637daac8c3048c8aaf9f41293ffc080febf384641f5fd779b98

Observation a16b5dda-0a53-42b1-9613-000a8c0c17a9 · inbound

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning cites this paper.

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning RewardDance: Reward Scaling in Visual Generation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-12T11:38:13.566600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:38:13.566600Z digest=sha256:ff1b9aef2a09c6f7754a3c9d82a9e9a0903e8dd8bb24c29e74187e67138f88d9

Observation c62e0581-9e0c-43ee-b72f-9b3456248803 · inbound

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning cites this paper.

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning RewardDance: Reward Scaling in Visual Generation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T09:56:25.929049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T09:56:25.929049Z digest=sha256:df86c9abb96b585d14c05bc2eeed8ed5ae498803c268f39ba0421cf8570aab03

Observation 99a15030-df7d-4c09-b32d-139e4a24111f · inbound

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation cites this paper.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation RewardDance: Reward Scaling in Visual Generation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.594851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.594851Z digest=sha256:7f9f9914bff641f9e1ab26616b9f1f2efc9d75ed212a8bfce63c1ca3929befbc

Observation 25d55f6b-63ee-43f3-b764-9312d1880f41 · inbound

SciForma: Structure-Faithful Generation of Scientific Diagrams cites this paper.

SciForma: Structure-Faithful Generation of Scientific Diagrams RewardDance: Reward Scaling in Visual Generation

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-01T16:12:23.899000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T16:12:23.899000Z digest=sha256:04bad9f137ca062475efc447ea4e4dc8ec313a2d47a857d93bdfcd51f5f864c5