Pith. sign in

Paper Citation Record · LEDGER

RewardDance: Reward Scaling in Visual Generation

As of 10 August 2026, this Paper Citation Record lists 72 of 72 outbound references and 30 inbound Pith citation observations for arXiv:2509.08826.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.08826 v1

Coverage vector

measured 72 of 72 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T20:09:01.822336Z

measured 102 of 102 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 30 of 30 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T21:47:20.112331Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:39:45.924030Z

Reference resolution

72 of 72 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved72
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 86905ea5-c022-43bb-8b50-c37819f87f6f · outbound

This paper cites Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models.

RewardDance: Reward Scaling in Visual Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:55.396895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:55.396895Z digest=sha256:951e86282ec5b379474dda8e0717d992a01ee343ea4369aa56e9501394fe7671

Observation ea855739-cf41-40b2-aa1e-1d10a6592d1f · outbound

This paper cites Improving image generation with better captions.Computer Science.

RewardDance: Reward Scaling in Visual Generation Improving image generation with better captions.Computer Science

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:55.529478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:55.529478Z digest=sha256:81a46f4527b2b5b215520b1182cd4cf7bc4c5db7014a9da2b05b26a4ad6ae29d

Observation 132c6830-d12c-43f7-9156-dddaf4ed4d40 · outbound

This paper cites Training Diffusion Models with Reinforcement Learning.

RewardDance: Reward Scaling in Visual Generation Training Diffusion Models with Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:55.650307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:55.650307Z digest=sha256:08d8be5ee1ea682cf796140a79fc55e5ceccb68d43e88a80a1b7de09e9c48e35

Observation 7efa567e-ef00-4989-a064-45e848d36bbf · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

RewardDance: Reward Scaling in Visual Generation Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:55.746553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:55.746553Z digest=sha256:753e29928072874ec369b80690caa8edd648d0cef6c34948669c1138323564c1

Observation 3abdd8b0-0899-4016-a961-6a5bcadc57df · outbound

This paper cites Rank analysis of incomplete block designs: I.

RewardDance: Reward Scaling in Visual Generation Rank analysis of incomplete block designs: I

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:55.865292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:55.865292Z digest=sha256:28ef9242b6cf8b7a02e51995a6e892c851844940395e92f5ac70153c50bc501a

Observation c3a37bb2-8a9b-4395-8d20-11ff4747b2f4 · outbound

This paper cites Video generation models as world simulators.OpenAI Blog, 1(8):1, 2024.

RewardDance: Reward Scaling in Visual Generation Video generation models as world simulators.OpenAI Blog, 1(8):1, 2024

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:55.944987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:55.944987Z digest=sha256:d3d3964f35fc341e025c22854c151ba744744dbaeab501ed1c4b5953f8d038c8

Observation 0f4bcd2d-bb19-4f13-b4b8-18379f110cfc · outbound

This paper cites PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis.

RewardDance: Reward Scaling in Visual Generation PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.030031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.030031Z digest=sha256:960eab273ffe6d564c4061a5bb54c7d289dc9dc3a52941256b614b24a9ea472c

Observation e91579b9-d6e5-4c0c-9e9d-818bfe04fc4c · outbound

This paper cites Control-a-video: Controllable text-to-video generation with diffusion models.CoRR, 2023.

RewardDance: Reward Scaling in Visual Generation Control-a-video: Controllable text-to-video generation with diffusion models.CoRR, 2023

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.108706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.108706Z digest=sha256:588441805fc46f19b42b3ef2670539ecb81d712d45c8d289d0cabb3c827c8581

Observation b272c883-8db7-4937-86eb-367ad3b86f30 · outbound

This paper cites The Accuracy Paradox in RLHF: When Better Reward Models Don't Yield Better Language Models.

RewardDance: Reward Scaling in Visual Generation The Accuracy Paradox in RLHF: When Better Reward Models Don't Yield Better Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.202728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.202728Z digest=sha256:8b89536e709e0cf8235b9c6d87fe481df2def93c01f438aeb26a610b1677c3b3

Observation 486d1163-e00a-4ea2-9d8a-5cf8cd809e93 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

RewardDance: Reward Scaling in Visual Generation Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.309210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.309210Z digest=sha256:677fcf86eb73b8b5d6c310b917de3578bb2921bf7f674615a572c3b70d0fee9c

Observation e2b42bf0-44f1-4691-ad5a-fc91b0429d1a · outbound

This paper cites RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment.

RewardDance: Reward Scaling in Visual Generation RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.415726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.415726Z digest=sha256:b38d2596e4ce771552df5b63d9384b2489e79207e9844a8ba71027d0d7491727

Observation e8c40eb9-35e7-4a24-aa42-90d44b3ff2ab · outbound

This paper cites Scaling rectified flow transformers for high-resolution image synthesis.

RewardDance: Reward Scaling in Visual Generation Scaling rectified flow transformers for high-resolution image synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.497198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.497198Z digest=sha256:d01b0fd7c14b8bd3b2be8a1b34654f181c309bb5c9a1c2a41f067df8f8336f73

Observation 1382ada2-b0c5-4221-8d9c-7225d5b6876a · outbound

This paper cites Reinforcement learning for fine-tuning text-to-image diffusion mod- els.

RewardDance: Reward Scaling in Visual Generation Reinforcement learning for fine-tuning text-to-image diffusion mod- els

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.572535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.572535Z digest=sha256:1c8e75b735370a1fdbfc74d48a11455c74018f318a20f4d482f6b020a83446ab

Observation c4bbc6c6-5620-430b-94b2-27eedc40a862 · outbound

This paper cites Seedream 3.0 Technical Report.

RewardDance: Reward Scaling in Visual Generation Seedream 3.0 Technical Report

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.678650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.678650Z digest=sha256:726a3f754c2802993c153630cb5a041952a8c0811bf88594dbac3a7ae97edacd

Observation e1ed9b66-e8ca-4b34-9f27-5433927eb701 · outbound

This paper cites Seedance 1.0: Exploring the Boundaries of Video Generation Models.

RewardDance: Reward Scaling in Visual Generation Seedance 1.0: Exploring the Boundaries of Video Generation Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.802057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.802057Z digest=sha256:9663d07af1417310ae862f63c67b4c063a24e93c7191ef5d10895d25082454e1

Observation da4d94ab-2adb-47a2-8c1f-a6dcf2fc3c85 · outbound

This paper cites Seedream 2.0: A Native Chinese-English Bilingual Image Generation Foundation Model.

RewardDance: Reward Scaling in Visual Generation Seedream 2.0: A Native Chinese-English Bilingual Image Generation Foundation Model

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.884737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.884737Z digest=sha256:97046a5cfbac7942c75dd648a6b844567ea647047e817ad925a33b3e12cb34c4

Observation d6e63974-c7c7-459b-a90b-de923f36774b · outbound

This paper cites Veo.https://deepmind.google/models/veo/, 2025.

RewardDance: Reward Scaling in Visual Generation Veo.https://deepmind.google/models/veo/, 2025

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.953258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.953258Z digest=sha256:1b088ca86302ca647411474bce53e5c96e845103ec4d5bbaf44799bdc7aa127f

Observation 301b6b30-f1d6-4a81-8021-787bedd2734f · outbound

This paper cites Multi-Reward as Condition for Instruction-based Image Editing.

RewardDance: Reward Scaling in Visual Generation Multi-Reward as Condition for Instruction-based Image Editing

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.041134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.041134Z digest=sha256:d29e62d0c3f50f11699cf805a02c39a20d251c867840923b91e79a971c94a238

Observation 9093d8ed-5bfb-4fb1-b9d5-4663fcf14436 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

RewardDance: Reward Scaling in Visual Generation AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.137515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.137515Z digest=sha256:8a5fb028612188d8264f646fba934106765aa2c2db6d5cec9ee61c6cea518c67

Observation 32d7c0f9-754f-47c4-b243-f3d6527613a3 · outbound

This paper cites A simple and effective reinforcement learning method for text-to-image diffusion fine-tuning.

RewardDance: Reward Scaling in Visual Generation A simple and effective reinforcement learning method for text-to-image diffusion fine-tuning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.228352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.228352Z digest=sha256:2f6bb04385e02b9d2d02f4ca7d086e703104bf0c95235326d86ec671d276987c

Observation 31caed26-a32f-4c1d-a531-0594ffda0736 · outbound

This paper cites Denoising diffusion probabilistic models.

RewardDance: Reward Scaling in Visual Generation Denoising diffusion probabilistic models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.297593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.297593Z digest=sha256:1854363bd0bea437a900b4aa8ccf446f5cb44af92ab5c205ca976d3f2efba023

Observation 79c2c946-32a1-47a5-a3c1-8ee6c77af671 · outbound

This paper cites Ideogram.https://about.ideogram.ai/1.0., 2024.

RewardDance: Reward Scaling in Visual Generation Ideogram.https://about.ideogram.ai/1.0., 2024

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.397897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.397897Z digest=sha256:9c2d93a633d7d3b62cd01e85d607428517704a98a005abb7e7d68e523a1f4ac3

Observation b88cd1fa-c9b6-41d9-aa71-bf3d89cbaa92 · outbound

This paper cites Pick-a-pic: An open dataset of user preferences for text-to-image generation.Advancesin neural information processing systems, 36:36652–36663, 2023.

RewardDance: Reward Scaling in Visual Generation Pick-a-pic: An open dataset of user preferences for text-to-image generation.Advancesin neural information processing systems, 36:36652–36663, 2023

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.483267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.483267Z digest=sha256:c48bb5b0ee8f2169fac4cc5a8aabc3c7808f9cf193e4483ec70723ab82484d5c

Observation 200e654a-6de4-4478-86be-48c758073b93 · outbound

This paper cites klingai.https://app.klingai.com/cn/, 2025.

RewardDance: Reward Scaling in Visual Generation klingai.https://app.klingai.com/cn/, 2025

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.585056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.585056Z digest=sha256:9f16ae68d9704b343d55866e1b29bca189e9ae2b57ccb8848733c97dc773fb41

Observation 836d2260-1b27-41f8-82c9-82b7380589d3 · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

RewardDance: Reward Scaling in Visual Generation HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.666163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.666163Z digest=sha256:0b2e1f96b3284d394cde8b060f4e7af4832f3a6f560cb00176a6e03ca0bdf659

Observation af1ebe2a-2734-4c43-a81c-231af56b104d · outbound

This paper cites Flux: Official inference repository for flux.1 models, 2024.

RewardDance: Reward Scaling in Visual Generation Flux: Official inference repository for flux.1 models, 2024

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.744104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.744104Z digest=sha256:0e9dc66023397a5b583d107297ceb9983f4580a0829c0e5655548b86aa6826f4

Observation 94a153bb-e4c9-46f1-bbac-6a8c7f108c32 · outbound

This paper cites Flux.https://github.com/black-forest-labs/flux, 2024.

RewardDance: Reward Scaling in Visual Generation Flux.https://github.com/black-forest-labs/flux, 2024

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.838549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.838549Z digest=sha256:3addf0373ef57fee4bd6c7745c021c51114f768d4a9067dd7eaba867db7be641

Observation 8e4c83b1-f2ed-44cd-91a9-df9a6b582506 · outbound

This paper cites Controlnet++: Improving conditional controls with efficient consistency feedback: Project page: liming-ai.

RewardDance: Reward Scaling in Visual Generation Controlnet++: Improving conditional controls with efficient consistency feedback: Project page: liming-ai

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.926797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.926797Z digest=sha256:2ce1d1b1dc54f54f3e122b680387b9cad7b17b1fdf1200285c7e7a24e0119ffd

Observation fba2b18b-871c-4917-8670-bbc8e802f8ee · outbound

This paper cites SuperEdit: Rectifying and Facilitating Supervision for Instruction-Based Image Editing.

RewardDance: Reward Scaling in Visual Generation SuperEdit: Rectifying and Facilitating Supervision for Instruction-Based Image Editing

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.003125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.003125Z digest=sha256:ae4ca749254f834883e073a1fba0df5202021cad324159b9135ab7fc2ce2f0ae

Observation 980525a9-64b4-4034-bc9c-1b258ed6ce7a · outbound

This paper cites Exploring How Generative MLLMs Perceive More Than CLIP with the Same Vision Encoder.

RewardDance: Reward Scaling in Visual Generation Exploring How Generative MLLMs Perceive More Than CLIP with the Same Vision Encoder

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.110097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.110097Z digest=sha256:9b22cfe4336cc9fd08e53b6a2fe3ae4d31ca939143c91996e6f3c33bfef9a251

Observation 40a2ebf7-fa29-4820-92fe-524f2545a508 · outbound

This paper cites An inverse scaling law for clip training.Advancesin Neural Information Processing Systems, 36:49068–49087, 2023.

RewardDance: Reward Scaling in Visual Generation An inverse scaling law for clip training.Advancesin Neural Information Processing Systems, 36:49068–49087, 2023

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.185503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.185503Z digest=sha256:da6edd410907adec508484291899e41e9724d6c56030b1247bc93e1137aa9ee4

Observation 9e56985d-a38c-4a4a-bc8a-387d8247ca8e · outbound

This paper cites Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation.

RewardDance: Reward Scaling in Visual Generation Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.270191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.270191Z digest=sha256:5b14a8f2c7c9a9f4d8bae52ba1ab505c1d84032db6283b8dfe9dca6e553bbeb0

Observation dbb474ed-47dd-4165-879f-85a61565a7b8 · outbound

This paper cites Flow Matching for Generative Modeling.

RewardDance: Reward Scaling in Visual Generation Flow Matching for Generative Modeling

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.373978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.373978Z digest=sha256:8d9371d542b6f89ed44b163ef8024729e01ca88262a166031327d7104ad89c62

Observation 8fa26720-2d27-44f9-ab9f-4fc28f386f7a · outbound

This paper cites Flow-GRPO: Training Flow Matching Models via Online RL.

RewardDance: Reward Scaling in Visual Generation Flow-GRPO: Training Flow Matching Models via Online RL

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.469893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.469893Z digest=sha256:088d10310e0ae4e9b859a69f8849098ca668cee90a4b5d49e8a0b01a7cc81f4c

Observation dcf4dd50-8905-4504-afa9-b61b94c65d46 · outbound

This paper cites Improving Video Generation with Human Feedback.

RewardDance: Reward Scaling in Visual Generation Improving Video Generation with Human Feedback

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.572097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.572097Z digest=sha256:cff5aadea14cee3c09ebd4d7a30e4e2d807b241c4940ca46194deaa8bacd2531

Observation b90777e3-0f54-4c9f-b360-3e758cd57151 · outbound

This paper cites Inference-time scaling for generalist reward modeling.arXiv preprint arXiv:2504.02495, 2025.

RewardDance: Reward Scaling in Visual Generation Inference-time scaling for generalist reward modeling.arXiv preprint arXiv:2504.02495, 2025

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.678868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.678868Z digest=sha256:d7013c7f075477de4bd758e8c9a668d3fc9dfc5bf1a758465a47d421cabd1b63

Observation d26e027a-12f0-4535-b7df-455007611c6f · outbound

This paper cites lumalabs.https://lumalabs.ai/, 2024.

RewardDance: Reward Scaling in Visual Generation lumalabs.https://lumalabs.ai/, 2024

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.744034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.744034Z digest=sha256:dcf74fb32090a647c073a360add363d59d36ef57409e28ef3c78e3992fbf6140

Observation e5445f7e-d9e2-48ea-89bf-662cbd1b64c2 · outbound

This paper cites Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model.

RewardDance: Reward Scaling in Visual Generation Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.832391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.832391Z digest=sha256:428eed64e8a401461149b5f5de79383c7badffc21e6bf8383ff71ae538512a5b

Observation 873f0fa2-ce1c-463e-81ea-5695252c4ed2 · outbound

This paper cites Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps.

RewardDance: Reward Scaling in Visual Generation Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.890930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.890930Z digest=sha256:371173e7f5bafce8acee0e26bb24b58b4d879889025645dad853ef4d3ddcd873

Observation 3930fdb3-dd8b-4d9a-abf8-1030b39d043a · outbound

This paper cites HPSv3: Towards Wide-Spectrum Human Preference Score.

RewardDance: Reward Scaling in Visual Generation HPSv3: Towards Wide-Spectrum Human Preference Score

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.995776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.995776Z digest=sha256:2382ebb808c3c25422c8f608e8c769982cfcbb0c407b3479f8b54e0e9dd768d2

Observation 69192038-94c5-445a-bf15-8916ef183f2f · outbound

This paper cites midjourney.https://www.midjourney.com/home, 2024.

RewardDance: Reward Scaling in Visual Generation midjourney.https://www.midjourney.com/home, 2024

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.094853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.094853Z digest=sha256:3d3d5508e446e8c02215d5315aacf148609b0bf4ce73d9d1a28e5438d4d4b61a

Observation d0664f12-d885-48ec-a8de-318928b50deb · outbound

This paper cites Inference-time text-to-video alignment with diffusion latent beam search.arXiv preprint arXiv:2501.19252, 2025.

RewardDance: Reward Scaling in Visual Generation Inference-time text-to-video alignment with diffusion latent beam search.arXiv preprint arXiv:2501.19252, 2025

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.185710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.185710Z digest=sha256:9658c9d0126ca790ebb8464f12a67dd1983885c2bf2c71ca7a0c0285141507d8

Observation 668c5b5a-e40c-4444-9aaf-14c4c7247767 · outbound

This paper cites Training language models to follow instructions with human feedback.

RewardDance: Reward Scaling in Visual Generation Training language models to follow instructions with human feedback

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.294363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.294363Z digest=sha256:c099149e3315834e7b633e3d241d02f31e2e7002a2a183539cf22c8f0fde6647

Observation 866901a6-8d1e-4f11-9f59-ecec90bfcced · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

RewardDance: Reward Scaling in Visual Generation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.377336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.377336Z digest=sha256:c578aa8de99b84e8fd2b05ee56dba89017d698af1faf75dd53ed3f3c5f3e6504

Observation 977bd8af-fed2-4656-ba94-2b087858437a · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

RewardDance: Reward Scaling in Visual Generation Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.470265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.470265Z digest=sha256:d53df68ee97b84d01eb26e151d6494634df788005273a6ee979b31ffa9c7d4c4

Observation efc721e4-fdb0-4f99-a5ea-fab919a1130e · outbound

This paper cites What makes a reward model a good teacher? an optimization perspective.arXiv preprint arXiv:2503.15477, 2025.

RewardDance: Reward Scaling in Visual Generation What makes a reward model a good teacher? an optimization perspective.arXiv preprint arXiv:2503.15477, 2025

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.549832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.549832Z digest=sha256:bf42b5c88df6f87549cc88507473c5347be3c94c31c7a4c923c3b70c85d7273f

Observation e3b317e8-9a99-42d8-87b3-6274f9ad808e · outbound

This paper cites recraft.https://www.recraft.ai/, 2024.

RewardDance: Reward Scaling in Visual Generation recraft.https://www.recraft.ai/, 2024

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.607627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.607627Z digest=sha256:3ccd29e71f341a493332db052f9815acd3789d76c9311973decc6e095d58b12f

Observation 81fb23ff-d3e2-4f00-99b9-c2f63a0f3a80 · outbound

This paper cites Byteedit: Boost, comply and accelerate generative image editing.

RewardDance: Reward Scaling in Visual Generation Byteedit: Boost, comply and accelerate generative image editing

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.703606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.703606Z digest=sha256:6c90ba1c6799dd5d43cf6596bf691da9b17dd95dc340ab1fa94856fb8f51fd12

Observation 5da2af02-3f3f-409f-baf2-707e228638a6 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

RewardDance: Reward Scaling in Visual Generation High-resolution image synthesis with latent diffusion models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.800075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.800075Z digest=sha256:a0a0ea54ca9b6095564b673625595bee2cc661dd523fd7e7f2a935cadbc39221

Observation f14a6ad6-6c8e-4c6e-bb67-41c035ef8863 · outbound

This paper cites Runway.https://runwayml.com/research/introducing-runway-gen-4, 2025.

RewardDance: Reward Scaling in Visual Generation Runway.https://runwayml.com/research/introducing-runway-gen-4, 2025

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.895610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.895610Z digest=sha256:c5ea2016a51d3bff185218e75be44f038e71db9fca34f90b295896d1fb675588

Observation 064fc98f-5828-48a3-ba5a-292bb297fa5d · outbound

This paper cites Seaweed-7B: Cost-Effective Training of Video Generation Foundation Model.

RewardDance: Reward Scaling in Visual Generation Seaweed-7B: Cost-Effective Training of Video Generation Foundation Model

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.968870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.968870Z digest=sha256:3db29357af0f002461a82ea2baf6b224bbca13bb880334dda5319d23ba8d3a0a

Observation d208184c-7791-4c8a-a987-fc85114cded9 · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

RewardDance: Reward Scaling in Visual Generation Deep unsupervised learning using nonequilibrium thermodynamics

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.052186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.052186Z digest=sha256:77eec35bc58529d5d8a5c5da832c48ad7bbc43d37478dc3e4d6adbf5f193f830

Observation 83058bb5-07f5-4989-a8b5-0a2f1f9925b6 · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

RewardDance: Reward Scaling in Visual Generation Score-Based Generative Modeling through Stochastic Differential Equations

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.112624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.112624Z digest=sha256:42a44921b472396c5254486becca1ab6dd76ba9ed2531774a1b331cb69b6d6ae

Observation a6df593d-ba6d-4eeb-a27b-ee3828f84c4b · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

RewardDance: Reward Scaling in Visual Generation Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.245755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.245755Z digest=sha256:f6ca526fd2a4935087bdea047c3b4fd0b8547220eb7169cac1f7d37ca2677327

Observation 4b61a4f2-c478-4e5b-bc2e-49d7706edeef · outbound

This paper cites Diffusion model alignment using direct preference optimization.

RewardDance: Reward Scaling in Visual Generation Diffusion model alignment using direct preference optimization

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.340485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.340485Z digest=sha256:c150d33183598d02fe074891c62cc4956ed5358966c414087a36204a29833603

Observation 02bedef1-7f2a-4ce3-bb51-59c74b8c2387 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

RewardDance: Reward Scaling in Visual Generation Wan: Open and Advanced Large-Scale Video Generative Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.409906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.409906Z digest=sha256:b1c405b0309703cd2e522f82690da63fd8af6004322c2b3acd32244499e34af5

Observation c6887f61-aa82-4a10-9f0e-16a07b78309b · outbound

This paper cites WorldPM: Scaling Human Preference Modeling.

RewardDance: Reward Scaling in Visual Generation WorldPM: Scaling Human Preference Modeling

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.504202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.504202Z digest=sha256:51c68368f22d41bf113b1030f116fd0b488bb449d6f89b57a8125c4f92738755

Observation de72a69d-f92c-4349-9115-e5f9f698a074 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

RewardDance: Reward Scaling in Visual Generation Emu3: Next-Token Prediction is All You Need

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.588199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.588199Z digest=sha256:6ded1062440af6f833f330fb1ed7fe9b1c0ff0829018462d1764b001332b2cc4

Observation 6efd3f03-2a1b-4d8e-b84f-c6ee7772ab80 · outbound

This paper cites Unified multimodal chain-of-thought reward model through reinforcement fine-tuning.arXiv preprint arXiv:2505.03318, 2025.

RewardDance: Reward Scaling in Visual Generation Unified multimodal chain-of-thought reward model through reinforcement fine-tuning.arXiv preprint arXiv:2505.03318, 2025

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.663870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.663870Z digest=sha256:53bbc2cf4723acac42186d272d8376937b3cf7bd96006b009cadb7ef9a4804b8

Observation efae6b3f-4095-4f72-8914-968501479d83 · outbound

This paper cites Rethinking Reward Model Evaluation: Are We Barking up the Wrong Tree?.

RewardDance: Reward Scaling in Visual Generation Rethinking Reward Model Evaluation: Are We Barking up the Wrong Tree?

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.786327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.786327Z digest=sha256:7235ce0f8469fc4516ffaf57ebccac714a7fdf168c737d1986bcb15f32047f8e

Observation eb924bd3-99c4-4e38-a602-3e7f9fda43e7 · outbound

This paper cites Qwen-Image Technical Report.

RewardDance: Reward Scaling in Visual Generation Qwen-Image Technical Report

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.883913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.883913Z digest=sha256:cdf13ccf859cfb71d382028848c7f8dfb5846b749482e83b4bba78a209ddf12e

Observation 46e3a5f9-eca2-47eb-a390-548ed9974217 · outbound

This paper cites Human Preference Score: Better Aligning Text-to-Image Models with Human Preference.

RewardDance: Reward Scaling in Visual Generation Human Preference Score: Better Aligning Text-to-Image Models with Human Preference

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.976315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.976315Z digest=sha256:fd387428a83bbf58142c8236ad21c365f020325dbca5d2b89d30ccd0b75b19b0

Observation 2d92b21d-ec74-43a2-b9b2-9b58ff1aa849 · outbound

This paper cites Human preference score: Better aligning text-to-image models with human preference.

RewardDance: Reward Scaling in Visual Generation Human preference score: Better aligning text-to-image models with human preference

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.048080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.048080Z digest=sha256:4ab794bbb770e04a5d3da6884d7b0e4a613258c1d5218076aa41d6faae86a62d

Observation aa80b625-a1c6-4678-a710-15e30fb59008 · outbound

This paper cites Imagereward: Learning and evaluating human preferences for text-to-image generation.

RewardDance: Reward Scaling in Visual Generation Imagereward: Learning and evaluating human preferences for text-to-image generation

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.130666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.130666Z digest=sha256:7c253f211aee37cc70a22fda5a4b12b7d28711aceeaad4db60d5ad26ba1e65c6

Observation 5101a8db-ce9f-4fd1-bb50-799474662b82 · outbound

This paper cites VisionReward: Fine-Grained Multi-Dimensional Human Preference Learning for Image and Video Generation.

RewardDance: Reward Scaling in Visual Generation VisionReward: Fine-Grained Multi-Dimensional Human Preference Learning for Image and Video Generation

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.179272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.179272Z digest=sha256:ec603923a8ef0a63a111662b865d7f2a80b16633b49a71a709fe09a9cb5bb69f

Observation 96ff8a98-aaa1-4426-84c9-7d72a24e8711 · outbound

This paper cites A Unified Pairwise Framework for RLHF: Bridging Generative Reward Modeling and Policy Optimization.

RewardDance: Reward Scaling in Visual Generation A Unified Pairwise Framework for RLHF: Bridging Generative Reward Modeling and Policy Optimization

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.240055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.240055Z digest=sha256:0b58ff197cc64b125ff27c0493ed49b963448b78f3627ab2b10ecae18e4524cc

Observation 04c91260-2b23-467c-90c4-d1a0b87aea58 · outbound

This paper cites DanceGRPO: Unleashing GRPO on Visual Generation.

RewardDance: Reward Scaling in Visual Generation DanceGRPO: Unleashing GRPO on Visual Generation

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.340171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.340171Z digest=sha256:5fe7f17defce741c2d79f0a60c09f9af5acab7de6575db3afcd15201a7b46e6b

Observation b0c408f1-1a71-4588-a01e-4be385e58372 · outbound

This paper cites Schedule on the fly: Diffusion time prediction for faster and better image generation.

RewardDance: Reward Scaling in Visual Generation Schedule on the fly: Diffusion time prediction for faster and better image generation

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.438578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.438578Z digest=sha256:7ab6738ed32a520dd194dec43950f61ac005a7bdfd9131731a000366e6a02d91

Observation a019f125-4279-4470-a980-f33dec0ff0f8 · outbound

This paper cites Make pixels dance: High-dynamic video generation.

RewardDance: Reward Scaling in Visual Generation Make pixels dance: High-dynamic video generation

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.521145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.521145Z digest=sha256:79a0a23e26ae2633a0e28cc88495257ec781a71874a4f71aa971137900844460

Observation 14d30f9e-c529-4d55-86ff-ac7fe409c8c4 · outbound

This paper cites Onlinevpo: Align video diffusion model with online video-centric preference optimization.arXiv preprint arXiv:2412.15159, 2024.

RewardDance: Reward Scaling in Visual Generation Onlinevpo: Align video diffusion model with online video-centric preference optimization.arXiv preprint arXiv:2412.15159, 2024

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.617649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.617649Z digest=sha256:feca5d7f2dac56e929e18d7e612df4560f974ee6341e4147feb88d85cd77658a

Observation 844e159e-4617-4f71-98e5-b22ea819c96f · outbound

This paper cites Unifl: Improve latent diffusion model via unified feedback learning.Advances in Neural Information Processing Systems, 37:67355–67382, 2024.

RewardDance: Reward Scaling in Visual Generation Unifl: Improve latent diffusion model via unified feedback learning.Advances in Neural Information Processing Systems, 37:67355–67382, 2024

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.735214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.735214Z digest=sha256:fff343d1baa91ffd3abe94d5fb8a124cf0613e20ea2e39093324f7f21c69832a

Observation a1fc5e6d-cd38-4165-878f-3d41f14e96dd · outbound

This paper cites Fine-Tuning Language Models from Human Preferences.

RewardDance: Reward Scaling in Visual Generation Fine-Tuning Language Models from Human Preferences

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.822336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.822336Z digest=sha256:6435be7d0dbb06ca0addeeb892bf9551d1e11dfe7667c82830a3f22c4bb72ca8

Pith citing papers

Observation aa41ecd4-f187-471c-8c5d-4737894c39c5 · inbound

MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE cites this paper.

MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE RewardDance: Reward Scaling in Visual Generation

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-13T13:27:50.078729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T13:27:50.031781Z digest=sha256:78492ff8d27194e054d6455385ec33a429114ce6b5a07df08b54d5420ecad90a

Observation 0f9c6109-5bb3-4599-9e82-6fd95e2f7a29 · inbound

Seedream 4.0: Toward Next-generation Multimodal Image Generation cites this paper.

Seedream 4.0: Toward Next-generation Multimodal Image Generation RewardDance: Reward Scaling in Visual Generation

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-12T16:39:00.896443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T16:39:00.828442Z digest=sha256:068c0b5e4bb786782b4d1a72adbddcbc5fc29f27741157bb9babf84feb0152d0

Observation 88b1eb00-efe6-485c-8474-5262e6456e6c · inbound

Uniworld-V2: Reinforce Image Editing with Diffusion Negative-aware Finetuning and MLLM Implicit Feedback cites this paper.

Uniworld-V2: Reinforce Image Editing with Diffusion Negative-aware Finetuning and MLLM Implicit Feedback RewardDance: Reward Scaling in Visual Generation

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:01:19.893657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T18:01:19.748677Z digest=sha256:79ce609b8e807ad0818b98aa0b79efe740f2c277fd85fdec12249ad19e305b6e

Observation 5bc90bef-5c59-4c41-accd-9ced908fd7a5 · inbound

Distribution Matching Distillation Meets Reinforcement Learning cites this paper.

Distribution Matching Distillation Meets Reinforcement Learning RewardDance: Reward Scaling in Visual Generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-03T21:47:20.112331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:47:20.112331Z digest=sha256:0256aefdd6fe408864ca8348c277d6d7be7cb250cfa1723aad475984d3991852

Observation 53be49cc-2695-4ae0-a27a-2e225b9edff9 · inbound

Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation cites this paper.

Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation RewardDance: Reward Scaling in Visual Generation

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-05-16T18:17:55.054498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T18:17:54.943863Z digest=sha256:070e2cdf6cb54242320526fdc8af22cf9a7e37dabdf76187e9912f0f01e25d56

Observation 8e67035b-89db-48f8-9eda-19f252874dad · inbound

Seedance 1.5 pro: A Native Audio-Visual Joint Generation Foundation Model cites this paper.

Seedance 1.5 pro: A Native Audio-Visual Joint Generation Foundation Model RewardDance: Reward Scaling in Visual Generation

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-16T01:35:37.891530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T01:35:37.817082Z digest=sha256:b5413d37a3893661a1b70649b82a7a835de90915f899357f5c3d95566a76c62d

Observation 38ee40f4-d9f5-422d-b436-c1dd5598246c · inbound

Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling cites this paper.

Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling RewardDance: Reward Scaling in Visual Generation

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:45:26.073726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-25T06:41:33.493927Z digest=sha256:352c25ad474cb4af4ecc90abb86d7873ab7ef6b641cfc62acd871f0a633bb011

Observation ee3ef21c-0e70-4519-894f-780fdc8f522d · inbound

Seedance 2.0: Advancing Video Generation for World Complexity cites this paper.

Seedance 2.0: Advancing Video Generation for World Complexity RewardDance: Reward Scaling in Visual Generation

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:35:26.276205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T13:34:36.248186Z digest=sha256:41c9bdffd16a945adae1d757aa89a98d106b3ba94d05a63aa3aa22c559fb7bcb

Observation e7328192-0136-46d3-8da4-955bba962ac0 · inbound

A Systematic Post-Train Framework for Video Generation cites this paper.

A Systematic Post-Train Framework for Video Generation RewardDance: Reward Scaling in Visual Generation

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:26:20.184788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T16:58:33.014401Z digest=sha256:8b499906c4d22e3c0a5b40c3a42d0e377428d857d0fa25c49e97dac40fa85bc5

Observation 055871ed-053e-4653-a29e-8ef0bf1af01f · inbound

Leveraging Verifier-Based Reinforcement Learning in Image Editing cites this paper.

Leveraging Verifier-Based Reinforcement Learning in Image Editing RewardDance: Reward Scaling in Visual Generation

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:06:27.520658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T08:00:33.307429Z digest=sha256:2571da3f8dcb2707dbd394254e088502f8e3f0aebb333db70c084a1d12ee042c

Observation 5eaba5f4-7dea-4564-b3bc-8bb8268a148c · inbound

Leveraging Verifier-Based Reinforcement Learning in Image Editing cites this paper.

Leveraging Verifier-Based Reinforcement Learning in Image Editing RewardDance: Reward Scaling in Visual Generation

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-21T09:14:05.973733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T09:11:02.183133Z digest=sha256:febbc1383f892f1c564341007f2e191fe18fb8478e11de1f364ee276b1f63899

Observation b1b1e6f1-ac7f-48aa-bf61-77fe883a442c · inbound

Think, then Score: Decoupled Reasoning and Scoring for Video Reward Modeling cites this paper.

Think, then Score: Decoupled Reasoning and Scoring for Video Reward Modeling RewardDance: Reward Scaling in Visual Generation

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:42:30.786573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T07:37:52.346280Z digest=sha256:0cac3366967eda4bde98a828d38ac6eb96ba8dcb67eb7b037bb48711972c7eac

Observation 18e9cbc2-15da-49eb-b302-ae263df069f0 · inbound

TMPO: Trajectory Matching Policy Optimization for Diverse and Efficient Diffusion Alignment cites this paper.

TMPO: Trajectory Matching Policy Optimization for Diverse and Efficient Diffusion Alignment RewardDance: Reward Scaling in Visual Generation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.181018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T07:36:16.811765Z digest=sha256:cafeac6656bd8b63e284051f18fbda324b82716fbf0ccf425f2e227123df456a

Observation 89a90e18-c6af-43be-ab91-cb40c0d2ff94 · inbound

TMPO: Trajectory Matching Policy Optimization for Diverse and Efficient Diffusion Alignment cites this paper.

TMPO: Trajectory Matching Policy Optimization for Diverse and Efficient Diffusion Alignment RewardDance: Reward Scaling in Visual Generation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:03:02.863836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-14T22:01:21.695270Z digest=sha256:d405b023dc6864a0f98d24084b0b7fd025cc7c0cb501067f37c820170ca3431b

Observation a3dbd52d-cab0-45e8-b875-8a9872158d8c · inbound

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating cites this paper.

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating RewardDance: Reward Scaling in Visual Generation

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:02:22.029828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T06:00:31.582714Z digest=sha256:fe031d6f14cdc30eb2b7c90a57ed06537b86d6e741969c8e07bce1186a39545a

Observation 68a52997-103e-422e-81c8-d98091d48fc1 · inbound

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating cites this paper.

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating RewardDance: Reward Scaling in Visual Generation

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:55:45.300168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T22:38:43.102769Z digest=sha256:2ad684e48ae7d0fb17ac1fde5f5c9c8c9856cf639198364983d90a443c92268c

Observation dadec1d0-7167-44be-bfb7-631db441c849 · inbound

DiffusionOPD: A Unified Perspective of On-Policy Distillation in Diffusion Models cites this paper.

DiffusionOPD: A Unified Perspective of On-Policy Distillation in Diffusion Models RewardDance: Reward Scaling in Visual Generation

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-01T14:25:46.853546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T21:20:23.420424Z digest=sha256:5e5cd7757a2401629b9b5389e625c0cdb93d1048febca39200c214d672ce730c

Observation cfaf8c48-414d-495c-b4fb-a7cb715e6ad4 · inbound

Edit-GRPO: A Locality-Preserving Policy Optimization Framework for Image Editing cites this paper.

Edit-GRPO: A Locality-Preserving Policy Optimization Framework for Image Editing RewardDance: Reward Scaling in Visual Generation

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:47:45.836625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T20:44:54.356658Z digest=sha256:514043967ba59705ab58fe052a50e59fa81df180e3501063246a87c22f25a54e

Observation bc19bdb5-f87f-4e3f-8b2b-b1779a76eed8 · inbound

StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement cites this paper.

StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement RewardDance: Reward Scaling in Visual Generation

Reference 118

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:25:59.831028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T22:45:58.629263Z digest=sha256:a8eb3deed1928b429d28aa9f9eb8266ffd7c0e512e9c07312e9b32e957291362

Observation 485db53f-e8ef-480a-bb21-91d596b0bdc5 · inbound

Improving Visual Representation Alignment Generation with GRPO cites this paper.

Improving Visual Representation Alignment Generation with GRPO RewardDance: Reward Scaling in Visual Generation

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T19:32:35.129428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T19:24:09.597127Z digest=sha256:019c9c4eb7b6eaefa6c2f6f96524597ef5db87ccde694e42f02a92ae1ed38685

Observation 876aee65-1061-4fbc-b3d5-b22c4f1ad8aa · inbound

Are we really tilting? The mechanics of reward guidance in flow and diffusion models cites this paper.

Are we really tilting? The mechanics of reward guidance in flow and diffusion models RewardDance: Reward Scaling in Visual Generation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:26:18.153780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T15:20:52.980868Z digest=sha256:56bc4f9512d9e4486bba640b54eb1a070cff7852f20bfeeb77b2cd491023ee46

Observation 1c45d939-a4dd-476c-9e57-1990a9cbc875 · inbound

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions cites this paper.

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions RewardDance: Reward Scaling in Visual Generation

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:07:27.944548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T17:30:57.001021Z digest=sha256:410ca6579eef824f73f96dc5c27890eecd3ffd166688bfb52e028bfa5d1e63f0

Observation 6e797adc-88a9-454b-89af-7a97da480d64 · inbound

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions cites this paper.

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions RewardDance: Reward Scaling in Visual Generation

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-15T10:53:37.186361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T10:53:37.186361Z digest=sha256:11d77ee12e967a8e43d4c1f4a0671821dc72b1cbf7289ac7928c6d0a2c88f8a3

Observation 494cc52e-0a45-4538-8d3d-0487d70aa0d5 · inbound

Each Judge Its Own Yardstick: Discovering Per-VLM Taxonomies for Physical Video Evaluation cites this paper.

Each Judge Its Own Yardstick: Discovering Per-VLM Taxonomies for Physical Video Evaluation RewardDance: Reward Scaling in Visual Generation

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:09:45.413050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T09:04:23.965554Z digest=sha256:4fd4f72c15ad09271ebc55a475921bc88bbfb4826609b5a923296cb2fd85d2a2

Observation 0b5713ab-4f56-45d7-9d42-b547b47d54e2 · inbound

DiT-Reward: Generative Representations for Text-to-Image Reward Modeling cites this paper.

DiT-Reward: Generative Representations for Text-to-Image Reward Modeling RewardDance: Reward Scaling in Visual Generation

Reference 118

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:39:45.925822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T08:35:04.038960Z digest=sha256:e625cf5732448b19e22c5acfd586af1b5178c58c1655581c88d32db548c1d47e

Observation 47831f1c-db81-48f8-b061-18d80119d225 · inbound

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning cites this paper.

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning RewardDance: Reward Scaling in Visual Generation

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:03:51.757351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T05:00:04.037802Z digest=sha256:d1872ca2af8d598497377a62f1466dc92602778130cfdaeb6264b427594813db

Observation a16b5dda-0a53-42b1-9613-000a8c0c17a9 · inbound

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning cites this paper.

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning RewardDance: Reward Scaling in Visual Generation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-12T11:38:13.566600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:38:13.566600Z digest=sha256:ff1b9aef2a09c6f7754a3c9d82a9e9a0903e8dd8bb24c29e74187e67138f88d9

Observation c62e0581-9e0c-43ee-b72f-9b3456248803 · inbound

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning cites this paper.

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning RewardDance: Reward Scaling in Visual Generation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T09:56:25.929049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T09:56:25.929049Z digest=sha256:df86c9abb96b585d14c05bc2eeed8ed5ae498803c268f39ba0421cf8570aab03

Observation 99a15030-df7d-4c09-b32d-139e4a24111f · inbound

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation cites this paper.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation RewardDance: Reward Scaling in Visual Generation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.594851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.594851Z digest=sha256:2c42072ed3636d0b0b1fb375d1b6a0979ce9e9e3219e7a0dc37da483526c48ca

Observation 25d55f6b-63ee-43f3-b764-9312d1880f41 · inbound

SciForma: Structure-Faithful Generation of Scientific Diagrams cites this paper.

SciForma: Structure-Faithful Generation of Scientific Diagrams RewardDance: Reward Scaling in Visual Generation

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-01T16:12:23.899000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T16:12:23.899000Z digest=sha256:6941f952a8eec2638b7dca93ce516f246b7c76cb6344dc2dfa4172c7bc516e1f