Pith. sign in

Paper Citation Record · LEDGER

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning

As of 15 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 5 inbound Pith citation observations for arXiv:2502.01819.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.01819 v3

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T14:29:12.477304Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T16:44:36.837980Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T18:38:53.000576Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact1
  • verified fuzzy3
  • unresolved30
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9ad45e5c-fefb-4c69-b1fe-a6645a2c7d85 · outbound

This paper cites Training Diffusion Models with Reinforcement Learning.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Training Diffusion Models with Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:11.815314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:11.815314Z digest=sha256:130059f8ffa96e52fd09f320175130a77cfa3c1772bee95ea50828714f95264a

Observation 8908ba33-e341-4015-9331-03746b77f918 · outbound

This paper cites Sampling is as easy as learning the score: theory for diffusion models with minimal data assumptions.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Sampling is as easy as learning the score: theory for diffusion models with minimal data assumptions

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:11.949366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:11.949366Z digest=sha256:63b1d4dedbff09e3224656eeb31abff7949b7eb53fdf75b43038f5de73785155

Observation 6e0fa41b-a1e7-49e7-9910-ac1c14611548 · outbound

This paper cites Optimizing DDPM Sampling with Shortcut Fine-Tuning.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Optimizing DDPM Sampling with Shortcut Fine-Tuning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.102481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.102481Z digest=sha256:75f7827d26dca14d37c1c62f9919d63cf56c1d4dd7318fd1258facf41dd50ae1

Observation 9d31d82f-5625-46b9-8dbc-719afcb9dcc9 · outbound

This paper cites DPOK: Reinforcement Learning for Fine-tuning Text-to-Image Diffusion Models.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning DPOK: Reinforcement Learning for Fine-tuning Text-to-Image Diffusion Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.106970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.106970Z digest=sha256:eddc4d7c6571967628b4b7f561b4dfefa124d31ada1dd2ee0faf12a7c0cdbdda

Observation 7b140705-6248-492b-a55b-9936f0285c30 · outbound

This paper cites Reward-Directed Score-Based Diffusion Models via q-Learning.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Reward-Directed Score-Based Diffusion Models via q-Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.111582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.111582Z digest=sha256:5c65930b0249499bdfbbe32c895b00ecd586edea7524f11e0bccc58b6fe89e5a

Observation 8c78c11f-a17a-4084-bb14-1053003b7e95 · outbound

This paper cites Optimizing Prompts for Text-to-Image Generation.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Optimizing Prompts for Text-to-Image Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.116193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.116193Z digest=sha256:ffbc58ae486bcfff88c0c6cf862a500c54ebd91c7c2d66b0a12809c8263042d3

Observation 318bb216-e0e6-40c9-ab84-9732adb0c691 · outbound

This paper cites an unresolved cited work.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-09T14:29:13.477786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-09T14:29:12.128862Z digest=sha256:48ce75001ac80d2efd9b15bf119df698c6e4fe137c87db13ea1bd39c19ce3da6

Observation 188b076e-de43-4217-8cb8-be549dd71121 · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.137957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.137957Z digest=sha256:b7e6c0b2ee8d2cccb808c31dd28f4a2d13b5fa167bdc44521a899f974295f6f8

Observation bd203227-8bd3-4fd9-a681-d2e6ad307b98 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.142340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.142340Z digest=sha256:2571d72d069e510ae82b91b3bb9a15311920f092e2f6ef057b3faf5034cfe813

Observation 13bba21e-6c31-49c1-a0a5-6c9196c3db35 · outbound

This paper cites Diffusion Policy Policy Optimization.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Diffusion Policy Policy Optimization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.146870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.146870Z digest=sha256:e9f35c13c9953bf15e9ab5caea97ef98627afaba7a5012b1121488e8f773ab4a

Observation b655a89a-93fb-4e3d-840d-8e4b5978885c · outbound

This paper cites Progressive Distillation for Fast Sampling of Diffusion Models.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Progressive Distillation for Fast Sampling of Diffusion Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.151666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.151666Z digest=sha256:1d98ea5d3e05458f7dcbad0951b560f2044ef99983094aa38895d0cb22d8f89d

Observation f1e74cfb-e291-45f9-a336-5b54ba8f2d88 · outbound

This paper cites Improving Image Captioning with Better Use of Captions.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Improving Image Captioning with Better Use of Captions

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.282159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.282159Z digest=sha256:7e28da5c7070a27b552f6e88ef906ad6943dab3f3ed584efbe76d352d6c418a4

Observation cefcffbe-160d-4714-8c1b-bc54d8e64a5e · outbound

This paper cites Denoising Diffusion Implicit Models.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Denoising Diffusion Implicit Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.359251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.359251Z digest=sha256:fe106f6872f1867f42369517635423774c80b04188f4e8fab4e3040c6aec620e

Observation 2eed287d-df1d-4f4f-b63e-981d1047143d · outbound

This paper cites Solving Inverse Problems in Medical Imaging with Score-Based Generative Models.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Solving Inverse Problems in Medical Imaging with Score-Based Generative Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.412419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.412419Z digest=sha256:08fcf2374e99bc0d6728b09155642d89f69d313be3b0601625e98934b6744301

Observation 547aceb9-289b-4e0c-8fd9-df5eac3c48b9 · outbound

This paper cites Score-based Diffusion Models via Stochastic Differential Equations -- a Technical Tutorial.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Score-based Diffusion Models via Stochastic Differential Equations -- a Technical Tutorial

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.423240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.423240Z digest=sha256:28b6c96fde4a9c4f9b049dbcce4c7ca5d3ce92fbf717558fe4f8a3487d852afd

Observation 68fbedcb-4df2-4a7d-8e21-43ba93e5d1a7 · outbound

This paper cites Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.427168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.427168Z digest=sha256:7efa3aacacf9e8a10e1c5ee3aa3edcd8e73145a0201849a6dbfae9145f9dad42

Observation 71d5a541-b706-4bb9-9431-0d7e780f0a85 · outbound

This paper cites Preference Tuning with Human Feedback on Language, Speech, and Vision Tasks: A Survey.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Preference Tuning with Human Feedback on Language, Speech, and Vision Tasks: A Survey

Reference 26

Resolution
metadata mismatch
local_arxiv, observed 2026-08-09T14:29:12.924859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-09T14:29:12.431395Z digest=sha256:98a42ba0b05cf6308137a157db8f9d166be405d0d6629afce19d4c37ea514f91

Observation 155c2cbe-35a1-495b-9c67-dcdddbc2b032 · outbound

This paper cites GeoDiff: a Geometric Diffusion Model for Molecular Conformation Generation.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning GeoDiff: a Geometric Diffusion Model for Molecular Conformation Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.435649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.435649Z digest=sha256:b2759faa26937724c8677ec5d9cf9d33b617d15d6d147b892cf0cac1de85f273

Observation 5a7f53eb-6642-4d76-9d21-89c598da81b0 · outbound

This paper cites Maximum Entropy Inverse Reinforcement Learning of Diffusion Models with Energy-Based Models.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Maximum Entropy Inverse Reinforcement Learning of Diffusion Models with Energy-Based Models

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-09T14:29:12.769275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-09T14:29:12.439940Z digest=sha256:eaa966f62e6b01887108cc71f2635257e31b5daefee25bc4c54233b2dcb531d0

Observation 00abb90a-3f36-4d01-a015-2da311d0e575 · outbound

This paper cites Self-Play Fine-Tuning of Diffusion Models for Text-to-Image Generation.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Self-Play Fine-Tuning of Diffusion Models for Text-to-Image Generation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.444589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.444589Z digest=sha256:90909cd5f0487d19296bb15a7128874944c22a079e55eb391dd4fa433babbb8b

Observation 6625498d-78cf-4fa1-9122-0b04808df995 · outbound

This paper cites Fast Sampling of Diffusion Models with Exponential Integrator.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Fast Sampling of Diffusion Models with Exponential Integrator

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.448838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.448838Z digest=sha256:391b8609f1c19bb095bd1199a1e3186ae9b146ab6d69413ba893ab1126d9adc3

Observation a6e1ead3-d69c-4063-80ae-077782e590de · outbound

This paper cites gDDIM: Generalized denoising diffusion implicit models.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning gDDIM: Generalized denoising diffusion implicit models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.454146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.454146Z digest=sha256:558d3e8ab928be2f8fae8dc892c90d6fa074f83af50bbf5a5c1274f326b6a8e4

Observation c27a3190-1fc2-4762-aa19-1da2beda719d · outbound

This paper cites Scores as Actions: a framework of fine-tuning diffusion models by continuous-time reinforcement learning.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Scores as Actions: a framework of fine-tuning diffusion models by continuous-time reinforcement learning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.459033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.459033Z digest=sha256:103d8f56bbdd2d0525f70f2db83310dee6601bf7242c7364710ea879155b61e0

Observation fd453468-5b45-4d12-9523-fdef139f399e · outbound

This paper cites predicted x0.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning predicted x0

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:29:13.463637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-09T14:29:12.464031Z digest=sha256:a8d34a5fc80ef47563ce497b204fb4e6e711a79cff1638e2c92e1ce2c1465417

Observation 561c8ab2-6f62-410f-9e1a-749fb9e94e0b · outbound

This paper cites predicted x0.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning predicted x0

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:29:13.447926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-09T14:29:12.468113Z digest=sha256:8db2b31bde7a141a182d2b00c66cf3da123ad1ccb058dfa7b580ab659d422965

Observation c14b68ca-c56b-414b-8899-dbbe163014b8 · outbound

This paper cites We adopt two ways of derivations: (a) Notice that, if we treat the ˆxθ (t, Xt) as a constant in (38) (or assume that it does not change w.r.p.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning We adopt two ways of derivations: (a) Notice that, if we treat the ˆxθ (t, Xt) as a constant in (38) (or assume that it does not change w.r.p

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T14:29:13.432515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-09T14:29:12.477304Z digest=sha256:bb1f0614ae3567bdf5bd5b6879d404a89a16513231b50de648f94ae8b3acdd2c

Observation 23460e6b-4a91-4898-9901-f3cdde308fc1 · outbound

This paper cites Fine-tuning of diffusion models via stochastic control: entropy regularization and beyond.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Fine-tuning of diffusion models via stochastic control: entropy regularization and beyond

Reference 1999

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.419252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.419252Z digest=sha256:ae33a6fd63815fc8362b1782a4fa786f686a73ff313a2d9e4d4866e334f5fba1

Observation 99790938-35c3-4b29-af97-657a3b53ebff · outbound

This paper cites Aligning Text-to-Image Models using Human Feedback.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Aligning Text-to-Image Models using Human Feedback

Reference 2009

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.133392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.133392Z digest=sha256:12640c8eeb6b6f30151d2e37098f6272c1ecd195c1cdcdea618980517a24aa30

Observation f003bc6f-d8b6-4a98-aeef-f1bde2c43243 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.192879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.192879Z digest=sha256:5efc2fa10d91425278158a81e993c82143818212e8d71c7cac0820e24733f1a6

Observation 198a0a69-340c-42b8-8ad4-d38c508af818 · outbound

This paper cites Directly Fine-Tuning Diffusion Models on Differentiable Rewards.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Directly Fine-Tuning Diffusion Models on Differentiable Rewards

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.019456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.019456Z digest=sha256:fc406c61780be4c487607d7fc841d8aa58c75f76b392cb0f5d63f12b61650deb

Observation 5097f194-b6d8-435f-a81e-d766c142c8af · outbound

This paper cites Imagen Video: High Definition Video Generation with Diffusion Models.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Imagen Video: High Definition Video Generation with Diffusion Models

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.120433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.120433Z digest=sha256:a4b9529a73e6f347595774f416ddf10f54cf97952e4586c28ade7e509db77d52

Observation d922e02d-ad8e-4824-bfa9-30962aeefab4 · outbound

This paper cites Adjoint Matching: Fine-tuning Flow and Diffusion Generative Models with Memoryless Stochastic Optimal Control.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Adjoint Matching: Fine-tuning Flow and Diffusion Generative Models with Memoryless Stochastic Optimal Control

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.091436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.091436Z digest=sha256:2f801d8a77c8219f8b6cf3bb9b101e89ca6a68f7d3f84acb5dfeb0a4f2aa90cb

Observation 9aa508f7-2248-432b-b9c5-1b476b2cb3ea · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning LoRA: Low-Rank Adaptation of Large Language Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.125068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.125068Z digest=sha256:3fb986fa5e457ba27676d28984f283ab090f9de7323ae4c07dc36ea1fd554e46

Observation 27a024c8-43f9-48a7-85ab-26a0c86d5cb3 · outbound

This paper cites Sparks of Artificial General Intelligence: Early experiments with GPT-4.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Sparks of Artificial General Intelligence: Early experiments with GPT-4

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:11.879830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:11.879830Z digest=sha256:c5eb918ad1e44cd1edda7d2e5c8859c19d19c91f441ad94f25ff03a211121e79

Observation 8082898e-1517-48f2-a77b-5a873d46bf2d · outbound

This paper cites Scaling Rectified Flow Transformers for High-Resolution Image Synthesis.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Scaling Rectified Flow Transformers for High-Resolution Image Synthesis

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.096528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.096528Z digest=sha256:ba85217577f073fcd4296325dbf3a75ed656c7f56373151a530267f87da3e069

Pith citing papers

Observation 4e45631e-d946-4b7a-b80d-c6d1cbbcb9e7 · inbound

Flow-GRPO: Training Flow Matching Models via Online RL cites this paper.

Flow-GRPO: Training Flow Matching Models via Online RL Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T18:45:16.728520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-11T18:45:16.641012Z digest=sha256:f8a3ec638dc470456ee517c083385de4184a3041eb5eb338786da2a340a71c27

Observation 6829bc90-f340-4ee5-89d3-ea9de10cd911 · inbound

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models cites this paper.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.837980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.837980Z digest=sha256:1bd0e1ed87338f139f74163823b798ed62659e8ccea85f51314b174137b6f859

Observation 6c6aee69-8f73-4a82-999d-d86651382aeb · inbound

Improving Text-to-Image Generation with Intrinsic Self-Confidence Rewards cites this paper.

Improving Text-to-Image Generation with Intrinsic Self-Confidence Rewards Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning

Reference 89

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T18:41:28.922983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-15T18:40:18.940889Z digest=sha256:7c2bb39a99a55bb026f40fabf86b20317091987f9ee9f213d3426f8364c4f9af

Observation 24c327ab-edde-46d3-8a47-fe166e38ee92 · inbound

Energy Generative Modeling: A Lyapunov-based Energy Matching Perspective cites this paper.

Energy Generative Modeling: A Lyapunov-based Energy Matching Perspective Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:36:08.216028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-09T16:05:00.985070Z digest=sha256:95f804ff86d8d026caeca9356e9bf913aa5dda67f6e90ac2ef376ad9b67333ac

Observation 0e4474a6-2ce3-46b3-9e89-59faa507c1ae · inbound

Embedding-perturbed Exploration Preference Optimization for Flow Models cites this paper.

Embedding-perturbed Exploration Preference Optimization for Flow Models Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-20T18:38:53.002168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-20T18:33:52.933672Z digest=sha256:f7b83ccfce640ff4cc19cce785231d7dcfa95969d1c9fe7fa90ff9cae03e2507