Pith. sign in

Paper Citation Record · LEDGER

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL

As of 21 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 2 inbound Pith citation observations for arXiv:2505.15791.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15791 v2

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:18:56.930347Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:21:12.263047Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T11:25:18.912960Z

Reference resolution

38 of 38 outbound references displayed

  • verified exact1
  • verified fuzzy4
  • unresolved33
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fa398351-ccc8-4599-a4fd-834ab65439bc · outbound

This paper cites Atom level enzyme active site scaffolding using rfdiffusion2.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Atom level enzyme active site scaffolding using rfdiffusion2

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:18:59.113875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T15:18:52.040527Z digest=sha256:b77a825bd4e133bda43e6a0a4d2cce365f6fe2c73bab0d12928836ff79a0217a

Observation d2c7f045-5afd-4730-9113-14e3e8151be9 · outbound

This paper cites Out of Many, One: Designing and Scaffolding Proteins at the Scale of the Structural Universe with Genie 2.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Out of Many, One: Designing and Scaffolding Proteins at the Scale of the Structural Universe with Genie 2

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:53.460195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:53.460195Z digest=sha256:2da1b837f8bdf7508d78ce25f0457edbeba612457fb3fb803d62fa46229c274b

Observation 5cf8b26f-86a2-47e3-979f-ef5b74fc1802 · outbound

This paper cites As referenced in the main text, Table 1 includes metrics from the DRaFT [Clark et al., 2023] and PRDP [Deng et al., 2024] papers.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL As referenced in the main text, Table 1 includes metrics from the DRaFT [Clark et al., 2023] and PRDP [Deng et al., 2024] papers

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:18:58.325029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T15:18:56.835558Z digest=sha256:ca60ac792784a5ad84d3ac5ba908d28cc4284ccf74cded53926b0d87902b9ab2

Observation 74c7321d-08c8-4c1b-aee3-a557fbd4b929 · outbound

This paper cites DiffuSeq: Sequence to Sequence Text Generation with Diffusion Models.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL DiffuSeq: Sequence to Sequence Text Generation with Diffusion Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:52.512770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:52.512770Z digest=sha256:e9786695cd3da3bda9d7747a15f268f3d660eb9c65cb4c3cda470116d44511f1

Observation 3dcc38c7-143c-4d6c-9845-a260871e95c9 · outbound

This paper cites Dealing with Sparse Rewards in Reinforcement Learning.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Dealing with Sparse Rewards in Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:52.713790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:52.713790Z digest=sha256:70fde6025ec57475eac83c463decc2df3148d824d3e14f50fb5483f4164bb63c

Observation 7fb24f74-efbe-459a-ac1b-38278052559d · outbound

This paper cites Sequence-Augmented SE(3)-Flow Matching For Conditional Protein Backbone Generation.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Sequence-Augmented SE(3)-Flow Matching For Conditional Protein Backbone Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:52.921055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:52.921055Z digest=sha256:cf4aa8214db0d21962b663d8baaf6b2fbfd26c85e466b1df16ea8aa8bcd6eeb9

Observation 007c5f7c-0714-4c6f-bc71-fb2220e2149c · outbound

This paper cites Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:52.995998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:52.995998Z digest=sha256:d9352359aa7fe9342621e3280e88535d4fd3fb084f6102460cd3bd6ce56d38a2

Observation 4c3004c2-7ed9-498e-b5a0-46ebcdb07874 · outbound

This paper cites Aligning Text-to-Image Models using Human Feedback.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Aligning Text-to-Image Models using Human Feedback

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:53.080608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:53.080608Z digest=sha256:9e99067b113590fe45cc4c99f69ca63e5395ace9ac5b85a795458a5dbaa9d281

Observation d3fe9064-24ba-441b-ba91-6e7565fbe53c · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:53.255145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:53.255145Z digest=sha256:88f70bac2c7a0cad9b9bd1b5b5d5e59c9b7b43cd604b8226abcd00233aeca680

Observation 778a4dbe-3938-45a6-b3fb-84e31e2ff8b8 · outbound

This paper cites Derivative-Free Guidance in Continuous and Discrete Diffusion Models with Soft Value-Based Decoding.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Derivative-Free Guidance in Continuous and Discrete Diffusion Models with Soft Value-Based Decoding

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:53.356102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:53.356102Z digest=sha256:49537c93ec3a718e16e6e507098e874cf76db8ef55de30ebac3e8ad60d9067b8

Observation 60b756a9-3543-4b66-b7a4-8eb2db9f1a9d · outbound

This paper cites Flow Matching for Generative Modeling.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Flow Matching for Generative Modeling

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:53.574219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:53.574219Z digest=sha256:4237877becb27c2baacfc62d546f69ba7868248ee5704a48a0bb59f51a8f128d

Observation 4c18afcc-799a-43ef-be36-8678337ee544 · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:53.950021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:53.950021Z digest=sha256:0fb9f9fca930bc9077f08c975b70c2d1e8af7e6d6e18341f8683a4fb8e056d13

Observation 51d5e71b-3704-4152-86e4-ff4274939898 · outbound

This paper cites Decoupled Weight Decay Regularization.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Decoupled Weight Decay Regularization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:54.458034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:54.458034Z digest=sha256:c55ad6cfe7993bb27d3158b3b09eabe9d6a05696eba572e83208c05d3c089aee

Observation 6a317298-8807-4244-bf64-974ce588f354 · outbound

This paper cites Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:54.669121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:54.669121Z digest=sha256:c0737a544a250e01ac874c3efc38e170a66c7372a5fa78aa920d41d2d2b1b0c2

Observation 927be3ae-e2e6-4ebf-9010-bb91d10dafed · outbound

This paper cites Reward Model Learning vs. Direct Policy Optimization: A Comparative Analysis of Learning from Human Preferences.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Reward Model Learning vs. Direct Policy Optimization: A Comparative Analysis of Learning from Human Preferences

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:18:57.467666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T15:18:54.854896Z digest=sha256:db29a26f06aadf7db5412b094d07446b6e301eeb97f19cc93a89044d70c42693

Observation b9ee0a31-7f21-4c32-99d2-569451a1693a · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:54.989604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:54.989604Z digest=sha256:b2c2f6d6d432d7b23b357320c61090518a017b8cced2ce42e64ea599f5f5c64b

Observation 9bd3559c-88e7-4184-bfb8-9c50e89dee06 · outbound

This paper cites Multisample Flow Matching: Straightening Flows with Minibatch Couplings.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Multisample Flow Matching: Straightening Flows with Minibatch Couplings

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:55.121823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:55.121823Z digest=sha256:0fcddf3b53fe9dad1e5f666849647714b5cd8f27806f9c002d527b9a7f70ed0a

Observation 99439a9f-a924-4e2a-87d4-1a291c60782e · outbound

This paper cites Proximal Policy Optimization Algorithms.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Proximal Policy Optimization Algorithms

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:55.223435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:55.223435Z digest=sha256:1327d04832e284c7012f0b8bfd78100d689193eca3b0e3c03497f28e47c71bfe

Observation b04cb345-3e1d-490c-a74b-3109af4dde62 · outbound

This paper cites Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:55.607841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:55.607841Z digest=sha256:051a24db9eda118e679051de0fe31824291545051878bf952a88053fa7579007

Observation 4bce29fa-00c9-40b7-b738-d00b81ff2a79 · outbound

This paper cites Diffusion Language Models Are Versatile Protein Learners.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Diffusion Language Models Are Versatile Protein Learners

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:55.707487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:55.707487Z digest=sha256:2c1680199f7c121459c58b5a8e7eed70917e25cda6149ff9b6adc6f25a3b3f88

Observation 1fcef7d8-965e-4abc-b44a-e5aa6c130d7f · outbound

This paper cites Focus-N-Fix: Region-Aware Fine-Tuning for Text-to-Image Generation.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Focus-N-Fix: Region-Aware Fine-Tuning for Text-to-Image Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:55.773131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:55.773131Z digest=sha256:13597e58b00b47c4af1125b2ed4c81757a4951abdaada60fe8934ae0aa77d508

Observation 0aa61977-60da-48fe-9f63-653bd8dbc57c · outbound

This paper cites VisionReward: Fine-Grained Multi-Dimensional Human Preference Learning for Image and Video Generation.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL VisionReward: Fine-Grained Multi-Dimensional Human Preference Learning for Image and Video Generation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:55.855992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:55.855992Z digest=sha256:82190a98c9e30267bf52c0780cf18f5004f5080630d44a3f1080456f218ed04a

Observation e1e16925-1af0-4627-bc8c-5cfc916a002e · outbound

This paper cites SE(3) diffusion model with application to protein backbone generation.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL SE(3) diffusion model with application to protein backbone generation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:56.049830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:56.049830Z digest=sha256:ab692b18fcbeb78225c74ad778118c4e330c7ef419377299d52cc832b7f47232

Observation 4316b84d-e748-4dfb-876f-d098d52ab0f4 · outbound

This paper cites Towards Controllable Diffusion Models via Reward-Guided Exploration.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Towards Controllable Diffusion Models via Reward-Guided Exploration

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:56.162933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:56.162933Z digest=sha256:c7808f9e6a44572a3c16b933c85545a51b07b6556314f89a64a6702e88dcf9ee

Observation 0740abd3-2e50-41dd-81ff-a55d42b6508d · outbound

This paper cites Large-scale reinforcement learning for diffusion models.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Large-scale reinforcement learning for diffusion models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:56.295167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:56.295167Z digest=sha256:ab000b8cd88fc51208c10031f0e82572d554122d54f12b9c82593b18df99b1a1

Observation 5068b278-5f7a-45ef-94e6-75e76d93dd16 · outbound

This paper cites Behavior Proximal Policy Optimization.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Behavior Proximal Policy Optimization

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:56.424024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:56.424024Z digest=sha256:d5051b08dc060387963dbc839371b33e81b3a0c19faf72c859ebacf2f1cc574c

Observation ecf48c83-7fc8-4538-8f51-18b38d8930e7 · outbound

This paper cites an unresolved cited work.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:18:58.916493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T15:18:56.551973Z digest=sha256:aa5363448dc3b081380a09dcd993c22c1413369e1a6d249292e0553fae29c7b9

Observation 0a16a69d-6f85-4532-9735-27775c67e570 · outbound

This paper cites Given these components, the flow matching objective for SO(3) can be formulated as: LSO(3)(θ) = Et∼U (0,1),q(R0,R1),Rt∼ρt(Rt|R0,R1) ∥vθ(t, Rt) − ut(Rt|R0, R1)∥2 SO(3).

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Given these components, the flow matching objective for SO(3) can be formulated as: LSO(3)(θ) = Et∼U (0,1),q(R0,R1),Rt∼ρt(Rt|R0,R1) ∥vθ(t, Rt) − ut(Rt|R0, R1)∥2 SO(3)

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:18:58.731174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T15:18:56.676501Z digest=sha256:cee06bfd4805d60cfc912c303801923c5e45c6c85e88c1278f8aff6ded5ba4e6

Observation c4c3a304-8b95-48f8-aed2-2517f9ea6ec9 · outbound

This paper cites an unresolved cited work.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:18:58.520325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T15:18:56.765412Z digest=sha256:ac77e69d62fa8e0383fc209b2df66c508137f345c11dc2526b4f0f9ad1b49b04

Observation 2c4718b8-dab6-4b0d-8d7b-384c6c591733 · outbound

This paper cites Color intensity correlates with η magnitude (darker corresponds to higher values).

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Color intensity correlates with η magnitude (darker corresponds to higher values)

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:18:58.060287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T15:18:56.930347Z digest=sha256:a57e6071a8305d9d1e121525ae481f8a933f4672c8d6c181e793d2e414c9770c

Observation 638dc0d6-b238-4d53-9b57-91a7984e2bad · outbound

This paper cites Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 1999

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:55.480756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:55.480756Z digest=sha256:25f2a5afa00ffddd2379807fcecc6203f256974cbf49903f3fc2ce767f49677c

Observation abeef27b-e5a6-4091-b5d6-461eee88ee95 · outbound

This paper cites Denoising Diffusion Implicit Models.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Denoising Diffusion Implicit Models

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:55.353442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:55.353442Z digest=sha256:28dcb21572eb9819db8b9a88dada268036e8e1bac7e36071de9e7bc23214857c

Observation 76dcb53b-f88c-45c0-953d-50265d3b067d · outbound

This paper cites Improve Mathematical Reasoning in Language Models by Automated Process Supervision.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Improve Mathematical Reasoning in Language Models by Automated Process Supervision

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:54.567472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:54.567472Z digest=sha256:2f9b91aebb94bc55eca22f574c0e33fe2e73a565fbea0243f89ead58e89ddb79

Observation dc7ce98c-5562-40c5-9f55-55aaeec638a8 · outbound

This paper cites Does RLHF Scale? Exploring the Impacts From Data, Model, and Method.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Does RLHF Scale? Exploring the Impacts From Data, Model, and Method

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:52.823260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:52.823260Z digest=sha256:2d1a2337bb8cf8731809a1aac1c3806671c821072239d9160031c1d5033bb15f

Observation a25b8486-174e-4f1d-8625-e5328860af49 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:52.576998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:52.576998Z digest=sha256:882078fa9a81e9ddda58529c0fa6438ba7209f82340f37191374ddd58bd3a76c

Observation e4f6683b-e3f5-4300-b6ec-303fcbf2bcc3 · outbound

This paper cites SE(3)-Stochastic Flow Matching for Protein Backbone Generation.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL SE(3)-Stochastic Flow Matching for Protein Backbone Generation

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:52.310087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:52.310087Z digest=sha256:499ed69e601d251d510d4e70308c5af74f0ff00a74bff19adfeba4c2530c6c69

Observation 78702dee-54bf-4302-adf0-977392a625cb · outbound

This paper cites Adjoint Matching: Fine-tuning Flow and Diffusion Generative Models with Memoryless Stochastic Optimal Control.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Adjoint Matching: Fine-tuning Flow and Diffusion Generative Models with Memoryless Stochastic Optimal Control

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:52.419212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:52.419212Z digest=sha256:ed01cea3720a520ab03691fca58a7d4414368bed0b312f553b9ad1c7c950d5d1

Observation a473a493-0b83-4763-8618-ad6be55cebf8 · outbound

This paper cites Training Diffusion Models with Reinforcement Learning.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Training Diffusion Models with Reinforcement Learning

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:52.161793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:52.161793Z digest=sha256:21d9ceaecf10bb4bb0f262ec24cf140edb33c42bb66b2284421551562e78596f

Pith citing papers

Observation d255a3be-70f8-42d9-8d24-f07cd6d4d8b4 · inbound

Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models cites this paper.

Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-15T23:21:12.263047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:21:12.263047Z digest=sha256:50722ef24ba8c156240e1b4edc615c62302c05b4a9d7ac6a1abace1fdb4949a8

Observation ba09d5f4-eef9-4d98-8718-36e276d1a124 · inbound

LeapAlign: Post-Training Flow Matching Models at Any Generation Step by Building Two-Step Trajectories cites this paper.

LeapAlign: Post-Training Flow Matching Models at Any Generation Step by Building Two-Step Trajectories VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:25:18.915785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T11:23:28.424453Z digest=sha256:1bbe49a48906e62c7bdee225247ada59a134ee6655d535ca62d3d78b4afdbb4b