Pith. sign in

Paper Citation Record · LEDGER

Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2406.04314.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.04314 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T18:55:14.721267Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T06:39:38.330848Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 10994389-e781-44da-9703-e15bf6d91809 · inbound

Improving Video Generation with Human Feedback cites this paper.

Improving Video Generation with Human Feedback Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:30:02.732151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T15:30:02.578430Z digest=sha256:bf9ce2a85de0498c33288f8b2b8f03e81834d420ec6c6f1f016a510f1b9f6b96

Observation 7ff31df5-0330-4e06-81aa-130ba46954b0 · inbound

Refining Alignment Framework for Diffusion Models with Intermediate-Step Preference Ranking cites this paper.

Refining Alignment Framework for Diffusion Models with Intermediate-Step Preference Ranking Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T18:55:14.721267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:55:14.721267Z digest=sha256:256b09dcaa312f6958f5c920fc9d71635a4e3c042b45c3d5ea635f367c42cd03

Observation 46008356-5e90-4b3f-92b1-4650c19aa5da · inbound

BalancedDPO: Adaptive Multi-Metric Alignment cites this paper.

BalancedDPO: Adaptive Multi-Metric Alignment Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-22T23:42:16.586536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T23:37:55.154902Z digest=sha256:36b923c7db26abc7f6cfb39e617f6a883e42810b9e525958be64cf045c145dd7

Observation 8701c302-a574-4a62-be56-481ec6092fd8 · inbound

Flow-GRPO: Training Flow Matching Models via Online RL cites this paper.

Flow-GRPO: Training Flow Matching Models via Online RL Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:45:16.971344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T18:45:16.641012Z digest=sha256:3a3d332310ab43fc68f2660a3014d6f5bdaab143844b7417e96278de7853f959

Observation 22a8705e-4e9f-4749-a376-3564740d0eea · inbound

DanceGRPO: Unleashing GRPO on Visual Generation cites this paper.

DanceGRPO: Unleashing GRPO on Visual Generation Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:28:28.248699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T22:28:24.929046Z digest=sha256:08d3958892dada04dd8612e78e577afaca585d1b5b68561ed569d0f5df847481

Observation 38dbfa4d-b793-48dc-a7c5-21f504a0b8cd · inbound

Scaling Image and Video Generation via Test-Time Evolutionary Search cites this paper.

Scaling Image and Video Generation via Test-Time Evolutionary Search Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:49:45.309170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:49:45.309170Z digest=sha256:7ed1670142dfe78fa92c7c34deff7b64a6ba8fc4b4f944bd73df57ec4ae29d8a

Observation 7ddc2152-916d-4b1e-b35e-103cb7023a17 · inbound

Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning cites this paper.

Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:26.144841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:26.144841Z digest=sha256:bb1dc210cdfce5885123d0c0f178ba3793e04a835d1beade6dc27a36b8899d6c

Observation c9b8effa-4baa-4c73-a76e-2eaec840d3e6 · inbound

Align-DA: Align Score-based Atmospheric Data Assimilation with Multiple Preferences cites this paper.

Align-DA: Align Score-based Atmospheric Data Assimilation with Multiple Preferences Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 1926

Resolution
unresolved
no resolver link, observed 2026-08-07T13:25:05.411402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:25:05.411402Z digest=sha256:f5e9c07113dc67bd147c83cf852c201b17ca4be0c695dd46971bf5252e2ed6ca

Observation 5c4bb27f-0cfe-4b52-9382-bcb7a4776523 · inbound

ImageReFL: Balancing Quality and Diversity in Human-Aligned Diffusion Models cites this paper.

ImageReFL: Balancing Quality and Diversity in Human-Aligned Diffusion Models Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:40.253118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:09:40.253118Z digest=sha256:10c0bacbd3fc854915082a4067a5481a2d1fa6e0c76880a69dbd423c0bf69421

Observation 04fd1a4d-7037-4407-a05a-267e226ce8c3 · inbound

Enhancing Diffusion-based Unrestricted Adversarial Attacks via Adversary Preferences Alignment cites this paper.

Enhancing Diffusion-based Unrestricted Adversarial Attacks via Adversary Preferences Alignment Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:00.600957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:47:00.600957Z digest=sha256:89b2f49b8252a3f0eaab4c38606f5ed189d72bc196c66101b75fe6f9614a2513

Observation 7b4eefa4-3d83-4a02-bd26-529ca2228b3a · inbound

GigaVideo-1: Advancing Video Generation via Automatic Feedback with 4 GPU-Hours Fine-Tuning cites this paper.

GigaVideo-1: Advancing Video Generation via Automatic Feedback with 4 GPU-Hours Fine-Tuning Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T04:29:38.396893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:29:38.396893Z digest=sha256:c82f47367a5fb34ab21e5f272d1f748c15413658f878002421d1bbdb478f7bf9

Observation 5ee3f4d7-01c7-4f6d-a7d5-70ed869f45d7 · inbound

Fake it till You Make it: Reward Modeling as Discriminative Prediction cites this paper.

Fake it till You Make it: Reward Modeling as Discriminative Prediction Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T00:30:41.767912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:30:41.767912Z digest=sha256:8f9afbea48e5c9e9c438be9e2bb9686ada6a5b7144b8a651760c518cfe72b4fc

Observation ab1c7732-9397-4f3f-8d79-0f73a6069d34 · inbound

Inversion-DPO: Precise and Efficient Post-Training for Diffusion Models cites this paper.

Inversion-DPO: Precise and Efficient Post-Training for Diffusion Models Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T17:54:45.840939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:54:45.840939Z digest=sha256:9caef6e1f76d77a1ed6b1d5ae7162fa0a871dc6d986f5b42796e14dd5f80543d

Observation 6649138f-042a-4dc8-9a84-75eda51b6484 · inbound

X-Omni: Reinforcement Learning Makes Discrete Autoregressive Image Generative Models Great Again cites this paper.

X-Omni: Reinforcement Learning Makes Discrete Autoregressive Image Generative Models Great Again Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T12:10:08.345491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:10:08.345491Z digest=sha256:1a44de509db86da36756b981c4167ba9ef45b77585b44805b042fb202fbae457

Observation f277eed4-bd71-49b1-b5b2-b9c8abba75c7 · inbound

ShortFT: Diffusion Model Alignment via Shortcut-based Fine-Tuning cites this paper.

ShortFT: Diffusion Model Alignment via Shortcut-based Fine-Tuning Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T11:32:41.949449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:32:41.949449Z digest=sha256:6178ce13ad9f0948d59a0178715d1be68ec2ede7867f63e527a997ff44690abb

Observation 3481ddd6-cc35-43e0-9868-f0e26542ff32 · inbound

FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation cites this paper.

FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T20:06:47.388338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:06:47.388338Z digest=sha256:0defa6e6f70816f687ffa9cdbb68d433562687000fca02f22ba6adf25b93c6eb

Observation 9e3f21ad-9081-4c4f-b264-8e22dd71db05 · inbound

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance cites this paper.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.280733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.280733Z digest=sha256:b89d2c4ab65852090789cdc37db8fe9deb99e233ddd6cf7b923a7b78e17fe978

Observation 5cc7a5a7-ad25-4918-8313-ab7ea734846f · inbound

Towards General Preference Alignment: Diffusion Models at Nash Equilibrium cites this paper.

Towards General Preference Alignment: Diffusion Models at Nash Equilibrium Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:56:06.868203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T16:54:58.732444Z digest=sha256:a8bfdac297f9dfd411afea59a185da5604d19a85705607191d332287e8fc8994

Observation 3b0eff0b-43ac-43ac-ab1c-0133c052b09e · inbound

Offline Preference Optimization for Rectified Flow with Noise-Tracked Pairs cites this paper.

Offline Preference Optimization for Rectified Flow with Noise-Tracked Pairs Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T02:11:15.534956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T02:10:27.595446Z digest=sha256:968f9bd7ddf89974a2f6cdaf60762d28d9c0e2984ed934f847e5a291bda70e96

Observation c77785a1-63fa-4b32-874d-9484b904a438 · inbound

Do Less, Achieve More: Do We Need Every-Step Optimization for RL Fine-tuning of Diffusion Models? cites this paper.

Do Less, Achieve More: Do We Need Every-Step Optimization for RL Fine-tuning of Diffusion Models? Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:48:57.326419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T19:48:17.049547Z digest=sha256:a6ed5ce718fd01bd9facd13f1f605b8e0c529f4cf387eca3b0171db36cb3601e

Observation 440fc27a-7370-4ef9-83a2-1ef7205665ef · inbound

Balancing Performance and Diversity in GRPO Autoregressive Text-to-Image Post-Training cites this paper.

Balancing Performance and Diversity in GRPO Autoregressive Text-to-Image Post-Training Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:39:38.332570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T14:14:45.547957Z digest=sha256:48ade18d98cdf26c0b6280faf6ed427a8dc52afcfe453477489d799e91bd0445