Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-25T05:21:52.748214Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 1 inbound Pith citation observation for arXiv:2605.23365.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-25T05:21:52.748214Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-10T14:36:10.359049Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T14:37:15.953758Z
33 of 33 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e16b1662-262e-4316-8611-47a11a2f0a76 · outbound
Score-Based One-step MeanFlow Policy Optimization Is Conditional Generative Modeling all you need for Decision-Making?
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation fe2a9ef8-634b-4dca-a98e-0ff79b082d4e · outbound
Score-Based One-step MeanFlow Policy Optimization Iterated denoising energy matching for sampling from boltzmann densities
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e835e503-70fd-45ff-b5ee-bae68fe6b10f · outbound
Score-Based One-step MeanFlow Policy Optimization Score regularized policy optimization through diffusion behavior
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation dac36487-d07e-49c8-9cf7-6fcfb7e5e17d · outbound
Score-Based One-step MeanFlow Policy Optimization Diffusion policy: Visuomotor policy learning via action diffusion
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c6566601-a61d-4164-a854-fefbedd48cc9 · outbound
Score-Based One-step MeanFlow Policy Optimization Diffusion-based reinforcement learning via q-weighted variational policy optimization
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c4fac339-7b83-440c-9ee2-ed955f3510f0 · outbound
Score-Based One-step MeanFlow Policy Optimization One step diffusion via shortcut models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 4aebde4d-8f20-4261-b539-b27e5d06e10a · outbound
Score-Based One-step MeanFlow Policy Optimization Mean flows for one-step generative modeling
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 80ca5bd5-9953-4cc2-b317-dd9e1246883c · outbound
Score-Based One-step MeanFlow Policy Optimization Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e145d41d-209f-4da1-aa6a-88860d1d8024 · outbound
Score-Based One-step MeanFlow Policy Optimization Denoising diffusion probabilistic models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation fb889358-7f67-45a3-a65c-8679d722f27d · outbound
Score-Based One-step MeanFlow Policy Optimization Planning with Diffusion for Flexible Behavior Synthesis
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b6ac2a15-732d-442c-b150-1d29aa4051a7 · outbound
Score-Based One-step MeanFlow Policy Optimization Prior-guided diffusion planning for offline reinforcement learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 492b20f6-9de9-4291-bda5-490666b1ca5e · outbound
Score-Based One-step MeanFlow Policy Optimization Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 07029e77-d08d-4d2a-bac3-0bbb973597dc · outbound
Score-Based One-step MeanFlow Policy Optimization Flow straight and fast: Learning to generate and transfer data with rectified flow
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d8c9407d-4691-446e-8b3d-fdb98f9f369a · outbound
Score-Based One-step MeanFlow Policy Optimization Simplifying, stabilizing and scaling continuous-time consistency models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e791643d-1a99-4c84-a9fc-2adffdbb1539 · outbound
Score-Based One-step MeanFlow Policy Optimization Efficient online reinforcement learning for diffusion policy
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0b95ac95-a5d2-49c9-ab48-91457004b53e · outbound
Score-Based One-step MeanFlow Policy Optimization Learning a diffusion model policy from rewards via q-score matching
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation dfd762db-b4d8-4092-912e-22a5e0bae85f · outbound
Score-Based One-step MeanFlow Policy Optimization Diffusion Policy Policy Optimization
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d6c1805b-c865-41ad-a931-fde69f2eb2e1 · outbound
Score-Based One-step MeanFlow Policy Optimization Progressive distillation for fast sampling of diffusion models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b8694883-1729-4fa4-a207-7653430550c3 · outbound
Score-Based One-step MeanFlow Policy Optimization Proximal Policy Optimization Algorithms
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f15f3bab-bf9b-4d07-bee2-f661700fecd3 · outbound
Score-Based One-step MeanFlow Policy Optimization DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7333592f-794c-4f02-bfc1-5aca43672304 · outbound
Score-Based One-step MeanFlow Policy Optimization Denoising Diffusion Implicit Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 763f04c7-b259-44b1-b986-4d16d979675a · outbound
Score-Based One-step MeanFlow Policy Optimization Consistency models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 8b9826e1-7c37-4249-807e-6aacddd5eddb · outbound
Score-Based One-step MeanFlow Policy Optimization Generative modeling by estimating gradients of the data distribution.Advances in Neural Information Processing Systems, 32
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 724ac68d-ab5e-4319-a08a-376234bc98d7 · outbound
Score-Based One-step MeanFlow Policy Optimization Score-based generative modeling through stochastic differential equations
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 589c7b6b-05c1-41fd-a090-a312414a711e · outbound
Score-Based One-step MeanFlow Policy Optimization MIT press Cambridge
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0d1e467d-bd2a-4a05-968b-fa0232021a69 · outbound
Score-Based One-step MeanFlow Policy Optimization Mujoco: A physics engine for model-based control
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2fb51365-d222-4e43-98f5-7ba26eb10f64 · outbound
Score-Based One-step MeanFlow Policy Optimization Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 8ccdbf9e-525a-4d85-b8e7-04ee27ca024f · outbound
Score-Based One-step MeanFlow Policy Optimization Diffusion actor-critic with entropy regulator
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6fcb2bcd-c38f-441e-b0a7-ca7b2faff9cb · outbound
Score-Based One-step MeanFlow Policy Optimization Diffusion policies as an expressive policy class for offline reinforcement learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c49ab6de-e540-40d5-97ee-f5641be5a5e2 · outbound
Score-Based One-step MeanFlow Policy Optimization Policy Representation via Diffusion Probability Model for Reinforcement Learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e1970c2d-ee3b-4ff8-a082-f1acb63f6d56 · outbound
Score-Based One-step MeanFlow Policy Optimization One-step diffusion with distribution matching distillation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1ffc8f5f-c9b4-4a12-86ba-e4dc119b479b · outbound
Score-Based One-step MeanFlow Policy Optimization Mean flow policy with instantaneous velocity constraint for one-step action generation
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 23a9e2cc-d14d-47f3-95c9-589d4565155f · outbound
Score-Based One-step MeanFlow Policy Optimization The final reward is given by the normalized mixture density, producing a smooth multimodal reward landscape with values in[0,1]
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c8575dd3-97fd-4132-989d-5617a15fd9d5 · inbound
Expressivity and Statistical Trade-offs in Diffusion Policy Learning Score-Based One-step MeanFlow Policy Optimization
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.