Pith. sign in

Paper Citation Record · LEDGER

Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2405.20555.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.20555 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T21:39:25.530151Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T04:17:36.912565Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c8cfc308-5b73-4999-aa3f-40b86b9361f0 · inbound

Behavior-Regularized Diffusion Policy Optimization for Offline Reinforcement Learning cites this paper.

Behavior-Regularized Diffusion Policy Optimization for Offline Reinforcement Learning Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T21:39:25.530151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T21:39:25.530151Z digest=sha256:77d607ef6feb36f203732f4799f4a12a489be0771c4358764d20ca23315d9ac4

Observation ca30f5ef-759b-47ad-97b6-76f6c65d8374 · inbound

FlowQ: Energy-Guided Flow Policies for Offline Reinforcement Learning cites this paper.

FlowQ: Energy-Guided Flow Policies for Offline Reinforcement Learning Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:49.396954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:49.396954Z digest=sha256:31d6b727b5fa2603d370b444319dc84c1cb53ece24998af841faff2203086603

Observation 385977de-291c-443e-9bf2-9b99ae46f638 · inbound

EXPO: Stable Reinforcement Learning with Expressive Policies cites this paper.

EXPO: Stable Reinforcement Learning with Expressive Policies Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-19T05:12:05.237520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T05:09:02.111308Z digest=sha256:0f5df9f2f3ae43251e853733596fc2d0bcee058b8798e591459222ff78a17969

Observation fd393471-b746-42dc-8718-6db43ac05ee8 · inbound

Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning cites this paper.

Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:41:49.076925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T02:17:25.783688Z digest=sha256:8aa6a946681b524d5f40c8b2b4bfb40de65f5170faf4b6319b973081254e13b4

Observation 96c4c744-305e-43fd-85f9-7dba25b841cb · inbound

Reinforcement Learning for Flow-Matching Policies with Density Transport cites this paper.

Reinforcement Learning for Flow-Matching Policies with Density Transport Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T22:27:25.842430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T18:55:02.040180Z digest=sha256:1a0f7346371020fa01d9cdc44f9bd8b70ee6e8cc149ebccfdea929245219e23c

Observation e71de311-483e-478e-a2ed-7fc2a0324cfb · inbound

Test-Time Gradient Guidance of Flow Policies in Reinforcement Learning cites this paper.

Test-Time Gradient Guidance of Flow Policies in Reinforcement Learning Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T04:17:36.914070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T14:05:01.073951Z digest=sha256:d27c67a631c950b68651b587689b94c69084d83bca475d686e103300ffcc3961