Pith. sign in

Paper Citation Record · LEDGER

Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2405.20555.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.20555 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T21:39:25.530151Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T04:17:36.912565Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c8cfc308-5b73-4999-aa3f-40b86b9361f0 · inbound

Behavior-Regularized Diffusion Policy Optimization for Offline Reinforcement Learning cites this paper.

Behavior-Regularized Diffusion Policy Optimization for Offline Reinforcement Learning Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T21:39:25.530151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T21:39:25.530151Z digest=sha256:54fb1ad4e0138b8ecfeef2cd9e97539a78525356513e277bdaffb975d050f162

Observation ca30f5ef-759b-47ad-97b6-76f6c65d8374 · inbound

FlowQ: Energy-Guided Flow Policies for Offline Reinforcement Learning cites this paper.

FlowQ: Energy-Guided Flow Policies for Offline Reinforcement Learning Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:49.396954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:49.396954Z digest=sha256:31d6b727b5fa2603d370b444319dc84c1cb53ece24998af841faff2203086603

Observation 385977de-291c-443e-9bf2-9b99ae46f638 · inbound

EXPO: Stable Reinforcement Learning with Expressive Policies cites this paper.

EXPO: Stable Reinforcement Learning with Expressive Policies Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-19T05:12:05.237520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T05:09:02.111308Z digest=sha256:2962c536c8128e12bd71219173133f7a6f491705532d6473f2a1ee84da8417f9

Observation fd393471-b746-42dc-8718-6db43ac05ee8 · inbound

Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning cites this paper.

Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:41:49.076925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-12T02:17:25.783688Z digest=sha256:93a15158b3b0f50cae6554f9de405f5142dec705093c9688b57c8b61db971310

Observation 96c4c744-305e-43fd-85f9-7dba25b841cb · inbound

Reinforcement Learning for Flow-Matching Policies with Density Transport cites this paper.

Reinforcement Learning for Flow-Matching Policies with Density Transport Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T22:27:25.842430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T18:55:02.040180Z digest=sha256:5d62dd99b207fa6860502c3cc57f7fd6b9f758ac698326956df3a9056486eb6d

Observation e71de311-483e-478e-a2ed-7fc2a0324cfb · inbound

Test-Time Gradient Guidance of Flow Policies in Reinforcement Learning cites this paper.

Test-Time Gradient Guidance of Flow Policies in Reinforcement Learning Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T04:17:36.914070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T14:05:01.073951Z digest=sha256:4a5e54b4240c4ac56368324d0a5f8f08829874fc65e8177e6097e544a4203e30