Pith. sign in

Paper Citation Record · LEDGER

D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

As of 10 August 2026, this Paper Citation Record lists 4 of 4 outbound references and 7 inbound Pith citation observations for arXiv:2508.06804.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.06804 v1

Coverage vector

measured 4 of 4 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T22:37:38.873857Z

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T05:50:32.796809Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

4 of 4 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved3
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 47af0d31-d22e-4a23-b863-c54299debc16 · outbound

This paper cites an unresolved cited work.

D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:37:38.937359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T22:37:38.870533Z digest=sha256:e0ec9a18423009751af5ac32416ecbe5b878d8cba1a60611204439b9a5289cfc

Observation 8485382e-27fc-4013-a240-d932e75c77e5 · outbound

This paper cites Obs dim (State).

D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning Obs dim (State)

Reference 2

Resolution
malformed identifier
raw_fallback, observed 2026-08-05T22:37:38.926673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T22:37:38.873857Z digest=sha256:b0897e32e980f871f4f893595ea7cd2c6870a292e67ea31f40b9cb2cd490889f

Observation e1c9652e-a846-42bb-b6cc-536c9be01cd3 · outbound

This paper cites Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning.

D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-05T22:37:38.862346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:37:38.862346Z digest=sha256:991837d01cfa50ceb5df7ebb92cbaba91db58008f1375fd13d60150c14306e3e

Observation e420a4f3-8086-455a-98ff-da8b22e1a528 · outbound

This paper cites Denoising Diffusion Implicit Models.

D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning Denoising Diffusion Implicit Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-05T22:37:38.866648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:37:38.866648Z digest=sha256:131027d7c2ede3334e0e792f24233c871c1a50ad302c7d33869877486ae65d64

Pith citing papers

Observation 91e47940-5acb-4039-b67d-68e1ce95c787 · inbound

Generative Actor-Critic with Soft Bridge Policies cites this paper.

Generative Actor-Critic with Soft Bridge Policies D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:46:28.347246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T04:01:08.896356Z digest=sha256:570892d8db219eb515d74a6ab1efa2d0b37ad9375a70bb0c9d3a29929a0eea0c

Observation 8f257ea9-f95f-44df-82b2-31a7586db65d · inbound

SANTS: A State-Adaptive Scheduler for World Action Models cites this paper.

SANTS: A State-Adaptive Scheduler for World Action Models D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:03:24.032257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T11:57:36.200662Z digest=sha256:ed50d7034f44e5dd59cf1ca97f07b9d4baf5bf732a3e284ddb317a87fc0aee16

Observation c5857eb3-2e15-494e-beaa-806a16ba233d · inbound

SANTS: A State-Adaptive Scheduler for World Action Models cites this paper.

SANTS: A State-Adaptive Scheduler for World Action Models D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-15T11:04:29.593725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T11:04:29.593725Z digest=sha256:734619d687711ce5d0c5fd05c11a0e4666079f63ebeb099b9e91b1f4f9c07a03

Observation 8c454ebb-7545-4570-ba9b-f2e777b59bbc · inbound

On-Device Robotic Planning: Eliminating Inference Redundancy for Efficient Decision-Making cites this paper.

On-Device Robotic Planning: Eliminating Inference Redundancy for Efficient Decision-Making D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:26:00.972117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T22:25:23.391802Z digest=sha256:bf185bad69311771ebc6579213a66ab099abf06fece03104807d0058423bfc57

Observation 6f1066a6-1ca8-40b6-a273-ad19b3f0396d · inbound

L-SDPPO: Policy Optimization of Spiking Diffusion Policy for Intra-vehicular Robotic Manipulation cites this paper.

L-SDPPO: Policy Optimization of Spiking Diffusion Policy for Intra-vehicular Robotic Manipulation D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:06:59.097365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T01:35:03.841581Z digest=sha256:d486a3fdd47f8ead642f168348960947c33ded1de9d7cbf1670a6f860706c7ef

Observation 039fcb9d-5d4b-43ca-b5c0-989b55dfd01b · inbound

ELASTIC: Efficiently Learning to Adaptively Scale Test-Time Compute for Generative Control Policies cites this paper.

ELASTIC: Efficiently Learning to Adaptively Scale Test-Time Compute for Generative Control Policies D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-01T05:45:25.222884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-01T05:44:43.682828Z digest=sha256:8da58bc1c0c4c69e62963fa1bca3d6595aac32a3ea8db1b44e7a9bf5046f3fa6

Observation f86f061b-1113-4748-8b5d-c12bbe3b94c6 · inbound

Learning When to Stop: Prefix-Optimal Dynamic Diffusion Policies for Continuous Control cites this paper.

Learning When to Stop: Prefix-Optimal Dynamic Diffusion Policies for Continuous Control D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T05:50:32.796809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:50:32.796809Z digest=sha256:0da867ebff67af572b15289889b236e07f38f21125b33f10c6bbfb250762ccd1