Pith. sign in

Paper Citation Record · LEDGER

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning

As of 18 August 2026, this Paper Citation Record lists 73 of 73 outbound references and 0 inbound Pith citation observations for arXiv:2607.03903.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.03903 v1

Coverage vector

measured 73 of 73 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-11T23:08:40.265657Z

measured 73 of 73 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

73 of 73 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved73
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 04215f3e-cbd5-4350-87d1-73ea8106ee6a · outbound

This paper cites Multi-task reinforcement learning in humans,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Multi-task reinforcement learning in humans,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:71cab6b9b5c164d9b35cb27dd73e60b29679596f26abb800ea01db5a0b08e289

Observation 78568ca2-ebcf-4eb2-b629-442bf8354ade · outbound

This paper cites Curriculum-based asymmetric multi-task reinforcement learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Curriculum-based asymmetric multi-task reinforcement learning,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:c49c1254328470e45ea57fc79007f70977d909a979bf1feb9c4e1e330188387c

Observation e8401553-28ba-4030-a384-fdecf499a8e6 · outbound

This paper cites Multi-task reinforcement learning with soft modularization,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Multi-task reinforcement learning with soft modularization,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:697a6015d57d3284d4140165cc275bfcfb51fc08fa88effa78a488d7f0a77d1d

Observation 85014611-7711-4779-a668-2f8ded09b135 · outbound

This paper cites Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:55bed116d79e17cb371d3cbfca4d678f2cdc6518542ab640b90289f1c2b8478d

Observation eca7f036-0ecf-486d-8e84-da5a1b94307d · outbound

This paper cites Lifelong robotic reinforcement learning by retain- ing experiences,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Lifelong robotic reinforcement learning by retain- ing experiences,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:46c8a141809b4274d621b67c03a0d1a18a3d1229f6a24faa65f8e8c0142ec494

Observation f0cde9cc-16e7-4061-a377-db73341ac6e4 · outbound

This paper cites Multi-task batch reinforcement learning with metric learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Multi-task batch reinforcement learning with metric learning,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:df6ab669740661eeefd3d3c28a342f5e435512130baca776af77e01a46f36812

Observation 7011936b-dbaf-4b89-9134-67177c8cf823 · outbound

This paper cites A discrete soft actor-critic decision-making strategy with sample filter for freeway autonomous driving,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning A discrete soft actor-critic decision-making strategy with sample filter for freeway autonomous driving,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:3fd5c26256017fbe5f725759fc31cd1d692bbb6453ca933f7ae2f8454eeb53b0

Observation 535df503-54ef-418a-a2e0-e8df750a437b · outbound

This paper cites Longitudinal speed control of autonomous vehicle based on a self-adaptive pid of radial basis function neural network,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Longitudinal speed control of autonomous vehicle based on a self-adaptive pid of radial basis function neural network,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:0e744ebc9cc63ce68fccc2424108631042f8dfbfa5959adee2697a4c6c16c63b

Observation d36a1702-7b0e-46cc-8e83-b534d062bb31 · outbound

This paper cites Skills regularized task decomposition for multi-task offline reinforcement learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Skills regularized task decomposition for multi-task offline reinforcement learning,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:adc7b035f8bf8679ff6e1327dc171c9a177ec99ffc18d5626e1687a70f62f614

Observation a3bb215a-fff6-4d4e-a7b4-40d26ce5967a · outbound

This paper cites Gnfactor: Multi-task real robot learning with generalizable neural feature fields,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Gnfactor: Multi-task real robot learning with generalizable neural feature fields,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:b23d6a8f33e685e7cdfb8ea8a45e21ca9d708f047333d8030036ae85684402b5

Observation 37bb97f4-f969-4205-a369-03a9161b1150 · outbound

This paper cites Multi-task learning with attention for end-to-end autonomous driving,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Multi-task learning with attention for end-to-end autonomous driving,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:3dc3a4a7d92ce817f81103fda556417c43b9f70bfcf2105b1619c8f09eabd27a

Observation e8db3227-1a37-4aaf-be21-081a44d449a1 · outbound

This paper cites Aoi-aware resource allocation for platoon-based c-v2x networks via multi-agent multi-task reinforcement learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Aoi-aware resource allocation for platoon-based c-v2x networks via multi-agent multi-task reinforcement learning,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:ce4ddb851d17ff938d9ff5cc10b2627e0ac049cfaf770d59a4015ea8dd9c23d0

Observation ad6befa9-10eb-4c8f-adfb-a5eeb7f5d7fe · outbound

This paper cites A multi-task-learning- based transfer deep reinforcement learning design for autonomic optical networks,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning A multi-task-learning- based transfer deep reinforcement learning design for autonomic optical networks,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:ddda094c775e65040ef0558868e897c66e67d7abe75e05d9fc4a6c4bbba7b2ba

Observation f357e30b-a7f7-4a1e-bec8-0906be7d7d0e · outbound

This paper cites Multi-task deep reinforcement learning with popart,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Multi-task deep reinforcement learning with popart,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:b545de97dcf9a57326b37fa311c40576a6856fb9bcd32eb9148051989ad3d57a

Observation e2dd1870-e2a1-4184-b4e8-62efd9255ec0 · outbound

This paper cites A survey of multi-task deep reinforcement learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning A survey of multi-task deep reinforcement learning,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:3001984bff3a5781ab2453e9906843d5d5de47a99d441073460ccdea05be6828

Observation a9a12da8-f332-4892-afb0-91690f5e9322 · outbound

This paper cites Provably efficient multi-task reinforcement learning with model transfer,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Provably efficient multi-task reinforcement learning with model transfer,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:3ba957fe59b8d8a45938fba0d5ee5d4219f204d23d48ae5e8a1d425747b2c84a

Observation 432fc9e5-ea89-46eb-9e2f-71f78d66c57a · outbound

This paper cites Toward trustworthy decision-making for autonomous vehicles: A robust reinforcement learning approach with safety guarantees,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Toward trustworthy decision-making for autonomous vehicles: A robust reinforcement learning approach with safety guarantees,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:6382f1f4f7c9578cb996dace7336b152e1b08896ebb34a53cbcd77974cb30854

Observation ec10fa17-d8fc-4038-8c3a-2f4508b5ad15 · outbound

This paper cites Residual policy learning facilitates efficient model-free autonomous racing,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Residual policy learning facilitates efficient model-free autonomous racing,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:f1e34bc0aa6187aaf54aab0c8aeb26e8cd3bc56fc4547169f7f2f68ef0a45997

Observation ed90efe3-db99-47b7-80bd-bc7de3155f43 · outbound

This paper cites Metadrive: Composing diverse driving scenarios for generalizable reinforcement learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Metadrive: Composing diverse driving scenarios for generalizable reinforcement learning,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:809d623c464f7b7626600f59acca8a777d99451dcedeb06cec4606690c539ccd

Observation 410eba7c-13b2-46d5-8036-3c62bb6d0dda · outbound

This paper cites Safe reinforce- ment learning for arm manipulation with constrained markov decision process,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Safe reinforce- ment learning for arm manipulation with constrained markov decision process,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:a204301fc1ea7c7ed323c4d56bab4b3de897432ed3b21b06906fc4a17ff07ddd

Observation a262a2f5-2a6d-4917-af90-3a0dc6bb5548 · outbound

This paper cites Safe exploration in reinforcement learning: A generalized formulation and algorithms,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Safe exploration in reinforcement learning: A generalized formulation and algorithms,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:a854e8b3ba9ea036ff325f4718ef0a6e5b091be6ba8396a3ec6f55c00cda23e2

Observation 63bce3eb-9a31-4a2b-aa39-349ec3226858 · outbound

This paper cites Constraints penalized q-learning for safe offline reinforcement learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Constraints penalized q-learning for safe offline reinforcement learning,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:f3cdbe708d59f6047e54ca67ade7680b3b5e36f4a1ca1c06ecef19baf93ed0dc

Observation 50f5dd15-e308-499a-904a-30b27a8c02de · outbound

This paper cites Poce: Primal policy optimization with conservative estimation for multi-constraint offline reinforcement learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Poce: Primal policy optimization with conservative estimation for multi-constraint offline reinforcement learning,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:bd5bbd1699206815068f42d528caf7c4533f8b93ad156170d2243421186676a3

Observation a707e98a-f7d7-4455-9fec-02180bf40a8b · outbound

This paper cites COptiDICE: Offline Constrained Reinforcement Learning via Stationary Distribution Correction Estimation.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning COptiDICE: Offline Constrained Reinforcement Learning via Stationary Distribution Correction Estimation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:e483682d15feda67f1087d15e176a6a4c9f1cef5400262ccc6db4ea1bf24f97f

Observation 22bffaf7-f2e7-45fc-bb6e-8303d334c4da · outbound

This paper cites Constrained offline policy optimization,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Constrained offline policy optimization,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:abbfddeaa8e9c179053f99097dd5cb187b2e0f0d45e058815d0dde71a9df74ea

Observation d8c62032-9432-4079-9e18-cfd32453242f · outbound

This paper cites VOCE: Variational optimization with conservative estimation for offline safe reinforcement learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning VOCE: Variational optimization with conservative estimation for offline safe reinforcement learning,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:eab5ee1f57ccc18482d885982eb53d218a28e0c754e384b241e54e936cfbdffe

Observation 8ba410fd-1e3b-4a29-970b-ee61f7dceb71 · outbound

This paper cites Sharing knowledge in multi-task deep reinforcement learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Sharing knowledge in multi-task deep reinforcement learning,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:6819e2d16386e0dff2e62f3192ffa8c378a07110cfea1d443dd89c1584f79671

Observation bd2a67a6-bb02-442c-a4d2-910be17c46db · outbound

This paper cites Optimization of deep rein- forcement learning with hybrid multi-task learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Optimization of deep rein- forcement learning with hybrid multi-task learning,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:b8d421c782bc0675e98c8893049916fb8b90ffb99abf250c91420b9d48116343

Observation f6a90f3c-1ff9-4842-be12-236436203226 · outbound

This paper cites Conservative data sharing for multi-task offline reinforcement learn- ing,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Conservative data sharing for multi-task offline reinforcement learn- ing,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:dd44fb1eb1d3e52954a6522889dbbee636f49287cfc14beb4d12e15acbd52aa4

Observation 041c4b91-1984-4753-8405-209afc81834a · outbound

This paper cites Multi-task reinforcement learning with context-based representations,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Multi-task reinforcement learning with context-based representations,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:44f48ac49c34b530ea0fe49217fa371c6ef9e3c580d92f238c47bfa147d9ae3f

Observation 1f08d56f-cf93-419b-bf43-2cd9065ce8a4 · outbound

This paper cites Just pick a sign: Optimizing deep multitask models 13 with gradient sign dropout,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Just pick a sign: Optimizing deep multitask models 13 with gradient sign dropout,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:b00541d6bb457015c596a70e0ea48e3e64e48f8c0ed91d3733d07d9fbd7efcae

Observation 3e90d6ca-8eaa-4595-a6b8-93c9793c159f · outbound

This paper cites Conflict-averse gradi- ent descent for multi-task learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Conflict-averse gradi- ent descent for multi-task learning,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:cdf7852b6c39aec9c3a92d8f4abd00a459e145601e6b46b26d784e91d3937d36

Observation 72639a85-8282-4e0f-b770-af0412fdcfd0 · outbound

This paper cites Multi-task reinforcement learning with soft modularization,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Multi-task reinforcement learning with soft modularization,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:a4471332de8807ff3b66919a0f1a9000d51b4b75ca9f4d9ef31bbe2a3ade7053

Observation 07039ef1-2186-463f-8601-900f8bd706cb · outbound

This paper cites Paco: Parameter- compositional multi-task reinforcement learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Paco: Parameter- compositional multi-task reinforcement learning,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:f69e2c125b65e959420fb540a912e47cc85af8a7b1944cc2e0eef2e9666f9285

Observation cba640cc-e760-4ed9-ae9a-8d03c92ca955 · outbound

This paper cites Recomposing the reinforcement learning building blocks with hypernetworks,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Recomposing the reinforcement learning building blocks with hypernetworks,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:dd546be6a8167fbfc921ebf4bd3f2dbd3770ccb58ed27cb37b2cd18126bb5001

Observation 9d6386d1-c764-4e8a-ac9a-e88b7048e76c · outbound

This paper cites Batch policy learning under con- straints,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Batch policy learning under con- straints,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:989439a1fe02bb9c2bfb81639977c32130a4805292d291db49a306df1ce04f9b

Observation 01b17824-fda4-4304-ba42-1011f98c43d8 · outbound

This paper cites Gendice: Generalized offline estimation of stationary values,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Gendice: Generalized offline estimation of stationary values,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:8780245d6756852b5a477f58013b58854e0aaaf3292b677631d0e8364a83e8a6

Observation 1d467ab3-19b5-418c-9c3b-9c3f119e17a8 · outbound

This paper cites Coindice: Off-policy confidence interval estimation,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Coindice: Off-policy confidence interval estimation,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:e242e348647c0464d7dba65ae0f173da39202b5bc5b6593c4ab468323e00dbde

Observation 63db115a-f3ff-4da7-86cd-0a051b679802 · outbound

This paper cites Optidice: Offline policy optimization via stationary distribution correction estimation,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Optidice: Offline policy optimization via stationary distribution correction estimation,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:3198cf13a67584f303949b231a0ea204e8526b8cf202fc66768a3d356d553e16

Observation 88b23f5a-8ef6-4337-bbd7-3ca1270b5bb3 · outbound

This paper cites Coptidice: Offline constrained reinforcement learning via stationary distribution correction estimation,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Coptidice: Offline constrained reinforcement learning via stationary distribution correction estimation,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:020accc2f058ffa3de0aec8fdbc096c1e9843568e34b0f344fba5386e95a118b

Observation 408dd544-0752-4cee-86e3-fcea984b2612 · outbound

This paper cites Con- strained decision transformer for offline safe reinforcement learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Con- strained decision transformer for offline safe reinforcement learning,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:665dc28c00f3373b01c1e8e84e4248a4142afde68b9cd922bcc7554076a500c8

Observation 261e6b07-d454-4bdc-b184-dd007f6f6389 · outbound

This paper cites SaFormer: A Conditional Sequence Modeling Approach to Offline Safe Reinforcement Learning.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning SaFormer: A Conditional Sequence Modeling Approach to Offline Safe Reinforcement Learning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:bde0ec439bcaeadefd2df71a098ec6ed64d77c84ad8c8bcbe17ffadaa6508d17

Observation 359ceb8b-ec38-48b3-b69e-99e2da34c471 · outbound

This paper cites Scaling up multi-task robotic reinforcement learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Scaling up multi-task robotic reinforcement learning,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:4f2115a087cf62fcd609d704a1816ba87e0ca09e0a922046313264006488d061

Observation 6c55a0ba-910f-4450-a2c6-f6142620ac55 · outbound

This paper cites Perceiver-actor: A multi-task transformer for robotic manipulation,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Perceiver-actor: A multi-task transformer for robotic manipulation,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:d2735555f1e63477ad9de523ae5804e4867d8a53235d5d036027a1170f004ede

Observation 51b76292-c5c2-4e61-8e2d-da4d55fae113 · outbound

This paper cites Effective adapta- tion in multi-task co-training for unified autonomous driving,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Effective adapta- tion in multi-task co-training for unified autonomous driving,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:fa746a5b499eb78a82a7466467816fc9c2548c92962812cfb37886d7c21b73c4

Observation e85b19b2-0c02-4001-a53a-16d4e8c33ebe · outbound

This paper cites Learning Shared Representations in Multi-task Reinforcement Learning.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Learning Shared Representations in Multi-task Reinforcement Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:d5d32aeb6d440507beaefb5f9326766d6734a6dd6670059ded6332cd161fd574

Observation fca99087-09fa-4cdc-be77-6356f9c207a9 · outbound

This paper cites Multi-task reinforcement learning with attention-based mixture of experts,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Multi-task reinforcement learning with attention-based mixture of experts,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:7f8456d0dfd5cadfc7a70e1c367faad47e6e8aba3cd5ae7577f1e8cb7cba082b

Observation dcb0bde3-8120-4158-8291-d522ed1947e4 · outbound

This paper cites Shar- ing knowledge in multi-task deep reinforcement learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Shar- ing knowledge in multi-task deep reinforcement learning,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:0fd8d3fff3e6c72ef1b4623018f6f482259d3e3acd77d5518d4b1c62c26a344a

Observation a23cecb9-da15-4f56-bcd0-1f997388c2d1 · outbound

This paper cites Impala: Scalable dis- tributed deep-rl with importance weighted actor-learner architectures,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Impala: Scalable dis- tributed deep-rl with importance weighted actor-learner architectures,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:cc88d2111f5eb47da0f5ec1a8a7923c1805a77d51c669e1d5ea20c76b2b891ee

Observation 71e56d1a-4a0f-4976-9896-25059432d99c · outbound

This paper cites Multi-Goal Reinforcement Learning: Challenging Robotics Environments and Request for Research.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Multi-Goal Reinforcement Learning: Challenging Robotics Environments and Request for Research

Reference 50

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:3fb307d212bbcfe1ff3ee04ec84ef38a8f9bd3bedbbc4b183151b0396d81ba51

Observation 0906c6cc-10fc-4c38-8215-bbc8217065aa · outbound

This paper cites Asynchronous methods for deep rein- forcement learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Asynchronous methods for deep rein- forcement learning,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:b57b30c5172ab627bd0bdec366a0d45f0c949f82e550153e0114fcdf2e864786

Observation dd044dfc-433f-4600-92a0-71c95af95851 · outbound

This paper cites Deterministic policy gradient algorithms,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Deterministic policy gradient algorithms,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:d5f048fb25301b543ed42c72e0c31709d90bb371e3734922e1bf1838aef8f7c5

Observation 0117f10b-a10a-4084-af08-890c7a2919cf · outbound

This paper cites Sharing knowledge in multi-task deep reinforcement learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Sharing knowledge in multi-task deep reinforcement learning,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:a6c2e9997b06a6c78ff6fa665960ef71fba13f39cffd9be7c1ed0abcc4910759

Observation 172b9bf1-720f-45e8-b4ac-7a499fa9a65c · outbound

This paper cites Multi-game decision transformers,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Multi-game decision transformers,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:7e406e5f79c974cd2723051434a055781fd55018642eed5318867a3b0ea309ba

Observation 874cb008-cd0a-4220-a280-7ee59b1d64a9 · outbound

This paper cites Diffusion model is an effective planner and data synthesizer for multi-task reinforcement learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Diffusion model is an effective planner and data synthesizer for multi-task reinforcement learning,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:06f8a59cf702836882ed1bc0800b95013e8f7feb3a4dfd600aaa69624816588f

Observation 7f5ff869-a440-4920-9332-faa881bf7fb0 · outbound

This paper cites Altman,Constrained Markov decision processes: stochastic modeling.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Altman,Constrained Markov decision processes: stochastic modeling

Reference 56

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:7697f69d4859e2689190bd59ed3abd7b2b74528dd808687bfd45b68d695a73d8

Observation 54aabce8-3c12-40d9-8c2f-22b2c1b44b20 · outbound

This paper cites Conservative q-learning for offline reinforcement learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Conservative q-learning for offline reinforcement learning,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:b787ded5a520ddcf1c6d7767af5432207537c8664d14fd97f02530755031a043

Observation 55e9c038-06c9-4532-b8e1-3b36e574fd41 · outbound

This paper cites Uac: Offline reinforcement learning with uncertain action constraint,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Uac: Offline reinforcement learning with uncertain action constraint,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:90180e5553293590e5feebd92cd3b143492c33beb89b79979476ec04d49e110a

Observation 438511c9-900e-43b2-9145-0fbcb758174d · outbound

This paper cites Generative modeling by estimating gradients of the data distribution,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Generative modeling by estimating gradients of the data distribution,

Reference 59

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:9cd40c2d8e73a9cc1b3b979e9fe5ef3e9ef3c1fa02f881750f14fbf4df1f1041

Observation b34c99d7-b0b7-4a2a-b573-88ead452794b · outbound

This paper cites Denoising diffusion probabilistic models,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Denoising diffusion probabilistic models,

Reference 60

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:65cb3e669b0facc6e58ecc28dc79e078fe2dc8c4b35997524f4c25a604e2e0c9

Observation c88020b8-d7b2-497d-8191-89017e501692 · outbound

This paper cites Variational inference: A review for statisticians,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Variational inference: A review for statisticians,

Reference 61

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:c8405a141d2ad0fe0a7ae86db04100c1d3025274af3c929a832686898ff18487

Observation b2aa7cf6-2102-4bca-9ddc-b2afb73c66d1 · outbound

This paper cites Efficient diffusion policies for offline reinforcement learning,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Efficient diffusion policies for offline reinforcement learning,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:e45008b248e3c984d5cc685001e32dd48599989ee4e677a0612c7e16862ee25b

Observation 2088c7e4-4445-47ea-90d5-6f4693828239 · outbound

This paper cites Is conditional generative modeling all you need for decision-making?.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Is conditional generative modeling all you need for decision-making?

Reference 63

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:b8998950f4fad84c1be28fd64886ed10e068efb87fc7b25360f60583ff3ee6c2

Observation 09d3b324-2deb-46b5-917f-bc4b34bc58a7 · outbound

This paper cites Efficient learning of safe driving policy via human-ai copilot optimization,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Efficient learning of safe driving policy via human-ai copilot optimization,

Reference 64

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:ddc478e369139f01223066fc14e6158fdc789f292f225da5eff2e2902ae82daa

Observation 8c49a55d-b920-49b0-8f79-caf907222248 · outbound

This paper cites Safe driving via expert guided policy optimization,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Safe driving via expert guided policy optimization,

Reference 65

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:3a92951241a43f5040b7019438ec230753898f7f80cda29efcbf5df22f4345f4

Observation f53a6df2-0e63-48bd-96f6-af342ceb79b9 · outbound

This paper cites Denoising diffusion probabilistic models,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Denoising diffusion probabilistic models,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:3bac89f08c0c7fbe3549d135758a941c66c7969f037783858c6a9126fb89e4d1

Observation 0c43d1b9-4e05-453a-a078-1e62b90a8aa4 · outbound

This paper cites Gradnorm: Gradient normalization for adaptive loss balancing in deep multitask networks,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Gradnorm: Gradient normalization for adaptive loss balancing in deep multitask networks,

Reference 67

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:62163c7e41b1e1415893f77917bb1ed21686a2a0a77c2260323ded874a4cea40

Observation 3c08cf8e-55d0-414e-b187-1e2135b5b822 · outbound

This paper cites Datasets and Benchmarks for Offline Safe Reinforcement Learning.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 68

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:b5de1aadee460fff159d9d64674bbdc66ec01741e3f41e1055e97b58eef5b737

Observation bdd159d0-4bfc-44fc-9bdd-cae2bc960446 · outbound

This paper cites Learning to simulate self-driven particles system with coordinated policy optimization,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Learning to simulate self-driven particles system with coordinated policy optimization,

Reference 69

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:f49e8a6822e91af7ef4afdabc91bbc471b1711bc2b6c9b36c3b9407bb4ead200

Observation 46128661-48c8-405d-a494-c4536aa2b14b · outbound

This paper cites Understanding Diffusion Models: A Unified Perspective.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Understanding Diffusion Models: A Unified Perspective

Reference 70

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:3e4422c842a45eb5936bf95508968ace9ebdee8b587d8a9fc4d4701d12d11846

Observation b1dfffcf-03fc-4dbc-93b4-ae055fee408a · outbound

This paper cites Denoising diffusion implicit models,.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Denoising diffusion implicit models,

Reference 71

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:c69914709deb6c49151b784f4778216ef37d7c431f3f15865851564701dff3c5

Observation 0c153df2-0c33-49fe-b757-f5384b2ed6f4 · outbound

This paper cites Table 2 outlines the primary network parameters of the model implemented by our CDCP algorithm.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Table 2 outlines the primary network parameters of the model implemented by our CDCP algorithm

Reference 72

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:3651754914977876b7f904610caf7f067923fd543edb4e202ca684748c3a50ef

Observation 0e01d0a9-bb72-4baf-a8da-d7ba5a9ae430 · outbound

This paper cites These scenarios include environments such as curves, intersections, T-junctions, and roundabouts.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning These scenarios include environments such as curves, intersections, T-junctions, and roundabouts

Reference 73

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:1d127cbf86d68a1f7153525eabbe72e90fff70efe6cc9f1b187726e19e8a66c3

Pith citing papers

No inbound Pith citation observations are available.