Pith. sign in

Paper Citation Record · LEDGER

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies

As of 8 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2505.19337.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19337 v2

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:21:29.435813Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

61 of 61 outbound references displayed

  • verified exact1
  • verified fuzzy45
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ae52ad05-4262-4008-a82f-9c60562da714 · outbound

This paper cites A review of reward functions for reinforcement learning in the context of autonomous driving.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies A review of reward functions for reinforcement learning in the context of autonomous driving

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:36.159398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:24.932108Z digest=sha256:2df36ca9308c3c4a75aa9569e145e2c5fee5c44f38b5fe436b289e8be8e5ffc2

Observation 1495beaf-23bc-4b1f-8e60-3db0b9161aa0 · outbound

This paper cites Hindsight experience replay.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Hindsight experience replay

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:36.051685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:24.991707Z digest=sha256:748881aa0d0211591a9d172b85f484382ede63f3e90f8c66f4530516ba2b4bc9

Observation 9c5b4ea9-7f87-4975-9c69-2d817661fef0 · outbound

This paper cites Offline Goal-Conditioned Reinforcement Learning for Safety-Critical Tasks with Recovery Policy.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline Goal-Conditioned Reinforcement Learning for Safety-Critical Tasks with Recovery Policy

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:21:29.770963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:25.084900Z digest=sha256:ea7958e6b744849128fd2ea8759d9b771f3294002d272642c1457347952280a8

Observation 933a4504-06f1-4922-9b93-77336d51342f · outbound

This paper cites Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.121008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.121008Z digest=sha256:19507f7914db03dd7ba15b46e84083d99a0cb1f0f1989a566b06739945237e7a

Observation d01ee715-bf00-4edb-909b-acfd60bf7f16 · outbound

This paper cites Decision Transformer: Reinforcement Learning via Sequence Modeling.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Decision Transformer: Reinforcement Learning via Sequence Modeling

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.163416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.163416Z digest=sha256:a1f88a825e9a8f6da71c8d7db6e91ec0431a60f8f3301228b57537ee52d1b0c2

Observation 457babd5-5a2a-42f9-b215-a1b260f6e891 · outbound

This paper cites Gymnasium robotics, 2024.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Gymnasium robotics, 2024

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.216093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.216093Z digest=sha256:97b73336067807cc89902a598ea20822ee2cc3c6d631e1437a3defeb670540e2

Observation 8bada3ed-c54a-4ff1-a252-9f16eeec7edf · outbound

This paper cites Contrastive Learning as Goal-Conditioned Reinforcement Learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Contrastive Learning as Goal-Conditioned Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.312416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.312416Z digest=sha256:61e544c59f62ebc9747b5a4f7103203700e1d82d2de953a1f7dd268c178e2100

Observation 5de72c4d-b6bf-4b36-a11f-df3dd372b951 · outbound

This paper cites Safe multi-agent navigation guided by goal- conditioned safe reinforcement learning, 2025.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Safe multi-agent navigation guided by goal- conditioned safe reinforcement learning, 2025

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.948907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:25.373485Z digest=sha256:d543499cd134955fa5cca1937472ffcdf631d7a92ccf08ff07017d33ccf519e0

Observation 7565752c-768f-4597-a351-4e6dc68adcab · outbound

This paper cites Curriculum reinforcement learning for complex reward functions.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Curriculum reinforcement learning for complex reward functions

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.792698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:25.413568Z digest=sha256:1c59787401562ca278757ebbf3472635af4aeb90ab1642a7b0d1e683e76169c7

Observation 4dff962f-8890-4939-bd0c-985c3d70ef3a · outbound

This paper cites Off-policy deep reinforcement learning without exploration.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Off-policy deep reinforcement learning without exploration

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.446509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.446509Z digest=sha256:73b21b09c8359eccb40e500afe7e551a7e0991ab1e14a4413eb2da6a00fe8c73

Observation 90562ba9-61e4-41f6-8a4f-6d9098548bb4 · outbound

This paper cites Integrating domain knowledge for handling limited data in offline RL.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Integrating domain knowledge for handling limited data in offline RL

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.704832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:25.528577Z digest=sha256:228c8f944e098699f91a45342cb9ed80aa46853d40ec3235c0ba53669e4d3d5a

Observation a2365892-1fd6-4786-9f4b-ca7f9256155e · outbound

This paper cites Learning to Reach Goals via Iterated Supervised Learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Learning to Reach Goals via Iterated Supervised Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.585866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.585866Z digest=sha256:7ff05e4a362a6de137430fd2c0d9d9bd1a3fc5f82d5e80351583e046d6882c5c

Observation 9970028b-05b8-44f1-b722-08857b85e611 · outbound

This paper cites Bullet-safety-gym: A framework for constrained reinforcement learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Bullet-safety-gym: A framework for constrained reinforcement learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.636687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.636687Z digest=sha256:36ba8f975a3439b4eec328b9291189dde61630721ec274bf5b47f5aedb8aece4

Observation 9b630f1c-0390-4132-9913-f3b28e69f4c0 · outbound

This paper cites Chemical reprogramming of human somatic cells to pluripotent stem cells.Nature, 605(7909):325–331, May 2022.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Chemical reprogramming of human somatic cells to pluripotent stem cells.Nature, 605(7909):325–331, May 2022

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.618483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:25.685668Z digest=sha256:464b765e8437d8854539c15c1fd7ad07e36e01386a7a72786eb7aaefbb9a627f

Observation 3dbeb6d6-8591-44b4-a41a-f01b776b3d6d · outbound

This paper cites A boolean model of the cardiac gene regulatory network determining first and second heart field identity.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies A boolean model of the cardiac gene regulatory network determining first and second heart field identity

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.510312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:25.765036Z digest=sha256:6d70882eef9427d83fd98f09d8f5a8ab0882c34b223b2d4b0aed4d95cd18cced

Observation 19a83738-c370-4fee-83f2-a72b0b7c191a · outbound

This paper cites Tomlin, and Jaime F.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Tomlin, and Jaime F

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.439543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:25.835613Z digest=sha256:17b236395e413a09094dd7c45f70f65f0cad3d187b28ff62b9eb95e688e6af79

Observation 41d37acb-7049-474a-8f5d-645c98de102d · outbound

This paper cites Offline reinforcement learning as one big sequence modeling problem.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline reinforcement learning as one big sequence modeling problem

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.333702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:25.870439Z digest=sha256:9de3842f2e05ab23a933d60411e51138b98a73d81a72f646730e1d4b62bcc38f

Observation a6ac1ca0-9b63-442a-b3b2-3336d349cf5f · outbound

This paper cites Bradley Knox, Alessandro Allievi, Holger Banzhaf, Felix Schmitt, and Peter Stone.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Bradley Knox, Alessandro Allievi, Holger Banzhaf, Felix Schmitt, and Peter Stone

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.300791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:25.913074Z digest=sha256:ab725cabf207a02945642e69f359d599917896bb1932a645c83ff9b7b4abf466

Observation 070e12ce-7d13-4cf8-a63d-73078647b43c · outbound

This paper cites Bradley Knox and James MacGlashan.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Bradley Knox and James MacGlashan

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.242897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:25.958197Z digest=sha256:e725f4348e65cbda20fcee4456ca6aa8dfdfab16212b2042b98d9e49ff62375f

Observation fde16a3c-67aa-45f5-a778-05c824c8086e · outbound

This paper cites Offline reinforcement learning with implicit q-learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline reinforcement learning with implicit q-learning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.061773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:26.027651Z digest=sha256:a97605ba174c333e621f45cf68978e31175c07ceb8571a076953d01734cfd1af

Observation b3a8f137-4bd1-45fb-aafd-d451f660ee51 · outbound

This paper cites Stabilizing off-policy q-learning via bootstrapping error reduction.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Stabilizing off-policy q-learning via bootstrapping error reduction

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.984672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:26.094429Z digest=sha256:c17f9e5e8b8f3adca773ec0c58d0b5ee6bc38615b53d5d485da5f5fa176d5048

Observation 64f57ae5-f4a7-4a64-9bd3-0a61aa870116 · outbound

This paper cites Should i run offline reinforcement learning or behavioral cloning? InInternational Conference on Learning Representations, 2022.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Should i run offline reinforcement learning or behavioral cloning? InInternational Conference on Learning Representations, 2022

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.867017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:26.141209Z digest=sha256:ade51d821c3bc65c2a1c61db4f0fbf3a22d7d9a17483062c2931ccf13b77c784

Observation f615fdb4-b6e0-4481-b0ae-1381711316b9 · outbound

This paper cites Batch policy learning under constraints.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Batch policy learning under constraints

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.794039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:26.214577Z digest=sha256:4095c1927c817db7b2cd118402581ccb55f0009af8ce8646803e72b153299121

Observation 29940060-8782-4a84-bcbe-a71e41b9ff96 · outbound

This paper cites COptiDICE: Offline constrained reinforcement learning via stationary distribution correction estimation.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies COptiDICE: Offline constrained reinforcement learning via stationary distribution correction estimation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.682331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:26.285379Z digest=sha256:2cd4d309b507ae40ec20cd6a7286310d47fa5167639785e00c1470e4a8916d30

Observation f6c4a3aa-cbd2-4447-9333-4b437eb5bc4e · outbound

This paper cites Possible strategies to reduce the tumorigenic risk of reprogrammed normal and cancer cells.Int.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Possible strategies to reduce the tumorigenic risk of reprogrammed normal and cancer cells.Int

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.616390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:26.370187Z digest=sha256:5965c410d90fce2f529b69b8600eba3bd253389afb755e1e209d5ef25799f493

Observation 3d0b3000-3afe-42dd-8183-142dda00f2bd · outbound

This paper cites Datasets and benchmarks for offline safe reinforcement learning.Journal of Data-centric Machine Learning Research, 2024.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Datasets and benchmarks for offline safe reinforcement learning.Journal of Data-centric Machine Learning Research, 2024

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:26.417572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:26.417572Z digest=sha256:cdea191e17763060f858579ee0a065e08aaad1b3b96b7734d9f5beb0d6cab2a9

Observation 9e2f4673-99b0-41e4-be50-de20088c4880 · outbound

This paper cites Learning latent plans from play.Conference on Robot Learning (CoRL), 2019.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Learning latent plans from play.Conference on Robot Learning (CoRL), 2019

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.508157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:26.461167Z digest=sha256:fac9dc7eb4fa7ec1dd81aaf0c4f1919c42ef3700321e9f6794cc3d5e48600f0a

Observation 19957aec-0bc0-4bf5-a110-ffb6ca7be307 · outbound

This paper cites Offline goal-conditioned reinforcement learning via $f$-advantage regression.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline goal-conditioned reinforcement learning via $f$-advantage regression

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.400791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:26.536036Z digest=sha256:18b7933910264dd03bc06d8c2cb7ffae69857e4aa50939c5f6a4c0bcb3a7854d

Observation 55ca60e0-2bd4-4da5-86fc-100ca10fd6ec · outbound

This paper cites Offline reinforcement learning with domain-unlabeled data.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline reinforcement learning with domain-unlabeled data

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.236623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:26.631423Z digest=sha256:7d30f095408bad9852db6fe16dd4edca22e7856d2a29dbe5bb47b2135b3cbbef

Observation 1e28bae9-e1bf-4a3f-8bdb-d59383ed5638 · outbound

This paper cites Partial cellular reprogramming: A deep dive into an emerging rejuvenation technology.Aging Cell, 23(2):e14039, February 2024.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Partial cellular reprogramming: A deep dive into an emerging rejuvenation technology.Aging Cell, 23(2):e14039, February 2024

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.199618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:26.674719Z digest=sha256:be1f49428d125f9353d6fcbd22659127f83e84fccc70430d573c4f39c4edf763

Observation 1505f4c1-666a-4ab6-bb5e-5a148f875d0b · outbound

This paper cites Chemical reprogramming takes the fast lane.Cell Stem Cell, 30(4):335–337, April 2023.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Chemical reprogramming takes the fast lane.Cell Stem Cell, 30(4):335–337, April 2023

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.059635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:26.718658Z digest=sha256:e507c5075177f659eea65d37f60fee5f109c7ee6f53d3242c666820df3094929

Observation fc144620-7a11-4d54-b5d4-a820045f2eeb · outbound

This paper cites Ogbench: Benchmarking offline goal-conditioned rl.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Ogbench: Benchmarking offline goal-conditioned rl

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:26.828085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:26.828085Z digest=sha256:32672200db3608907f7458343e2884e948d691b037462e6966e1323fd82db9a0

Observation f16168be-be60-4fb6-9d09-e87d9224aff0 · outbound

This paper cites HIQL: Offline goal- conditioned RL with latent states as actions.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies HIQL: Offline goal- conditioned RL with latent states as actions

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.941328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:26.912477Z digest=sha256:8f18da285b6e65f32ca96dce270efe24a67525525240d52d0de60dee6fd1715e

Observation 3863d5e1-c4f1-480e-9a33-3a17062b9eb1 · outbound

This paper cites Epigenetic reprogramming as a key to reverse ageing and increase longevity.Ageing Res.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Epigenetic reprogramming as a key to reverse ageing and increase longevity.Ageing Res

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.772143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:26.964931Z digest=sha256:8075f236c44f1fb0b0b9a4129573da58eaed4f446233fca95544de72ec667217

Observation c81ebcd8-9b26-41a0-90d5-05a1ab410a42 · outbound

This paper cites Benchmarking Safe Exploration in Deep Reinforcement Learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Benchmarking Safe Exploration in Deep Reinforcement Learning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:27.035109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:27.035109Z digest=sha256:141a12fad1fda795533b96b55d095c211e19c98fd4234204df2cb383b410ad62

Observation 0c29f1ad-9d93-4042-a881-897d0b151ce5 · outbound

This paper cites Optimizing sequential gene expression modulation for cellular reprogramming - coupled boolean modeling and reinforcement learning based method.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Optimizing sequential gene expression modulation for cellular reprogramming - coupled boolean modeling and reinforcement learning based method

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.626843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:27.153418Z digest=sha256:f4fe6784e5a084ed7e88e2c99c93f67745c0f2f43339ec7063bf037d44ed72e1

Observation 719c873d-b636-49ae-8ca9-83c0dc06cab5 · outbound

This paper cites Solving minimum-cost reach avoid using reinforcement learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Solving minimum-cost reach avoid using reinforcement learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.455537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:27.204231Z digest=sha256:417d5b1eb3148c41376d1b6293f9e373e39154bebaef6f86ec4ebb7465425b7a

Observation d198c106-36b2-43fc-bf20-ee396a656646 · outbound

This paper cites Responsive safety in reinforcement learning by PID lagrangian methods.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Responsive safety in reinforcement learning by PID lagrangian methods

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.354472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:27.244936Z digest=sha256:2ea980027f865dde2baf88536d39e3018f3ba8ac7d4ffd7f00bba327b16768fe

Observation 22f5e84f-d15b-4419-998d-7b650fcd3c16 · outbound

This paper cites Induction of pluripotent stem cells from mouse embryonic and adult fibroblast cultures by defined factors.Cell, 126(4):663–676, August 2006.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Induction of pluripotent stem cells from mouse embryonic and adult fibroblast cultures by defined factors.Cell, 126(4):663–676, August 2006

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.169280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:27.332020Z digest=sha256:10186fde7ada9d3b9333a3f2381608a426eebfbf07f892d938e65c681a699ba2

Observation 81fc6768-94aa-4fd5-b5ba-dffb72dfc4dd · outbound

This paper cites Direct neuronal reprogramming: Bridging the gap between basic science and clinical application.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Direct neuronal reprogramming: Bridging the gap between basic science and clinical application

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.046446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:27.430495Z digest=sha256:02c56a0cdc3a1cd96ac752d626abf997924c64035d4c7fdc20a869e883fc4b4d

Observation a897a0c7-abd7-45af-acaf-cb2cb7a1331d · outbound

This paper cites Strategies and mechanisms of neuronal reprogramming.Brain Res.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Strategies and mechanisms of neuronal reprogramming.Brain Res

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:32.851143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:27.552585Z digest=sha256:48c803cba7409fcba9f1b415a10ba24b4192d8d636d5f57116898ff44e6aa8a6

Observation f5b567f7-8873-4b9a-8df8-4648b951a71c · outbound

This paper cites Safe decision transformer with learning-based constraints.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Safe decision transformer with learning-based constraints

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:32.666599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:27.649651Z digest=sha256:629a0935d4a95e76711f88880d8d31e586440bc824063d1c4530f9ca94b78533

Observation 33e47295-9900-4203-ac32-39d4403a5549 · outbound

This paper cites Elastic decision transformer.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Elastic decision transformer

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:32.485416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:27.751195Z digest=sha256:c94091e5287bc8a01e540464cc68c1109c261e5da849cc2ed5d998d50d22770b

Observation d6f17874-2bc7-4a92-b511-726a498cc945 · outbound

This paper cites Prevention of tumor risk associated with the reprogramming of human pluripotent stem cells.J.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Prevention of tumor risk associated with the reprogramming of human pluripotent stem cells.J

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:32.350234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:27.829982Z digest=sha256:24c76f94c09ce92f794929a7a413c28a85912ab4a723d932dd6e950b9a07f1e5

Observation ec8446de-a943-44b6-974b-92a2e14c78f0 · outbound

This paper cites Constraints penalized q-learning for safe offline reinforcement learning.Proc.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Constraints penalized q-learning for safe offline reinforcement learning.Proc

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:32.172365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:27.880716Z digest=sha256:428c7a9f513445d4258c8b38ec229df39471e5817bd64ab6a5b5e411e4f46edc

Observation 645f00e5-c53b-458b-a07e-601d437ed242 · outbound

This paper cites Joshua Tenenbaum, and Chuang Gan.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Joshua Tenenbaum, and Chuang Gan

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:32.014577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:27.974547Z digest=sha256:1a2d01ce1d3837093dd5305946a44268a10d2d94c5000f869333a843c3dbcc50

Observation a09187af-56c3-43c8-87dd-ddc1188ae56d · outbound

This paper cites Rethinking goal-conditioned supervised learning and its connection to offline RL.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Rethinking goal-conditioned supervised learning and its connection to offline RL

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:28.122271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:28.122271Z digest=sha256:a1aed8c4ed3b5b7b7089669a4a33c3c7ba1960d30cb1a9d053399622fc9077bd

Observation aa775184-0685-471b-b5f0-9354e9cb721b · outbound

This paper cites Swapped goal-conditioned offline reinforcement learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Swapped goal-conditioned offline reinforcement learning

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:31.829713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:28.248875Z digest=sha256:0440bb8b01165aa0bf2c5be6f23b473f823a3d82a9dc823cde578071eac2d3e2

Observation 364f41cf-dbec-44fa-af09-eebf4479af4b · outbound

This paper cites Pre-trained multi-goal transformers with prompt optimization for efficient online adaptation.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Pre-trained multi-goal transformers with prompt optimization for efficient online adaptation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:31.676433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:28.327895Z digest=sha256:5ef6c9b04b3bab63ebdf21f069aab4cea65c96ba5619e6e9d98eee5766001318

Observation 26d04b46-3e7c-471b-b2be-b6dac9ba78a6 · outbound

This paper cites Online Decision Transformer.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Online Decision Transformer

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:28.427899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:28.427899Z digest=sha256:e1651591bf12597f01570a41effa1ec9f0f005b17cf0b9712c98e2d1c6d45b57

Observation b0b4617a-db0e-4b89-849b-5180e6f0b2ad · outbound

This paper cites attempting.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies attempting

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:31.480166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:28.513558Z digest=sha256:cbec814bc53c4e9cc48163974f64aa35bbeaafcd3774db5506cd81e2a714a6e0

Observation 91714414-649b-4091-84e7-e04cc85556d9 · outbound

This paper cites Constrained markov decision processes with total cost criteria: Occupation measures and primal LP.Math.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Constrained markov decision processes with total cost criteria: Occupation measures and primal LP.Math

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:31.305262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:28.593570Z digest=sha256:65b19b391a1b44ac5a4df1cd667d8836dcaadc82506cfdf36f623c95f437edea

Observation c90cf0b8-f87b-4efb-a9bf-4f07173c00f7 · outbound

This paper cites An efficient algorith for determining the convex hull of a finite planar set.Inf.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies An efficient algorith for determining the convex hull of a finite planar set.Inf

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:31.151567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:28.735155Z digest=sha256:c85e7b5b04978faa4ba1ff458667b86ef86811a295261fcce08a010f94bc6b45

Observation 3c4ae2a4-c47f-41a2-a97f-f7b787c03a27 · outbound

This paper cites Cosine annealing with warmup for pytorch.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Cosine annealing with warmup for pytorch

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:31.055318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:28.825129Z digest=sha256:9f3c566e0c1e42e5e1c27316576534ab7706502733e38c2ebdcab0f601d3effc

Observation fc3ae9f4-ed21-4f3f-a0fd-1ad405207a86 · outbound

This paper cites Tune: A Research Platform for Distributed Model Selection and Training.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Tune: A Research Platform for Distributed Model Selection and Training

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:28.915801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:28.915801Z digest=sha256:f359ae1e36f58b2c0fa530517982e57b085eec3e03f3eff3e88ac1052a8d845d

Observation 49aa24ea-f7f6-4f17-a854-2475887477c4 · outbound

This paper cites A new concave hull algorithm and concaveness measure for n-dimensional datasets.Journal of Information Science and Engineering, 29:379–392, 03 2013.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies A new concave hull algorithm and concaveness measure for n-dimensional datasets.Journal of Information Science and Engineering, 29:379–392, 03 2013

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:30.950557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:28.989109Z digest=sha256:ea03a6023ba4f146be08683a9c11cbc31769a830da1cc9a0beeec289f7870219

Observation 27de630f-f718-4b98-bc1a-2acde3be25cf · outbound

This paper cites Language models are unsupervised multitask learners.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Language models are unsupervised multitask learners

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:29.069477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:29.069477Z digest=sha256:c4022ce9fe95148922f1d66f498cd05b1e17e3f739d33b86e83b60c72bf5e63b

Observation ee2c9eec-e71c-4554-afb9-ca7904b261f8 · outbound

This paper cites Pay attention to what matters.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Pay attention to what matters

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:30.825656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:29.159136Z digest=sha256:16889996dbfa134198a0a2446359f43e3f7d3d028a6ef062e6c9f7a5b080e309

Observation d3273ad3-ad56-4f19-ada0-778141a333ca · outbound

This paper cites concave_hull.https://github.com/cubao/concave_hull.git, 2022.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies concave_hull.https://github.com/cubao/concave_hull.git, 2022

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:30.566152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:29.274220Z digest=sha256:a0eaaa9c3437005cc350149855d4701e28ee1b06a6c6dae4af781eeb910b6c5a

Observation 41df043f-01aa-43ac-b4bb-418911854e30 · outbound

This paper cites 11th EAI International Conference, ICCASA 2022 Vinh Long, Vietnam, October 27–28, 2022 Proceedings.04 2023.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies 11th EAI International Conference, ICCASA 2022 Vinh Long, Vietnam, October 27–28, 2022 Proceedings.04 2023

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:30.322274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:29.355717Z digest=sha256:5f5364144547351af61c9bb7bcea2c2b2f45dd767f3cc6b73676858625ca5e53

Observation 320e1e58-922e-4c2d-8f19-5b5401d74c1e · outbound

This paper cites an unresolved cited work.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:21:30.024476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:21:29.435813Z digest=sha256:179aa4d1589e144bff0a227d12d251f31aab1f8d67b92f18620c2574858773ff

Pith citing papers

No inbound Pith citation observations are available.