Pith. sign in

Paper Citation Record · LEDGER

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies

As of 10 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2505.19337.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19337 v2

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:21:29.435813Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

61 of 61 outbound references displayed

  • verified exact1
  • verified fuzzy45
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ae52ad05-4262-4008-a82f-9c60562da714 · outbound

This paper cites A review of reward functions for reinforcement learning in the context of autonomous driving.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies A review of reward functions for reinforcement learning in the context of autonomous driving

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:36.159398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:24.932108Z digest=sha256:d32161c7477d0fe6fc70b2b1af4dbb396143189e469eab4f4c67742cce088970

Observation 1495beaf-23bc-4b1f-8e60-3db0b9161aa0 · outbound

This paper cites Hindsight experience replay.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Hindsight experience replay

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:36.051685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:24.991707Z digest=sha256:197e6c0f001c783b4967f6e7792e1917604dfb83082bd3b25bf752be2f10637b

Observation 9c5b4ea9-7f87-4975-9c69-2d817661fef0 · outbound

This paper cites Offline Goal-Conditioned Reinforcement Learning for Safety-Critical Tasks with Recovery Policy.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline Goal-Conditioned Reinforcement Learning for Safety-Critical Tasks with Recovery Policy

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:21:29.770963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:25.084900Z digest=sha256:f0aad0db47ab03f280549e014812d2c2ebe19450243102abae00ca08ff054057

Observation 933a4504-06f1-4922-9b93-77336d51342f · outbound

This paper cites Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.121008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.121008Z digest=sha256:19507f7914db03dd7ba15b46e84083d99a0cb1f0f1989a566b06739945237e7a

Observation d01ee715-bf00-4edb-909b-acfd60bf7f16 · outbound

This paper cites Decision Transformer: Reinforcement Learning via Sequence Modeling.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Decision Transformer: Reinforcement Learning via Sequence Modeling

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.163416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.163416Z digest=sha256:a1f88a825e9a8f6da71c8d7db6e91ec0431a60f8f3301228b57537ee52d1b0c2

Observation 457babd5-5a2a-42f9-b215-a1b260f6e891 · outbound

This paper cites Gymnasium robotics, 2024.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Gymnasium robotics, 2024

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.216093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.216093Z digest=sha256:97b73336067807cc89902a598ea20822ee2cc3c6d631e1437a3defeb670540e2

Observation 8bada3ed-c54a-4ff1-a252-9f16eeec7edf · outbound

This paper cites Contrastive Learning as Goal-Conditioned Reinforcement Learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Contrastive Learning as Goal-Conditioned Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.312416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.312416Z digest=sha256:f97c7b2734f49b2811e4358059202f41e16f1b7a5db61d599579f61ccddbcba9

Observation 5de72c4d-b6bf-4b36-a11f-df3dd372b951 · outbound

This paper cites Safe multi-agent navigation guided by goal- conditioned safe reinforcement learning, 2025.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Safe multi-agent navigation guided by goal- conditioned safe reinforcement learning, 2025

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.948907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:25.373485Z digest=sha256:d9a37162e3a4699700bf4d2c907f7674be5a5b7201abda8d75e2683bfd676e1a

Observation 7565752c-768f-4597-a351-4e6dc68adcab · outbound

This paper cites Curriculum reinforcement learning for complex reward functions.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Curriculum reinforcement learning for complex reward functions

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.792698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:25.413568Z digest=sha256:88b985fa81d171055ff6f2cd2f727700b0aa14241ed2b8634677001ebb8ff02a

Observation 4dff962f-8890-4939-bd0c-985c3d70ef3a · outbound

This paper cites Off-policy deep reinforcement learning without exploration.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Off-policy deep reinforcement learning without exploration

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.446509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.446509Z digest=sha256:73b21b09c8359eccb40e500afe7e551a7e0991ab1e14a4413eb2da6a00fe8c73

Observation 90562ba9-61e4-41f6-8a4f-6d9098548bb4 · outbound

This paper cites Integrating domain knowledge for handling limited data in offline RL.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Integrating domain knowledge for handling limited data in offline RL

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.704832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:25.528577Z digest=sha256:dd837ec8e6d2d918e1a63110e36ff74c2f1ac2f9428d061c7a50f7dd6e20c278

Observation a2365892-1fd6-4786-9f4b-ca7f9256155e · outbound

This paper cites Learning to Reach Goals via Iterated Supervised Learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Learning to Reach Goals via Iterated Supervised Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.585866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.585866Z digest=sha256:7ff05e4a362a6de137430fd2c0d9d9bd1a3fc5f82d5e80351583e046d6882c5c

Observation 9970028b-05b8-44f1-b722-08857b85e611 · outbound

This paper cites Bullet-safety-gym: A framework for constrained reinforcement learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Bullet-safety-gym: A framework for constrained reinforcement learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.636687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.636687Z digest=sha256:36ba8f975a3439b4eec328b9291189dde61630721ec274bf5b47f5aedb8aece4

Observation 9b630f1c-0390-4132-9913-f3b28e69f4c0 · outbound

This paper cites Chemical reprogramming of human somatic cells to pluripotent stem cells.Nature, 605(7909):325–331, May 2022.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Chemical reprogramming of human somatic cells to pluripotent stem cells.Nature, 605(7909):325–331, May 2022

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.618483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:25.685668Z digest=sha256:f92cca65090aaec7910061204ed37356b84fdcb36aef3b9deff957755430ad7d

Observation 3dbeb6d6-8591-44b4-a41a-f01b776b3d6d · outbound

This paper cites A boolean model of the cardiac gene regulatory network determining first and second heart field identity.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies A boolean model of the cardiac gene regulatory network determining first and second heart field identity

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.510312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:25.765036Z digest=sha256:10a6c83a860389a58e53487200f44533a0d12728ee0304cc8472b2417a7400dc

Observation 19a83738-c370-4fee-83f2-a72b0b7c191a · outbound

This paper cites Tomlin, and Jaime F.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Tomlin, and Jaime F

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.439543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:25.835613Z digest=sha256:7584b5227194456ed4da6304ffd5879d392ac866c271c07a432d80d8be19c3bb

Observation 41d37acb-7049-474a-8f5d-645c98de102d · outbound

This paper cites Offline reinforcement learning as one big sequence modeling problem.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline reinforcement learning as one big sequence modeling problem

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.333702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:25.870439Z digest=sha256:1b53be460a0aca1b99ce7bbf07dfa54c0aa2369dd9595643182b981612ab21d0

Observation a6ac1ca0-9b63-442a-b3b2-3336d349cf5f · outbound

This paper cites Bradley Knox, Alessandro Allievi, Holger Banzhaf, Felix Schmitt, and Peter Stone.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Bradley Knox, Alessandro Allievi, Holger Banzhaf, Felix Schmitt, and Peter Stone

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.300791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:25.913074Z digest=sha256:d59810ad352ff594323e277eba5062e43b20e434970b69f23fef8dcea31e8195

Observation 070e12ce-7d13-4cf8-a63d-73078647b43c · outbound

This paper cites Bradley Knox and James MacGlashan.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Bradley Knox and James MacGlashan

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.242897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:25.958197Z digest=sha256:e2388c91b7f534ff05a279143d6b58a0afa08b47a52290a81696f1fd951991fc

Observation fde16a3c-67aa-45f5-a778-05c824c8086e · outbound

This paper cites Offline reinforcement learning with implicit q-learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline reinforcement learning with implicit q-learning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.061773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:26.027651Z digest=sha256:58b5decda622e548ee0d105834999ec45a02993096ba4860a48d8dae4c576a8d

Observation b3a8f137-4bd1-45fb-aafd-d451f660ee51 · outbound

This paper cites Stabilizing off-policy q-learning via bootstrapping error reduction.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Stabilizing off-policy q-learning via bootstrapping error reduction

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.984672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:26.094429Z digest=sha256:7446b8d5496bf394a57e47be865b50a71fdc9d97debdb99fb8272f49d6fbf52f

Observation 64f57ae5-f4a7-4a64-9bd3-0a61aa870116 · outbound

This paper cites Should i run offline reinforcement learning or behavioral cloning? InInternational Conference on Learning Representations, 2022.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Should i run offline reinforcement learning or behavioral cloning? InInternational Conference on Learning Representations, 2022

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.867017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:26.141209Z digest=sha256:61dc58b046254e7403acc22baae24a792fcb74149abe82ffed8d26c0fdaba362

Observation f615fdb4-b6e0-4481-b0ae-1381711316b9 · outbound

This paper cites Batch policy learning under constraints.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Batch policy learning under constraints

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.794039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:26.214577Z digest=sha256:49d1327ded2643faea4e0ab00fbc1628d1d3c7dc9474d568c9ef3fff9310020e

Observation 29940060-8782-4a84-bcbe-a71e41b9ff96 · outbound

This paper cites COptiDICE: Offline constrained reinforcement learning via stationary distribution correction estimation.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies COptiDICE: Offline constrained reinforcement learning via stationary distribution correction estimation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.682331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:26.285379Z digest=sha256:b7fabcbbccd14e63a8189d77d4585cd2320e0da2f7fad284e328852be1361301

Observation f6c4a3aa-cbd2-4447-9333-4b437eb5bc4e · outbound

This paper cites Possible strategies to reduce the tumorigenic risk of reprogrammed normal and cancer cells.Int.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Possible strategies to reduce the tumorigenic risk of reprogrammed normal and cancer cells.Int

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.616390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:26.370187Z digest=sha256:58ea05d1ebfc0c870c557bde15a6496cba1dae10244650c483bfb28f5a58ae46

Observation 3d0b3000-3afe-42dd-8183-142dda00f2bd · outbound

This paper cites Datasets and benchmarks for offline safe reinforcement learning.Journal of Data-centric Machine Learning Research, 2024.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Datasets and benchmarks for offline safe reinforcement learning.Journal of Data-centric Machine Learning Research, 2024

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:26.417572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:26.417572Z digest=sha256:cdea191e17763060f858579ee0a065e08aaad1b3b96b7734d9f5beb0d6cab2a9

Observation 9e2f4673-99b0-41e4-be50-de20088c4880 · outbound

This paper cites Learning latent plans from play.Conference on Robot Learning (CoRL), 2019.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Learning latent plans from play.Conference on Robot Learning (CoRL), 2019

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.508157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:26.461167Z digest=sha256:30d3cc7137b904d2a63a018322c340cc4286051e1a4a599e3421b3ace63bfebd

Observation 19957aec-0bc0-4bf5-a110-ffb6ca7be307 · outbound

This paper cites Offline goal-conditioned reinforcement learning via $f$-advantage regression.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline goal-conditioned reinforcement learning via $f$-advantage regression

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.400791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:26.536036Z digest=sha256:89a5c993bd1d1a019ed403e7aedd82d9ac369f0b9b894d95449e8685e73ccb7e

Observation 55ca60e0-2bd4-4da5-86fc-100ca10fd6ec · outbound

This paper cites Offline reinforcement learning with domain-unlabeled data.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline reinforcement learning with domain-unlabeled data

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.236623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:26.631423Z digest=sha256:2c58cbe83833ceddf7dd7d9a97cd98349ef5b4baf09e0f33e46737e07c00e34d

Observation 1e28bae9-e1bf-4a3f-8bdb-d59383ed5638 · outbound

This paper cites Partial cellular reprogramming: A deep dive into an emerging rejuvenation technology.Aging Cell, 23(2):e14039, February 2024.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Partial cellular reprogramming: A deep dive into an emerging rejuvenation technology.Aging Cell, 23(2):e14039, February 2024

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.199618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:26.674719Z digest=sha256:4c895b6303cd2531d9da32c2585ccc1935326264d080dd6bee7e1d7d1bacc8b2

Observation 1505f4c1-666a-4ab6-bb5e-5a148f875d0b · outbound

This paper cites Chemical reprogramming takes the fast lane.Cell Stem Cell, 30(4):335–337, April 2023.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Chemical reprogramming takes the fast lane.Cell Stem Cell, 30(4):335–337, April 2023

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.059635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:26.718658Z digest=sha256:ffaea3675d369528934bcbf30c5a83abbf2888da37791ec77ff27f64ab59de17

Observation fc144620-7a11-4d54-b5d4-a820045f2eeb · outbound

This paper cites Ogbench: Benchmarking offline goal-conditioned rl.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Ogbench: Benchmarking offline goal-conditioned rl

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:26.828085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:26.828085Z digest=sha256:32672200db3608907f7458343e2884e948d691b037462e6966e1323fd82db9a0

Observation f16168be-be60-4fb6-9d09-e87d9224aff0 · outbound

This paper cites HIQL: Offline goal- conditioned RL with latent states as actions.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies HIQL: Offline goal- conditioned RL with latent states as actions

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.941328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:26.912477Z digest=sha256:a7f9c29f2bb351e645c07f298803b7eb5dc4e4a61220615f8e89901f23ef4657

Observation 3863d5e1-c4f1-480e-9a33-3a17062b9eb1 · outbound

This paper cites Epigenetic reprogramming as a key to reverse ageing and increase longevity.Ageing Res.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Epigenetic reprogramming as a key to reverse ageing and increase longevity.Ageing Res

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.772143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:26.964931Z digest=sha256:ff0610156379167d0bd169b086745de4ec5edfb2a4e2c2fe411f3a93ed0f6de2

Observation c81ebcd8-9b26-41a0-90d5-05a1ab410a42 · outbound

This paper cites Benchmarking Safe Exploration in Deep Reinforcement Learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Benchmarking Safe Exploration in Deep Reinforcement Learning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:27.035109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:27.035109Z digest=sha256:141a12fad1fda795533b96b55d095c211e19c98fd4234204df2cb383b410ad62

Observation 0c29f1ad-9d93-4042-a881-897d0b151ce5 · outbound

This paper cites Optimizing sequential gene expression modulation for cellular reprogramming - coupled boolean modeling and reinforcement learning based method.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Optimizing sequential gene expression modulation for cellular reprogramming - coupled boolean modeling and reinforcement learning based method

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.626843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:27.153418Z digest=sha256:8da39b4120a9b8fc0a11f64018aa140c809af594eb96c24adce9a2609eb4e74d

Observation 719c873d-b636-49ae-8ca9-83c0dc06cab5 · outbound

This paper cites Solving minimum-cost reach avoid using reinforcement learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Solving minimum-cost reach avoid using reinforcement learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.455537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:27.204231Z digest=sha256:715ead2d06d5ebe02f6d1ecb4201b13224bc256fa409f6c934c4532f1b2efe57

Observation d198c106-36b2-43fc-bf20-ee396a656646 · outbound

This paper cites Responsive safety in reinforcement learning by PID lagrangian methods.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Responsive safety in reinforcement learning by PID lagrangian methods

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.354472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:27.244936Z digest=sha256:a53974b3492c3fc3dac51210d39224246b17048e910112d18788ef3892deef7e

Observation 22f5e84f-d15b-4419-998d-7b650fcd3c16 · outbound

This paper cites Induction of pluripotent stem cells from mouse embryonic and adult fibroblast cultures by defined factors.Cell, 126(4):663–676, August 2006.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Induction of pluripotent stem cells from mouse embryonic and adult fibroblast cultures by defined factors.Cell, 126(4):663–676, August 2006

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.169280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:27.332020Z digest=sha256:342b5d1770ea702343780194e08c357a0f5dcd333007f1e2fa347aa73fcd5a74

Observation 81fc6768-94aa-4fd5-b5ba-dffb72dfc4dd · outbound

This paper cites Direct neuronal reprogramming: Bridging the gap between basic science and clinical application.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Direct neuronal reprogramming: Bridging the gap between basic science and clinical application

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.046446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:27.430495Z digest=sha256:c4cf98e0f7702c8c18dd1031cd3e84bed669ed4b7050a25748fd0c8ede5357b8

Observation a897a0c7-abd7-45af-acaf-cb2cb7a1331d · outbound

This paper cites Strategies and mechanisms of neuronal reprogramming.Brain Res.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Strategies and mechanisms of neuronal reprogramming.Brain Res

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:32.851143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:27.552585Z digest=sha256:8efc5643d7cbcff90a770a086f688b7794c80448174502bac845a706e0d6c618

Observation f5b567f7-8873-4b9a-8df8-4648b951a71c · outbound

This paper cites Safe decision transformer with learning-based constraints.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Safe decision transformer with learning-based constraints

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:32.666599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:27.649651Z digest=sha256:c18a70b47b5bc0ee73d061d1f4c4781ed46582e0c7eeb89a3ae28d7b43b9dccb

Observation 33e47295-9900-4203-ac32-39d4403a5549 · outbound

This paper cites Elastic decision transformer.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Elastic decision transformer

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:32.485416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:27.751195Z digest=sha256:5b6a0064e811841555f8231f10ca9d4f25b70d5141fb9752cc2ae5c447b89322

Observation d6f17874-2bc7-4a92-b511-726a498cc945 · outbound

This paper cites Prevention of tumor risk associated with the reprogramming of human pluripotent stem cells.J.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Prevention of tumor risk associated with the reprogramming of human pluripotent stem cells.J

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:32.350234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:27.829982Z digest=sha256:9a6ead561284d37381315a6f9e2631c2fde720143d656965d06227c1b4ebeb63

Observation ec8446de-a943-44b6-974b-92a2e14c78f0 · outbound

This paper cites Constraints penalized q-learning for safe offline reinforcement learning.Proc.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Constraints penalized q-learning for safe offline reinforcement learning.Proc

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:32.172365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:27.880716Z digest=sha256:5f5878608381d9a9a4b7711f05d14f3512d0adbc9141569a4aeb63315d08e616

Observation 645f00e5-c53b-458b-a07e-601d437ed242 · outbound

This paper cites Joshua Tenenbaum, and Chuang Gan.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Joshua Tenenbaum, and Chuang Gan

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:32.014577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:27.974547Z digest=sha256:fd73332f56c77a55e91101decbf88dbd9ca493955833b65972411f9e63928e9c

Observation a09187af-56c3-43c8-87dd-ddc1188ae56d · outbound

This paper cites Rethinking goal-conditioned supervised learning and its connection to offline RL.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Rethinking goal-conditioned supervised learning and its connection to offline RL

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:28.122271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:28.122271Z digest=sha256:a1aed8c4ed3b5b7b7089669a4a33c3c7ba1960d30cb1a9d053399622fc9077bd

Observation aa775184-0685-471b-b5f0-9354e9cb721b · outbound

This paper cites Swapped goal-conditioned offline reinforcement learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Swapped goal-conditioned offline reinforcement learning

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:31.829713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:28.248875Z digest=sha256:c9b865672a738f7ab5eeac4ed5ec6753772073c5af78679f6dce6f403a9f41cd

Observation 364f41cf-dbec-44fa-af09-eebf4479af4b · outbound

This paper cites Pre-trained multi-goal transformers with prompt optimization for efficient online adaptation.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Pre-trained multi-goal transformers with prompt optimization for efficient online adaptation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:31.676433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:28.327895Z digest=sha256:793f5cadf6b774d61d674aad4d62311e5f2c52a3f56739f3ffad5a792f0b75b2

Observation 26d04b46-3e7c-471b-b2be-b6dac9ba78a6 · outbound

This paper cites Online Decision Transformer.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Online Decision Transformer

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:28.427899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:28.427899Z digest=sha256:b850615dba9633324cb99d9e46a07dd2e5c6a5481ee9cfd86c5cfdee602b311f

Observation b0b4617a-db0e-4b89-849b-5180e6f0b2ad · outbound

This paper cites attempting.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies attempting

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:31.480166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:28.513558Z digest=sha256:09820e2b4e26d789ed8aa778c92a873d3309487fee396323cfc179795b768328

Observation 91714414-649b-4091-84e7-e04cc85556d9 · outbound

This paper cites Constrained markov decision processes with total cost criteria: Occupation measures and primal LP.Math.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Constrained markov decision processes with total cost criteria: Occupation measures and primal LP.Math

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:31.305262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:28.593570Z digest=sha256:c604117f474a3ebfade97f145534fbc1f6a95673d42a359360c96f0fd03671e0

Observation c90cf0b8-f87b-4efb-a9bf-4f07173c00f7 · outbound

This paper cites An efficient algorith for determining the convex hull of a finite planar set.Inf.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies An efficient algorith for determining the convex hull of a finite planar set.Inf

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:31.151567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:28.735155Z digest=sha256:ec450156083490b3d4d6532ee5f98c8f41f73fd386fb886bea750bf8ee4f765d

Observation 3c4ae2a4-c47f-41a2-a97f-f7b787c03a27 · outbound

This paper cites Cosine annealing with warmup for pytorch.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Cosine annealing with warmup for pytorch

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:31.055318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:28.825129Z digest=sha256:959a913064b87a8859b136177c249f2c2b6453535a54bcab10aa3ee38661a825

Observation fc3ae9f4-ed21-4f3f-a0fd-1ad405207a86 · outbound

This paper cites Tune: A Research Platform for Distributed Model Selection and Training.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Tune: A Research Platform for Distributed Model Selection and Training

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:28.915801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:28.915801Z digest=sha256:f359ae1e36f58b2c0fa530517982e57b085eec3e03f3eff3e88ac1052a8d845d

Observation 49aa24ea-f7f6-4f17-a854-2475887477c4 · outbound

This paper cites A new concave hull algorithm and concaveness measure for n-dimensional datasets.Journal of Information Science and Engineering, 29:379–392, 03 2013.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies A new concave hull algorithm and concaveness measure for n-dimensional datasets.Journal of Information Science and Engineering, 29:379–392, 03 2013

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:30.950557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:28.989109Z digest=sha256:2794d7f7fdd120e79382373a09d635faf52eaef5b7ee7e6561d0cf85cd909670

Observation 27de630f-f718-4b98-bc1a-2acde3be25cf · outbound

This paper cites Language models are unsupervised multitask learners.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Language models are unsupervised multitask learners

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:29.069477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:29.069477Z digest=sha256:c4022ce9fe95148922f1d66f498cd05b1e17e3f739d33b86e83b60c72bf5e63b

Observation ee2c9eec-e71c-4554-afb9-ca7904b261f8 · outbound

This paper cites Pay attention to what matters.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Pay attention to what matters

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:30.825656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:29.159136Z digest=sha256:99bf8d2f894ac3dfb10ecffda409c6ddda03fd37aee58ab33ef6fa11c1b8d3bc

Observation d3273ad3-ad56-4f19-ada0-778141a333ca · outbound

This paper cites concave_hull.https://github.com/cubao/concave_hull.git, 2022.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies concave_hull.https://github.com/cubao/concave_hull.git, 2022

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:30.566152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:29.274220Z digest=sha256:9426ec8434b0cc9e33527b77c162e080f56af05f4e409147385ca11113cc99f3

Observation 41df043f-01aa-43ac-b4bb-418911854e30 · outbound

This paper cites 11th EAI International Conference, ICCASA 2022 Vinh Long, Vietnam, October 27–28, 2022 Proceedings.04 2023.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies 11th EAI International Conference, ICCASA 2022 Vinh Long, Vietnam, October 27–28, 2022 Proceedings.04 2023

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:30.322274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:29.355717Z digest=sha256:8d2f689588f7d882a564ca11965b227a5adc48eca6988de2ddef0680a89436ed

Observation 320e1e58-922e-4c2d-8f19-5b5401d74c1e · outbound

This paper cites an unresolved cited work.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:21:30.024476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T14:21:29.435813Z digest=sha256:50ac31a1ebbfcd7399e683b347afe1e57c3494fcd9346896ce4c7a7e1eefc74c

Pith citing papers

No inbound Pith citation observations are available.