Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:21:29.435813Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2505.19337.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:21:29.435813Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
61 of 61 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ae52ad05-4262-4008-a82f-9c60562da714 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies A review of reward functions for reinforcement learning in the context of autonomous driving
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1495beaf-23bc-4b1f-8e60-3db0b9161aa0 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Hindsight experience replay
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9c5b4ea9-7f87-4975-9c69-2d817661fef0 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline Goal-Conditioned Reinforcement Learning for Safety-Critical Tasks with Recovery Policy
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 933a4504-06f1-4922-9b93-77336d51342f · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d01ee715-bf00-4edb-909b-acfd60bf7f16 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Decision Transformer: Reinforcement Learning via Sequence Modeling
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 457babd5-5a2a-42f9-b215-a1b260f6e891 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Gymnasium robotics, 2024
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bada3ed-c54a-4ff1-a252-9f16eeec7edf · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Contrastive Learning as Goal-Conditioned Reinforcement Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5de72c4d-b6bf-4b36-a11f-df3dd372b951 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Safe multi-agent navigation guided by goal- conditioned safe reinforcement learning, 2025
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7565752c-768f-4597-a351-4e6dc68adcab · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Curriculum reinforcement learning for complex reward functions
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4dff962f-8890-4939-bd0c-985c3d70ef3a · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Off-policy deep reinforcement learning without exploration
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90562ba9-61e4-41f6-8a4f-6d9098548bb4 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Integrating domain knowledge for handling limited data in offline RL
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a2365892-1fd6-4786-9f4b-ca7f9256155e · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Learning to Reach Goals via Iterated Supervised Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9970028b-05b8-44f1-b722-08857b85e611 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Bullet-safety-gym: A framework for constrained reinforcement learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b630f1c-0390-4132-9913-f3b28e69f4c0 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Chemical reprogramming of human somatic cells to pluripotent stem cells.Nature, 605(7909):325–331, May 2022
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3dbeb6d6-8591-44b4-a41a-f01b776b3d6d · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies A boolean model of the cardiac gene regulatory network determining first and second heart field identity
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 19a83738-c370-4fee-83f2-a72b0b7c191a · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Tomlin, and Jaime F
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 41d37acb-7049-474a-8f5d-645c98de102d · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline reinforcement learning as one big sequence modeling problem
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a6ac1ca0-9b63-442a-b3b2-3336d349cf5f · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Bradley Knox, Alessandro Allievi, Holger Banzhaf, Felix Schmitt, and Peter Stone
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 070e12ce-7d13-4cf8-a63d-73078647b43c · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Bradley Knox and James MacGlashan
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fde16a3c-67aa-45f5-a778-05c824c8086e · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline reinforcement learning with implicit q-learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b3a8f137-4bd1-45fb-aafd-d451f660ee51 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Stabilizing off-policy q-learning via bootstrapping error reduction
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 64f57ae5-f4a7-4a64-9bd3-0a61aa870116 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Should i run offline reinforcement learning or behavioral cloning? InInternational Conference on Learning Representations, 2022
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f615fdb4-b6e0-4481-b0ae-1381711316b9 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Batch policy learning under constraints
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 29940060-8782-4a84-bcbe-a71e41b9ff96 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies COptiDICE: Offline constrained reinforcement learning via stationary distribution correction estimation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f6c4a3aa-cbd2-4447-9333-4b437eb5bc4e · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Possible strategies to reduce the tumorigenic risk of reprogrammed normal and cancer cells.Int
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3d0b3000-3afe-42dd-8183-142dda00f2bd · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Datasets and benchmarks for offline safe reinforcement learning.Journal of Data-centric Machine Learning Research, 2024
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e2f4673-99b0-41e4-be50-de20088c4880 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Learning latent plans from play.Conference on Robot Learning (CoRL), 2019
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 19957aec-0bc0-4bf5-a110-ffb6ca7be307 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline goal-conditioned reinforcement learning via $f$-advantage regression
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 55ca60e0-2bd4-4da5-86fc-100ca10fd6ec · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline reinforcement learning with domain-unlabeled data
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1e28bae9-e1bf-4a3f-8bdb-d59383ed5638 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Partial cellular reprogramming: A deep dive into an emerging rejuvenation technology.Aging Cell, 23(2):e14039, February 2024
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1505f4c1-666a-4ab6-bb5e-5a148f875d0b · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Chemical reprogramming takes the fast lane.Cell Stem Cell, 30(4):335–337, April 2023
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fc144620-7a11-4d54-b5d4-a820045f2eeb · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Ogbench: Benchmarking offline goal-conditioned rl
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f16168be-be60-4fb6-9d09-e87d9224aff0 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies HIQL: Offline goal- conditioned RL with latent states as actions
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3863d5e1-c4f1-480e-9a33-3a17062b9eb1 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Epigenetic reprogramming as a key to reverse ageing and increase longevity.Ageing Res
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c81ebcd8-9b26-41a0-90d5-05a1ab410a42 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Benchmarking Safe Exploration in Deep Reinforcement Learning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c29f1ad-9d93-4042-a881-897d0b151ce5 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Optimizing sequential gene expression modulation for cellular reprogramming - coupled boolean modeling and reinforcement learning based method
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 719c873d-b636-49ae-8ca9-83c0dc06cab5 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Solving minimum-cost reach avoid using reinforcement learning
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d198c106-36b2-43fc-bf20-ee396a656646 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Responsive safety in reinforcement learning by PID lagrangian methods
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 22f5e84f-d15b-4419-998d-7b650fcd3c16 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Induction of pluripotent stem cells from mouse embryonic and adult fibroblast cultures by defined factors.Cell, 126(4):663–676, August 2006
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 81fc6768-94aa-4fd5-b5ba-dffb72dfc4dd · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Direct neuronal reprogramming: Bridging the gap between basic science and clinical application
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a897a0c7-abd7-45af-acaf-cb2cb7a1331d · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Strategies and mechanisms of neuronal reprogramming.Brain Res
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f5b567f7-8873-4b9a-8df8-4648b951a71c · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Safe decision transformer with learning-based constraints
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 33e47295-9900-4203-ac32-39d4403a5549 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Elastic decision transformer
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d6f17874-2bc7-4a92-b511-726a498cc945 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Prevention of tumor risk associated with the reprogramming of human pluripotent stem cells.J
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ec8446de-a943-44b6-974b-92a2e14c78f0 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Constraints penalized q-learning for safe offline reinforcement learning.Proc
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 645f00e5-c53b-458b-a07e-601d437ed242 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Joshua Tenenbaum, and Chuang Gan
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a09187af-56c3-43c8-87dd-ddc1188ae56d · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Rethinking goal-conditioned supervised learning and its connection to offline RL
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa775184-0685-471b-b5f0-9354e9cb721b · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Swapped goal-conditioned offline reinforcement learning
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 364f41cf-dbec-44fa-af09-eebf4479af4b · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Pre-trained multi-goal transformers with prompt optimization for efficient online adaptation
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 26d04b46-3e7c-471b-b2be-b6dac9ba78a6 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Online Decision Transformer
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0b4617a-db0e-4b89-849b-5180e6f0b2ad · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies attempting
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 91714414-649b-4091-84e7-e04cc85556d9 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Constrained markov decision processes with total cost criteria: Occupation measures and primal LP.Math
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c90cf0b8-f87b-4efb-a9bf-4f07173c00f7 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies An efficient algorith for determining the convex hull of a finite planar set.Inf
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3c4ae2a4-c47f-41a2-a97f-f7b787c03a27 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Cosine annealing with warmup for pytorch
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fc3ae9f4-ed21-4f3f-a0fd-1ad405207a86 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Tune: A Research Platform for Distributed Model Selection and Training
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49aa24ea-f7f6-4f17-a854-2475887477c4 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies A new concave hull algorithm and concaveness measure for n-dimensional datasets.Journal of Information Science and Engineering, 29:379–392, 03 2013
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 27de630f-f718-4b98-bc1a-2acde3be25cf · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Language models are unsupervised multitask learners
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee2c9eec-e71c-4554-afb9-ca7904b261f8 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Pay attention to what matters
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d3273ad3-ad56-4f19-ada0-778141a333ca · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies concave_hull.https://github.com/cubao/concave_hull.git, 2022
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 41df043f-01aa-43ac-b4bb-418911854e30 · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies 11th EAI International Conference, ICCASA 2022 Vinh Long, Vietnam, October 27–28, 2022 Proceedings.04 2023
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 320e1e58-922e-4c2d-8f19-5b5401d74c1e · outbound
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Unresolved cited work
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
No inbound Pith citation observations are available.