Pith. sign in

Paper Citation Record · LEDGER

Upside-Down Reinforcement Learning for More Interpretable Optimal Control

As of 20 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 0 inbound Pith citation observations for arXiv:2411.11457.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.11457 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T18:35:59.796489Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

50 of 50 outbound references displayed

  • verified exact1
  • verified fuzzy10
  • unresolved38
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2bdfe559-f9ee-4bae-a6cd-366433b95602 · outbound

This paper cites write newline.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.608437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.608437Z digest=sha256:67834aeae80fa178a694b601f4b57fdb2c2ef8da4b7e44f502209a75c0cb66d5

Observation abb3a62d-6a64-41a0-8675-4821a0df9f8f · outbound

This paper cites and Mishra, S.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Mishra, S

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.399198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.614212Z digest=sha256:4b4e37707c13198442fc6ab16a22b066ee934ae74836bf788cf38bace8e81e3b

Observation a29ac238-2445-4f64-a265-a146a823f43f · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.387963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.619203Z digest=sha256:bdc737d7331985d2a9738d60545ffd5bebdaeb4ed16b45ca1b83c19d84115667

Observation 5f06d254-b2f7-44b0-9c53-1105dd80dc89 · outbound

This paper cites All You Need Is Supervised Learning: From Imitation Learning to Meta-RL With Upside Down RL.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control All You Need Is Supervised Learning: From Imitation Learning to Meta-RL With Upside Down RL

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.623576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.623576Z digest=sha256:f05c5cb35d1276a81e42a37c868f13fd506515a9790fbe48d3bbf65991a10c36

Observation d8711a58-671b-4cde-9fb0-e2c312369abd · outbound

This paper cites Learning Relative Return Policies With Upside-Down Reinforcement Learning.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Learning Relative Return Policies With Upside-Down Reinforcement Learning

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-08-12T18:35:59.922846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.627958Z digest=sha256:4c832fb04cf5dd729d9e63736b3707aa2aeb660877c03b1c436c81ae4f9f5fa7

Observation 4d3df4d2-cb04-4778-a7ac-f2f997510db1 · outbound

This paper cites G., Sutton, R.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control G., Sutton, R

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.375535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.632421Z digest=sha256:3c0b4a8e716d9e09f575306d50505fe015a7b2eb7eeb7c8575cf2c1d40812bde

Observation 356fc9e3-a00e-43f0-8fe1-b43b19252053 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.636180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.636180Z digest=sha256:31e3aeec499f0561617aef591e7bc829fa6a7b8be5a57f2df89f4a9fb3fe7e2b

Observation fb3827f9-b6af-4954-a12f-8b0c649e1fbb · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.356335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.640291Z digest=sha256:a85413dd4a7de0bbe0ec089a5cf978dd52467dfad8932e526687edac61dc0c21

Observation 7f6b6e94-cfb7-4d46-9236-d7c4e1798352 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.343117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.644021Z digest=sha256:1f6d506f61242a2910c6ffa6bfd6fe73cee48fbab8541917913bd9f40c2f2a72

Observation d21af92e-d3c8-4ea6-b0ca-c53355df626d · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.330895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.647621Z digest=sha256:22da3732ba13e1cd8cd82d481dc041eb6c8327487c0aa653ac1fe6ea9edf5add

Observation 805e1cfb-34f3-4478-b0a6-caa6085156b5 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.318334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.651418Z digest=sha256:0be5c5b68ce0e5e4080517f74f5a899d14c347bfe430bba9bc00d2917259088c

Observation c303bfe3-5043-4dc1-bc5e-94e1c89e15a4 · outbound

This paper cites and Hart, P.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Hart, P

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.305489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.656560Z digest=sha256:b7f86c76519e419ee3df7019052298feb34adeb1bfee1f8feabbb01f9701a36e

Observation 0291fb40-5715-4138-b17e-695f96393908 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.661138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.661138Z digest=sha256:a69af39585b681db431259cd8dc0e44d0728eab408a676afd6a352fa565cf7f8

Observation a8b6ea7f-fc98-4de4-9809-652779e16cb0 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.284611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.664712Z digest=sha256:03c603142b4fbcf26ded55dd36ffc1af83a18f5bb2d1571e3ea48bcd65eca7cd

Observation 5c374c7d-2122-4422-8d62-101f6fee61e1 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.272324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.668061Z digest=sha256:95c5f27686345f127dac9517b197343333e675356fdc7251f0c737ebbac60ae5

Observation e2ae35d7-bfd2-4d5d-b482-327d410a24d3 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.260146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.671467Z digest=sha256:0ebd900f3ea89583b84265e3b5b14fd3bea1612bceee8d5b4e90c5b4b32d2df2

Observation afa1af9c-2172-4f6a-a10f-d1fcd82c4cfb · outbound

This paper cites E., et al.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control E., et al

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.247060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.675111Z digest=sha256:251079f72e85b63e25dad7fd0011da76eef58e2c1caabacc9c277a4408f1d49d

Observation c747cf82-8295-4991-a24c-1f977f7eccfe · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.678651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.678651Z digest=sha256:e5718809b94c1daacab7ed189df20779dfda71df73954fb71d756c00aa4ca696

Observation 946b6186-11c2-4f8a-97e6-e932345ed70f · outbound

This paper cites Generalized Decision Transformer for Offline Hindsight Information Matching.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Generalized Decision Transformer for Offline Hindsight Information Matching

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.682047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.682047Z digest=sha256:13401c2f858dbb4fc31d749e6793061dbfa5000925fd54a29955d82dad0142e2

Observation becba9f6-d4e7-47b3-b584-eaca283b99b6 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.686181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.686181Z digest=sha256:aec88638e565901383814b128bca79de1c4f4ffbb354dc2f993b187e4d0fd238

Observation fdef25be-0db2-4f12-a82d-bc3adcb58a4e · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.221200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.689567Z digest=sha256:c39b11072ce867e6984aef9b161bf3564465769f4fc96d122769da6fdcce4979

Observation 302823e9-b521-4a8f-919c-2d8d23f4a813 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.208953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.692976Z digest=sha256:91c13eeeace8bfaada745db2d63d656e63b1b9ae4534a003503d01ed1eb3efdb

Observation 8ca6e5eb-2068-4c52-aa7a-acf85b2f1f75 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.197096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.696297Z digest=sha256:13c6d60b13bfa587643d975cdb972e01d551ac96938172e589e1cedccd585eda

Observation 5a6a0b13-8e32-4fc3-b387-3ee9ae08df0d · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.699667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.699667Z digest=sha256:5a5c2e12f4aa69c4c398a9de8d0158edfb559a27fd80a579ac052fdbf4458f83

Observation 203b2ee3-f44c-41ef-8dac-d0c729cbe1bd · outbound

This paper cites G., Pisane, J., Kolios, A., and Ernst, D.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control G., Pisane, J., Kolios, A., and Ernst, D

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.176905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.703145Z digest=sha256:64f5f07f59882a562dac76c506b7636175a5b8a0ddfcd287ca67f0193d362b9a

Observation 4b876560-7f52-4f3a-b377-c66222a2351a · outbound

This paper cites Goal-Conditioned Reinforcement Learning: Problems and Solutions.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.706912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.706912Z digest=sha256:b8acec569ccf0b8ca89148a6b96f9ddf7798a1458b4d3c5727c48eba45cae399

Observation 00a8459a-452e-49f8-934d-8147ed3222b6 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.164422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.711006Z digest=sha256:bcd9574ad7a3e66da4f24fc29241394ff1ca9a514ac169177609ad4e0c19437d

Observation b4402b34-2019-4601-9143-c974431e8000 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.152815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.714854Z digest=sha256:58346c4878692f0b7cdc14cd05b74fc2e88c9966461a0c89ae848257b7d158bc

Observation 4d233def-ac71-4df7-b44a-b982ce9c86c8 · outbound

This paper cites A., de Lope, J., and Maravall, D.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control A., de Lope, J., and Maravall, D

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.140298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.718392Z digest=sha256:c96956ccf10b01c0c47a7e23b2a118d134dc921e2f58d411bda89a00117c2497

Observation 1a1af645-6662-47f5-99ce-265f2350f2bd · outbound

This paper cites Q-learning with online random forests.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Q-learning with online random forests

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-12T18:35:59.882991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.722299Z digest=sha256:341fb31e621fe3b1d36bac26c5e97f69c88c2a3619a2c792c771174d4291f5b5

Observation 7ceaf3c9-5804-4fa8-b4d8-1ad46ccde549 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.128061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.726465Z digest=sha256:a656550d0d8a95a6a25c9341e308d5445fee484a3ee9690554b43ffb193cc626

Observation d3a0bc44-5ee8-4a3e-addf-3c405e51ff49 · outbound

This paper cites J., and Moyà-Alcover, G.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control J., and Moyà-Alcover, G

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.116517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.730152Z digest=sha256:d18d2743cc2f966cc37efe5f2cde5adfe20839da5a5093be39a48a9895472d76

Observation 39603694-50ef-4e27-affb-57c7eb754f85 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.104713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.733683Z digest=sha256:5d6c3d16a996460758c2c4ea432ca870a9df9de1426c2471670253aa13635d1e

Observation 0d9da2dd-e1a4-49e1-b420-d93758669cf1 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.092837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.736993Z digest=sha256:a4c4bf6c7333da27046bde6afe833d07f318d92c03c48bffa2ad7419457e1731

Observation fcd2e4fa-b7db-4bfe-98e2-fe534b1db697 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.740227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.740227Z digest=sha256:47126d9d2c649a82dff740e0b03c0a766001b77305e2f188610f45daa43d2609

Observation 4666b078-bf88-4c11-bec8-d7bc8cd7ed62 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.743996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.743996Z digest=sha256:8ddf65014ec69dc007f247ce7e45243af9c049f64ee20cfb849abc2b9306ac48

Observation d470a829-5ff8-4a2f-8206-074ddbd89a5f · outbound

This paper cites A Reinforcement Learning Approach to Weaning of Mechanical Ventilation in Intensive Care Units.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control A Reinforcement Learning Approach to Weaning of Mechanical Ventilation in Intensive Care Units

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.747396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.747396Z digest=sha256:939a9a5014fdfe86abedc089b6eb073949078db340113561272e074a46a78034

Observation d6724efb-68e3-4366-b7e2-ccf4fcc47c4b · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.751444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.751444Z digest=sha256:fe28bd5bcad594a4270b86df845183f020831c14bafe2c43522a65552d7a24cf

Observation e0c38d81-767c-44ec-ab96-e3a035356c7f · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.755016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.755016Z digest=sha256:7d568153165e7f84b7a74b72853e3c5ba744d7a121ef19129e6d1b3a9064b519

Observation 7fce7cf5-155f-4588-9a64-85b414a3262b · outbound

This paper cites Deep Reinforcement Learning framework for Autonomous Driving.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Deep Reinforcement Learning framework for Autonomous Driving

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.758519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.758519Z digest=sha256:82749f2409d54b6869407f5634db482060ebbda77de27bddf107a18f7dd19b47

Observation 32f142d3-c417-4019-bb8a-747a4a27778c · outbound

This paper cites Reinforcement Learning Upside Down: Don't Predict Rewards -- Just Map Them to Actions.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Reinforcement Learning Upside Down: Don't Predict Rewards -- Just Map Them to Actions

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.762452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.762452Z digest=sha256:a04d39c2023ee04fa31a3a754f722c21022b5ac0c46c3b67cc9f530558b71a26

Observation 428271ff-1325-4d6a-a264-5acf5a065854 · outbound

This paper cites and Xie, Q.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Xie, Q

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.048189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.766504Z digest=sha256:75b3ba06d1640967de47f3ec90fa82a5b6d37e08946bf263796718749e46ba47

Observation 78b945d5-c1ce-468f-8318-7bd9334bd59b · outbound

This paper cites and Armon, A.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Armon, A

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.771028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.771028Z digest=sha256:aa340dcf65eb741d3f08b376def042add52e91f069d460f911dc8830ad1f7130

Observation 2b5cb02b-cd30-4de8-b819-8532c5117eba · outbound

This paper cites and Wang, L.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Wang, L

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.006631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.774377Z digest=sha256:cebd68096f437a9e5b6ecf97fa45ffba6b66a58b21007a6a83a25579e79a63f5

Observation 0ff21285-d901-4b47-901b-4138ce934c81 · outbound

This paper cites Training Agents using Upside-Down Reinforcement Learning.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Training Agents using Upside-Down Reinforcement Learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.778062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.778062Z digest=sha256:78b04ab5995f236005b432cec3fb7f0a5cd09c4ce13e874326f4fb51429511b4

Observation a4c67d7f-c38a-4553-b64b-7f11ccb22ce9 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:35:59.992730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.781808Z digest=sha256:d7b406570d42b78b8627b572d460b1d60653d4fd31e40de3cf07ce5fc3d61734

Observation 1ce1ff3d-6701-4dc1-8af7-e86ce1fd4fad · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:35:59.981068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.785384Z digest=sha256:dda7e25617b72fbf459733dccb8e9396da573745c66e07a1259e77e2e010c175

Observation d4032b09-f706-4d38-8406-cbed089e4fba · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:35:59.968977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.789040Z digest=sha256:35492046c44a0164d58a06d8aa8bbca28a52926891a67bbdb0e2467de42625a1

Observation 5818cd91-1ac4-4f32-9642-6f5116d0e862 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:35:59.957638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.792631Z digest=sha256:466a32dcd1db70d9ccadcd3c4d244473f1f8c7b3c8c0f3324db253ba442240d5

Observation 684369a4-35c1-47a2-8e24-5e7eecc5e9eb · outbound

This paper cites R., and Zeng, D.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control R., and Zeng, D

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:35:59.946338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.796489Z digest=sha256:998fe739054091a3a9740b8d1e44a599ca9d8fd816e24757e464fe4b2a8f5890

Pith citing papers

No inbound Pith citation observations are available.