Pith. sign in

Paper Citation Record · LEDGER

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes

As of 20 August 2026, this Paper Citation Record lists 84 of 84 outbound references and 0 inbound Pith citation observations for arXiv:2607.12924.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.12924 v2

Coverage vector

measured 84 of 84 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T06:16:29.925713Z

measured 84 of 84 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

84 of 84 outbound references displayed

  • verified exact11
  • verified fuzzy0
  • unresolved72
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a96d4f15-4c4d-486f-8fc8-1ff09fe79671 · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:19.444015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:19.444015Z digest=sha256:b86619d5d48bf3590387ba797344f2dbbf31fe96e9a53392633f109e98cdb740

Observation 3dc5bfa3-2704-40c1-88d7-0791e85a07b0 · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:19.480939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:19.480939Z digest=sha256:50c45db1730e89e46158d04b02caa0608ac0278054e0fd5e30d1bd0ddb56e64b

Observation 8078c979-7bcc-4d8a-b8e8-08acd73d62d3 · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:19.558912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:19.558912Z digest=sha256:328a6be709398c10f2475802b65a959ada85fb528ceb5885026c76b1ae6f3021

Observation 35835c41-0a11-4a8a-a688-1ed4aa8a383d · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:19.634720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:19.634720Z digest=sha256:83d0462a024a3a193973e6e2808f30877b7d4bb514bed17a2f66044878007568

Observation ce8b05d9-9f2f-48cc-8630-6abd798d3c86 · outbound

This paper cites Multi-Pass Q-Networks for Deep Reinforcement Learning with Parameterised Action Spaces.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Multi-Pass Q-Networks for Deep Reinforcement Learning with Parameterised Action Spaces

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:19.698520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:19.698520Z digest=sha256:771566d07929e4b3c25a465bcda1fa2bb78ad9f0342e0cc85d97723ba0c58a52

Observation f87c370a-9cba-433d-b6cf-238ad8bb8eb9 · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:19.773478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:19.773478Z digest=sha256:2225e78fd884762a40792a3b381a283fdda0054e2afe5fc676c4890a86f09eef

Observation a4e27dda-4282-4f84-9ecf-98343b37310a · outbound

This paper cites data-hungry.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes data-hungry

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:19.941789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:19.941789Z digest=sha256:b372e94b65dc965c917d2ee63ebcdaf36ec962c412bab08ac5cdb696bbcfeb93

Observation 60ea5feb-01f3-440c-9520-bced1c32d286 · outbound

This paper cites J.; Li, J.; Paduraru, C.; Gowal, S.; and Hester, T.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes J.; Li, J.; Paduraru, C.; Gowal, S.; and Hester, T

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:20.028042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:20.028042Z digest=sha256:3a6408d4e1c7020f4be80b60d5a5a94fefe7f06bef4cf6420684946ad5c8ff59

Observation 9fb323e3-776b-4437-ae55-d41466b7f6d1 · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:20.099420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:20.099420Z digest=sha256:bbc48af5e51517e73a04124304eed0e3cf4171c00c1b58bfa11804813a70d205

Observation be3e02c1-5b74-4cd5-9755-e9bb1c43c829 · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:20.311841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:20.311841Z digest=sha256:12badb2444914f5e8a72556a4bc3b7db8e558bf2d213ccccce8ddc0163e09bf5

Observation 918d3b6b-a748-49e8-9430-4734676c2faf · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:20.488358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:20.488358Z digest=sha256:4f0e957de46517ded90dcd576488df4f4f574a25dfcadc46ebf0e7e96067043e

Observation df943df9-9c68-481c-a845-d47ddaeedfda · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:20.649466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:20.649466Z digest=sha256:163ff4ed776287de5f77a8ceef10f039149a79351c0d0a8184a67f65a22e5e1f

Observation 9ef2f7b3-90d8-4ec0-9eba-234b8d4c0e64 · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:20.804208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:20.804208Z digest=sha256:b7288b96e138e4ee4a36fec4b828e38f867811288a089ac0843ebb64435d6330

Observation 5423079e-e498-467e-8fe6-a06266c21887 · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:20.977635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:20.977635Z digest=sha256:c06a3403c69e2c809fa67d74b9b934194b3f7a2c2eb1b66e15b83a763f5002fa

Observation f45041d7-5062-4094-858d-044fece5e7a3 · outbound

This paper cites J., and Stone, P.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes J., and Stone, P

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:21.062412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:21.062412Z digest=sha256:69f217d69ff501a0836dede5d8b1cd201dc473b088fa00a9a72f2fa2419d0f43

Observation 7c6a1d6f-1485-4665-90bd-d882b23c12db · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:21.220920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:21.220920Z digest=sha256:9eb76e2b980677c2affb06f9228f92af04345057687c5d173f27cb61da693819

Observation 57f0c4ee-f039-45d9-9c84-10c3a22b3268 · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:21.392358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:21.392358Z digest=sha256:7f55606ef4c31bd4ad549e132ad2ccf82a15b7bd0accf2fac8d7cf2f38304e6c

Observation 85438c08-c775-4f3f-9127-f213164eff0e · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:21.514470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:21.514470Z digest=sha256:d202ea965d666e77ffdaccecbb7aa8e4b1105ededf87cce3b96f70f166a1ea8d

Observation 5f2a4d72-3b0f-425f-974c-37de9596fcf2 · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:21.655508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:21.655508Z digest=sha256:5697dec679fd4000144e7ba6fecf777e382a121c213e36eb29c533a04d5545ef

Observation 9b5442d6-aba4-40ce-a811-4456d40a425e · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:21.979923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:21.979923Z digest=sha256:0c0754779b6aa4c63c0d6d65e5d8ed586169c76292f0cd5e16548b75d2d05b48

Observation 08bebbee-f6c1-4159-8e5c-7cd66d46a44b · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:22.165564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:22.165564Z digest=sha256:9f34570636d511edcf34fc236e7e3af0833635a1dbc40964159d487110a20f14

Observation 27b4a562-cc1a-4164-abd1-0a23904c8994 · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:22.337500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:22.337500Z digest=sha256:e7c21cf3f64b45b0a5bf25ab2d8cb6be080daf85a6b8ec96037013916c9dfdc0

Observation 23a88363-8bdc-4909-806e-08ba7b6bf1f1 · outbound

This paper cites B.; and Wu, J.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes B.; and Wu, J

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:22.524377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:22.524377Z digest=sha256:ca6292442b6aea628eb4b2a42e44ad048e16e8b48c8f1bc30289fa91e9f298ba

Observation e8a1d8a9-9865-4ab8-a90f-3ee01aa706a6 · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:22.716252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:22.716252Z digest=sha256:610321cf33a7bb8f3413f52792e13ae4ed0cbcbe80172f3ef013634430e4be17

Observation 451b72cd-fa6b-44b9-ae31-3533644bcdc9 · outbound

This paper cites Knowledge-Guided Exploration in Deep Reinforcement Learning.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Knowledge-Guided Exploration in Deep Reinforcement Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:22.879192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:22.879192Z digest=sha256:4976c4ad6e9b04ea330cc09280f9cc5e5b0e04992ba9b5cf5f8c0edb370dd994

Observation 269ee04a-0e9a-4421-8344-1bf0d7cbcd23 · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:22.983548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:22.983548Z digest=sha256:58f92de6086347601af8df23e8d7b79a81cd90bebc2f3611b112282c54d1d446

Observation d835f6e5-1116-4615-97cf-d9d329787255 · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:23.153134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:23.153134Z digest=sha256:6fbdff085a739d527a5f716e586fd8820a39f1eecead370a6dc4acd395e73998

Observation 4326291d-7bf1-4731-8e00-ea0ad9c8bbca · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:23.215896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:23.215896Z digest=sha256:1366d8236e967769f61822906bf34bf7140fb0a5dc109934ff32c67a22a27b9e

Observation 80444a42-96e7-49e4-bcc2-a8fe6f385efb · outbound

This paper cites A.; Veness, J.; Bellemare, M.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes A.; Veness, J.; Bellemare, M

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:23.339290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:23.339290Z digest=sha256:31c2a11965fd8af5effe3d84775fc19c3e779cb0a5dabb55ba2bfd102982167c

Observation fac7cb76-e5e9-4d1b-9cf3-4cd4aa5649bd · outbound

This paper cites M.; Broekens, J.; Plaat, A.; and Jonker, C.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes M.; Broekens, J.; Plaat, A.; and Jonker, C

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:23.475080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:23.475080Z digest=sha256:c56c863d79ad9525c5b96ab3ceabeec79f5a9da1570143bc9d5988553de8bf55

Observation 561bd2d6-8844-4fae-ae3c-125243bbfb4b · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:23.608300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:23.608300Z digest=sha256:e2decf5cb71171b1adce458551b26a4490bd4d63cd782f92d725e9870fa0ffac

Observation 91953609-23a7-4ee6-b20b-5b4c9c6bb8fe · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:23.688353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:23.688353Z digest=sha256:d6a872f1036a31a9066638986c9d2af7d0f6c6e9d0bf6c2ad036e00bf4e3d2c7

Observation 87d6f797-e2fb-40b5-9509-48c65f1fe487 · outbound

This paper cites S., and Barto, A.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes S., and Barto, A

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:23.733220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:23.733220Z digest=sha256:0bc7d1a17f37cc22e5550c63317c5f4ad0a719e9f0e4b8611af7c77a5635b9ea

Observation 4e56d103-0d48-41f6-8c1b-6f4e62d7658d · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:23.790550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:23.790550Z digest=sha256:84751c8e629c117358c96691c4a3411868f242534ec540014f1acf580bbcc35e

Observation 3eea92e5-a730-4043-b1f1-7c50dd87108e · outbound

This paper cites S.; and Niggemann, O.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes S.; and Niggemann, O

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:23.891998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:23.891998Z digest=sha256:47df09a00af231c54a626847fddf4fbc78681a42f65e0d50c291971d21c648dd

Observation 3381eb43-e10c-46f3-b64b-22f4fc8e7de9 · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:24.003850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:24.003850Z digest=sha256:b5cb25a98a91109e8fb4f68815682ab70e0a2e2084421adb1fae4c0427cc301f

Observation 1fb530e4-e3d6-49ea-8d8b-971a5eb84db9 · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:24.136790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:24.136790Z digest=sha256:a8bee2e1e342fca66d8fc6f328fed62faba82f210253dbf934d06499263230c0

Observation b7985c58-66c0-4960-ae84-dfdf34c16574 · outbound

This paper cites an unresolved cited work.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:24.272334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:24.272334Z digest=sha256:23d320df55ede8d44e1edf05bc25a427684dcef5527a28ac1d0fa53dec80e9f7

Observation f4ff7624-8ce6-48e6-b3c9-5e7f4dea3a9b · outbound

This paper cites Proceedings of the 34th International Conference on Machine Learning , pages =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Proceedings of the 34th International Conference on Machine Learning , pages =

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:24.431436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:24.431436Z digest=sha256:71016c06a585944adc288d8be2fa8b78237803fd96afd0871c2d28c7fb00980f

Observation 992ebebe-d565-4898-9526-54f6fe02cf9b · outbound

This paper cites Safe Reinforcement Learning via Shielding , volume =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Safe Reinforcement Learning via Shielding , volume =

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:24.590486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:24.590486Z digest=sha256:7ad35b3e4d36d02773069ebd011d63297311e9d916cd474962cb5c07d60b907a

Observation 47475d88-02c8-48f7-8dd3-af0f3de0cad1 · outbound

This paper cites Neurosymbolic Reinforcement Learning with Formally Verified Exploration , url =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Neurosymbolic Reinforcement Learning with Formally Verified Exploration , url =

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:24.788631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:24.788631Z digest=sha256:0b1d0070095dad34fe49548413f044429bc320d710b5ea657e72f25490e54177

Observation e221cde2-a9b0-4700-a659-241645727dea · outbound

This paper cites A review of learning planning action models , volume =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes A review of learning planning action models , volume =

Reference 43

Resolution
verified exact
doi, observed 2026-08-02T06:18:28.656625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-02T06:16:24.958560Z digest=sha256:08742c4e18c2d5846ea5b80156ccfb4aff52397a17c1eca84826a180f20d95f6

Observation c304ca6e-aadc-422d-aaee-4fbf4eee039c · outbound

This paper cites Multi-Pass Q-Networks for Deep Reinforcement Learning with Parameterised Action Spaces.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Multi-Pass Q-Networks for Deep Reinforcement Learning with Parameterised Action Spaces

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:25.077337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:25.077337Z digest=sha256:6bf171de8be0b8f97b1e620c7173290013f4425549c7b2a3468b550a79ffdd7b

Observation 79a79fca-b125-4c28-8e3f-6dbd3ff88a7b · outbound

This paper cites and Moré, Jorge J.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes and Moré, Jorge J

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:25.192724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:25.192724Z digest=sha256:4d8f82c1e23e40e1abc4b454afdfb15005f31ca01ddcb8843684e974602782b8

Observation 560045bf-1543-4dd2-a646-c1f6a84ea73f · outbound

This paper cites Symbolic Knowledge Extraction and Injection with Sub-symbolic Predictors: A Systematic Literature Review , volume =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Symbolic Knowledge Extraction and Injection with Sub-symbolic Predictors: A Systematic Literature Review , volume =

Reference 46

Resolution
verified exact
doi, observed 2026-08-02T06:18:28.466446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-02T06:16:25.306105Z digest=sha256:cf4b6799eb41e7e725e9581ac56e933fb612fcb56e806f31fe3926a27e535f4f

Observation d1e4d12e-d6fb-49d5-add2-624d34d4268e · outbound

This paper cites data-hungry.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes data-hungry

Reference 47

Resolution
verified exact
doi, observed 2026-08-02T06:18:28.259437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-02T06:16:25.404038Z digest=sha256:0ad3f5c1380874ed23c9a3f1b49bc529ce8d103b36fbc8459751d3cac22ea092

Observation 779befcc-7b3f-4535-86be-1efc7230512e · outbound

This paper cites and Li, Jerry and Paduraru, Cosmin and Gowal, Sven and Hester, Todd , year =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes and Li, Jerry and Paduraru, Cosmin and Gowal, Sven and Hester, Todd , year =

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:25.514244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:25.514244Z digest=sha256:69518b1bb9bb7b4f7ba93e30192d04296576279e92ebbabc86719b4c3760ea73

Observation f000c661-7d86-4194-835e-539452f0a346 · outbound

This paper cites Proceedings of the ECAI Workshop on AI-based Planning for Complex Real-World Applications (CAIPI 2025) , url=.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Proceedings of the ECAI Workshop on AI-based Planning for Complex Real-World Applications (CAIPI 2025) , url=

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:25.629556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:25.629556Z digest=sha256:70299b3c56b6d75432f9b08cd642eb59626dc7b85f0513508ad010cca9cf6cb9

Observation 1cee0515-d49d-48c6-a0ea-a44492009ef1 · outbound

This paper cites Proceedings of the Twenty-Eighth International Joint Conference on Artificial Intelligence,.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Proceedings of the Twenty-Eighth International Joint Conference on Artificial Intelligence,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:25.704991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:25.704991Z digest=sha256:c394e510131dd4b2baaa6d0bb5d767e10175434eccccb79e311ac935a7fb9a0e

Observation 122aad10-f1f9-4094-b765-e997a262a661 · outbound

This paper cites Adaptive Shielding via Parametric Safety Proofs , volume =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Adaptive Shielding via Parametric Safety Proofs , volume =

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:25.887044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:25.887044Z digest=sha256:33a898a4b4297269585a988a9690d1246dfd55f2a35d18e16ca5187e378ae1db

Observation 951ffbf3-47d9-4571-907e-f21fc5d9b83d · outbound

This paper cites Proceedings of the 36th International Conference on Machine Learning , pages =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Proceedings of the 36th International Conference on Machine Learning , pages =

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:26.036285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:26.036285Z digest=sha256:8b1b1683d766234946446658088081f3b60b8d269975604c38383bb0a9e505e1

Observation a035d244-546a-4870-8fab-d26738e2c81c · outbound

This paper cites Safe Reinforcement Learning via Formal Methods: Toward Safe Control Through Proof and Learning , volume=.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Safe Reinforcement Learning via Formal Methods: Toward Safe Control Through Proof and Learning , volume=

Reference 53

Resolution
verified exact
doi, observed 2026-08-02T06:18:28.001108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-02T06:16:26.129164Z digest=sha256:e2190ea33f6bc4c5369e3456edb261a0eb4f914108f61b0565f7a5b9e36cd4c8

Observation e97f83c0-42ad-471b-a34d-24c238010ca5 · outbound

This paper cites Using ontology to guide reinforcement learning agents in unseen situations: A traffic signal control system case study , volume =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Using ontology to guide reinforcement learning agents in unseen situations: A traffic signal control system case study , volume =

Reference 54

Resolution
verified exact
doi, observed 2026-08-02T06:18:27.695250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-02T06:16:26.240893Z digest=sha256:cdc5ac6be6924c6334b9c4f9597763516864a841c2a5b063acce2aa12289a8bc

Observation 7ada9c0f-48e0-463a-9c55-e537abfa362e · outbound

This paper cites Hausknecht and Peter Stone , editor =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Hausknecht and Peter Stone , editor =

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:26.410190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:26.410190Z digest=sha256:fb84136cf4289adfab0be481bbb98c6ce1ddc60f0419cb4f9f12f9f7a7504f41

Observation 3bed9127-f4eb-4c48-b31b-b39b394049d6 · outbound

This paper cites Neuro-symbolic Action Masking for Deep Reinforcement Learning , year =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Neuro-symbolic Action Masking for Deep Reinforcement Learning , year =

Reference 56

Resolution
verified exact
doi, observed 2026-08-02T06:18:27.461124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-02T06:16:26.628297Z digest=sha256:e1093aa4473e6152e273867a7d6af9a1ad0e0ada700614e5509f47e208ca6825

Observation f3815475-323d-41ad-bb2c-37c14110fe14 · outbound

This paper cites A Lazy Approach to Neural Numerical Planning with Control Parameters , ISBN =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes A Lazy Approach to Neural Numerical Planning with Control Parameters , ISBN =

Reference 57

Resolution
verified exact
doi, observed 2026-08-02T06:18:27.165978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-02T06:16:26.775548Z digest=sha256:545e687b15591c4d314cbf734a5bf697833cd4b1287706797d3f623e73e12cbe

Observation a585880f-dbf6-47e0-81c1-aada86032a37 · outbound

This paper cites Proceedings of the ECAI Workshop on AI-based Planning for Complex Real-World Applications (CAIPI 2025) , url=.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Proceedings of the ECAI Workshop on AI-based Planning for Complex Real-World Applications (CAIPI 2025) , url=

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:26.910657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:26.910657Z digest=sha256:7cea5b55e1c2cd511d90aef4435fd510f0ad5c5a82bd1a6e8ad1e59f332b5c30

Observation f177cc4d-d0ec-4d79-b29e-fac52d2edcad · outbound

This paper cites International Conference on Learning Representations , year=.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes International Conference on Learning Representations , year=

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:27.059896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:27.059896Z digest=sha256:d6e1e1fe116fd180639d644b8252c4021fc415f2ed72b8e80a95cf6b00efac0c

Observation ba70b725-516b-407c-b7bd-582019a71520 · outbound

This paper cites Shield Synthesis for Reinforcement Learning , ISBN =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Shield Synthesis for Reinforcement Learning , ISBN =

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:27.130097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:27.130097Z digest=sha256:dfdf1e6810df34376cd9028875395782253a743ed777593c9a0dba10b83571ea

Observation fa150aa4-8cba-4c5e-83b0-e7a9e87d865e · outbound

This paper cites Shields for Safe Reinforcement Learning , volume =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Shields for Safe Reinforcement Learning , volume =

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:27.202300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:27.202300Z digest=sha256:72f728560806c7d3ff843ce436fd0331405d3b0fddebae6fa4d83828ab6f0d2f

Observation 42223e18-61e3-4197-8074-54c8d05d5b4a · outbound

This paper cites ICAPS Workshop on Explainable AI Planning (XAIP) , year=.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes ICAPS Workshop on Explainable AI Planning (XAIP) , year=

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:27.275990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:27.275990Z digest=sha256:6e2a18045ae1a87d3be232a792290d846b7938345c18a6342a0e6008ed8f797e

Observation 84b6a43e-4df7-48a0-bd97-4dfa51c0d2e2 · outbound

This paper cites The ILASP system for Inductive Learning of Answer Set Programs.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes The ILASP system for Inductive Learning of Answer Set Programs

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:27.401938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:27.401938Z digest=sha256:9adf2b0380da212110fe54b06b9b5eba9410000e95d2b3de44ddeab7dcfe46a9

Observation 6ea3b983-0169-4d32-b2c4-1e5231a0f3cf · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:27.569876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:27.569876Z digest=sha256:3cf424150ddb3699e541786c92f0e34b9262a306caa9d6af9ac9c7c21cb5cb06

Observation b438d2ee-9b71-4756-8c27-276e9ca971a2 · outbound

This paper cites and Polyak, B.T.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes and Polyak, B.T

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:27.694741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:27.694741Z digest=sha256:6e1d179c2ba22b540ae1260120bc6ac21e5b7e7fb49402c858d4a3ad82d81f8c

Observation 604c1204-0829-42c5-b9a8-9c6cde9f9dda · outbound

This paper cites International Conference on Learning Representations , year =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes International Conference on Learning Representations , year =

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:27.813781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:27.813781Z digest=sha256:1a3ad04f16a3515569b28695550823405d6c2cedaf31256f16eb49e3ce67698c

Observation 57cf1f45-c90b-41a9-9b9c-55f7fb9abaab · outbound

This paper cites International Conference on Learning Representations , year=.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes International Conference on Learning Representations , year=

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:27.967488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:27.967488Z digest=sha256:da42edc86da9ef650c5ae041597d7dd12dec1df6e46e5be7575e5684af65ed17

Observation 4dad749b-3201-498c-92ba-6dddda2f6c0d · outbound

This paper cites Reinforcement Learning with Parameterized Actions , volume =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Reinforcement Learning with Parameterized Actions , volume =

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:28.074111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:28.074111Z digest=sha256:88ee6d3e6f89fb8cbe6d52b752d0d2adb9dd8e7e19d84e29aa351de8ed939ed8

Observation d3cb9ad3-dcff-461d-bb02-cc607587e66d · outbound

This paper cites Proceedings of the 2023 International Conference on Autonomous Agents and Multiagent Systems , year =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Proceedings of the 2023 International Conference on Autonomous Agents and Multiagent Systems , year =

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:28.180794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:28.180794Z digest=sha256:305c2062fa0b22192ee43dcf0f044bd5fc8d4b489ceee5eaaec902a0ba475d92

Observation 0f9e25bf-d952-4d3c-8e66-38d4fc7ee8de · outbound

This paper cites Learning Logic Specifications for Policy Guidance in POMDPs: an Inductive Logic Programming Approach , volume =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Learning Logic Specifications for Policy Guidance in POMDPs: an Inductive Logic Programming Approach , volume =

Reference 70

Resolution
verified exact
doi, observed 2026-08-02T06:18:26.868858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-02T06:16:28.321484Z digest=sha256:e92233144d01c780c9ae44a08c7b8db41307d6cd506ff92821fd9c0025eab5d7

Observation 7d4e75d7-d900-4e54-80db-dd11873a4df3 · outbound

This paper cites Knowledge-Guided Exploration in Deep Reinforcement Learning.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Knowledge-Guided Exploration in Deep Reinforcement Learning

Reference 71

Resolution
metadata mismatch
local_arxiv, observed 2026-08-02T06:18:26.576456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-02T06:16:28.549326Z digest=sha256:8dc02d6a14142834375dcb609d7b3759bc7402901140f3dbe8169b803d7306ef

Observation dc257678-15f4-45ae-bf05-9d65b10fce5d · outbound

This paper cites Explainable Reinforcement Learning: A Survey and Comparative Review , volume =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Explainable Reinforcement Learning: A Survey and Comparative Review , volume =

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:28.648918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:28.648918Z digest=sha256:ef376395c825faafb7c53f0668a1f83e1525968ac6e32b6d4f38deff6b18825f

Observation 0697f083-79e2-4358-9872-9be75676e8ce · outbound

This paper cites and Veness, Joel and Bellemare, Marc G.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes and Veness, Joel and Bellemare, Marc G

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:28.669550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:28.669550Z digest=sha256:496d1b4a6ca682abeca0f909ee0c59f0451a35ace310de22edaf3963aa5ce2ed

Observation 4f33fa6c-9036-415c-b4b2-9913daadbdf8 · outbound

This paper cites Learning Safe Numeric Action Models , volume =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Learning Safe Numeric Action Models , volume =

Reference 74

Resolution
verified exact
doi, observed 2026-08-02T06:18:26.271528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-02T06:16:28.805897Z digest=sha256:8eaa1debaf4f5dc39b6c03ec1688dbd6d40bbf23a7185e0dbcd5bfda264b9982

Observation 0c845f05-4996-46cb-8f8a-d11cd1b1e138 · outbound

This paper cites and Broekens, Joost and Plaat, Aske and Jonker, Catholijn M.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes and Broekens, Joost and Plaat, Aske and Jonker, Catholijn M

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:28.916245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:28.916245Z digest=sha256:16660d32230bb0f3490fe362a9045d589ac151175888e8ffc6bba80dcc4c4134

Observation 9f4442bf-8f33-43f3-abdd-a8a5b53ca883 · outbound

This paper cites Mastering the game of Go without human knowledge , volume =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Mastering the game of Go without human knowledge , volume =

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:29.095114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:29.095114Z digest=sha256:6366a52cd2cbd5424b22ccbbe6bcbb48cdca9514982ca36d9be41d08c0fc5325

Observation fb08ccc2-b2ec-4b55-94e6-d7c9053b5eb0 · outbound

This paper cites A general reinforcement learning algorithm that masters chess, shogi, and Go through self-play , volume =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes A general reinforcement learning algorithm that masters chess, shogi, and Go through self-play , volume =

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:29.234694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:29.234694Z digest=sha256:8687a73ec777953859d3d2a9b1b2b9787330ac558dd0aceff279df2bf0450d17

Observation 099a3f27-6003-4204-8cb5-176d26418cd5 · outbound

This paper cites 2018 , publisher=.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes 2018 , publisher=

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:29.294079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:29.294079Z digest=sha256:fb8a354a2315b25755e23720662ebf8ed3a82448683f4489114bea165e15fe44

Observation cdc73c9d-f5f0-4e1c-8755-31b746536327 · outbound

This paper cites Sample-Efficient Neurosymbolic Deep Reinforcement Learning.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Sample-Efficient Neurosymbolic Deep Reinforcement Learning

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:29.386639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:29.386639Z digest=sha256:9732d4864b3fa0679cf50c7ce352fa7c14af34e99c3b116e561c1787c5623714

Observation 635f1335-3841-41b9-89cd-efcbcfb9fb12 · outbound

This paper cites 35th International Conference on Principles of Diagnosis and Resilient Systems (DX 2024) , pages =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes 35th International Conference on Principles of Diagnosis and Resilient Systems (DX 2024) , pages =

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:29.493025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:29.493025Z digest=sha256:c4d59fd620ba90a08dede9bc9d804a718c144ad61b1986412137ffa75243398c

Observation 5f576683-d002-4caf-9902-fb5c2d8f6ff0 · outbound

This paper cites Scalable Planning with Tensorflow for Hybrid Nonlinear Domains , volume =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Scalable Planning with Tensorflow for Hybrid Nonlinear Domains , volume =

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:29.534594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:29.534594Z digest=sha256:242a07836042a005639f8d5f92c60530f23844d64f5ba5e30d8818e33f8a9975

Observation b208660a-98ef-4f51-8e7e-9db66ce85f45 · outbound

This paper cites Parametrized Deep Q-Networks Learning: Reinforcement Learning with Discrete-Continuous Hybrid Action Space.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Parametrized Deep Q-Networks Learning: Reinforcement Learning with Discrete-Continuous Hybrid Action Space

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:29.630219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:29.630219Z digest=sha256:a78ebae4f0bcd66db3740d9df2a7779e3adfffb6082ded695fd377caa97f64b2

Observation d92f5e27-9ed6-4433-984c-23bc4131e333 · outbound

This paper cites Towards Sample Efficient Reinforcement Learning , url =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Towards Sample Efficient Reinforcement Learning , url =

Reference 83

Resolution
verified exact
doi, observed 2026-08-02T06:18:25.930208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-02T06:16:29.734104Z digest=sha256:3a48933c3f8329f76631678f42f6b79774baa8f01dc68da48f8354fe4bd3b17d

Observation c8795eb2-fc94-49b1-9d7e-db6c3f8384be · outbound

This paper cites An Overview of the Action Space for Deep Reinforcement Learning , DOI =.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes An Overview of the Action Space for Deep Reinforcement Learning , DOI =

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:29.846227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:16:29.846227Z digest=sha256:fc0866712db8adca889ffc5d567851be50fc6f532170e179794eb8d05d802cbb

Observation 08161f32-55c0-462b-8fc2-95cea9a50f10 · outbound

This paper cites Model-lite planning: Case-based vs.

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Model-lite planning: Case-based vs

Reference 85

Resolution
verified exact
doi, observed 2026-08-02T06:18:25.607267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-02T06:16:29.925713Z digest=sha256:63b1c16c4ae59e64d2840cc6d8c800405c705596d7cec8e7c6e9f0f77bf98793

Pith citing papers

No inbound Pith citation observations are available.