Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T06:16:29.925713Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 84 of 84 outbound references and 0 inbound Pith citation observations for arXiv:2607.12924.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T06:16:29.925713Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
84 of 84 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a96d4f15-4c4d-486f-8fc8-1ff09fe79671 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dc5bfa3-2704-40c1-88d7-0791e85a07b0 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8078c979-7bcc-4d8a-b8e8-08acd73d62d3 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35835c41-0a11-4a8a-a688-1ed4aa8a383d · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce8b05d9-9f2f-48cc-8630-6abd798d3c86 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Multi-Pass Q-Networks for Deep Reinforcement Learning with Parameterised Action Spaces
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f87c370a-9cba-433d-b6cf-238ad8bb8eb9 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4e27dda-4282-4f84-9ecf-98343b37310a · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes data-hungry
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60ea5feb-01f3-440c-9520-bced1c32d286 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes J.; Li, J.; Paduraru, C.; Gowal, S.; and Hester, T
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fb323e3-776b-4437-ae55-d41466b7f6d1 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be3e02c1-5b74-4cd5-9755-e9bb1c43c829 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 918d3b6b-a748-49e8-9430-4734676c2faf · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df943df9-9c68-481c-a845-d47ddaeedfda · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ef2f7b3-90d8-4ec0-9eba-234b8d4c0e64 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5423079e-e498-467e-8fe6-a06266c21887 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f45041d7-5062-4094-858d-044fece5e7a3 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes J., and Stone, P
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c6a1d6f-1485-4665-90bd-d882b23c12db · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57f0c4ee-f039-45d9-9c84-10c3a22b3268 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85438c08-c775-4f3f-9127-f213164eff0e · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f2a4d72-3b0f-425f-974c-37de9596fcf2 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b5442d6-aba4-40ce-a811-4456d40a425e · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08bebbee-f6c1-4159-8e5c-7cd66d46a44b · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27b4a562-cc1a-4164-abd1-0a23904c8994 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23a88363-8bdc-4909-806e-08ba7b6bf1f1 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes B.; and Wu, J
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8a1d8a9-9865-4ab8-a90f-3ee01aa706a6 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 451b72cd-fa6b-44b9-ae31-3533644bcdc9 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Knowledge-Guided Exploration in Deep Reinforcement Learning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 269ee04a-0e9a-4421-8344-1bf0d7cbcd23 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d835f6e5-1116-4615-97cf-d9d329787255 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4326291d-7bf1-4731-8e00-ea0ad9c8bbca · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80444a42-96e7-49e4-bcc2-a8fe6f385efb · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes A.; Veness, J.; Bellemare, M
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fac7cb76-e5e9-4d1b-9cf3-4cd4aa5649bd · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes M.; Broekens, J.; Plaat, A.; and Jonker, C
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 561bd2d6-8844-4fae-ae3c-125243bbfb4b · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91953609-23a7-4ee6-b20b-5b4c9c6bb8fe · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87d6f797-e2fb-40b5-9509-48c65f1fe487 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes S., and Barto, A
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e56d103-0d48-41f6-8c1b-6f4e62d7658d · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3eea92e5-a730-4043-b1f1-7c50dd87108e · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes S.; and Niggemann, O
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3381eb43-e10c-46f3-b64b-22f4fc8e7de9 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fb530e4-e3d6-49ea-8d8b-971a5eb84db9 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7985c58-66c0-4960-ae84-dfdf34c16574 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Unresolved cited work
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4ff7624-8ce6-48e6-b3c9-5e7f4dea3a9b · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Proceedings of the 34th International Conference on Machine Learning , pages =
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 992ebebe-d565-4898-9526-54f6fe02cf9b · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Safe Reinforcement Learning via Shielding , volume =
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47475d88-02c8-48f7-8dd3-af0f3de0cad1 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Neurosymbolic Reinforcement Learning with Formally Verified Exploration , url =
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e221cde2-a9b0-4700-a659-241645727dea · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes A review of learning planning action models , volume =
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c304ca6e-aadc-422d-aaee-4fbf4eee039c · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Multi-Pass Q-Networks for Deep Reinforcement Learning with Parameterised Action Spaces
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79a79fca-b125-4c28-8e3f-6dbd3ff88a7b · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes and Moré, Jorge J
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 560045bf-1543-4dd2-a646-c1f6a84ea73f · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Symbolic Knowledge Extraction and Injection with Sub-symbolic Predictors: A Systematic Literature Review , volume =
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d1e4d12e-d6fb-49d5-add2-624d34d4268e · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes data-hungry
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 779befcc-7b3f-4535-86be-1efc7230512e · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes and Li, Jerry and Paduraru, Cosmin and Gowal, Sven and Hester, Todd , year =
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f000c661-7d86-4194-835e-539452f0a346 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Proceedings of the ECAI Workshop on AI-based Planning for Complex Real-World Applications (CAIPI 2025) , url=
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1cee0515-d49d-48c6-a0ea-a44492009ef1 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Proceedings of the Twenty-Eighth International Joint Conference on Artificial Intelligence,
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 122aad10-f1f9-4094-b765-e997a262a661 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Adaptive Shielding via Parametric Safety Proofs , volume =
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 951ffbf3-47d9-4571-907e-f21fc5d9b83d · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Proceedings of the 36th International Conference on Machine Learning , pages =
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a035d244-546a-4870-8fab-d26738e2c81c · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Safe Reinforcement Learning via Formal Methods: Toward Safe Control Through Proof and Learning , volume=
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e97f83c0-42ad-471b-a34d-24c238010ca5 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Using ontology to guide reinforcement learning agents in unseen situations: A traffic signal control system case study , volume =
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7ada9c0f-48e0-463a-9c55-e537abfa362e · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Hausknecht and Peter Stone , editor =
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bed9127-f4eb-4c48-b31b-b39b394049d6 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Neuro-symbolic Action Masking for Deep Reinforcement Learning , year =
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f3815475-323d-41ad-bb2c-37c14110fe14 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes A Lazy Approach to Neural Numerical Planning with Control Parameters , ISBN =
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a585880f-dbf6-47e0-81c1-aada86032a37 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Proceedings of the ECAI Workshop on AI-based Planning for Complex Real-World Applications (CAIPI 2025) , url=
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f177cc4d-d0ec-4d79-b29e-fac52d2edcad · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes International Conference on Learning Representations , year=
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba70b725-516b-407c-b7bd-582019a71520 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Shield Synthesis for Reinforcement Learning , ISBN =
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa150aa4-8cba-4c5e-83b0-e7a9e87d865e · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Shields for Safe Reinforcement Learning , volume =
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42223e18-61e3-4197-8074-54c8d05d5b4a · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes ICAPS Workshop on Explainable AI Planning (XAIP) , year=
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84b6a43e-4df7-48a0-bd97-4dfa51c0d2e2 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes The ILASP system for Inductive Learning of Answer Set Programs
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ea3b983-0169-4d32-b2c4-1e5231a0f3cf · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b438d2ee-9b71-4756-8c27-276e9ca971a2 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes and Polyak, B.T
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 604c1204-0829-42c5-b9a8-9c6cde9f9dda · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes International Conference on Learning Representations , year =
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57cf1f45-c90b-41a9-9b9c-55f7fb9abaab · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes International Conference on Learning Representations , year=
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dad749b-3201-498c-92ba-6dddda2f6c0d · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Reinforcement Learning with Parameterized Actions , volume =
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3cb9ad3-dcff-461d-bb02-cc607587e66d · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Proceedings of the 2023 International Conference on Autonomous Agents and Multiagent Systems , year =
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f9e25bf-d952-4d3c-8e66-38d4fc7ee8de · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Learning Logic Specifications for Policy Guidance in POMDPs: an Inductive Logic Programming Approach , volume =
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7d4e75d7-d900-4e54-80db-dd11873a4df3 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Knowledge-Guided Exploration in Deep Reinforcement Learning
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation dc257678-15f4-45ae-bf05-9d65b10fce5d · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Explainable Reinforcement Learning: A Survey and Comparative Review , volume =
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0697f083-79e2-4358-9872-9be75676e8ce · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes and Veness, Joel and Bellemare, Marc G
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f33fa6c-9036-415c-b4b2-9913daadbdf8 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Learning Safe Numeric Action Models , volume =
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0c845f05-4996-46cb-8f8a-d11cd1b1e138 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes and Broekens, Joost and Plaat, Aske and Jonker, Catholijn M
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f4442bf-8f33-43f3-abdd-a8a5b53ca883 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Mastering the game of Go without human knowledge , volume =
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb08ccc2-b2ec-4b55-94e6-d7c9053b5eb0 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes A general reinforcement learning algorithm that masters chess, shogi, and Go through self-play , volume =
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 099a3f27-6003-4204-8cb5-176d26418cd5 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes 2018 , publisher=
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdc73c9d-f5f0-4e1c-8755-31b746536327 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Sample-Efficient Neurosymbolic Deep Reinforcement Learning
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 635f1335-3841-41b9-89cd-efcbcfb9fb12 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes 35th International Conference on Principles of Diagnosis and Resilient Systems (DX 2024) , pages =
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f576683-d002-4caf-9902-fb5c2d8f6ff0 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Scalable Planning with Tensorflow for Hybrid Nonlinear Domains , volume =
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b208660a-98ef-4f51-8e7e-9db66ce85f45 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Parametrized Deep Q-Networks Learning: Reinforcement Learning with Discrete-Continuous Hybrid Action Space
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d92f5e27-9ed6-4433-984c-23bc4131e333 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Towards Sample Efficient Reinforcement Learning , url =
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c8795eb2-fc94-49b1-9d7e-db6c3f8384be · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes An Overview of the Action Space for Deep Reinforcement Learning , DOI =
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08161f32-55c0-462b-8fc2-95cea9a50f10 · outbound
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes Model-lite planning: Case-based vs
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
No inbound Pith citation observations are available.