Pith. sign in

Paper Citation Record · LEDGER

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity

As of 9 August 2026, this Paper Citation Record lists 100 of 154 outbound references and 1 inbound Pith citation observation for arXiv:2602.08690.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2602.08690 v2

Coverage vector

measured 100 of 154 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T03:15:10.053731Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T16:46:41.137112Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T08:11:03.819865Z

Reference resolution

100 of 154 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved100
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation df3ff68c-83b4-483a-a115-e33081f5252d · outbound

This paper cites Created by Maxwell Standen, David Bowman, Son Hoang, Toby Richer, Martin Lucas, Richard Van Tassel.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Created by Maxwell Standen, David Bowman, Son Hoang, Toby Richer, Martin Lucas, Richard Van Tassel

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.778365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.778365Z digest=sha256:d3785c032e71c2298d8a626e98bf2e0fee9cdbeb2aaccc9244630d9b29e8410b

Observation 5f82f4dc-a06c-491f-afa0-bd4a1bb0f44e · outbound

This paper cites Created by Maxwell Standen, David Bowman, Son Hoang, Toby Richer, Martin Lucas, Richard Van Tassel, Phillip Vu, Mitchell Kiely.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Created by Maxwell Standen, David Bowman, Son Hoang, Toby Richer, Martin Lucas, Richard Van Tassel, Phillip Vu, Mitchell Kiely

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.782041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.782041Z digest=sha256:4395aba1a87cbdfe3747d3afdcee031e285ca9fdb33ddb54e59ba21e9f6f7aef

Observation e358b6e8-44b0-46a8-b5fe-68fade36656e · outbound

This paper cites Adkins, M.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Adkins, M

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.785117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.785117Z digest=sha256:a3a6f36ded03e8073341b6a13be1bfe47f4b3152e86f6b704f513f385dea3d2b

Observation 07bbf896-35a3-4a88-910e-3de8b5b78183 · outbound

This paper cites Agarwal, M.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Agarwal, M

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.788291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.788291Z digest=sha256:483ed046e523e9163a3c8d62c2e434adc2ad6ef40b2225b20767c79f5ec27e7f

Observation 08db024e-e004-4268-a85c-37e4b7c4647a · outbound

This paper cites Agarwal, M.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Agarwal, M

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.791249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.791249Z digest=sha256:8a51e5a0c4d32d7ef6306cf85686ace07e8fb8b00ece50bd65a3b7554ea469c6

Observation 1b74d2c5-6c07-45c4-a1ac-c35ac2306ca6 · outbound

This paper cites Al-Fawa’reh, J.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Al-Fawa’reh, J

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.794347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.794347Z digest=sha256:aa814be9c74498ecfbc429fc409e64f8156b82bd744b31867b9d69ea8e56f65e

Observation 7527c350-ba1d-4ed0-ab93-a3d072ea3a37 · outbound

This paper cites Al Wahaibi, M.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Al Wahaibi, M

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.797853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.797853Z digest=sha256:9bca3e0e31d420b5bf02cc2ce557ae3a1b24e7c21ddb4dd1d73520ab7d48ce39

Observation 5bd499f2-18ac-4c15-a8a3-0eab06669f29 · outbound

This paper cites Learning to Evade Static PE Machine Learning Malware Models via Reinforcement Learning.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Learning to Evade Static PE Machine Learning Malware Models via Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.800647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.800647Z digest=sha256:68eef72442b32284bd1d379754b00e75a7fdb41f85aeb060ff0d89f2412b5d27

Observation ff357e28-9a52-4ed9-ae1f-9e75fe0a80ac · outbound

This paper cites Andrew, S.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Andrew, S

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.803956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.803956Z digest=sha256:4381f6bf75589d54bae0ca1733477ea5739c966dfe21605b927082205965106f

Observation 5afa8bc7-632d-40c6-bd22-af541538804b · outbound

This paper cites Applebaum, C.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Applebaum, C

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.807023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.807023Z digest=sha256:45500d4fecd09fc3cac9171d697637f4e9ce5ee81593ddc40eb5cf0ed505498f

Observation ae3597b7-bf08-4695-86ce-7ec574d648ce · outbound

This paper cites Apruzzese, M.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Apruzzese, M

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.810020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.810020Z digest=sha256:b49a9b2c4bca6009dc7a5fcd229e1c6afcd176dcb806519583fb68c9a06e3b97

Observation bbc9e51b-782e-4a13-ba27-972c6d438c79 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.812800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.812800Z digest=sha256:9660c305bc7b023ef9084011b3df8eb6a7bb3b5ce735df1723443012c7345539

Observation 23a0fefa-e378-4f03-8f07-a097450a28f1 · outbound

This paper cites Arulkumaran, M.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Arulkumaran, M

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.815379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.815379Z digest=sha256:30bddff6536d5671234c80c34b46a5327b7d02e01647b66cf143bed7895225c6

Observation 1836a92d-8f62-4881-ac2b-e0eba830df20 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.818088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.818088Z digest=sha256:7dae9a8cf0ac81f14629ae5226d97ad5ceb410a26e5387f7801b3617b15817a9

Observation b3399342-1a07-4405-9853-ccab3a6ba344 · outbound

This paper cites Bates, V.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Bates, V

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.821336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.821336Z digest=sha256:d19a4a90ed190de15fa766aa22bb5fc2ba36f9bc27444d012308cf63e2e2feff

Observation f6d8a231-e664-49cf-b92b-ece5b12bad61 · outbound

This paper cites Beyond rewards in reinforcement learning for cyber de- fence, 2026.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Beyond rewards in reinforcement learning for cyber de- fence, 2026

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.823908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.823908Z digest=sha256:6b5af6065ddb7a5177ff0544beb28ed03a5c51127147e24355ca479913e9e6e0

Observation ce9b109d-b0b8-4101-bb63-94cd96f05656 · outbound

This paper cites Dota 2 with Large Scale Deep Reinforcement Learning.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Dota 2 with Large Scale Deep Reinforcement Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.826615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.826615Z digest=sha256:c8a75ecdda6ce160713aafab0d62ac4464777448a2298c88f6c46b107f1d3167

Observation 30fef07e-8202-4a08-b6a0-0b1019e8c4e0 · outbound

This paper cites B ¨ottinger, P.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity B ¨ottinger, P

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.830167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.830167Z digest=sha256:e463aa7c799c23345750b2f20c81c7c9c66cd8c10bc05994cf4c7a7526ce3a31

Observation 88b7e6d1-e67a-4242-9d50-0f3d91e35306 · outbound

This paper cites Boutilier, T.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Boutilier, T

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.832904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.832904Z digest=sha256:20f1c3d4ba524c4963f2b369173956e159646f2c3587057ccf45974540269400

Observation 272143ee-e566-4bab-af68-261851b6c40a · outbound

This paper cites Exploration by Random Network Distillation.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Exploration by Random Network Distillation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.835966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.835966Z digest=sha256:b5605eff074d22f69d46736b208fb3a177cbdc438c77e761de8122abb3faeb0d

Observation 9d1c3226-d612-4f55-a1a9-a882299d361d · outbound

This paper cites Caminero, M.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Caminero, M

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.839239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.839239Z digest=sha256:555415e04f6b78b3fad024cae2507ed008a1e79b8df06db1f89881c0cec1db88

Observation f8edf2b0-821f-4bb1-a06c-9f9d4706b2fe · outbound

This paper cites Cavenaghi, G.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Cavenaghi, G

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.842252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.842252Z digest=sha256:2c7442b86d3324eac61c23f55de0b375baf68acdf1ed07c4d54806d7350a3265

Observation 339e7f89-55ec-4de5-a959-0e9b60b36113 · outbound

This paper cites Measuring the Reliability of Reinforcement Learning Algorithms.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Measuring the Reliability of Reinforcement Learning Algorithms

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.845886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.845886Z digest=sha256:240e13b9f845cee816d32d49d08723f47c6b4129345be668211509d5a2fdb77c

Observation 36b3c8db-827a-4a43-a6fd-849410eace45 · outbound

This paper cites Chatterjee and Akbar-Siami Namin.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Chatterjee and Akbar-Siami Namin

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.849093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.849093Z digest=sha256:d3b25cde8c9ad0f326519962219fe15bb951e2972497b52dae74cd2db7de7cb8

Observation 815dfd8d-bee5-4a0e-9154-8ae45eb75856 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.851992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.851992Z digest=sha256:d03e982778e345f175ea85aa798a92cd3d572bf766cfa40ba2fef60c25c3f5f1

Observation 004f208c-dc06-4be1-92e2-8dca16de6482 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.854750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.854750Z digest=sha256:d512175200a4b860882aa91a478dee5d439bf9d6c0dc630ade2b49753e017e8a

Observation 93d20245-72c5-46ae-852f-0997fb39a085 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.857706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.857706Z digest=sha256:0dd687aea7e02c5e5f4abfb929ea049e61a14a5cb6b0c35bd6f0be6fe460185e

Observation 443f22e1-53e7-45f2-a35c-fedc600bf774 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.860411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.860411Z digest=sha256:30cdab7de0fe4d5890e6f636f90961bb43380e17e64837523ad1264bd9e9ca18

Observation 1b75d5b0-1b77-4ad5-93c2-a7f4565b0f57 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.863577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.863577Z digest=sha256:546cb6f122eb65887c3069fe3603765f21cf20b594c50253006841537077f02c

Observation d7de28be-5ed1-4ebf-8e4c-49584431cb1a · outbound

This paper cites Colas, O.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Colas, O

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.866140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.866140Z digest=sha256:16fe1f55403f6b9f52275aa35b42109a09b576e921e97260ccf052185066b87b

Observation 82caaba0-c05f-4012-a02c-65932f2ba4ef · outbound

This paper cites Collyer, A.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Collyer, A

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.868781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.868781Z digest=sha256:260ff6800646cd7fdc397ca519540f6fb80121b53b19611b1a8d790c98710122

Observation 6b77f1e3-17c1-49d9-933b-9e476596d17c · outbound

This paper cites Corradini, Z.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Corradini, Z

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.871654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.871654Z digest=sha256:52c3f6f28f47bebf4fc27f41dc8162c377e87767057e5c04c51d3e929db1e3f6

Observation bf1931f4-3d43-414b-becf-b0611c8bb1c1 · outbound

This paper cites Dabney, G.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Dabney, G

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.874169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.874169Z digest=sha256:2584ded3c66da7a2ad3c13be2daab8d55c10a42d3b1f587cf95906f0b9827b91

Observation c42b846f-f353-4b8c-bbac-bb5dc198ed72 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.877023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.877023Z digest=sha256:1446dbd6d3e21b936307804992b22929fa335f1be4b3cfa7e1f2c22d24cff9b8

Observation c733322b-881b-4f87-afa1-031b5d43b1c2 · outbound

This paper cites De Silva, W.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity De Silva, W

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.879906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.879906Z digest=sha256:2794adfa8f3cd2554a992cc8b1004d20fb7cb7b8782d257f81a17084b71089c8

Observation fe5d7787-dbdc-4255-9eef-2bb2f4dcd305 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.882494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.882494Z digest=sha256:85dd7a5ea279c23869660ca62bde2f60e0e94e09ea3599d78d6f8c064567d6b1

Observation 59dbbaaf-e995-4f0d-a1b5-1fcf4c512824 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.885279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.885279Z digest=sha256:55bf55d26f7803774e5a420120925c8d3e696dfec368dc2b74520afe4f208668

Observation 0e5b0c62-485a-461c-b3d6-dd93fc63e6eb · outbound

This paper cites Challenges of Real-World Reinforcement Learning.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Challenges of Real-World Reinforcement Learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.888113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.888113Z digest=sha256:f14d48ceeee48d46989f88a38a62601e581c6cd2c26e8b877dec75a976603831

Observation 4de0b8f9-2d2c-4464-a7e0-d9b4dadb10d6 · outbound

This paper cites Eimer, M.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Eimer, M

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.891389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.891389Z digest=sha256:5b55e61eb6370e7ec4d45f8c7c62ac48a3742b5f81568f84a2a206207c25c439

Observation 0b989531-b9b7-41d0-b865-2e93c2fa0101 · outbound

This paper cites Emerson, L.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Emerson, L

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.894002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.894002Z digest=sha256:94db5c992812329c27e8b70e14ec410fe0ba92a32e944ab8893491f8d49dce9c

Observation e5b96df9-4520-4db0-bba1-4ac31eb51971 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.896674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.896674Z digest=sha256:125bff8562c692024893b2cc39c3af3123858da1b13050d495c20def1d1f8a96

Observation f6106954-54f2-42e2-9b3c-18ee0b27b29f · outbound

This paper cites Erd ˝odi, ˚A Sommervoll, and F.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Erd ˝odi, ˚A Sommervoll, and F

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.899929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.899929Z digest=sha256:e6a5b94ffe3ec09c5d58b043875182944c7956a0fa117e392499d90877c8cb45

Observation 17dfed62-7d7e-4d59-96ac-20b310881ec1 · outbound

This paper cites Evertz, N.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Evertz, N

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.902517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.902517Z digest=sha256:7080de67c776bce9b3f9875bcff88674e6e351c7fd43c021e396676380ab875b

Observation b60b782b-77d6-4c3d-a7b9-d94eeb30fcba · outbound

This paper cites Faillon, B.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Faillon, B

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.905271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.905271Z digest=sha256:078fd3d445b1627ffbb8e9e5fb5be513e6053b609cc256c7d615e08d066486b8

Observation 311b314e-ea77-4a55-9fdd-cfbfe1f452e7 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.907675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.907675Z digest=sha256:75ca397811e85767791a1f4741dc001f12a3940bb46925f4b314ce83adf1cb88

Observation d50d30e8-a9c6-4897-938a-a12a75f4724c · outbound

This paper cites Fawzi, M.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Fawzi, M

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.910359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.910359Z digest=sha256:11d30f9e7eddc3d3465362b29939c39823fd190ad1136530a557488faa664414

Observation e23ab54d-1dfa-41f4-ae43-0e867b8b7d5c · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.913094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.913094Z digest=sha256:16b3741c3c5ac75893192ebe5fca88362e62873426bf94ad324204f34e0ca434

Observation b0b97e3a-dcb0-46e4-9173-12ecd4aad9f3 · outbound

This paper cites Foley, C.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Foley, C

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.915649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.915649Z digest=sha256:1f8ae2e006d337a503a473859fe74430f44df51491ac72e9b518a025629cf854

Observation 9693a93e-58f0-41d9-8c55-faeb13ec7fd2 · outbound

This paper cites Foley and S.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Foley and S

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.918347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.918347Z digest=sha256:66c59549df1bdedf0d2ab9bfdaa06b653e00a8ad8c28f5844e3e7b9439f646c1

Observation c23b3602-fed9-44fd-abb6-6777788c3164 · outbound

This paper cites Foley and S.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Foley and S

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.921523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.921523Z digest=sha256:7a7df2134339a624dd8b4749b04144555356926230e30a3e9aeba71a254885d5

Observation d1c7d651-99e6-4063-8663-4a24e29f2029 · outbound

This paper cites Foley, M.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Foley, M

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.923907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.923907Z digest=sha256:581fd54397ba217d87c5a18d6a3fc050ed8f5f388c2e36472ccc5869ea7303f7

Observation d453ac1f-647b-4e35-b91f-25bc5bf7ef0d · outbound

This paper cites Gangupantulu, T.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Gangupantulu, T

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.926619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.926619Z digest=sha256:878fddb5d49ff32dfffdb548375fd537037f2bf1c69de34cde06b41b5c24f7e2

Observation 15c259b0-7f1e-4ef1-94e3-72f758120ec0 · outbound

This paper cites Adversarial Policies: Attacking Deep Reinforcement Learning.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Adversarial Policies: Attacking Deep Reinforcement Learning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.929179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.929179Z digest=sha256:3c63d78f791ca7fd5d43056cc540d2d055731dd9e775ba7e19fa22ba1a66be57

Observation 53a628a1-05c1-4bfc-9129-8d081c9dedc5 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.932100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.932100Z digest=sha256:90b4a8daecc7786b77ecb4b33204bd532117b479c3c686d93996d27e68141f80

Observation 13c21b26-3899-4fb8-9a2b-ddf96b19d400 · outbound

This paper cites Gohil, H.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Gohil, H

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.934501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.934501Z digest=sha256:934329007d89278df66c67b5e31e0074cfce1a68c3a2d37da88e94fc8f6a77ba

Observation fb79d636-6b3b-4f44-aa0a-4d3f9e5bbf3d · outbound

This paper cites Gohil, S.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Gohil, S

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.936931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.936931Z digest=sha256:ed13f8020d965cdf7a4537ed060985b0af32f41efbe981f0f415d023ce6bac91

Observation 3b322cc0-c95e-459a-8b3c-cff8e684c6fa · outbound

This paper cites Ttcp cage chal- lenge 3.https://github.com/cage-challenge/ cage-challenge-3, 2022.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Ttcp cage chal- lenge 3.https://github.com/cage-challenge/ cage-challenge-3, 2022

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.939350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.939350Z digest=sha256:888963e7fee8661c004b74f13027b692e90637c763088d07d7ee78930e77fc60

Observation c9595248-4214-4c35-999c-2554e6e30663 · outbound

This paper cites Ttcp cage chal- lenge 4.https://github.com/cage-challenge/ cage-challenge-4, 2023.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Ttcp cage chal- lenge 4.https://github.com/cage-challenge/ cage-challenge-4, 2023

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.941848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.941848Z digest=sha256:fc009236dc442a7a1a83429778741795bc2e066769a6d5a37fcfd71a26f85d77

Observation 26e1da50-7ac6-4a62-b337-9eebe51556e7 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.944258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.944258Z digest=sha256:29b9e54fac2a5245da5adcbc293e3374947c4f0409b1bf78a51e79ec3d83bc64

Observation 5bd72e83-e30f-4f22-aaa1-e79a8f968694 · outbound

This paper cites Hausknecht and P.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Hausknecht and P

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.946632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.946632Z digest=sha256:1e2bf442401b79db61f537bcec819cad36400b9fefffcdb7071d6da39215df98

Observation e83c17e8-5f2e-4fc3-acb2-267050f33ad5 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.949150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.949150Z digest=sha256:a48cd6b236c9a228591aeef8bf784b1cc7da7c11ed2ee9a88d667f001fd1c8ef

Observation 376e0d53-2a32-4fa0-a83e-b9623a867c6c · outbound

This paper cites Henderson, R.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Henderson, R

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.951461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.951461Z digest=sha256:38301464828b98c43c64c5dcb0145a693083d39d7880064e313b67235062dcd5

Observation 92cce13c-e720-41bc-969f-9264d1278750 · outbound

This paper cites On Inductive Biases in Deep Reinforcement Learning.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity On Inductive Biases in Deep Reinforcement Learning

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.953851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.953851Z digest=sha256:deadd294bdf1255a2d25bc307bb790d98969db5cd8ef4293ad96f287db3bbda9

Observation 370e90af-38dc-4d56-b6af-f97d61b9508c · outbound

This paper cites Hicks, V.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Hicks, V

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.956469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.956469Z digest=sha256:94c474f383d6d1fbd3dca5b237c059caf0c0c0a1a3fab1c3c277cdf14f94bf31

Observation f721aac8-6fd6-4010-89f7-af069076397f · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.958827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.958827Z digest=sha256:469e3960fadb74d45648547ea5b79aeceb3e158f205788d7710f3b00708fb99a

Observation e9d00cfd-ccce-4741-ae37-0925c1efdb6b · outbound

This paper cites SquirRL: Automating Attack Analysis on Blockchain Incentive Mechanisms with Deep Reinforcement Learning.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity SquirRL: Automating Attack Analysis on Blockchain Incentive Mechanisms with Deep Reinforcement Learning

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.961548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.961548Z digest=sha256:a4f337d789a178f013b9eff4b8148fcdb6f9e66c0f8e4abde0c559f4719f84e8

Observation 7a52a833-700e-4b64-b623-46c3be3511cf · outbound

This paper cites Hsu and M.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Hsu and M

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.964133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.964133Z digest=sha256:d8bf52a3601d8091545a2351b358c3d6280c4476d4688e9c5ba49a201922d972

Observation 6d276a13-3349-4858-95be-d866acd43687 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.966775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.966775Z digest=sha256:d76007b1e5aefe48ab8789b565be8e03744b692d52393cc3c18bae39401c8fdc

Observation 12a65293-f836-4f4d-afed-8f89ca7eca1f · outbound

This paper cites Jumper, R.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Jumper, R

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.969146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.969146Z digest=sha256:8cd00465a047145133c8b666b5abb709d1e80180455c27afa4923f4fab31f5e2

Observation a4bcdbcd-8485-4da5-a2c6-808fce26cc4e · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.971677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.971677Z digest=sha256:42c4ccf4c071ded8ec8d368c15c19456144ee174ec8040c918f9db53c6618a33

Observation 0e6e0af5-d3a0-47eb-a7d8-b781195bb902 · outbound

This paper cites Model-Based Reinforcement Learning for Atari.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Model-Based Reinforcement Learning for Atari

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.974165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.974165Z digest=sha256:a607d96529be79aa01c450d64b4632b2ecda7c67972b9b4993a73de8170770c8

Observation 21715484-6bf0-4f8f-83db-ba06b7fc81bd · outbound

This paper cites TESSERACT: Eliminating Experimental Bias in Malware Classification across Space and Time (Extended Version).

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity TESSERACT: Eliminating Experimental Bias in Malware Classification across Space and Time (Extended Version)

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.977417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.977417Z digest=sha256:efefd0e09d37a8f9361fc08918d8e544c92fcb653bb7dc72007e2939770ac99a

Observation efe29ff9-21e5-41c0-8750-d9f3598b59d9 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.980231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.980231Z digest=sha256:09feb1fa431217b4d51bef8ae3528af8c50736dc1e46868cdb1ac61fac749baf

Observation 93ac4450-3308-4b38-aa0b-72c8c9f44030 · outbound

This paper cites Klees, A.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Klees, A

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.983232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.983232Z digest=sha256:26915979c86d87a36f25ce75036cc6e77769fde8715179ac188919f25d64294a

Observation 9bfbc5c4-a4c2-4c2d-9143-242c30171f47 · outbound

This paper cites Kvasov, M.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Kvasov, M

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.985951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.985951Z digest=sha256:41a42e90df4e25fe5bf641f232a95bc9f87199756d10cc2a6144fd6f9c03599c

Observation 50fd35dc-31f6-4a05-9c18-d287b9e89273 · outbound

This paper cites Landen, K.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Landen, K

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.988381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.988381Z digest=sha256:d0e5fd2089edf3f811668ab75ab4f466ddc54ec4ae0fbad746c8c139e246d533

Observation 63909332-5fd1-4e86-a25f-0ec99801f06d · outbound

This paper cites Landis and G.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Landis and G

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.990821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.990821Z digest=sha256:90dbce9a254438a41050ba23cc57924254f78b79675ffc1d36ec8531076525bf

Observation 0d59880a-4cd0-4121-a5f5-c3c7a7aa86e2 · outbound

This paper cites Le Tolguenec, E.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Le Tolguenec, E

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.993430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.993430Z digest=sha256:bcda47294e848a5cc6f3030d814babce9193c086266250dd04b5de17b0dc3d5f

Observation 13b8c343-7b53-4224-b476-661a6160b989 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.995739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.995739Z digest=sha256:a1105bd3233d438b9e9b8789e920a863a3de506ba03226e59a11d3acbb98c3ab

Observation 163bdb5b-5330-4c5c-8f4b-84f03e772680 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:09.998821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:09.998821Z digest=sha256:dcc0700c60bbbee24b9a30e865cff23bde785a2b2dc93e888ce87ecf1b49c580

Observation 3f8c22e3-ed8b-4157-9277-0364c1aa7e09 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.001377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.001377Z digest=sha256:98dabf8ece5db772c07cfe54588bc3f969fbfb6fb6bc304a90cc07262b4f91eb

Observation b7e036e0-e1f2-477d-92bc-78f48574ff2d · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.004278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.004278Z digest=sha256:4361306286b4c4c2ef9be43eeb779658916cb9a7d636aa52bb0ca62707e04be0

Observation fd98dc55-f369-4fc3-b0b8-821e8affd97c · outbound

This paper cites Lopez-Martin, B.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Lopez-Martin, B

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.006789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.006789Z digest=sha256:7fe466220b2f16ddb050891dbe3db3f96af7ae90f542688ef691dc1a2b816dcd

Observation 533f6006-bc2d-4fa5-be57-4040536b5969 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.009390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.009390Z digest=sha256:239c45218107af965a5b0947a304e130f2fdc13c67679c3d922ee050582d3527

Observation 80ea6da6-5e9f-4549-ad8c-3a97544d8071 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.011827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.011827Z digest=sha256:024b7ca5cf0e1353c625fc45d13345f1fb8e70a5fddb39c1f8aa2e71c2ee789a

Observation 78e4e4b4-d6aa-463d-8111-932f1c90d82a · outbound

This paper cites Maeda and M.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Maeda and M

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.014268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.014268Z digest=sha256:64a8ea61b996fb785fbbfc37c9e008c3c1bc3cc743f78520916458eeeeca1187

Observation 4e85cdea-4499-4c2c-9106-42eea1647fe5 · outbound

This paper cites Guidelines for Applying RL and MARL in Cybersecurity Applications.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Guidelines for Applying RL and MARL in Cybersecurity Applications

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.016619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.016619Z digest=sha256:650cb7cf31ff54866bb7ffafc22ab45970ca5dd22c6a129304871533ce28cccc

Observation 604565b2-5d9c-48eb-8bef-d525f02b2ea5 · outbound

This paper cites McFadden, M.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity McFadden, M

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.019651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.019651Z digest=sha256:f6db88cb3d6eabca9ba193fd7a5cdcbeed319e6bf29e3e37257716a939e1b209

Observation cb99c67d-c6e4-4655-a385-fdb9db72d5a3 · outbound

This paper cites McFadden, M.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity McFadden, M

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.022298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.022298Z digest=sha256:dd9582551042da8a545b14fa2f9c1ea3aead40e919e7e5f73547c1a97c7d3cad

Observation 64926738-6233-48f9-a7bb-02fc8c37a9d9 · outbound

This paper cites McFadden, Z.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity McFadden, Z

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.025213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.025213Z digest=sha256:1b744d49f7caced73dfe1dd373c9a660e87113711506e6bab6f60981e258e5d2

Observation 9543db87-378c-4914-901f-1be6c661b2a4 · outbound

This paper cites McFadden, M.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity McFadden, M

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.027896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.027896Z digest=sha256:5a601f3802ba93999bb1b3000a4cb3f94b02ff800cbbbceb7cfaf515a6fc3b8c

Observation bc519fd2-d6b0-4a02-b7b3-f8b71ad9fe43 · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.030583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.030583Z digest=sha256:62493b5298261372a44caddc1f149b37fdb86948ac3ba0dee5c886b24d7ae615

Observation d0451681-c554-45ae-8676-8945d4f013bb · outbound

This paper cites Mirhoseini, A.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Mirhoseini, A

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.033069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.033069Z digest=sha256:d890b0f4f0717a13e49dff4721c74f86dc5c06155e5e373dafee492885c69d8e

Observation 26611a8e-a274-44e4-acab-daf7203933fd · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.035822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.035822Z digest=sha256:1e54e495f7d86c72cc6fcd01dc42ddf7e0ff7fae5a1d8e04d301812e83f76210

Observation 8c574b0a-088d-4cd2-82db-047394280cf8 · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Playing Atari with Deep Reinforcement Learning

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.039266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.039266Z digest=sha256:0a9b37097b445d1ea09639d146e13671a197442b316c23b32ba5a87c9c313f73

Observation 46404a93-0be5-443a-a942-39be683e530b · outbound

This paper cites Mohamed and R.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Mohamed and R

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.042701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.042701Z digest=sha256:44e2377f0f2bf37b1352f47b67ff84978ac5125bd3ee37590cb0d19d28d4532b

Observation 01633a73-c818-4141-8021-836978c8ec75 · outbound

This paper cites Moore, A.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Moore, A

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.045495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.045495Z digest=sha256:8f47fb034c9f2f94ab401c80efddbabfac93e17db12e2747c92a680113432e18

Observation ec6e792f-3697-4fa0-9cf0-8b7ee909d90b · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.048370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.048370Z digest=sha256:99bbd202ed558b5d55e019a302a20236ea76aee2347d1f8d071bd7e75037b2b4

Observation 9f009ee0-b67a-46ac-9c19-f1d03c13b64e · outbound

This paper cites an unresolved cited work.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Unresolved cited work

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.051172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.051172Z digest=sha256:5201439c3f5c0bdfbaf0accccc43084f1e0e43751b291f180510724a48ee8563

Observation 28ba3372-47f8-4388-90f8-72602b90d690 · outbound

This paper cites Nikishin, M.

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity Nikishin, M

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-03T03:15:10.053731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:15:10.053731Z digest=sha256:89b347fc79c497a2456191fe75d01b30f239845bdd8303b706ce1b01987c838f

Pith citing papers

Observation da4a0ec6-b9eb-4379-b446-073292b874ea · inbound

Building Better Environments for Autonomous Cyber Defence cites this paper.

Building Better Environments for Autonomous Cyber Defence SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-06-23T04:13:44.442522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T16:46:41.137112Z digest=sha256:78711a20ee21f20cc7782b9aa79d31cced9eaf1cf8c1a10cac8dcc30aeda2e69