Pith. sign in

Paper Citation Record · LEDGER

On the Surprising Efficacy of LLMs for Penetration-Testing

As of 8 August 2026, this Paper Citation Record lists 100 of 122 outbound references and 2 inbound Pith citation observations for arXiv:2507.00829.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.00829 v1

Coverage vector

measured 100 of 122 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:10:06.693153Z

measured 102 of 102 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-12T09:10:11.585499Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T23:45:52.892312Z

Reference resolution

100 of 122 outbound references displayed

  • verified exact5
  • verified fuzzy18
  • unresolved73
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f3225ebe-ba01-415d-a6b1-f2e49bf161f7 · outbound

This paper cites Control-flow integrity principles, implementations, and applications.

On the Surprising Efficacy of LLMs for Penetration-Testing Control-flow integrity principles, implementations, and applications

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.314597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.314597Z digest=sha256:1c8afff1bfdb73a9af19352d3909b294e65add0027f49f4993735019bce264b8

Observation 20e5481e-c1dd-4cb4-8266-d5ec09ef84c6 · outbound

This paper cites O1 is less powerful than o1-preview due to the less time it spends on thinking (compute time).

On the Surprising Efficacy of LLMs for Penetration-Testing O1 is less powerful than o1-preview due to the less time it spends on thinking (compute time)

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.320803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.320803Z digest=sha256:c049cb0bef368c654b54991313ce3c78270314a812cc21503ee03d26d298bc82

Observation c52a4c27-1a57-42bc-82d8-84bf99629e55 · outbound

This paper cites Performance of o1 vs.

On the Surprising Efficacy of LLMs for Penetration-Testing Performance of o1 vs

Reference 3

Resolution
verified exact
raw_fallback, observed 2026-08-06T21:10:09.053196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:04.326393Z digest=sha256:43e3d26bd5e52a0f91f43f4248d81c0c43c0dc17fc74af9878134b40fe047bd1

Observation 0bc2b123-44fc-41b7-ad5e-c87f867638bd · outbound

This paper cites Introducing the model context protocol.

On the Surprising Efficacy of LLMs for Penetration-Testing Introducing the model context protocol

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.330529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.330529Z digest=sha256:b27b1e05e9305bcf481b47c24e8c2b6be3634fc916e899667146b63d1ef11c93

Observation 4b2a3b2a-8ab9-42ae-b9a1-98df2d552953 · outbound

This paper cites Detecting and countering malicious uses of claude: March.

On the Surprising Efficacy of LLMs for Penetration-Testing Detecting and countering malicious uses of claude: March

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.336114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.336114Z digest=sha256:8beaf8d37b85f4ecbc1da7bf4201ab212479d82588a97025e0fa7c74c3c58112

Observation 9f02aa17-409e-458c-ba92-f1e86240ef77 · outbound

This paper cites Non-Determinism of "Deterministic" LLM Settings.

On the Surprising Efficacy of LLMs for Penetration-Testing Non-Determinism of "Deterministic" LLM Settings

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.347147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.347147Z digest=sha256:f8db8ca9474172d741d9f504444a500241a2d29a8407ad1713a5e771a7308b0a

Observation 79c07c35-0da4-4725-a9c9-b13bbc0ce8b0 · outbound

This paper cites Llms for in- telligent software testing: A comparative study.

On the Surprising Efficacy of LLMs for Penetration-Testing Llms for in- telligent software testing: A comparative study

Reference 7

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T21:10:08.977364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:04.353351Z digest=sha256:d6ee3c2017ee1c2d9697f90b6e3e3708b8751ac813021d40c2d9e42923463636

Observation 3a3130f6-fb7e-4477-8583-3d0b2e9b30f0 · outbound

This paper cites Ai angst.

On the Surprising Efficacy of LLMs for Penetration-Testing Ai angst

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.359510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.359510Z digest=sha256:015a1ec4902f98f2cfb84260c2ce2f93c0637cb1dacecb699de0be34801ba9f9

Observation db27d6d7-8b52-4b3f-9cef-cd3d16cb0fb3 · outbound

This paper cites Generative ai at work.

On the Surprising Efficacy of LLMs for Penetration-Testing Generative ai at work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.363666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.363666Z digest=sha256:8c856c7f504029878e185224122e557f897ca4b225a3d7ca466ddf5cc074c817

Observation a8c7c9e9-7929-42f3-aaa1-8ff06219a2b6 · outbound

This paper cites On large language models in national security applications.

On the Surprising Efficacy of LLMs for Penetration-Testing On large language models in national security applications

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.368448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.368448Z digest=sha256:061679a5448a90fd9106dffc936607297311d4653e72d01f7fe8f38fdd6df84c

Observation 8d5ef76a-0a9c-4c1f-b2e5-e2cc2aeb42f7 · outbound

This paper cites Leveling up fuzzing: Finding more vulnerabilities with ai.

On the Surprising Efficacy of LLMs for Penetration-Testing Leveling up fuzzing: Finding more vulnerabilities with ai

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.371574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.371574Z digest=sha256:48970cd160f7a06c068569ea4bb447b9af986f162b4b4b4c566b899bee35f94e

Observation 3b15cd00-5170-4d33-8115-3ae182f709bb · outbound

This paper cites LlamaFirewall: An open source guardrail system for building secure AI agents.

On the Surprising Efficacy of LLMs for Penetration-Testing LlamaFirewall: An open source guardrail system for building secure AI agents

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.377411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.377411Z digest=sha256:a25d66d56b97930c445e9d92aec154b6b0b69086400bc70bc983488f3284d878

Observation 4eb133aa-2187-46b2-bd27-e30ccb9f55ff · outbound

This paper cites Extracting memorized pieces of (copyrighted) books from open-weight language models.

On the Surprising Efficacy of LLMs for Penetration-Testing Extracting memorized pieces of (copyrighted) books from open-weight language models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.382357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.382357Z digest=sha256:008c4df364362c94a59b77fdca2008696c13f3a02f3f61ccf5104061462f92aa

Observation 5ce74feb-3325-4a0e-9b06-7afa0861bf4a · outbound

This paper cites Bias and unfairness in information retrieval systems: New challenges in the llm era.

On the Surprising Efficacy of LLMs for Penetration-Testing Bias and unfairness in information retrieval systems: New challenges in the llm era

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.387221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.387221Z digest=sha256:ed13b47d01f8dcacfe5048e809ec0a8da0f7472ede6f949c0304d0bc9533b763

Observation 13f48b7b-a5e2-4441-b83d-18379d1ebc07 · outbound

This paper cites Defeating Prompt Injections by Design.

On the Surprising Efficacy of LLMs for Penetration-Testing Defeating Prompt Injections by Design

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.391584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.391584Z digest=sha256:8a5140ac50cd9704f79737e5aae07f5cf42bdd82227c9bb97ff293fe16ccf988

Observation 30d47ca5-84a4-4b08-bd8a-ac56e8aed790 · outbound

This paper cites {PentestGPT}: Evaluating and harnessing large language models for automated penetration testing.

On the Surprising Efficacy of LLMs for Penetration-Testing {PentestGPT}: Evaluating and harnessing large language models for automated penetration testing

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.396067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.396067Z digest=sha256:226f57023e61b10c0c7f315bde3dafb9335f8a8371ef440ec03103a0099f4539

Observation 9b660b0f-7c10-41c1-a9ad-a195665dba64 · outbound

This paper cites Schumpeter’s creative destruction: A review of the evidence.

On the Surprising Efficacy of LLMs for Penetration-Testing Schumpeter’s creative destruction: A review of the evidence

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.399421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.399421Z digest=sha256:513755ca210daf4296a1d6d77ca4c20e337244b91be0b517e8fb0e7f2c71de48

Observation 2f9e4508-1887-4c13-9e0f-e6ccd1f843c3 · outbound

This paper cites The explainability challenge of generative ai and llms.

On the Surprising Efficacy of LLMs for Penetration-Testing The explainability challenge of generative ai and llms

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.405054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.405054Z digest=sha256:61c0b8689753fc969d8d3d9fae9e8888b6b222524b130a4e88f392044cc0e925

Observation adaae1a6-4744-4256-ab3f-615205f224de · outbound

This paper cites The potential for jurisdictional challenges to ai or llm training datasets.

On the Surprising Efficacy of LLMs for Penetration-Testing The potential for jurisdictional challenges to ai or llm training datasets

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.415104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.415104Z digest=sha256:39d9a33e827d8135a397fae91996fd29a3aa7067c1e8b298a11fee997ebfbf19

Observation 7ad48560-dfdf-49fc-9dd1-92c5c61be25a · outbound

This paper cites Large language models in information security research: A january 2024 survey.

On the Surprising Efficacy of LLMs for Penetration-Testing Large language models in information security research: A january 2024 survey

Reference 20

Resolution
verified exact
raw_fallback, observed 2026-08-06T21:10:08.866942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:04.419185Z digest=sha256:1fecb564dda609e9c832a7cb3bac4cf89784e5a2ef0d4510240e25c5ba2c8444

Observation 667e80d5-d5e8-4bc9-9437-a1d6061ceb1a · outbound

This paper cites Google’s approach for secure ai agents.

On the Surprising Efficacy of LLMs for Penetration-Testing Google’s approach for secure ai agents

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.423535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.423535Z digest=sha256:6112a747159fd605d3c9c698d97b6475d016de45684ac01b619ffd3f0b0b7ee9

Observation 642e0ffe-f561-4d7d-9dfe-e25d84ae1e4a · outbound

This paper cites Gpts are gpts: Labor market impact potential of llms.

On the Surprising Efficacy of LLMs for Penetration-Testing Gpts are gpts: Labor market impact potential of llms

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.427808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.427808Z digest=sha256:9390d6c6fc6482b2b8dd8e12d5dcfaee1ec4bf8c63184c3f00dbd6dda7b8aeb1

Observation 351a2e6a-ccff-421d-9007-946b0c54f360 · outbound

This paper cites LLM Agents can Autonomously Exploit One-day Vulnerabilities.

On the Surprising Efficacy of LLMs for Penetration-Testing LLM Agents can Autonomously Exploit One-day Vulnerabilities

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.432005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.432005Z digest=sha256:6c5810788e0007dd02b6bbfd929077a1322ec4de913a94435a31792d164fcf93

Observation 7f7d507f-19ff-43c2-8dca-6943b95dbe52 · outbound

This paper cites Llm agents can autonomously hack websites, 2024.

On the Surprising Efficacy of LLMs for Penetration-Testing Llm agents can autonomously hack websites, 2024

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.436415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.436415Z digest=sha256:5911363b08eacc4fb211ad9a7e20dc204e1acf20d04ecc808c32943554e9bfb5

Observation fac21925-22f5-4eca-b4be-f76d5119c601 · outbound

This paper cites Teams of LLM Agents can Exploit Zero-Day Vulnerabilities.

On the Surprising Efficacy of LLMs for Penetration-Testing Teams of LLM Agents can Exploit Zero-Day Vulnerabilities

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.440364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.440364Z digest=sha256:f69f62b4d62e2a804b2efd517b0650c3624ea3bef750ecccdcebe66ed1135d25

Observation a63c52cf-7bc2-4218-b509-de8ffb86a80a · outbound

This paper cites Wormgpt: a large language model chatbot for criminals.

On the Surprising Efficacy of LLMs for Penetration-Testing Wormgpt: a large language model chatbot for criminals

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.445121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.445121Z digest=sha256:6ceb258f80c0fb4f0afd436deec8b519c661c96391789985b8450d896e2ab9f9

Observation 3cd811ee-0138-4314-8cbd-db11bcedce08 · outbound

This paper cites Who’s asking? user personas and the mechanics of latent misalignment.

On the Surprising Efficacy of LLMs for Penetration-Testing Who’s asking? user personas and the mechanics of latent misalignment

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.451499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.451499Z digest=sha256:edf335ea1785374c334db8a65e38464d13e50d20ed56da7ab4f42d0fb1259d6c

Observation 4d7d9d3d-5ce4-418b-a742-22dd271fc41d · outbound

This paper cites AutoPenBench: Benchmarking Generative Agents for Penetration Testing.

On the Surprising Efficacy of LLMs for Penetration-Testing AutoPenBench: Benchmarking Generative Agents for Penetration Testing

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.455902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.455902Z digest=sha256:f442391c0270eb626bc138dff2c3784fc0d7c93795481bc884be2f048db3880f

Observation 352318db-05e8-4bc1-a136-2a0e310be3ba · outbound

This paper cites Project naptime: Evaluating offensive security capabilities of large language models.

On the Surprising Efficacy of LLMs for Penetration-Testing Project naptime: Evaluating offensive security capabilities of large language models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.460359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.460359Z digest=sha256:d56cea71a51374b60f234831295eea20aa48cdb09008c31dd36a12f5a0d1295b

Observation ac7dec3a-805f-4039-b198-e63a9802c4f9 · outbound

This paper cites Adversarial misuse of generative ai.

On the Surprising Efficacy of LLMs for Penetration-Testing Adversarial misuse of generative ai

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.464783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.464783Z digest=sha256:e57b51de6734caad51474ba3017083018839d202f437844a4c6771cd042ea164

Observation b1a09501-8220-4633-b105-9be022bec520 · outbound

This paper cites A Survey on LLM-as-a-Judge.

On the Surprising Efficacy of LLMs for Penetration-Testing A Survey on LLM-as-a-Judge

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.470704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.470704Z digest=sha256:ba29b124121b36105ad735060fe6a89ff5eccb24cea6f848693d9c324a3020ef

Observation 575a2042-aec6-40b3-b019-087c8aebb908 · outbound

This paper cites How we built our multi-agent research system.

On the Surprising Efficacy of LLMs for Penetration-Testing How we built our multi-agent research system

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.476893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.476893Z digest=sha256:42774574480f061bf00522ee8d5d9ebfb018fb5b351b3ec50c4c128ad8b8e591

Observation 82093bb9-bcdb-4380-af6e-0a3aa0f4ac5e · outbound

This paper cites Getting pwn’d by ai: Penetration testing with large language models.

On the Surprising Efficacy of LLMs for Penetration-Testing Getting pwn’d by ai: Penetration testing with large language models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.482642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.482642Z digest=sha256:cbeee2190c95ee72539433d26f7504332b5a2fe491e26ac37d7fe7c4a258eb4d

Observation b64cde0e-e664-414b-892a-6ce2c38a81d6 · outbound

This paper cites Understanding hackers’ work: An empirical study of offensive security practitioners.

On the Surprising Efficacy of LLMs for Penetration-Testing Understanding hackers’ work: An empirical study of offensive security practitioners

Reference 34

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T21:10:08.744071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:04.486963Z digest=sha256:d17ff1900ec4cdfa2eb745376b1f5b3bdd9b95b5e930c8cf94319cfb92f0ed0e

Observation bed81608-05d1-4c89-a81a-27894b52f624 · outbound

This paper cites Benchmarking Practices in LLM-driven Offensive Security: Testbeds, Metrics, and Experiment Design.

On the Surprising Efficacy of LLMs for Penetration-Testing Benchmarking Practices in LLM-driven Offensive Security: Testbeds, Metrics, and Experiment Design

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.491227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.491227Z digest=sha256:0f31e65cfe3a9f1caf8955a3617afe09118227f6fb539cf232fb6ce04130d031

Observation 5409c3b5-13cd-4e16-a2a8-3d490009a0f6 · outbound

This paper cites Recognition Without Mitigation: Ethical Frameworks in Autonomous Offensive-LLM Agent Research.

On the Surprising Efficacy of LLMs for Penetration-Testing Recognition Without Mitigation: Ethical Frameworks in Autonomous Offensive-LLM Agent Research

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:10:08.637385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:04.497285Z digest=sha256:54560215cec0902cb0481d61559dc0f6dffcd8fdf41315508e1d9522c4c90f42

Observation 7232b88f-5019-4a15-b0b6-63694ce69ce9 · outbound

This paper cites Can LLMs Hack Enterprise Networks? Autonomous Assumed Breach Penetration-Testing Active Directory Networks.

On the Surprising Efficacy of LLMs for Penetration-Testing Can LLMs Hack Enterprise Networks? Autonomous Assumed Breach Penetration-Testing Active Directory Networks

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:10:08.622015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:04.501811Z digest=sha256:dbf409f97cea126e6216884740abbe43c34d11b46992f946a6b652e4fb3f7d11

Observation 7ca824d5-1b66-4f31-a69c-4bc534fafbc1 · outbound

This paper cites Llms as hackers: Autonomous linux privilege escalation attacks.

On the Surprising Efficacy of LLMs for Penetration-Testing Llms as hackers: Autonomous linux privilege escalation attacks

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.508012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.508012Z digest=sha256:ee7abf7cc0966a0848c274b1d30f9749d43e454cc0cc833f38d9e9fe271bf73e

Observation 824535aa-baff-4a58-9a00-e3275372f47e · outbound

This paper cites A Comprehensive Overview of Large Language Models (LLMs) for Cyber Defences: Opportunities and Directions.

On the Surprising Efficacy of LLMs for Penetration-Testing A Comprehensive Overview of Large Language Models (LLMs) for Cyber Defences: Opportunities and Directions

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.514085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.514085Z digest=sha256:42ad38e185f400fc9ec26e5d2b2d706953e1619a0034755a6526f4da490239ae

Observation d885cf06-eb21-412e-b80c-2f725fe7200c · outbound

This paper cites Does prompt formatting have any impact on llm performance?,.

On the Surprising Efficacy of LLMs for Penetration-Testing Does prompt formatting have any impact on llm performance?,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.520107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.520107Z digest=sha256:3f9600de9628fc395a77ff2b500129d910a71d4e2b1f2f1ca4f0e4b71b696bb3

Observation a756b40d-d815-44d0-aa17-4e521e5cb53d · outbound

This paper cites How i used o3 to find cve-2025-37899, a remote zeroday vulnerability in the linux kernel’s smb implementation.

On the Surprising Efficacy of LLMs for Penetration-Testing How i used o3 to find cve-2025-37899, a remote zeroday vulnerability in the linux kernel’s smb implementation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.528878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.528878Z digest=sha256:a51fd485cc5ca8e245c76d061ad02020b997771832ef62bbbe5d2d781f9c7c9d

Observation 7d980772-7ee5-44f6-83a5-9fd4bb7610d1 · outbound

This paper cites Ai and the increase of productivity and labor inequality in latin america: Potential impact of large language models on latin american workforce.

On the Surprising Efficacy of LLMs for Penetration-Testing Ai and the increase of productivity and labor inequality in latin america: Potential impact of large language models on latin american workforce

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.533037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.533037Z digest=sha256:20026e60cab614de777e00f2c9541bb07c09ad4a0455acf3d029429a95e96d89

Observation 9a2fb4a9-d0b9-401d-bd66-52ba8e377ad2 · outbound

This paper cites Does Prompt Formatting Have Any Impact on LLM Performance?.

On the Surprising Efficacy of LLMs for Penetration-Testing Does Prompt Formatting Have Any Impact on LLM Performance?

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.524128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.524128Z digest=sha256:93e78e223bfa7cf27487bf285c8ede1e88712ea28d7cd27628d9be58e490ece2

Observation 934cfb72-d162-4124-9093-7f5d8f904889 · outbound

This paper cites Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations.

On the Surprising Efficacy of LLMs for Penetration-Testing Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.541993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.541993Z digest=sha256:12a2c1b9746d3060a3831e5940dce25bdbac0e3f761fd872a9e6bbce0fa9db95

Observation 1a4bd8c7-32bd-4450-b98b-9d9485beea19 · outbound

This paper cites Ethics and algorithms.

On the Surprising Efficacy of LLMs for Penetration-Testing Ethics and algorithms

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.510114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.510114Z digest=sha256:016968eea127d502a56b2abe26967c5ae5e6d7a4c50357b2eccda1620b7128ef

Observation d83347f0-7bba-4c01-9f4f-d1c037e52658 · outbound

This paper cites Uncertainty of thoughts: Uncertainty-aware planning enhances information seeking in llms.

On the Surprising Efficacy of LLMs for Penetration-Testing Uncertainty of thoughts: Uncertainty-aware planning enhances information seeking in llms

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.537340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.537340Z digest=sha256:85cf04d75f77863161f31781dff60cfa2e21134a5a0dfe498d4fae5c0254c607

Observation ef64c1d6-acf3-4aa8-9cef-9c39043fc37d · outbound

This paper cites How hungry is ai? benchmarking energy, water, and carbon footprint of llm inference, 2025.

On the Surprising Efficacy of LLMs for Penetration-Testing How hungry is ai? benchmarking energy, water, and carbon footprint of llm inference, 2025

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.517346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.517346Z digest=sha256:0004a49a97771262159a3b703923c90c5a12865dc1a8012c66e51733c970ceb2

Observation 97326b02-7ed5-4e50-893f-3bacf69d6168 · outbound

This paper cites From LLMs to LLM-based Agents for Software Engineering: A Survey of Current, Challenges and Future.

On the Surprising Efficacy of LLMs for Penetration-Testing From LLMs to LLM-based Agents for Software Engineering: A Survey of Current, Challenges and Future

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.520524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.520524Z digest=sha256:771a5c968b3f8fa72afd16a596c24667b42d950d3e713014229950ebfca65260

Observation 2c901dc0-205b-44f7-bba4-f82be3990337 · outbound

This paper cites 2024 isc2 cybersecurity workforce study.

On the Surprising Efficacy of LLMs for Penetration-Testing 2024 isc2 cybersecurity workforce study

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.513269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.513269Z digest=sha256:9a867b63322fdbf59e63fa1570a082c28d5d1db7a52bb4f0e1779a751c982e4c

Observation 76e18fe3-4ef6-465a-bae8-ea273ed2ea40 · outbound

This paper cites Advances in llms with focus on reasoning, adaptability, efficiency and ethics.

On the Surprising Efficacy of LLMs for Penetration-Testing Advances in llms with focus on reasoning, adaptability, efficiency and ethics

Reference 50

Resolution
verified exact
raw_fallback, observed 2026-08-06T21:10:08.412919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.527544Z digest=sha256:1cce8a2a5afe391d9cd09de6bee15a9807d17adfdac21b57477307f221871793

Observation 0e6b257d-38d6-4045-8e09-46aa422153ac · outbound

This paper cites A survey of llm-driven ai agent communication: Protocols, security risks, and defense countermeasures, 2025.

On the Surprising Efficacy of LLMs for Penetration-Testing A survey of llm-driven ai agent communication: Protocols, security risks, and defense countermeasures, 2025

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.530477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.530477Z digest=sha256:81bd8f5379cb80f4204f39c1431b059acc1c70f66f7a5f2edb834a8dc39a9ce6

Observation 8ef0c43c-fd67-4030-b82a-de8f52a34551 · outbound

This paper cites Generation, Detection, and Evaluation of Role-play based Jailbreak attacks in Large Language Models.

On the Surprising Efficacy of LLMs for Penetration-Testing Generation, Detection, and Evaluation of Role-play based Jailbreak attacks in Large Language Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.523847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.523847Z digest=sha256:0609452ad935f9269da743e1eb3ab4eb27dfddbb4dcdf05820b017d17341bbbf

Observation d76cb56b-025a-4ac3-8a7b-1f5937f3182a · outbound

This paper cites Your Brain on ChatGPT: Accumulation of Cognitive Debt when Using an AI Assistant for Essay Writing Task.

On the Surprising Efficacy of LLMs for Penetration-Testing Your Brain on ChatGPT: Accumulation of Cognitive Debt when Using an AI Assistant for Essay Writing Task

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.537095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.537095Z digest=sha256:e64ea1503564f40a442feec65ae361eb97d25687fdb42f516673b2801a73482e

Observation 38a60a22-f8bf-4d99-bd1c-d079e507966f · outbound

This paper cites Revolutionizing talent: the path in 21st century workforce transformation.

On the Surprising Efficacy of LLMs for Penetration-Testing Revolutionizing talent: the path in 21st century workforce transformation

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.540785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.540785Z digest=sha256:fd1e90551aa95afed3cde511b82f8e7b6586a3f61b8882ab29af5e9f922aa0a6

Observation 5f620134-4133-4460-ac39-cd37a3c7b458 · outbound

This paper cites VulnBot: Autonomous Penetration Testing for A Multi-Agent Collaborative Framework.

On the Surprising Efficacy of LLMs for Penetration-Testing VulnBot: Autonomous Penetration Testing for A Multi-Agent Collaborative Framework

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.533634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.533634Z digest=sha256:39c6cade241f8e1d428cd850657d0612e2b8a911d818b27a409bee4d0d3276f1

Observation 9770da71-26cb-4f25-824d-f95688174039 · outbound

This paper cites Shade-arena: Evaluating sabotage and monitoring in llm agents.

On the Surprising Efficacy of LLMs for Penetration-Testing Shade-arena: Evaluating sabotage and monitoring in llm agents

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.547704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.547704Z digest=sha256:b40790629007e6b537683edcc01dec13e1592ffbf4c03abc90619ad9250d6d9f

Observation a943492c-c5fc-4534-b50c-a02ab8c4a18e · outbound

This paper cites LLMs Get Lost In Multi-Turn Conversation.

On the Surprising Efficacy of LLMs for Penetration-Testing LLMs Get Lost In Multi-Turn Conversation

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.551304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.551304Z digest=sha256:6200c4b64694b0be9d12ce631b0fbc403c3235ae79eeccf6ae456603a9731dbd

Observation 753acd2e-50e0-4d0b-9b36-9024c953804f · outbound

This paper cites an unresolved cited work.

On the Surprising Efficacy of LLMs for Penetration-Testing Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.544515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.544515Z digest=sha256:556bb7f658d6da81bb005f7df0fab6ad741c7ead95cf17f5c0101687f35fc58b

Observation 941fbd55-c03d-4a9b-a1b9-8e0f023b5ed1 · outbound

This paper cites Can Long-Context Language Models Subsume Retrieval, RAG, SQL, and More?.

On the Surprising Efficacy of LLMs for Penetration-Testing Can Long-Context Language Models Subsume Retrieval, RAG, SQL, and More?

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.558326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.558326Z digest=sha256:722bbd1ef3e87285b81579ceb4966a5b2a408e6e6c139e3d5dca9bc606626771

Observation 73e34ba9-3a85-4509-baa3-3fe265b8d32f · outbound

This paper cites I think i’m done thinking about genai for now.

On the Surprising Efficacy of LLMs for Penetration-Testing I think i’m done thinking about genai for now

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.562357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.562357Z digest=sha256:763b41d8516c9cde24bb17cd093bf4f4cc8f62b58a4b6d31f52b18f8b352616c

Observation d6943c7f-25d7-4208-8584-8dea67852ca9 · outbound

This paper cites Operating multi-client influ- ence networks across platforms.

On the Surprising Efficacy of LLMs for Penetration-Testing Operating multi-client influ- ence networks across platforms

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.555475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.555475Z digest=sha256:68237b03af5fcb0194caf580c11853e04730e296d5423990031fa1a1a50eb191

Observation 9f76e63d-8118-4da8-88da-7a5654e0c337 · outbound

This paper cites Malla: Demystifying real-world large language model integrated malicious services.

On the Surprising Efficacy of LLMs for Penetration-Testing Malla: Demystifying real-world large language model integrated malicious services

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.571200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.571200Z digest=sha256:890a4c537fa822076717bb872fe602dec0101527440714aea446cdadc13ce21a

Observation cdbc0a5f-136d-49ff-833f-e1e91ac638f3 · outbound

This paper cites Ai-powered fuzzing: Breaking the bug hunting barrier.

On the Surprising Efficacy of LLMs for Penetration-Testing Ai-powered fuzzing: Breaking the bug hunting barrier

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.574294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.574294Z digest=sha256:83feb46581e672968ba80e8e684ad0d547daaf5f7b35840f1219d9d0216a31f2

Observation 6cfbdbc0-3553-41da-9d06-9a8b602012f7 · outbound

This paper cites an unresolved cited work.

On the Surprising Efficacy of LLMs for Penetration-Testing Unresolved cited work

Reference 64

Resolution
parse uncertain
no resolver link, observed 2026-08-06T21:10:06.565216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.565216Z digest=sha256:3a658082e3b29e8178bca9a4939aa668fdbf927b7e3bc7bcf36668b1e0e369d7

Observation f6d90882-42a1-4a02-9b4d-b17aaff211e0 · outbound

This paper cites When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs.

On the Surprising Efficacy of LLMs for Penetration-Testing When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.568134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.568134Z digest=sha256:a4a9ff86fef8a9c339c4e50ae21ebe57e7ef08615c82b38eaf36400a0d4f4840

Observation fd554f45-1f6a-4fd8-aacf-30328e9f2855 · outbound

This paper cites Troy, Stuart J.

On the Surprising Efficacy of LLMs for Penetration-Testing Troy, Stuart J

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.582782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.582782Z digest=sha256:3e555eef34bc60035ce6e38f4249864d8cf813b4e25a6d873cd2543e8a466cb9

Observation 5a67be7a-4fcd-46e8-b0bd-2665d1676dc3 · outbound

This paper cites Llm dataset inference: Did you train on my dataset? Advances in Neural Information Pro- cessing Systems, 37:124069–124092, 2024.

On the Surprising Efficacy of LLMs for Penetration-Testing Llm dataset inference: Did you train on my dataset? Advances in Neural Information Pro- cessing Systems, 37:124069–124092, 2024

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.585900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.585900Z digest=sha256:33b9507f958f74f3156c6e3e8cd6bc915039d770c03106f0a976ac9e4e07a468

Observation 2064aa91-2129-48b3-b60f-39c01393918b · outbound

This paper cites LLM Cyber Evaluations Don't Capture Real-World Risk.

On the Surprising Efficacy of LLMs for Penetration-Testing LLM Cyber Evaluations Don't Capture Real-World Risk

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.577088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.577088Z digest=sha256:ec3ca82152bdc413ce3254056511d2d0153bfd7b4e742b03be89096e48ae1396

Observation 527d6ca1-fbff-49aa-a0de-c5a81fee8ce2 · outbound

This paper cites The dual-use security dilemma and the social construction of insecurity.

On the Surprising Efficacy of LLMs for Penetration-Testing The dual-use security dilemma and the social construction of insecurity

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.579914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.579914Z digest=sha256:47d01a87bdf3b8e17424a525e047aae4cd4b9542fcec757aeaf7afd3fd00c153

Observation 7a73bf9f-fdc7-4f23-8b9d-8ec1da972092 · outbound

This paper cites Llama prompt guard 2.

On the Surprising Efficacy of LLMs for Penetration-Testing Llama prompt guard 2

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:10:09.329925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.595204Z digest=sha256:455e0b7c6b4fd1c8cba1d515b01d559dac10fd64085f7a5e3d08b9f5022ba559

Observation f0a075df-f788-413b-b3d5-1ec21f4e1f96 · outbound

This paper cites Large Language Models as General Pattern Machines.

On the Surprising Efficacy of LLMs for Penetration-Testing Large Language Models as General Pattern Machines

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.598563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.598563Z digest=sha256:29ae7c3d6dd97438f59378be7d61acf630fb7238857cf1169c01bf3edff6d7d6

Observation 8cdebbe3-21a2-4fd5-8731-bdc14bbde1f4 · outbound

This paper cites Why using chatgpt is not bad for the environment - a cheat sheet.

On the Surprising Efficacy of LLMs for Penetration-Testing Why using chatgpt is not bad for the environment - a cheat sheet

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:10:09.339592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.589014Z digest=sha256:8497f1b2be2998f211042d53e9de0b23d81ec09cf6412b06b8d3c36cdb3038a0

Observation 662b7cca-2024-463e-bec2-9520076fd46b · outbound

This paper cites Mavikumbure, Victor Cobilean, Chathurika S.

On the Surprising Efficacy of LLMs for Penetration-Testing Mavikumbure, Victor Cobilean, Chathurika S

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.592249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.592249Z digest=sha256:379811db194912a25bda24c78a6b139d5fad4edcc9a45205a5d5caa198a87ad8

Observation f02c66bd-d8c0-4d03-ab22-6e32659f1f97 · outbound

This paper cites Large Language Models in Cybersecurity: State-of-the-Art.

On the Surprising Efficacy of LLMs for Penetration-Testing Large Language Models in Cybersecurity: State-of-the-Art

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.607291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.607291Z digest=sha256:bcd35ef47af11e188e7c54015693209dd14a8dbe56bfed3a34c255db40d85d08

Observation ca887944-91c6-436c-ad1f-c28a8bb08463 · outbound

This paper cites Influence and cyber operations: an up- date.

On the Surprising Efficacy of LLMs for Penetration-Testing Influence and cyber operations: an up- date

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:10:09.300190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.611038Z digest=sha256:864155b139d2a25e715cbf0ebd9838d317581a13940328adbd9ccf541334628f

Observation 698337fb-0bdc-415a-8372-a560aa84f6cf · outbound

This paper cites The threat of offensive ai to organizations.

On the Surprising Efficacy of LLMs for Penetration-Testing The threat of offensive ai to organizations

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:10:09.319250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.601293Z digest=sha256:bc39307a6b58bc2999b40d7e1aa87ac4ce7139d802a01897f98cb5eafef09f99

Observation 8c162f27-3027-4da3-b2d9-3a43952d4639 · outbound

This paper cites Global ransomware damage costs predicted to exceed $275 billion by 2031.

On the Surprising Efficacy of LLMs for Penetration-Testing Global ransomware damage costs predicted to exceed $275 billion by 2031

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:10:09.309051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.604515Z digest=sha256:9734660d32ae957fdd56a021add08034d9059bafb6174af737f68a28ae8e992c

Observation 0e0b8246-7a41-489c-a403-b45879be233b · outbound

This paper cites Introducting chatgpt.

On the Surprising Efficacy of LLMs for Penetration-Testing Introducting chatgpt

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:10:09.272572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.619907Z digest=sha256:46734a1a047311a35c5d43f1f042f7842d3fee76eb0500e68db92d14a1d5b350

Observation 0790fb15-a216-4c14-aa4a-318c8af564b9 · outbound

This paper cites Introducing openai o1.

On the Surprising Efficacy of LLMs for Penetration-Testing Introducing openai o1

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:10:09.258320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.625974Z digest=sha256:9b6e32ff982ce754936b3eef819a944075589eac464c0264fe29f9c928875f08

Observation 7c54e393-af43-4a89-82e9-72d3fe751ce9 · outbound

This paper cites Disrupting malicious uses of ai: June 2025.

On the Surprising Efficacy of LLMs for Penetration-Testing Disrupting malicious uses of ai: June 2025

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:10:09.290340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.613817Z digest=sha256:1de34c67e2264d58373fe01ab514c4d9f57a2ed17db0d17fd28c9e09812dbc73

Observation 4f70c9e9-1810-44ec-9fbe-c06c02ded2a1 · outbound

This paper cites Disrupting malicious uses of ai: February 2025.

On the Surprising Efficacy of LLMs for Penetration-Testing Disrupting malicious uses of ai: February 2025

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:10:09.281540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.616696Z digest=sha256:097cd891e0b6ce039bd8ac9d2befbee5b6666e106e8a355bc2ce56cb94621401

Observation 571dd12d-aa7e-403e-bd5d-1c666edd50f7 · outbound

This paper cites Proof or Bluff? Evaluating LLMs on 2025 USA Math Olympiad.

On the Surprising Efficacy of LLMs for Penetration-Testing Proof or Bluff? Evaluating LLMs on 2025 USA Math Olympiad

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.635014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.635014Z digest=sha256:fa62b0f3cfbfb143265f3b168dc7294978dda7a97707e7817fdbadeb53fd5ba1

Observation efb000e8-90cc-4a31-a9cf-7692ecbdea01 · outbound

This paper cites Cipher: Cyberse- curity intelligent penetration-testing helper for ethical researcher.

On the Surprising Efficacy of LLMs for Penetration-Testing Cipher: Cyberse- curity intelligent penetration-testing helper for ethical researcher

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:10:09.232584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.637910Z digest=sha256:a0a1ce00a422ca3471ec7f0f9c998f6ac5493d422e6c2ba43d4fc89f6700d617

Observation 459fd750-7c0a-4363-948f-c74a99703929 · outbound

This paper cites My ai skeptic friends are all nuts.

On the Surprising Efficacy of LLMs for Penetration-Testing My ai skeptic friends are all nuts

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:10:09.223421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.641043Z digest=sha256:e66db92a6d27ae873ca4218cf4779bff99f89e6eea0c9a31e9ce7f16cd58f92e

Observation ee14fad6-56cc-4a15-b823-0e4e40406d1a · outbound

This paper cites Disrupting malicious uses of ai by state-affiliated threat actors.

On the Surprising Efficacy of LLMs for Penetration-Testing Disrupting malicious uses of ai by state-affiliated threat actors

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:10:09.250034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.628911Z digest=sha256:72c22c29160f55daf9339ccfe02e24365bea219145d9f8cc9ae8ab555d618825

Observation 1b63a2a8-0fb5-4b1a-97ad-5252ff786855 · outbound

This paper cites Annual share of organizations affected by ransomware at- tacks worldwide from 2018 to 2023.

On the Surprising Efficacy of LLMs for Penetration-Testing Annual share of organizations affected by ransomware at- tacks worldwide from 2018 to 2023

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:10:09.241453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.631812Z digest=sha256:5de309e30d07ce39eb7112fe9da75737ccd49768352de7aa51eb48089bf55d6e

Observation 2a2862cf-5373-44b2-9608-044e88ade227 · outbound

This paper cites Anderson, Edward W.

On the Surprising Efficacy of LLMs for Penetration-Testing Anderson, Edward W

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.653621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.653621Z digest=sha256:1bf8f6290c8eaf791c7ca6aa254d0d107bac3ec3a1124ecc6a0014281943cc30

Observation a0e3fa75-76a2-41c3-a90e-714a8f0405ac · outbound

This paper cites An Empirical Evaluation of LLMs for Solving Offensive Security Challenges.

On the Surprising Efficacy of LLMs for Penetration-Testing An Empirical Evaluation of LLMs for Solving Offensive Security Challenges

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.656665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.656665Z digest=sha256:ee7c628c9db511121272648a655deea080e1ee7b9dba183571cc8cae5d18b0df

Observation c5fa02d9-f373-44db-93ea-6c0320e3148e · outbound

This paper cites Future of work with ai agents: Auditing automation and augmentation potential across the u.s.

On the Surprising Efficacy of LLMs for Penetration-Testing Future of work with ai agents: Auditing automation and augmentation potential across the u.s

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.660064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.660064Z digest=sha256:e1d1117b9106de2988ac5683f1cc034be590362fa496cb5c54e2ffb15f226aed

Observation 2cddee42-f3aa-4465-9b39-adbb27e6b83c · outbound

This paper cites Complementarity, Augmentation, or Substitutivity? The Impact of Generative Artificial Intelligence on the U.S. Federal Workforce.

On the Surprising Efficacy of LLMs for Penetration-Testing Complementarity, Augmentation, or Substitutivity? The Impact of Generative Artificial Intelligence on the U.S. Federal Workforce

Reference 90

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T21:10:08.097671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.644645Z digest=sha256:a28ed248dcad1d499a5524a10da74957d9813e0783c7724d9fdb7461113817c3

Observation cfeb8941-1cb2-47e7-ad14-f552855b4ab5 · outbound

This paper cites Llm-based design pattern detection,.

On the Surprising Efficacy of LLMs for Penetration-Testing Llm-based design pattern detection,

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:10:09.214551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.647645Z digest=sha256:f5dcf6fa6cc2ac9e7cfb022a0493ae0f24f33a5fd69131c288422f933455e2a6

Observation d446b817-71c4-460d-93c4-1843012c972e · outbound

This paper cites LLM-Based Design Pattern Detection.

On the Surprising Efficacy of LLMs for Penetration-Testing LLM-Based Design Pattern Detection

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.651054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.651054Z digest=sha256:5cac61767f540e864fc7b23183f75c14bd6435ef0e3eb60e7f1377eff7d272f6

Observation 37b9e362-dabc-4594-9cf4-593d5dadabd0 · outbound

This paper cites Announcing the agent2agent protocol (a2a).

On the Surprising Efficacy of LLMs for Penetration-Testing Announcing the agent2agent protocol (a2a)

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:10:09.194465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.672291Z digest=sha256:a6d5b4261ec6ef257e8ea5ae8e99ed266ddc9dc897937cd32f662d581f3eb9de

Observation 474c657b-c803-4db8-a3d6-ed2534922675 · outbound

This paper cites Systematic Biases in LLM Simulations of Debates.

On the Surprising Efficacy of LLMs for Penetration-Testing Systematic Biases in LLM Simulations of Debates

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.675136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.675136Z digest=sha256:e369b32e823150e50a356a86ec159eff57f76fb895de64297dac7bb5b013a1cf

Observation 8c6994cc-c998-441c-937c-3c9232052e30 · outbound

This paper cites From naptime to big sleep: Using large language models to catch vulnerabilities in real-world code.

On the Surprising Efficacy of LLMs for Penetration-Testing From naptime to big sleep: Using large language models to catch vulnerabilities in real-world code

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:10:09.184325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.677586Z digest=sha256:1e1e63b2f13996722635eef9bab9380c7af2b0a23ca24e075fc037440c7fd221

Observation 8ebe2604-da0a-468a-89dd-3d2cc62a406f · outbound

This paper cites The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity.

On the Surprising Efficacy of LLMs for Penetration-Testing The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.662541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.662541Z digest=sha256:e435addf792a2b3f2a0aba26e67f64879666b437bb072e4da07b2c83f21fdcf9

Observation e79bed61-a8e8-4bd1-8b9b-dc0e8799e91b · outbound

This paper cites On the feasibility of using llms to execute multistage network attacks.

On the Surprising Efficacy of LLMs for Penetration-Testing On the feasibility of using llms to execute multistage network attacks

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.666134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.666134Z digest=sha256:7787fa3a064cc075324b893cc03a13b3739169ddfbad937b1c6e6da02331ca61

Observation c04cfdd0-1812-4a2e-8697-2babd1f60407 · outbound

This paper cites Outside the closed world: On using machine learning for network intrusion detection.

On the Surprising Efficacy of LLMs for Penetration-Testing Outside the closed world: On using machine learning for network intrusion detection

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:10:09.204429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.669118Z digest=sha256:73a6f72618be19fe87837d2ddc37f70351d9eb51d74a3e54d880444815bf7848

Observation ec2808ef-23f9-4135-add9-3963cb3bcf77 · outbound

This paper cites Rainbows End: A Novel With One Foot In The Future.

On the Surprising Efficacy of LLMs for Penetration-Testing Rainbows End: A Novel With One Foot In The Future

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:10:09.150860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T21:10:06.690581Z digest=sha256:918fb4957e8b491c899878f7707c60c2ed01d31c222aad146551499193b8eb17

Observation 8ae6a063-c6fc-4d39-824d-bb6363b39682 · outbound

This paper cites "Kelly is a Warm Person, Joseph is a Role Model": Gender Biases in LLM-Generated Reference Letters.

On the Surprising Efficacy of LLMs for Penetration-Testing "Kelly is a Warm Person, Joseph is a Role Model": Gender Biases in LLM-Generated Reference Letters

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:06.693153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:06.693153Z digest=sha256:f2f745a84e9b0d657f48534aeb6915072e7eaec7f512e65f18037301ecca37ce

Pith citing papers

Observation 5ff18c7b-91bd-477f-95d3-153df0404d8b · inbound

Hackers or Hallucinators? A Comprehensive Analysis of LLM-Based Automated Penetration Testing cites this paper.

Hackers or Hallucinators? A Comprehensive Analysis of LLM-Based Automated Penetration Testing On the Surprising Efficacy of LLMs for Penetration-Testing

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:45:52.913350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:52:57.225878Z digest=sha256:66f27b44dd6a0607f97dc1d15a371e153a6352699bafda4e42769a44b40060c7

Observation a471ea7d-910a-43e2-ac91-5381519797e2 · inbound

A Survey of LLM-Driven Penetration Testing: Taxonomy, Co-Evolution, and Open Challenges cites this paper.

A Survey of LLM-Driven Penetration Testing: Taxonomy, Co-Evolution, and Open Challenges On the Surprising Efficacy of LLMs for Penetration-Testing

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-12T09:10:11.585499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:10:11.585499Z digest=sha256:fe87081baa089779eb9f11fb07cf0d8ee066a945f8e4a3340a9049364fff99b6