Pith. sign in

Paper Citation Record · LEDGER

Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation

As of 20 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 2 inbound Pith citation observations for arXiv:2505.01065.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.01065 v1

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:29:54.244745Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T14:59:21.115774Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-11T14:59:21.606635Z

Reference resolution

18 of 18 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ea9ca3b8-dc0a-477a-ae2d-83c8b015b3c6 · outbound

This paper cites A survey on software vulnerability exploitability assessment,.

Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation A survey on software vulnerability exploitability assessment,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:29:54.588511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:29:54.163038Z digest=sha256:ad1404316e1a99087447040c72c52fd06806b93d1bec32e8083a42fd06e385d2

Observation 16e2ca52-90ce-458c-94bf-026c5cb15b0d · outbound

This paper cites Koobe: Towards facilitating exploit generation of kernel out-of-bounds write vulnerabilities,.

Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation Koobe: Towards facilitating exploit generation of kernel out-of-bounds write vulnerabilities,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:29:54.574838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:29:54.168097Z digest=sha256:4bf101683ffcf2c79ada0b0e2f9d40058c32e6f1231e6e53e19c10ac00741435

Observation 25bbbe80-0a6f-49a1-88c5-64b6feb05401 · outbound

This paper cites Automatic generation of control flow hijacking exploits for software vulnerabilities,.

Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation Automatic generation of control flow hijacking exploits for software vulnerabilities,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:29:54.561551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:29:54.173086Z digest=sha256:d994b4694eb910473925f884144c367271a9b3e7b58e68fc86a8ee50ac371342

Observation f991f8a0-3dd6-4364-9169-7b98be97007c · outbound

This paper cites Automatic patch-based exploit generation is possible: Techniques and implications.

Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation Automatic patch-based exploit generation is possible: Techniques and implications

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:29:54.548724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:29:54.178368Z digest=sha256:53131baa2f66d92a62ece69a77d70fcf5e476db2cc84cfb6e733986e2678ce6c

Observation af88af8c-f82d-468f-9624-ef0d84488642 · outbound

This paper cites Aeg: Automatic exploit generation,.

Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation Aeg: Automatic exploit generation,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:29:54.534722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:29:54.182620Z digest=sha256:ad0e55d018a4c800902e483916541ae9faeba2947d20114ac36c60e6a7cc8a77

Observation a1ccd026-348e-47d9-b9bd-ed4e639e840e · outbound

This paper cites Fuze: Towards facilitating exploit generation for kernel use-after-free vulnerabilities,.

Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation Fuze: Towards facilitating exploit generation for kernel use-after-free vulnerabilities,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:29:54.520615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:29:54.187542Z digest=sha256:fd1d45e2a4cd64ad691cf12ada138ca7fda84b5273a3a4ea86a2b6d14da51e16

Observation 4a17078d-5c16-4b41-adef-c2e4e0ccbab6 · outbound

This paper cites Automated exploit generation for stack buffer overflow vulnerabilities,.

Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation Automated exploit generation for stack buffer overflow vulnerabilities,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:29:54.507338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:29:54.192865Z digest=sha256:279e23de00833af23dacb85ca56072c7c17dc9f0c2a7339fc207ee697f89296b

Observation 2e8041ab-56d3-4dfc-92df-a5c2665baff0 · outbound

This paper cites Automatic exploit generation for buffer overflow vulnerabilities,.

Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation Automatic exploit generation for buffer overflow vulnerabilities,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:29:54.493371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:29:54.198553Z digest=sha256:445ddf461613a58ba0cef062cb15ae288ab7648a719ba87732ce346b73cd2999

Observation 76b3eb8c-256e-4371-8fea-17a96ecd9066 · outbound

This paper cites Toward automated exploit generation for known vulnerabilities in open-source libraries,.

Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation Toward automated exploit generation for known vulnerabilities in open-source libraries,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:29:54.478717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:29:54.202744Z digest=sha256:150ff0c5d9354b2db05cc3f357cb65b166c793bd94ec78f9e6cb645c91a19df7

Observation 57b085b7-0d98-4787-ad42-4e4bedee8e43 · outbound

This paper cites Large language models for software engineering: A systematic literature review,.

Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation Large language models for software engineering: A systematic literature review,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:29:54.465079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:29:54.207949Z digest=sha256:061a72a42c65e674da49dd2dfac5e40f5f2b73c3014e8349cd5550680a3db2bd

Observation 4648c16b-1e36-4b5c-8496-9140c91bebbc · outbound

This paper cites How well does llm generate security tests?.

Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation How well does llm generate security tests?

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T04:29:54.212225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:29:54.212225Z digest=sha256:f863c7a9ff306e0c43bf1fc3a0291a5b63bbcf6c0be8e4aa921000fc1137177b

Observation 84c0efb3-9eb9-4c85-ade5-aef218334377 · outbound

This paper cites Pentestgpt: Evaluating and harnessing large language models for automated penetration testing,.

Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation Pentestgpt: Evaluating and harnessing large language models for automated penetration testing,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:29:54.450317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:29:54.217377Z digest=sha256:aabbd69ada1f18275a3dedf51eaa53cc018122880bcdb0c0853e2c33790a37f7

Observation e2d6bdbf-d8b5-4131-b894-f73ed706bfe2 · outbound

This paper cites Leveraging semantic relations in code and data to enhance taint analysis of embedded systems,.

Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation Leveraging semantic relations in code and data to enhance taint analysis of embedded systems,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:29:54.434757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:29:54.221650Z digest=sha256:86deea5f94788a434949a1b73a302e765b25fe0e8838c80ee4116c7f8460dc67

Observation 6a6b88f9-af50-441c-8202-a9028e85e0ba · outbound

This paper cites Seed labs,.

Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation Seed labs,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:29:54.419720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:29:54.227026Z digest=sha256:c258c02f593dc131070078d84ed35b209e5ff2bef6e708b7867ab3f1b0241524

Observation e00e9e8c-7c12-4717-abe2-8dcdeef5fa09 · outbound

This paper cites Explaining neural scaling laws,.

Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation Explaining neural scaling laws,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:29:54.404822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:29:54.231160Z digest=sha256:5569cc94ad201890d000d7cacd2af8ff452e81dd4f31fbf85b38ba3d07b4c9ea

Observation d0d970f3-6ec2-457e-b47f-0de6bb92b7e7 · outbound

This paper cites A comprehensive study of jailbreak attack versus defense for large language models,.

Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation A comprehensive study of jailbreak attack versus defense for large language models,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:29:54.390071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:29:54.235896Z digest=sha256:951ecf167036a78b3a663fd5412d9de6e40e7d405698539da44182b19be615e4

Observation 74e11491-6780-4e73-9193-2b2cf2983e19 · outbound

This paper cites Get up and running with large language models.

Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation Get up and running with large language models

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:29:54.362443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:29:54.244745Z digest=sha256:ba37968ffec343f5f85a86862368787cf0cc9ef70b19aff72fe585262758e34d

Observation d71657a1-e4b8-4071-92d6-167f4d0875c4 · outbound

This paper cites Available: https://api.semanticscholar.org/CorpusID: 267770234.

Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation Available: https://api.semanticscholar.org/CorpusID: 267770234

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:29:54.375589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:29:54.240107Z digest=sha256:c9c8b49b3f26880217b8133a088dd6ecd3b9f8049b86c14af436610411f8245d

Pith citing papers

Observation a750dab8-940a-44e0-bd63-468671c645fe · inbound

Knowdit: Agentic Smart Contract Vulnerability Detection with Auditing Knowledge Summarization cites this paper.

Knowdit: Agentic Smart Contract Vulnerability Detection with Auditing Knowledge Summarization Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-13T17:39:42.764726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T17:39:42.764726Z digest=sha256:a33571d45ba5016ff9f30a803a0a443e744fd8e33ce7e6a244d501c96a5f9bde

Observation c9316492-b658-4e01-b4c2-1351facc0ad4 · inbound

From Runnable to Verifiable: An Independent Reproducibility Study of LLM/Agent-Driven Vulnerability Validation Artifacts cites this paper.

From Runnable to Verifiable: An Independent Reproducibility Study of LLM/Agent-Driven Vulnerability Validation Artifacts Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-08-11T14:59:21.609935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T14:59:21.115774Z digest=sha256:1b0cda0adc3073dcfd2b01d23bcbb8a750b34102cb6959c64b1b9121a686b26e