Pith. sign in

Paper Citation Record · LEDGER

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities

As of 19 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 3 inbound Pith citation observations for arXiv:2509.03331.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.03331 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T11:02:54.374219Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T00:36:36.484575Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T21:10:09.278283Z

Reference resolution

38 of 38 outbound references displayed

  • verified exact2
  • verified fuzzy6
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1fe78373-78ef-4580-897d-c3f6e6f9c67d · outbound

This paper cites git-apply Documentation.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities git-apply Documentation

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:02:55.504115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T11:02:54.215896Z digest=sha256:b9c7df30dea9fad237dd9c738ea50f7da9f7e778b244e51ae857358506c852f6

Observation aae13449-36d6-4e44-82e8-f001940e9357 · outbound

This paper cites patch Command – IBM AIX Documentation.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities patch Command – IBM AIX Documentation

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:02:55.489752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T11:02:54.221430Z digest=sha256:aff2dd3d9db6817b6f42f902e0012207ab14268328f447ad12157ed2b8ccd20b

Observation 3335d6c9-997a-4a55-abf1-39b619b0fd84 · outbound

This paper cites SecVulEval: Benchmarking LLMs for Real-World C/C++ Vulnerability Detection.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities SecVulEval: Benchmarking LLMs for Real-World C/C++ Vulnerability Detection

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.227103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.227103Z digest=sha256:d2aa547b1fe822440bcce171abfff92435b3a9ed1e288dddb1f5b8da582ce021

Observation 6aa13417-3567-4b24-87c2-647a4f7da1e2 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:02:55.474274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T11:02:54.232277Z digest=sha256:a63adf39fe838d6de995d8c3c7b97878ecc4453969a400c12597f5970b7f74c3

Observation fcf7e0e2-84fd-43bc-b1f0-4d209fe7434f · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 5

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T11:02:55.159079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T11:02:54.240818Z digest=sha256:0ae89c34ef7f7fbbb380d2024432329e44d35969f9ab93892756c4f5a436974f

Observation 2160a1e4-ca25-49b7-84aa-d9015e8f31ae · outbound

This paper cites Smith, Donald Stufft.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Smith, Donald Stufft

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:02:55.457829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T11:02:54.244993Z digest=sha256:0533d38157586be7e638c455c24307d59b342cb2bb186df94740023692426fe9

Observation 51f3abce-3622-4da9-baa8-ec4d4fca65e7 · outbound

This paper cites Díaz Ferreyra.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Díaz Ferreyra

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:02:55.441769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T11:02:54.249614Z digest=sha256:95f438e4e016f7064467688bdc9443b84f82b41c6c9f1c269d213a57c50f14ec

Observation fa455ae8-adb9-4dee-9914-3948c20a4cd3 · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.257899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.257899Z digest=sha256:80765fa347718e0ea2d89a20ff1bfff08e9c5b83652098958549e3945cc6fe34

Observation 32531a87-5e05-4bb2-9c9f-a93a7e93b598 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:02:55.426094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T11:02:54.262484Z digest=sha256:f0a363937f653223cfb3d58db537c3fd0d74af30d922b8dc22dbb640b16f4720

Observation ccacb866-ac59-4ebe-a5de-aad7dd7a32c3 · outbound

This paper cites 2024.{PentestGPT}: Evaluating and harnessing large language models for automated penetration testing.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities 2024.{PentestGPT}: Evaluating and harnessing large language models for automated penetration testing

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:02:55.410264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T11:02:54.266534Z digest=sha256:c122bbaa996d099a839b177e60d84bc35f84ec7d32413825ca0b99cd7e060418

Observation 9975aef7-e0ce-43ae-bcf8-fa8bb060bcba · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.270368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.270368Z digest=sha256:839f06d463c9e6e247ea8706e84c144bf10d82920c46229d400a1f44eb865f61

Observation eda7427e-65f4-4ccb-a686-4742b5b712b8 · outbound

This paper cites SoK: Automated Vulnerability Repair: Methods, Tools, and Assessments.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities SoK: Automated Vulnerability Repair: Methods, Tools, and Assessments

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.274440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.274440Z digest=sha256:10d6a8db34e1998c492469034cb2217c8a07c4cda4affcee2ff045d07997a402

Observation 78c2840b-a16b-462f-8517-8453d680abe7 · outbound

This paper cites GPT-4o System Card.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities GPT-4o System Card

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.278555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.278555Z digest=sha256:19ee7e0a20699740d81ab60d21e77e6b536b390d1a0c505f738b784b17444e81

Observation fa2a095b-71c5-4e85-9e5c-17d13bc37afb · outbound

This paper cites SWE-bench: Can Language Models Resolve Real-World GitHub Issues?.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities SWE-bench: Can Language Models Resolve Real-World GitHub Issues?

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.282800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.282800Z digest=sha256:b26c32d78b92fbb5b007937a9b85c7c83f38b2e190a4a2e3d0f00950f3c38c92

Observation c5ce7e3e-d9f2-482e-8a85-40c6ff2a6991 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 15

Resolution
verified exact
doi, observed 2026-08-05T11:02:54.433629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T11:02:54.286766Z digest=sha256:4eae61d99fe604e83314e591929c62321ba7e9852d954485a84a6cb99cfb1925

Observation 94b8e3ec-25f3-4208-a029-5e9f5af06b14 · outbound

This paper cites DeepSeek-V3 Technical Report.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities DeepSeek-V3 Technical Report

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.290488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.290488Z digest=sha256:00670921e5509bfd57c5debc8b2464195485a7b003454705bae65bb89c771e8f

Observation b19cda7d-ec03-4ea9-a087-29efc38f48da · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 17

Resolution
verified exact
raw_fallback, observed 2026-08-05T11:02:54.868133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T11:02:54.294508Z digest=sha256:b04007347c54528b2aa3fd8fa9b7952211d5bd4dfe300e5cae9f643c04dd414a

Observation daf00636-cbd1-45d5-af4b-823bc58e582e · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.298556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.298556Z digest=sha256:4e61ac74e4af489b99a35569264dac72d0098b9f5553beb1ceead1aabd90a2b6

Observation 5c25037c-6566-4125-9537-c9bb6f45048b · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.302738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.302738Z digest=sha256:a65c0620b8a0e29421d7a5c6ef3a44a3713953a5a6c0e4ad15dbdecd227a1871

Observation f87cfc28-6fc0-4aef-b2b0-e21b1884ea57 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:02:55.394487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T11:02:54.306710Z digest=sha256:a9295bf86cb37936b38698ab1c388bb87c9650d22b06e3b287b1f832041b5c41

Observation 638c9e38-958c-4a99-8720-cfd443933dee · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:02:55.378799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T11:02:54.310590Z digest=sha256:ebf2e537fe6f5e23146f7dbc304d5680733943c7c012e71d8655f28581219261

Observation dd2a5acd-51d3-4125-9988-298a3e56e04c · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:02:55.360546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T11:02:54.314475Z digest=sha256:e819872169c5ed6b63b7fc786d6b084d2a22ee9c4ada3269ff39e02aa99b9dfc

Observation dae2b195-b9f8-4b0a-bf7a-c0d0fe8dace1 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:02:55.345392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T11:02:54.318435Z digest=sha256:63e376781e9ea0ce79120bb223d7d2802d8b6b7a6243d259bd7e01b76a709eca

Observation 894ea19f-28fb-40ab-82a3-0d38da2b47f9 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:02:55.330514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T11:02:54.322331Z digest=sha256:9dfd97156be859e06bd560f4d9c9f2132a8d26be7211fc612cf404cc7344fa4a

Observation d68be404-6158-48e5-84aa-bd3386653d12 · outbound

This paper cites Are Large Language Models Robust in Understanding Code Against Semantics-Preserving Mutations?.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Are Large Language Models Robust in Understanding Code Against Semantics-Preserving Mutations?

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.326649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.326649Z digest=sha256:2177ff2fff9281107d7ec0c9d770f83e524b7244cef29eae0ff887f67f5a3836

Observation 4027879f-411b-4ce2-82d8-470bf078a7db · outbound

This paper cites PoCGen: Generating Proof-of-Concept Exploits for Vulnerabilities in Npm Packages.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities PoCGen: Generating Proof-of-Concept Exploits for Vulnerabilities in Npm Packages

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.330975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.330975Z digest=sha256:5b38233cfabec6819cec25ec10c1d52e621eff74fc904e1584f3bc8708950510

Observation 93938fae-9036-4db6-8c12-d5deee521d12 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:02:55.314977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T11:02:54.335305Z digest=sha256:d778f5b83516510afd6c3ec5da5e22aa0bbbeed9f494f15930c94a2f0e9fe32d

Observation d0b7014b-3e98-4d3c-af4d-aa9436733ee1 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:02:55.296795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T11:02:54.339819Z digest=sha256:6ccfb0daa854641b416336d16aeb0d4aca4c6b90da27eaebeb2ac11b701100d8

Observation 8a98f66b-e71a-4e1a-ac0c-fdf38f97bfa9 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.344033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.344033Z digest=sha256:feb70ce7051898d470906b1f1fd168f9550b897a5e1769cc5cc8d3bb013806b0

Observation 9cd4b76a-358a-491e-9a76-0e1296531cd4 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 30

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T11:02:54.738213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T11:02:54.348211Z digest=sha256:10e2db4ca158a563937c01fbc96a2dba894e0e7c109dfc23480cb3017f0c7d86

Observation fbd3072f-6bfa-4d4c-a6dc-f74e563abc10 · outbound

This paper cites Chi, Quoc V.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Chi, Quoc V

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:02:55.268973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T11:02:54.352409Z digest=sha256:f9042b46ec53abdb2c92152e566f821a9399458b3af26914c13e8a87a8f16e5c

Observation 9acab7d4-dcaa-44f8-956e-afa3feefab96 · outbound

This paper cites SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.356699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.356699Z digest=sha256:9b0d12f4de3c44ec5fc0400774c952c19df063581147fead2ac2c72abf45b9a6

Observation 4510945b-4e26-41f5-82ba-5f17b02228d2 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.361313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.361313Z digest=sha256:c6d69e4086e782ff202ef2484558d4c198fe59996f655cadcdcbf4a224430a26

Observation 6bec30c8-092d-490b-9c66-b16ad0df4e5b · outbound

This paper cites Qwen3 Technical Report.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Qwen3 Technical Report

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.365718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.365718Z digest=sha256:3c964ed82c3a2d85f2c98709fe6bafe15c2f9642e0fc48ed75d667fd183a1a29

Observation 70468bcb-b603-4346-9b14-e217fe97db3e · outbound

This paper cites SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.369975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.369975Z digest=sha256:883be997d7346d5bff33e62214aeabd0a2acd4c3adc5c4e3f528d7563aabea4c

Observation 4eee2536-09a9-429d-a48e-5bb0e504d29c · outbound

This paper cites BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.374219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.374219Z digest=sha256:ee8da344c22da76c93bff99de8b622b5e1e51157974545a9c8683dd6969ce19e

Observation 5ae84726-e274-437f-98ff-e1aace6cfe89 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 468

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.253699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.253699Z digest=sha256:34110fc9daf8502a1925f61fa88f934c1ab0b10241e0162ed67dd54c32861bdb

Observation 0e0f64d6-918a-46ff-974b-6e73324bd0a2 · outbound

This paper cites In Proceedings of the 20th International Confer- ence on Predictive Models and Data Analytics in Software Engineering (PROMISE 2024).

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities In Proceedings of the 20th International Confer- ence on Predictive Models and Data Analytics in Software Engineering (PROMISE 2024)

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.236609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.236609Z digest=sha256:dd53ee7f239910f9f367613c770ca6d102494127284079b3416ab256354cad08

Pith citing papers

Observation b7c6bc9d-3adf-4169-8b9d-9aa773a4d679 · inbound

Towards Demystifying and Repairing LLM-in-the-Loop Vulnerabilities cites this paper.

Towards Demystifying and Repairing LLM-in-the-Loop Vulnerabilities VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-06-29T11:23:20.652423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-29T11:21:33.012202Z digest=sha256:f4370f846f99b00579e3847a4d2a8b6864b3369509ec24ae476fe5298ab44592

Observation d7933b54-82de-45c3-a3a0-052065f36990 · inbound

Helpful or Harmful? Evaluating LLM-Assisted Vulnerability Patching via a Human Study cites this paper.

Helpful or Harmful? Evaluating LLM-Assisted Vulnerability Patching via a Human Study VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-04T21:10:09.279825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-25T19:02:45.109478Z digest=sha256:63b0e00a23b887e8457128085b2e2a848e4829fcce3801eba5da76c5c59bf587

Observation 61821484-3ea5-4605-8978-8c3a73778f7f · inbound

Vul4Py: Benchmarking Automated Vulnerability Repair in Python with Paired Exploit and Functional Oracles cites this paper.

Vul4Py: Benchmarking Automated Vulnerability Repair in Python with Paired Exploit and Functional Oracles VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T00:36:36.484575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:36:36.484575Z digest=sha256:71e2481dddf12fc8128201616f308c9e3b6d90751aaa392221fb8018d5c81533