Pith. sign in

Paper Citation Record · LEDGER

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities

As of 17 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 3 inbound Pith citation observations for arXiv:2509.03331.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.03331 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T11:02:54.374219Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T00:36:36.484575Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T21:10:09.278283Z

Reference resolution

38 of 38 outbound references displayed

  • verified exact2
  • verified fuzzy6
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1fe78373-78ef-4580-897d-c3f6e6f9c67d · outbound

This paper cites git-apply Documentation.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities git-apply Documentation

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:02:55.504115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T11:02:54.215896Z digest=sha256:4b4ea308eed2f3a50525dba3c5e5ef2f563b995b07f81dbe72a2019a40e1f8ba

Observation aae13449-36d6-4e44-82e8-f001940e9357 · outbound

This paper cites patch Command – IBM AIX Documentation.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities patch Command – IBM AIX Documentation

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:02:55.489752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T11:02:54.221430Z digest=sha256:6ed409a363dbad813e4dd1c9f9050545344c38ed72081e1107d2fbbd69673e43

Observation 3335d6c9-997a-4a55-abf1-39b619b0fd84 · outbound

This paper cites SecVulEval: Benchmarking LLMs for Real-World C/C++ Vulnerability Detection.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities SecVulEval: Benchmarking LLMs for Real-World C/C++ Vulnerability Detection

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.227103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.227103Z digest=sha256:d2aa547b1fe822440bcce171abfff92435b3a9ed1e288dddb1f5b8da582ce021

Observation 6aa13417-3567-4b24-87c2-647a4f7da1e2 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:02:55.474274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T11:02:54.232277Z digest=sha256:4e86bb4e67298b91ae3bbb28cb6606cc1fad330706df98510bd9887177081c41

Observation fcf7e0e2-84fd-43bc-b1f0-4d209fe7434f · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 5

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T11:02:55.159079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T11:02:54.240818Z digest=sha256:12ea7bda757247ae637a855ff1c0f6674c4fc6710c086cbd301d3116540106e6

Observation 2160a1e4-ca25-49b7-84aa-d9015e8f31ae · outbound

This paper cites Smith, Donald Stufft.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Smith, Donald Stufft

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:02:55.457829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T11:02:54.244993Z digest=sha256:f053b060cd343f1e50aebadbae7c77fb419897d25f8a3202330d05d116df00de

Observation 51f3abce-3622-4da9-baa8-ec4d4fca65e7 · outbound

This paper cites Díaz Ferreyra.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Díaz Ferreyra

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:02:55.441769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T11:02:54.249614Z digest=sha256:86c2b1fcfa01c3c5c9e67b84a7d78647d679dc70d77bd8b749bf3d13f6745272

Observation fa455ae8-adb9-4dee-9914-3948c20a4cd3 · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.257899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.257899Z digest=sha256:80765fa347718e0ea2d89a20ff1bfff08e9c5b83652098958549e3945cc6fe34

Observation 32531a87-5e05-4bb2-9c9f-a93a7e93b598 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:02:55.426094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T11:02:54.262484Z digest=sha256:045b5205959a667417c419f8ff34a69858ec87579f9c604db365d6c1fdef4d27

Observation ccacb866-ac59-4ebe-a5de-aad7dd7a32c3 · outbound

This paper cites 2024.{PentestGPT}: Evaluating and harnessing large language models for automated penetration testing.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities 2024.{PentestGPT}: Evaluating and harnessing large language models for automated penetration testing

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:02:55.410264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T11:02:54.266534Z digest=sha256:71cc98ae57179b4fec7627897023accf1a4866a9582b1a0e6dd9e3c591b21282

Observation 9975aef7-e0ce-43ae-bcf8-fa8bb060bcba · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.270368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.270368Z digest=sha256:839f06d463c9e6e247ea8706e84c144bf10d82920c46229d400a1f44eb865f61

Observation eda7427e-65f4-4ccb-a686-4742b5b712b8 · outbound

This paper cites SoK: Automated Vulnerability Repair: Methods, Tools, and Assessments.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities SoK: Automated Vulnerability Repair: Methods, Tools, and Assessments

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.274440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.274440Z digest=sha256:e3b96b722fb8128a05f15e508380f7601b000318dab890ad6c50dcfca61b7ddc

Observation 78c2840b-a16b-462f-8517-8453d680abe7 · outbound

This paper cites GPT-4o System Card.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities GPT-4o System Card

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.278555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.278555Z digest=sha256:19ee7e0a20699740d81ab60d21e77e6b536b390d1a0c505f738b784b17444e81

Observation fa2a095b-71c5-4e85-9e5c-17d13bc37afb · outbound

This paper cites SWE-bench: Can Language Models Resolve Real-World GitHub Issues?.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities SWE-bench: Can Language Models Resolve Real-World GitHub Issues?

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.282800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.282800Z digest=sha256:3729fe901662195de6240b804edfffd0875254427707628dd98076b1c42f9408

Observation c5ce7e3e-d9f2-482e-8a85-40c6ff2a6991 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 15

Resolution
verified exact
doi, observed 2026-08-05T11:02:54.433629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T11:02:54.286766Z digest=sha256:f0a7722ee4f11ab3f0175032f721335ff2e6c8bb321543079e71ffb700ddcd86

Observation 94b8e3ec-25f3-4208-a029-5e9f5af06b14 · outbound

This paper cites DeepSeek-V3 Technical Report.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities DeepSeek-V3 Technical Report

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.290488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.290488Z digest=sha256:cf61ce70201851eb1762bc3882892c1879b3505405969ac838b0f56631a4cc7b

Observation b19cda7d-ec03-4ea9-a087-29efc38f48da · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 17

Resolution
verified exact
raw_fallback, observed 2026-08-05T11:02:54.868133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T11:02:54.294508Z digest=sha256:261fd1e73983892f6b71d72efd2ae9ca6de741df8045f2bfc892f6ad03c15db0

Observation daf00636-cbd1-45d5-af4b-823bc58e582e · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.298556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.298556Z digest=sha256:4e61ac74e4af489b99a35569264dac72d0098b9f5553beb1ceead1aabd90a2b6

Observation 5c25037c-6566-4125-9537-c9bb6f45048b · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.302738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.302738Z digest=sha256:a65c0620b8a0e29421d7a5c6ef3a44a3713953a5a6c0e4ad15dbdecd227a1871

Observation f87cfc28-6fc0-4aef-b2b0-e21b1884ea57 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:02:55.394487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T11:02:54.306710Z digest=sha256:a8b74a2153ddb5a063df12b8c4dccdac61089eb62fdbfac90a7f3fcecbbf3061

Observation 638c9e38-958c-4a99-8720-cfd443933dee · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:02:55.378799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T11:02:54.310590Z digest=sha256:1959703c48f8867faf2a28603cd05d23af25f2702a51d18f23cf90f10285ca76

Observation dd2a5acd-51d3-4125-9988-298a3e56e04c · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:02:55.360546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T11:02:54.314475Z digest=sha256:3b4c27830763c7c6559e5c1a1d38c52c493469467e6f17130e18bb5ced98de49

Observation dae2b195-b9f8-4b0a-bf7a-c0d0fe8dace1 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:02:55.345392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T11:02:54.318435Z digest=sha256:9b3f4f2f0c1c82772bbb00f483a475c7410b9e738f781dd104ebcb6f8184be63

Observation 894ea19f-28fb-40ab-82a3-0d38da2b47f9 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:02:55.330514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T11:02:54.322331Z digest=sha256:5bf291f20e2867c2c9077e95537df17b67ff2e4b56e83b43c6edc5cfad4b2055

Observation d68be404-6158-48e5-84aa-bd3386653d12 · outbound

This paper cites Are Large Language Models Robust in Understanding Code Against Semantics-Preserving Mutations?.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Are Large Language Models Robust in Understanding Code Against Semantics-Preserving Mutations?

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.326649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.326649Z digest=sha256:2177ff2fff9281107d7ec0c9d770f83e524b7244cef29eae0ff887f67f5a3836

Observation 4027879f-411b-4ce2-82d8-470bf078a7db · outbound

This paper cites PoCGen: Generating Proof-of-Concept Exploits for Vulnerabilities in Npm Packages.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities PoCGen: Generating Proof-of-Concept Exploits for Vulnerabilities in Npm Packages

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.330975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.330975Z digest=sha256:5b38233cfabec6819cec25ec10c1d52e621eff74fc904e1584f3bc8708950510

Observation 93938fae-9036-4db6-8c12-d5deee521d12 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:02:55.314977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T11:02:54.335305Z digest=sha256:46f94dbed96fe857f2a1cfa7f396f2eb2fb22b1963d639cf77b897f1d9c67ffd

Observation d0b7014b-3e98-4d3c-af4d-aa9436733ee1 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:02:55.296795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T11:02:54.339819Z digest=sha256:6581c502a981f6e140d84427491e6cccd40f771bcf23318529a490dfd5d8c32f

Observation 8a98f66b-e71a-4e1a-ac0c-fdf38f97bfa9 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.344033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.344033Z digest=sha256:feb70ce7051898d470906b1f1fd168f9550b897a5e1769cc5cc8d3bb013806b0

Observation 9cd4b76a-358a-491e-9a76-0e1296531cd4 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 30

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T11:02:54.738213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T11:02:54.348211Z digest=sha256:f605a4b0ef7a5dd027e6577024800448f906e85139d5de8b12af811ca2ee6a0e

Observation fbd3072f-6bfa-4d4c-a6dc-f74e563abc10 · outbound

This paper cites Chi, Quoc V.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Chi, Quoc V

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:02:55.268973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T11:02:54.352409Z digest=sha256:a32794f3fe565a543e6f20a26c90433a5ac8ff136b2234bf77063d03bfcff7a4

Observation 9acab7d4-dcaa-44f8-956e-afa3feefab96 · outbound

This paper cites SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.356699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.356699Z digest=sha256:9b0d12f4de3c44ec5fc0400774c952c19df063581147fead2ac2c72abf45b9a6

Observation 4510945b-4e26-41f5-82ba-5f17b02228d2 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.361313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.361313Z digest=sha256:c6d69e4086e782ff202ef2484558d4c198fe59996f655cadcdcbf4a224430a26

Observation 6bec30c8-092d-490b-9c66-b16ad0df4e5b · outbound

This paper cites Qwen3 Technical Report.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Qwen3 Technical Report

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.365718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.365718Z digest=sha256:3c964ed82c3a2d85f2c98709fe6bafe15c2f9642e0fc48ed75d667fd183a1a29

Observation 70468bcb-b603-4346-9b14-e217fe97db3e · outbound

This paper cites SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.369975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.369975Z digest=sha256:883be997d7346d5bff33e62214aeabd0a2acd4c3adc5c4e3f528d7563aabea4c

Observation 4eee2536-09a9-429d-a48e-5bb0e504d29c · outbound

This paper cites BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.374219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.374219Z digest=sha256:ee8da344c22da76c93bff99de8b622b5e1e51157974545a9c8683dd6969ce19e

Observation 5ae84726-e274-437f-98ff-e1aace6cfe89 · outbound

This paper cites an unresolved cited work.

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities Unresolved cited work

Reference 468

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.253699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.253699Z digest=sha256:34110fc9daf8502a1925f61fa88f934c1ab0b10241e0162ed67dd54c32861bdb

Observation 0e0f64d6-918a-46ff-974b-6e73324bd0a2 · outbound

This paper cites In Proceedings of the 20th International Confer- ence on Predictive Models and Data Analytics in Software Engineering (PROMISE 2024).

VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities In Proceedings of the 20th International Confer- ence on Predictive Models and Data Analytics in Software Engineering (PROMISE 2024)

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-05T11:02:54.236609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:02:54.236609Z digest=sha256:dd53ee7f239910f9f367613c770ca6d102494127284079b3416ab256354cad08

Pith citing papers

Observation b7c6bc9d-3adf-4169-8b9d-9aa773a4d679 · inbound

Towards Demystifying and Repairing LLM-in-the-Loop Vulnerabilities cites this paper.

Towards Demystifying and Repairing LLM-in-the-Loop Vulnerabilities VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-06-29T11:23:20.652423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-29T11:21:33.012202Z digest=sha256:4379cc6268ee255ec9ce78068976df9ecc07964ff723fd4098d0cbd3e9d72448

Observation d7933b54-82de-45c3-a3a0-052065f36990 · inbound

Helpful or Harmful? Evaluating LLM-Assisted Vulnerability Patching via a Human Study cites this paper.

Helpful or Harmful? Evaluating LLM-Assisted Vulnerability Patching via a Human Study VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-04T21:10:09.279825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-25T19:02:45.109478Z digest=sha256:39492fbeebec34d45f21f79930569fa692ef7ab49890b7e373b0cae69ac99762

Observation 61821484-3ea5-4605-8978-8c3a73778f7f · inbound

Vul4Py: Benchmarking Automated Vulnerability Repair in Python with Paired Exploit and Functional Oracles cites this paper.

Vul4Py: Benchmarking Automated Vulnerability Repair in Python with Paired Exploit and Functional Oracles VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T00:36:36.484575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:36:36.484575Z digest=sha256:71e2481dddf12fc8128201616f308c9e3b6d90751aaa392221fb8018d5c81533