Pith. sign in

Paper Citation Record · LEDGER

Empirical Evaluation of Large Language Models in Automated Program Repair

As of 19 August 2026, this Paper Citation Record lists 79 of 79 outbound references and 2 inbound Pith citation observations for arXiv:2506.13186.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.13186 v1

Coverage vector

measured 79 of 79 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:42:05.472880Z

measured 81 of 81 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T15:06:55.050376Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T05:57:24.432794Z

Reference resolution

79 of 79 outbound references displayed

  • verified exact0
  • verified fuzzy56
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b167f023-3aeb-4ead-87bb-3be516ef7ee1 · outbound

This paper cites The debugging mindset: Understanding the psychology of learning strategies leads to effective problem-solving skills.

Empirical Evaluation of Large Language Models in Automated Program Repair The debugging mindset: Understanding the psychology of learning strategies leads to effective problem-solving skills

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.240703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:41:59.444522Z digest=sha256:a110fde6c2823f130068cd20a9a16243efeb3e78063e7dd132bcc7e3d783e254

Observation b0258e0a-e192-4a1c-8c68-00beae8a946a · outbound

This paper cites Genprog: A generic method for automatic software repair,.

Empirical Evaluation of Large Language Models in Automated Program Repair Genprog: A generic method for automatic software repair,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.230666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:41:59.501952Z digest=sha256:caa999ea530c1bf2974224b29d779c5bcb98bc866e6dd9527b62766434f54ef5

Observation fec9e93c-374f-43be-ab4d-ccb6c5fdc2b6 · outbound

This paper cites Nopol: Automatic repair of conditional statement bugs in java programs,.

Empirical Evaluation of Large Language Models in Automated Program Repair Nopol: Automatic repair of conditional statement bugs in java programs,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.221383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:41:59.555989Z digest=sha256:e5528a66b4b880f2a6b759e7f138ab2b898a2eb2351da1438cb3bc75ae2b55d9

Observation 1be5c364-e6c2-4dd1-9dff-5d0d56b98bf4 · outbound

This paper cites S3: syntax- and semantic-guided repair synthesis via programming by examples,.

Empirical Evaluation of Large Language Models in Automated Program Repair S3: syntax- and semantic-guided repair synthesis via programming by examples,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.211545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:41:59.663118Z digest=sha256:8212dca74feaf0ba99aa00f2658223ad058add591be32bf560ae078683c676ec

Observation f3684160-823d-471d-aabd-bf85a36a4d0c · outbound

This paper cites Staged program repair with condition synthesis,.

Empirical Evaluation of Large Language Models in Automated Program Repair Staged program repair with condition synthesis,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.202384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:41:59.728744Z digest=sha256:1955e6cdc615ab1021d3f84ed4856d72a8ba57941531cfb5e499e4f54cae3937

Observation e5142586-429b-4d87-8f3b-e8613068e4c7 · outbound

This paper cites Angelix: Scalable multiline program patch synthesis via symbolic analysis,.

Empirical Evaluation of Large Language Models in Automated Program Repair Angelix: Scalable multiline program patch synthesis via symbolic analysis,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.192887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:41:59.800504Z digest=sha256:1d0aa916f6951ed6131f8d89fafe68e8d79697f8c8ea9430ceb0611f96013639

Observation 12bfc4d2-6a0e-4392-9338-213d63a946a7 · outbound

This paper cites Astor: A program repair library for java,.

Empirical Evaluation of Large Language Models in Automated Program Repair Astor: A program repair library for java,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.183114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:41:59.850334Z digest=sha256:7336453abac6c0395873589646f221f2ebdf8f9b3c7d4f5f49ecb0ae60272ffc

Observation dcbd5c05-4553-4527-a135-7da6f642e4f1 · outbound

This paper cites History driven program repair,.

Empirical Evaluation of Large Language Models in Automated Program Repair History driven program repair,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.173408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:41:59.908963Z digest=sha256:b0b0f343ef1b4b44def74ba138ff47bd7e7b8b418609e9cc8f0e2d0b8f81b308

Observation d4f33bc5-e716-4ef2-a8db-f104b7270513 · outbound

This paper cites Automatic patch generation by learning correct code,.

Empirical Evaluation of Large Language Models in Automated Program Repair Automatic patch generation by learning correct code,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.163121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:41:59.998933Z digest=sha256:2d96eac6d3ffcfcd3e46688ff1602ca73e2d218af57bba529af56776f7711095

Observation a82ae89a-5d64-4f65-8722-3a611339bd54 · outbound

This paper cites Leveraging syntax-related code for automated program repair,.

Empirical Evaluation of Large Language Models in Automated Program Repair Leveraging syntax-related code for automated program repair,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.152650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:00.044997Z digest=sha256:5db440d4e3367d81ae38ad16f9029283cd2876f92d2ca95f69b109e47294c712

Observation 1a9e125e-47fc-43ad-9b21-e394ad4819b3 · outbound

This paper cites Precise condition synthesis for program repair,.

Empirical Evaluation of Large Language Models in Automated Program Repair Precise condition synthesis for program repair,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.143327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:00.134660Z digest=sha256:22bf139924b8799fef89305edf778dc570d317709ef1324e726e7a3aea7ec88a

Observation 8daf1de8-8a74-4f81-8fd6-cd5eb1935dc9 · outbound

This paper cites Automatic inference of code transforms for patch generation,.

Empirical Evaluation of Large Language Models in Automated Program Repair Automatic inference of code transforms for patch generation,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.133299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:00.179038Z digest=sha256:4f455e62dc5a5ce0245b502da2e38960a4f0dd01de5da773a47cc00a03cb2774

Observation a91899b0-3bc4-4ce6-bf47-fa49b316d55d · outbound

This paper cites Towards practical program repair with on-demand candidate generation,.

Empirical Evaluation of Large Language Models in Automated Program Repair Towards practical program repair with on-demand candidate generation,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.124032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:00.235932Z digest=sha256:ee5e034a6c245252f4ac7461073e7b4e4077dd8aeaacb0aeeffbca4c80f365e3

Observation 2419a248-dd07-4c16-9866-400c3d19ee80 · outbound

This paper cites Context-aware patch generation for better automated program repair,.

Empirical Evaluation of Large Language Models in Automated Program Repair Context-aware patch generation for better automated program repair,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.114649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:00.296258Z digest=sha256:c7b82f3115912f646904510bc62713ac1db8e49afc79349cfba4fd89769997e2

Observation 14cee7f8-6e84-4240-8655-7ae6d43ef020 · outbound

This paper cites Shaping program repair space with existing patches and similar code,.

Empirical Evaluation of Large Language Models in Automated Program Repair Shaping program repair space with existing patches and similar code,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.105045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:00.355358Z digest=sha256:bc4c9cb1bb94f8159cd9a6e2fae10aeb70c5f3387ca1e9f3913b5e3e81952fa8

Observation eab405de-641b-47f9-8578-f29250fee8fc · outbound

This paper cites Tbar: Revisiting template-based automated program repair,.

Empirical Evaluation of Large Language Models in Automated Program Repair Tbar: Revisiting template-based automated program repair,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.095272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:00.412296Z digest=sha256:72550b6da1feea85533e0945e93b1905e15ce7b976a6f1259f313fa6315a20cf

Observation df8f5b3c-5fdc-460c-a4b8-91ddb4702f95 · outbound

This paper cites Avatar: Fixing semantic bugs with fix patterns of static analysis violations,.

Empirical Evaluation of Large Language Models in Automated Program Repair Avatar: Fixing semantic bugs with fix patterns of static analysis violations,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.085697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:00.508136Z digest=sha256:fa25c2b740fc18d8bebfb1c90ae1c201fb020dbe6e1cd3c3cd927664772317a5

Observation 824fb81a-f85f-4cdd-be5d-7612145a97ea · outbound

This paper cites Practical program repair via bytecode mutation,.

Empirical Evaluation of Large Language Models in Automated Program Repair Practical program repair via bytecode mutation,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.075497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:00.563390Z digest=sha256:edf891a423b4c91b3406dd810d088adf456c06ad987c23c17e08f7c7cfba0369

Observation 40a98f43-ce8c-4078-a88a-9c3e76c300d4 · outbound

This paper cites Inferring program transfor- mations from singular examples via big code,.

Empirical Evaluation of Large Language Models in Automated Program Repair Inferring program transfor- mations from singular examples via big code,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.064987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:00.614514Z digest=sha256:100592ac4673c47be785624af9dbec4143e3e0feb2efe4d37b54a6e6558b4bd5

Observation 6e5ff99e-aa98-4872-9f3d-93c009ec6061 · outbound

This paper cites The plastic surgery hypothesis in the era of large language models,.

Empirical Evaluation of Large Language Models in Automated Program Repair The plastic surgery hypothesis in the era of large language models,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.055015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:00.669053Z digest=sha256:9ba3b4da06deff6df322a914c6726c7f7bffd4a9249910b591ea8d60981d3621

Observation 76c5b301-c4e8-4f53-b2ca-f42df3cd92d7 · outbound

This paper cites Sequencer: Sequence-to-sequence learning for end- to-end program repair,.

Empirical Evaluation of Large Language Models in Automated Program Repair Sequencer: Sequence-to-sequence learning for end- to-end program repair,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.044297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:00.722403Z digest=sha256:e0bfb3b1956c5f7f9daec817bc2c34b746faa0c6c573a217be90fe7ff4fedf43

Observation 20d6c679-fdae-4ccb-854e-c531c40ebe60 · outbound

This paper cites Coconut: combining context-aware neural translation models using ensemble for program repair,.

Empirical Evaluation of Large Language Models in Automated Program Repair Coconut: combining context-aware neural translation models using ensemble for program repair,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.035055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:00.775241Z digest=sha256:89c90d87a7ed5dcc07ddc7db8764c406e8f638971c726f840664aef33b8b7e30

Observation 83986091-dad1-4b05-8a6e-9e7f4ee67b54 · outbound

This paper cites Dlfix: Context-based code transformation learning for automated program repair,.

Empirical Evaluation of Large Language Models in Automated Program Repair Dlfix: Context-based code transformation learning for automated program repair,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.025782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:00.856107Z digest=sha256:c625a868cacbd6d32bfa2a527b03ddb8bb8e329ee328da79bb1777737983efac

Observation e6de4167-2e91-408d-a26e-4926315cfdb1 · outbound

This paper cites A syntax-guided edit decoder for neural program repair,.

Empirical Evaluation of Large Language Models in Automated Program Repair A syntax-guided edit decoder for neural program repair,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.016369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:00.924694Z digest=sha256:b6b289176ed50c723c6dd872c0b82db947f81b548d896f0212b30c4b6e9f9c11

Observation 8796f69d-9ee4-44d4-bb53-ffaca7f768e3 · outbound

This paper cites Cure: Code-aware neural machine translation for automatic program repair,.

Empirical Evaluation of Large Language Models in Automated Program Repair Cure: Code-aware neural machine translation for automatic program repair,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:06.005991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:00.976528Z digest=sha256:66ee63784080c80a0b28e635cb17c5cdc2bdc271372c7ad9a585e312337b9cc9

Observation be5ffd9d-7362-4a38-95c6-0f8edb534c66 · outbound

This paper cites Neural program repair with execution-based backpropagation,.

Empirical Evaluation of Large Language Models in Automated Program Repair Neural program repair with execution-based backpropagation,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.995572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:01.051254Z digest=sha256:9d58f69a991b32a10623d6d61487a488315d8a9480f6aaaedcfbb823fdbc45f6

Observation eb43304e-50c7-41e2-a187-5ca3ef30ac27 · outbound

This paper cites Tare: Type-aware neural program repair,.

Empirical Evaluation of Large Language Models in Automated Program Repair Tare: Type-aware neural program repair,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.984769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:01.187004Z digest=sha256:f745fa5ad5a2be7f6bc80e10da4d87899a38fe258cab0e8b963e7b0611c1eb21

Observation 39adb762-c39f-4fb0-8f79-b49f7e1bc8df · outbound

This paper cites Vulrepair: a t5-based automated software vulnerability repair,.

Empirical Evaluation of Large Language Models in Automated Program Repair Vulrepair: a t5-based automated software vulnerability repair,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.975049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:01.293848Z digest=sha256:57ebc12333947e1bfbebee6d58394356bb457c7a4beab3652a0214e14313c896

Observation 45ed4b8a-9196-4c11-93e4-7d97271e9cf8 · outbound

This paper cites Less training, more repairing please: revisiting automated program repair via zero-shot learning,.

Empirical Evaluation of Large Language Models in Automated Program Repair Less training, more repairing please: revisiting automated program repair via zero-shot learning,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.964939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:01.341834Z digest=sha256:9bc7d5113f5b0240c42e4cd9e6a1b64b741b6233f30b89e80315eabcca2c1c99

Observation bd9d3dd8-2e81-4187-a3f6-eadb708e6b81 · outbound

This paper cites Prompting is all you need: Automated android bug replay with large language models,.

Empirical Evaluation of Large Language Models in Automated Program Repair Prompting is all you need: Automated android bug replay with large language models,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:01.431179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:01.431179Z digest=sha256:2278bcf88a150a728edf31e45b90edfa588e36d26fb9d565654f31da358ed0d9

Observation f30c9134-5bde-46a6-83ae-1f22e9b1e495 · outbound

This paper cites UniXcoder: Unified Cross-Modal Pre-training for Code Representation.

Empirical Evaluation of Large Language Models in Automated Program Repair UniXcoder: Unified Cross-Modal Pre-training for Code Representation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:01.663531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:01.663531Z digest=sha256:3807c92be736aa928eb07d7cbeed5aad0dff5f95dfc0f42f064239c28ff37de2

Observation 87540288-3fae-4b86-ad57-5ffbc7223095 · outbound

This paper cites CodeGen: An Open Large Language Model for Code with Multi-Turn Program Synthesis.

Empirical Evaluation of Large Language Models in Automated Program Repair CodeGen: An Open Large Language Model for Code with Multi-Turn Program Synthesis

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:01.795662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:01.795662Z digest=sha256:66972f29901432094461f3d18bcef9bc5cd553954b731b42d4d267dddbcafe4d

Observation 72d2a4e6-3669-4d0b-8395-a636b64f53b1 · outbound

This paper cites CodeT5+: Open Code Large Language Models for Code Understanding and Generation.

Empirical Evaluation of Large Language Models in Automated Program Repair CodeT5+: Open Code Large Language Models for Code Understanding and Generation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:01.938924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:01.938924Z digest=sha256:399b399fae7c0e1e07ed1a53c8e6e2b337ac2fd6c5ea38b0d03c5591642b6294

Observation a14e427d-ca39-4cc4-aa8d-5d032d4c6a55 · outbound

This paper cites Few-shot training llms for project-specific code-summarization,.

Empirical Evaluation of Large Language Models in Automated Program Repair Few-shot training llms for project-specific code-summarization,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.948557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:02.070297Z digest=sha256:f760cd3ed80a7f1835c867d1003dfdcba55a9b1eea326b239337ef6f33ca8abc

Observation 11988c79-de0a-4047-b547-158d6c01a4cf · outbound

This paper cites Can openai’s codex fix bugs? an evaluation on quixbugs,.

Empirical Evaluation of Large Language Models in Automated Program Repair Can openai’s codex fix bugs? an evaluation on quixbugs,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.938979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:02.236582Z digest=sha256:0638d2d211641a12c9b43247fe9ec5f4a0d076fa00dbed342cbafbffdfa4d7ce

Observation 71212d70-bae2-4a25-a985-2dae3756a3e8 · outbound

This paper cites Impact of code language models on automated program repair,.

Empirical Evaluation of Large Language Models in Automated Program Repair Impact of code language models on automated program repair,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.929747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:02.352819Z digest=sha256:81e1e4bbf8faecd15946d17747dce76253e886df188e2cc1911d59b72e9b34ca

Observation a5dc196a-4e68-4f5f-b48d-92ac96460765 · outbound

This paper cites Automated program repair in the era of large pre-trained language models,.

Empirical Evaluation of Large Language Models in Automated Program Repair Automated program repair in the era of large pre-trained language models,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.920026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:02.553542Z digest=sha256:065dd7965a643caa1e9a2f2ec4bf6ffe5962345e6b9893dd034066e00022e21a

Observation b6741863-1deb-43b7-90ac-7eb0d1fbccaa · outbound

This paper cites Automated repair of programs from large language models,.

Empirical Evaluation of Large Language Models in Automated Program Repair Automated repair of programs from large language models,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.909849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:02.700004Z digest=sha256:4b8c2c839a7a29ff522e0272bc393e6463a6806dda309eef0c39c2a65a9a0845

Observation 96adef9e-365e-494a-b863-4b0c845df6ea · outbound

This paper cites Gamma: Revisiting template-based automated program repair via mask prediction,.

Empirical Evaluation of Large Language Models in Automated Program Repair Gamma: Revisiting template-based automated program repair via mask prediction,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.899311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:02.816392Z digest=sha256:d20de7f603307eabeeb944bbf0816a46619be63f1d27d41e75962414a8e76c19

Observation ce07b5f5-d85c-4cb4-8702-e0107a8877f3 · outbound

This paper cites Thinkrepair: Self-directed automated program repair,.

Empirical Evaluation of Large Language Models in Automated Program Repair Thinkrepair: Self-directed automated program repair,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.888905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:02.981678Z digest=sha256:b9aa73e9bb4e5add36533348d4e890df063853ddb133e7815ae8b919f85e251a

Observation d77097f1-f099-446a-84a6-34e63dfd4890 · outbound

This paper cites Keep the Conversation Going: Fixing 162 out of 337 bugs for $0.42 each using ChatGPT.

Empirical Evaluation of Large Language Models in Automated Program Repair Keep the Conversation Going: Fixing 162 out of 337 bugs for $0.42 each using ChatGPT

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:03.138296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:03.138296Z digest=sha256:6521edccc114b40d46196083610f2b10d66a6735c66dfb59bbcdf6fccc09e3ba

Observation 98fd2f8f-62b1-4077-a1e6-4354f99d6d4e · outbound

This paper cites An empirical study on fine-tuning large language models of code for automated program repair,.

Empirical Evaluation of Large Language Models in Automated Program Repair An empirical study on fine-tuning large language models of code for automated program repair,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.879214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:03.322281Z digest=sha256:8eaa5528c5f2520763546f7ba248218b127c279918205318ecb2392676c33071

Observation 21468d90-0805-48e3-aec4-fca00d321f3b · outbound

This paper cites How Far Can We Go with Practical Function-Level Program Repair?.

Empirical Evaluation of Large Language Models in Automated Program Repair How Far Can We Go with Practical Function-Level Program Repair?

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:03.480214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:03.480214Z digest=sha256:736d201ec67ecbf85574d6b8d722901a5157c59ac12ae773445a8f6806d2c6b2

Observation b0f7b167-a413-410c-ad5b-17c321a70e03 · outbound

This paper cites CodeBERT: A Pre-Trained Model for Programming and Natural Languages.

Empirical Evaluation of Large Language Models in Automated Program Repair CodeBERT: A Pre-Trained Model for Programming and Natural Languages

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:03.627340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:03.627340Z digest=sha256:f10127b679ac4500d0209274a808206f88eb9a1b347772a88595d332fd38c2cd

Observation 1afb6817-ed90-44b2-8915-55e1e791a7f3 · outbound

This paper cites CodeT5: Identifier-aware Unified Pre-trained Encoder-Decoder Models for Code Understanding and Generation.

Empirical Evaluation of Large Language Models in Automated Program Repair CodeT5: Identifier-aware Unified Pre-trained Encoder-Decoder Models for Code Understanding and Generation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:03.731652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:03.731652Z digest=sha256:232764367afb1ed8330058d9927e452157a648a5761f5215a3a70204e70d26eb

Observation 758c4715-52aa-4cd4-8b03-1727af3945f7 · outbound

This paper cites Defects4j: A database of existing faults to enable controlled testing studies for java programs,.

Empirical Evaluation of Large Language Models in Automated Program Repair Defects4j: A database of existing faults to enable controlled testing studies for java programs,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.868542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:03.832274Z digest=sha256:9e8614021870551abfbd12a45f4191fbe793e1a1849ed4045e2b6df711070cd7

Observation c1fc1f57-f631-4ef0-b5f2-d8b7186e9d57 · outbound

This paper cites The manybugs and introclass benchmarks for automated repair of c programs,.

Empirical Evaluation of Large Language Models in Automated Program Repair The manybugs and introclass benchmarks for automated repair of c programs,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:03.902785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:03.902785Z digest=sha256:4c73b1c5c6db5b1d4e6bc946563bef75b9fa6b00aea19c7fa67f2004b5746e51

Observation 26afc1d8-52b2-4708-8300-61f244381d91 · outbound

This paper cites Quixbugs: A multi- lingual program repair benchmark set based on the quixey challenge,.

Empirical Evaluation of Large Language Models in Automated Program Repair Quixbugs: A multi- lingual program repair benchmark set based on the quixey challenge,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.853231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:04.032946Z digest=sha256:45739a3a465085b526dc5f52d20cce21277d7b399563ac09b87c28f86dda6c9c

Observation 49dfc44b-32a7-4799-bdc1-6ade43900b54 · outbound

This paper cites An overview of large ai models and their applications,.

Empirical Evaluation of Large Language Models in Automated Program Repair An overview of large ai models and their applications,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.844090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:04.114251Z digest=sha256:8abf0d7a2f682444a191751279b136fe35babe42a7fda909fe16123dd71eee63

Observation f8eff1b3-e612-4aba-8e72-316a59496d27 · outbound

This paper cites The Cost of Training NLP Models: A Concise Overview.

Empirical Evaluation of Large Language Models in Automated Program Repair The Cost of Training NLP Models: A Concise Overview

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:04.244106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:04.244106Z digest=sha256:959582037031283c717cb665843949030788e66b2bdbcfaea8b655992b0025d9

Observation f86b1894-1265-47fa-8cb9-f80bd7d7d7b4 · outbound

This paper cites Code Llama: Open Foundation Models for Code.

Empirical Evaluation of Large Language Models in Automated Program Repair Code Llama: Open Foundation Models for Code

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:04.329887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:04.329887Z digest=sha256:077d30364fbb7623ba0c8680839edb872def5089402293d36986bfd6f8e616dc

Observation 7accd6b2-a4e5-4940-b73c-383e0f4eeffc · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Empirical Evaluation of Large Language Models in Automated Program Repair Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:04.420135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:04.420135Z digest=sha256:2255bb1f6ae046162ebccb6168b104d4aba9072576cb6f023caf485d8492c379

Observation 24a46f90-1e97-467e-88ba-1ade845c20f6 · outbound

This paper cites StarCoder: may the source be with you!.

Empirical Evaluation of Large Language Models in Automated Program Repair StarCoder: may the source be with you!

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:04.536895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:04.536895Z digest=sha256:65f8516bb8cb45da1f22a97ca7da3770dade40882c734b57601985b5c60d6204

Observation 72fb79e8-f04b-431e-a3ca-aa36c8637845 · outbound

This paper cites DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence.

Empirical Evaluation of Large Language Models in Automated Program Repair DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:04.649827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:04.649827Z digest=sha256:e205c31c67b700df3d09d77962e65b8e61daa9d083bec80fd0d3207315438c3e

Observation 7f35cd4c-3c94-4e94-94f1-96e2ebdd80f4 · outbound

This paper cites Bugsc++: A highly usable real world defect benchmark for c/c++,.

Empirical Evaluation of Large Language Models in Automated Program Repair Bugsc++: A highly usable real world defect benchmark for c/c++,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.834621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:04.737314Z digest=sha256:f2113a183a4d8edae0c010efc6b625237b2370adb1441dc521284923365b2a63

Observation 16bcf1d0-f2e3-4c55-9c95-50eea4e4a965 · outbound

This paper cites Introclassjava: A benchmark of 297 small and buggy java programs,.

Empirical Evaluation of Large Language Models in Automated Program Repair Introclassjava: A benchmark of 297 small and buggy java programs,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.825380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:04.810975Z digest=sha256:4407a0a67901041645981e4b2abfab58e5f443352eb6491784d746ee9bf0d8b0

Observation 85559a4d-8504-45af-a6eb-cbf78c345fe6 · outbound

This paper cites ConDefects: A New Dataset to Address the Data Leakage Concern for LLM-based Fault Localization and Program Repair.

Empirical Evaluation of Large Language Models in Automated Program Repair ConDefects: A New Dataset to Address the Data Leakage Concern for LLM-based Fault Localization and Program Repair

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:04.948194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:04.948194Z digest=sha256:9c5375ffb71cf17596d2de0d2414cb2dbe4feafd1a5fc50d87902d1a1923b7ba

Observation 8b11fc15-23a8-40b4-a76a-184428472385 · outbound

This paper cites Attention is all you need,.

Empirical Evaluation of Large Language Models in Automated Program Repair Attention is all you need,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:05.051611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:05.051611Z digest=sha256:10a3ff7d86cae0f3a6da265005577f85b19398b2f496652416b0372392079e5e

Observation 94df2463-7a28-411c-816f-51d15f78510f · outbound

This paper cites Scaling Laws for Neural Language Models.

Empirical Evaluation of Large Language Models in Automated Program Repair Scaling Laws for Neural Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:05.171629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:05.171629Z digest=sha256:d9cec9f491498de10d4145dba779861650f5501a339fa6e608a3663294cf47d6

Observation 20d51a2a-2b9f-41d1-a7e5-9320708393e1 · outbound

This paper cites How to fine-tune bert for text classification?.

Empirical Evaluation of Large Language Models in Automated Program Repair How to fine-tune bert for text classification?

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.809993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:05.216894Z digest=sha256:715281394bc8efa106bae4961b249f84785e1f3b2b7e087a53540141f9b3c96b

Observation 395ba91c-dd99-45f1-a74b-ac54d3b2a4a9 · outbound

This paper cites Parameter-efficient fine-tuning of large-scale pre-trained language models,.

Empirical Evaluation of Large Language Models in Automated Program Repair Parameter-efficient fine-tuning of large-scale pre-trained language models,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.800845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:05.248130Z digest=sha256:3a07319ea03a8f8ff70eefd4fde4b727f8a1150cefc0f8c7bde9c3a792050b79

Observation 6a83e52f-062a-4de2-a231-7cfe1cd31a74 · outbound

This paper cites The Power of Scale for Parameter-Efficient Prompt Tuning.

Empirical Evaluation of Large Language Models in Automated Program Repair The Power of Scale for Parameter-Efficient Prompt Tuning

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:05.362416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:05.362416Z digest=sha256:0876554a2e4a9b8769342a8de3cce0a8ad9b83328dada9b20bd5a2118d9bdbab

Observation 6fb7979d-aa5e-4487-87df-695dc7e371e8 · outbound

This paper cites Visual prompt tuning,.

Empirical Evaluation of Large Language Models in Automated Program Repair Visual prompt tuning,

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:05.421426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:05.421426Z digest=sha256:9e018ff05d6c9f61f2ee8c26296467efe460ceee7636d2c8e5fba77bad20df0b

Observation 5b3dc69d-4f59-42dd-94e0-98c7e9cde194 · outbound

This paper cites RepairLLaMA: Efficient Representations and Fine-Tuned Adapters for Program Repair.

Empirical Evaluation of Large Language Models in Automated Program Repair RepairLLaMA: Efficient Representations and Fine-Tuned Adapters for Program Repair

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:05.424919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:05.424919Z digest=sha256:b45872f67f8a75c27443b26058c863a9bb47fd1ab44964302177eb6d81f3c1bd

Observation 0e006f2e-67e4-42de-9531-2a63aae568e5 · outbound

This paper cites Lora: Low-rank adaptation of large language models.

Empirical Evaluation of Large Language Models in Automated Program Repair Lora: Low-rank adaptation of large language models

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:05.428384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:05.428384Z digest=sha256:e1e54a2ed1f02a60c22d10e8bb09a25214ffed3ffd6e14a18c996d5fcfaeb18d

Observation 0aa587ea-34f5-40c7-ad32-a661661efd1d · outbound

This paper cites Language models are few-shot learners,.

Empirical Evaluation of Large Language Models in Automated Program Repair Language models are few-shot learners,

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.778200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:05.431659Z digest=sha256:5da088f46c5f05a3b6edba732509abcfee5ddf5e440ad369e4095e1884564a89

Observation a15d5cee-f17f-48b3-a7f5-87d85476bd06 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models,.

Empirical Evaluation of Large Language Models in Automated Program Repair Chain-of-thought prompting elicits reasoning in large language models,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.768606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:05.435339Z digest=sha256:a9345dc72ab2da60c73fb996f6f7faf7a411657a9ffcbdf6c6519cb8fa3437da

Observation 7e50a810-d84d-4cd0-8eb6-d0b0ab34c358 · outbound

This paper cites Automatic Chain of Thought Prompting in Large Language Models.

Empirical Evaluation of Large Language Models in Automated Program Repair Automatic Chain of Thought Prompting in Large Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:05.441867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:05.441867Z digest=sha256:886776f3370ba4bdbfd925510be1a35fdb2d281719f00575ca523476b3e2b3c2

Observation b6094612-9c13-4d54-b8cd-ceddc3b278a3 · outbound

This paper cites Towards understanding chain-of-thought prompting: An empirical study of what matters,.

Empirical Evaluation of Large Language Models in Automated Program Repair Towards understanding chain-of-thought prompting: An empirical study of what matters,

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.750080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:05.445261Z digest=sha256:d0606aefa29f5ad080e749309af17e58c66ec0825b9dd4d011fe06f9cc922299

Observation ad81bf8e-c8f9-48a0-ba15-8930e9b2725f · outbound

This paper cites Copiloting the copilots: Fusing large language models with completion engines for automated program repair,.

Empirical Evaluation of Large Language Models in Automated Program Repair Copiloting the copilots: Fusing large language models with completion engines for automated program repair,

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.740858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:05.447945Z digest=sha256:aa4907416e77a615d805484d1220c38e4957561ac49313561ca4ee6f7c1b044b

Observation cc987921-c713-4283-8f42-bf026044ea93 · outbound

This paper cites Language models are few-shot learners,.

Empirical Evaluation of Large Language Models in Automated Program Repair Language models are few-shot learners,

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.730831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:05.451148Z digest=sha256:51b54761fa072b4b99ba378b41c3148f96856a10a724c95fb16de150e6fe6ba8

Observation e44f86b7-e733-4626-839b-b8b9fde34cc9 · outbound

This paper cites Hybrid automated program repair by combining large language models and program analysis,.

Empirical Evaluation of Large Language Models in Automated Program Repair Hybrid automated program repair by combining large language models and program analysis,

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.715419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:05.457284Z digest=sha256:534537cfa0305892d704fed443755a27f838b54c8e5e9c88cb4e83af00c3116c

Observation a3686d76-ead4-4241-b460-ee9eebcd86b5 · outbound

This paper cites Atcoder,.

Empirical Evaluation of Large Language Models in Automated Program Repair Atcoder,

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.703761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:05.460165Z digest=sha256:97ccf08ac573cd92f747fc497067f70ce4e259cdae001c88c88c823af3e01217

Observation 55508178-2240-479c-b518-8719019d69cd · outbound

This paper cites Benchmarking automated program repair: An extensive study on both real-world and artificial bugs,.

Empirical Evaluation of Large Language Models in Automated Program Repair Benchmarking automated program repair: An extensive study on both real-world and artificial bugs,

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.694450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:05.463552Z digest=sha256:26adb5426de27eeeeaa60b9f93a3daa4399ecaaee32c639b17790744c43a4451

Observation 3742ddc5-6f61-4290-956f-49029b7ae247 · outbound

This paper cites A large-scale empirical review of patch correctness checking approaches,.

Empirical Evaluation of Large Language Models in Automated Program Repair A large-scale empirical review of patch correctness checking approaches,

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.684923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:05.466363Z digest=sha256:04a3d36242eac95049ea3a478fef3455acd88c36ab39f0b9018c4b8a991d38ab

Observation 22b224c3-c84b-4ba9-be58-302809035acd · outbound

This paper cites Fine-grained and accurate source code differencing,.

Empirical Evaluation of Large Language Models in Automated Program Repair Fine-grained and accurate source code differencing,

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.675707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:05.469686Z digest=sha256:7e989d61ddbf5e44a04881abd76f7296656bdd3182a98744135c4d5a5d287ae6

Observation 31906295-cc80-4353-beac-7113e8e9715b · outbound

This paper cites Hyperparameter optimiza- tion for ast differencing,.

Empirical Evaluation of Large Language Models in Automated Program Repair Hyperparameter optimiza- tion for ast differencing,

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.665562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:05.472880Z digest=sha256:b8fa16e5567169211d75c1ad8e740cc66df75f062c0b9016bbbf996719db425e

Observation e5d92750-57d9-4d2c-a8aa-24c3189299a2 · outbound

This paper cites Available: https://proceedings.neurips.cc/paper files/ paper/2020/file/1457c0d6bfcb4967418bfb8ac142f64a-Paper.pdf.

Empirical Evaluation of Large Language Models in Automated Program Repair Available: https://proceedings.neurips.cc/paper files/ paper/2020/file/1457c0d6bfcb4967418bfb8ac142f64a-Paper.pdf

Reference 1901

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:05.454169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:05.454169Z digest=sha256:57f9eb0a4c8e4a9334c10754f4ef90a72ba93d788e3e36603e0e15a8c8206ea3

Observation 59f72b53-4217-48de-9a4c-7f0376aba95d · outbound

This paper cites Available: https://proceedings.neurips.cc/paper/2022/ hash/9d5609613524ecf4f15af0f7b31abca4-Abstract-Conference.html.

Empirical Evaluation of Large Language Models in Automated Program Repair Available: https://proceedings.neurips.cc/paper/2022/ hash/9d5609613524ecf4f15af0f7b31abca4-Abstract-Conference.html

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.758987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:42:05.438664Z digest=sha256:0188a8575228bb81c9259b526fa3df3de9e0fdf2fc0dc2213d4e7be782a788c5

Pith citing papers

Observation 6827e5a6-d000-403b-bcf7-ea2837457152 · inbound

Automating Computational Reproducibility in Social Science: Comparing Prompt-Based and Agent-Based Approaches cites this paper.

Automating Computational Reproducibility in Social Science: Comparing Prompt-Based and Agent-Based Approaches Empirical Evaluation of Large Language Models in Automated Program Repair

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:57:24.435042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-16T05:55:48.203213Z digest=sha256:95f506d6f92d3bbf8b4c0a8f1b20065fb2aa64d69280dd7a6ae0711bb4a28da9

Observation 431b66bc-afd3-450e-bfd8-2ebc102b0793 · inbound

Semantic Drift in Bug Resolution: How Behavioral Signals Propagate from Reports to Tests and Patches cites this paper.

Semantic Drift in Bug Resolution: How Behavioral Signals Propagate from Reports to Tests and Patches Empirical Evaluation of Large Language Models in Automated Program Repair

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T15:06:55.050376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:06:55.050376Z digest=sha256:e886eca3e26cdd60f8c939e7df004ff3c622440031e7c4df077e4768ced79c69