Pith. sign in

Paper Citation Record · LEDGER

Quality Assessment of Python Tests Generated by Large Language Models

As of 22 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 2 inbound Pith citation observations for arXiv:2506.14297.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.14297 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:22:57.229784Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-08T17:37:51.790000Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T20:46:10.824386Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact5
  • verified fuzzy3
  • unresolved28
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f2ce75d2-4b3d-4acf-8427-82c68f5519f3 · outbound

This paper cites Structured Code Representations Enable Data-Efficient Adaptation of Code Language Models.

Quality Assessment of Python Tests Generated by Large Language Models Structured Code Representations Enable Data-Efficient Adaptation of Code Language Models

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:23:00.522246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T00:22:52.907011Z digest=sha256:12d565f50780350edacdb96059199042a219313c480cebe45a77c218f717d2b1

Observation 2baf4883-dc09-4d31-ad66-46f3a5a069e6 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:52.956816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:52.956816Z digest=sha256:2a94bf5fe0cce432d7771a60482544539135710b60db0bfd4217780b8b58fd32

Observation b74754e8-2605-493f-add0-a45212bee263 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 3

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T00:23:00.320326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T00:22:53.166885Z digest=sha256:3be41eb6c726f27d9279f6c93152617a3b3059928854e94390c367d3c2cd3798

Observation c992e3e7-edcd-42eb-aa68-d3d4ea0763e0 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 5

Resolution
malformed identifier
no resolver link, observed 2026-08-07T00:22:53.383429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:53.383429Z digest=sha256:6a480db4dbb426e1999d8776a5147b56479913bcb870e1eea60bc88958c66cc7

Observation a591b136-a682-4b39-8399-501e867fa2c6 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 6

Resolution
verified exact
doi, observed 2026-08-07T00:22:57.527939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T00:22:53.465341Z digest=sha256:f2fa8b5d1ea856d68fc0dddd1063d025dce9eb9ef46aaaea38b087f33431331c

Observation 3492489c-369f-471f-aa77-e76fca365e09 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

Quality Assessment of Python Tests Generated by Large Language Models Evaluating Large Language Models Trained on Code

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:53.528702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:53.528702Z digest=sha256:720a2e2a3acba823fae984d46a230ecf9414a69d53465987656bad4072483736

Observation 7c700471-b715-4be9-932c-072a26a67464 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:53.620863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:53.620863Z digest=sha256:45f2bfc20d1e9e20a4cb8c1ccf881ebcc4678ad4773423faa996e022d78a987a

Observation f53d4363-7104-4d7a-be28-8c88d3f2c9e3 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:53.777526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:53.777526Z digest=sha256:252eb735e720093d603a0243d0f1d51be3942461786bffc45f72ccad38466098

Observation 3dbac20d-5c75-4c0f-86c4-021eaf678633 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 11

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:22:59.852913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T00:22:53.861129Z digest=sha256:77b8acd1396ec7cf965c9fdb1b5e814394bb12f4853faccc0e8a4c6b7c9b09cf

Observation 05025f6e-0f17-4f8f-9cf2-b90afd9365ab · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 12

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T00:22:59.627204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T00:22:53.926091Z digest=sha256:2b51b9cd6925f5a764a027033ecd8db8532085ce85966844edb80b4eb7dfbd7e

Observation f0773172-c8a0-4b84-ab6c-cc268ddd6c2d · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.000232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.000232Z digest=sha256:b835876a4eb8fe4289298efbcf1d21f3b6d8afd3b1ddbf90f45d18e47990d873

Observation 6f2a4774-4e72-495a-88aa-4d5328114048 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.048447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.048447Z digest=sha256:24747af61f8fd0434f695e0c3197a906c7e8758859ae120fc3f9079df3bab23d

Observation 406b2339-ec23-49fb-ae9d-635d5e903e2f · outbound

This paper cites Santos, Andrew Popovich, Mehdi Mirakhorli, and Mei Nagappan.

Quality Assessment of Python Tests Generated by Large Language Models Santos, Andrew Popovich, Mehdi Mirakhorli, and Mei Nagappan

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.117566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.117566Z digest=sha256:e20a78b06969ead2330bda63e0bd0231c835823c2789f19d0183183bc64ff193

Observation 652449a9-dd3d-4787-b770-88486c2f2ec8 · outbound

This paper cites Graham, R.

Quality Assessment of Python Tests Generated by Large Language Models Graham, R

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:23:01.553263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T00:22:54.188448Z digest=sha256:37b8502b116d966106e3d0ef073b55ed97046a8f1b80af76e304bc224d1044be

Observation 4d2b26f3-3ec7-4b32-8016-91961208dc66 · outbound

This paper cites 2023.Code Correctness and Quality in the Era of AI Code Generation: Examining ChatGPT and GitHub Copilot.

Quality Assessment of Python Tests Generated by Large Language Models 2023.Code Correctness and Quality in the Era of AI Code Generation: Examining ChatGPT and GitHub Copilot

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:23:01.495272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T00:22:54.285734Z digest=sha256:291ca6bb04cf88d7f64c823c2c372026e0f0a9aa9f202609660b639279f55d84

Observation 714c059f-df9c-4ebc-82a6-8e96fafcbce3 · outbound

This paper cites Khorikov.

Quality Assessment of Python Tests Generated by Large Language Models Khorikov

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:23:01.388956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T00:22:54.290955Z digest=sha256:78846a6b952709365df4f09e0740e48c31706346c6907b5faa821204dc5f577c

Observation 30ff293e-bddf-4993-aa49-e27b5f4f3b5f · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:01.267075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T00:22:54.366203Z digest=sha256:72351dd757572a7489d984082475dad3a7a580dc0c9cd2120c3180fed8e23d6e

Observation 733ac602-bcdc-44aa-8cb6-c23bccdb624f · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.510354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.510354Z digest=sha256:dd1da902ef9e96bf5c6054a467260858f15ae14fc12fcb86db962dd01827454c

Observation 7b5b833e-b19b-4dc0-a875-998f9e59c5a5 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.593557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.593557Z digest=sha256:bf07ea7b61c8e1e20682543e540e15f5f84f6ac890c055c9a016315f8b03549c

Observation e591b293-72b1-4234-9c51-0df5c84c48a8 · outbound

This paper cites CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation.

Quality Assessment of Python Tests Generated by Large Language Models CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.681694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.681694Z digest=sha256:42b06673943ae3fdef9f8a85c16dccb9dd9e93f8a88c25bacf242a730c5355d2

Observation ac2fe842-4109-4667-8115-26f6bad1abd4 · outbound

This paper cites Search-based software test data generation using evolutionary computation.

Quality Assessment of Python Tests Generated by Large Language Models Search-based software test data generation using evolutionary computation

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:22:59.126324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T00:22:54.843366Z digest=sha256:791320cf26799eb48c11d1d746b95d5b10929bf4859aadb3b90787a0e41e5b40

Observation d7e546bd-95d7-4e47-9ef4-c4df30528740 · outbound

This paper cites Marvin, N.

Quality Assessment of Python Tests Generated by Large Language Models Marvin, N

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.900003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.900003Z digest=sha256:3ce4015d5504365323b02a7ab2a54deecef7db835814778e0f4f29e00dca3296

Observation 7f82476d-740b-4305-80da-8fb4cd17db78 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:01.190255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T00:22:55.164951Z digest=sha256:f2e8003aa83ccf9469fadc622c68c2ccc0564c32a0a22a8459fa170efb0fc317

Observation 2eae1305-74c4-4636-a1fb-1b625e8ab18a · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 27

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:22:58.832867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T00:22:55.292198Z digest=sha256:ccde1bd548a432b7525379a53a5a7b91b146ed824149e202ef952e31ef6b57f4

Observation 11045527-012f-422f-8f98-653f9768ffc2 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.412598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.412598Z digest=sha256:dc46b34c6bc0cca9f77e54acce1d9c8fcb9fb2529ff506f10becc94622ba84e4

Observation 45045dd9-da6f-4acb-b9dd-f5b383544d3f · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.499037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.499037Z digest=sha256:4f44e4dcf48c466ad2c13a103fd24ff0cb0368c498b92f36cd21b5d0e3222d08

Observation ef0d7043-afe7-4524-90f7-08061ce8e248 · outbound

This paper cites Becker, Arto Hellas, Bailey Kimmel, Garrett Powell, and Juho Leinonen.

Quality Assessment of Python Tests Generated by Large Language Models Becker, Arto Hellas, Bailey Kimmel, Garrett Powell, and Juho Leinonen

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.606582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.606582Z digest=sha256:9b2590c83dc8cc6e3dda5860a66b90c1cb49386ceb5180bf9c32aba252980a05

Observation d8e185d7-c9ad-4032-97ed-2d9d2b5b99b6 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.733894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.733894Z digest=sha256:7329cf1b349ce2f0ee8090cc3054e60df4ccc5c9283ad82b6d818d665824461a

Observation b5459b6b-81ea-482d-94b1-af2126f754a0 · outbound

This paper cites A Systematic Survey of Prompt Engineering in Large Language Models: Techniques and Applications.

Quality Assessment of Python Tests Generated by Large Language Models A Systematic Survey of Prompt Engineering in Large Language Models: Techniques and Applications

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.803426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.803426Z digest=sha256:a1cca70e62097317809ef7bd713dcb81ca20285646d4ab661deec41226bfa429

Observation 43a2ce56-4b73-4c8c-b842-8b7dc576f190 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:01.023171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T00:22:55.889640Z digest=sha256:0111d875b75ba3a19d6278814dc22fdc5d69352c35c6fde1b88c38454fb46c5b

Observation 8b96de58-5399-4419-ad44-ca45813a2857 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.959286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.959286Z digest=sha256:050a51661342a8ee1f99f8d6847163f953feb6d2bda4ec85dde99c70057548f5

Observation b7803329-62f0-4a4b-871e-aff5426195d1 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 35

Resolution
malformed identifier
doi_truncated, observed 2026-08-07T00:22:58.159750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T00:22:56.099483Z digest=sha256:9d6e7222ef81914ea8c75a927589c7c5ac859457571ac3c5e823b34f15fa95b5

Observation f630d06a-d59c-4f07-8e21-b1f445bbf636 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 36

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T00:23:00.129940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T00:22:56.206882Z digest=sha256:8c1546aeedd431f8269215fba5a0fb9c7f0cc2af0107384c5928c3d7ab91bba7

Observation a362e75d-eb7b-4d96-a0fe-f4c181f31073 · outbound

This paper cites Unit Test Case Generation with Transformers and Focal Context.

Quality Assessment of Python Tests Generated by Large Language Models Unit Test Case Generation with Transformers and Focal Context

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:56.344629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:56.344629Z digest=sha256:cd8d4610e117dbca6daf0accca6d6c7acd0520bbd9fa098703d7c1fc4eab978b

Observation b5f5786d-522c-4a62-917c-d3003a413ac0 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:56.473597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:56.473597Z digest=sha256:3d8c51d97f72a878f4ce570420afb1ed32d6fde176d85eb4a7353fd1e51637bd

Observation f45fb04c-3f11-425c-ac28-c4f22366833d · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:00.923542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T00:22:56.564574Z digest=sha256:9020e60eec387be1c722009db800de3d43577dc9679641134ff30899f150c211

Observation 8f52b362-fb2f-43c2-9899-2532c4bca62b · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 40

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T00:22:57.843768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T00:22:56.694578Z digest=sha256:fa33942abeca034d2a1ddd9cb8e20a8784758cde22616ce545342d52bc8db97a

Observation 4a8b768d-a97c-4a77-8ef5-bcbe7c4bc016 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:56.791389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:56.791389Z digest=sha256:7b889a1af167c579986c3770749cf5f0cf7e01fcc0ed7af4ad7fd9108bd38685

Observation 4a1f0b84-16fa-41e4-bf57-aae274d446a0 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:00.781962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T00:22:56.914404Z digest=sha256:e0cc48bacec2d8ae54f53eb03655cf59d4a40c658f56734290ae3aba424577f6

Observation 20852e48-d589-4406-8d08-7d1439c0ea86 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:57.025241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:57.025241Z digest=sha256:aa1ccb56bb7567de15a9a8215d0b484659c48892e2f2f470a1ad4e80645b6f7d

Observation 9da3e49d-a12e-44d9-b41c-d7f49378169f · outbound

This paper cites Evaluating the Code Quality of AI-Assisted Code Generation Tools: An Empirical Study on GitHub Copilot, Amazon CodeWhisperer, and ChatGPT.

Quality Assessment of Python Tests Generated by Large Language Models Evaluating the Code Quality of AI-Assisted Code Generation Tools: An Empirical Study on GitHub Copilot, Amazon CodeWhisperer, and ChatGPT

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:57.143131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:57.143131Z digest=sha256:69141d1004fd0c89bd528bf40706fba85531553249f2d9c6c9c78fdf8cc1b036

Observation 32ddb553-d030-44a7-b15e-9ce263555191 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:00.629614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T00:22:57.229784Z digest=sha256:002787aa5332651ca1b647a81553ccb2fa888665a8c866a432d6eeb0bd0fdd0f

Pith citing papers

Observation 039d6874-8fbe-4269-9744-543facd57397 · inbound

Can LLMs be Effective Code Contributors? A Study on Open-source Projects cites this paper.

Can LLMs be Effective Code Contributors? A Study on Open-source Projects Quality Assessment of Python Tests Generated by Large Language Models

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:46:10.830944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-08T08:09:33.211692Z digest=sha256:e0861276291f02bdb2cf40ec8e381383350d6ff38c89b89b0aedba06207c2697

Observation a220b4a9-3197-41ba-95b2-36e94fffb31a · inbound

Bridging Generation and Training: A Systematic Review of Quality Issues in LLMs for Code cites this paper.

Bridging Generation and Training: A Systematic Review of Quality Issues in LLMs for Code Quality Assessment of Python Tests Generated by Large Language Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:26:04.393629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-08T17:37:51.790000Z digest=sha256:e3c614a6bd0498cb705e7c567c965f7fc79b24f84a63bbc4fe2b434b5e0e0c7d