Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T10:15:37.403984Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 1 inbound Pith citation observation for arXiv:2508.00408.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T10:15:37.403984Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-13T21:36:14.478007Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-13T21:38:18.794785Z
47 of 47 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 909ff480-c21a-4d51-8576-4053065cded6 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions An orchestrated survey of methodologies for automated software test case generation,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ea7abe96-a976-4ac4-9c89-dbba8875d68b · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions A survey on model-based testing tools for test case generation,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 800f75a8-58ec-41cb-b846-af53ddb8ef4b · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions An empirical evaluation of using large language models for automated unit test generation,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 20765719-ae57-483c-b8bc-a4c241f1158b · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Measuring the Influence of Incorrect Code on Test Generation
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29603517-e653-4063-81b1-b10fab08a9be · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions TESTEVAL: Benchmarking Large Language Models for Test Case Generation
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 049e1bf8-c0f5-4740-8a45-f9f4d428a18b · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions TestGenEval: A Real World Unit Test Generation and Test Completion Benchmark
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae226fdf-6f5f-40f1-8101-baf4c422bc97 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions A survey on unit testing practices and problems,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eb0478e7-befc-4ddf-8116-e9c4b62f79d2 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions DeCon: Detecting Incorrect Assertions via Postconditions Generated by a Large Language Model
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0a4d2216-eb5c-4bda-9b3d-eba1c356a96d · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions CodeT: Code Generation with Generated Tests
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d65a213-27c3-4f68-8908-30991d5de08c · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions CodeCoT: Tackling Code Syntax Errors in CoT Reasoning for Code Generation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ae9a8b8-c5cc-405d-a5bb-098ae768dbdd · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions AgentCoder: Multi-Agent-based Code Generation with Iterative Testing and Optimisation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d6aaf3f-2819-4815-bc29-b6e02d53aab1 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Mercury: A Code Efficiency Benchmark for Code Large Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e438847-c64c-46cb-b037-17e03cdcd579 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Reflexion: Language agents with verbal reinforcement learning,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a9d9dcd7-daaf-4541-90a9-e715e922b3fe · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Kernelgpt: Enhanced kernel fuzzing via large language models,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 71555cd2-b041-4621-b1d4-4a882b6319a2 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Universal fuzzing via large language models,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5694b4f1-d54c-491b-9b6a-c62ab0566d60 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Large language models are edge-case generators: Crafting unusual programs for fuzzing deep learning libraries,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3e6756b4-c2fc-46d2-a198-6f38cec3352f · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Whitefox: White-box compiler fuzzing empowered by large language models,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation af101ac6-f2ce-4913-b4ee-00732bf7ef31 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Large language models are zero-shot fuzzers: Fuzzing deep-learning libraries via large language models,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b4f36b78-d7bc-454c-b966-ae36971a537d · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Large Language Models are Edge-Case Fuzzers: Testing Deep Learning Libraries via FuzzGPT
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2705c7b-8835-4789-9913-b08f624d00c1 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Swt-bench: Testing and validating real-world bug-fixes with code agents,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cf2f71a3-4925-48f5-82d3-fce2de10f34b · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions TestBench: Evaluating Class-Level Test Case Generation Capability of Large Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83562be3-cda6-4a35-a212-1d2279ecdaa8 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions SWE-bench: Can Language Models Resolve Real-World GitHub Issues?
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5ec29d1-8d54-4483-b0bf-bfbcd156afee · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions LessLeak-Bench: A First Investigation of Data Leakage in LLMs Across 83 Software Engineering Benchmarks
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff5e4032-eae5-4b3f-b5ab-053d5a41b488 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Large-scale, independent and comprehensive study of the power of llms for test case generation,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02536c0a-22c9-4bfa-bade-7b3fda8814aa · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions StarCoder 2 and The Stack v2: The Next Generation
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41ed3b87-906b-47ac-ad8d-ac26dfdcda2c · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Software engineering (ed.),
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0f116071-e015-4802-95ed-7e87226c6279 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ff03e296-9140-4fed-a2d7-f911343c657d · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions A survey of unit testing practices,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dd1ddfd1-1f53-49bc-a13d-38bbb1428f03 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Meszaros, xUnit test patterns: Refactoring test code
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 924cc713-f094-4d54-8340-4c419023b4d5 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Automated unit test generation for evolving software,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5c3066a3-18ad-4b10-a7b0-67c9f92525d7 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Symbolic execution and program testing,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b384488a-e4aa-44a9-a408-8a17902f543d · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions The s2e platform: Design, implementation, and applications,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d1d81501-34b4-491b-b7f3-9ee0d69f9d12 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Search-based software testing: Past, present and future,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 83da7827-a392-4379-9e6e-6a8fd88a1d87 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions An empirical study of the reliability of unix utilities,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e86156dc-c893-471e-be4e-40841930b032 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Large language models for software engineering: A systematic literature review,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 93e459f2-955d-469c-9a37-1fe3996eea03 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Mutation testing advances: an analysis and survey,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2c7014cb-cf71-48d7-a63b-a9737d135658 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions An analysis and survey of the development of mutation testing,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c2fd3083-4b86-4351-bcac-e386a66b740e · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa05e55a-0e56-4278-a613-8fced4f459b2 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Using large language models to generate junit tests: An empirical study,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6cbb6b53-f7a5-4458-88b7-b559f9465ca1 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Using github copilot for test generation in python: An empirical study,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f0f93134-cd37-4036-be2d-f5ed052d9b88 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Testspark: Intellij idea’s ultimate test generation companion,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bed90727-2ed0-47ba-af57-ab730671c8fc · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Code Generation Tools (Almost) for Free? A Study of Few-Shot, Pre-Trained Language Models on Code
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a015a862-9b21-4eda-8a81-7bf7dcde4a6c · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Retrieval-based prompt selection for code-related few-shot learning,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 85510848-0265-4bfc-be73-f7a94fdf1be4 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Codamosa: Escaping coverage plateaus in test generation with pre-trained large language models,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2b5f2b33-ba1b-4734-b40c-8a316f39248c · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Testart: Improving llm-based unit test via co-evolution of automated generation and repair iteration,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e54c550-ad65-43a7-9ee1-ceffe40b417a · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions ASTER: Natural and Multi-language Unit Test Generation with LLMs
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e0e4403-ed56-48be-aa69-1e2fce5c4585 · outbound
Benchmarking LLMs for Unit Test Generation from Real-World Functions Effective test generation using pre-trained large language models and mutation testing,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3c6eb1fb-54b4-4124-8224-b7e758e4d17a · inbound
TestDecision: Sequential Test Suite Generation via Greedy Optimization and Reinforcement Learning Benchmarking LLMs for Unit Test Generation from Real-World Functions
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.