Pith. sign in

Paper Citation Record · LEDGER

RLTF: Reinforcement Learning from Unit Test Feedback

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2307.04349.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.04349 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:52:20.405213Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T03:24:12.929087Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2b093690-3f8b-4070-b3d9-5140612b6408 · inbound

A Survey on Large Language Models for Code Generation cites this paper.

A Survey on Large Language Models for Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 167

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:18:06.419569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T20:18:06.304134Z digest=sha256:d98cf7e9d2ad94f481d703c591506e935c66a365e298fa7e3264b9c2ace7e6d3

Observation 410df87a-1370-44be-a2a7-a39f93b1637b · inbound

Dr. Boot: Bootstrapping Program Synthesis Language Models to Perform Repairing cites this paper.

Dr. Boot: Bootstrapping Program Synthesis Language Models to Perform Repairing RLTF: Reinforcement Learning from Unit Test Feedback

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T15:52:20.405213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:52:20.405213Z digest=sha256:0e990390af7d54d4441c66c27bd14560e21ca566c8db44d598d1af58a24b470b

Observation 1d7ba4df-15ae-416f-baeb-94ba70bf9c15 · inbound

BitsAI-Fix: LLM-Driven Approach for Automated Lint Error Resolution in Practice cites this paper.

BitsAI-Fix: LLM-Driven Approach for Automated Lint Error Resolution in Practice RLTF: Reinforcement Learning from Unit Test Feedback

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T04:31:01.806003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:31:01.806003Z digest=sha256:5c8d9cf767d694bd01cf305878ac36adc73ac838c3750b74864862b68c1f7be9

Observation ba9eba82-17ef-4d24-bc79-9752b4484a38 · inbound

Efficiency of turbulence cites this paper.

Efficiency of turbulence RLTF: Reinforcement Learning from Unit Test Feedback

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T00:58:28.756567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:58:28.756567Z digest=sha256:3194628cc3c57f401c70bec0d24180a410283ad99783728a690d337bbe866dee

Observation b94c0dae-105f-471e-bd15-ed0e24365f63 · inbound

InfoSynth: Information-Guided Benchmark Synthesis for LLMs cites this paper.

InfoSynth: Information-Guided Benchmark Synthesis for LLMs RLTF: Reinforcement Learning from Unit Test Feedback

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T13:06:57.310399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:06:57.310399Z digest=sha256:5d7a0221d41978c9e41695916ca6c774b9764c923167837f4994120a547b69d0

Observation 89d4fab7-2e22-4c66-b56d-19dfef7c43d8 · inbound

Beyond Binary: Turning Partial Success into Dense Verifiable Rewards for Reinforcement Learning in Code Generation cites this paper.

Beyond Binary: Turning Partial Success into Dense Verifiable Rewards for Reinforcement Learning in Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T12:19:32.688253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T12:19:32.688253Z digest=sha256:43849645c4d4bffc52ea4f234cd523813c68262214b46096609a8abbbcc01638

Observation f5ca73ae-328a-4a89-acac-7700313e646d · inbound

An Iterative Test-and-Repair Framework for Competitive Code Generation cites this paper.

An Iterative Test-and-Repair Framework for Competitive Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:35:48.667824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T19:44:32.950977Z digest=sha256:fdcb7ada8c6664787012d04f2f34579b23b10d65536bde48ce1a4bc20c212866

Observation d0c75916-e9c6-455f-b7d1-672ee32b87d1 · inbound

An Iterative Test-and-Repair Framework for Competitive Code Generation cites this paper.

An Iterative Test-and-Repair Framework for Competitive Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-13T09:22:44.557413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T09:22:44.557413Z digest=sha256:fc1e80230166ebc89dc7a316f968faa40bb5770e4d28f21fef530618bf793d55

Observation 4f369d24-d0c0-4a41-95d2-97c4d85e1afe · inbound

Generate, Filter, Control, Replay: A Comprehensive Survey of Rollout Strategies for LLM Reinforcement Learning cites this paper.

Generate, Filter, Control, Replay: A Comprehensive Survey of Rollout Strategies for LLM Reinforcement Learning RLTF: Reinforcement Learning from Unit Test Feedback

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:15:49.220299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T19:15:27.406778Z digest=sha256:0cf602da40aa44ecfaa06c71af2c97fa9757e5e20e796d6aab283815402977e7

Observation 3ee44391-847f-4928-965a-a1d915aefd0e · inbound

ShapeCodeBench: A Renewable Benchmark for Perception-to-Program Reconstruction of Synthetic Shape Scenes cites this paper.

ShapeCodeBench: A Renewable Benchmark for Perception-to-Program Reconstruction of Synthetic Shape Scenes RLTF: Reinforcement Learning from Unit Test Feedback

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:22:02.191104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T01:19:04.708277Z digest=sha256:0baad4c3783a0b51ef12ecf376114eae30a1a3d9f82c369f59d24d3c3eb1d273

Observation 28e01ef3-ae78-48b6-b3ec-8b69ab5954d6 · inbound

Code as Agent Harness cites this paper.

Code as Agent Harness RLTF: Reinforcement Learning from Unit Test Feedback

Reference 102

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:58:14.222864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T10:54:54.558241Z digest=sha256:55e072a1a7cc82563b8da3e3eb265a7977053e5c0e72aca0c685049e88813ba1

Observation 60dda99f-c8f2-44f6-b592-999725e379dd · inbound

Building to the Test: Coding Agents Deliver What You Check, Not What You Requested cites this paper.

Building to the Test: Coding Agents Deliver What You Check, Not What You Requested RLTF: Reinforcement Learning from Unit Test Feedback

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-30T03:24:12.930587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T01:40:24.510828Z digest=sha256:fcba8a7bfa714688fb8b6997a1a95ceadc8192bf2a6ad3576e078972e6812ae1

Observation 512f06d4-a10c-4f0c-bb94-79332a221459 · inbound

Beyond the Need for Speed: Energy-Aware Code Generation via Simulation-Guided Reinforcement Learning cites this paper.

Beyond the Need for Speed: Energy-Aware Code Generation via Simulation-Guided Reinforcement Learning RLTF: Reinforcement Learning from Unit Test Feedback

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-11T17:00:48.664985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T17:00:48.664985Z digest=sha256:f6ae90d09ae3f4716d9cca81cb6ebd80aea25a479455c107227666516a380bc0

Observation 59afdc66-0d0a-4922-b340-a70b187a9293 · inbound

RLPF: Reinforcement Learning from Performance Feedback for Code Generation cites this paper.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.905721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.905721Z digest=sha256:4cdec977d526f062b592935663658f0270c7561d90bd7b20bf8fbd2cf67d949f