Pith. sign in

Paper Citation Record · LEDGER

RLTF: Reinforcement Learning from Unit Test Feedback

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2307.04349.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.04349 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T15:10:38.868972Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T03:24:12.929087Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2b093690-3f8b-4070-b3d9-5140612b6408 · inbound

A Survey on Large Language Models for Code Generation cites this paper.

A Survey on Large Language Models for Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 167

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:18:06.419569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T20:18:06.304134Z digest=sha256:f4f7f41faf00cec852d432ad16960fb235defc1296f579dbcc7c809eb3d8db23

Observation 399b8bb8-37a4-4c0a-aa2f-376e220713c2 · inbound

Process-Supervised Reinforcement Learning for Code Generation cites this paper.

Process-Supervised Reinforcement Learning for Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T15:10:38.868972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T15:10:38.868972Z digest=sha256:fc63f0a8ce9d0341b3b5afc3718d3271aabdbd46468fb95a7d1d2f200cb06bc3

Observation 01e10bc9-eb54-4dab-ac59-af1746945dcc · inbound

e-SimFT: Alignment of Generative Models with Simulation Feedback for Pareto-Front Design Exploration cites this paper.

e-SimFT: Alignment of Generative Models with Simulation Feedback for Pareto-Front Design Exploration RLTF: Reinforcement Learning from Unit Test Feedback

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T12:11:21.953185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T12:11:21.953185Z digest=sha256:d377397fd4ffe5c7037b5007215de2118971038f15b9235006554069f451bd5f

Observation 410df87a-1370-44be-a2a7-a39f93b1637b · inbound

Dr. Boot: Bootstrapping Program Synthesis Language Models to Perform Repairing cites this paper.

Dr. Boot: Bootstrapping Program Synthesis Language Models to Perform Repairing RLTF: Reinforcement Learning from Unit Test Feedback

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T15:52:20.405213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:52:20.405213Z digest=sha256:640c4178fa4449a9610f7bd38d28ab381e489e5d52a3e8e12102878cafbda2d0

Observation 1d7ba4df-15ae-416f-baeb-94ba70bf9c15 · inbound

BitsAI-Fix: LLM-Driven Approach for Automated Lint Error Resolution in Practice cites this paper.

BitsAI-Fix: LLM-Driven Approach for Automated Lint Error Resolution in Practice RLTF: Reinforcement Learning from Unit Test Feedback

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T04:31:01.806003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:31:01.806003Z digest=sha256:f62fa226aab38dd571b1a3ddfb44037bbaa6fe449adb645cb847bf08d03f4cee

Observation ba9eba82-17ef-4d24-bc79-9752b4484a38 · inbound

Efficiency of turbulence cites this paper.

Efficiency of turbulence RLTF: Reinforcement Learning from Unit Test Feedback

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T00:58:28.756567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:58:28.756567Z digest=sha256:a30ff46d61e7a226ac32bf54ebdaed49ab4aa3068a484e824b6ef02edb9683ee

Observation b94c0dae-105f-471e-bd15-ed0e24365f63 · inbound

InfoSynth: Information-Guided Benchmark Synthesis for LLMs cites this paper.

InfoSynth: Information-Guided Benchmark Synthesis for LLMs RLTF: Reinforcement Learning from Unit Test Feedback

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T13:06:57.310399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:06:57.310399Z digest=sha256:5d7a0221d41978c9e41695916ca6c774b9764c923167837f4994120a547b69d0

Observation 89d4fab7-2e22-4c66-b56d-19dfef7c43d8 · inbound

Beyond Binary: Turning Partial Success into Dense Verifiable Rewards for Reinforcement Learning in Code Generation cites this paper.

Beyond Binary: Turning Partial Success into Dense Verifiable Rewards for Reinforcement Learning in Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T12:19:32.688253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T12:19:32.688253Z digest=sha256:43849645c4d4bffc52ea4f234cd523813c68262214b46096609a8abbbcc01638

Observation f5ca73ae-328a-4a89-acac-7700313e646d · inbound

An Iterative Test-and-Repair Framework for Competitive Code Generation cites this paper.

An Iterative Test-and-Repair Framework for Competitive Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:35:48.667824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T19:44:32.950977Z digest=sha256:3c3bf9da71ce969dce09e918cdaf3e13c502f558b93815dbee197433291c8293

Observation d0c75916-e9c6-455f-b7d1-672ee32b87d1 · inbound

An Iterative Test-and-Repair Framework for Competitive Code Generation cites this paper.

An Iterative Test-and-Repair Framework for Competitive Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-13T09:22:44.557413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T09:22:44.557413Z digest=sha256:fc1e80230166ebc89dc7a316f968faa40bb5770e4d28f21fef530618bf793d55

Observation 4f369d24-d0c0-4a41-95d2-97c4d85e1afe · inbound

Generate, Filter, Control, Replay: A Comprehensive Survey of Rollout Strategies for LLM Reinforcement Learning cites this paper.

Generate, Filter, Control, Replay: A Comprehensive Survey of Rollout Strategies for LLM Reinforcement Learning RLTF: Reinforcement Learning from Unit Test Feedback

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:15:49.220299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T19:15:27.406778Z digest=sha256:46f6cfe5ca479f25159e349cb97f5c0403b3393bef7c9fcf536a337c18575985

Observation 3ee44391-847f-4928-965a-a1d915aefd0e · inbound

ShapeCodeBench: A Renewable Benchmark for Perception-to-Program Reconstruction of Synthetic Shape Scenes cites this paper.

ShapeCodeBench: A Renewable Benchmark for Perception-to-Program Reconstruction of Synthetic Shape Scenes RLTF: Reinforcement Learning from Unit Test Feedback

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:22:02.191104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T01:19:04.708277Z digest=sha256:b62318da35abef96f7f5ad09798eb8ae34c049dba25125d9335ded350fce81dd

Observation 28e01ef3-ae78-48b6-b3ec-8b69ab5954d6 · inbound

Code as Agent Harness cites this paper.

Code as Agent Harness RLTF: Reinforcement Learning from Unit Test Feedback

Reference 102

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:58:14.222864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T10:54:54.558241Z digest=sha256:64a7e9746a37e1d45fe6582743e63248fe241021acb2e786e122d880106b17ee

Observation 60dda99f-c8f2-44f6-b592-999725e379dd · inbound

Building to the Test: Coding Agents Deliver What You Check, Not What You Requested cites this paper.

Building to the Test: Coding Agents Deliver What You Check, Not What You Requested RLTF: Reinforcement Learning from Unit Test Feedback

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-30T03:24:12.930587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T01:40:24.510828Z digest=sha256:9b2e07ac52229e7a6ecf67fa55e61d929533a9139fbdf9a2c8280d4160b89c0d

Observation 512f06d4-a10c-4f0c-bb94-79332a221459 · inbound

Beyond the Need for Speed: Energy-Aware Code Generation via Simulation-Guided Reinforcement Learning cites this paper.

Beyond the Need for Speed: Energy-Aware Code Generation via Simulation-Guided Reinforcement Learning RLTF: Reinforcement Learning from Unit Test Feedback

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-11T17:00:48.664985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T17:00:48.664985Z digest=sha256:f6ae90d09ae3f4716d9cca81cb6ebd80aea25a479455c107227666516a380bc0

Observation 59afdc66-0d0a-4922-b340-a70b187a9293 · inbound

RLPF: Reinforcement Learning from Performance Feedback for Code Generation cites this paper.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.905721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.905721Z digest=sha256:4cdec977d526f062b592935663658f0270c7561d90bd7b20bf8fbd2cf67d949f