Pith. sign in

Paper Citation Record · LEDGER

RLTF: Reinforcement Learning from Unit Test Feedback

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2307.04349.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.04349 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:52:20.405213Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T03:24:12.929087Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2b093690-3f8b-4070-b3d9-5140612b6408 · inbound

A Survey on Large Language Models for Code Generation cites this paper.

A Survey on Large Language Models for Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 167

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:18:06.419569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T20:18:06.304134Z digest=sha256:de85d4cf8d9990129efe5ab784e1d4464d2b87566a8b856099060a8fc9c8d8d2

Observation 410df87a-1370-44be-a2a7-a39f93b1637b · inbound

Dr. Boot: Bootstrapping Program Synthesis Language Models to Perform Repairing cites this paper.

Dr. Boot: Bootstrapping Program Synthesis Language Models to Perform Repairing RLTF: Reinforcement Learning from Unit Test Feedback

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T15:52:20.405213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:52:20.405213Z digest=sha256:f0693850c24ab1cb1f16e7f715af524b6cc789a2045c144f668423e5712bea43

Observation 1d7ba4df-15ae-416f-baeb-94ba70bf9c15 · inbound

BitsAI-Fix: LLM-Driven Approach for Automated Lint Error Resolution in Practice cites this paper.

BitsAI-Fix: LLM-Driven Approach for Automated Lint Error Resolution in Practice RLTF: Reinforcement Learning from Unit Test Feedback

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T04:31:01.806003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:31:01.806003Z digest=sha256:5c8d9cf767d694bd01cf305878ac36adc73ac838c3750b74864862b68c1f7be9

Observation ba9eba82-17ef-4d24-bc79-9752b4484a38 · inbound

Efficiency of turbulence cites this paper.

Efficiency of turbulence RLTF: Reinforcement Learning from Unit Test Feedback

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T00:58:28.756567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:58:28.756567Z digest=sha256:9c6103ff8f1a4c8c6bb084838669eb9291f3a1cb6e43bec4558cc7d8bf8ff92f

Observation b94c0dae-105f-471e-bd15-ed0e24365f63 · inbound

InfoSynth: Information-Guided Benchmark Synthesis for LLMs cites this paper.

InfoSynth: Information-Guided Benchmark Synthesis for LLMs RLTF: Reinforcement Learning from Unit Test Feedback

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T13:06:57.310399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:06:57.310399Z digest=sha256:5d7a0221d41978c9e41695916ca6c774b9764c923167837f4994120a547b69d0

Observation 89d4fab7-2e22-4c66-b56d-19dfef7c43d8 · inbound

Beyond Binary: Turning Partial Success into Dense Verifiable Rewards for Reinforcement Learning in Code Generation cites this paper.

Beyond Binary: Turning Partial Success into Dense Verifiable Rewards for Reinforcement Learning in Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T12:19:32.688253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T12:19:32.688253Z digest=sha256:43849645c4d4bffc52ea4f234cd523813c68262214b46096609a8abbbcc01638

Observation f5ca73ae-328a-4a89-acac-7700313e646d · inbound

An Iterative Test-and-Repair Framework for Competitive Code Generation cites this paper.

An Iterative Test-and-Repair Framework for Competitive Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:35:48.667824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T19:44:32.950977Z digest=sha256:296071a86ab0f318119b885992e286edeef0edbf3805d51eee6eabb65846207e

Observation d0c75916-e9c6-455f-b7d1-672ee32b87d1 · inbound

An Iterative Test-and-Repair Framework for Competitive Code Generation cites this paper.

An Iterative Test-and-Repair Framework for Competitive Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-13T09:22:44.557413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T09:22:44.557413Z digest=sha256:fc1e80230166ebc89dc7a316f968faa40bb5770e4d28f21fef530618bf793d55

Observation 4f369d24-d0c0-4a41-95d2-97c4d85e1afe · inbound

Generate, Filter, Control, Replay: A Comprehensive Survey of Rollout Strategies for LLM Reinforcement Learning cites this paper.

Generate, Filter, Control, Replay: A Comprehensive Survey of Rollout Strategies for LLM Reinforcement Learning RLTF: Reinforcement Learning from Unit Test Feedback

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:15:49.220299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T19:15:27.406778Z digest=sha256:75d013a38982e2aa8c8d033e06e21a462c654149bc6421ec2973477c6658638c

Observation 3ee44391-847f-4928-965a-a1d915aefd0e · inbound

ShapeCodeBench: A Renewable Benchmark for Perception-to-Program Reconstruction of Synthetic Shape Scenes cites this paper.

ShapeCodeBench: A Renewable Benchmark for Perception-to-Program Reconstruction of Synthetic Shape Scenes RLTF: Reinforcement Learning from Unit Test Feedback

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:22:02.191104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T01:19:04.708277Z digest=sha256:45fd5d595f783c57e0c0b5780d89d15b981082d40b92f36e91dbf3ed422ba32d

Observation 28e01ef3-ae78-48b6-b3ec-8b69ab5954d6 · inbound

Code as Agent Harness cites this paper.

Code as Agent Harness RLTF: Reinforcement Learning from Unit Test Feedback

Reference 102

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:58:14.222864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T10:54:54.558241Z digest=sha256:bdd3d5fd2d1612cdd71355c3984d026c289fff995db75c6a5febb4d3ca1b2df3

Observation 60dda99f-c8f2-44f6-b592-999725e379dd · inbound

Building to the Test: Coding Agents Deliver What You Check, Not What You Requested cites this paper.

Building to the Test: Coding Agents Deliver What You Check, Not What You Requested RLTF: Reinforcement Learning from Unit Test Feedback

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-30T03:24:12.930587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T01:40:24.510828Z digest=sha256:e31ac11311ab7ee683e036e8562ac29719fa3dbf06061a1b77606be1404bf792

Observation 512f06d4-a10c-4f0c-bb94-79332a221459 · inbound

Beyond the Need for Speed: Energy-Aware Code Generation via Simulation-Guided Reinforcement Learning cites this paper.

Beyond the Need for Speed: Energy-Aware Code Generation via Simulation-Guided Reinforcement Learning RLTF: Reinforcement Learning from Unit Test Feedback

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-11T17:00:48.664985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T17:00:48.664985Z digest=sha256:f6ae90d09ae3f4716d9cca81cb6ebd80aea25a479455c107227666516a380bc0

Observation 59afdc66-0d0a-4922-b340-a70b187a9293 · inbound

RLPF: Reinforcement Learning from Performance Feedback for Code Generation cites this paper.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.905721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.905721Z digest=sha256:4cdec977d526f062b592935663658f0270c7561d90bd7b20bf8fbd2cf67d949f