Pith. sign in

Paper Citation Record · LEDGER

RLTF: Reinforcement Learning from Unit Test Feedback

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2307.04349.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.04349 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:53:28.190874Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T03:24:12.929087Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2b093690-3f8b-4070-b3d9-5140612b6408 · inbound

A Survey on Large Language Models for Code Generation cites this paper.

A Survey on Large Language Models for Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 167

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:18:06.419569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-13T20:18:06.304134Z digest=sha256:30164ebce43f7b7afac007183bc68781cc4da66ae917c1eac7c6b426a89928fa

Observation 2563ed22-ba8d-474b-b04a-b7e066147a20 · inbound

Search, Verify and Feedback: Towards Next Generation Post-training Paradigm of Foundation Models via Verifier Engineering cites this paper.

Search, Verify and Feedback: Towards Next Generation Post-training Paradigm of Foundation Models via Verifier Engineering RLTF: Reinforcement Learning from Unit Test Feedback

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-12T18:28:47.888890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:28:47.888890Z digest=sha256:080bc1c247d913e2b5c73d3ed1a5ed035b086a87e2b563b9ba1a73182a7255f8

Observation e9159c6d-9c66-4344-94a0-774407c77f20 · inbound

DSTC: Direct Preference Learning with Only Self-Generated Tests and Code to Improve Code LMs cites this paper.

DSTC: Direct Preference Learning with Only Self-Generated Tests and Code to Improve Code LMs RLTF: Reinforcement Learning from Unit Test Feedback

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T17:05:15.141795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:05:15.141795Z digest=sha256:1a70ebb71da122810f34ed92f06541b5fbba8c16b98eb4b6c254200f6a3a3c44

Observation 89267435-b1af-4c59-85f4-772331d5d4af · inbound

GenX: Mastering Code and Test Generation with Execution Feedback cites this paper.

GenX: Mastering Code and Test Generation with Execution Feedback RLTF: Reinforcement Learning from Unit Test Feedback

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T13:09:26.845497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:09:26.845497Z digest=sha256:06811646e203299e36065beadde047188239aaadfd533e802e2eaedc21c1f4bb

Observation 399b8bb8-37a4-4c0a-aa2f-376e220713c2 · inbound

Process-Supervised Reinforcement Learning for Code Generation cites this paper.

Process-Supervised Reinforcement Learning for Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T15:10:38.868972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T15:10:38.868972Z digest=sha256:ee4c3d0e479ca7390bc06d7b4ebcb8e026a7c33c9e29642e7e9a19107375f1bb

Observation 01e10bc9-eb54-4dab-ac59-af1746945dcc · inbound

e-SimFT: Alignment of Generative Models with Simulation Feedback for Pareto-Front Design Exploration cites this paper.

e-SimFT: Alignment of Generative Models with Simulation Feedback for Pareto-Front Design Exploration RLTF: Reinforcement Learning from Unit Test Feedback

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T12:11:21.953185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T12:11:21.953185Z digest=sha256:5dc19fdb84a27a7d91764533718c68bd4b1eadbf681088195830f14579653c88

Observation d67fe6a5-165f-4abe-981a-cdc3bc643bc2 · inbound

Improving RL Exploration for LLM Reasoning through Retrospective Replay cites this paper.

Improving RL Exploration for LLM Reasoning through Retrospective Replay RLTF: Reinforcement Learning from Unit Test Feedback

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T11:53:28.190874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:53:28.190874Z digest=sha256:6ba297e4e526758f12c73423b124fbb89664fb98a81c2f826e669c7603966de3

Observation 41d0aa29-416c-468f-a611-61a732325f4d · inbound

Integrating Symbolic Execution into the Fine-Tuning of Code-Generating LLMs cites this paper.

Integrating Symbolic Execution into the Fine-Tuning of Code-Generating LLMs RLTF: Reinforcement Learning from Unit Test Feedback

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:01.086282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:01.086282Z digest=sha256:5c980b2279077427b40bd40cd6c72a269c4ce3eafd6a3163030d39e076cf5597

Observation c36b0a2e-38c1-4221-9c57-c893235bb0a3 · inbound

Knowledge Augmented Complex Problem Solving with Large Language Models: A Survey cites this paper.

Knowledge Augmented Complex Problem Solving with Large Language Models: A Survey RLTF: Reinforcement Learning from Unit Test Feedback

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-15T23:56:02.142812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:56:02.142812Z digest=sha256:f1855064d2c5a7a74968b3555c436e5eddff83d2437a1c191afcde21c2b668c8

Observation c5bb4075-ad1b-40fd-b894-284cdbc164f6 · inbound

CRPE: Expanding The Reasoning Capability of Large Language Model for Code Generation cites this paper.

CRPE: Expanding The Reasoning Capability of Large Language Model for Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T21:22:19.051890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:22:19.051890Z digest=sha256:c353bbb290023a29ff9587000ca4f17ed0cc6efb6667e3665490098a78851b4f

Observation 410df87a-1370-44be-a2a7-a39f93b1637b · inbound

Dr. Boot: Bootstrapping Program Synthesis Language Models to Perform Repairing cites this paper.

Dr. Boot: Bootstrapping Program Synthesis Language Models to Perform Repairing RLTF: Reinforcement Learning from Unit Test Feedback

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T15:52:20.405213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:52:20.405213Z digest=sha256:40b758cef7abeb9a57e528370b51c64cae7a29ce058ed44b3ff1d9d083f7c869

Observation 1d7ba4df-15ae-416f-baeb-94ba70bf9c15 · inbound

BitsAI-Fix: LLM-Driven Approach for Automated Lint Error Resolution in Practice cites this paper.

BitsAI-Fix: LLM-Driven Approach for Automated Lint Error Resolution in Practice RLTF: Reinforcement Learning from Unit Test Feedback

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T04:31:01.806003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:31:01.806003Z digest=sha256:995ff0c8020aa5fc47a7fecb92a3e17235d751db32b0bb0a5980ef2e4001532e

Observation ba9eba82-17ef-4d24-bc79-9752b4484a38 · inbound

Efficiency of turbulence cites this paper.

Efficiency of turbulence RLTF: Reinforcement Learning from Unit Test Feedback

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T00:58:28.756567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:58:28.756567Z digest=sha256:71d71297e8e25da2a7037f6cf9e71c47f450469cb574ba566f0546b17cbf2a53

Observation b94c0dae-105f-471e-bd15-ed0e24365f63 · inbound

InfoSynth: Information-Guided Benchmark Synthesis for LLMs cites this paper.

InfoSynth: Information-Guided Benchmark Synthesis for LLMs RLTF: Reinforcement Learning from Unit Test Feedback

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T13:06:57.310399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:06:57.310399Z digest=sha256:345efb5ec4821891e32ac1227a35df8835ddba5a3a888c081b2229946b308bb8

Observation 89d4fab7-2e22-4c66-b56d-19dfef7c43d8 · inbound

Beyond Binary: Turning Partial Success into Dense Verifiable Rewards for Reinforcement Learning in Code Generation cites this paper.

Beyond Binary: Turning Partial Success into Dense Verifiable Rewards for Reinforcement Learning in Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T12:19:32.688253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T12:19:32.688253Z digest=sha256:35dbd54f8cc16b5519940df38866b8e7df57dde4dccec5a6fe045f54402900b4

Observation f5ca73ae-328a-4a89-acac-7700313e646d · inbound

An Iterative Test-and-Repair Framework for Competitive Code Generation cites this paper.

An Iterative Test-and-Repair Framework for Competitive Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:35:48.667824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T19:44:32.950977Z digest=sha256:0db8e71f4a2c3f76ccefde38463d6a1eea0cb56d087bc396074c8d7b2df27291

Observation d0c75916-e9c6-455f-b7d1-672ee32b87d1 · inbound

An Iterative Test-and-Repair Framework for Competitive Code Generation cites this paper.

An Iterative Test-and-Repair Framework for Competitive Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-13T09:22:44.557413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T09:22:44.557413Z digest=sha256:8d6491fa5ead5f263cc375e2a728e54c1245aff47247eff8e557878cf0f49bec

Observation 4f369d24-d0c0-4a41-95d2-97c4d85e1afe · inbound

Generate, Filter, Control, Replay: A Comprehensive Survey of Rollout Strategies for LLM Reinforcement Learning cites this paper.

Generate, Filter, Control, Replay: A Comprehensive Survey of Rollout Strategies for LLM Reinforcement Learning RLTF: Reinforcement Learning from Unit Test Feedback

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:15:49.220299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-10T19:15:27.406778Z digest=sha256:81d37e8b6ca54e67b1294c331a41ec037c4513f13a2215dba4c284dbfb9a76fb

Observation 3ee44391-847f-4928-965a-a1d915aefd0e · inbound

ShapeCodeBench: A Renewable Benchmark for Perception-to-Program Reconstruction of Synthetic Shape Scenes cites this paper.

ShapeCodeBench: A Renewable Benchmark for Perception-to-Program Reconstruction of Synthetic Shape Scenes RLTF: Reinforcement Learning from Unit Test Feedback

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:22:02.191104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-13T01:19:04.708277Z digest=sha256:a880c0579c73785392198917106e642ce9bc4a5d1e5392befd12c43f111f214c

Observation 28e01ef3-ae78-48b6-b3ec-8b69ab5954d6 · inbound

Code as Agent Harness cites this paper.

Code as Agent Harness RLTF: Reinforcement Learning from Unit Test Feedback

Reference 102

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:58:14.222864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-20T10:54:54.558241Z digest=sha256:ff637bd709de60f655522ebef8a3e16bc5b773a217db6b544e35488c99e75249

Observation 60dda99f-c8f2-44f6-b592-999725e379dd · inbound

Building to the Test: Coding Agents Deliver What You Check, Not What You Requested cites this paper.

Building to the Test: Coding Agents Deliver What You Check, Not What You Requested RLTF: Reinforcement Learning from Unit Test Feedback

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-30T03:24:12.930587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T01:40:24.510828Z digest=sha256:951ffdcdcb8cfd0f0d8c2b4e6651d37f6356155754a18b676c249a3b5158752a

Observation 512f06d4-a10c-4f0c-bb94-79332a221459 · inbound

Beyond the Need for Speed: Energy-Aware Code Generation via Simulation-Guided Reinforcement Learning cites this paper.

Beyond the Need for Speed: Energy-Aware Code Generation via Simulation-Guided Reinforcement Learning RLTF: Reinforcement Learning from Unit Test Feedback

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-11T17:00:48.664985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T17:00:48.664985Z digest=sha256:f9c80be1b42040869c163c1f02b5e74f5234f2fa28f5bd76bf0692f4b9e93f4f

Observation 59afdc66-0d0a-4922-b340-a70b187a9293 · inbound

RLPF: Reinforcement Learning from Performance Feedback for Code Generation cites this paper.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.905721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.905721Z digest=sha256:51d69547fdd84d114990263fb2bdd98698bc30267bf3ce2be6efe4c3ffe4d0f2