Pith. sign in

Paper Citation Record · LEDGER

RLPF: Reinforcement Learning from Performance Feedback for Code Generation

As of 7 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 0 inbound Pith citation observations for arXiv:2607.27271.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.27271 v1

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T10:53:08.933314Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

24 of 24 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e9330891-4e2e-4c05-9b43-746cf8730ead · outbound

This paper cites Evaluating Large Language Models Trained on Code.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation Evaluating Large Language Models Trained on Code

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.874539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.874539Z digest=sha256:e6fc777cc41cb70c220a678e0bc1083f5bb7c135e8575902f23c220bcb1e31b4

Observation bef4cad0-d52b-4b87-83e6-eded72b0f436 · outbound

This paper cites EffiPair: Improving the Efficiency of LLM-generated Code with Relative Contrastive Feedback.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation EffiPair: Improving the Efficiency of LLM-generated Code with Relative Contrastive Feedback

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.883214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.883214Z digest=sha256:a357e32d6a4d5bcbe8a5b8513429f0130377719d18a83b9133eb186531959f82

Observation 47a8dcb2-3c67-45f0-a525-b78c749fc98f · outbound

This paper cites InAdvances in Neural Information Pro- cessing Systems.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation InAdvances in Neural Information Pro- cessing Systems

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.885978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.885978Z digest=sha256:c5f4a3e3f8346956121975d420ebb57fdedd170073aba082b05d6293681ffc55

Observation 60951eb3-aa24-4776-857b-6cb220a5d79c · outbound

This paper cites PerfCodeBench: Benchmarking LLMs for System-Level High-Performance Code Optimization.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation PerfCodeBench: Benchmarking LLMs for System-Level High-Performance Code Optimization

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.891638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.891638Z digest=sha256:da1069dbf8a24da8872a01892dd8c803814146e158647cab8e4355363f560737

Observation 9a5b7dc2-b0aa-4a0d-9c6f-b02a1f79626c · outbound

This paper cites DevBench: A Realistic, Developer-Informed Benchmark for Code Generation Models.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation DevBench: A Realistic, Developer-Informed Benchmark for Code Generation Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.895092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.895092Z digest=sha256:6d2c2a81b959de83bf5eeccc301f06e9d633229fdbca64615411b7dffc60f8ba

Observation 1b53eb0d-ebe6-4be9-9aee-8b7a819913ef · outbound

This paper cites Rethinking Code Performance Benchmarks for LLMs.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation Rethinking Code Performance Benchmarks for LLMs

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.900571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.900571Z digest=sha256:11e1b11061499922e1fe3bfd7990c94e4f5e5ac525380a43fa9c236a7f4232d5

Observation d62ac326-57e1-4974-9c40-44d477d16e37 · outbound

This paper cites CUDA-L1: Improving CUDA Optimization via Contrastive Reinforcement Learning.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation CUDA-L1: Improving CUDA Optimization via Contrastive Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.903173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.903173Z digest=sha256:d71274d36d2601bf2c0bc8e4f175baf754dbf5fcc56d1c15d7c52ac6449b7f4e

Observation 59afdc66-0d0a-4922-b340-a70b187a9293 · outbound

This paper cites RLTF: Reinforcement Learning from Unit Test Feedback.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation RLTF: Reinforcement Learning from Unit Test Feedback

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.905721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.905721Z digest=sha256:4cdec977d526f062b592935663658f0270c7561d90bd7b20bf8fbd2cf67d949f

Observation 4b5be7ab-7c0d-4ee0-b2c6-a0aaa95bcc30 · outbound

This paper cites OpenAI.2026.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation OpenAI.2026

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.908244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.908244Z digest=sha256:11116ccb7def6dd6832b900338c94d96cabc517cbedb1e895858c601cb4be561

Observation f0cb727b-2dc9-4458-938c-0b93bc2c43a9 · outbound

This paper cites Ouyang, A.; Guo, S.; Arora, S.; Zhang, A.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation Ouyang, A.; Guo, S.; Arora, S.; Zhang, A

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.910684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.910684Z digest=sha256:99f3502a3254b64f8971a395d5c76497277ddd03a49e1f717bf8801c39caa96c

Observation 0fb95a22-b8e5-43fb-b3b0-24e334074edd · outbound

This paper cites PerfCodeGen: Improving Performance of LLM Generated Code with Execution Feedback.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation PerfCodeGen: Improving Performance of LLM Generated Code with Execution Feedback

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.913097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.913097Z digest=sha256:0e2840821f81c0c65ea4400cc237513da5ef7859dd9b4c50f5403a9641fe27a5

Observation 32ec4a62-5626-4d13-b3c9-133104ea53c4 · outbound

This paper cites EffiBench-X: A Multi-Language Benchmark for Measuring Efficiency of LLM-Generated Code.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation EffiBench-X: A Multi-Language Benchmark for Measuring Efficiency of LLM-Generated Code

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.915570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.915570Z digest=sha256:5cc20e960b64a0ca75939cfe27da00be450fa0a8cea879c7493ae8a0c2472ae6

Observation 78ce596a-61be-40fb-bff8-8c7f36077942 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.918161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.918161Z digest=sha256:783d423ffd112aced67f04d58c9b4f145cc278ef28ed7f1dafeb0bbc9efca934

Observation 93a0138e-48bf-473a-bf74-2024bf7e6166 · outbound

This paper cites Improving Small Language Models for Code Generation with Reinforcement Learning from Verification Feedback.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation Improving Small Language Models for Code Generation with Reinforcement Learning from Verification Feedback

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.923482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.923482Z digest=sha256:4e5182d866aceab4bf19ab0854ada93bbf796949bb6088e0b3378b02dc02c0b7

Observation bef56da3-3259-4401-a3f7-c5ecf94863d5 · outbound

This paper cites an unresolved cited work.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.925914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.925914Z digest=sha256:1e681ab84edbf07570707cce3ddd6af5cdd165bbd9306c8f3bc43f003a4cb542

Observation 35865d5a-124b-4e43-81bc-dbcf1707ed6d · outbound

This paper cites Qwen3 Technical Report.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation Qwen3 Technical Report

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.928329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.928329Z digest=sha256:f9a6dbc295ffbbebe0512e5af470ac9e13cbee5c4b3c1611889f880b7855f896

Observation 59c1f194-136d-4baf-a680-679c52266d59 · outbound

This paper cites Zhu, X.; Zhou, X.; Zhu, B.; Hu, H.; Du, M.; Zhang, H.; Wang, H.; and Guo, Z.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation Zhu, X.; Zhou, X.; Zhu, B.; Hu, H.; Du, M.; Zhang, H.; Wang, H.; and Guo, Z

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.930892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.930892Z digest=sha256:a9fb292cfa38d6af1c286c4617252960dbd1aef9f8216f853e0636336bceea05

Observation 00c04b33-b9d4-4edc-b362-566fde1b79ba · outbound

This paper cites CodeScaler: Scaling Code LLM Training and Test-Time Inference via Reward Models.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation CodeScaler: Scaling Code LLM Training and Test-Time Inference via Reward Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.933314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.933314Z digest=sha256:4bd1227742200d1168713c86c0891552355c13271abc4da9ecbbc614e68cf1fa

Observation 4b3ae5e2-6859-4763-aed5-c6d71a0ab1cb · outbound

This paper cites Program Synthesis with Large Language Models.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation Program Synthesis with Large Language Models

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.870724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.870724Z digest=sha256:d00dcd05e3aa1ec4d034ffb32e567a64498dbf9d664d8fbb05cdb091cd156348

Observation a8fb92f1-ad13-46ae-8ac5-4b2272b2de53 · outbound

This paper cites CodeRL: Mastering Code Generation through Pretrained Models and Deep Reinforcement Learning.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation CodeRL: Mastering Code Generation through Pretrained Models and Deep Reinforcement Learning

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.897721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.897721Z digest=sha256:dacc45f24116d2a75499e708a11874d9aeb121ed4437643765d43f2f6809c9b6

Observation 05535d48-e427-44af-b35c-e4acd4361a96 · outbound

This paper cites Learning Performance-Improving Code Edits.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation Learning Performance-Improving Code Edits

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.920912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.920912Z digest=sha256:a29f7f0bce7bd2437fe4380a160afa669f0f6da404ce85a57ee8d735962f23b8

Observation 9989336a-ed70-49c6-9097-e3ccdfc965fc · outbound

This paper cites RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.880515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.880515Z digest=sha256:d5f1d6de7428be55188dba21b8eb04d464c76cc4fb4ae6c2a8352dab1eab3c33

Observation e337d119-be4c-46a8-b7c0-d02df08c9801 · outbound

This paper cites CodeRL+: Improving Code Generation via Reinforcement with Execution Semantics Alignment.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation CodeRL+: Improving Code Generation via Reinforcement with Execution Semantics Alignment

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.888767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.888767Z digest=sha256:49703ae19d7884cedd9083f846c9787b9cc0dc2ab7c98e34434ba5aefe524f76

Observation db153e02-5e0c-400f-b798-53df2e0cc360 · outbound

This paper cites CoRe-Code: Collaborative Reinforcement Learning for Code Generation.

RLPF: Reinforcement Learning from Performance Feedback for Code Generation CoRe-Code: Collaborative Reinforcement Learning for Code Generation

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-01T10:53:08.877543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:53:08.877543Z digest=sha256:7bc8801a32a35dc2a6f6d4551beb53fbd91d273d3a3a9289a4dad2f591d97ed8

Pith citing papers

No inbound Pith citation observations are available.