Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 40 inbound Pith citation observations for arXiv:2402.01391.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:39:09.585165Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
3
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 493ee7ce-f59d-4a86-87b1-2c8bae5e5a44 · inbound
A Survey on Large Language Models for Code Generation StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a6fe8069-4af6-42a9-b8a4-475274dab7e2 · inbound
MR-Adopt: Automatic Deduction of Input Transformation Function for Metamorphic Testing StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation dae15ffe-6ac2-4676-832a-1328af38d70a · inbound
DSTC: Direct Preference Learning with Only Self-Generated Tests and Code to Improve Code LMs StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bbc4a7e-b8cc-4b9e-86d0-5d060bdf148d · inbound
Preference Optimization for Reasoning with Pseudo Feedback StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f1aaa35-7be3-47ad-a19d-409ca0333f57 · inbound
Trading Devil RL: Backdoor attack via Stock market, Bayesian Optimization and Reinforcement Learning StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d034c76-dece-451c-96ae-c69ff824e8a2 · inbound
Distilling Desired Comments for Enhanced Code Review with Large Language Models StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a48100f-4cdd-435d-8ebb-f657497a38e3 · inbound
ACECODER: Acing Coder RL via Automated Test-Case Synthesis StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c172a57-6eb3-497a-9cd2-7e8bdbb84290 · inbound
Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 164
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 3403b63f-cad3-4b34-a036-c1401269f0ff · inbound
Themisto: Jupyter-Based Runtime Benchmark StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea30a76c-e77e-40e4-b4e9-06aaaf99fb45 · inbound
Integrating Symbolic Execution into the Fine-Tuning of Code-Generating LLMs StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44bd3ffe-f9ed-4004-ad79-63854f7f6591 · inbound
CRPE: Expanding The Reasoning Capability of Large Language Model for Code Generation StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59474258-9f83-4571-bd6e-e0f7e94a8363 · inbound
SPA-RL: Reinforcing LLM Agents via Stepwise Progress Attribution StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bb10acb-174c-40f0-81d2-ae3bcbc23691 · inbound
Training Language Models to Generate Quality Code with Program Analysis Feedback StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 457bd8db-44ed-4218-8753-a5053c9b1667 · inbound
Afterburner: Reinforcement Learning Facilitates Self-Improving Code Efficiency Optimization StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4887bd46-fc39-4dd5-b312-36bb30a7e0e8 · inbound
CRScore++: Reinforcement Learning with Verifiable Tool and AI Feedback for Code Review StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27ef6c89-a7dc-41ec-84a4-cefc545ee0fd · inbound
Improving LLM-Generated Code Quality with GRPO StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14c18af4-3176-4b4b-a3dc-f9fdb4b4e445 · inbound
D-LiFT: Improving LLM-based Decompiler Backend via Code Quality-driven Fine-tuning StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c073181-d46a-45d0-908e-9845c553223b · inbound
SysTemp: A Multi-Agent System for Template-Based Generation of SysML v2 StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a511d87-afd3-4702-83c6-9d560eabbca4 · inbound
ParaStudent: Generating and Evaluating Realistic Student Code by Teaching LLMs to Struggle StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e11cf801-ef24-4569-aa48-584de858c0cc · inbound
ChemDFM-R: A Chemical Reasoning LLM Enhanced with Atomized Chemical Knowledge StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1182d1ba-ef9a-4b20-a296-7ae0aea84ed5 · inbound
Repair-R1: Better Test Before Repair StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 983fec0b-a38e-41d8-aba5-5a554a3e9642 · inbound
EyeMulator: Improving Code Language Models by Mimicking Human Visual Attention StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation bbe70dc4-8e78-4d4e-88a8-b072a15b3a7e · inbound
AR$^2$: Adversarial Reinforcement Learning for Abstract Reasoning in Large Language Models StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26cfeec9-3740-40b3-b940-d4cd569333d8 · inbound
Outcome-Based RL Provably Leads Transformers to Reason, but Only With the Right Data StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b21c6ace-e1f4-4553-8d50-6b21cb945071 · inbound
The Past Is Not Past: Memory-Enhanced Dynamic Reward Shaping StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9d74a1e8-b880-4c2e-bf55-8761bab3834c · inbound
Towards Enabling An Artificial Self-Construction Software Life-cycle via Autopoietic Architectures StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ae70bde9-1c87-4c90-aff3-4b911f425f8f · inbound
AutoOR: Scalably Post-training LLMs to Autoformalize Operations Research Problems StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 52d9851c-dd6b-438e-b8a0-86ca5b61a069 · inbound
WebGen-R1: Incentivizing Large Language Models to Generate Functional and Aesthetic Websites with Reinforcement Learning StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 81432182-5677-4745-bd85-ee10950b3dc0 · inbound
BoostLoRA: Growing Effective Rank by Boosting Adapters StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c8fd815c-4d19-4ff3-9422-04b899c04fca · inbound
Improving LLM Code Generation via Requirement-Aware Curriculum Reinforcement Learning StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2836c5b3-c6a5-4726-9aa4-7c5e83278da3 · inbound
Beyond Execution: Static-Analysis Rewards and Hint-Conditioned Diffusion RL for Code Generation StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f96a4253-1842-4991-b3d9-ad54b92d3e15 · inbound
Distilling Game Code World Model Generation into Lightweight Large Language Models StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c3ffb77b-e6c1-4d7a-89fc-d70816408c83 · inbound
Efficient Post-training of LLMs for Code Generation With Offline Reinforcement Learning StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 819bc428-11f8-4236-a84a-3f8b2dd030b1 · inbound
Improving Small Language Models for Code Generation with Reinforcement Learning from Verification Feedback StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1e40299e-9fa3-4d54-a36f-a848fbcf333f · inbound
Attention Amnesia in Hybrid LLMs: When CoT Fine-Tuning Breaks Long-Range Recall, and How to Fix It StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8b70a479-b2a7-40a8-b68c-7f5172e8d7ab · inbound
From Trainee to Trainer: LLM-Designed Training Environment for RL with Multi-Agent Reasoning StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a8185407-e9a5-4fd2-906a-d6a1d7386175 · inbound
MAGNIFIED: RL Fine-tuning of Multimodal Large Language Models for Motion Planning StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e5986654-96d5-413b-8126-212f1c8c847e · inbound
When Do Intrinsic Rewards Work for Code Reasoning? A Comprehensive Study StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 396d5770-455f-4181-9e4a-5c07690b25d9 · inbound
DHRCL:Training Code LLMs with Dense Hierarchical Rewards and Curriculum Learning StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab486f05-3228-4ee3-9670-56dd3dd2c8af · inbound
DHRCL:Training Code LLMs with Dense Hierarchical Rewards and Curriculum Learning StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.