Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:31:54.358241Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 2 inbound Pith citation observations for arXiv:2506.02211.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:31:54.358241Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T20:14:04.101113Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T17:26:04.783434Z
14 of 14 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e7d8571f-bed6-4320-94eb-c67e61bc8125 · outbound
Improving LLM-Generated Code Quality with GRPO Program Synthesis with Large Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27ef6c89-a7dc-41ec-84a4-cefc545ee0fd · outbound
Improving LLM-Generated Code Quality with GRPO StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42e8fd50-6085-47b2-9879-b8b77ad352ed · outbound
Improving LLM-Generated Code Quality with GRPO RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9bfec68-2cfd-40d0-a16a-60fa32f7e1e5 · outbound
Improving LLM-Generated Code Quality with GRPO 2 OLMo 2 Furious
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e323f52-8cda-464e-850c-c18facdd882c · outbound
Improving LLM-Generated Code Quality with GRPO Qwen2.5 Technical Report
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db786141-1b76-41c9-9c68-35409a456dcc · outbound
Improving LLM-Generated Code Quality with GRPO DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18e4674b-e365-46e7-8ec7-e59fadbb27af · outbound
Improving LLM-Generated Code Quality with GRPO Process-Supervised Reinforcement Learning for Code Generation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 378849d3-2eeb-4410-9d09-7f42e5337343 · outbound
Improving LLM-Generated Code Quality with GRPO DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16b8eb3b-5ed2-474b-9cc0-dc6b5efd9bd6 · outbound
Improving LLM-Generated Code Quality with GRPO Measuring Coding Challenge Competence With APPS
Reference 1977
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f03c37c0-91bf-4042-876b-76f091be84d8 · outbound
Improving LLM-Generated Code Quality with GRPO CodeRL: Mastering Code Generation through Pretrained Models and Deep Reinforcement Learning
Reference 2007
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ca98cb9-9959-472b-be3d-e676d2059e9d · outbound
Improving LLM-Generated Code Quality with GRPO Proximal Policy Optimization Algorithms
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7aa5ae01-ac28-44cb-aa26-bbf9606695d6 · outbound
Improving LLM-Generated Code Quality with GRPO Iterative Self-Training for Code Generation via Reinforced Re-Ranking
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6bbf3403-b156-47ef-9a44-b9cca6cd9fd6 · outbound
Improving LLM-Generated Code Quality with GRPO The Llama 3 Herd of Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70f3149e-36fb-4255-b48f-bdcf2e5f911e · outbound
Improving LLM-Generated Code Quality with GRPO Process Supervision-Guided Policy Optimization for Code Generation
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4a76982-bab2-4716-b3a1-911f1a06ad74 · inbound
EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models Improving LLM-Generated Code Quality with GRPO
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 053b9dab-34da-4aa9-834e-b03b938dc901 · inbound
Bridging Generation and Training: A Systematic Review of Quality Issues in LLMs for Code Improving LLM-Generated Code Quality with GRPO
Reference 106
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.