Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2305.08844.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T20:02:27.304144Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T20:57:23.916786Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 1c01ad14-062f-45ea-84c3-d22bc840301e · inbound
Training Language Models to Self-Correct via Reinforcement Learning RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2ffd8493-2a16-4758-99c7-854ef7f7e73c · inbound
AlphaVerus: Bootstrapping Formally Verified Code Generation through Self-Improving Translation and Treefinement RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb22d84d-125e-49fe-a389-7002506c671a · inbound
Refining Answer Distributions for Improved Large Language Model Reasoning RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1d2656c-05a8-4140-983c-75f78cff1f32 · inbound
Understanding the Dark Side of LLMs' Intrinsic Self-Correction RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 502c660b-9abc-472c-9b00-c8827d543e98 · inbound
Error-driven Data-efficient Large Multimodal Model Tuning RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 716b68ea-4806-4bab-b082-434262358aff · inbound
Towards Intrinsic Self-Correction Enhancement in Monte Carlo Tree Search Boosted Reasoning via Iterative Preference Learning RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aca12052-b038-4c5e-ac9e-45d7fa2bc64b · inbound
When Data Manipulation Meets Attack Goals: An In-depth Survey of Attacks for VLMs RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5a81be3-6384-4657-9ab5-a0f67f876325 · inbound
Boosting LLM Reasoning via Spontaneous Self-Correction RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dce8310b-6091-4582-b261-78014fd6a7fb · inbound
Formalizing Learning from Language Feedback with Provable Guarantees RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d058e917-99fb-4147-a0fb-f5940a049c4d · inbound
SGIC: A Self-Guided Iterative Calibration Framework for RAG RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58597724-9d59-4848-82a8-41f668bbafe8 · inbound
I2CR: Intra- and Inter-modal Collaborative Reflections for Multimodal Entity Linking RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0eeae88-b57d-4f47-a746-03c3ec01fc2d · inbound
Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6a39c6f5-0200-496d-8ec4-931d615ea600 · inbound
DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.