Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2409.13156.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:09:53.954441Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T03:19:30.665306Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 31781822-1fb8-43b9-b342-af1f9554021d · inbound
Think-RM: Enabling Long-Horizon Reasoning in Generative Reward Models RRM: Robust Reward Model Training Mitigates Reward Hacking
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e72c6b5-9edd-4f59-b902-d1674575148e · inbound
Learning a Pessimistic Reward Model in RLHF RRM: Robust Reward Model Training Mitigates Reward Hacking
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e836238a-5e6c-4473-87fd-580985d69e41 · inbound
Teach a Reward Model to Correct Itself: Reward Guided Adversarial Failure Discovery for Robust Reward Modeling RRM: Robust Reward Model Training Mitigates Reward Hacking
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9236ba57-eb05-42d2-a5e3-a1d44d921573 · inbound
Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities RRM: Robust Reward Model Training Mitigates Reward Hacking
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f11790f-2146-4f4b-b693-e3c2f5d2a95e · inbound
Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges RRM: Robust Reward Model Training Mitigates Reward Hacking
Reference 246
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 151a3599-716c-4f98-ad09-f762cebc36b2 · inbound
Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction RRM: Robust Reward Model Training Mitigates Reward Hacking
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38324b6b-7ff6-4626-aae3-e722bec70186 · inbound
Encouraging Good Processes Without the Need for Good Answers: Reinforcement Learning for LLM Agent Planning RRM: Robust Reward Model Training Mitigates Reward Hacking
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a28b8faa-a12c-4835-9c20-62263506f7aa · inbound
Factored Causal Representation Learning for Robust Reward Modeling in RLHF RRM: Robust Reward Model Training Mitigates Reward Hacking
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6ef86500-44be-4aa3-b01e-412e34b3bb22 · inbound
MARS: Margin and Semantic-Aware Data Augmentation for Reward Modeling RRM: Robust Reward Model Training Mitigates Reward Hacking
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f864ed2-9d40-4b23-ba63-4e6f83bd6e96 · inbound
Optimal Transport for LLM Reward Modeling from Noisy Preference RRM: Robust Reward Model Training Mitigates Reward Hacking
Reference 212
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fa2ac8bf-e749-4359-8e2b-63111cc12bbb · inbound
The Hidden Bias of Process Reward Models:PRISM for Rewarding the Right Reasoning RRM: Robust Reward Model Training Mitigates Reward Hacking
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7d0e2e98-01af-477f-a1c4-d3df1ae905df · inbound
Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization RRM: Robust Reward Model Training Mitigates Reward Hacking
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4f0df22f-a659-4a08-9687-546640696a52 · inbound
Uncertainty-Aware Reward Modeling for Stable RLHF RRM: Robust Reward Model Training Mitigates Reward Hacking
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 979b804f-5af7-42f0-8aa5-5ca53274541e · inbound
Multi-Turn On-Policy Distillation with Prefix Replay RRM: Robust Reward Model Training Mitigates Reward Hacking
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c81f251-c07b-4059-a2c3-79fcf80c5546 · inbound
Multi-Turn On-Policy Distillation with Prefix Replay RRM: Robust Reward Model Training Mitigates Reward Hacking
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62b0f854-1733-48f8-9274-0cffbcf5f435 · inbound
Test-Time Scaling via Error Localization RRM: Robust Reward Model Training Mitigates Reward Hacking
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.