Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2402.07319.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T10:01:57.056053Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T17:58:46.669571Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 45abfca5-0ae6-49c3-943e-6ac160b9dfdd · inbound
Exploring the Secondary Risks of Large Language Models ODIN: Disentangled Reward Mitigates Hacking in RLHF
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2e00a555-b139-4013-8ce2-8655891135f1 · inbound
Teach a Reward Model to Correct Itself: Reward Guided Adversarial Failure Discovery for Robust Reward Modeling ODIN: Disentangled Reward Mitigates Hacking in RLHF
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 163fec65-6cc6-462f-b789-08baac871e75 · inbound
CoLD: Counterfactually-Guided Length Debiasing for Process Reward Models in Mathematical Reasoning ODIN: Disentangled Reward Mitigates Hacking in RLHF
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 674e5b0e-29ed-44e1-8a23-8126f4dbeea9 · inbound
Rubrics as Rewards: Reinforcement Learning Beyond Verifiable Domains ODIN: Disentangled Reward Mitigates Hacking in RLHF
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 29b958c3-ab11-43ba-bbbc-8b7b9dceb1b5 · inbound
Factored Causal Representation Learning for Robust Reward Modeling in RLHF ODIN: Disentangled Reward Mitigates Hacking in RLHF
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0ae14f58-bc00-4810-96d7-48ceae5e544f · inbound
RVPO: Risk-Sensitive Alignment via Variance Regularization ODIN: Disentangled Reward Mitigates Hacking in RLHF
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e6bc8d08-6f0e-4077-aa87-a0c5207aa80b · inbound
AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment ODIN: Disentangled Reward Mitigates Hacking in RLHF
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ec71211d-e745-4e88-9e62-9c33004311fc · inbound
AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment ODIN: Disentangled Reward Mitigates Hacking in RLHF
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5eca7ad6-eb1a-4f45-9ac3-6ff0ab240b1b · inbound
General Preference Reinforcement Learning ODIN: Disentangled Reward Mitigates Hacking in RLHF
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e9f9d502-d305-4086-9927-bedb017f52d8 · inbound
General Preference Reinforcement Learning ODIN: Disentangled Reward Mitigates Hacking in RLHF
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 184967ed-9231-46dc-8553-08bb083bfc90 · inbound
General Preference Reinforcement Learning ODIN: Disentangled Reward Mitigates Hacking in RLHF
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 86b8fada-495c-4ee4-8766-514451a80c4e · inbound
Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization ODIN: Disentangled Reward Mitigates Hacking in RLHF
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 16d00a7b-976c-4d46-909b-80e40d11902d · inbound
Many Voices, One Reward: Multi-Role Rubric Generation for LLM Judging and Reward Modeling ODIN: Disentangled Reward Mitigates Hacking in RLHF
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e09f281b-e0e9-4d96-8bb5-3d7d060122fa · inbound
Multi-Turn On-Policy Distillation with Prefix Replay ODIN: Disentangled Reward Mitigates Hacking in RLHF
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 768ec640-82d1-4148-9f71-238efeaac97c · inbound
Multi-Turn On-Policy Distillation with Prefix Replay ODIN: Disentangled Reward Mitigates Hacking in RLHF
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de695394-6551-43ea-8b6d-e119fa33956d · inbound
Normalized Rewards for Preference Optimization ODIN: Disentangled Reward Mitigates Hacking in RLHF
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.