Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2501.09620.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T15:17:58.970049Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T01:37:30.312996Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation c5f310af-a24e-4a49-bc3c-7c61467b464f · inbound
Token-Level LLM Collaboration via FusionRoute Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1652529e-63d4-48c6-b9f1-18c2e2152821 · inbound
Factored Causal Representation Learning for Robust Reward Modeling in RLHF Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 179f691f-920f-4e6c-940f-1558f0e63770 · inbound
Towards Generalizable Reasoning: Group Causal Counterfactual Policy Optimization for LLM Reasoning Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation cfc6e527-9550-452b-97d5-ed00910a49c1 · inbound
Robust Reward Modeling for Large Language Models via Causal Decomposition Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 13eda639-c3b6-4123-b9ec-b038e7e81e47 · inbound
Preference Instability in Reward Models: Detection and Mitigation via Sparse Autoencoders Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e9c2965c-b75e-43b0-9ebb-2b896aa8d476 · inbound
General Preference Reinforcement Learning Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2cb5d842-b19a-408f-ae1c-5c2321cf1067 · inbound
General Preference Reinforcement Learning Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6141da83-70ab-495c-873f-1f9cadf5b143 · inbound
General Preference Reinforcement Learning Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1e74ceca-1986-4219-9dab-cf06241319fe · inbound
Causality as the Statistical Conscience of Artificial Intelligence: From Pearl's Ladder to Trustworthy Machines Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bcaadb17-13e9-4e69-b5b7-dd25a1675cfd · inbound
Reward Bias Substitution: Single-Axis Bias Mitigations Redirect Optimization Pressure Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 187781a2-cf6c-4319-a5dc-22d5ffae41ed · inbound
Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment
Reference 135
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0112c743-f68e-4928-a191-06466a8189c1 · inbound
Style over Substance: A Shortcut Audit of Emotion-Description Preference Evaluation Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 043a5c66-d191-449a-afca-960a7c2be8ca · inbound
What do Reward Models Memorize? Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.