Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:1908.04734.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:31:58.520812Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 7f8677c2-5b30-41ac-a233-fa9c0bdc1e58 · inbound
RIVAL: Reinforcement Learning with Iterative and Adversarial Optimization for Machine Translation Reward Tampering Problems and Solutions in Reinforcement Learning: A Causal Influence Diagram Perspective
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d85b3b8b-e0fe-4a47-8372-e790844142c7 · inbound
AI Alignment via Incentives and Correction Reward Tampering Problems and Solutions in Reinforcement Learning: A Causal Influence Diagram Perspective
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation df910814-39e0-40df-9a3c-f5d7c16a6846 · inbound
AI Alignment via Incentives and Correction Reward Tampering Problems and Solutions in Reinforcement Learning: A Causal Influence Diagram Perspective
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 20557408-85e8-4ef3-bea9-6014f7709f51 · inbound
Do Androids Dream of Breaking the Game? Systematically Auditing AI Agent Benchmarks with BenchJack Reward Tampering Problems and Solutions in Reinforcement Learning: A Causal Influence Diagram Perspective
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6450a2a7-a3cb-4be2-9690-80086c5e63eb · inbound
Hide to Guide: Learning via Semantic Masking Reward Tampering Problems and Solutions in Reinforcement Learning: A Causal Influence Diagram Perspective
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 40b214d8-0f47-4531-811d-777aff5d58fc · inbound
Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization Reward Tampering Problems and Solutions in Reinforcement Learning: A Causal Influence Diagram Perspective
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7a88bb85-e460-4ef1-b92f-2f796a95c04a · inbound
Reframing AGI Confrontation with Off Earth Autonomy Reward Tampering Problems and Solutions in Reinforcement Learning: A Causal Influence Diagram Perspective
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f1e3aded-70be-4758-abd1-013060a60e25 · inbound
Stop Hand-Holding Your Coding Agent: Engineering the Loops that Replace Step-by-Step Prompting Reward Tampering Problems and Solutions in Reinforcement Learning: A Causal Influence Diagram Perspective
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.