Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2501.13011.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:03:52.729172Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T08:35:34.312098Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation cad4fade-f35b-4ddc-a36e-bd417372e751 · inbound
Teaching Models to Verbalize Reward Hacking in Chain-of-Thought Reasoning MONA: Myopic Optimization with Non-myopic Approval Can Mitigate Multi-step Reward Hacking
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fc44be3-48ff-4081-9111-0771b43f8d76 · inbound
A Survey on Autonomy-Induced Security Risks in Large Model-Based Agents MONA: Myopic Optimization with Non-myopic Approval Can Mitigate Multi-step Reward Hacking
Reference 131
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7cba60c-98a3-4fd5-88dd-8e6a6cc988b5 · inbound
NEST: Nascent Encoded Steganographic Thoughts MONA: Myopic Optimization with Non-myopic Approval Can Mitigate Multi-step Reward Hacking
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7374811-0504-457b-8b7f-a7891146e2cd · inbound
Evaluating Plan Compliance in Autonomous Programming Agents MONA: Myopic Optimization with Non-myopic Approval Can Mitigate Multi-step Reward Hacking
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 118ac02c-fcf8-4144-8f0c-1527929d9062 · inbound
Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challenges MONA: Myopic Optimization with Non-myopic Approval Can Mitigate Multi-step Reward Hacking
Reference 181
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7e703ea7-08f4-4b0c-b520-f63868b1e0c7 · inbound
Reframing AGI Confrontation with Off Earth Autonomy MONA: Myopic Optimization with Non-myopic Approval Can Mitigate Multi-step Reward Hacking
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a783e649-5861-4c56-97a2-c9cc85a32ab9 · inbound
Draining the Energy Commons: Self-Defeating Over-Appropriation as a Coordination Failure in Agentic LLM Collectives MONA: Myopic Optimization with Non-myopic Approval Can Mitigate Multi-step Reward Hacking
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.