Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2212.04717.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T15:55:14.418381Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-24T05:56:01.775640Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation c59668c5-e4a3-475b-ae45-4ce0524dd658 · inbound
Towards Understanding Sycophancy in Language Models On the Sensitivity of Reward Inference to Misspecified Human Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation cac51f1b-d807-403b-9007-7626c20493f2 · inbound
Active teacher selection for reward learning On the Sensitivity of Reward Inference to Misspecified Human Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 135f7846-d08d-45df-a25d-07eb56de92bc · inbound
Solving the Inverse Alignment Problem for Efficient RLHF On the Sensitivity of Reward Inference to Misspecified Human Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3960d9f-4779-478e-8d5b-ad6648f7c391 · inbound
Multi-Turn On-Policy Distillation with Prefix Replay On the Sensitivity of Reward Inference to Misspecified Human Models
Reference 278
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9477a98-f093-4cfa-ba7e-ede5ca1cd741 · inbound
Multi-Turn On-Policy Distillation with Prefix Replay On the Sensitivity of Reward Inference to Misspecified Human Models
Reference 279
Source-reported events for the cited work
Unavailable: canonical work link unavailable.