Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2312.08358.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:49:10.543390Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation c0531dae-9d5b-44ea-8ec1-4e8c62143004 · inbound
Test-Time Alignment via Hypothesis Reweighting Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5d19d2f2-1c7b-4921-a9fd-08f0ea002ba9 · inbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 699e2fc7-0744-4fe0-9bf4-189127a6bc72 · inbound
Theoretical Tensions in RLHF: Reconciling Empirical Success with Inconsistencies in Social Choice Theory Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 129db947-a33a-4513-93e2-67f981934587 · inbound
Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e20a2f2a-12a6-4e2e-8637-5519f87425ea · inbound
Active Query Selection for Crowd-Based Reinforcement Learning Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1abf0f85-8f73-4ef6-801e-8190f4448a73 · inbound
RLHF May Not Reflect Genuine Preferences Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4fc5fe13-7443-4442-8716-83ee117809d9 · inbound
Efficient Personalization of Generative User Interfaces Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 06cf540a-7373-4691-b261-15b869a81079 · inbound
Relative Principals, Pluralistic Alignment, and the Structural Value Alignment Problem Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a4e556fa-f511-4cbe-b4cd-41ca01a10e4c · inbound
Personalization Meets Safety:Mechanisms,Risks,and Mitigations in Personalized LLMs Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 186
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7c1484f3-54f5-4e27-b71b-48038f27c286 · inbound
The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 215
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.