Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2405.19316.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:20:00.991880Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-21T13:10:10.421442Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 797bfae3-4e3f-4202-ba83-994136bb68fa · inbound
Frictional Agent Alignment Framework: Slow Down and Don't Break Things Robust Preference Optimization through Reward Model Distillation
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1acd4098-80ed-4623-bbf1-898a97b2c34f · inbound
Risk-aware Direct Preference Optimization under Nested Risk Measure Robust Preference Optimization through Reward Model Distillation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f499fe4f-e977-411a-88ef-e517db30b68f · inbound
Learning a Pessimistic Reward Model in RLHF Robust Preference Optimization through Reward Model Distillation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05043c1f-ced1-415c-9229-4faccb5a5452 · inbound
On Symmetric Losses for Robust Policy Optimization with Noisy Preferences Robust Preference Optimization through Reward Model Distillation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d036d27-18f1-40b4-9b72-d2a38a1bdf8a · inbound
Evaluating the Effectiveness of Direct Preference Optimization for Personalizing German Automatic Text Simplifications for Persons with Intellectual Disabilities Robust Preference Optimization through Reward Model Distillation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 803e3cde-bdbf-4e6e-849f-6e2179ecd459 · inbound
Robust Single-Stage Fully Sparse 3D Object Detection via Detachable Latent Diffusion Robust Preference Optimization through Reward Model Distillation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8134630d-ed0f-48c1-b116-42f1f1a20d78 · inbound
MPO: Multidimensional Preference Optimization for Language Model-based Text-to-Speech Robust Preference Optimization through Reward Model Distillation
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76f6a24b-e98e-4788-ac89-2b54f12811da · inbound
Multiplayer Nash Preference Optimization Robust Preference Optimization through Reward Model Distillation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ee0af6d8-2c8e-43f7-9456-befc13bb6f26 · inbound
LLM Harms: A Taxonomy and Discussion Robust Preference Optimization through Reward Model Distillation
Reference 217
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d8754ce3-7ed1-4c12-b701-f451f8677142 · inbound
LLM Harms: A Taxonomy and Discussion Robust Preference Optimization through Reward Model Distillation
Reference 217
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e62ac6f-5cce-443d-a546-a63ae0fd0e8c · inbound
Provably avoiding over-optimization in Direct Preference Optimization without knowing the data distribution Robust Preference Optimization through Reward Model Distillation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 76156812-159e-4015-bb05-c44fe90c4f7e · inbound
Provably avoiding over-optimization in Direct Preference Optimization without knowing the data distribution Robust Preference Optimization through Reward Model Distillation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8da4c509-4e9c-464b-b81b-2995af2b488f · inbound
Generating Place-Based Compromises Between Two Points of View Robust Preference Optimization through Reward Model Distillation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 52382d59-bd63-41ce-87da-ffbec13e15e0 · inbound
Normalized Rewards for Preference Optimization Robust Preference Optimization through Reward Model Distillation
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.