Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2405.19332.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T10:54:13.592468Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T00:35:10.408592Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 243ea628-e798-41c4-acb3-3fda90368e0b · inbound
SharedRep-RLHF: A Shared Representation Approach to RLHF with Diverse Preferences Self-Exploring Language Models: Active Preference Elicitation for Online Alignment
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07e30d42-0401-4a3e-98d7-9f776f3d8e3e · inbound
Outcome-based Exploration for LLM Reasoning Self-Exploring Language Models: Active Preference Elicitation for Online Alignment
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22b91954-b84e-4ac6-a24d-098a88669763 · inbound
Data-dependent Exploration for Online Reinforcement Learning from Human Feedback Self-Exploring Language Models: Active Preference Elicitation for Online Alignment
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 982bf544-798f-40b6-b81c-1003b1591837 · inbound
Data-dependent Exploration for Online Reinforcement Learning from Human Feedback Self-Exploring Language Models: Active Preference Elicitation for Online Alignment
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 757cb0ef-18f2-4480-a604-30e6e4b799b9 · inbound
Recall Isn't Enough: Bounding Commitments in Personalized Language Systems Self-Exploring Language Models: Active Preference Elicitation for Online Alignment
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 04e9207b-5337-4d7f-af8e-b06cb0c08b39 · inbound
Spectral Souping: A Unified Framework for Online Preference Alignment Self-Exploring Language Models: Active Preference Elicitation for Online Alignment
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.