Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2503.22480.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:30:32.243421Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T07:59:40.146385Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 05a0c8d6-973a-41d6-9df9-e5c73c95ca7f · inbound
Variance-aware Reward Modeling with Anchor Guidance Probabilistic Uncertain Reward Model
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 211795d3-9128-45df-a8d4-2c412bc6daa8 · inbound
A Unifying Lens on Reward Uncertainty in RLHF Probabilistic Uncertain Reward Model
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2f649fe5-fadc-4ca0-b36b-0c601a011908 · inbound
Modularized Reinforcement Learning on LLMs: From MDP Creation to Exploration and Learning Probabilistic Uncertain Reward Model
Reference 190
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f7ab9048-6927-4354-97c3-af51283e1d62 · inbound
Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Probabilistic Uncertain Reward Model
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.