Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2410.00847.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:20:39.668066Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 3a96b802-7114-41b3-929f-8842b77fe6fc · inbound
Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs Uncertainty-aware Reward Model: Teaching Reward Models to Know What is Unknown
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bfb96da5-a6e1-454f-9a76-5ed920a99fc4 · inbound
Decision Flow Policy Optimization Uncertainty-aware Reward Model: Teaching Reward Models to Know What is Unknown
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f36f7cb-1e25-4f99-9ba8-6c51aac05e88 · inbound
Amulet: Putting Complex Multi-Turn Conversations on the Stand with LLM Juries Uncertainty-aware Reward Model: Teaching Reward Models to Know What is Unknown
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21bd6289-0180-41c4-8cd2-23152b44ea54 · inbound
Test-Time Scaling with Repeated Sampling Improves Multilingual Text Generation Uncertainty-aware Reward Model: Teaching Reward Models to Know What is Unknown
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 902799f6-fb92-4983-98a6-59f7a3fff674 · inbound
Towards Reliable, Uncertainty-Aware Alignment Uncertainty-aware Reward Model: Teaching Reward Models to Know What is Unknown
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73bda382-13c6-4c90-8b42-9ab188b7af6c · inbound
Post-Training Large Language Models via Reinforcement Learning from Self-Feedback Uncertainty-aware Reward Model: Teaching Reward Models to Know What is Unknown
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11f61d99-e7ed-46bd-9690-0283bd6cae50 · inbound
Uncertainty Propagation in LLM-Based Systems Uncertainty-aware Reward Model: Teaching Reward Models to Know What is Unknown
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5ed72434-0f99-4ea8-bfbd-fd3b2932c86e · inbound
Test-Time Personalization: A Diagnostic Framework and Probabilistic Fix for Scaling Failures Uncertainty-aware Reward Model: Teaching Reward Models to Know What is Unknown
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 671935e5-6850-4189-85fc-f0ab8475c13a · inbound
Variance-aware Reward Modeling with Anchor Guidance Uncertainty-aware Reward Model: Teaching Reward Models to Know What is Unknown
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 361b6382-5bf2-4435-8965-e6c977f5c9bc · inbound
World Models: A Comprehensive Survey of Architectures, Methodologies, Reasoning Paradigms, and Applications Uncertainty-aware Reward Model: Teaching Reward Models to Know What is Unknown
Reference 248
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 03aed5eb-fb96-49c5-929d-cde6b515c9ee · inbound
DriveReward: A Comprehensive Dataset and Generative Vision-Language Reward Model for Autonomous Driving Uncertainty-aware Reward Model: Teaching Reward Models to Know What is Unknown
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 844bd634-e661-417c-a104-7e262df5aa57 · inbound
DynaCF: Mitigating Shortcut Learning in Reward Models via Dynamic Counterfactual Sensitivity Uncertainty-aware Reward Model: Teaching Reward Models to Know What is Unknown
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 607f6be3-24da-466f-a1df-51fbdc437eac · inbound
CSPF: A Constrained Shared-Private Fusion Method for Non-Verifiable Preference Evaluation Uncertainty-aware Reward Model: Teaching Reward Models to Know What is Unknown
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.