Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2409.10164.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T16:18:40.819124Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T00:27:29.852309Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 5d4c7c56-409a-4a5e-9726-ca7b20976493 · inbound
On Almost Surely Safe Alignment of Large Language Models at Inference-Time Quantile Regression for Distributional Reward Models in RLHF
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53ffa0d5-a87a-4185-98c0-eff96bbce587 · inbound
Multi-Domain Explainability of Preferences Quantile Regression for Distributional Reward Models in RLHF
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e14e7958-b5e2-40df-8bca-e24c582c732f · inbound
Amulet: Putting Complex Multi-Turn Conversations on the Stand with LLM Juries Quantile Regression for Distributional Reward Models in RLHF
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 605a7477-2042-4cab-a746-468a1fb2b8d0 · inbound
PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Quantile Regression for Distributional Reward Models in RLHF
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33048014-125a-43ab-99e2-e3c720b8ab88 · inbound
Post-Training Large Language Models via Reinforcement Learning from Self-Feedback Quantile Regression for Distributional Reward Models in RLHF
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1abfe77c-dd23-4c1d-94a2-7941aff43713 · inbound
Safety Game: Inference-Time Alignment of Black-Box LLMs via Constrained Optimization Quantile Regression for Distributional Reward Models in RLHF
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff967312-685a-45e0-8d41-ac7f87920a54 · inbound
DVPO: Distributional Value Modeling-based Policy Optimization for LLM Post-Training Quantile Regression for Distributional Reward Models in RLHF
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 607041f9-2297-447a-8ecb-117f80fea90b · inbound
Hyperfastrl: Hypernetwork-based reinforcement learning for unified control of parametric chaotic PDEs Quantile Regression for Distributional Reward Models in RLHF
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 219033b7-07ce-4dd3-aef9-48b3bbbb5b42 · inbound
Text-to-Distribution Prediction with Quantile Tokens and Neighbor Context Quantile Regression for Distributional Reward Models in RLHF
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation af1524c6-ab14-47d6-980b-fc1416bfd25c · inbound
StoryAlign: Evaluating and Training Reward Models for Story Generation Quantile Regression for Distributional Reward Models in RLHF
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 51585fec-0444-44c3-baa0-d5dc12cf2a50 · inbound
A Unifying Lens on Reward Uncertainty in RLHF Quantile Regression for Distributional Reward Models in RLHF
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2646d0b2-5b6e-49fb-a256-db1c4c2be5ad · inbound
Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges Quantile Regression for Distributional Reward Models in RLHF
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.