Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2208.12584.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T15:48:52.367061Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T12:29:52.055883Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 9cae0340-d051-47f5-a34d-f4eeeb3f4ca1 · inbound
p-Mean Regret for Stochastic Bandits Socially Fair Reinforcement Learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cca406cb-314b-4baf-b432-62cb5f4153e4 · inbound
Fairness in Reinforcement Learning with Bisimulation Metrics Socially Fair Reinforcement Learning
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a870e05-a123-40e1-ad4d-9e1d458f7817 · inbound
Multi-User Dueling Bandits: A Fair Approach using Nash Social Welfare Socially Fair Reinforcement Learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 8ba50916-e4d7-4757-9530-79e66ab7eb87 · inbound
Learning Fair Pareto-Optimal Policies in Multi-Objective Reinforcement Learning Socially Fair Reinforcement Learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 385d0b49-f910-4ce8-ac99-6a870df8a9dd · inbound
Welfarist Control Design -- How to fulfill the societal mandate in multi-agent control? Socially Fair Reinforcement Learning
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.