Pith. sign in

REVIEW 1 cited by

Remembering to Be Fair: Non-Markovian Fairness in Sequential Decision Making

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2312.04772 v4 pith:C45CH2HJ submitted 2023-12-08 cs.AI cs.CYcs.LG

classification cs.AIcs.CYcs.LG
keywords fairnessdecisionnon-markovianfairmakingsequentialprocesscontext
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Fair decision making has largely been studied with respect to a single decision. Here we investigate the notion of fairness in the context of sequential decision making where multiple stakeholders can be affected by the outcomes of decisions. We observe that fairness often depends on the history of the sequential decision-making process, and in this sense that it is inherently non-Markovian. We further observe that fairness often needs to be assessed at time points within the process, not just at the end of the process. To advance our understanding of this class of fairness problems, we explore the notion of non-Markovian fairness in the context of sequential decision making. We identify properties of non-Markovian fairness, including notions of long-term, anytime, periodic, and bounded fairness. We explore the interplay between non-Markovian fairness and memory and how memory can support construction of fair policies. Finally, we introduce the FairQCM algorithm, which can automatically augment its training data to improve sample efficiency in the synthesis of fair policies via reinforcement learning.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Fair Resource Allocation in Weakly Coupled Markov Decision Processes

    cs.LG 2024-11 accept novelty 6.0 of 10

    For symmetric weakly coupled MDPs, maximizing a generalized Gini fairness objective reduces to solving a standard average-reward (utilitarian) problem over permutation-invariant policies.

Pith tools