REVIEW 1 cited by
Markov Decision Processes under Ambiguity
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
We consider statistical Markov Decision Processes where the decision maker is risk averse against model ambiguity. The latter is given by an unknown parameter which influences the transition law and the cost functions. Risk aversion is either measured by the entropic risk measure or by the Average Value at Risk. We show how to solve these kind of problems using a general minimax theorem. Under some continuity and compactness assumptions we prove the existence of an optimal (deterministic) policy and discuss its computation. We illustrate our results using an example from statistical decision theory.
Forward citations
Cited by 1 Pith paper
-
Discrete time portfolio optimisation managing value at risk under heavy tail return distribution
A dynamic programming framework for VaR-constrained portfolio choice under heavy-tailed returns, but the numerical implementation fits prices instead of returns and omits the actual optimization.
Discussion (0). Continue with ORCID to comment.