For MDPs with possibly infinite rewards, the minimal and maximal total expected rewards equal the least fixed points of the corresponding Bellman operators.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.LO 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
J-P: MDP. FP. PP.: Characterizing Total Expected Rewards in Markov Decision Processes as Least Fixed Points with an Application to Operational Semantics of Probabilistic Programs (Technical Report)
For MDPs with possibly infinite rewards, the minimal and maximal total expected rewards equal the least fixed points of the corresponding Bellman operators.