REVIEW 1 cited by
Epistemic Risk-Sensitive Reinforcement Learning
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
abstract
We develop a framework for interacting with uncertain environments in reinforcement learning (RL) by leveraging preferences in the form of utility functions. We claim that there is value in considering different risk measures during learning. In this framework, the preference for risk can be tuned by variation of the parameter $\beta$ and the resulting behavior can be risk-averse, risk-neutral or risk-taking depending on the parameter choice. We evaluate our framework for learning problems with model uncertainty. We measure and control for \emph{epistemic} risk using dynamic programming (DP) and policy gradient-based algorithms. The risk-averse behavior is then compared with the behavior of the optimal risk-neutral policy in environments with epistemic risk.
Forward citations
Cited by 1 Pith paper
-
Diffusion-Modeled Reinforcement Learning for Carbon and Risk-Aware Microgrid Optimization
DiffCarl, a diffusion-actor variant of SAC with carbon pricing and CVaR risk terms, is reported to lower microgrid operating cost by 2.3-30.1% versus baselines, though the paper's own numbers contradict its 28.7% carb...
Discussion (0). Continue with ORCID to comment.