Pith. sign in

REVIEW

Representation of Reinforcement Learning Policies in Reproducing Kernel Hilbert Spaces

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2002.02863 v2 pith:Y6SQ6RWR submitted 2020-02-07 cs.LG stat.ML

classification cs.LGstat.ML
keywords policyembeddedframeworkguaranteeshilbertkernellearninglow-dimensional
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We propose a general framework for policy representation for reinforcement learning tasks. This framework involves finding a low-dimensional embedding of the policy on a reproducing kernel Hilbert space (RKHS). The usage of RKHS based methods allows us to derive strong theoretical guarantees on the expected return of the reconstructed policy. Such guarantees are typically lacking in black-box models, but are very desirable in tasks requiring stability. We conduct several experiments on classic RL domains. The results confirm that the policies can be robustly embedded in a low-dimensional space while the embedded policy incurs almost no decrease in return.

Discussion (0). Sign in to comment.

Pith tools