pith. sign in

arxiv: 1811.07201 · v1 · pith:MLDKJS2Bnew · submitted 2018-11-17 · 💻 cs.LG · stat.ML

Recursive Sparse Pseudo-input Gaussian Process SARSA

classification 💻 cs.LG stat.ML
keywords gaussianprocesslearningpseudo-inputrecursivesarsasparsealgorithm
0
0 comments X
read the original abstract

The class of Gaussian Process (GP) methods for Temporal Difference learning has shown promise for data-efficient model-free Reinforcement Learning. In this paper, we consider a recent variant of the GP-SARSA algorithm, called Sparse Pseudo-input Gaussian Process SARSA (SPGP-SARSA), and derive recursive formulas for its predictive moments. This extension promotes greater memory efficiency, since previous computations can be reused and, interestingly, it provides a technique for updating value estimates on a multiple timescales

This paper has not been read by Pith yet.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.