Pith. sign in

Sample-efficient reinforcement learning with stochastic ensemble value expansion

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.LG 1

years

2019 1

verdicts

CONDITIONAL 1

representative citing papers

Model-based Lookahead Reinforcement Learning

cs.LG · 2019-08-15 · conditional · novelty 5.0

The authors propose MPC-MFRL, which couples a TRPO policy with model predictive control and reports model-free-level scores with fewer samples on MuJoCo tasks.

citing papers explorer

Showing 1 of 1 citing paper.

  • Model-based Lookahead Reinforcement Learning cs.LG · 2019-08-15 · conditional · none · ref 3

    The authors propose MPC-MFRL, which couples a TRPO policy with model predictive control and reports model-free-level scores with fewer samples on MuJoCo tasks.