REVIEW 1 cited by
MPC-Inspired Neural Network Policies for Sequential Decision Making
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
In this paper we investigate the use of MPC-inspired neural network policies for sequential decision making. We introduce an extension to the DAgger algorithm for training such policies and show how they have improved training performance and generalization capabilities. We take advantage of this extension to show scalable and efficient training of complex planning policy architectures in continuous state and action spaces. We provide an extensive comparison of neural network policies by considering feed forward policies, recurrent policies, and recurrent policies with planning structure inspired by the Path Integral control framework. Our results suggest that MPC-type recurrent policies have better robustness to disturbances and modeling error.
Forward citations
Cited by 1 Pith paper
-
DiLQR: Differentiable Iterative Linear Quadratic Regulator via Implicit Differentiation
DiLQR computes gradients of a converged iLQR controller with implicit differentiation, making the backward pass cost constant in the number of solver iterations.
Discussion (0). Continue with ORCID to comment.