Pith. sign in

REVIEW

Stochastic Action Prediction for Imitation Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2101.01055 v1 pith:IY5E3FRW submitted 2020-12-26 cs.LG cs.RO

classification cs.LGcs.RO
keywords demonstrationsstochasticityactiondataexpertimitationincludinglearning
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Imitation learning is a data-driven approach to acquiring skills that relies on expert demonstrations to learn a policy that maps observations to actions. When performing demonstrations, experts are not always consistent and might accomplish the same task in slightly different ways. In this paper, we demonstrate inherent stochasticity in demonstrations collected for tasks including line following with a remote-controlled car and manipulation tasks including reaching, pushing, and picking and placing an object. We model stochasticity in the data distribution using autoregressive action generation, generative adversarial nets, and variational prediction and compare the performance of these approaches. We find that accounting for stochasticity in the expert data leads to substantial improvement in the success rate of task completion.

Discussion (0). Sign in to comment.

Pith tools