Pith. sign in

REVIEW

Imitation Learning with Stability and Safety Guarantees

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2012.09293 v2 pith:DKDNIM2M submitted 2020-12-16 eess.SY cs.SY

classification eess.SYcs.SY
keywords dynamicsmethodsafetystabilityconditionscontrollersguaranteesimitation
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

A method is presented to learn neural network (NN) controllers with stability and safety guarantees through imitation learning (IL). Convex stability and safety conditions are derived for linear time-invariant plant dynamics with NN controllers by merging Lyapunov theory with local quadratic constraints to bound the nonlinear activation functions in the NN. These conditions are incorporated in the IL process, which minimizes the IL loss, and maximizes the volume of the region of attraction associated with the NN controller simultaneously. An alternating direction method of multipliers based algorithm is proposed to solve the IL problem. The method is illustrated on an inverted pendulum system, aircraft longitudinal dynamics, and vehicle lateral dynamics.

Discussion (0). Continue with ORCID to comment.

Pith tools