Pith. sign in

REVIEW 1 cited by

Convergence Analysis of Gradient Flow for Overparameterized LQR Formulations

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2408.15456 v3 pith:D6FA2VLY submitted 2024-08-28 eess.SY cs.SYmath.DS

classification eess.SYcs.SYmath.DS
keywords convergencegradientanalysiscaseresultstrainingconservationcontrol
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Motivated by the growing use of artificial intelligence (AI) tools in control design, this paper analyses the intersection between results from gradient methods for the model-free linear quadratic regulator (LQR), and linear feedforward neural networks (LFFNNs), More specifically, it looks into the case where one wants to find a LFFNN feedback that minimizes a LQR cost. This paper starts by analyzing the structure of the gradient expression for the parameters of each layer, which implies a key conservation law of the system. This conservation law is then leveraged to generalize existing results on boundedness and global convergence of solutions to critical points, and invariance of the set of stabilizing networks under the training dynamics. This is followed by an analysis of the case where the LFFNN has a single hidden layer, for which the paper proves that the training converges not only to the critical points, but to the optimal feedback control law for all but a set of Lebesgue measure zero of the initializations. These theoretical results are followed by an extensive analysis of a simple version of the problem -- the ``vector case'' -- proving the theoretical properties of accelerated convergence and small-input input-to-state stability (ISS) for this simpler example. Finally, the paper presents numerical evidence of faster convergence of the training of general LFFNNs when compared to non-overparameterized formulations, showing that the acceleration of the solution is observable even when the gradient is not explicitly computed, but estimated from evaluations of the cost function.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Similarity Matching Networks: Hebbian Learning and Convergence Over Multiple Time Scales

    q-bio.NC 2025-06 conditional novelty 6.0 of 10

    A continuous-time Hebbian/anti-Hebbian similarity matching network is shown to converge layer by layer to the principal subspace solution, with the slow layer convergence relying on two unproven conjectures.

Pith tools