REVIEW 2 cited by
Online Optimization of Switched LTI Systems Using Continuous-Time and Hybrid Accelerated Gradient Flows
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
This paper studies the design of feedback controllers to steer a switching linear time-invariant dynamical system towards the solution trajectory of a time-varying convex optimization problem. We propose two types of controllers: (i) a continuous controller inspired by the online gradient descent method, and (ii) a hybrid controller that can be interpreted as an online version of Nesterov's accelerated gradient method with restarts of the state variables. By design, the controllers continuously steer the system towards the time-varying optimizer without requiring knowledge of exogenous disturbances affecting the system. For cost functions that are smooth and satisfy the Polyak-\L ojasiewicz inequality, we demonstrate that the online gradient-flow controller ensures uniform global exponential stability when the time scales of the system and controller are sufficiently separated and the switching signal of the system varies slowly on average. For cost functions that are strongly convex, we show that the hybrid accelerated controller outperforms the continuous gradient descent method. When the cost function is not strongly convex, we show that the the hybrid accelerated method guarantees global practical asymptotic stability.
Forward citations
Cited by 2 Pith papers
-
MARCO: Hardware-Aware Neural Architecture Search for Edge Devices with Multi-Agent Reinforcement Learning and Conformal Prediction Filtering
A two-agent RL search with a conformal prediction filter designs mixed-precision networks for microcontrollers, cutting search time by 3-4x versus once-for-all with near-equal accuracy.
-
Some remarks on gradient dominance and LQR policy optimization
Continuous-time LQR policy optimization satisfies a saturated PŁI condition that yields input-to-state stability of perturbed gradient flows, and overparametrization can restore global exponential convergence.
Discussion (0). Continue with ORCID to comment.