Endpoint null-space projection yields a critic-free off-policy update for continuous-time LQR that recovers the Kleinman gain under a projected actor rank condition.
Reinforcement learning-based adaptive optimal exponent ial tracking control of linear systems with unknown dynamics,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
eess.SY 1years
2026 1verdicts
ACCEPT 1representative citing papers
citing papers explorer
-
Data-Driven Critic-Free Policy Iteration for Continuous-Time Linear Quadratic Regulation
Endpoint null-space projection yields a critic-free off-policy update for continuous-time LQR that recovers the Kleinman gain under a projected actor rank condition.