An adaptive Koopman MPC with historical safety constraints is proposed for tobacco conditioning, but the claimed Cpk improvements come from model-generated advisor-mode trajectories, not real closed-loop control.
Data-Driven Inverse Optimal Control for Continuous-Time Nonlinear Systems
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
This paper introduces a novel model-free and a partially model-free algorithm for inverse optimal control (IOC), also known as inverse reinforcement learning (IRL), aimed at estimating the cost function of continuous-time nonlinear deterministic systems. Using the input-state trajectories of an expert agent, the proposed algorithms separately utilize control policy information and the Hamilton-Jacobi-Bellman equation to estimate different sets of cost function parameters. This approach allows the algorithms to achieve broader applicability while maintaining a model-free framework. Also, the model-free algorithm reduces complexity compared to existing methods, as it requires solving a forward optimal control problem only once during initialization. Furthermore, in our partially model-free algorithm, this step can be bypassed entirely for systems with known input dynamics. Simulation results demonstrate the effectiveness and efficiency of our algorithms, highlighting their potential for real-world deployment in autonomous systems and robotics.
citation-role summary
citation-polarity summary
fields
eess.SY 1years
2025 1verdicts
REJECT 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
Online Learning Control Strategies for Industrial Processes with Application for Loosening and Conditioning
An adaptive Koopman MPC with historical safety constraints is proposed for tobacco conditioning, but the claimed Cpk improvements come from model-generated advisor-mode trajectories, not real closed-loop control.