REVIEW 1 cited by
Complementary Meta-Reinforcement Learning for Fault-Adaptive Control
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Faults are endemic to all systems. Adaptive fault-tolerant control maintains degraded performance when faults occur as opposed to unsafe conditions or catastrophic events. In systems with abrupt faults and strict time constraints, it is imperative for control to adapt quickly to system changes to maintain system operations. We present a meta-reinforcement learning approach that quickly adapts its control policy to changing conditions. The approach builds upon model-agnostic meta learning (MAML). The controller maintains a complement of prior policies learned under system faults. This "library" is evaluated on a system after a new fault to initialize the new policy. This contrasts with MAML, where the controller derives intermediate policies anew, sampled from a distribution of similar systems, to initialize a new policy. Our approach improves sample efficiency of the reinforcement learning process. We evaluate our approach on an aircraft fuel transfer system under abrupt faults.
Forward citations
Cited by 1 Pith paper
-
A Study of the Efficacy of Generative Flow Networks for Robotics and Machine Fault-Adaptation
On a simulated Reacher arm with four injected faults, CFlowNets matches or beats DDPG, TD3, PPO, and SAC on adaptation speed and asymptotic reward, while using far more GPU memory and wall-clock time.
Discussion (0). Continue with ORCID to comment.