Pith. sign in

REVIEW

Steady-State Error Compensation for Reinforcement Learning with Quadratic Rewards

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2402.09075 v2 pith:P2XGDRTN submitted 2024-02-14 eess.SY cs.LGcs.SY

classification eess.SYcs.LGcs.SY
keywords rewardsteady-statesystemerrorsfunctionssignificantintegrallearning
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The selection of a reward function in Reinforcement Learning (RL) has garnered significant attention because of its impact on system performance. Issues of significant steady-state errors often manifest when quadratic reward functions are employed. Although absolute-value-type reward functions alleviate this problem, they tend to induce substantial fluctuations in specific system states, leading to abrupt changes. In response to this challenge, this study proposes an approach that introduces an integral term. By integrating this integral term into quadratic-type reward functions, the RL algorithm is adeptly tuned, augmenting the system's consideration of reward history, and consequently alleviates concerns related to steady-state errors. Through experiments and performance evaluations on the Adaptive Cruise Control (ACC) and lane change models, we validate that the proposed method effectively diminishes steady-state errors and does not cause significant spikes in some system states.

Discussion (0). Continue with ORCID to comment.

Pith tools