REVIEW 2 cited by
RLPP: A Residual Method for Zero-Shot Real-World Autonomous Racing on Scaled Platforms
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Autonomous racing presents a complex environment requiring robust controllers capable of making rapid decisions under dynamic conditions. While traditional controllers based on tire models are reliable, they often demand extensive tuning or system identification. Reinforcement Learning (RL) methods offer significant potential due to their ability to learn directly from interaction, yet they typically suffer from the sim-to-real gap, where policies trained in simulation fail to perform effectively in the real world. In this paper, we propose RLPP, a residual RL framework that enhances a Pure Pursuit (PP) controller with an RL-based residual. This hybrid approach leverages the reliability and interpretability of PP while using RL to fine-tune the controller's performance in real-world scenarios. Extensive testing on the F1TENTH platform demonstrates that RLPP improves lap times of the baseline controllers by up to 6.37 %, closing the gap to the State-of-the-Art methods by more than 52 % and providing reliable performance in zero-shot real-world deployment, overcoming key challenges associated with the sim-to-real transfer and reducing the performance gap from simulation to reality by more than 8-fold when compared to the baseline RL controller. The RLPP framework is made available as an open-source tool, encouraging further exploration and advancement in autonomous racing research. The code is available at: www.github.com/forzaeth/rlpp.
Forward citations
Cited by 2 Pith papers
-
New Scheme Adaption Strategy for Hyperbolic Conservation Laws
Continuously varying an SBM-type limiter parameter yields a smooth rough-to-smooth transition that improves resolution and cuts dissipation versus threshold-based adaptive schemes for Euler equations.
-
Drive Fast, Learn Faster: On-Board RL for High Performance Autonomous Racing
A residual reinforcement learning controller trained entirely on a physical 1:10 race car beats a pursuit controller's lap time by up to 11.5% after about 20 minutes of on-track practice.
Discussion (0). Continue with ORCID to comment.