REVIEW 3 cited by
Learning to Fly in Seconds
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Learning-based methods, particularly Reinforcement Learning (RL), hold great promise for streamlining deployment, enhancing performance, and achieving generalization in the control of autonomous multirotor aerial vehicles. Deep RL has been able to control complex systems with impressive fidelity and agility in simulation but the simulation-to-reality transfer often brings a hard-to-bridge reality gap. Moreover, RL is commonly plagued by prohibitively long training times. In this work, we propose a novel asymmetric actor-critic-based architecture coupled with a highly reliable RL-based training paradigm for end-to-end quadrotor control. We show how curriculum learning and a highly optimized simulator enhance sample complexity and lead to fast training times. To precisely discuss the challenges related to low-level/end-to-end multirotor control, we also introduce a taxonomy that classifies the existing levels of control abstractions as well as non-linearities and domain parameters. Our framework enables Simulation-to-Reality (Sim2Real) transfer for direct RPM control after only 18 seconds of training on a consumer-grade laptop as well as its deployment on microcontrollers to control a multirotor under real-time guarantees. Finally, our solution exhibits competitive performance in trajectory tracking, as demonstrated through various experimental comparisons with existing state-of-the-art control solutions using a real Crazyflie nano quadrotor. We open source the code including a very fast multirotor dynamics simulator that can simulate about 5 months of flight per second on a laptop GPU. The fast training times and deployment to a cheap, off-the-shelf quadrotor lower the barriers to entry and help democratize the research and development of these systems.
Forward citations
Cited by 3 Pith papers
-
A Neural Network Mode for PX4 on Embedded Flight Controllers
A neural network controller trained in simulation runs directly on the PX4 flight controller's microcontroller and tracks a square path on a real quadrotor with behavior similar to simulation.
-
Quadrotor Morpho-Transition: Learning vs Model-Based Control Strategies
A reinforcement learning policy trained in a randomized simulator with motor dynamics and observation delays transfers to hardware and lands a morphing quadrotor through mid-air transformation, beating an MPC baseline...
-
One Net to Rule Them All: Domain Randomization in Quadcopter Racing Across Different Platforms
A domain-randomized neural network policy trained in simulation races both a 3-inch and a 5-inch quadcopter in the real world, and randomization level trades speed for sim-to-real robustness.
Discussion (0). Continue with ORCID to comment.