Using Simulation Optimization to Improve Zero-shot Policy Transfer of Quadrotors

2022-01-04 22:32:05

Sven Gronauer, Matthias Kissel, Luca Sacchetto, Mathias Korte, Klaus Diepold

arXiv_AI

Abstract
Abstract (translated)
URL
PDF

Abstract

In this work, we show that it is possible to train low-level control policies with reinforcement learning entirely in simulation and, then, deploy them on a quadrotor robot without using real-world data to fine-tune. To render zero-shot policy transfers feasible, we apply simulation optimization to narrow the reality gap. Our neural network-based policies use only onboard sensor data and run entirely on the embedded drone hardware. In extensive real-world experiments, we compare three different control structures ranging from low-level pulse-width-modulated motor commands to high-level attitude control based on nested proportional-integral-derivative controllers. Our experiments show that low-level controllers trained with reinforcement learning require a more accurate simulation than higher-level control policies.

Abstract (translated)

URL

https://arxiv.org/abs/2201.01369

PDF

https://arxiv.org/pdf/2201.01369.pdf