Sample-efficient Reinforcement Learning in Robotic Table Tennis

2020-11-06 10:42:41

Jonas Tebbe, Lukas Krauch, Yapeng Gao, Andreas Zell

arXiv_AI

arXiv_AI Reinforcement_Learning Action

Abstract
Abstract (translated)
URL
PDF

Abstract

Reinforcement learning (RL) has recently shown impressive success in various computer games and simulations. Most of these successes are based on numerous episodes to be learned from. For typical robotic applications, however, the number of feasible attempts is very limited. In this paper we present a sample-efficient RL algorithm applied to the example of a table tennis robot. In table tennis every stroke is different, of varying placement, speed and spin. Therefore, an accurate return has be found depending on a high-dimensional continuous state space. To make learning in few trials possible the method is embedded into our robot system. In this way we can use a one-step environment. The state space depends on the ball at hitting time (position, velocity, spin) and the action is the racket state (orientation, velocity) at hitting. An actor-critic based deterministic policy gradient algorithm was developed for accelerated learning. Our approach shows competitive performance in both simulation and on the real robot in different challenging scenarios. Accurate results are always obtained within under 200 episodes of training. A demonstration video is provided as supplementary material.

Abstract (translated)

URL

https://arxiv.org/abs/2011.03275

PDF

https://arxiv.org/pdf/2011.03275.pdf