We introduce a reinforcement learning algorithm assisted by a feedback controller. The idea is to enhance tabular learning algorithms by means of a control strategy with limited knowledge of the system model. We show that, by tutoring the learning process, the algorithm converges more quickly than the tabular Q-learning strategy. We use the classical problem of stabilizing an inverted pendulum as a benchmark to numerically illustrate the advantages and disadvantages of the approach.
Paper
References (23)
09OpenAI GymGreg Brockman, Vicki Cheung, Ludwig Pettersson et al.2016 · arXiv.org · 5.6k citations In Library
Scroll for more · 11 remaining