Direct Policy Search Reinforcement Learning for Robot Control

El-Fakdi, Andres; Carreras, Marc; Palomeras, Narc&#237;s

Abstract

In this paper, we present Policy Methods as an alternative to Value Methods to solve Reinforcement Learning problems. The paper proposes a Direct Policy Search algorithm that uses a Neural Network to represent the control policies. Details about the algorithm and the update rules are given. The main application of the proposed algorithm is to implement robot control systems, in which the generalization problem usually arises. In this paper, we point out the suitability of our algorithm in a RL benchmark, that was specially designed to test the generalization capability of RL algorithms. Results check out that policy methods obtain better results than value methods in these situations.

This website uses cookies

This website uses cookies