Gradient based algorithms with loss functions and kernels for improved on-policy control
Loading...
Date
Authors
Robards, Matthew
Sunehag, Peter
Journal Title
Journal ISSN
Volume Title
Publisher
Access Statement
Abstract
We introduce and empirically evaluate two novel online gradient-based reinforcement learning algorithms with function approximation - one model based, and the other model free. These algorithms come with the possibility of having non-squared loss functions which is novel in reinforcement learning, and seems to come with empirical advantages. We further extend a previous gradient based algorithm to the case of full control, by using generalized policy iteration. Theoretical properties of these algorithms are studied in a companion paper.
Description
Keywords
Citation
Collections
Source
Type
Book Title
Recent Advances in Reinforcement Learning - 9th European Workshop, EWRL 2011, Revised Selected Papers
Entity type
Publication