login

A Survey of Actor-Critic Reinforcement Learning: Standard and Natural Policy Gradients

IEEE Transactions on Systems Man and Cybernetics Part C (Applications and Reviews)Published 1 November 2012Open access
I. Grondman, Lucian Buşoniu, Gabriel A. D. Lopes, Robert Babuška
Citations1,005
View PDF

TL;DR

The workings of the natural gradient is described, which has made its way into many actor-critic algorithms over the past few years, and a review of several standard and natural actor-critic algorithms is given.

Abstract

International audience

Keywords

Computer ScienceEngineering