Policy gradient methods

Peters, J

doi:10.4249/scholarpedia.3698

Item

ITEM ACTIONSEXPORT

DownloadE-Mail

Please note that a newer version of this item is available:
https://pure.mpg.de/pubman/item/item_1788428_2

DetailsSummary

Policy gradient methods

Peters, J. (2010). Policy gradient methods. Scholarpedia, 5(11), 3698. doi:10.4249/scholarpedia.3698.

Item is Released

show all hide all

Basic

show hide

Item Permalink: https://hdl.handle.net/11858/00-001M-0000-0013-BD68-3 Version Permalink: https://hdl.handle.net/11858/00-001M-0000-0013-BD69-1

Genre: Journal Article

Files

show Files

Locators

show

Creators

show

hide

Creators:
Peters, J^{1, 2}, Author

Affiliations:
1Department Empirical Inference, Max Planck Institute for Biological Cybernetics, Max Planck Society, ou_1497795
2Dept. Empirical Inference, Max Planck Institute for Intelligent Systems, Max Planck Society, ou_1497647

Content

show

hide

Free keywords: -

Abstract: Policy gradient methods are a type of reinforcement learning techniques that rely upon optimizing parametrized policies with respect to the expected return (long-term cumulative reward) by gradient descent. They do not suffer from many of the problems that have been marring traditional reinforcement learning approaches such as the lack of guarantees of a value function, the intractability problem resulting from uncertain state information and the complexity arising from continuous states actions.

Details

show

hide

Language(s):

Dates: Date issued: 2010-11

Publication Status: Issued

Pages: -

Publishing info: -

Table of Contents: -

Rev. Type: -

Identifiers: URI: http://www.scholarpedia.org/article/Policy_gradient_methods
DOI: 10.4249/scholarpedia.3698
BibTex Citekey: 6940

Degree: -

Event

show

Legal Case

show

Project information

show

Source 1

show

hide

Title: Scholarpedia

Source Genre: Journal

Creator(s):

Affiliations:

Publ. Info: -

Pages: - Volume / Issue: 5 (11) Sequence Number: - Start / End Page: 3698 Identifier: -