PDF4PRO ⚡AMP

Modern search engine that looking for books and documents around the web

Example: biology

Policy Gradient Methods for Reinforcement Learning with ...

Back to document page

Advances in Neural Information Processing Systems 12, pp. 1057{1063, MIT Press, 2000Policy Gradient Methods forReinforcement Learning with FunctionApproximationRichard S. Sutton, David McAllester, Satinder Singh, Yishay MansourAT&T Labs { Research, 180 Park Avenue, Florham Park, NJ 07932AbstractFunction approximation is essential to Reinforcement Learning , butthe standard approach of approximating a value function and deter-mining a Policy from it has so far proven theoretically this paper we explore an alternative approach in which the policyis explicitly represented by its own function approximator, indepen-dent of the value function, and is updated according to the gradientof expected reward with respect to the Policy parameters.}}

policy parameters are updated approximately proportional to the gradient: ... Learning a value function and using it to reduce the variance of the gradient estimate appears to be essential for rapid learning. Jaakkola, Singh and Jordan (1995) proved a result very similar to ours for the special case of function ...

  Policy, Methods, Learning, Reinforcement, Derating, Policy gradient methods for reinforcement learning

Download Policy Gradient Methods for Reinforcement Learning with ...


Information

Domain:

Source:

Link to this page:

Please notify us if you found a problem with this document:

Spam in document Broken preview Other abuse

Related search queries