Logo elodees  elodees

A caring AI for a better world













Only alphabetic characters accented or not as well as the space are accepted

Logo IA




Policy gradient method





No account yet ?

Sign up to access all content




In reinforcement learning, a policy gradient method is an algorithm that learns a policy directly by being directly interested in it.

The policy gradient method allows the optimization of the parameterized policy with respect to the expected performance with the gradient descent method.

At the end of a certain number of iterations, the objective is to obtain a maximization of the performance of the policy for a studied model.

Policy gradient methods are therefore opposed to value-based methods which optimize values in order to then define the optimal policy for these values.



Reinforcement learning













Welcome, my name is Eric Soupet and I am the administrator of the site elodees.com. elodees.com is a state of the art of Artificial Intelligence and aims to be collaborative, you can now offer content such as articles, events, tutorials, ... so don't hesitate !

Platform images credit : Pixabay - Pixabay License | Pexels - Pexels License