January 2020
Intermediate to advanced
826 pages
21h 1m
English
In this first chapter of part three of the book, we will consider an alternative way to handle Markov decision process (MDP) problems, which forms a full family of methods called policy gradient methods.
In this chapter, we will:
Before we start talking about policy gradients, let's refresh our minds with the common characteristics of the methods covered in part two of this book. The central ...
Read now
Unlock full access