April 2019
Intermediate to advanced
212 pages
5h 34m
English
Let's discuss one primary difference between the model-free agent that we built in Chapter 4, Teaching a Smartcab to Drive Using Q-Learning, and the Q-networks we'll be building here and in the next chapter. We've brought up this topic briefly before, but it's important to understand it at a higher level:
The primary practical difference is that policy agents, unlike value agents, develop the ability ...
Read now
Unlock full access