April 2019
Intermediate to advanced
212 pages
5h 34m
English
Recall our discussion of the three major hyperparameters of a Q-learning model:
What values should we choose for these hyperparameters to optimize the performance of our taxi agent? We will discover these values through experimentation once we have constructed our game environment, and we can also take advantage of existing research on the taxi problem and set the variables to known optimal values.
A large part of our model-tuning and optimization phase will consist of comparing the performance of different combinations of these three hyperparamenters together.
One option that we have is the ...
Read now
Unlock full access