April 2019
Intermediate to advanced
212 pages
5h 34m
English
What does it mean to be in a state, to take an action, or to receive a reward? These are the most important concepts for us to understand intuitively, so let's dig deeper into them. The following diagram depicts the agent-environment interaction in an MDP:

The agent interacts with the environment through actions, and it receives rewards and state information from the environment. In other words, the states and rewards are feedback from the environment, and the actions are inputs to the environment from the agent.
Going back to our simple driving simulator example, our agent might be moving or stopped at a red ...
Read now
Unlock full access