In this new video we begin understanding the Markov decision process (MDP) in the context of reinforcement learning! I explain the probabilities involved behind stochastic state transitions and stochastic reward placements. 🙂
💻 The code that we develop in this series can be accessed on GitHub:
