Outline
Reinforcement Learning (Review)
Q-Learning
A simple example of Q-Learning
SARSA
E-SARSA
OpenAI Gym toolkit
A paper that presented with Q-Learning
Empirical example in Python and gym toolkit (Q-learning, SARSA, ESARSA)
*Update in Slide 24, minute 54:28: The Q-value will be 34.74*