Chapter 3

Tabular RL and Q-Learning

Use small finite environments to understand learning from experience, balancing exploration and exploitation, and updating action values.