Chapter 3
Tabular RL and Q-Learning
Use small finite environments to understand learning from experience, balancing exploration and exploitation, and updating action values.
Chapter 3
Use small finite environments to understand learning from experience, balancing exploration and exploitation, and updating action values.