MC: Monte Carlo Learning
摘要
We finally start to learn some real RL algorithms. RL algorithms can learn from the experience of interactions between agents and environments, and do not necessarily require the environment models.