Fig. I — A run, in order
Reinforcement Learning
4 parts, read in order — each one assumes the one before it. Back toall runs orthe full archive.
The parts
Supervised learning felt intuitive to me from day one: here's the input, here's the right answer, minimize the difference.