Vol. IV · 22
Reinforcement Learning from Zero: Agents, Rewards, and the Loop That Learns from Consequences
Reinforcement Learning · Part 1
Supervised learning felt intuitive to me from day one: here's the input, here's the right answer, minimize the difference.
Read entry