open Reinforcement learning From first principles to Rainbow DQN. Everything that shows why RL works and why it breaks, one problem and one fix at a time. 1 of 3 written