A korrentour readingWhat is a korrent?
Reinforcement learning is terrible, and it only looks good because everything we had before it was much worse.
Drawn from what Andrej Karpathy said
What this subject means
reinforcement learning Training by reward: environments, verifiable tasks, value functions, and whether it works or merely beats what came before.