A korrentour readingWhat is a korrent?
Learning from a reward ten years away is a solved problem: a value function trained by temporal-difference learning rewards the steps along the way.
Drawn from what Richard Sutton said
A korrentour readingWhat is a korrent?
Drawn from what Richard Sutton said
Word for word, with the source under each one. They did not write this page.
Computer scientist who founded the field of reinforcement learning
when you learn to play chess you have the grand the long-term goal is winning the game and yet you you can't you um you want to be able to learn from shorter term things like you know taking the your opponent's pieces um and so you do that by having a value function which predicts the long-term outcome
Where this was said
Watch from 0:28:39 plays here↗Richard Sutton – Father of RL thinks LLMs are a dead endyoutube.com
All 25 korrents from this recordingthis one is 13th
Added to korrents 26 Sept 2025 · How quotes work · Something wrong? Tell us
Our reading — they may agree, disagree or merely touch the same thing. Closest first: a shared subject counts for most, then how near the wording is.