korrents

On the map

Tap a claim on the ring to put it at the centre.

← Value functions only make reinforcement learning faster: anything you…

8 connected korrents · 8 moments on record from 11 Jan 2021 to 12 Aug 2026.

Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: Value functions only make reinforcement learning faster: anything you can do with one you can also do without it, just more slowly. Value functions only makereinforcement learningfaster: anything you can do… IS Ilya Sutskever — holds since 2025-11-25 Same subject: Learning from a reward ten years away is a solved problem: a value function trained by temporal-difference learning rewards the steps along the way. — tap to centre the map on it Learning from a rewardten years away is a… RS Richard Sutton — holds since 2025-09-26 Same subject: The goal was never good simulation; it was answering counterfactuals, and a value function does that job as well as a simulator does. — tap to centre the map on it The goal was never goodsimulation; it was… SL Sergey Levine — holds since 2025-09-12 Same subject: RLHF is not a tax on capability: preference tuning also raises maths and code scores, which is why the labs keep reaching for it. — tap to centre the map on it RLHF is not a tax oncapability: preference… NL Nathan Lambert — holds since 2025-02-03 Same subject: Humans barely use reinforcement learning for intelligence — what RL they do use goes into motor tasks, not problem solving. — tap to centre the map on it Humans barely usereinforcement learning… AK Andrej Karpathy — holds since 2025-10-17 Same subject: The human value function is extraordinarily robust — addiction aside — which is why a person can judge their own performance at a skill they have only just started. — tap to centre the map on it The human valuefunction is… IS Ilya Sutskever — holds since 2025-11-25 Same subject: Reinforcement learning cannot be scaled for robots the way it was for language models, because every attempt spends real robot hours instead of data-centre compute. — tap to centre the map on it Reinforcement learningcannot be scaled for… CF Chelsea Finn — holds since 2026-08-12 Same subject: The value-versus-growth distinction does not serve investors well in a fast-changing world. — tap to centre the map on it The value-versus-growthdistinction does not… HM Howard Marks — holds since 2021-01-11 Same subject: Even the most advanced AI cannot be followed blindly in investing, where value added is zero-sum and what is widely known is therefore worth little. — tap to centre the map on it Even the most advancedAI cannot be followed… RD Ray Dalio — holds since 2026-06-10
same subject or similar wording

At the centre Value functions only make reinforcement learning faster: anything you can do with one you can also do without it, just more slowly. Holds ISIlya Sutskever Read this korrent →