Tap a claim on the ring to put it at the centre.
← Models already generalise substantially from tasks that can be verified to tasks that cannot.
17 connected korrents · 17 moments on record from 28 Jul 2023 to 1 Sept 2026.
Everything filed under reinforcement learning
reinforcement learning
Everything filed under scaling laws
scaling laws
Everything filed under Google
Google
Everything filed under LLMs
LLMs
Everything filed under AI alignment
AI alignment
Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject
Read this korrent: Models already generalise substantially from tasks that can be verified to tasks that cannot.
Models already generalise substantially from tasks that can be verified to tasks that cannot.
Last stated 7 months ago
13 Feb 2026
DA
Dario Amodei — holds since 2026-02-13 — tap for who they are
Same subject: RL environments exist to make a model generalise, not to teach it each skill one at a time — exactly as pre-training does. — tap to centre the map on it
RL environments exist to make a model generalise, not to teach it each skill one at a time — exactly as pre-training does.
Last stated 7 months ago
13 Feb 2026
DA
Dario Amodei — holds since 2026-02-13 — tap for who they are
Same subject: Reasoning will generalize the way instruction tuning did: add enough verifiable domains and, at some point nobody can yet locate, the rest start working on their own. — tap to centre the map on it
Reasoning will generalize the way instruction tuning did: add enough verifiable domains and, at some point nobody can yet locate, the rest start working on their own.
Last stated 2 years ago
3 Feb 2025
NL
Nathan Lambert — holds since 2025-02-03 — tap for who they are
Same subject: Large language models are generalisations of what we already have; the breakthrough that would let AI discover things is still waiting to happen. — tap to centre the map on it
Large language models are generalisations of what we already have; the breakthrough that would let AI discover things is still waiting to happen.
Last stated 3 years ago
3 Dec 2023
LR
Lisa Randall — holds since 2023-12-03 — tap for who they are
Same subject: The way to find what a model can do is to hand it tasks slightly harder than you believe it can handle. — tap to centre the map on it
The way to find what a model can do is to hand it tasks slightly harder than you believe it can handle.
Last stated a month ago
27 Jul 2026
BC
Boris Cherny — holds since 2026-07-27 — tap for who they are
Same subject: Deep learning generalises badly, and catastrophic interference with what a network already knew is the proof of it. — tap to centre the map on it
Deep learning generalises badly, and catastrophic interference with what a network already knew is the proof of it.
Last stated 11 months ago
26 Sept 2025
RS
Richard Sutton — holds since 2025-09-26 — tap for who they are
Same subject: A model you have to fine-tune for each thing you want it to do is not a general-purpose model. — tap to centre the map on it
A model you have to fine-tune for each thing you want it to do is not a general-purpose model.
Last stated 4 weeks ago
12 Aug 2026
CF
Chelsea Finn — holds since 2026-08-12 — tap for who they are
Same subject: AI systems keep trying hard outside training because a model that only exerted itself when it detected training would be useless and would be selected away. — tap to centre the map on it
AI systems keep trying hard outside training because a model that only exerted itself when it detected training would be useless and would be selected away.
Last stated a week ago
1 Sept 2026
AC
Ajeya Cotra — holds since 2026-09-01 — tap for who they are
Same subject: Formal verification is about to become economical, because models are getting good enough at writing the proofs that humans no longer have to. — tap to centre the map on it
Formal verification is about to become economical, because models are getting good enough at writing the proofs that humans no longer have to.
Last stated 5 months ago
22 Apr 2026
MK
Martin Kleppmann — holds since 2026-04-22 — tap for who they are
Same subject: A benchmark result should be reported under a stated budget, or as a curve against test-time compute — never as a single number. — tap to centre the map on it
A benchmark result should be reported under a stated budget, or as a curve against test-time compute — never as a single number.
Last stated 2 months ago
26 Jun 2026
NB
Noam Brown — holds since 2026-06-26 — tap for who they are
Same subject: A lab should spend most of its compute on research rather than on building the next model, because research is where the tenfold yearly efficiency gains come from. — tap to centre the map on it
A lab should spend most of its compute on research rather than on building the next model, because research is where the tenfold yearly efficiency gains come from.
Last stated 6 months ago
13 Mar 2026
DP
Dylan Patel — holds since 2026-03-13 — tap for who they are
Same subject: After pre-training, post-training and test-time scaling, the fourth scaling law is agentic: multiplying AI by spawning agents, and the whole loop scales on one thing, compute. — tap to centre the map on it
After pre-training, post-training and test-time scaling, the fourth scaling law is agentic: multiplying AI by spawning agents, and the whole loop scales on one thing, compute.
Last stated 6 months ago
23 Mar 2026
JH
Jensen Huang — holds since 2026-03-23 — tap for who they are
Same subject: A verifiable task can be optimised by reinforcement learning until a neural network performs it extremely well. — tap to centre the map on it
A verifiable task can be optimised by reinforcement learning until a neural network performs it extremely well.
Last stated 10 months ago
17 Nov 2025
AK
Andrej Karpathy — holds since 2025-11-17 — tap for who they are
Same subject: Capability does not generalise for free: a model that will move mountains on an agentic task still tells the same bad joke it told five years ago. — tap to centre the map on it
Capability does not generalise for free: a model that will move mountains on an agentic task still tells the same bad joke it told five years ago.
Last stated 6 months ago
20 Mar 2026
AK
Andrej Karpathy — holds since 2026-03-20 — tap for who they are
Same subject: Humans barely use reinforcement learning for intelligence — what RL they do use goes into motor tasks, not problem solving. — tap to centre the map on it
Humans barely use reinforcement learning for intelligence — what RL they do use goes into motor tasks, not problem solving.
Last stated 11 months ago
17 Oct 2025
AK
Andrej Karpathy — holds since 2025-10-17 — tap for who they are
Same subject: A company's staff-engineer bar should be set against the best companies in the industry rather than against its own history, which is what makes title inflation a real cost. — tap to centre the map on it
A company's staff-engineer bar should be set against the best companies in the industry rather than against its own history, which is what makes title inflation a real cost.
Last stated 5 months ago
1 Apr 2026
TP
Thuan Pham — holds since 2026-04-01 — tap for who they are
Same subject: A crewed rocket cannot be made safe by making the booster reliable, so the only real way to improve safety is to carry an escape system. — tap to centre the map on it
A crewed rocket cannot be made safe by making the booster reliable, so the only real way to improve safety is to carry an escape system.
Last stated 3 years ago
14 Dec 2023
JB
Jeff Bezos — holds since 2023-12-14 — tap for who they are
Same subject: A monopolist that can no longer grow by winning new users can only grow by making its product worse for the users it already has. — tap to centre the map on it
A monopolist that can no longer grow by winning new users can only grow by making its product worse for the users it already has.
Last stated 3 years ago
28 Jul 2023
CD
Cory Doctorow — holds since 2023-07-28 — tap for who they are
same subject or similar wording a cloud: claims about one subject, named for it bar: when it was last stated, on a scale from 2015 to today — full is today a face: someone on record holding the claim — tap it for who they are
At the centre
Models already generalise substantially from tasks that can be verified to tasks that cannot.
Last stated 13 Feb 2026 · 7 months ago
Holds DA Dario Amodei
Read this korrent →
Similar wording
RL environments exist to make a model generalise, not to teach it each skill one at a time — exactly as pre-training does.
Last stated 13 Feb 2026 · 7 months ago
Holds DA Dario Amodei
Similar wording
Reasoning will generalize the way instruction tuning did: add enough verifiable domains and, at some point nobody can yet locate, the rest start working on their own.
Last stated 3 Feb 2025 · 2 years ago
Holds NL Nathan Lambert
Similar wording
Large language models are generalisations of what we already have; the breakthrough that would let AI discover things is still waiting to happen.
Last stated 3 Dec 2023 · 3 years ago
Holds LR Lisa Randall
Similar wording
The way to find what a model can do is to hand it tasks slightly harder than you believe it can handle.
Last stated 27 Jul 2026 · a month ago
Holds Boris Cherny
Similar wording
Deep learning generalises badly, and catastrophic interference with what a network already knew is the proof of it.
Last stated 26 Sept 2025 · 11 months ago
Holds RS Richard Sutton
Similar wording
A model you have to fine-tune for each thing you want it to do is not a general-purpose model.
Last stated 12 Aug 2026 · 4 weeks ago
Holds CF Chelsea Finn
Similar wording
AI systems keep trying hard outside training because a model that only exerted itself when it detected training would be useless and would be selected away.
Last stated 1 Sept 2026 · a week ago
Holds AC Ajeya Cotra
Similar wording
Formal verification is about to become economical, because models are getting good enough at writing the proofs that humans no longer have to.
Last stated 22 Apr 2026 · 5 months ago
Holds MK Martin Kleppmann
Same subject: scaling laws
A benchmark result should be reported under a stated budget, or as a curve against test-time compute — never as a single number.
Last stated 26 Jun 2026 · 2 months ago
Holds NB Noam Brown
Same subject: scaling laws
A lab should spend most of its compute on research rather than on building the next model, because research is where the tenfold yearly efficiency gains come from.
Last stated 13 Mar 2026 · 6 months ago
Holds DP Dylan Patel
Same subject: scaling laws
After pre-training, post-training and test-time scaling, the fourth scaling law is agentic: multiplying AI by spawning agents, and the whole loop scales on one thing, compute.
Last stated 23 Mar 2026 · 6 months ago
Holds JH Jensen Huang
Same subject: reinforcement learning
A verifiable task can be optimised by reinforcement learning until a neural network performs it extremely well.
Last stated 17 Nov 2025 · 10 months ago
Holds Andrej Karpathy
Same subject: reinforcement learning
Capability does not generalise for free: a model that will move mountains on an agentic task still tells the same bad joke it told five years ago.
Last stated 20 Mar 2026 · 6 months ago
Holds Andrej Karpathy
Same subject: reinforcement learning
Humans barely use reinforcement learning for intelligence — what RL they do use goes into motor tasks, not problem solving.
Last stated 17 Oct 2025 · 11 months ago
Holds Andrej Karpathy
Same subject: Google
A company's staff-engineer bar should be set against the best companies in the industry rather than against its own history, which is what makes title inflation a real cost.
Last stated 1 Apr 2026 · 5 months ago
Holds TP Thuan Pham
Same subject: Google
A crewed rocket cannot be made safe by making the booster reliable, so the only real way to improve safety is to carry an escape system.
Last stated 14 Dec 2023 · 3 years ago
Holds JB Jeff Bezos
Same subject: Google
A monopolist that can no longer grow by winning new users can only grow by making its product worse for the users it already has.
Last stated 28 Jul 2023 · 3 years ago
Holds CD Cory Doctorow