Tap a claim on the ring to put it at the centre.
← Document similarity scores across a corpus tend to cluster bimodally…
16 connected korrents · 15 moments on record from 11 Oct 2018 to 15 Sept 2026.
Everything filed under LLMs
LLMs
Everything filed under benchmarks
benchmarks
Everything filed under reinforcement learning
reinforcement learning
Everything filed under measuring intelligence
measuring intelligence
Everything filed under AI alignment
AI alignment
Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject
Read this korrent: Document similarity scores across a corpus tend to cluster bimodally near identical and near unrelated, rather than spreading evenly.
Document similarity scores across a corpus tend to cluster bimodally near identical and near unrelated, rather than spreading evenly.
Last stated 2 years ago
3 Jul 2024
NE
Nelson Elhage — holds since 2024-07-03 — tap for who they are
Same subject: The next-token language modeling objective is misaligned with following user instructions helpfully and safely. — tap to centre the map on it
The next-token language modeling objective is misaligned with following user instructions helpfully and safely.
Last stated 5 years ago
4 Mar 2022
JL
Jan Leike — holds since 2022-03-04 — tap for who they are
Same subject: A neural network's latent space is closer to an uncopyrightable syntax than to copyrightable expression. — tap to centre the map on it
A neural network's latent space is closer to an uncopyrightable syntax than to copyrightable expression.
Last stated 2 weeks ago
7 Sept 2026
KK
Kevin Kelly — holds since 2026-09-07 — tap for who they are
Same subject: Current techniques restrict pre-trained representation power because standard language models are unidirectional. — tap to centre the map on it
Current techniques restrict pre-trained representation power because standard language models are unidirectional.
Last stated 8 years ago
11 Oct 2018
KT
Kristina Toutanova — holds since 2018-10-11 — tap for who they are
MC
Ming-Wei Chang — holds since 2018-10-11 — tap for who they are
KL
Kenton Lee — holds since 2018-10-11 — tap for who they are
JD
Jacob Devlin — holds since 2018-10-11 — tap for who they are
Same subject: Vector embeddings will not solve search: a decades-old term-frequency algorithm still beats most of them at ranking. — tap to centre the map on it
Vector embeddings will not solve search: a decades-old term-frequency algorithm still beats most of them at ranking.
Last stated 2 years ago
19 Jun 2024
AS
Aravind Srinivas — holds since 2024-06-19 — tap for who they are
Same subject: Once a language model reads the results, search can trade precision for recall, because the model does not care that the right link came ninth. — tap to centre the map on it
Once a language model reads the results, search can trade precision for recall, because the model does not care that the right link came ninth.
Last stated 2 years ago
19 Jun 2024
AS
Aravind Srinivas — holds since 2024-06-19 — tap for who they are
Same subject: Large language models could still plateau, and that possibility should be held open even though no evidence of it has appeared. — tap to centre the map on it
Large language models could still plateau, and that possibility should be held open even though no evidence of it has appeared.
Last stated 4 weeks ago
26 Aug 2026
DH
David Heinemeier Hansson — holds since 2026-08-26 — tap for who they are
Same subject: Text embeddings produced by different large language models converge on a shared, universal geometric structure regardless of architecture or training data. — tap to centre the map on it
Text embeddings produced by different large language models converge on a shared, universal geometric structure regardless of architecture or training data.
Last stated a year ago
5 Jun 2025
SH
Samuel Hammond — holds since 2025-06-05 — tap for who they are
Same subject: A true artificial general intelligence cannot exist without being recognized as a moral subject. — tap to centre the map on it
A true artificial general intelligence cannot exist without being recognized as a moral subject.
Last stated a year ago
10 Jun 2025
SH
Samuel Hammond — holds since 2025-06-10 — tap for who they are
Same subject: A unit of AI inference needs to be defined, for example via a chain of increasingly hard problems where each consecutive pair is solvable by one model. — tap to centre the map on it
A unit of AI inference needs to be defined, for example via a chain of increasingly hard problems where each consecutive pair is solvable by one model.
Last stated 6 days ago
15 Sept 2026
PG
Paul Graham — holds since 2026-09-15 — tap for who they are
Same subject: Both the special-purpose-programs view and the blank-slate view of human intelligence are likely incorrect. — tap to centre the map on it
Both the special-purpose-programs view and the blank-slate view of human intelligence are likely incorrect.
Last stated 7 years ago
5 Nov 2019
FC
François Chollet — holds since 2019-11-05 — tap for who they are
Same subject: A physical AI company needs three AIs, not one — the agent, the simulator and the critic — turning deployment into a flywheel. — tap to centre the map on it
A physical AI company needs three AIs, not one — the agent, the simulator and the critic — turning deployment into a flywheel.
Last stated 2 months ago
3 Aug 2026
DD
Dmitri Dolgov — holds since 2026-08-03 — tap for who they are
Same subject: A verifiable task can be optimised by reinforcement learning until a neural network performs it extremely well. — tap to centre the map on it
A verifiable task can be optimised by reinforcement learning until a neural network performs it extremely well.
Last stated 10 months ago
17 Nov 2025
AK
Andrej Karpathy — holds since 2025-11-17 — tap for who they are
Same subject: Building a realistic simulator is exactly as hard as building the agent, because the simulator is itself a large AI model. — tap to centre the map on it
Building a realistic simulator is exactly as hard as building the agent, because the simulator is itself a large AI model.
Last stated 2 months ago
3 Aug 2026
DD
Dmitri Dolgov — holds since 2026-08-03 — tap for who they are
Same subject: A benchmark that ranks Claude Code last while it stays first in use is measuring the wrong thing, and has been for a year. — tap to centre the map on it
A benchmark that ranks Claude Code last while it stays first in use is measuring the wrong thing, and has been for a year.
Last stated 3 weeks ago
3 Sept 2026
DR
Dax Raad — holds since 2026-09-03 — tap for who they are
Same subject: A carmaker's claim to be the safest is mostly an artefact of comparing a new car against a fleet average twelve years old. — tap to centre the map on it
A carmaker's claim to be the safest is mostly an artefact of comparing a new car against a fleet average twelve years old.
Last stated 3 years ago
19 Dec 2023
PK
Philip Koopman — holds since 2023-12-19 — tap for who they are
Same subject: A company's staff-engineer bar should be set against the best companies in the industry rather than against its own history, which is what makes title inflation a real cost. — tap to centre the map on it
A company's staff-engineer bar should be set against the best companies in the industry rather than against its own history, which is what makes title inflation a real cost.
Last stated 6 months ago
1 Apr 2026
TP
Thuan Pham — holds since 2026-04-01 — tap for who they are
same subject or similar wording a cloud: claims about one subject, named for it bar: when it was last stated, on a scale from 2015 to today — full is today a face: someone on record holding the claim — tap it for who they are
At the centre
Document similarity scores across a corpus tend to cluster bimodally near identical and near unrelated, rather than spreading evenly.
Last stated 3 Jul 2024 · 2 years ago
Holds Nelson Elhage
Read this korrent →
Similar wording
The next-token language modeling objective is misaligned with following user instructions helpfully and safely.
Last stated 4 Mar 2022 · 5 years ago
Holds Jan Leike
Similar wording
A neural network's latent space is closer to an uncopyrightable syntax than to copyrightable expression.
Last stated 7 Sept 2026 · 2 weeks ago
Holds Kevin Kelly
Similar wording
Current techniques restrict pre-trained representation power because standard language models are unidirectional.
Last stated 11 Oct 2018 · 8 years ago
Holds Kristina Toutanova Ming-Wei Chang Kenton Lee Jacob Devlin
Similar wording
Vector embeddings will not solve search: a decades-old term-frequency algorithm still beats most of them at ranking.
Last stated 19 Jun 2024 · 2 years ago
Holds Aravind Srinivas
Similar wording
Once a language model reads the results, search can trade precision for recall, because the model does not care that the right link came ninth.
Last stated 19 Jun 2024 · 2 years ago
Holds Aravind Srinivas
Similar wording
Large language models could still plateau, and that possibility should be held open even though no evidence of it has appeared.
Last stated 26 Aug 2026 · 4 weeks ago
Holds David Heinemeier Hansson
Similar wording
Text embeddings produced by different large language models converge on a shared, universal geometric structure regardless of architecture or training data.
Last stated 5 Jun 2025 · a year ago
Holds Samuel Hammond
Same subject: measuring intelligence
A true artificial general intelligence cannot exist without being recognized as a moral subject.
Last stated 10 Jun 2025 · a year ago
Holds Samuel Hammond
Same subject: measuring intelligence
A unit of AI inference needs to be defined, for example via a chain of increasingly hard problems where each consecutive pair is solvable by one model.
Last stated 15 Sept 2026 · 6 days ago
Holds Paul Graham
Same subject: measuring intelligence
Both the special-purpose-programs view and the blank-slate view of human intelligence are likely incorrect.
Last stated 5 Nov 2019 · 7 years ago
Holds François Chollet
Same subject: reinforcement learning
A physical AI company needs three AIs, not one — the agent, the simulator and the critic — turning deployment into a flywheel.
Last stated 3 Aug 2026 · 2 months ago
Holds Dmitri Dolgov
Same subject: reinforcement learning
A verifiable task can be optimised by reinforcement learning until a neural network performs it extremely well.
Last stated 17 Nov 2025 · 10 months ago
Holds Andrej Karpathy
Same subject: reinforcement learning
Building a realistic simulator is exactly as hard as building the agent, because the simulator is itself a large AI model.
Last stated 3 Aug 2026 · 2 months ago
Holds Dmitri Dolgov
Same subject: benchmarks
A benchmark that ranks Claude Code last while it stays first in use is measuring the wrong thing, and has been for a year.
Last stated 3 Sept 2026 · 3 weeks ago
Holds Dax Raad
Same subject: benchmarks
A carmaker's claim to be the safest is mostly an artefact of comparing a new car against a fleet average twelve years old.
Last stated 19 Dec 2023 · 3 years ago
Holds Philip Koopman
Same subject: benchmarks
A company's staff-engineer bar should be set against the best companies in the industry rather than against its own history, which is what makes title inflation a real cost.
Last stated 1 Apr 2026 · 6 months ago
Holds Thuan Pham