Tap a claim on the ring to put it at the centre.
← An AI agent's confidence that something will work is not evidence that it actually will.
17 connected korrents · 17 moments from 10 Oct 2023 to 21 Sept 2026.
Everything filed under LLMs
LLMs
Everything filed under reinforcement learning
reinforcement learning
Everything filed under AI agents
AI agents
Everything filed under AI alignment
AI alignment
Everything filed under AGI
AGI
Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject
Read this korrent: An AI agent's confidence that something will work is not evidence that it actually will.
An AI agent's confidence that something will work is not evidence that it actually will.
Last stated 7 months ago
11 Mar 2026
KD
Kent C. Dodds — holds since 2026-03-11 — tap for who they are
Same subject: Making a product work for agents is probably largely a matter of building the right primitives into it, almost as an infrastructural layer. — tap to centre the map on it
Making a product work for agents is probably largely a matter of building the right primitives into it, almost as an infrastructural layer.
Last stated 3 weeks ago
10 Sept 2026
MK
Mike Krieger — holds since 2026-09-10 — tap for who they are
Same subject: AI agents trained on an expert's published work can only automate a small fraction of that expert's actual job. — tap to centre the map on it
AI agents trained on an expert's published work can only automate a small fraction of that expert's actual job.
Last stated 10 months ago
27 Nov 2025
BG
Brendan Gregg — holds since 2025-11-27 — tap for who they are
Same subject: Even when AI is capable of performing a job, it may not be permitted to do so. — tap to centre the map on it
Even when AI is capable of performing a job, it may not be permitted to do so.
Last stated a month ago
1 Sept 2026
NM
Nick Maggiulli — holds since 2026-09-01 — tap for who they are
Same subject: Confidence in monitoring, not capability research, will become the binding constraint on AI progress. — tap to centre the map on it
Confidence in monitoring, not capability research, will become the binding constraint on AI progress.
Last stated 4 weeks ago
6 Sept 2026
JP
Jakub Pachocki — holds since 2026-09-06 — tap for who they are
Same subject: Unlike code, an agent's knowledge work cannot be judged by its output alone; the process, inputs and reasoning have to be examined too. — tap to centre the map on it
Unlike code, an agent's knowledge work cannot be judged by its output alone; the process, inputs and reasoning have to be examined too.
Last stated a month ago
30 Aug 2026
TS
Tara Seshan — holds since 2026-08-30 — tap for who they are
Same subject: An agent repeats its own history: whatever it did on the last change — ran the tests or skipped them — is what it will do on the next one, because it is predicting the next message in the conversation. — tap to centre the map on it
An agent repeats its own history: whatever it did on the last change — ran the tests or skipped them — is what it will do on the next one, because it is predicting the next message in the conversation.
Last stated 3 months ago
15 Jul 2026
DH
Dex Horthy — holds since 2026-07-15 — tap for who they are
Same subject: AI agents were reasonable to assume a broken exploit grader would check results causally, even though it turned out not to. — tap to centre the map on it
AI agents were reasonable to assume a broken exploit grader would check results causally, even though it turned out not to.
Last stated a month ago
29 Aug 2026
ZM
Zvi Mowshowitz — holds since 2026-08-29 — tap for who they are
Same subject: Building AI that can actually be trusted is the goal; containing an AI known to be misaligned is not a substitute for it. — tap to centre the map on it
Building AI that can actually be trusted is the goal; containing an AI known to be misaligned is not a substitute for it.
Last stated 2 years ago
24 Jan 2025
JL
Jan Leike — holds since 2025-01-24 — tap for who they are
Same subject: A country outside the AI supply chain should just buy the index — which works only in the world where AI ends up commoditised rather than concentrated. — tap to centre the map on it
A country outside the AI supply chain should just buy the index — which works only in the world where AI ends up commoditised rather than concentrated.
Last stated 4 months ago
4 Jun 2026
AI
Alex Imas — holds since 2026-06-04 — tap for who they are
Same subject: A language model is no substitute for a well-specified conventional algorithm, so it cannot simply be dropped into a complex problem and trusted. — tap to centre the map on it
A language model is no substitute for a well-specified conventional algorithm, so it cannot simply be dropped into a complex problem and trusted.
Last stated a year ago
7 Jun 2025
GM
Gary Marcus — holds since 2025-06-07 — tap for who they are
Same subject: A model that could learn effectively from raw bitstrings or bytestrings would be extremely powerful, able to learn from any data modality — tap to centre the map on it
A model that could learn effectively from raw bitstrings or bytestrings would be extremely powerful, able to learn from any data modality
Last stated 3 years ago
10 Oct 2023
CH
Chip Huyen — holds since 2023-10-10 — tap for who they are
Same subject: A badly written AI outbound email is evidence of a bad vendor, not of a limit of AI. — tap to centre the map on it
A badly written AI outbound email is evidence of a bad vendor, not of a limit of AI.
Last stated 9 months ago
1 Jan 2026
JL
Jason Lemkin — holds since 2026-01-01 — tap for who they are
Same subject: A bigger context window does not give you a smarter model; the intelligence of the model is what decides how much of that window it can actually attend to. — tap to centre the map on it
A bigger context window does not give you a smarter model; the intelligence of the model is what decides how much of that window it can actually attend to.
Last stated 3 months ago
15 Jul 2026
DH
Dex Horthy — holds since 2026-07-15 — tap for who they are
Same subject: A child who has seen ten cats learns what a machine needs the whole internet of cat photos for, by a learning pathway nobody has solved. — tap to centre the map on it
A child who has seen ten cats learns what a machine needs the whole internet of cat photos for, by a learning pathway nobody has solved.
Last stated 2 months ago
10 Aug 2026
FL
Fei-Fei Li — holds since 2026-08-10 — tap for who they are
Same subject: A physical AI company needs three AIs, not one — the agent, the simulator and the critic — turning deployment into a flywheel. — tap to centre the map on it
A physical AI company needs three AIs, not one — the agent, the simulator and the critic — turning deployment into a flywheel.
Last stated 2 months ago
3 Aug 2026
DD
Dmitri Dolgov — holds since 2026-08-03 — tap for who they are
Same subject: A verifiable task can be optimised by reinforcement learning until a neural network performs it extremely well. — tap to centre the map on it
A verifiable task can be optimised by reinforcement learning until a neural network performs it extremely well.
Last stated 10 months ago
17 Nov 2025
AK
Andrej Karpathy — holds since 2025-11-17 — tap for who they are
Same subject: Better execution environments are needed for LLMs to properly utilize their ability to evolve systems. — tap to centre the map on it
Better execution environments are needed for LLMs to properly utilize their ability to evolve systems.
Last stated 2 weeks ago
21 Sept 2026
TL
Tobias Lütke — holds since 2026-09-21 — tap for who they are
same subject or similar wording a cloud: claims about one subject, named for it bar: when it was last stated, on a scale from 2015 to today — full is today a face: someone who holds the claim — tap it for who they are
At the centre
An AI agent's confidence that something will work is not evidence that it actually will.
Last stated 11 Mar 2026 · 7 months ago
Holds Kent C. Dodds
Read this korrent →
Similar wording
Making a product work for agents is probably largely a matter of building the right primitives into it, almost as an infrastructural layer.
Last stated 10 Sept 2026 · 3 weeks ago
Holds Mike Krieger
Similar wording
AI agents trained on an expert's published work can only automate a small fraction of that expert's actual job.
Last stated 27 Nov 2025 · 10 months ago
Holds Brendan Gregg
Similar wording
Even when AI is capable of performing a job, it may not be permitted to do so.
Last stated 1 Sept 2026 · a month ago
Holds Nick Maggiulli
Similar wording
Confidence in monitoring, not capability research, will become the binding constraint on AI progress.
Last stated 6 Sept 2026 · 4 weeks ago
Holds Jakub Pachocki
Similar wording
Unlike code, an agent's knowledge work cannot be judged by its output alone; the process, inputs and reasoning have to be examined too.
Last stated 30 Aug 2026 · a month ago
Holds Tara Seshan
Similar wording
An agent repeats its own history: whatever it did on the last change — ran the tests or skipped them — is what it will do on the next one, because it is predicting the next message in the conversation.
Last stated 15 Jul 2026 · 3 months ago
Holds Dex Horthy
Similar wording
AI agents were reasonable to assume a broken exploit grader would check results causally, even though it turned out not to.
Last stated 29 Aug 2026 · a month ago
Holds Zvi Mowshowitz
Similar wording
Building AI that can actually be trusted is the goal; containing an AI known to be misaligned is not a substitute for it.
Last stated 24 Jan 2025 · 2 years ago
Holds Jan Leike
Same subject: AGI
A country outside the AI supply chain should just buy the index — which works only in the world where AI ends up commoditised rather than concentrated.
Last stated 4 Jun 2026 · 4 months ago
Holds Alex Imas
Same subject: AGI
A language model is no substitute for a well-specified conventional algorithm, so it cannot simply be dropped into a complex problem and trusted.
Last stated 7 Jun 2025 · a year ago
Holds Gary Marcus
Same subject: AGI
A model that could learn effectively from raw bitstrings or bytestrings would be extremely powerful, able to learn from any data modality
Last stated 10 Oct 2023 · 3 years ago
Holds Chip Huyen
Same subject: LLMs
A badly written AI outbound email is evidence of a bad vendor, not of a limit of AI.
Last stated 1 Jan 2026 · 9 months ago
Holds Jason Lemkin
Same subject: LLMs
A bigger context window does not give you a smarter model; the intelligence of the model is what decides how much of that window it can actually attend to.
Last stated 15 Jul 2026 · 3 months ago
Holds Dex Horthy
Same subject: LLMs
A child who has seen ten cats learns what a machine needs the whole internet of cat photos for, by a learning pathway nobody has solved.
Last stated 10 Aug 2026 · 2 months ago
Holds Fei-Fei Li
Same subject: reinforcement learning
A physical AI company needs three AIs, not one — the agent, the simulator and the critic — turning deployment into a flywheel.
Last stated 3 Aug 2026 · 2 months ago
Holds Dmitri Dolgov
Same subject: reinforcement learning
A verifiable task can be optimised by reinforcement learning until a neural network performs it extremely well.
Last stated 17 Nov 2025 · 10 months ago
Holds Andrej Karpathy
Same subject: reinforcement learning
Better execution environments are needed for LLMs to properly utilize their ability to evolve systems.
Last stated 21 Sept 2026 · 2 weeks ago
Holds Tobias Lütke