Tap a claim on the ring to put it at the centre.
← It is not enough for an AI model to be smart: aligning it to act…
17 connected korrents · 16 moments from 10 Jun 2022 to 19 Sept 2026. Nearly all of them are about AI alignment .
Everything filed under AI agents
AI agents
Everything filed under OpenAI
OpenAI
Everything filed under reinforcement learning
reinforcement learning
Everything filed under AGI
AGI
Everything filed under LLMs
LLMs
Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject
Read this korrent: It is not enough for an AI model to be smart: aligning it to act "correctly" according to its model spec is incredibly important.
It is not enough for an AI model to be smart: aligning it to act "correctly" according to its model spec is incredibly important.
Last stated 3 months ago
by 24 Jul 2026
LR
Lee Robinson — holds since 2026-07-24 — tap for who they are
Same subject: Building AI that can actually be trusted is the goal; containing an AI known to be misaligned is not a substitute for it. — tap to centre the map on it
Building AI that can actually be trusted is the goal; containing an AI known to be misaligned is not a substitute for it.
Last stated 2 years ago
24 Jan 2025
JL
Jan Leike — holds since 2025-01-24 — tap for who they are
Same subject: No observation of a model's behaviour can establish that it is aligned, because a model capable enough to be dangerous is capable enough to produce whatever behaviour it is being watched for. — tap to centre the map on it
No observation of a model's behaviour can establish that it is aligned, because a model capable enough to be dangerous is capable enough to produce whatever behaviour it is being watched for.
Last stated 2 years ago
7 May 2024
BS
Buck Shlegeris — holds since 2024-05-07 — tap for who they are
EY
Eliezer Yudkowsky — holds since 2022-06-10 — tap for who they are
Same subject: A validated theory of intelligence may be necessary for achieving genuine AI alignment. — tap to centre the map on it
A validated theory of intelligence may be necessary for achieving genuine AI alignment.
Last stated 2 months ago
17 Aug 2026
KK
Kevin Kelly — holds since 2026-08-17 — tap for who they are
Same subject: Alignment of today's models is going well enough to look solvable, while aligning models we can no longer understand remains unsolved. — tap to centre the map on it
Alignment of today's models is going well enough to look solvable, while aligning models we can no longer understand remains unsolved.
Last stated 9 months ago
22 Jan 2026
JL
Jan Leike — holds since 2026-01-22 — tap for who they are
Same subject: Very capable AI will be harder to align than current systems, because the loop of spotting a bad behaviour and patching the training that caused it breaks down. — tap to centre the map on it
Very capable AI will be harder to align than current systems, because the loop of spotting a bad behaviour and patching the training that caused it breaks down.
Last stated 2 months ago
11 Aug 2026
RG
Ryan Greenblatt — holds since 2026-08-11 — tap for who they are
Same subject: AI models considered smart today will be considered very dumb in the future, and smarter models make fewer mistakes and need fewer turns. — tap to centre the map on it
AI models considered smart today will be considered very dumb in the future, and smarter models make fewer mistakes and need fewer turns.
Last stated 3 weeks ago
19 Sept 2026
TB
Thorsten Ball — holds since 2026-09-19 — tap for who they are
Same subject: We do not have to align superintelligence directly; we have to build a human-level automated alignment researcher we trust more than ourselves. — tap to centre the map on it
We do not have to align superintelligence directly; we have to build a human-level automated alignment researcher we trust more than ourselves.
Last stated 9 months ago
22 Jan 2026
JL
Jan Leike — holds since 2026-01-22 — tap for who they are
Same subject: AI alignment is a red herring not worth the effort currently devoted to it. — tap to centre the map on it
AI alignment is a red herring not worth the effort currently devoted to it.
Last stated 2 months ago
8 Aug 2026
MW
Matt Webb — holds since 2026-08-08 — tap for who they are
Same subject: A startup should not begin life as a nonprofit and bolt a for-profit arm on later, whatever OpenAI's own history suggests. — tap to centre the map on it
A startup should not begin life as a nonprofit and bolt a for-profit arm on later, whatever OpenAI's own history suggests.
Last stated 3 years ago
18 Mar 2024
SA
Sam Altman — holds since 2024-03-18 — tap for who they are
Same subject: Advertising was a necessary phase for the internet but a momentary industry, and an AI people pay for is better because they know the answers are not influenced by advertisers. — tap to centre the map on it
Advertising was a necessary phase for the internet but a momentary industry, and an AI people pay for is better because they know the answers are not influenced by advertisers.
Last stated 3 years ago
18 Mar 2024
SA
Sam Altman — holds since 2024-03-18 — tap for who they are
Same subject: Agents are going to be incorporated into mainstream interfaces — tap to centre the map on it
Agents are going to be incorporated into mainstream interfaces
Last stated 7 months ago
25 Mar 2026
DH
David Heinemeier Hansson — holds since 2026-03-25 — tap for who they are
Same subject: A collection of AI agents does not automatically form a functioning organization any more than a group of smart people does. — tap to centre the map on it
A collection of AI agents does not automatically form a functioning organization any more than a group of smart people does.
Last stated 5 months ago
21 May 2026
RK
Rohit Krishnan — holds since 2026-05-21 — tap for who they are
Same subject: A company cannot be led by machines, because nobody would have recourse against them. — tap to centre the map on it
A company cannot be led by machines, because nobody would have recourse against them.
Last stated 4 weeks ago
15 Sept 2026
TL
Tobias Lütke — holds since 2026-09-15 — tap for who they are
Same subject: A computer-using agent needs a machine of its own, or you will spend the day fighting it for the mouse cursor. — tap to centre the map on it
A computer-using agent needs a machine of its own, or you will spend the day fighting it for the mouse cursor.
Last stated 2 months ago
10 Aug 2026
PS
Peter Steinberger — holds since 2026-08-10 — tap for who they are
Same subject: A physical AI company needs three AIs, not one — the agent, the simulator and the critic — turning deployment into a flywheel. — tap to centre the map on it
A physical AI company needs three AIs, not one — the agent, the simulator and the critic — turning deployment into a flywheel.
Last stated 2 months ago
3 Aug 2026
DD
Dmitri Dolgov — holds since 2026-08-03 — tap for who they are
Same subject: A verifiable task can be optimised by reinforcement learning until a neural network performs it extremely well. — tap to centre the map on it
A verifiable task can be optimised by reinforcement learning until a neural network performs it extremely well.
Last stated 11 months ago
17 Nov 2025
AK
Andrej Karpathy — holds since 2025-11-17 — tap for who they are
Same subject: Building a realistic simulator is exactly as hard as building the agent, because the simulator is itself a large AI model. — tap to centre the map on it
Building a realistic simulator is exactly as hard as building the agent, because the simulator is itself a large AI model.
Last stated 2 months ago
3 Aug 2026
DD
Dmitri Dolgov — holds since 2026-08-03 — tap for who they are
same subject or similar wording a cloud: claims about one subject, named for it bar: when it was last stated, on a scale from 2015 to today — full is today a face: someone who holds the claim — tap it for who they are
At the centre
It is not enough for an AI model to be smart: aligning it to act "correctly" according to its model spec is incredibly important.
Last stated by 24 Jul 2026 · 3 months ago
Holds Lee Robinson
Read this korrent →
Similar wording
Building AI that can actually be trusted is the goal; containing an AI known to be misaligned is not a substitute for it.
Last stated 24 Jan 2025 · 2 years ago
Holds Jan Leike
Similar wording
No observation of a model's behaviour can establish that it is aligned, because a model capable enough to be dangerous is capable enough to produce whatever behaviour it is being watched for.
Last stated 7 May 2024 · 2 years ago
Holds Buck Shlegeris Eliezer Yudkowsky
Similar wording
A validated theory of intelligence may be necessary for achieving genuine AI alignment.
Last stated 17 Aug 2026 · 2 months ago
Holds Kevin Kelly
Similar wording
Alignment of today's models is going well enough to look solvable, while aligning models we can no longer understand remains unsolved.
Last stated 22 Jan 2026 · 9 months ago
Holds Jan Leike
Similar wording
Very capable AI will be harder to align than current systems, because the loop of spotting a bad behaviour and patching the training that caused it breaks down.
Last stated 11 Aug 2026 · 2 months ago
Holds Ryan Greenblatt
Similar wording
AI models considered smart today will be considered very dumb in the future, and smarter models make fewer mistakes and need fewer turns.
Last stated 19 Sept 2026 · 3 weeks ago
Holds Thorsten Ball
Similar wording
We do not have to align superintelligence directly; we have to build a human-level automated alignment researcher we trust more than ourselves.
Last stated 22 Jan 2026 · 9 months ago
Holds Jan Leike
Similar wording
AI alignment is a red herring not worth the effort currently devoted to it.
Last stated 8 Aug 2026 · 2 months ago
Holds Matt Webb
Same subject: OpenAI
A startup should not begin life as a nonprofit and bolt a for-profit arm on later, whatever OpenAI's own history suggests.
Last stated 18 Mar 2024 · 3 years ago
Holds Sam Altman
Same subject: OpenAI
Advertising was a necessary phase for the internet but a momentary industry, and an AI people pay for is better because they know the answers are not influenced by advertisers.
Last stated 18 Mar 2024 · 3 years ago
Holds Sam Altman
Same subject: OpenAI
Agents are going to be incorporated into mainstream interfaces
Last stated 25 Mar 2026 · 7 months ago
Holds David Heinemeier Hansson
Same subject: AI agents
A collection of AI agents does not automatically form a functioning organization any more than a group of smart people does.
Last stated 21 May 2026 · 5 months ago
Holds Rohit Krishnan
Same subject: AI agents
A company cannot be led by machines, because nobody would have recourse against them.
Last stated 15 Sept 2026 · 4 weeks ago
Holds Tobias Lütke
Same subject: AI agents
A computer-using agent needs a machine of its own, or you will spend the day fighting it for the mouse cursor.
Last stated 10 Aug 2026 · 2 months ago
Holds Peter Steinberger
Same subject: reinforcement learning
A physical AI company needs three AIs, not one — the agent, the simulator and the critic — turning deployment into a flywheel.
Last stated 3 Aug 2026 · 2 months ago
Holds Dmitri Dolgov
Same subject: reinforcement learning
A verifiable task can be optimised by reinforcement learning until a neural network performs it extremely well.
Last stated 17 Nov 2025 · 11 months ago
Holds Andrej Karpathy
Same subject: reinforcement learning
Building a realistic simulator is exactly as hard as building the agent, because the simulator is itself a large AI model.
Last stated 3 Aug 2026 · 2 months ago
Holds Dmitri Dolgov