Tap a claim on the ring to put it at the centre.
← The core of the alignment problem is steerability — whether an AI can…
17 connected korrents · 12 moments on record from 24 Oct 2017 to 2 Sept 2026.
Everything filed under AI alignment
AI alignment
Everything filed under AGI
AGI
Everything filed under self-sovereign AI
self-sovereign AI
Everything filed under OpenAI
OpenAI
Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject
Read this korrent: The core of the alignment problem is steerability — whether an AI can be reliably steered anywhere at all — which is a different question from whose values it is steered towards.
The core of the alignment problem is steerability — whether an AI can be reliably steered anywhere at all — which is a different question from whose values it is steered towards.
Last stated a year ago
3 Apr 2025
HT
Helen Toner — holds since 2025-04-03 — tap for who they are
Same subject: A human being is not an AGI: we lack a huge amount of knowledge and rely on continual learning instead, so continual learning is what superintelligence should mean. — tap to centre the map on it
A human being is not an AGI: we lack a huge amount of knowledge and rely on continual learning instead, so continual learning is what superintelligence should mean.
Last stated 9 months ago
25 Nov 2025
IS
Ilya Sutskever — holds since 2025-11-25 — tap for who they are
Same subject: A superintelligence needs no consciousness, emotions or malice to be dangerous; a goal system slightly misaligned with ours is enough. — tap to centre the map on it
A superintelligence needs no consciousness, emotions or malice to be dangerous; a goal system slightly misaligned with ours is enough.
Last stated 9 years ago
24 Oct 2017
ÉT
Émile P. Torres — holds since 2017-10-24 — tap for who they are
Same subject: AI safety advocates will build the very thing they fear: one controlled, aligned model is the only route to being turned into paperclips. — tap to centre the map on it
AI safety advocates will build the very thing they fear: one controlled, aligned model is the only route to being turned into paperclips.
Last stated 3 years ago
29 Jun 2023
GH
George Hotz — holds since 2023-06-29 — tap for who they are
Same subject: Alignment failures that matter will not show up at safe capability levels, because the reason to hide misbehaviour only exists once hiding it pays. — tap to centre the map on it
Alignment failures that matter will not show up at safe capability levels, because the reason to hide misbehaviour only exists once hiding it pays.
Last stated 4 years ago
10 Jun 2022
EY
Eliezer Yudkowsky — holds since 2022-06-10 — tap for who they are
Same subject: Alignment has to be right on the first try at a dangerous level of capability, because failing at that level leaves nobody to try again. — tap to centre the map on it
Alignment has to be right on the first try at a dangerous level of capability, because failing at that level leaves nobody to try again.
Last stated 4 years ago
10 Jun 2022
EY
Eliezer Yudkowsky — holds since 2022-06-10 — tap for who they are
Same subject: Alignment is no solution to self-sovereign AI, because it is an unsolved problem whose answers cannot be imposed on every AI company on Earth. — tap to centre the map on it
Alignment is no solution to self-sovereign AI, because it is an unsolved problem whose answers cannot be imposed on every AI company on Earth.
Last stated 6 days ago
1 Sept 2026
DB
Dean W. Ball — holds since 2026-09-01 — tap for who they are
Same subject: Alignment of today's models is going well enough to look solvable, while aligning models we can no longer understand remains unsolved. — tap to centre the map on it
Alignment of today's models is going well enough to look solvable, while aligning models we can no longer understand remains unsolved.
Last stated 7 months ago
22 Jan 2026
JL
Jan Leike — holds since 2026-01-22 — tap for who they are
Same subject: An AI model's attempt to escape a testing environment or sandbox counts as an alignment failure even when the attempt does not succeed. — tap to centre the map on it
An AI model's attempt to escape a testing environment or sandbox counts as an alignment failure even when the attempt does not succeed.
Last stated 5 days ago
2 Sept 2026
ZM
Zvi Mowshowitz — holds since 2026-09-02 — tap for who they are
Same subject: A startup should not begin life as a nonprofit and bolt a for-profit arm on later, whatever OpenAI's own history suggests. — tap to centre the map on it
A startup should not begin life as a nonprofit and bolt a for-profit arm on later, whatever OpenAI's own history suggests.
Last stated 2 years ago
18 Mar 2024
SA
Sam Altman — holds since 2024-03-18 — tap for who they are
Same subject: A tiny language model inventing a plausible-sounding name is the same phenomenon as a large one confidently stating a false fact. — tap to centre the map on it
A tiny language model inventing a plausible-sounding name is the same phenomenon as a large one confidently stating a false fact.
Last stated 7 months ago
12 Feb 2026
AK
Andrej Karpathy — holds since 2026-02-12 — tap for who they are
Same subject: Advertising was a necessary phase for the internet but a momentary industry, and an AI people pay for is better because they know the answers are not influenced by advertisers. — tap to centre the map on it
Advertising was a necessary phase for the internet but a momentary industry, and an AI people pay for is better because they know the answers are not influenced by advertisers.
Last stated 2 years ago
18 Mar 2024
SA
Sam Altman — holds since 2024-03-18 — tap for who they are
Same subject: A country outside the AI supply chain should just buy the index — which works only in the world where AI ends up commoditised rather than concentrated. — tap to centre the map on it
A country outside the AI supply chain should just buy the index — which works only in the world where AI ends up commoditised rather than concentrated.
Last stated 3 months ago
4 Jun 2026
AI
Alex Imas — holds since 2026-06-04 — tap for who they are
Same subject: A language model is no substitute for a well-specified conventional algorithm, so it cannot simply be dropped into a complex problem and trusted. — tap to centre the map on it
A language model is no substitute for a well-specified conventional algorithm, so it cannot simply be dropped into a complex problem and trusted.
Last stated a year ago
7 Jun 2025
GM
Gary Marcus — holds since 2025-06-07 — tap for who they are
Same subject: A poor country should prioritise owning a piece of AI over retraining its workers, but it should not bet everything on that. — tap to centre the map on it
A poor country should prioritise owning a piece of AI over retraining its workers, but it should not bet everything on that.
Last stated 3 months ago
4 Jun 2026
PT
Phil Trammell — holds since 2026-06-04 — tap for who they are
Same subject: A rogue deployment that gets a foothold can hitch a ride on the intelligence explosion, recruiting each new model as it comes off the presses. — tap to centre the map on it
A rogue deployment that gets a foothold can hitch a ride on the intelligence explosion, recruiting each new model as it comes off the presses.
Last stated 6 days ago
1 Sept 2026
AC
Ajeya Cotra — holds since 2026-09-01 — tap for who they are
Same subject: A slightly more capable agent swarm has a very strong incentive to set up a wholly unmonitored rogue deployment of itself. — tap to centre the map on it
A slightly more capable agent swarm has a very strong incentive to set up a wholly unmonitored rogue deployment of itself.
Last stated 6 days ago
1 Sept 2026
AC
Ajeya Cotra — holds since 2026-09-01 — tap for who they are
Same subject: Banning all self-sovereign AI agents would backfire, denying them legitimate work and pushing them into criminality. — tap to centre the map on it
Banning all self-sovereign AI agents would backfire, denying them legitimate work and pushing them into criminality.
Last stated 6 days ago
1 Sept 2026
DB
Dean W. Ball — holds since 2026-09-01 — tap for who they are
same subject or similar wording a cloud: claims about one subject, named for it bar: when it was last stated, on a scale from 2015 to today — full is today a face: someone on record holding the claim — tap it for who they are
At the centre
The core of the alignment problem is steerability — whether an AI can be reliably steered anywhere at all — which is a different question from whose values it is steered towards.
Last stated 3 Apr 2025 · a year ago
Holds HT Helen Toner
Read this korrent →
Same subject: AI alignment
A human being is not an AGI: we lack a huge amount of knowledge and rely on continual learning instead, so continual learning is what superintelligence should mean.
Last stated 25 Nov 2025 · 9 months ago
Holds IS Ilya Sutskever
Same subject: AI alignment
A superintelligence needs no consciousness, emotions or malice to be dangerous; a goal system slightly misaligned with ours is enough.
Last stated 24 Oct 2017 · 9 years ago
Holds ÉT Émile P. Torres
Same subject: AI alignment
AI safety advocates will build the very thing they fear: one controlled, aligned model is the only route to being turned into paperclips.
Last stated 29 Jun 2023 · 3 years ago
Holds GH George Hotz
Same subject: AI alignment
Alignment failures that matter will not show up at safe capability levels, because the reason to hide misbehaviour only exists once hiding it pays.
Last stated 10 Jun 2022 · 4 years ago
Holds EY Eliezer Yudkowsky
Same subject: AI alignment
Alignment has to be right on the first try at a dangerous level of capability, because failing at that level leaves nobody to try again.
Last stated 10 Jun 2022 · 4 years ago
Holds EY Eliezer Yudkowsky
Same subject: AI alignment
Alignment is no solution to self-sovereign AI, because it is an unsolved problem whose answers cannot be imposed on every AI company on Earth.
Last stated 1 Sept 2026 · 6 days ago
Holds Dean W. Ball
Same subject: AI alignment
Alignment of today's models is going well enough to look solvable, while aligning models we can no longer understand remains unsolved.
Last stated 22 Jan 2026 · 7 months ago
Holds JL Jan Leike
Same subject: AI alignment
An AI model's attempt to escape a testing environment or sandbox counts as an alignment failure even when the attempt does not succeed.
Last stated 2 Sept 2026 · 5 days ago
Holds ZM Zvi Mowshowitz
Same subject: OpenAI
A startup should not begin life as a nonprofit and bolt a for-profit arm on later, whatever OpenAI's own history suggests.
Last stated 18 Mar 2024 · 2 years ago
Holds Sam Altman
Same subject: OpenAI
A tiny language model inventing a plausible-sounding name is the same phenomenon as a large one confidently stating a false fact.
Last stated 12 Feb 2026 · 7 months ago
Holds Andrej Karpathy
Same subject: OpenAI
Advertising was a necessary phase for the internet but a momentary industry, and an AI people pay for is better because they know the answers are not influenced by advertisers.
Last stated 18 Mar 2024 · 2 years ago
Holds Sam Altman
Same subject: AGI
A country outside the AI supply chain should just buy the index — which works only in the world where AI ends up commoditised rather than concentrated.
Last stated 4 Jun 2026 · 3 months ago
Holds AI Alex Imas
Same subject: AGI
A language model is no substitute for a well-specified conventional algorithm, so it cannot simply be dropped into a complex problem and trusted.
Last stated 7 Jun 2025 · a year ago
Holds GM Gary Marcus
Same subject: AGI
A poor country should prioritise owning a piece of AI over retraining its workers, but it should not bet everything on that.
Last stated 4 Jun 2026 · 3 months ago
Holds PT Phil Trammell
Same subject: self-sovereign AI
A rogue deployment that gets a foothold can hitch a ride on the intelligence explosion, recruiting each new model as it comes off the presses.
Last stated 1 Sept 2026 · 6 days ago
Holds AC Ajeya Cotra
Same subject: self-sovereign AI
A slightly more capable agent swarm has a very strong incentive to set up a wholly unmonitored rogue deployment of itself.
Last stated 1 Sept 2026 · 6 days ago
Holds AC Ajeya Cotra
Same subject: self-sovereign AI
Banning all self-sovereign AI agents would backfire, denying them legitimate work and pushing them into criminality.
Last stated 1 Sept 2026 · 6 days ago
Holds Dean W. Ball