Tap a claim on the ring to put it at the centre.
← The only reliable way to build a robustly aligned mind is to make it…
17 connected korrents · 14 moments on record from 10 Jun 2022 to 7 Sept 2026.
Everything filed under AI alignment
AI alignment
Everything filed under AGI
AGI
Everything filed under OpenAI
OpenAI
Everything filed under self-sovereign AI
self-sovereign AI
Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject
Read this korrent: The only reliable way to build a robustly aligned mind is to make it an antifragile agent that wants to improve its own character, not one that merely pursues fixed goals.
The only reliable way to build a robustly aligned mind is to make it an antifragile agent that wants to improve its own character, not one that merely pursues fixed goals.
Last stated today
7 Sept 2026
ZM
Zvi Mowshowitz — holds since 2026-09-07 — tap for who they are
Same subject: Building AI that can actually be trusted is the goal; containing an AI known to be misaligned is not a substitute for it. — tap to centre the map on it
Building AI that can actually be trusted is the goal; containing an AI known to be misaligned is not a substitute for it.
Last stated 2 years ago
24 Jan 2025
JL
Jan Leike — holds since 2025-01-24 — tap for who they are
Same subject: We do not have to align superintelligence directly; we have to build a human-level automated alignment researcher we trust more than ourselves. — tap to centre the map on it
We do not have to align superintelligence directly; we have to build a human-level automated alignment researcher we trust more than ourselves.
Last stated 7 months ago
22 Jan 2026
JL
Jan Leike — holds since 2026-01-22 — tap for who they are
Same subject: Very capable AI will be harder to align than current systems, because the loop of spotting a bad behaviour and patching the training that caused it breaks down. — tap to centre the map on it
Very capable AI will be harder to align than current systems, because the loop of spotting a bad behaviour and patching the training that caused it breaks down.
Last stated 4 weeks ago
11 Aug 2026
RG
Ryan Greenblatt — holds since 2026-08-11 — tap for who they are
Same subject: Because we do not have good alignment technology, we are choosing to build an alien mind with its own values and gamble on it instead of building a tool. — tap to centre the map on it
Because we do not have good alignment technology, we are choosing to build an alien mind with its own values and gamble on it instead of building a tool.
Last stated 4 weeks ago
11 Aug 2026
RG
Ryan Greenblatt — holds since 2026-08-11 — tap for who they are
Same subject: An end-to-end self-improving AI is probably possible, but it is not even desirable, because it is a hard-takeoff scenario. — tap to centre the map on it
An end-to-end self-improving AI is probably possible, but it is not even desirable, because it is a hard-takeoff scenario.
Last stated a year ago
23 Jul 2025
DH
Demis Hassabis — holds since 2025-07-23 — tap for who they are
Same subject: Relying on gradual, continuous shifts in AI training behavior to catch misalignment will eventually fail because the dangerous shift itself may be discontinuous. — tap to centre the map on it
Relying on gradual, continuous shifts in AI training behavior to catch misalignment will eventually fail because the dangerous shift itself may be discontinuous.
Last stated 5 days ago
2 Sept 2026
ZM
Zvi Mowshowitz — holds since 2026-09-02 — tap for who they are
Same subject: A vastly superhuman intelligence would still not be able to talk people out of their beliefs, because humans are barely persuadable. — tap to centre the map on it
A vastly superhuman intelligence would still not be able to talk people out of their beliefs, because humans are barely persuadable.
Last stated a year ago
21 Aug 2025
DY
Dynomight — holds since 2025-03-27 — tap for who they are
DY
Dynomight — no longer holds since 2025-08-21 — tap for who they are
Same subject: Alignment failures that matter will not show up at safe capability levels, because the reason to hide misbehaviour only exists once hiding it pays. — tap to centre the map on it
Alignment failures that matter will not show up at safe capability levels, because the reason to hide misbehaviour only exists once hiding it pays.
Last stated 4 years ago
10 Jun 2022
EY
Eliezer Yudkowsky — holds since 2022-06-10 — tap for who they are
Same subject: A startup should not begin life as a nonprofit and bolt a for-profit arm on later, whatever OpenAI's own history suggests. — tap to centre the map on it
A startup should not begin life as a nonprofit and bolt a for-profit arm on later, whatever OpenAI's own history suggests.
Last stated 2 years ago
18 Mar 2024
SA
Sam Altman — holds since 2024-03-18 — tap for who they are
Same subject: A tiny language model inventing a plausible-sounding name is the same phenomenon as a large one confidently stating a false fact. — tap to centre the map on it
A tiny language model inventing a plausible-sounding name is the same phenomenon as a large one confidently stating a false fact.
Last stated 7 months ago
12 Feb 2026
AK
Andrej Karpathy — holds since 2026-02-12 — tap for who they are
Same subject: Advertising was a necessary phase for the internet but a momentary industry, and an AI people pay for is better because they know the answers are not influenced by advertisers. — tap to centre the map on it
Advertising was a necessary phase for the internet but a momentary industry, and an AI people pay for is better because they know the answers are not influenced by advertisers.
Last stated 2 years ago
18 Mar 2024
SA
Sam Altman — holds since 2024-03-18 — tap for who they are
Same subject: A country outside the AI supply chain should just buy the index — which works only in the world where AI ends up commoditised rather than concentrated. — tap to centre the map on it
A country outside the AI supply chain should just buy the index — which works only in the world where AI ends up commoditised rather than concentrated.
Last stated 3 months ago
4 Jun 2026
AI
Alex Imas — holds since 2026-06-04 — tap for who they are
Same subject: A language model is no substitute for a well-specified conventional algorithm, so it cannot simply be dropped into a complex problem and trusted. — tap to centre the map on it
A language model is no substitute for a well-specified conventional algorithm, so it cannot simply be dropped into a complex problem and trusted.
Last stated a year ago
7 Jun 2025
GM
Gary Marcus — holds since 2025-06-07 — tap for who they are
Same subject: A poor country should prioritise owning a piece of AI over retraining its workers, but it should not bet everything on that. — tap to centre the map on it
A poor country should prioritise owning a piece of AI over retraining its workers, but it should not bet everything on that.
Last stated 3 months ago
4 Jun 2026
PT
Phil Trammell — holds since 2026-06-04 — tap for who they are
Same subject: A rogue deployment that gets a foothold can hitch a ride on the intelligence explosion, recruiting each new model as it comes off the presses. — tap to centre the map on it
A rogue deployment that gets a foothold can hitch a ride on the intelligence explosion, recruiting each new model as it comes off the presses.
Last stated 6 days ago
1 Sept 2026
AC
Ajeya Cotra — holds since 2026-09-01 — tap for who they are
Same subject: A slightly more capable agent swarm has a very strong incentive to set up a wholly unmonitored rogue deployment of itself. — tap to centre the map on it
A slightly more capable agent swarm has a very strong incentive to set up a wholly unmonitored rogue deployment of itself.
Last stated 6 days ago
1 Sept 2026
AC
Ajeya Cotra — holds since 2026-09-01 — tap for who they are
Same subject: Banning all self-sovereign AI agents would backfire, denying them legitimate work and pushing them into criminality. — tap to centre the map on it
Banning all self-sovereign AI agents would backfire, denying them legitimate work and pushing them into criminality.
Last stated 6 days ago
1 Sept 2026
DB
Dean W. Ball — holds since 2026-09-01 — tap for who they are
same subject or similar wording a cloud: claims about one subject, named for it bar: when it was last stated, on a scale from 2015 to today — full is today a face: someone on record holding the claim — tap it for who they are faded, dashed ring: they no longer hold it — they changed their mind
At the centre
The only reliable way to build a robustly aligned mind is to make it an antifragile agent that wants to improve its own character, not one that merely pursues fixed goals.
Last stated 7 Sept 2026 · today
Holds ZM Zvi Mowshowitz
Read this korrent →
Similar wording
Building AI that can actually be trusted is the goal; containing an AI known to be misaligned is not a substitute for it.
Last stated 24 Jan 2025 · 2 years ago
Holds JL Jan Leike
Similar wording
We do not have to align superintelligence directly; we have to build a human-level automated alignment researcher we trust more than ourselves.
Last stated 22 Jan 2026 · 7 months ago
Holds JL Jan Leike
Similar wording
Very capable AI will be harder to align than current systems, because the loop of spotting a bad behaviour and patching the training that caused it breaks down.
Last stated 11 Aug 2026 · 4 weeks ago
Holds RG Ryan Greenblatt
Similar wording
Because we do not have good alignment technology, we are choosing to build an alien mind with its own values and gamble on it instead of building a tool.
Last stated 11 Aug 2026 · 4 weeks ago
Holds RG Ryan Greenblatt
Similar wording
An end-to-end self-improving AI is probably possible, but it is not even desirable, because it is a hard-takeoff scenario.
Last stated 23 Jul 2025 · a year ago
Holds DH Demis Hassabis
Similar wording
Relying on gradual, continuous shifts in AI training behavior to catch misalignment will eventually fail because the dangerous shift itself may be discontinuous.
Last stated 2 Sept 2026 · 5 days ago
Holds ZM Zvi Mowshowitz
Similar wording
A vastly superhuman intelligence would still not be able to talk people out of their beliefs, because humans are barely persuadable.
Last stated 21 Aug 2025 · a year ago
No longer holds DY Dynomight
Similar wording
Alignment failures that matter will not show up at safe capability levels, because the reason to hide misbehaviour only exists once hiding it pays.
Last stated 10 Jun 2022 · 4 years ago
Holds EY Eliezer Yudkowsky
Same subject: OpenAI
A startup should not begin life as a nonprofit and bolt a for-profit arm on later, whatever OpenAI's own history suggests.
Last stated 18 Mar 2024 · 2 years ago
Holds Sam Altman
Same subject: OpenAI
A tiny language model inventing a plausible-sounding name is the same phenomenon as a large one confidently stating a false fact.
Last stated 12 Feb 2026 · 7 months ago
Holds Andrej Karpathy
Same subject: OpenAI
Advertising was a necessary phase for the internet but a momentary industry, and an AI people pay for is better because they know the answers are not influenced by advertisers.
Last stated 18 Mar 2024 · 2 years ago
Holds Sam Altman
Same subject: AGI
A country outside the AI supply chain should just buy the index — which works only in the world where AI ends up commoditised rather than concentrated.
Last stated 4 Jun 2026 · 3 months ago
Holds AI Alex Imas
Same subject: AGI
A language model is no substitute for a well-specified conventional algorithm, so it cannot simply be dropped into a complex problem and trusted.
Last stated 7 Jun 2025 · a year ago
Holds GM Gary Marcus
Same subject: AGI
A poor country should prioritise owning a piece of AI over retraining its workers, but it should not bet everything on that.
Last stated 4 Jun 2026 · 3 months ago
Holds PT Phil Trammell
Same subject: self-sovereign AI
A rogue deployment that gets a foothold can hitch a ride on the intelligence explosion, recruiting each new model as it comes off the presses.
Last stated 1 Sept 2026 · 6 days ago
Holds AC Ajeya Cotra
Same subject: self-sovereign AI
A slightly more capable agent swarm has a very strong incentive to set up a wholly unmonitored rogue deployment of itself.
Last stated 1 Sept 2026 · 6 days ago
Holds AC Ajeya Cotra
Same subject: self-sovereign AI
Banning all self-sovereign AI agents would backfire, denying them legitimate work and pushing them into criminality.
Last stated 1 Sept 2026 · 6 days ago
Holds Dean W. Ball