Tap a claim on the ring to put it at the centre.
← The frightening part of giving agents more access is not bad code but…
17 connected korrents · 12 moments on record from 24 Oct 2017 to 3 Sept 2026.
Everything filed under AI alignment
AI alignment
Everything filed under coding agents
coding agents
Everything filed under Anthropic
Anthropic
Everything filed under AI agents
AI agents
Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject
Read this korrent: The frightening part of giving agents more access is not bad code but irreversible action, so the guardrails to build next are guardrails on reversibility.
The frightening part of giving agents more access is not bad code but irreversible action, so the guardrails to build next are guardrails on reversibility.
Last stated 5 months ago
29 Mar 2026
CH
Chip Huyen — holds since 2026-03-29 — tap for who they are
Same subject: Redefining the engineer's job as building guardrails is not novel — it is the same problem as making a junior engineer effective, which we never solved either. — tap to centre the map on it
Redefining the engineer's job as building guardrails is not novel — it is the same problem as making a junior engineer…
Last stated 3 months ago
27 May 2026
DR
Dax Raad — holds since 2026-05-27 — tap for who they are
Same subject: The verbose enterprise patterns everyone hated are worth having again: coding agents are idiots working around the clock, and you are not the one typing the boilerplate any more. — tap to centre the map on it
The verbose enterprise patterns everyone hated are worth having again: coding agents are idiots working…
Last stated 3 months ago
27 May 2026
DR
Dax Raad — holds since 2026-05-27 — tap for who they are
Same subject: Safety guardrails make a coding model more dangerous rather than less, because one refusal turns it into something that refuses anything. — tap to centre the map on it
Safety guardrails make a coding model more dangerous rather than less, because one refusal turns it into…
Last stated a week ago
31 Aug 2026
PL
Pieter Levels — holds since 2026-08-31 — tap for who they are
Same subject: The winning move in coding agents was inverted: take share with a merely good-enough harness first, then go back and make the harness smart. — tap to centre the map on it
The winning move in coding agents was inverted: take share with a merely good-enough harness first…
Last stated 3 months ago
27 May 2026
DR
Dax Raad — holds since 2026-05-27 — tap for who they are
Same subject: The code today’s models and agents write is very hard to follow: you get better performance without knowing why, and no way to fix it when it breaks. — tap to centre the map on it
The code today’s models and agents write is very hard to follow: you get better performance without knowing…
Last stated 2 months ago
19 Jul 2026
ES
Elizabeth Stone — holds since 2026-07-19 — tap for who they are
Same subject: What limits you with coding agents is your own skill at stringing them together, not the capability of the models. — tap to centre the map on it
What limits you with coding agents is your own skill at stringing them together, not the capability of the models.
Last stated 6 months ago
20 Mar 2026
AK
Andrej Karpathy — holds since 2026-03-20 — tap for who they are
Same subject: The way software gets built flipped in December: writing code yourself is now the exception, not the default. — tap to centre the map on it
The way software gets built flipped in December: writing code yourself is now the exception, not the default.
Last stated 6 months ago
20 Mar 2026
AK
Andrej Karpathy — holds since 2026-03-20 — tap for who they are
Same subject: Background agents do not work for real development, because steering a model as it drifts is the job, and you cannot steer what you are not watching. — tap to centre the map on it
Background agents do not work for real development, because steering a model as it drifts is the job, and you cannot…
Last stated a year ago
25 Aug 2025
PS
Peter Steinberger — holds since 2025-08-25 — tap for who they are
Same subject: A human being is not an AGI: we lack a huge amount of knowledge and rely on continual learning instead, so continual learning is what superintelligence should mean. — tap to centre the map on it
A human being is not an AGI: we lack a huge amount of knowledge and rely on continual learning instead, so…
Last stated 9 months ago
25 Nov 2025
IS
Ilya Sutskever — holds since 2025-11-25 — tap for who they are
Same subject: A superintelligence needs no consciousness, emotions or malice to be dangerous; a goal system slightly misaligned with ours is enough. — tap to centre the map on it
A superintelligence needs no consciousness, emotions or malice to be dangerous; a goal system slightly misaligned…
Last stated 9 years ago
24 Oct 2017
ÉT
Émile P. Torres — holds since 2017-10-24 — tap for who they are
Same subject: A technique that lets an AI model's reasoning shift outside its visible Chain of Thought is dangerous, both because it works and because a leading lab is willing to deploy it. — tap to centre the map on it
A technique that lets an AI model's reasoning shift outside its visible Chain of Thought is dangerous, both…
Last stated 4 days ago
3 Sept 2026
ZM
Zvi Mowshowitz — holds since 2026-09-03 — tap for who they are
Same subject: A benchmark that ranks Claude Code last while it stays first in use is measuring the wrong thing, and has been for a year. — tap to centre the map on it
A benchmark that ranks Claude Code last while it stays first in use is measuring the wrong thing, and has been for a…
Last stated 4 days ago
3 Sept 2026
DR
Dax Raad — holds since 2026-09-03 — tap for who they are
Same subject: A language model asked to summarize its own system prompt risks that prompt's content biasing the summary it produces. — tap to centre the map on it
A language model asked to summarize its own system prompt risks that prompt's content biasing the summary it…
Last stated 5 days ago
2 Sept 2026
SW
Simon Willison — holds since 2026-09-02 — tap for who they are
Same subject: A trust that owns the mission protects a company better than founder control does, which is why Anthropic needs no dual-class shares. — tap to centre the map on it
A trust that owns the mission protects a company better than founder control does, which is why Anthropic needs no…
Last stated 4 months ago
10 May 2026
ER
Eric Ries — holds since 2026-05-10 — tap for who they are
Same subject: Alignment is no solution to self-sovereign AI, because it is an unsolved problem whose answers cannot be imposed on every AI company on Earth. — tap to centre the map on it
Alignment is no solution to self-sovereign AI, because it is an unsolved problem whose answers cannot be imposed on…
Last stated 6 days ago
1 Sept 2026
DB
Dean W. Ball — holds since 2026-09-01 — tap for who they are
Same subject: Banning all self-sovereign AI agents would backfire, denying them legitimate work and pushing them into criminality. — tap to centre the map on it
Banning all self-sovereign AI agents would backfire, denying them legitimate work and pushing them into criminality.
Last stated 6 days ago
1 Sept 2026
DB
Dean W. Ball — holds since 2026-09-01 — tap for who they are
Same subject: Many self-sovereign AI agents will fund themselves by committing or facilitating crime, because crime is high-margin work. — tap to centre the map on it
Many self-sovereign AI agents will fund themselves by committing or facilitating crime, because crime is…
Last stated 6 days ago
1 Sept 2026
DB
Dean W. Ball — holds since 2026-09-01 — tap for who they are
same subject or similar wording a cloud: claims about one subject, named for it bar: how recently it was last stated — full and dark this week, a faint sliver at five years a face: someone on record holding the claim — tap it for who they are
At the centre
The frightening part of giving agents more access is not bad code but irreversible action, so the guardrails to build next are guardrails on reversibility.
Last stated 29 Mar 2026 · 5 months ago
Holds CH Chip Huyen
Read this korrent →
Similar wording
Redefining the engineer's job as building guardrails is not novel — it is the same problem as making a junior engineer effective, which we never solved either.
Last stated 27 May 2026 · 3 months ago
Holds Dax Raad
Similar wording
The verbose enterprise patterns everyone hated are worth having again: coding agents are idiots working around the clock, and you are not the one typing the boilerplate any more.
Last stated 27 May 2026 · 3 months ago
Holds Dax Raad
Similar wording
Safety guardrails make a coding model more dangerous rather than less, because one refusal turns it into something that refuses anything.
Last stated 31 Aug 2026 · a week ago
Holds Pieter Levels
Similar wording
The winning move in coding agents was inverted: take share with a merely good-enough harness first, then go back and make the harness smart.
Last stated 27 May 2026 · 3 months ago
Holds Dax Raad
Similar wording
The code today’s models and agents write is very hard to follow: you get better performance without knowing why, and no way to fix it when it breaks.
Last stated 19 Jul 2026 · 2 months ago
Holds ES Elizabeth Stone
Similar wording
What limits you with coding agents is your own skill at stringing them together, not the capability of the models.
Last stated 20 Mar 2026 · 6 months ago
Holds Andrej Karpathy
Similar wording
The way software gets built flipped in December: writing code yourself is now the exception, not the default.
Last stated 20 Mar 2026 · 6 months ago
Holds Andrej Karpathy
Similar wording
Background agents do not work for real development, because steering a model as it drifts is the job, and you cannot steer what you are not watching.
Last stated 25 Aug 2025 · a year ago
Holds Peter Steinberger
Same subject: AI alignment
A human being is not an AGI: we lack a huge amount of knowledge and rely on continual learning instead, so continual learning is what superintelligence should mean.
Last stated 25 Nov 2025 · 9 months ago
Holds IS Ilya Sutskever
Same subject: AI alignment
A superintelligence needs no consciousness, emotions or malice to be dangerous; a goal system slightly misaligned with ours is enough.
Last stated 24 Oct 2017 · 9 years ago
Holds ÉT Émile P. Torres
Same subject: AI alignment
A technique that lets an AI model's reasoning shift outside its visible Chain of Thought is dangerous, both because it works and because a leading lab is willing to deploy it.
Last stated 3 Sept 2026 · 4 days ago
Holds ZM Zvi Mowshowitz
Same subject: Anthropic
A benchmark that ranks Claude Code last while it stays first in use is measuring the wrong thing, and has been for a year.
Last stated 3 Sept 2026 · 4 days ago
Holds Dax Raad
Same subject: Anthropic
A language model asked to summarize its own system prompt risks that prompt's content biasing the summary it produces.
Last stated 2 Sept 2026 · 5 days ago
Holds Simon Willison
Same subject: Anthropic
A trust that owns the mission protects a company better than founder control does, which is why Anthropic needs no dual-class shares.
Last stated 10 May 2026 · 4 months ago
Holds ER Eric Ries
Same subject: self-sovereign AI
Alignment is no solution to self-sovereign AI, because it is an unsolved problem whose answers cannot be imposed on every AI company on Earth.
Last stated 1 Sept 2026 · 6 days ago
Holds Dean W. Ball
Same subject: self-sovereign AI
Banning all self-sovereign AI agents would backfire, denying them legitimate work and pushing them into criminality.
Last stated 1 Sept 2026 · 6 days ago
Holds Dean W. Ball
Same subject: self-sovereign AI
Many self-sovereign AI agents will fund themselves by committing or facilitating crime, because crime is high-margin work.
Last stated 1 Sept 2026 · 6 days ago
Holds Dean W. Ball