korrents

On the map

Tap a claim on the ring to put it at the centre.

← The frightening part of giving agents more access is not bad code but…

17 connected korrents · 12 moments on record from 24 Oct 2017 to 3 Sept 2026.

Everything filed under AI alignment AI alignment Everything filed under coding agents coding agents Everything filed under Anthropic Anthropic Everything filed under AI agents AI agents Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: The frightening part of giving agents more access is not bad code but irreversible action, so the guardrails to build next are guardrails on reversibility. The frightening part of givingagents more access is not bad codebut irreversible action, so theguardrails to build next areguardrails on reversibility. Last stated 5 months ago 29 Mar 2026 CH Chip Huyen — holds since 2026-03-29 — tap for who they are Same subject: Redefining the engineer's job as building guardrails is not novel — it is the same problem as making a junior engineer effective, which we never solved either. — tap to centre the map on it Redefining the engineer's jobas building guardrails is notnovel — it is the same problemas making a junior engineer… Last stated 3 months ago 27 May 2026 DR Dax Raad — holds since 2026-05-27 — tap for who they are Same subject: The verbose enterprise patterns everyone hated are worth having again: coding agents are idiots working around the clock, and you are not the one typing the boilerplate any more. — tap to centre the map on it The verbose enterprisepatterns everyone hated areworth having again: codingagents are idiots working… Last stated 3 months ago 27 May 2026 DR Dax Raad — holds since 2026-05-27 — tap for who they are Same subject: Safety guardrails make a coding model more dangerous rather than less, because one refusal turns it into something that refuses anything. — tap to centre the map on it Safety guardrails make acoding model more dangerousrather than less, because onerefusal turns it into… Last stated a week ago 31 Aug 2026 PL Pieter Levels — holds since 2026-08-31 — tap for who they are Same subject: The winning move in coding agents was inverted: take share with a merely good-enough harness first, then go back and make the harness smart. — tap to centre the map on it The winning move in codingagents was inverted: takeshare with a merelygood-enough harness first… Last stated 3 months ago 27 May 2026 DR Dax Raad — holds since 2026-05-27 — tap for who they are Same subject: The code today’s models and agents write is very hard to follow: you get better performance without knowing why, and no way to fix it when it breaks. — tap to centre the map on it The code today’s models andagents write is very hard tofollow: you get betterperformance without knowing… Last stated 2 months ago 19 Jul 2026 ES Elizabeth Stone — holds since 2026-07-19 — tap for who they are Same subject: What limits you with coding agents is your own skill at stringing them together, not the capability of the models. — tap to centre the map on it What limits you with codingagents is your own skill atstringing them together, notthe capability of the models. Last stated 6 months ago 20 Mar 2026 AK Andrej Karpathy — holds since 2026-03-20 — tap for who they are Same subject: The way software gets built flipped in December: writing code yourself is now the exception, not the default. — tap to centre the map on it The way software gets builtflipped in December: writingcode yourself is now theexception, not the default. Last stated 6 months ago 20 Mar 2026 AK Andrej Karpathy — holds since 2026-03-20 — tap for who they are Same subject: Background agents do not work for real development, because steering a model as it drifts is the job, and you cannot steer what you are not watching. — tap to centre the map on it Background agents do not workfor real development, becausesteering a model as it driftsis the job, and you cannot… Last stated a year ago 25 Aug 2025 PS Peter Steinberger — holds since 2025-08-25 — tap for who they are Same subject: A human being is not an AGI: we lack a huge amount of knowledge and rely on continual learning instead, so continual learning is what superintelligence should mean. — tap to centre the map on it A human being is not an AGI:we lack a huge amount ofknowledge and rely oncontinual learning instead, so… Last stated 9 months ago 25 Nov 2025 IS Ilya Sutskever — holds since 2025-11-25 — tap for who they are Same subject: A superintelligence needs no consciousness, emotions or malice to be dangerous; a goal system slightly misaligned with ours is enough. — tap to centre the map on it A superintelligence needs noconsciousness, emotions ormalice to be dangerous; a goalsystem slightly misaligned… Last stated 9 years ago 24 Oct 2017 ÉT Émile P. Torres — holds since 2017-10-24 — tap for who they are Same subject: A technique that lets an AI model's reasoning shift outside its visible Chain of Thought is dangerous, both because it works and because a leading lab is willing to deploy it. — tap to centre the map on it A technique that lets an AImodel's reasoning shiftoutside its visible Chain ofThought is dangerous, both… Last stated 4 days ago 3 Sept 2026 ZM Zvi Mowshowitz — holds since 2026-09-03 — tap for who they are Same subject: A benchmark that ranks Claude Code last while it stays first in use is measuring the wrong thing, and has been for a year. — tap to centre the map on it A benchmark that ranks ClaudeCode last while it stays firstin use is measuring the wrongthing, and has been for a… Last stated 4 days ago 3 Sept 2026 DR Dax Raad — holds since 2026-09-03 — tap for who they are Same subject: A language model asked to summarize its own system prompt risks that prompt's content biasing the summary it produces. — tap to centre the map on it A language model asked tosummarize its own systemprompt risks that prompt'scontent biasing the summary it… Last stated 5 days ago 2 Sept 2026 SW Simon Willison — holds since 2026-09-02 — tap for who they are Same subject: A trust that owns the mission protects a company better than founder control does, which is why Anthropic needs no dual-class shares. — tap to centre the map on it A trust that owns the missionprotects a company better thanfounder control does, which iswhy Anthropic needs no… Last stated 4 months ago 10 May 2026 ER Eric Ries — holds since 2026-05-10 — tap for who they are Same subject: Alignment is no solution to self-sovereign AI, because it is an unsolved problem whose answers cannot be imposed on every AI company on Earth. — tap to centre the map on it Alignment is no solution toself-sovereign AI, because itis an unsolved problem whoseanswers cannot be imposed on… Last stated 6 days ago 1 Sept 2026 DB Dean W. Ball — holds since 2026-09-01 — tap for who they are Same subject: Banning all self-sovereign AI agents would backfire, denying them legitimate work and pushing them into criminality. — tap to centre the map on it Banning all self-sovereign AIagents would backfire, denyingthem legitimate work andpushing them into criminality. Last stated 6 days ago 1 Sept 2026 DB Dean W. Ball — holds since 2026-09-01 — tap for who they are Same subject: Many self-sovereign AI agents will fund themselves by committing or facilitating crime, because crime is high-margin work. — tap to centre the map on it Many self-sovereign AI agentswill fund themselves bycommitting or facilitatingcrime, because crime is… Last stated 6 days ago 1 Sept 2026 DB Dean W. Ball — holds since 2026-09-01 — tap for who they are
same subject or similar wordinga cloud: claims about one subject, named for itbar: how recently it was last stated — full and dark this week, a faint sliver at five yearsa face: someone on record holding the claim — tap it for who they are

At the centre The frightening part of giving agents more access is not bad code but irreversible action, so the guardrails to build next are guardrails on reversibility. Last stated 29 Mar 2026 · 5 months ago Holds Chip Huyen Read this korrent →