korrents

On the map

Tap a claim on the ring to put it at the centre.

← It is not enough for an AI model to be smart: aligning it to act…

17 connected korrents · 16 moments from 10 Jun 2022 to 19 Sept 2026. Nearly all of them are about AI alignment.

Everything filed under AI agents AI agents Everything filed under OpenAI OpenAI Everything filed under reinforcement learning reinforcement learning Everything filed under AGI AGI Everything filed under LLMs LLMs Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: It is not enough for an AI model to be smart: aligning it to act "correctly" according to its model spec is incredibly important. It is not enough for an AI model to besmart: aligning it to act "correctly"according to its model spec is incrediblyimportant. Last stated 3 months ago by 24 Jul 2026 LR Lee Robinson — holds since 2026-07-24 — tap for who they are Same subject: Building AI that can actually be trusted is the goal; containing an AI known to be misaligned is not a substitute for it. — tap to centre the map on it Building AI that can actually betrusted is the goal; containing anAI known to be misaligned is not asubstitute for it. Last stated 2 years ago 24 Jan 2025 JL Jan Leike — holds since 2025-01-24 — tap for who they are Same subject: No observation of a model's behaviour can establish that it is aligned, because a model capable enough to be dangerous is capable enough to produce whatever behaviour it is being watched for. — tap to centre the map on it No observation of a model'sbehaviour can establish that it isaligned, because a model capableenough to be dangerous is capableenough to produce whatever behaviourit is being watched for. Last stated 2 years ago 7 May 2024 BS Buck Shlegeris — holds since 2024-05-07 — tap for who they are EY Eliezer Yudkowsky — holds since 2022-06-10 — tap for who they are Same subject: A validated theory of intelligence may be necessary for achieving genuine AI alignment. — tap to centre the map on it A validated theory of intelligencemay be necessary for achievinggenuine AI alignment. Last stated 2 months ago 17 Aug 2026 KK Kevin Kelly — holds since 2026-08-17 — tap for who they are Same subject: Alignment of today's models is going well enough to look solvable, while aligning models we can no longer understand remains unsolved. — tap to centre the map on it Alignment of today's models is goingwell enough to look solvable, whilealigning models we can no longerunderstand remains unsolved. Last stated 9 months ago 22 Jan 2026 JL Jan Leike — holds since 2026-01-22 — tap for who they are Same subject: Very capable AI will be harder to align than current systems, because the loop of spotting a bad behaviour and patching the training that caused it breaks down. — tap to centre the map on it Very capable AI will be harder toalign than current systems, becausethe loop of spotting a bad behaviourand patching the training thatcaused it breaks down. Last stated 2 months ago 11 Aug 2026 RG Ryan Greenblatt — holds since 2026-08-11 — tap for who they are Same subject: AI models considered smart today will be considered very dumb in the future, and smarter models make fewer mistakes and need fewer turns. — tap to centre the map on it AI models considered smart todaywill be considered very dumb in thefuture, and smarter models makefewer mistakes and need fewer turns. Last stated 3 weeks ago 19 Sept 2026 TB Thorsten Ball — holds since 2026-09-19 — tap for who they are Same subject: We do not have to align superintelligence directly; we have to build a human-level automated alignment researcher we trust more than ourselves. — tap to centre the map on it We do not have to alignsuperintelligence directly; we haveto build a human-level automatedalignment researcher we trust morethan ourselves. Last stated 9 months ago 22 Jan 2026 JL Jan Leike — holds since 2026-01-22 — tap for who they are Same subject: AI alignment is a red herring not worth the effort currently devoted to it. — tap to centre the map on it AI alignment is a red herring notworth the effort currently devotedto it. Last stated 2 months ago 8 Aug 2026 MW Matt Webb — holds since 2026-08-08 — tap for who they are Same subject: A startup should not begin life as a nonprofit and bolt a for-profit arm on later, whatever OpenAI's own history suggests. — tap to centre the map on it A startup should not begin life as anonprofit and bolt a for-profit armon later, whatever OpenAI's ownhistory suggests. Last stated 3 years ago 18 Mar 2024 SA Sam Altman — holds since 2024-03-18 — tap for who they are Same subject: Advertising was a necessary phase for the internet but a momentary industry, and an AI people pay for is better because they know the answers are not influenced by advertisers. — tap to centre the map on it Advertising was a necessary phasefor the internet but a momentaryindustry, and an AI people pay foris better because they know theanswers are not influenced byadvertisers. Last stated 3 years ago 18 Mar 2024 SA Sam Altman — holds since 2024-03-18 — tap for who they are Same subject: Agents are going to be incorporated into mainstream interfaces — tap to centre the map on it Agents are going to be incorporatedinto mainstream interfaces Last stated 7 months ago 25 Mar 2026 DH David Heinemeier Hansson — holds since 2026-03-25 — tap for who they are Same subject: A collection of AI agents does not automatically form a functioning organization any more than a group of smart people does. — tap to centre the map on it A collection of AI agents does notautomatically form a functioningorganization any more than a groupof smart people does. Last stated 5 months ago 21 May 2026 RK Rohit Krishnan — holds since 2026-05-21 — tap for who they are Same subject: A company cannot be led by machines, because nobody would have recourse against them. — tap to centre the map on it A company cannot be led by machines,because nobody would have recourseagainst them. Last stated 4 weeks ago 15 Sept 2026 TL Tobias Lütke — holds since 2026-09-15 — tap for who they are Same subject: A computer-using agent needs a machine of its own, or you will spend the day fighting it for the mouse cursor. — tap to centre the map on it A computer-using agent needs amachine of its own, or you willspend the day fighting it for themouse cursor. Last stated 2 months ago 10 Aug 2026 PS Peter Steinberger — holds since 2026-08-10 — tap for who they are Same subject: A physical AI company needs three AIs, not one — the agent, the simulator and the critic — turning deployment into a flywheel. — tap to centre the map on it A physical AI company needs threeAIs, not one — the agent, thesimulator and the critic — turningdeployment into a flywheel. Last stated 2 months ago 3 Aug 2026 DD Dmitri Dolgov — holds since 2026-08-03 — tap for who they are Same subject: A verifiable task can be optimised by reinforcement learning until a neural network performs it extremely well. — tap to centre the map on it A verifiable task can be optimisedby reinforcement learning until aneural network performs it extremelywell. Last stated 11 months ago 17 Nov 2025 AK Andrej Karpathy — holds since 2025-11-17 — tap for who they are Same subject: Building a realistic simulator is exactly as hard as building the agent, because the simulator is itself a large AI model. — tap to centre the map on it Building a realistic simulator isexactly as hard as building theagent, because the simulator isitself a large AI model. Last stated 2 months ago 3 Aug 2026 DD Dmitri Dolgov — holds since 2026-08-03 — tap for who they are
same subject or similar wordinga cloud: claims about one subject, named for itbar: when it was last stated, on a scale from 2015 to today — full is todaya face: someone who holds the claim — tap it for who they are

At the centre It is not enough for an AI model to be smart: aligning it to act "correctly" according to its model spec is incredibly important. Last stated by 24 Jul 2026 · 3 months ago Holds Lee Robinson Read this korrent →