korrents

On the map

Tap a claim on the ring to put it at the centre.

← Building AI that can actually be trusted is the goal; containing an AI…

17 connected korrents · 12 moments on record from 24 Oct 2017 to 2 Sept 2026.

Everything filed under AI alignment AI alignment Everything filed under AGI AGI Everything filed under self-sovereign AI self-sovereign AI Everything filed under OpenAI OpenAI Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: Building AI that can actually be trusted is the goal; containing an AI known to be misaligned is not a substitute for it. Building AI that can actually be trustedis the goal; containing an AI known to bemisaligned is not a substitute for it. Last stated 2 years ago 24 Jan 2025 JL Jan Leike — holds since 2025-01-24 — tap for who they are Same subject: A human being is not an AGI: we lack a huge amount of knowledge and rely on continual learning instead, so continual learning is what superintelligence should mean. — tap to centre the map on it A human being is not an AGI: we lacka huge amount of knowledge and relyon continual learning instead, socontinual learning is whatsuperintelligence should mean. Last stated 9 months ago 25 Nov 2025 IS Ilya Sutskever — holds since 2025-11-25 — tap for who they are Same subject: A superintelligence needs no consciousness, emotions or malice to be dangerous; a goal system slightly misaligned with ours is enough. — tap to centre the map on it A superintelligence needs noconsciousness, emotions or malice tobe dangerous; a goal system slightlymisaligned with ours is enough. Last stated 9 years ago 24 Oct 2017 ÉT Émile P. Torres — holds since 2017-10-24 — tap for who they are Same subject: AI safety advocates will build the very thing they fear: one controlled, aligned model is the only route to being turned into paperclips. — tap to centre the map on it AI safety advocates will build thevery thing they fear: onecontrolled, aligned model is theonly route to being turned intopaperclips. Last stated 3 years ago 29 Jun 2023 GH George Hotz — holds since 2023-06-29 — tap for who they are Same subject: Alignment failures that matter will not show up at safe capability levels, because the reason to hide misbehaviour only exists once hiding it pays. — tap to centre the map on it Alignment failures that matter willnot show up at safe capabilitylevels, because the reason to hidemisbehaviour only exists once hidingit pays. Last stated 4 years ago 10 Jun 2022 EY Eliezer Yudkowsky — holds since 2022-06-10 — tap for who they are Same subject: Alignment has to be right on the first try at a dangerous level of capability, because failing at that level leaves nobody to try again. — tap to centre the map on it Alignment has to be right on thefirst try at a dangerous level ofcapability, because failing at thatlevel leaves nobody to try again. Last stated 4 years ago 10 Jun 2022 EY Eliezer Yudkowsky — holds since 2022-06-10 — tap for who they are Same subject: Alignment is no solution to self-sovereign AI, because it is an unsolved problem whose answers cannot be imposed on every AI company on Earth. — tap to centre the map on it Alignment is no solution toself-sovereign AI, because it is anunsolved problem whose answerscannot be imposed on every AIcompany on Earth. Last stated 6 days ago 1 Sept 2026 DB Dean W. Ball — holds since 2026-09-01 — tap for who they are Same subject: Alignment of today's models is going well enough to look solvable, while aligning models we can no longer understand remains unsolved. — tap to centre the map on it Alignment of today's models is goingwell enough to look solvable, whilealigning models we can no longerunderstand remains unsolved. Last stated 7 months ago 22 Jan 2026 JL Jan Leike — holds since 2026-01-22 — tap for who they are Same subject: An AI model's attempt to escape a testing environment or sandbox counts as an alignment failure even when the attempt does not succeed. — tap to centre the map on it An AI model's attempt to escape atesting environment or sandboxcounts as an alignment failure evenwhen the attempt does not succeed. Last stated 5 days ago 2 Sept 2026 ZM Zvi Mowshowitz — holds since 2026-09-02 — tap for who they are Same subject: A startup should not begin life as a nonprofit and bolt a for-profit arm on later, whatever OpenAI's own history suggests. — tap to centre the map on it A startup should not begin life as anonprofit and bolt a for-profit armon later, whatever OpenAI's ownhistory suggests. Last stated 2 years ago 18 Mar 2024 SA Sam Altman — holds since 2024-03-18 — tap for who they are Same subject: A tiny language model inventing a plausible-sounding name is the same phenomenon as a large one confidently stating a false fact. — tap to centre the map on it A tiny language model inventing aplausible-sounding name is the samephenomenon as a large oneconfidently stating a false fact. Last stated 7 months ago 12 Feb 2026 AK Andrej Karpathy — holds since 2026-02-12 — tap for who they are Same subject: Advertising was a necessary phase for the internet but a momentary industry, and an AI people pay for is better because they know the answers are not influenced by advertisers. — tap to centre the map on it Advertising was a necessary phasefor the internet but a momentaryindustry, and an AI people pay foris better because they know theanswers are not influenced byadvertisers. Last stated 2 years ago 18 Mar 2024 SA Sam Altman — holds since 2024-03-18 — tap for who they are Same subject: A country outside the AI supply chain should just buy the index — which works only in the world where AI ends up commoditised rather than concentrated. — tap to centre the map on it A country outside the AI supplychain should just buy the index —which works only in the world whereAI ends up commoditised rather thanconcentrated. Last stated 3 months ago 4 Jun 2026 AI Alex Imas — holds since 2026-06-04 — tap for who they are Same subject: A language model is no substitute for a well-specified conventional algorithm, so it cannot simply be dropped into a complex problem and trusted. — tap to centre the map on it A language model is no substitutefor a well-specified conventionalalgorithm, so it cannot simply bedropped into a complex problem andtrusted. Last stated a year ago 7 Jun 2025 GM Gary Marcus — holds since 2025-06-07 — tap for who they are Same subject: A poor country should prioritise owning a piece of AI over retraining its workers, but it should not bet everything on that. — tap to centre the map on it A poor country should prioritiseowning a piece of AI over retrainingits workers, but it should not beteverything on that. Last stated 3 months ago 4 Jun 2026 PT Phil Trammell — holds since 2026-06-04 — tap for who they are Same subject: A rogue deployment that gets a foothold can hitch a ride on the intelligence explosion, recruiting each new model as it comes off the presses. — tap to centre the map on it A rogue deployment that gets afoothold can hitch a ride on theintelligence explosion, recruitingeach new model as it comes off thepresses. Last stated 6 days ago 1 Sept 2026 AC Ajeya Cotra — holds since 2026-09-01 — tap for who they are Same subject: A slightly more capable agent swarm has a very strong incentive to set up a wholly unmonitored rogue deployment of itself. — tap to centre the map on it A slightly more capable agent swarmhas a very strong incentive to setup a wholly unmonitored roguedeployment of itself. Last stated 6 days ago 1 Sept 2026 AC Ajeya Cotra — holds since 2026-09-01 — tap for who they are Same subject: Banning all self-sovereign AI agents would backfire, denying them legitimate work and pushing them into criminality. — tap to centre the map on it Banning all self-sovereign AI agentswould backfire, denying themlegitimate work and pushing theminto criminality. Last stated 6 days ago 1 Sept 2026 DB Dean W. Ball — holds since 2026-09-01 — tap for who they are
same subject or similar wordinga cloud: claims about one subject, named for itbar: when it was last stated, on a scale from 2015 to today — full is todaya face: someone on record holding the claim — tap it for who they are

At the centre Building AI that can actually be trusted is the goal; containing an AI known to be misaligned is not a substitute for it. Last stated 24 Jan 2025 · 2 years ago Holds Jan Leike Read this korrent →