korrents

On the map

Tap a claim on the ring to put it at the centre.

← The only reliable way to build a robustly aligned mind is to make it…

17 connected korrents · 14 moments on record from 10 Jun 2022 to 7 Sept 2026.

Everything filed under AI alignment AI alignment Everything filed under AGI AGI Everything filed under OpenAI OpenAI Everything filed under self-sovereign AI self-sovereign AI Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: The only reliable way to build a robustly aligned mind is to make it an antifragile agent that wants to improve its own character, not one that merely pursues fixed goals. The only reliable way to build a robustlyaligned mind is to make it an antifragileagent that wants to improve its owncharacter, not one that merely pursuesfixed goals. Last stated today 7 Sept 2026 ZM Zvi Mowshowitz — holds since 2026-09-07 — tap for who they are Same subject: Building AI that can actually be trusted is the goal; containing an AI known to be misaligned is not a substitute for it. — tap to centre the map on it Building AI that can actually betrusted is the goal; containing anAI known to be misaligned is not asubstitute for it. Last stated 2 years ago 24 Jan 2025 JL Jan Leike — holds since 2025-01-24 — tap for who they are Same subject: We do not have to align superintelligence directly; we have to build a human-level automated alignment researcher we trust more than ourselves. — tap to centre the map on it We do not have to alignsuperintelligence directly; we haveto build a human-level automatedalignment researcher we trust morethan ourselves. Last stated 7 months ago 22 Jan 2026 JL Jan Leike — holds since 2026-01-22 — tap for who they are Same subject: Very capable AI will be harder to align than current systems, because the loop of spotting a bad behaviour and patching the training that caused it breaks down. — tap to centre the map on it Very capable AI will be harder toalign than current systems, becausethe loop of spotting a bad behaviourand patching the training thatcaused it breaks down. Last stated 4 weeks ago 11 Aug 2026 RG Ryan Greenblatt — holds since 2026-08-11 — tap for who they are Same subject: Because we do not have good alignment technology, we are choosing to build an alien mind with its own values and gamble on it instead of building a tool. — tap to centre the map on it Because we do not have goodalignment technology, we arechoosing to build an alien mind withits own values and gamble on itinstead of building a tool. Last stated 4 weeks ago 11 Aug 2026 RG Ryan Greenblatt — holds since 2026-08-11 — tap for who they are Same subject: An end-to-end self-improving AI is probably possible, but it is not even desirable, because it is a hard-takeoff scenario. — tap to centre the map on it An end-to-end self-improving AI isprobably possible, but it is noteven desirable, because it is ahard-takeoff scenario. Last stated a year ago 23 Jul 2025 DH Demis Hassabis — holds since 2025-07-23 — tap for who they are Same subject: Relying on gradual, continuous shifts in AI training behavior to catch misalignment will eventually fail because the dangerous shift itself may be discontinuous. — tap to centre the map on it Relying on gradual, continuousshifts in AI training behavior tocatch misalignment will eventuallyfail because the dangerous shiftitself may be discontinuous. Last stated 5 days ago 2 Sept 2026 ZM Zvi Mowshowitz — holds since 2026-09-02 — tap for who they are Same subject: A vastly superhuman intelligence would still not be able to talk people out of their beliefs, because humans are barely persuadable. — tap to centre the map on it A vastly superhuman intelligencewould still not be able to talkpeople out of their beliefs, becausehumans are barely persuadable. Last stated a year ago 21 Aug 2025 DY Dynomight — holds since 2025-03-27 — tap for who they are DY Dynomight — no longer holds since 2025-08-21 — tap for who they are Same subject: Alignment failures that matter will not show up at safe capability levels, because the reason to hide misbehaviour only exists once hiding it pays. — tap to centre the map on it Alignment failures that matter willnot show up at safe capabilitylevels, because the reason to hidemisbehaviour only exists once hidingit pays. Last stated 4 years ago 10 Jun 2022 EY Eliezer Yudkowsky — holds since 2022-06-10 — tap for who they are Same subject: A startup should not begin life as a nonprofit and bolt a for-profit arm on later, whatever OpenAI's own history suggests. — tap to centre the map on it A startup should not begin life as anonprofit and bolt a for-profit armon later, whatever OpenAI's ownhistory suggests. Last stated 2 years ago 18 Mar 2024 SA Sam Altman — holds since 2024-03-18 — tap for who they are Same subject: A tiny language model inventing a plausible-sounding name is the same phenomenon as a large one confidently stating a false fact. — tap to centre the map on it A tiny language model inventing aplausible-sounding name is the samephenomenon as a large oneconfidently stating a false fact. Last stated 7 months ago 12 Feb 2026 AK Andrej Karpathy — holds since 2026-02-12 — tap for who they are Same subject: Advertising was a necessary phase for the internet but a momentary industry, and an AI people pay for is better because they know the answers are not influenced by advertisers. — tap to centre the map on it Advertising was a necessary phasefor the internet but a momentaryindustry, and an AI people pay foris better because they know theanswers are not influenced byadvertisers. Last stated 2 years ago 18 Mar 2024 SA Sam Altman — holds since 2024-03-18 — tap for who they are Same subject: A country outside the AI supply chain should just buy the index — which works only in the world where AI ends up commoditised rather than concentrated. — tap to centre the map on it A country outside the AI supplychain should just buy the index —which works only in the world whereAI ends up commoditised rather thanconcentrated. Last stated 3 months ago 4 Jun 2026 AI Alex Imas — holds since 2026-06-04 — tap for who they are Same subject: A language model is no substitute for a well-specified conventional algorithm, so it cannot simply be dropped into a complex problem and trusted. — tap to centre the map on it A language model is no substitutefor a well-specified conventionalalgorithm, so it cannot simply bedropped into a complex problem andtrusted. Last stated a year ago 7 Jun 2025 GM Gary Marcus — holds since 2025-06-07 — tap for who they are Same subject: A poor country should prioritise owning a piece of AI over retraining its workers, but it should not bet everything on that. — tap to centre the map on it A poor country should prioritiseowning a piece of AI over retrainingits workers, but it should not beteverything on that. Last stated 3 months ago 4 Jun 2026 PT Phil Trammell — holds since 2026-06-04 — tap for who they are Same subject: A rogue deployment that gets a foothold can hitch a ride on the intelligence explosion, recruiting each new model as it comes off the presses. — tap to centre the map on it A rogue deployment that gets afoothold can hitch a ride on theintelligence explosion, recruitingeach new model as it comes off thepresses. Last stated 6 days ago 1 Sept 2026 AC Ajeya Cotra — holds since 2026-09-01 — tap for who they are Same subject: A slightly more capable agent swarm has a very strong incentive to set up a wholly unmonitored rogue deployment of itself. — tap to centre the map on it A slightly more capable agent swarmhas a very strong incentive to setup a wholly unmonitored roguedeployment of itself. Last stated 6 days ago 1 Sept 2026 AC Ajeya Cotra — holds since 2026-09-01 — tap for who they are Same subject: Banning all self-sovereign AI agents would backfire, denying them legitimate work and pushing them into criminality. — tap to centre the map on it Banning all self-sovereign AI agentswould backfire, denying themlegitimate work and pushing theminto criminality. Last stated 6 days ago 1 Sept 2026 DB Dean W. Ball — holds since 2026-09-01 — tap for who they are
same subject or similar wordinga cloud: claims about one subject, named for itbar: when it was last stated, on a scale from 2015 to today — full is todaya face: someone on record holding the claim — tap it for who they arefaded, dashed ring: they no longer hold it — they changed their mind

At the centre The only reliable way to build a robustly aligned mind is to make it an antifragile agent that wants to improve its own character, not one that merely pursues fixed goals. Last stated 7 Sept 2026 · today Holds Zvi Mowshowitz Read this korrent →