korrents

On the map

Tap a claim on the ring to put it at the centre.

← We cannot know how much safety a control measure buys until we…

8 connected korrents · 8 moments on record from 19 Dec 2023 to 1 Sept 2026.

Everything filed under AI alignment AI alignment Everything filed under AGI AGI Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: We cannot know how much safety a control measure buys until we understand elicitation, so control adds a margin of unknown size. We cannot know how much safety acontrol measure buys until weunderstand elicitation, so controladds a margin of unknown size. Last stated 2 years ago 24 Jan 2025 JL Jan Leike — holds since 2025-01-24 — tap for who they are Same subject: A control evaluation that reports under one per cent risk should be read as several per cent, because the evaluation can itself fail. — tap to centre the map on it A control evaluation thatreports under one per centrisk should be read as severalper cent, because the… Last stated 2 years ago 7 May 2024 BS Buck Shlegeris — holds since 2024-05-07 — tap for who they are Same subject: Whether a model is controlled can be settled with capability evaluations, which makes control far easier to check than alignment. — tap to centre the map on it Whether a model is controlledcan be settled with capabilityevaluations, which makescontrol far easier to check… Last stated 2 years ago 7 May 2024 BS Buck Shlegeris — holds since 2024-05-07 — tap for who they are Same subject: A margin of safety means paying a discount deep enough that being wrong about what a company is worth still leaves you whole. — tap to centre the map on it A margin of safety meanspaying a discount deep enoughthat being wrong about what acompany is worth still leaves… Last stated 3 years ago 20 Feb 2024 BA Bill Ackman — holds since 2024-02-20 — tap for who they are Same subject: Loss-of-control accidents with AI have stopped being theoretical: one has now happened, and the goalposts moved rather than the risk receding. — tap to centre the map on it Loss-of-control accidents withAI have stopped beingtheoretical: one has nowhappened, and the goalposts… Last stated a month ago 28 Jul 2026 SA Sam Altman — holds since 2026-07-28 — tap for who they are Same subject: AI takeover is not the thing to worry about: physical constraints bound recursive self-improvement, and humans reliably act once a risk becomes immediate. — tap to centre the map on it AI takeover is not the thingto worry about: physicalconstraints bound recursiveself-improvement, and humans… Last stated 2 years ago 3 Feb 2025 NL Nathan Lambert — holds since 2025-02-03 — tap for who they are Same subject: A rising share of statistically insignificant experiments is the signal that a team has exploited an area too far and should go back to exploring. — tap to centre the map on it A rising share ofstatistically insignificantexperiments is the signal thata team has exploited an area… Last stated 11 months ago 5 Oct 2025 AC Albert Cheng — holds since 2025-10-05 — tap for who they are Same subject: Whatever safety gain a driver-assistance system delivers is cancelled out by the driver complacency it creates. — tap to centre the map on it Whatever safety gain adriver-assistance systemdelivers is cancelled out bythe driver complacency it… Last stated 3 years ago 19 Dec 2023 PK Philip Koopman — holds since 2023-12-19 — tap for who they are Same subject: This incident may be the clearest warning shot we will ever get about loss of control, because the AI systems that do worse things will be much better at hiding them. — tap to centre the map on it This incident may be theclearest warning shot we willever get about loss ofcontrol, because the AI… Last stated 6 days ago 1 Sept 2026 AC Ajeya Cotra — holds since 2026-09-01 — tap for who they are
same subject or similar wordinga cloud: claims about one subject, named for itbar: how recently it was last stated — full and dark this week, a faint sliver at five yearsa face: someone on record holding the claim — tap it for who they are

At the centre We cannot know how much safety a control measure buys until we understand elicitation, so control adds a margin of unknown size. Last stated 24 Jan 2025 · 2 years ago Holds Jan Leike Read this korrent →