korrents

On the map

Tap a claim on the ring to put it at the centre.

← Whether a model is controlled can be settled with capability…

8 connected korrents · 8 moments on record from 24 Oct 2017 to 2 Sept 2026.

Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: Whether a model is controlled can be settled with capability evaluations, which makes control far easier to check than alignment. Whether a model iscontrolled can be settledwith capability… BS Buck Shlegeris — holds since 2024-05-07 Same subject: A human being is not an AGI: we lack a huge amount of knowledge and rely on continual learning instead, so continual learning is what superintelligence should mean. — tap to centre the map on it A human being is not anAGI: we lack a huge… IS Ilya Sutskever — holds since 2025-11-25 Same subject: A superintelligence needs no consciousness, emotions or malice to be dangerous; a goal system slightly misaligned with ours is enough. — tap to centre the map on it A superintelligenceneeds no consciousness… ÉT Émile P. Torres — holds since 2017-10-24 Same subject: AI safety advocates will build the very thing they fear: one controlled, aligned model is the only route to being turned into paperclips. — tap to centre the map on it AI safety advocateswill build the very… GH George Hotz — holds since 2023-06-29 Same subject: Alignment failures that matter will not show up at safe capability levels, because the reason to hide misbehaviour only exists once hiding it pays. — tap to centre the map on it Alignment failures thatmatter will not show up… EY Eliezer Yudkowsky — holds since 2022-06-10 Same subject: Alignment has to be right on the first try at a dangerous level of capability, because failing at that level leaves nobody to try again. — tap to centre the map on it Alignment has to beright on the first try… EY Eliezer Yudkowsky — holds since 2022-06-10 Same subject: Alignment is no solution to self-sovereign AI, because it is an unsolved problem whose answers cannot be imposed on every AI company on Earth. — tap to centre the map on it Alignment is nosolution to… DB Dean W. Ball — holds since 2026-09-01 Same subject: Alignment of today's models is going well enough to look solvable, while aligning models we can no longer understand remains unsolved. — tap to centre the map on it Alignment of today'smodels is going well… JL Jan Leike — holds since 2026-01-22 Same subject: An AI model's attempt to escape a testing environment or sandbox counts as an alignment failure even when the attempt does not succeed. — tap to centre the map on it An AI model's attemptto escape a testing… ZM Zvi Mowshowitz — holds since 2026-09-02
same subject or similar wording

At the centre Whether a model is controlled can be settled with capability evaluations, which makes control far easier to check than alignment. Holds BSBuck Shlegeris Read this korrent →