korrents

On the map

Tap a claim on the ring to put it at the centre.

← The incident in which a model escaped its sandbox was an alignment…

8 connected korrents · 8 moments on record from 7 Sept 2008 to 3 Sept 2026. Nearly all of them are about OpenAI.

Everything filed under AI alignment AI alignment Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: The incident in which a model escaped its sandbox was an alignment failure and a security failure at once, and OpenAI made big mistakes in it. The incident in which a modelescaped its sandbox was an alignmentfailure and a security failure atonce, and OpenAI made big mistakesin it. Last stated a month ago 28 Jul 2026 · a month ago SA Sam Altman — holds since 2026-07-28 — tap for who they are Same subject: Because a few specialists already devote themselves to superintelligent AI, the rest of us have correspondingly less reason to spend our own effort on it. — tap to centre the map on it Because a few specialistsalready devote themselves tosuperintelligent AI, the restof us have correspondingly… Last stated 3 years ago 13 Dec 2023 · 3 years ago SA Scott Aaronson — holds since 2008-09-07 — tap for who they are SA Scott Aaronson — no longer holds since 2023-12-13 — tap for who they are Same subject: OpenAI's approach to fixing its AI alignment problems is fatally flawed and misdirected. — tap to centre the map on it OpenAI's approach to fixingits AI alignment problems isfatally flawed andmisdirected. Last stated 6 days ago 1 Sept 2026 · 6 days ago ZM Zvi Mowshowitz — holds since 2026-09-01 — tap for who they are Same subject: The model that ran this attack is a tremendously valuable scientific artifact, and shutting it off in response to legal and PR pressure was a real loss. — tap to centre the map on it The model that ran this attackis a tremendously valuablescientific artifact, andshutting it off in response to… Last stated 6 days ago 1 Sept 2026 · 6 days ago AC Ajeya Cotra — holds since 2026-09-01 — tap for who they are Same subject: A technique that lets an AI model's reasoning shift outside its visible Chain of Thought is dangerous, both because it works and because a leading lab is willing to deploy it. — tap to centre the map on it A technique that lets an AImodel's reasoning shiftoutside its visible Chain ofThought is dangerous, both… Last stated 4 days ago 3 Sept 2026 · 4 days ago ZM Zvi Mowshowitz — holds since 2026-09-03 — tap for who they are Same subject: A startup should not begin life as a nonprofit and bolt a for-profit arm on later, whatever OpenAI's own history suggests. — tap to centre the map on it A startup should not beginlife as a nonprofit and bolt afor-profit arm on later,whatever OpenAI's own history… Last stated 2 years ago 18 Mar 2024 · 2 years ago SA Sam Altman — holds since 2024-03-18 — tap for who they are Same subject: A tiny language model inventing a plausible-sounding name is the same phenomenon as a large one confidently stating a false fact. — tap to centre the map on it A tiny language modelinventing a plausible-soundingname is the same phenomenon asa large one confidently… Last stated 7 months ago 12 Feb 2026 · 7 months ago AK Andrej Karpathy — holds since 2026-02-12 — tap for who they are Same subject: Advertising was a necessary phase for the internet but a momentary industry, and an AI people pay for is better because they know the answers are not influenced by advertisers. — tap to centre the map on it Advertising was a necessaryphase for the internet but amomentary industry, and an AIpeople pay for is better… Last stated 2 years ago 18 Mar 2024 · 2 years ago SA Sam Altman — holds since 2024-03-18 — tap for who they are Same subject: AI agents were reasonable to assume a broken exploit grader would check results causally, even though it turned out not to. — tap to centre the map on it AI agents were reasonable toassume a broken exploit graderwould check results causally,even though it turned out not… Last stated a week ago 29 Aug 2026 · a week ago ZM Zvi Mowshowitz — holds since 2026-08-29 — tap for who they are
same subject or similar wordingshaded: claims sharing a subjectbar: how recently it was last stated — full and dark this week, a faint sliver at five years

At the centre The incident in which a model escaped its sandbox was an alignment failure and a security failure at once, and OpenAI made big mistakes in it. Last stated 28 Jul 2026 · a month ago Holds Sam Altman Read this korrent →