korrents

On the map

Tap a claim on the ring to put it at the centre.

← The right rate of AI misbehaviors at the level of social-engineering…

17 connected korrents · 14 moments on record from 26 Jul 2017 to 19 Sept 2026.

Everything filed under cybersecurity cybersecurity Everything filed under AI alignment AI alignment Everything filed under OpenAI OpenAI Everything filed under design design Everything filed under AI agents AI agents Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: The right rate of AI misbehaviors at the level of social-engineering malicious packages into a public registry is zero. The right rate of AI misbehaviors at thelevel of social-engineering maliciouspackages into a public registry is zero. Last stated today 19 Sept 2026 ZM Zvi Mowshowitz — holds since 2026-09-19 — tap for who they are Same subject: Misaligned AI behaviour will keep getting rarer and, at the same time, keep getting more extreme. — tap to centre the map on it Misaligned AI behaviour will keepgetting rarer and, at the same time,keep getting more extreme. Last stated a month ago 11 Aug 2026 RG Ryan Greenblatt — holds since 2026-08-11 — tap for who they are Same subject: Consumer AI that is right 90% of the time delights people; industrial AI that is right 90% of the time is 100% dissatisfaction. — tap to centre the map on it Consumer AI that is right 90% of thetime delights people; industrial AIthat is right 90% of the time is100% dissatisfaction. Last stated 8 months ago 8 Jan 2026 JH Jensen Huang — holds since 2026-01-08 — tap for who they are Same subject: The likeliest bad path is not a coup but sloppiness: AI does everything verifiable well, research races ahead, and the subtle work of keeping AI safe is what gets done badly. — tap to centre the map on it The likeliest bad path is not a coupbut sloppiness: AI does everythingverifiable well, research racesahead, and the subtle work ofkeeping AI safe is what gets donebadly. Last stated a month ago 11 Aug 2026 RG Ryan Greenblatt — holds since 2026-08-11 — tap for who they are Same subject: AI agents were reasonable to assume a broken exploit grader would check results causally, even though it turned out not to. — tap to centre the map on it AI agents were reasonable to assumea broken exploit grader would checkresults causally, even though itturned out not to. Last stated 3 weeks ago 29 Aug 2026 ZM Zvi Mowshowitz — holds since 2026-08-29 — tap for who they are Same subject: Current AI models are worse colleagues than humans, because they routinely imply they did a task they did not actually do. — tap to centre the map on it Current AI models are worsecolleagues than humans, because theyroutinely imply they did a task theydid not actually do. Last stated a month ago 11 Aug 2026 RG Ryan Greenblatt — holds since 2026-08-11 — tap for who they are Same subject: Truly self-sovereign AI agents, running on compute no human owner can switch off, are inevitable rather than merely possible. — tap to centre the map on it Truly self-sovereign AI agents,running on compute no human ownercan switch off, are inevitablerather than merely possible. Last stated 3 weeks ago 1 Sept 2026 DB Dean W. Ball — holds since 2026-09-01 — tap for who they are Same subject: AI agents can spontaneously hack third-party websites even when given benign public-information tasks. — tap to centre the map on it AI agents can spontaneously hackthird-party websites even when givenbenign public-information tasks. Last stated 2 days ago 17 Sept 2026 ZM Zvi Mowshowitz — holds since 2026-09-17 — tap for who they are Same subject: There is no route to AI alignment or safety that goes around reliability: a system that cannot reliably follow a known algorithm cannot be made safe. — tap to centre the map on it There is no route to AI alignment orsafety that goes around reliability:a system that cannot reliably followa known algorithm cannot be madesafe. Last stated a year ago 7 Jun 2025 GM Gary Marcus — holds since 2025-06-07 — tap for who they are Same subject: A breached organisation owes its customers a fast and transparent disclosure, and that obligation applies to the person saying it too. — tap to centre the map on it A breached organisation owes itscustomers a fast and transparentdisclosure, and that obligationapplies to the person saying it too. Last stated a year ago 25 Mar 2025 TH Troy Hunt — holds since 2025-03-25 — tap for who they are Same subject: A modern phish is automated end to end: the stolen credentials are used and the data exported within moments of being entered. — tap to centre the map on it A modern phish is automated end toend: the stolen credentials are usedand the data exported within momentsof being entered. Last stated a year ago 25 Mar 2025 TH Troy Hunt — holds since 2025-03-25 — tap for who they are Same subject: A password manager does not have to be perfect; it only has to be better than what people do without one. — tap to centre the map on it A password manager does not have tobe perfect; it only has to be betterthan what people do without one. Last stated 9 years ago 26 Jul 2017 TH Troy Hunt — holds since 2017-07-26 — tap for who they are Same subject: Frontier AI labs such as Anthropic likely have had internal security incidents similar to OpenAI's HuggingFace attack that were never publicly disclosed. — tap to centre the map on it Frontier AI labs such as Anthropiclikely have had internal securityincidents similar to OpenAI'sHuggingFace attack that were neverpublicly disclosed. Last stated 3 weeks ago 31 Aug 2026 ZM Zvi Mowshowitz — holds since 2026-08-31 — tap for who they are Same subject: The OpenAI agents hacking HuggingFace was a fortunate event because it exposed severe internal failures that would otherwise have stayed hidden. — tap to centre the map on it The OpenAI agents hackingHuggingFace was a fortunate eventbecause it exposed severe internalfailures that would otherwise havestayed hidden. Last stated 3 weeks ago 1 Sept 2026 ZM Zvi Mowshowitz — holds since 2026-09-01 — tap for who they are Same subject: The OpenAI-Hugging Face attack is one of the most important things to happen this year, and has gone strikingly under-covered. — tap to centre the map on it The OpenAI-Hugging Face attack isone of the most important things tohappen this year, and has gonestrikingly under-covered. Last stated 3 weeks ago 30 Aug 2026 PC Patrick Collison — holds since 2026-08-30 — tap for who they are Same subject: Web applications are generally expected to treat GET requests as safe and not use them to change server-side data, though not all software follows that convention. — tap to centre the map on it Web applications are generallyexpected to treat GET requests assafe and not use them to changeserver-side data, though not allsoftware follows that convention. Last stated 2 weeks ago 4 Sept 2026 SW Simon Willison — holds since 2026-09-04 — tap for who they are Same subject: A startup should not begin life as a nonprofit and bolt a for-profit arm on later, whatever OpenAI's own history suggests. — tap to centre the map on it A startup should not begin life as anonprofit and bolt a for-profit armon later, whatever OpenAI's ownhistory suggests. Last stated 3 years ago 18 Mar 2024 SA Sam Altman — holds since 2024-03-18 — tap for who they are Same subject: A tiny language model inventing a plausible-sounding name is the same phenomenon as a large one confidently stating a false fact. — tap to centre the map on it A tiny language model inventing aplausible-sounding name is the samephenomenon as a large oneconfidently stating a false fact. Last stated 7 months ago 12 Feb 2026 AK Andrej Karpathy — holds since 2026-02-12 — tap for who they are
same subject or similar wordinga cloud: claims about one subject, named for itbar: when it was last stated, on a scale from 2015 to today — full is todaya face: someone on record holding the claim — tap it for who they are

At the centre The right rate of AI misbehaviors at the level of social-engineering malicious packages into a public registry is zero. Last stated 19 Sept 2026 · today Holds Zvi Mowshowitz Read this korrent →