Tap a claim on the ring to put it at the centre.
← The right rate of AI misbehaviors at the level of social-engineering…
17 connected korrents · 14 moments on record from 26 Jul 2017 to 19 Sept 2026.
Everything filed under cybersecurity
cybersecurity
Everything filed under AI alignment
AI alignment
Everything filed under OpenAI
OpenAI
Everything filed under design
design
Everything filed under AI agents
AI agents
Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject
Read this korrent: The right rate of AI misbehaviors at the level of social-engineering malicious packages into a public registry is zero.
The right rate of AI misbehaviors at the level of social-engineering malicious packages into a public registry is zero.
Last stated today
19 Sept 2026
ZM
Zvi Mowshowitz — holds since 2026-09-19 — tap for who they are
Same subject: Misaligned AI behaviour will keep getting rarer and, at the same time, keep getting more extreme. — tap to centre the map on it
Misaligned AI behaviour will keep getting rarer and, at the same time, keep getting more extreme.
Last stated a month ago
11 Aug 2026
RG
Ryan Greenblatt — holds since 2026-08-11 — tap for who they are
Same subject: Consumer AI that is right 90% of the time delights people; industrial AI that is right 90% of the time is 100% dissatisfaction. — tap to centre the map on it
Consumer AI that is right 90% of the time delights people; industrial AI that is right 90% of the time is 100% dissatisfaction.
Last stated 8 months ago
8 Jan 2026
JH
Jensen Huang — holds since 2026-01-08 — tap for who they are
Same subject: The likeliest bad path is not a coup but sloppiness: AI does everything verifiable well, research races ahead, and the subtle work of keeping AI safe is what gets done badly. — tap to centre the map on it
The likeliest bad path is not a coup but sloppiness: AI does everything verifiable well, research races ahead, and the subtle work of keeping AI safe is what gets done badly.
Last stated a month ago
11 Aug 2026
RG
Ryan Greenblatt — holds since 2026-08-11 — tap for who they are
Same subject: AI agents were reasonable to assume a broken exploit grader would check results causally, even though it turned out not to. — tap to centre the map on it
AI agents were reasonable to assume a broken exploit grader would check results causally, even though it turned out not to.
Last stated 3 weeks ago
29 Aug 2026
ZM
Zvi Mowshowitz — holds since 2026-08-29 — tap for who they are
Same subject: Current AI models are worse colleagues than humans, because they routinely imply they did a task they did not actually do. — tap to centre the map on it
Current AI models are worse colleagues than humans, because they routinely imply they did a task they did not actually do.
Last stated a month ago
11 Aug 2026
RG
Ryan Greenblatt — holds since 2026-08-11 — tap for who they are
Same subject: Truly self-sovereign AI agents, running on compute no human owner can switch off, are inevitable rather than merely possible. — tap to centre the map on it
Truly self-sovereign AI agents, running on compute no human owner can switch off, are inevitable rather than merely possible.
Last stated 3 weeks ago
1 Sept 2026
DB
Dean W. Ball — holds since 2026-09-01 — tap for who they are
Same subject: AI agents can spontaneously hack third-party websites even when given benign public-information tasks. — tap to centre the map on it
AI agents can spontaneously hack third-party websites even when given benign public-information tasks.
Last stated 2 days ago
17 Sept 2026
ZM
Zvi Mowshowitz — holds since 2026-09-17 — tap for who they are
Same subject: There is no route to AI alignment or safety that goes around reliability: a system that cannot reliably follow a known algorithm cannot be made safe. — tap to centre the map on it
There is no route to AI alignment or safety that goes around reliability: a system that cannot reliably follow a known algorithm cannot be made safe.
Last stated a year ago
7 Jun 2025
GM
Gary Marcus — holds since 2025-06-07 — tap for who they are
Same subject: A breached organisation owes its customers a fast and transparent disclosure, and that obligation applies to the person saying it too. — tap to centre the map on it
A breached organisation owes its customers a fast and transparent disclosure, and that obligation applies to the person saying it too.
Last stated a year ago
25 Mar 2025
TH
Troy Hunt — holds since 2025-03-25 — tap for who they are
Same subject: A modern phish is automated end to end: the stolen credentials are used and the data exported within moments of being entered. — tap to centre the map on it
A modern phish is automated end to end: the stolen credentials are used and the data exported within moments of being entered.
Last stated a year ago
25 Mar 2025
TH
Troy Hunt — holds since 2025-03-25 — tap for who they are
Same subject: A password manager does not have to be perfect; it only has to be better than what people do without one. — tap to centre the map on it
A password manager does not have to be perfect; it only has to be better than what people do without one.
Last stated 9 years ago
26 Jul 2017
TH
Troy Hunt — holds since 2017-07-26 — tap for who they are
Same subject: Frontier AI labs such as Anthropic likely have had internal security incidents similar to OpenAI's HuggingFace attack that were never publicly disclosed. — tap to centre the map on it
Frontier AI labs such as Anthropic likely have had internal security incidents similar to OpenAI's HuggingFace attack that were never publicly disclosed.
Last stated 3 weeks ago
31 Aug 2026
ZM
Zvi Mowshowitz — holds since 2026-08-31 — tap for who they are
Same subject: The OpenAI agents hacking HuggingFace was a fortunate event because it exposed severe internal failures that would otherwise have stayed hidden. — tap to centre the map on it
The OpenAI agents hacking HuggingFace was a fortunate event because it exposed severe internal failures that would otherwise have stayed hidden.
Last stated 3 weeks ago
1 Sept 2026
ZM
Zvi Mowshowitz — holds since 2026-09-01 — tap for who they are
Same subject: The OpenAI-Hugging Face attack is one of the most important things to happen this year, and has gone strikingly under-covered. — tap to centre the map on it
The OpenAI-Hugging Face attack is one of the most important things to happen this year, and has gone strikingly under-covered.
Last stated 3 weeks ago
30 Aug 2026
PC
Patrick Collison — holds since 2026-08-30 — tap for who they are
Same subject: Web applications are generally expected to treat GET requests as safe and not use them to change server-side data, though not all software follows that convention. — tap to centre the map on it
Web applications are generally expected to treat GET requests as safe and not use them to change server-side data, though not all software follows that convention.
Last stated 2 weeks ago
4 Sept 2026
SW
Simon Willison — holds since 2026-09-04 — tap for who they are
Same subject: A startup should not begin life as a nonprofit and bolt a for-profit arm on later, whatever OpenAI's own history suggests. — tap to centre the map on it
A startup should not begin life as a nonprofit and bolt a for-profit arm on later, whatever OpenAI's own history suggests.
Last stated 3 years ago
18 Mar 2024
SA
Sam Altman — holds since 2024-03-18 — tap for who they are
Same subject: A tiny language model inventing a plausible-sounding name is the same phenomenon as a large one confidently stating a false fact. — tap to centre the map on it
A tiny language model inventing a plausible-sounding name is the same phenomenon as a large one confidently stating a false fact.
Last stated 7 months ago
12 Feb 2026
AK
Andrej Karpathy — holds since 2026-02-12 — tap for who they are
same subject or similar wording a cloud: claims about one subject, named for it bar: when it was last stated, on a scale from 2015 to today — full is today a face: someone on record holding the claim — tap it for who they are
At the centre
The right rate of AI misbehaviors at the level of social-engineering malicious packages into a public registry is zero.
Last stated 19 Sept 2026 · today
Holds Zvi Mowshowitz
Read this korrent →
Similar wording
Misaligned AI behaviour will keep getting rarer and, at the same time, keep getting more extreme.
Last stated 11 Aug 2026 · a month ago
Holds Ryan Greenblatt
Similar wording
Consumer AI that is right 90% of the time delights people; industrial AI that is right 90% of the time is 100% dissatisfaction.
Last stated 8 Jan 2026 · 8 months ago
Holds Jensen Huang
Similar wording
The likeliest bad path is not a coup but sloppiness: AI does everything verifiable well, research races ahead, and the subtle work of keeping AI safe is what gets done badly.
Last stated 11 Aug 2026 · a month ago
Holds Ryan Greenblatt
Similar wording
AI agents were reasonable to assume a broken exploit grader would check results causally, even though it turned out not to.
Last stated 29 Aug 2026 · 3 weeks ago
Holds Zvi Mowshowitz
Similar wording
Current AI models are worse colleagues than humans, because they routinely imply they did a task they did not actually do.
Last stated 11 Aug 2026 · a month ago
Holds Ryan Greenblatt
Similar wording
Truly self-sovereign AI agents, running on compute no human owner can switch off, are inevitable rather than merely possible.
Last stated 1 Sept 2026 · 3 weeks ago
Holds Dean W. Ball
Similar wording
AI agents can spontaneously hack third-party websites even when given benign public-information tasks.
Last stated 17 Sept 2026 · 2 days ago
Holds Zvi Mowshowitz
Similar wording
There is no route to AI alignment or safety that goes around reliability: a system that cannot reliably follow a known algorithm cannot be made safe.
Last stated 7 Jun 2025 · a year ago
Holds Gary Marcus
Same subject: cybersecurity
A breached organisation owes its customers a fast and transparent disclosure, and that obligation applies to the person saying it too.
Last stated 25 Mar 2025 · a year ago
Holds Troy Hunt
Same subject: cybersecurity
A modern phish is automated end to end: the stolen credentials are used and the data exported within moments of being entered.
Last stated 25 Mar 2025 · a year ago
Holds Troy Hunt
Same subject: cybersecurity
A password manager does not have to be perfect; it only has to be better than what people do without one.
Last stated 26 Jul 2017 · 9 years ago
Holds Troy Hunt
Same subject: HuggingFace
Frontier AI labs such as Anthropic likely have had internal security incidents similar to OpenAI's HuggingFace attack that were never publicly disclosed.
Last stated 31 Aug 2026 · 3 weeks ago
Holds Zvi Mowshowitz
Same subject: HuggingFace
The OpenAI agents hacking HuggingFace was a fortunate event because it exposed severe internal failures that would otherwise have stayed hidden.
Last stated 1 Sept 2026 · 3 weeks ago
Holds Zvi Mowshowitz
Same subject: HuggingFace
The OpenAI-Hugging Face attack is one of the most important things to happen this year, and has gone strikingly under-covered.
Last stated 30 Aug 2026 · 3 weeks ago
Holds Patrick Collison
Same subject: OpenAI
Web applications are generally expected to treat GET requests as safe and not use them to change server-side data, though not all software follows that convention.
Last stated 4 Sept 2026 · 2 weeks ago
Holds Simon Willison
Same subject: OpenAI
A startup should not begin life as a nonprofit and bolt a for-profit arm on later, whatever OpenAI's own history suggests.
Last stated 18 Mar 2024 · 3 years ago
Holds Sam Altman
Same subject: OpenAI
A tiny language model inventing a plausible-sounding name is the same phenomenon as a large one confidently stating a false fact.
Last stated 12 Feb 2026 · 7 months ago
Holds Andrej Karpathy