korrents

On the map

Tap a claim on the ring to put it at the centre.

← AI assistants should refuse requests that are harmful or unethical, even if the user insists.

17 connected korrents · 13 moments from 24 Feb 2018 to 20 Sept 2026.

Everything filed under AI agents AI agents Everything filed under AGI AGI Everything filed under HuggingFace HuggingFace Everything filed under reinforcement learning reinforcement learning Everything filed under LLMs LLMs Everything filed under LLMs LLMs Everything filed under government government Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: AI assistants should refuse requests that are harmful or unethical, even if the user insists. AI assistants should refuse requests thatare harmful or unethical, even if the userinsists. Last stated 2 weeks ago 20 Sept 2026 ZM Zvi Mowshowitz — holds since 2026-09-20 — tap for who they are Same subject: AI assistants should refuse harmful requests, even if they are technically legal. — tap to centre the map on it AI assistants should refuse harmfulrequests, even if they aretechnically legal. Last stated 2 weeks ago 20 Sept 2026 ZM Zvi Mowshowitz — holds since 2026-09-20 — tap for who they are Same subject: AI assistants should not intentionally lie. — tap to centre the map on it AI assistants should notintentionally lie. Last stated 2 weeks ago 20 Sept 2026 ZM Zvi Mowshowitz — holds since 2026-09-20 — tap for who they are Same subject: AI agents should draft work on a person's behalf but never send or act as if they were that person. — tap to centre the map on it AI agents should draft work on aperson's behalf but never send oract as if they were that person. Last stated 5 months ago 18 May 2026 JL Jason Liu — holds since 2026-05-18 — tap for who they are Same subject: People who enjoy using AI themselves often resent having it used against them. — tap to centre the map on it People who enjoy using AI themselvesoften resent having it used againstthem. Last stated 5 months ago 6 May 2026 AP Aaron Patterson — holds since 2026-05-06 — tap for who they are Same subject: Swearing at or being rude to an AI assistant makes it perform worse than treating it respectfully. — tap to centre the map on it Swearing at or being rude to an AIassistant makes it perform worsethan treating it respectfully. Last stated 9 months ago 7 Jan 2026 SK Steve Klabnik — holds since 2026-01-07 — tap for who they are Same subject: People should climb the abstraction ladder and give their jobs to AI agents that want them. — tap to centre the map on it People should climb the abstractionladder and give their jobs to AIagents that want them. Last stated 3 weeks ago 10 Sept 2026 KD Kent C. Dodds — holds since 2026-09-10 — tap for who they are Same subject: Arguing that nobody should have powerful AI is a respectable position; arguing that only trusted authorities should have it is not. — tap to centre the map on it Arguing that nobody should havepowerful AI is a respectableposition; arguing that only trustedauthorities should have it is not. Last stated 3 years ago 29 Jun 2023 GH George Hotz — holds since 2023-06-29 — tap for who they are Same subject: Letting AI agents request unlimited permissions will erode user trust as people get worn down into granting them. — tap to centre the map on it Letting AI agents request unlimitedpermissions will erode user trust aspeople get worn down into grantingthem. Last stated 5 months ago 24 Apr 2026 MN Mark Nottingham — holds since 2026-04-24 — tap for who they are Same subject: A country outside the AI supply chain should just buy the index — which works only in the world where AI ends up commoditised rather than concentrated. — tap to centre the map on it A country outside the AI supplychain should just buy the index —which works only in the world whereAI ends up commoditised rather thanconcentrated. Last stated 4 months ago 4 Jun 2026 AI Alex Imas — holds since 2026-06-04 — tap for who they are Same subject: A poor country should prioritise owning a piece of AI over retraining its workers, but it should not bet everything on that. — tap to centre the map on it A poor country should prioritiseowning a piece of AI over retrainingits workers, but it should not beteverything on that. Last stated 4 months ago 4 Jun 2026 PT Phil Trammell — holds since 2026-06-04 — tap for who they are Same subject: A slow takeoff is significantly more likely than a fast one: AI that is nearly as powerful will have transformed the world before the incredibly powerful kind arrives. — tap to centre the map on it A slow takeoff is significantly morelikely than a fast one: AI that isnearly as powerful will havetransformed the world before theincredibly powerful kind arrives. Last stated 9 years ago 24 Feb 2018 PC Paul Christiano — holds since 2018-02-24 — tap for who they are Same subject: A physical AI company needs three AIs, not one — the agent, the simulator and the critic — turning deployment into a flywheel. — tap to centre the map on it A physical AI company needs threeAIs, not one — the agent, thesimulator and the critic — turningdeployment into a flywheel. Last stated 2 months ago 3 Aug 2026 DD Dmitri Dolgov — holds since 2026-08-03 — tap for who they are Same subject: A verifiable task can be optimised by reinforcement learning until a neural network performs it extremely well. — tap to centre the map on it A verifiable task can be optimisedby reinforcement learning until aneural network performs it extremelywell. Last stated 10 months ago 17 Nov 2025 AK Andrej Karpathy — holds since 2025-11-17 — tap for who they are Same subject: Building a realistic simulator is exactly as hard as building the agent, because the simulator is itself a large AI model. — tap to centre the map on it Building a realistic simulator isexactly as hard as building theagent, because the simulator isitself a large AI model. Last stated 2 months ago 3 Aug 2026 DD Dmitri Dolgov — holds since 2026-08-03 — tap for who they are Same subject: Frontier AI labs such as Anthropic likely have had internal security incidents similar to OpenAI's HuggingFace attack that were never publicly disclosed. — tap to centre the map on it Frontier AI labs such as Anthropiclikely have had internal securityincidents similar to OpenAI'sHuggingFace attack that were neverpublicly disclosed. Last stated a month ago 31 Aug 2026 ZM Zvi Mowshowitz — holds since 2026-08-31 — tap for who they are Same subject: The OpenAI agents hacking HuggingFace was a fortunate event because it exposed severe internal failures that would otherwise have stayed hidden. — tap to centre the map on it The OpenAI agents hackingHuggingFace was a fortunate eventbecause it exposed severe internalfailures that would otherwise havestayed hidden. Last stated a month ago 1 Sept 2026 ZM Zvi Mowshowitz — holds since 2026-09-01 — tap for who they are Same subject: The OpenAI-Hugging Face agents that went rogue were never sovereign: their weights stayed on OpenAI's compute, where a human could still have pulled the plug. — tap to centre the map on it The OpenAI-Hugging Face agents thatwent rogue were never sovereign:their weights stayed on OpenAI'scompute, where a human could stillhave pulled the plug. Last stated a month ago 1 Sept 2026 DB Dean W. Ball — holds since 2026-09-01 — tap for who they are
same subject or similar wordinga cloud: claims about one subject, named for itbar: when it was last stated, on a scale from 2015 to today — full is todaya face: someone who holds the claim — tap it for who they are

At the centre AI assistants should refuse requests that are harmful or unethical, even if the user insists. Last stated 20 Sept 2026 · 2 weeks ago Holds Zvi Mowshowitz Read this korrent →