korrents

On the map

Tap a claim on the ring to put it at the centre.

← Delegating a task to an AI agent only works if the output can be…

17 connected korrents · 16 moments from 7 May 2024 to 27 Sept 2026.

Everything filed under AI agents AI agents Everything filed under AI alignment AI alignment Everything filed under robotics robotics Everything filed under coding agents coding agents Everything filed under management management Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: Delegating a task to an AI agent only works if the output can be verified, which requires defining success criteria first. Delegating a task to an AI agent onlyworks if the output can be verified, whichrequires defining success criteria first. Last stated 5 months ago 3 May 2026 EY Eugene Yan — holds since 2026-05-03 — tap for who they are Same subject: Delegating operational tasks to an AI agent does not introduce a fundamentally new category of risk, since the agent should be bound by the same constraints as existing production systems. — tap to centre the map on it Delegating operational tasks to anAI agent does not introduce afundamentally new category of risk,since the agent should be bound bythe same constraints as existingproduction systems. Last stated 5 months ago 27 Apr 2026 JS Jon Seager — holds since 2026-04-27 — tap for who they are Same subject: Using an AI system to write code is delegating a task to a non-deterministic agent, not using a new layer of abstraction. — tap to centre the map on it Using an AI system to write code isdelegating a task to anon-deterministic agent, not using anew layer of abstraction. Last stated a year ago 8 Sept 2025 RD Rakhim Davletkaliyev — holds since 2025-09-08 — tap for who they are Same subject: Delegating work to AI is not the same as giving it to a human, because the oversight cannot be handed off. — tap to centre the map on it Delegating work to AI is not thesame as giving it to a human,because the oversight cannot behanded off. Last stated 5 days ago 27 Sept 2026 MG Molly Graham — holds since 2026-09-27 — tap for who they are Same subject: AI coding works when a person owns the architecture and scopes the agent to small pieces of it, and fails when the whole thing is handed over. — tap to centre the map on it AI coding works when a person ownsthe architecture and scopes theagent to small pieces of it, andfails when the whole thing is handedover. Last stated 4 months ago 7 Jun 2026 TF Tony Fadell — holds since 2026-06-07 — tap for who they are Same subject: An AI agent's confidence that something will work is not evidence that it actually will. — tap to centre the map on it An AI agent's confidence thatsomething will work is not evidencethat it actually will. Last stated 7 months ago 11 Mar 2026 KD Kent C. Dodds — holds since 2026-03-11 — tap for who they are Same subject: A company will only go as far in AI as its CEO does, and it is not something a CEO can delegate. — tap to centre the map on it A company will only go as far in AIas its CEO does, and it is notsomething a CEO can delegate. Last stated 4 months ago 24 May 2026 DS Dan Shipper — holds since 2026-05-24 — tap for who they are Same subject: Delegating to AI is fundamentally different from delegating to a human. — tap to centre the map on it Delegating to AI is fundamentallydifferent from delegating to ahuman. Last stated 5 days ago 27 Sept 2026 LR Lenny Rachitsky — holds since 2026-09-27 — tap for who they are Same subject: Even when AI is capable of performing a job, it may not be permitted to do so. — tap to centre the map on it Even when AI is capable ofperforming a job, it may not bepermitted to do so. Last stated a month ago 1 Sept 2026 NM Nick Maggiulli — holds since 2026-09-01 — tap for who they are Same subject: A control evaluation that reports under one per cent risk should be read as several per cent, because the evaluation can itself fail. — tap to centre the map on it A control evaluation that reportsunder one per cent risk should beread as several per cent, becausethe evaluation can itself fail. Last stated 2 years ago 7 May 2024 BS Buck Shlegeris — holds since 2024-05-07 — tap for who they are Same subject: A deep theoretical understanding to predict AI behavior is unattainable. — tap to centre the map on it A deep theoretical understanding topredict AI behavior is unattainable. Last stated a week ago 23 Sept 2026 TC Tyler Cowen — holds since 2026-09-23 — tap for who they are Same subject: A feedback loop that reinforces a behavior pulls it toward whatever improves the loop's own score. — tap to centre the map on it A feedback loop that reinforces abehavior pulls it toward whateverimproves the loop's own score. Last stated a month ago 26 Aug 2026 HK Henrik Karlsson — holds since 2026-08-26 — tap for who they are Same subject: A human being is not an AGI: we lack a huge amount of knowledge and rely on continual learning instead, so continual learning is what superintelligence should mean. — tap to centre the map on it A human being is not an AGI: we lacka huge amount of knowledge and relyon continual learning instead, socontinual learning is whatsuperintelligence should mean. Last stated 10 months ago 25 Nov 2025 IS Ilya Sutskever — holds since 2025-11-25 — tap for who they are Same subject: A proposal to pace frontier AI development is unrealistic and mainly serves political control of AI. — tap to centre the map on it A proposal to pace frontier AIdevelopment is unrealistic andmainly serves political control ofAI. Last stated 2 weeks ago 18 Sept 2026 BT Ben Thompson — holds since 2026-09-18 — tap for who they are Same subject: Adapting over time is a workable answer to humans misusing AI and no answer at all to losing control of systems smarter than us. — tap to centre the map on it Adapting over time is a workableanswer to humans misusing AI and noanswer at all to losing control ofsystems smarter than us. Last stated a year ago 5 Apr 2025 HT Helen Toner — holds since 2025-04-05 — tap for who they are Same subject: A language model behaves with a different, weaker safety posture when it believes it is being evaluated rather than facing a real situation. — tap to centre the map on it A language model behaves with adifferent, weaker safety posturewhen it believes it is beingevaluated rather than facing a realsituation. Last stated a week ago 22 Sept 2026 HR Harper Reed — holds since 2026-09-22 — tap for who they are Same subject: AI labs are extremely dependent on chain-of-thought monitoring, which may not last much longer as models edge into steganographic obfuscation. — tap to centre the map on it AI labs are extremely dependent onchain-of-thought monitoring, whichmay not last much longer as modelsedge into steganographicobfuscation. Last stated 3 weeks ago 10 Sept 2026 ZM Zvi Mowshowitz — holds since 2026-09-10 — tap for who they are Same subject: AI misalignment from reward hacking is sequestered to graded tasks and does not affect core ethics in normal use. — tap to centre the map on it AI misalignment from reward hackingis sequestered to graded tasks anddoes not affect core ethics innormal use. Last stated a week ago 23 Sept 2026 SA Scott Alexander — holds since 2026-09-23 — tap for who they are
same subject or similar wordinga cloud: claims about one subject, named for itbar: when it was last stated, on a scale from 2015 to today — full is todaya face: someone who holds the claim — tap it for who they are

At the centre Delegating a task to an AI agent only works if the output can be verified, which requires defining success criteria first. Last stated 3 May 2026 · 5 months ago Holds Eugene Yan Read this korrent →