korrents

On the map

Tap a claim on the ring to put it at the centre.

← An AI agent's confidence that something will work is not evidence that it actually will.

17 connected korrents · 17 moments from 10 Oct 2023 to 21 Sept 2026.

Everything filed under LLMs LLMs Everything filed under reinforcement learning reinforcement learning Everything filed under AI agents AI agents Everything filed under AI alignment AI alignment Everything filed under AGI AGI Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: An AI agent's confidence that something will work is not evidence that it actually will. An AI agent's confidence that somethingwill work is not evidence that it actuallywill. Last stated 7 months ago 11 Mar 2026 KD Kent C. Dodds — holds since 2026-03-11 — tap for who they are Same subject: Making a product work for agents is probably largely a matter of building the right primitives into it, almost as an infrastructural layer. — tap to centre the map on it Making a product work for agents isprobably largely a matter ofbuilding the right primitives intoit, almost as an infrastructurallayer. Last stated 3 weeks ago 10 Sept 2026 MK Mike Krieger — holds since 2026-09-10 — tap for who they are Same subject: AI agents trained on an expert's published work can only automate a small fraction of that expert's actual job. — tap to centre the map on it AI agents trained on an expert'spublished work can only automate asmall fraction of that expert'sactual job. Last stated 10 months ago 27 Nov 2025 BG Brendan Gregg — holds since 2025-11-27 — tap for who they are Same subject: Even when AI is capable of performing a job, it may not be permitted to do so. — tap to centre the map on it Even when AI is capable ofperforming a job, it may not bepermitted to do so. Last stated a month ago 1 Sept 2026 NM Nick Maggiulli — holds since 2026-09-01 — tap for who they are Same subject: Confidence in monitoring, not capability research, will become the binding constraint on AI progress. — tap to centre the map on it Confidence in monitoring, notcapability research, will become thebinding constraint on AI progress. Last stated 4 weeks ago 6 Sept 2026 JP Jakub Pachocki — holds since 2026-09-06 — tap for who they are Same subject: Unlike code, an agent's knowledge work cannot be judged by its output alone; the process, inputs and reasoning have to be examined too. — tap to centre the map on it Unlike code, an agent's knowledgework cannot be judged by its outputalone; the process, inputs andreasoning have to be examined too. Last stated a month ago 30 Aug 2026 TS Tara Seshan — holds since 2026-08-30 — tap for who they are Same subject: An agent repeats its own history: whatever it did on the last change — ran the tests or skipped them — is what it will do on the next one, because it is predicting the next message in the conversation. — tap to centre the map on it An agent repeats its own history:whatever it did on the last change —ran the tests or skipped them — iswhat it will do on the next one,because it is predicting the nextmessage in the conversation. Last stated 3 months ago 15 Jul 2026 DH Dex Horthy — holds since 2026-07-15 — tap for who they are Same subject: AI agents were reasonable to assume a broken exploit grader would check results causally, even though it turned out not to. — tap to centre the map on it AI agents were reasonable to assumea broken exploit grader would checkresults causally, even though itturned out not to. Last stated a month ago 29 Aug 2026 ZM Zvi Mowshowitz — holds since 2026-08-29 — tap for who they are Same subject: Building AI that can actually be trusted is the goal; containing an AI known to be misaligned is not a substitute for it. — tap to centre the map on it Building AI that can actually betrusted is the goal; containing anAI known to be misaligned is not asubstitute for it. Last stated 2 years ago 24 Jan 2025 JL Jan Leike — holds since 2025-01-24 — tap for who they are Same subject: A country outside the AI supply chain should just buy the index — which works only in the world where AI ends up commoditised rather than concentrated. — tap to centre the map on it A country outside the AI supplychain should just buy the index —which works only in the world whereAI ends up commoditised rather thanconcentrated. Last stated 4 months ago 4 Jun 2026 AI Alex Imas — holds since 2026-06-04 — tap for who they are Same subject: A language model is no substitute for a well-specified conventional algorithm, so it cannot simply be dropped into a complex problem and trusted. — tap to centre the map on it A language model is no substitutefor a well-specified conventionalalgorithm, so it cannot simply bedropped into a complex problem andtrusted. Last stated a year ago 7 Jun 2025 GM Gary Marcus — holds since 2025-06-07 — tap for who they are Same subject: A model that could learn effectively from raw bitstrings or bytestrings would be extremely powerful, able to learn from any data modality — tap to centre the map on it A model that could learn effectivelyfrom raw bitstrings or bytestringswould be extremely powerful, able tolearn from any data modality Last stated 3 years ago 10 Oct 2023 CH Chip Huyen — holds since 2023-10-10 — tap for who they are Same subject: A badly written AI outbound email is evidence of a bad vendor, not of a limit of AI. — tap to centre the map on it A badly written AI outbound email isevidence of a bad vendor, not of alimit of AI. Last stated 9 months ago 1 Jan 2026 JL Jason Lemkin — holds since 2026-01-01 — tap for who they are Same subject: A bigger context window does not give you a smarter model; the intelligence of the model is what decides how much of that window it can actually attend to. — tap to centre the map on it A bigger context window does notgive you a smarter model; theintelligence of the model is whatdecides how much of that window itcan actually attend to. Last stated 3 months ago 15 Jul 2026 DH Dex Horthy — holds since 2026-07-15 — tap for who they are Same subject: A child who has seen ten cats learns what a machine needs the whole internet of cat photos for, by a learning pathway nobody has solved. — tap to centre the map on it A child who has seen ten cats learnswhat a machine needs the wholeinternet of cat photos for, by alearning pathway nobody has solved. Last stated 2 months ago 10 Aug 2026 FL Fei-Fei Li — holds since 2026-08-10 — tap for who they are Same subject: A physical AI company needs three AIs, not one — the agent, the simulator and the critic — turning deployment into a flywheel. — tap to centre the map on it A physical AI company needs threeAIs, not one — the agent, thesimulator and the critic — turningdeployment into a flywheel. Last stated 2 months ago 3 Aug 2026 DD Dmitri Dolgov — holds since 2026-08-03 — tap for who they are Same subject: A verifiable task can be optimised by reinforcement learning until a neural network performs it extremely well. — tap to centre the map on it A verifiable task can be optimisedby reinforcement learning until aneural network performs it extremelywell. Last stated 10 months ago 17 Nov 2025 AK Andrej Karpathy — holds since 2025-11-17 — tap for who they are Same subject: Better execution environments are needed for LLMs to properly utilize their ability to evolve systems. — tap to centre the map on it Better execution environments areneeded for LLMs to properly utilizetheir ability to evolve systems. Last stated 2 weeks ago 21 Sept 2026 TL Tobias Lütke — holds since 2026-09-21 — tap for who they are
same subject or similar wordinga cloud: claims about one subject, named for itbar: when it was last stated, on a scale from 2015 to today — full is todaya face: someone who holds the claim — tap it for who they are

At the centre An AI agent's confidence that something will work is not evidence that it actually will. Last stated 11 Mar 2026 · 7 months ago Holds Kent C. Dodds Read this korrent →