korrents

On the map

Tap a claim on the ring to put it at the centre.

← Work produced by one frontier model should be reviewed by a different…

17 connected korrents · 16 moments on record from 20 Jan 2021 to 1 Sept 2026.

Everything filed under LLMs LLMs Everything filed under open source open source Everything filed under software quality software quality Everything filed under coding agents coding agents Everything filed under physics physics Everything filed under formal proof formal proof Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: Work produced by one frontier model should be reviewed by a different one, for the same reason a good programmer's work improves under a good peer's review. Work produced by one frontier modelshould be reviewed by a differentone, for the same reason a goodprogrammer's work improves under agood peer's review. Last stated 2 weeks ago 26 Aug 2026 DH David Heinemeier Hansson — holds since 2026-08-26 — tap for who they are Same subject: Open-source models cannot keep frontier systems in check, because they will always be much dumber than the frontier. — tap to centre the map on it Open-source models cannot keepfrontier systems in check,because they will always bemuch dumber than the frontier. Last stated 6 days ago 1 Sept 2026 AC Ajeya Cotra — holds since 2026-09-01 — tap for who they are Same subject: Agentic code review raises the floor but cannot be trusted, because the model reading the code is the same model that wrote it, and it will tell you the code is great. — tap to centre the map on it Agentic code review raises thefloor but cannot be trusted,because the model reading thecode is the same model that… Last stated 2 months ago 15 Jul 2026 DH Dex Horthy — holds since 2026-07-15 — tap for who they are Same subject: Line-by-line code review has become an outdated ritual: senior engineers should be giving feedback on how their team instructs AI, not on the code AI wrote. — tap to centre the map on it Line-by-line code review hasbecome an outdated ritual:senior engineers should begiving feedback on how their… Last stated 5 months ago 29 Mar 2026 CH Chip Huyen — holds since 2026-03-29 — tap for who they are Same subject: AI-validated pull requests beat human review at exactly the things humans are bad at, which frees the humans to argue about direction instead. — tap to centre the map on it AI-validated pull requestsbeat human review at exactlythe things humans are bad at,which frees the humans to… Last stated 4 weeks ago 12 Aug 2026 CM Charity Majors — holds since 2026-08-12 — tap for who they are Same subject: Today's models are the student who drilled 10,000 hours for the programming contest, which is exactly why what they learn does not carry anywhere else. — tap to centre the map on it Today's models are the studentwho drilled 10,000 hours forthe programming contest, whichis exactly why what they learn… Last stated 9 months ago 25 Nov 2025 IS Ilya Sutskever — holds since 2025-11-25 — tap for who they are Same subject: Lines of code are the wrong artifact to review; what we should be storing and reviewing is an architecture that generates the code to spec. — tap to centre the map on it Lines of code are the wrongartifact to review; what weshould be storing andreviewing is an architecture… Last stated 4 weeks ago 12 Aug 2026 CM Charity Majors — holds since 2026-08-12 — tap for who they are Same subject: Reviewers can tell the obviously good from the obviously bad, but there is a messy middle where they cannot tell at all. — tap to centre the map on it Reviewers can tell theobviously good from theobviously bad, but there is amessy middle where they cannot… Last stated 6 years ago 20 Jan 2021 JR José Luis Ricón — holds since 2021-01-20 — tap for who they are Same subject: A frontier lab should not spend its researchers harvesting results out of today's models; the payoff is in building the next one. — tap to centre the map on it A frontier lab should notspend its researchersharvesting results out oftoday's models; the payoff is… Last stated 2 months ago 26 Jun 2026 NB Noam Brown — holds since 2026-06-26 — tap for who they are Same subject: A language model is no substitute for a well-specified conventional algorithm, so it cannot simply be dropped into a complex problem and trusted. — tap to centre the map on it A language model is nosubstitute for awell-specified conventionalalgorithm, so it cannot simply… Last stated a year ago 7 Jun 2025 GM Gary Marcus — holds since 2025-06-07 — tap for who they are Same subject: A language model is not using language at all, because language requires an intention to communicate. — tap to centre the map on it A language model is not usinglanguage at all, becauselanguage requires an intentionto communicate. Last stated 2 years ago 31 Aug 2024 TC Ted Chiang — holds since 2024-08-31 — tap for who they are Same subject: A language model's apparent mind is mostly our own bias: it predicts text, and leverages our evolved habit of attributing intentionality to anything that acts human. — tap to centre the map on it A language model's apparentmind is mostly our own bias:it predicts text, andleverages our evolved habit of… Last stated 2 years ago 22 Apr 2024 SC Sean Carroll — holds since 2024-04-22 — tap for who they are Same subject: AI-written code makes formal proof necessary, because human review of all that generated code becomes the bottleneck. — tap to centre the map on it AI-written code makes formalproof necessary, because humanreview of all that generatedcode becomes the bottleneck. Last stated 5 months ago 22 Apr 2026 MK Martin Kleppmann — holds since 2026-04-22 — tap for who they are Same subject: Formal verification is about to become economical, because models are getting good enough at writing the proofs that humans no longer have to. — tap to centre the map on it Formal verification is aboutto become economical, becausemodels are getting good enoughat writing the proofs that… Last stated 5 months ago 22 Apr 2026 MK Martin Kleppmann — holds since 2026-04-22 — tap for who they are Same subject: Formalising a proof in Lean currently takes about ten times the effort of writing it out: doable, but annoying. — tap to centre the map on it Formalising a proof in Leancurrently takes about tentimes the effort of writing itout: doable, but annoying. Last stated a year ago 14 Jun 2025 TT Terence Tao — holds since 2025-06-14 — tap for who they are Same subject: LLMs can write a large fraction of the tedious code a developer will ever need to write, and most code on most projects is tedious. — tap to centre the map on it LLMs can write a largefraction of the tedious code adeveloper will ever need towrite, and most code on most… Last stated a year ago 2 Jun 2025 TP Thomas Ptacek — holds since 2025-06-02 — tap for who they are Same subject: Once a language model reads the results, search can trade precision for recall, because the model does not care that the right link came ninth. — tap to centre the map on it Once a language model readsthe results, search can tradeprecision for recall, becausethe model does not care that… Last stated 2 years ago 19 Jun 2024 AS Aravind Srinivas — holds since 2024-06-19 — tap for who they are Same subject: The best use of an LLM for learning is as a souped-up search engine that points you at the right human-written resource. — tap to centre the map on it The best use of an LLM forlearning is as a souped-upsearch engine that points youat the right human-written… Last stated 2 months ago 30 Jun 2026 GS Grant Sanderson — holds since 2026-06-30 — tap for who they are
same subject or similar wordinga cloud: claims about one subject, named for itbar: how recently it was last stated — full and dark this week, a faint sliver at five yearsa face: someone on record holding the claim — tap it for who they are

At the centre Work produced by one frontier model should be reviewed by a different one, for the same reason a good programmer's work improves under a good peer's review. Last stated 26 Aug 2026 · 2 weeks ago Holds David Heinemeier Hansson Read this korrent →