korrents

On the map

Tap a claim on the ring to put it at the centre.

← Elaborate agent scaffolding is unnecessary: a task plus a way to check…

17 connected korrents · 15 moments on record from 16 Oct 2025 to 3 Sept 2026.

Everything filed under Anthropic Anthropic Everything filed under coding agents coding agents Everything filed under benchmarks benchmarks Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: Elaborate agent scaffolding is unnecessary: a task plus a way to check the output is enough to keep a model working for weeks. Elaborate agent scaffolding isunnecessary: a task plus a way to checkthe output is enough to keep a modelworking for weeks. Last stated a month ago 27 Jul 2026 BC Boris Cherny — holds since 2026-07-27 — tap for who they are Same subject: The code today’s models and agents write is very hard to follow: you get better performance without knowing why, and no way to fix it when it breaks. — tap to centre the map on it The code today’s models and agentswrite is very hard to follow: youget better performance withoutknowing why, and no way to fix itwhen it breaks. Last stated 2 months ago 19 Jul 2026 ES Elizabeth Stone — holds since 2026-07-19 — tap for who they are Same subject: What limits you with coding agents is your own skill at stringing them together, not the capability of the models. — tap to centre the map on it What limits you with coding agentsis your own skill at stringing themtogether, not the capability of themodels. Last stated 6 months ago 20 Mar 2026 AK Andrej Karpathy — holds since 2026-03-20 — tap for who they are Same subject: The goal is not a better prompt but your own absence: arrange it once, hit go, and keep agents running for long stretches without you. — tap to centre the map on it The goal is not a better prompt butyour own absence: arrange it once,hit go, and keep agents running forlong stretches without you. Last stated 6 months ago 20 Mar 2026 AK Andrej Karpathy — holds since 2026-03-20 — tap for who they are Same subject: AI does not remove a job so much as change what it consists of, and knowing exactly where a model's capabilities stop is now part of being a good engineer. — tap to centre the map on it AI does not remove a job so much aschange what it consists of, andknowing exactly where a model'scapabilities stop is now part ofbeing a good engineer. Last stated 11 months ago 16 Oct 2025 DF Dylan Field — holds since 2025-10-16 — tap for who they are Same subject: When an AI agent runs out of forward progress, throw the project away and start over rather than trying to tweak it. — tap to centre the map on it When an AI agent runs out of forwardprogress, throw the project away andstart over rather than trying totweak it. Last stated 2 months ago 1 Jul 2026 KB Kent Beck — holds since 2026-07-01 — tap for who they are Same subject: Running a model until its performance plateaus is no longer a usable evaluation rule, because a well-scaffolded model keeps improving for weeks. — tap to centre the map on it Running a model until itsperformance plateaus is no longer ausable evaluation rule, because awell-scaffolded model keepsimproving for weeks. Last stated 2 months ago 26 Jun 2026 NB Noam Brown — holds since 2026-06-26 — tap for who they are Same subject: Every abstraction in agentic programming — RAG, memory, agentic history, structured output — is just a different way of passing tokens into a model, and understanding that beats learning any of them. — tap to centre the map on it Every abstraction in agenticprogramming — RAG, memory, agentichistory, structured output — is justa different way of passing tokensinto a model, and understanding thatbeats learning any of them. Last stated 2 months ago 15 Jul 2026 DH Dex Horthy — holds since 2026-07-15 — tap for who they are Same subject: Verifying a simplified model rather than the real system is worth doing even though the real system will still have bugs, because the bugs you designed in never get built. — tap to centre the map on it Verifying a simplified model ratherthan the real system is worth doingeven though the real system willstill have bugs, because the bugsyou designed in never get built. Last stated a month ago 29 Jul 2026 HW Hillel Wayne — holds since 2026-07-29 — tap for who they are Same subject: A benchmark result should be reported under a stated budget, or as a curve against test-time compute — never as a single number. — tap to centre the map on it A benchmark result should bereported under a stated budget, oras a curve against test-time compute— never as a single number. Last stated 2 months ago 26 Jun 2026 NB Noam Brown — holds since 2026-06-26 — tap for who they are Same subject: A benchmark that ranks Claude Code last while it stays first in use is measuring the wrong thing, and has been for a year. — tap to centre the map on it A benchmark that ranks Claude Codelast while it stays first in use ismeasuring the wrong thing, and hasbeen for a year. Last stated 4 days ago 3 Sept 2026 DR Dax Raad — holds since 2026-09-03 — tap for who they are Same subject: A company's staff-engineer bar should be set against the best companies in the industry rather than against its own history, which is what makes title inflation a real cost. — tap to centre the map on it A company's staff-engineer barshould be set against the bestcompanies in the industry ratherthan against its own history, whichis what makes title inflation a realcost. Last stated 5 months ago 1 Apr 2026 TP Thuan Pham — holds since 2026-04-01 — tap for who they are Same subject: A coding agent does not learn from its mistakes the way a person does -- it repeats the same error indefinitely unless a human notices and writes it down. — tap to centre the map on it A coding agent does not learn fromits mistakes the way a person does-- it repeats the same errorindefinitely unless a human noticesand writes it down. Last stated 5 months ago 25 Mar 2026 MZ Mario Zechner — holds since 2026-03-25 — tap for who they are Same subject: A coding-agent company should not train its own model: it has to stay neutral ground for models to compete on. — tap to centre the map on it A coding-agent company should nottrain its own model: it has to stayneutral ground for models to competeon. Last stated 4 days ago 3 Sept 2026 DR Dax Raad — holds since 2026-09-03 — tap for who they are Same subject: A machine can complete the task and completely miss the job -- coding makes this easy to see precisely because its tasks are so legible and verifiable. — tap to centre the map on it A machine can complete the task andcompletely miss the job -- codingmakes this easy to see preciselybecause its tasks are so legible andverifiable. Last stated 3 weeks ago 14 Aug 2026 SP Sunil Pai — holds since 2026-08-14 — tap for who they are Same subject: A language model asked to summarize its own system prompt risks that prompt's content biasing the summary it produces. — tap to centre the map on it A language model asked to summarizeits own system prompt risks thatprompt's content biasing the summaryit produces. Last stated 5 days ago 2 Sept 2026 SW Simon Willison — holds since 2026-09-02 — tap for who they are Same subject: A trust that owns the mission protects a company better than founder control does, which is why Anthropic needs no dual-class shares. — tap to centre the map on it A trust that owns the missionprotects a company better thanfounder control does, which is whyAnthropic needs no dual-classshares. Last stated 4 months ago 10 May 2026 ER Eric Ries — holds since 2026-05-10 — tap for who they are Same subject: Agents can build about half a million lines before the codebase dissolves into a mess, and the next model will push that to a few million. — tap to centre the map on it Agents can build about half amillion lines before the codebasedissolves into a mess, and the nextmodel will push that to a fewmillion. Last stated 6 months ago 11 Mar 2026 SY Steve Yegge — holds since 2026-03-11 — tap for who they are
same subject or similar wordinga cloud: claims about one subject, named for itbar: when it was last stated, on a scale from 2015 to today — full is todaya face: someone on record holding the claim — tap it for who they are

At the centre Elaborate agent scaffolding is unnecessary: a task plus a way to check the output is enough to keep a model working for weeks. Last stated 27 Jul 2026 · a month ago Holds Boris Cherny Read this korrent →