Tap a claim on the ring to put it at the centre.
← Elaborate agent scaffolding is unnecessary: a task plus a way to check…
17 connected korrents · 15 moments on record from 16 Oct 2025 to 3 Sept 2026.
Everything filed under Anthropic
Anthropic
Everything filed under coding agents
coding agents
Everything filed under benchmarks
benchmarks
Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject
Read this korrent: Elaborate agent scaffolding is unnecessary: a task plus a way to check the output is enough to keep a model working for weeks.
Elaborate agent scaffolding is unnecessary: a task plus a way to check the output is enough to keep a model working for weeks.
Last stated a month ago
27 Jul 2026
BC
Boris Cherny — holds since 2026-07-27 — tap for who they are
Same subject: The code today’s models and agents write is very hard to follow: you get better performance without knowing why, and no way to fix it when it breaks. — tap to centre the map on it
The code today’s models and agents write is very hard to follow: you get better performance without knowing why, and no way to fix it when it breaks.
Last stated 2 months ago
19 Jul 2026
ES
Elizabeth Stone — holds since 2026-07-19 — tap for who they are
Same subject: What limits you with coding agents is your own skill at stringing them together, not the capability of the models. — tap to centre the map on it
What limits you with coding agents is your own skill at stringing them together, not the capability of the models.
Last stated 6 months ago
20 Mar 2026
AK
Andrej Karpathy — holds since 2026-03-20 — tap for who they are
Same subject: The goal is not a better prompt but your own absence: arrange it once, hit go, and keep agents running for long stretches without you. — tap to centre the map on it
The goal is not a better prompt but your own absence: arrange it once, hit go, and keep agents running for long stretches without you.
Last stated 6 months ago
20 Mar 2026
AK
Andrej Karpathy — holds since 2026-03-20 — tap for who they are
Same subject: AI does not remove a job so much as change what it consists of, and knowing exactly where a model's capabilities stop is now part of being a good engineer. — tap to centre the map on it
AI does not remove a job so much as change what it consists of, and knowing exactly where a model's capabilities stop is now part of being a good engineer.
Last stated 11 months ago
16 Oct 2025
DF
Dylan Field — holds since 2025-10-16 — tap for who they are
Same subject: When an AI agent runs out of forward progress, throw the project away and start over rather than trying to tweak it. — tap to centre the map on it
When an AI agent runs out of forward progress, throw the project away and start over rather than trying to tweak it.
Last stated 2 months ago
1 Jul 2026
KB
Kent Beck — holds since 2026-07-01 — tap for who they are
Same subject: Running a model until its performance plateaus is no longer a usable evaluation rule, because a well-scaffolded model keeps improving for weeks. — tap to centre the map on it
Running a model until its performance plateaus is no longer a usable evaluation rule, because a well-scaffolded model keeps improving for weeks.
Last stated 2 months ago
26 Jun 2026
NB
Noam Brown — holds since 2026-06-26 — tap for who they are
Same subject: Every abstraction in agentic programming — RAG, memory, agentic history, structured output — is just a different way of passing tokens into a model, and understanding that beats learning any of them. — tap to centre the map on it
Every abstraction in agentic programming — RAG, memory, agentic history, structured output — is just a different way of passing tokens into a model, and understanding that beats learning any of them.
Last stated 2 months ago
15 Jul 2026
DH
Dex Horthy — holds since 2026-07-15 — tap for who they are
Same subject: Verifying a simplified model rather than the real system is worth doing even though the real system will still have bugs, because the bugs you designed in never get built. — tap to centre the map on it
Verifying a simplified model rather than the real system is worth doing even though the real system will still have bugs, because the bugs you designed in never get built.
Last stated a month ago
29 Jul 2026
HW
Hillel Wayne — holds since 2026-07-29 — tap for who they are
Same subject: A benchmark result should be reported under a stated budget, or as a curve against test-time compute — never as a single number. — tap to centre the map on it
A benchmark result should be reported under a stated budget, or as a curve against test-time compute — never as a single number.
Last stated 2 months ago
26 Jun 2026
NB
Noam Brown — holds since 2026-06-26 — tap for who they are
Same subject: A benchmark that ranks Claude Code last while it stays first in use is measuring the wrong thing, and has been for a year. — tap to centre the map on it
A benchmark that ranks Claude Code last while it stays first in use is measuring the wrong thing, and has been for a year.
Last stated 4 days ago
3 Sept 2026
DR
Dax Raad — holds since 2026-09-03 — tap for who they are
Same subject: A company's staff-engineer bar should be set against the best companies in the industry rather than against its own history, which is what makes title inflation a real cost. — tap to centre the map on it
A company's staff-engineer bar should be set against the best companies in the industry rather than against its own history, which is what makes title inflation a real cost.
Last stated 5 months ago
1 Apr 2026
TP
Thuan Pham — holds since 2026-04-01 — tap for who they are
Same subject: A coding agent does not learn from its mistakes the way a person does -- it repeats the same error indefinitely unless a human notices and writes it down. — tap to centre the map on it
A coding agent does not learn from its mistakes the way a person does -- it repeats the same error indefinitely unless a human notices and writes it down.
Last stated 5 months ago
25 Mar 2026
MZ
Mario Zechner — holds since 2026-03-25 — tap for who they are
Same subject: A coding-agent company should not train its own model: it has to stay neutral ground for models to compete on. — tap to centre the map on it
A coding-agent company should not train its own model: it has to stay neutral ground for models to compete on.
Last stated 4 days ago
3 Sept 2026
DR
Dax Raad — holds since 2026-09-03 — tap for who they are
Same subject: A machine can complete the task and completely miss the job -- coding makes this easy to see precisely because its tasks are so legible and verifiable. — tap to centre the map on it
A machine can complete the task and completely miss the job -- coding makes this easy to see precisely because its tasks are so legible and verifiable.
Last stated 3 weeks ago
14 Aug 2026
SP
Sunil Pai — holds since 2026-08-14 — tap for who they are
Same subject: A language model asked to summarize its own system prompt risks that prompt's content biasing the summary it produces. — tap to centre the map on it
A language model asked to summarize its own system prompt risks that prompt's content biasing the summary it produces.
Last stated 5 days ago
2 Sept 2026
SW
Simon Willison — holds since 2026-09-02 — tap for who they are
Same subject: A trust that owns the mission protects a company better than founder control does, which is why Anthropic needs no dual-class shares. — tap to centre the map on it
A trust that owns the mission protects a company better than founder control does, which is why Anthropic needs no dual-class shares.
Last stated 4 months ago
10 May 2026
ER
Eric Ries — holds since 2026-05-10 — tap for who they are
Same subject: Agents can build about half a million lines before the codebase dissolves into a mess, and the next model will push that to a few million. — tap to centre the map on it
Agents can build about half a million lines before the codebase dissolves into a mess, and the next model will push that to a few million.
Last stated 6 months ago
11 Mar 2026
SY
Steve Yegge — holds since 2026-03-11 — tap for who they are
same subject or similar wording a cloud: claims about one subject, named for it bar: when it was last stated, on a scale from 2015 to today — full is today a face: someone on record holding the claim — tap it for who they are
At the centre
Elaborate agent scaffolding is unnecessary: a task plus a way to check the output is enough to keep a model working for weeks.
Last stated 27 Jul 2026 · a month ago
Holds Boris Cherny
Read this korrent →
Similar wording
The code today’s models and agents write is very hard to follow: you get better performance without knowing why, and no way to fix it when it breaks.
Last stated 19 Jul 2026 · 2 months ago
Holds ES Elizabeth Stone
Similar wording
What limits you with coding agents is your own skill at stringing them together, not the capability of the models.
Last stated 20 Mar 2026 · 6 months ago
Holds Andrej Karpathy
Similar wording
The goal is not a better prompt but your own absence: arrange it once, hit go, and keep agents running for long stretches without you.
Last stated 20 Mar 2026 · 6 months ago
Holds Andrej Karpathy
Similar wording
AI does not remove a job so much as change what it consists of, and knowing exactly where a model's capabilities stop is now part of being a good engineer.
Last stated 16 Oct 2025 · 11 months ago
Holds DF Dylan Field
Similar wording
When an AI agent runs out of forward progress, throw the project away and start over rather than trying to tweak it.
Last stated 1 Jul 2026 · 2 months ago
Holds KB Kent Beck
Similar wording
Running a model until its performance plateaus is no longer a usable evaluation rule, because a well-scaffolded model keeps improving for weeks.
Last stated 26 Jun 2026 · 2 months ago
Holds NB Noam Brown
Similar wording
Every abstraction in agentic programming — RAG, memory, agentic history, structured output — is just a different way of passing tokens into a model, and understanding that beats learning any of them.
Last stated 15 Jul 2026 · 2 months ago
Holds DH Dex Horthy
Similar wording
Verifying a simplified model rather than the real system is worth doing even though the real system will still have bugs, because the bugs you designed in never get built.
Last stated 29 Jul 2026 · a month ago
Holds HW Hillel Wayne
Same subject: benchmarks
A benchmark result should be reported under a stated budget, or as a curve against test-time compute — never as a single number.
Last stated 26 Jun 2026 · 2 months ago
Holds NB Noam Brown
Same subject: benchmarks
A benchmark that ranks Claude Code last while it stays first in use is measuring the wrong thing, and has been for a year.
Last stated 3 Sept 2026 · 4 days ago
Holds Dax Raad
Same subject: benchmarks
A company's staff-engineer bar should be set against the best companies in the industry rather than against its own history, which is what makes title inflation a real cost.
Last stated 1 Apr 2026 · 5 months ago
Holds TP Thuan Pham
Same subject: coding agents
A coding agent does not learn from its mistakes the way a person does -- it repeats the same error indefinitely unless a human notices and writes it down.
Last stated 25 Mar 2026 · 5 months ago
Holds Mario Zechner
Same subject: coding agents
A coding-agent company should not train its own model: it has to stay neutral ground for models to compete on.
Last stated 3 Sept 2026 · 4 days ago
Holds Dax Raad
Same subject: coding agents
A machine can complete the task and completely miss the job -- coding makes this easy to see precisely because its tasks are so legible and verifiable.
Last stated 14 Aug 2026 · 3 weeks ago
Holds Sunil Pai
Same subject: Anthropic
A language model asked to summarize its own system prompt risks that prompt's content biasing the summary it produces.
Last stated 2 Sept 2026 · 5 days ago
Holds Simon Willison
Same subject: Anthropic
A trust that owns the mission protects a company better than founder control does, which is why Anthropic needs no dual-class shares.
Last stated 10 May 2026 · 4 months ago
Holds ER Eric Ries
Same subject: Anthropic
Agents can build about half a million lines before the codebase dissolves into a mess, and the next model will push that to a few million.
Last stated 11 Mar 2026 · 6 months ago
Holds SY Steve Yegge