korrents

On the map

Tap a claim on the ring to put it at the centre.

← The leading chatbots are undifferentiated: in a blind test most people…

17 connected korrents · 14 moments on record from 1 May 2023 to 3 Sept 2026.

Everything filed under chat interfaces chat interfaces Everything filed under benchmarks benchmarks Everything filed under coding agents coding agents Everything filed under OpenAI OpenAI Everything filed under Google Google Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: The leading chatbots are undifferentiated: in a blind test most people could not tell one model's output from another's. The leading chatbots are undifferentiated:in a blind test most people could not tellone model's output from another's. Last stated a year ago 2 Sept 2025 BE Benedict Evans — holds since 2025-09-02 — tap for who they are Same subject: The skill that matters now is not prompt engineering but giving a model a way to verify its own work, and it is the thing people most often get wrong. — tap to centre the map on it The skill that matters now is notprompt engineering but giving amodel a way to verify its own work,and it is the thing people mostoften get wrong. Last stated a month ago 27 Jul 2026 BC Boris Cherny — holds since 2026-07-27 — tap for who they are Same subject: Chatbot interfaces are the wrong tool for meaningful coding work, because you are mostly hoping the model recalls the right answer and correcting it requires a human in the loop. — tap to centre the map on it Chatbot interfaces are the wrongtool for meaningful coding work,because you are mostly hoping themodel recalls the right answer andcorrecting it requires a human inthe loop. Last stated 7 months ago 5 Feb 2026 MH Mitchell Hashimoto — holds since 2026-02-05 — tap for who they are Same subject: The chat box is not the final form of working with models; it is radio shows recorded onto television. — tap to centre the map on it The chat box is not the final formof working with models; it is radioshows recorded onto television. Last stated 7 months ago 12 Feb 2026 PS Peter Steinberger — holds since 2026-02-12 — tap for who they are Same subject: Anthropic saw the coming shortage of chips before Google did, and bought Google's own TPUs out from under it. — tap to centre the map on it Anthropic saw the coming shortage ofchips before Google did, and boughtGoogle's own TPUs out from under it. Last stated 6 months ago 13 Mar 2026 DP Dylan Patel — holds since 2026-03-13 — tap for who they are Same subject: Google currently has no leading frontier AI model and no agentic coding tool comparable to Codex or Claude Code — tap to centre the map on it Google currently has no leadingfrontier AI model and no agenticcoding tool comparable to Codex orClaude Code Last stated 2 months ago 23 Jul 2026 EM Ethan Mollick — holds since 2026-07-23 — tap for who they are Same subject: A blank chat box is lazy: it violates the first rule of a good user experience, that it is obvious what you can do. — tap to centre the map on it A blank chat box is lazy: itviolates the first rule of a gooduser experience, that it is obviouswhat you can do. Last stated a year ago 12 May 2025 JZ Julie Zhuo — holds since 2025-05-12 — tap for who they are Same subject: Around a tenth of people use a chatbot daily and another fifteen to twenty per cent weekly, far below the headline adoption figures. — tap to centre the map on it Around a tenth of people use achatbot daily and another fifteen totwenty per cent weekly, far belowthe headline adoption figures. Last stated a year ago 2 Sept 2025 BE Benedict Evans — holds since 2025-09-02 — tap for who they are Same subject: Chatbots are a terrible interface for large language models, because a text box has no affordances. — tap to centre the map on it Chatbots are a terrible interfacefor large language models, because atext box has no affordances. Last stated 3 years ago by 1 May 2023 AW Amelia Wattenberger — holds since 2023-05-01 — tap for who they are Same subject: A startup should not begin life as a nonprofit and bolt a for-profit arm on later, whatever OpenAI's own history suggests. — tap to centre the map on it A startup should not begin life as anonprofit and bolt a for-profit armon later, whatever OpenAI's ownhistory suggests. Last stated 2 years ago 18 Mar 2024 SA Sam Altman — holds since 2024-03-18 — tap for who they are Same subject: A technique that lets an AI model's reasoning shift outside its visible Chain of Thought is dangerous, both because it works and because a leading lab is willing to deploy it. — tap to centre the map on it A technique that lets an AI model'sreasoning shift outside its visibleChain of Thought is dangerous, bothbecause it works and because aleading lab is willing to deploy it. Last stated 5 days ago 3 Sept 2026 ZM Zvi Mowshowitz — holds since 2026-09-03 — tap for who they are Same subject: A tiny language model inventing a plausible-sounding name is the same phenomenon as a large one confidently stating a false fact. — tap to centre the map on it A tiny language model inventing aplausible-sounding name is the samephenomenon as a large oneconfidently stating a false fact. Last stated 7 months ago 12 Feb 2026 AK Andrej Karpathy — holds since 2026-02-12 — tap for who they are Same subject: A coding agent does not learn from its mistakes the way a person does -- it repeats the same error indefinitely unless a human notices and writes it down. — tap to centre the map on it A coding agent does not learn fromits mistakes the way a person does-- it repeats the same errorindefinitely unless a human noticesand writes it down. Last stated 5 months ago 25 Mar 2026 MZ Mario Zechner — holds since 2026-03-25 — tap for who they are Same subject: A coding-agent company should not train its own model: it has to stay neutral ground for models to compete on. — tap to centre the map on it A coding-agent company should nottrain its own model: it has to stayneutral ground for models to competeon. Last stated 5 days ago 3 Sept 2026 DR Dax Raad — holds since 2026-09-03 — tap for who they are Same subject: A machine can complete the task and completely miss the job -- coding makes this easy to see precisely because its tasks are so legible and verifiable. — tap to centre the map on it A machine can complete the task andcompletely miss the job -- codingmakes this easy to see preciselybecause its tasks are so legible andverifiable. Last stated 4 weeks ago 14 Aug 2026 SP Sunil Pai — holds since 2026-08-14 — tap for who they are Same subject: A benchmark result should be reported under a stated budget, or as a curve against test-time compute — never as a single number. — tap to centre the map on it A benchmark result should bereported under a stated budget, oras a curve against test-time compute— never as a single number. Last stated 2 months ago 26 Jun 2026 NB Noam Brown — holds since 2026-06-26 — tap for who they are Same subject: A model's capability is now a function of how much money you spend on it, so asking what a model can do means nothing until you name a budget. — tap to centre the map on it A model's capability is now afunction of how much money you spendon it, so asking what a model can domeans nothing until you name abudget. Last stated 2 months ago 26 Jun 2026 NB Noam Brown — holds since 2026-06-26 — tap for who they are Same subject: A speed difference between languages that share the LLVM backend measures the benchmark author, not the languages. — tap to centre the map on it A speed difference between languagesthat share the LLVM backend measuresthe benchmark author, not thelanguages. Last stated a year ago 22 Mar 2025 TH ThePrimeagen — holds since 2025-03-22 — tap for who they are
same subject or similar wordinga cloud: claims about one subject, named for itbar: when it was last stated, on a scale from 2015 to today — full is todaya face: someone on record holding the claim — tap it for who they are

At the centre The leading chatbots are undifferentiated: in a blind test most people could not tell one model's output from another's. Last stated 2 Sept 2025 · a year ago Holds Benedict Evans Read this korrent →