korrents

On the map

Tap a claim on the ring to put it at the centre.

← Today's models are severely under-elicited: they are far more capable…

17 connected korrents · 15 moments on record from 12 Feb 2010 to 25 Aug 2026.

Everything filed under LLMs LLMs Everything filed under taste taste Everything filed under ambition ambition Everything filed under reinforcement learning reinforcement learning Everything filed under formal proof formal proof Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: Today's models are severely under-elicited: they are far more capable than the way we currently ask them makes them look. Today's models are severelyunder-elicited: they are far more capablethan the way we currently ask them makesthem look. Last stated 2 years ago 8 Nov 2024 JL Jan Leike — holds since 2024-11-08 — tap for who they are Same subject: Models have no research taste yet, which is why they complement researchers rather than replace them. — tap to centre the map on it Models have no research taste yet,which is why they complementresearchers rather than replacethem. Last stated 2 months ago 26 Jun 2026 NB Noam Brown — holds since 2026-06-26 — tap for who they are Same subject: Newer model releases have got worse at producing text that is actually pleasant to read and understand. — tap to centre the map on it Newer model releases have got worseat producing text that is actuallypleasant to read and understand. Last stated 2 weeks ago 25 Aug 2026 AR Armin Ronacher — holds since 2026-08-25 — tap for who they are Same subject: Today’s models are already powerful enough to add many points of GDP growth; what is missing is people with the vision and ambition to apply them. — tap to centre the map on it Today’s models are already powerfulenough to add many points of GDPgrowth; what is missing is peoplewith the vision and ambition toapply them. Last stated a month ago 29 Jul 2026 AW Alexandr Wang — holds since 2026-07-29 — tap for who they are Same subject: There are dozens or hundreds of things today's models can already do that nobody has discovered yet, and they are found by playing rather than by planning. — tap to centre the map on it There are dozens or hundreds ofthings today's models can already dothat nobody has discovered yet, andthey are found by playing ratherthan by planning. Last stated a month ago 27 Jul 2026 BC Boris Cherny — holds since 2026-07-27 — tap for who they are Same subject: Models look far better on evals than they are in the world because researchers, inadvertently, take inspiration from the evals when they build RL environments. — tap to centre the map on it Models look far better on evals thanthey are in the world becauseresearchers, inadvertently, takeinspiration from the evals when theybuild RL environments. Last stated 9 months ago 25 Nov 2025 IS Ilya Sutskever — holds since 2025-11-25 — tap for who they are Same subject: A model can produce version two of a product that already exists, but the genuinely new version one still has to be made by people. — tap to centre the map on it A model can produce version two of aproduct that already exists, but thegenuinely new version one still hasto be made by people. Last stated 3 months ago 7 Jun 2026 TF Tony Fadell — holds since 2026-06-07 — tap for who they are Same subject: A frontier model is measurably more intelligent with no system prompt at all; the prompts that remain are there for the product, not the model. — tap to centre the map on it A frontier model is measurably moreintelligent with no system prompt atall; the prompts that remain arethere for the product, not themodel. Last stated a month ago 27 Jul 2026 BC Boris Cherny — holds since 2026-07-27 — tap for who they are Same subject: The most fundamental thing wrong with today's models is not scale or efficiency but that they generalize dramatically worse than people do. — tap to centre the map on it The most fundamental thing wrongwith today's models is not scale orefficiency but that they generalizedramatically worse than people do. Last stated 9 months ago 25 Nov 2025 IS Ilya Sutskever — holds since 2025-11-25 — tap for who they are Same subject: A language model is no substitute for a well-specified conventional algorithm, so it cannot simply be dropped into a complex problem and trusted. — tap to centre the map on it A language model is no substitutefor a well-specified conventionalalgorithm, so it cannot simply bedropped into a complex problem andtrusted. Last stated a year ago 7 Jun 2025 GM Gary Marcus — holds since 2025-06-07 — tap for who they are Same subject: A language model is not using language at all, because language requires an intention to communicate. — tap to centre the map on it A language model is not usinglanguage at all, because languagerequires an intention tocommunicate. Last stated 2 years ago 31 Aug 2024 TC Ted Chiang — holds since 2024-08-31 — tap for who they are Same subject: A language model's apparent mind is mostly our own bias: it predicts text, and leverages our evolved habit of attributing intentionality to anything that acts human. — tap to centre the map on it A language model's apparent mind ismostly our own bias: it predictstext, and leverages our evolvedhabit of attributing intentionalityto anything that acts human. Last stated 2 years ago 22 Apr 2024 SC Sean Carroll — holds since 2024-04-22 — tap for who they are Same subject: A first draft should not be graded good or bad; it is only the material that taste then gets to act on. — tap to centre the map on it A first draft should not be gradedgood or bad; it is only the materialthat taste then gets to act on. Last stated a month ago 6 Aug 2026 GS George Saunders — holds since 2026-08-06 — tap for who they are Same subject: A lot of people can match a framework for taste; almost nobody can create one, and creating one is the rare skill. — tap to centre the map on it A lot of people can match aframework for taste; almost nobodycan create one, and creating one isthe rare skill. Last stated 11 months ago 16 Oct 2025 DF Dylan Field — holds since 2025-10-16 — tap for who they are Same subject: A product needs a soul, and for that it needs one person of great taste who is its living, breathing aspect and gets furious about every small detail. — tap to centre the map on it A product needs a soul, and for thatit needs one person of great tastewho is its living, breathing aspectand gets furious about every smalldetail. Last stated 17 years ago 12 Feb 2010 KS Karri Saarinen — holds since 2010-02-12 — tap for who they are Same subject: AI-written code makes formal proof necessary, because human review of all that generated code becomes the bottleneck. — tap to centre the map on it AI-written code makes formal proofnecessary, because human review ofall that generated code becomes thebottleneck. Last stated 5 months ago 22 Apr 2026 MK Martin Kleppmann — holds since 2026-04-22 — tap for who they are Same subject: Formal verification is about to become economical, because models are getting good enough at writing the proofs that humans no longer have to. — tap to centre the map on it Formal verification is about tobecome economical, because modelsare getting good enough at writingthe proofs that humans no longerhave to. Last stated 5 months ago 22 Apr 2026 MK Martin Kleppmann — holds since 2026-04-22 — tap for who they are Same subject: Formalising a proof in Lean currently takes about ten times the effort of writing it out: doable, but annoying. — tap to centre the map on it Formalising a proof in Leancurrently takes about ten times theeffort of writing it out: doable,but annoying. Last stated a year ago 14 Jun 2025 TT Terence Tao — holds since 2025-06-14 — tap for who they are
same subject or similar wordinga cloud: claims about one subject, named for itbar: when it was last stated, on a scale from 2010 to today (stretched back to the oldest claim here) — full is todaya face: someone on record holding the claim — tap it for who they are

At the centre Today's models are severely under-elicited: they are far more capable than the way we currently ask them makes them look. Last stated 8 Nov 2024 · 2 years ago Holds Jan Leike Read this korrent →