korrents

On the map

Tap a claim on the ring to put it at the centre.

← Screen technology cannot yet match the dynamic range and contrast of the real world.

17 connected korrents · 16 moments from 5 Nov 2019 to 2 Oct 2026.

Everything filed under benchmarks benchmarks Everything filed under ARC-AGI ARC-AGI Everything filed under Anthropic Anthropic Everything filed under LLMs LLMs Everything filed under value investing value investing Everything filed under AGI AGI Everything filed under video games video games Everything filed under startups startups Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: Screen technology cannot yet match the dynamic range and contrast of the real world. Screen technology cannot yet match thedynamic range and contrast of the realworld. Last stated 4 months ago 29 May 2026 JS John Siracusa — holds since 2026-05-29 — tap for who they are Same subject: Anthropic's model may not be good enough — tap to centre the map on it Anthropic's model may not be goodenough Last stated yesterday 2 Oct 2026 MS M.G. Siegler — holds since 2026-10-02 — tap for who they are Same subject: A widely adopted technology's ultimate effects never match the early expectations — tap to centre the map on it A widely adopted technology'sultimate effects never match theearly expectations Last stated a year ago 31 Aug 2025 NC Nicholas Carr — holds since 2025-08-31 — tap for who they are Same subject: Current AI models still fall short of expert human performance. — tap to centre the map on it Current AI models still fall shortof expert human performance. Last stated 2 years ago 27 Jan 2025 NE Nelson Elhage — holds since 2025-01-27 — tap for who they are Same subject: The technology side keeps changing, but the human side does not change that fast: people have the same problems. — tap to centre the map on it The technology side keeps changing,but the human side does not changethat fast: people have the sameproblems. Last stated 3 weeks ago 10 Sept 2026 AV Ami Vora — holds since 2026-09-10 — tap for who they are Same subject: The value-versus-growth distinction does not serve investors well in a fast-changing world. — tap to centre the map on it The value-versus-growth distinctiondoes not serve investors well in afast-changing world. Last stated 6 years ago 11 Jan 2021 HM Howard Marks — holds since 2021-01-11 — tap for who they are Same subject: AI may never reach the promised land: nobody understands how it works, and VR shows a vivid future can simply fail to arrive. — tap to centre the map on it AI may never reach the promisedland: nobody understands how itworks, and VR shows a vivid futurecan simply fail to arrive. Last stated 2 years ago 18 Sept 2024 DH David Heinemeier Hansson — holds since 2024-09-18 — tap for who they are Same subject: Games are hard to see clearly, and the usual tools for looking at them distort as much as they reveal. — tap to centre the map on it Games are hard to see clearly, andthe usual tools for looking at themdistort as much as they reveal. Last stated 2 years ago 25 Jul 2024 FL Frank Lantz — holds since 2024-07-25 — tap for who they are Same subject: A product can be technically impressive yet fail because it launched before a real market for it existed. — tap to centre the map on it A product can be technicallyimpressive yet fail because itlaunched before a real market for itexisted. Last stated 3 weeks ago 11 Sept 2026 MS M.G. Siegler — holds since 2026-09-11 — tap for who they are Same subject: A benchmark result should be reported under a stated budget, or as a curve against test-time compute — never as a single number. — tap to centre the map on it A benchmark result should bereported under a stated budget, oras a curve against test-time compute— never as a single number. Last stated 3 months ago 26 Jun 2026 NB Noam Brown — holds since 2026-06-26 — tap for who they are Same subject: A benchmark that ranks Claude Code last while it stays first in use is measuring the wrong thing, and has been for a year. — tap to centre the map on it A benchmark that ranks Claude Codelast while it stays first in use ismeasuring the wrong thing, and hasbeen for a year. Last stated 4 weeks ago 3 Sept 2026 DR Dax Raad — holds since 2026-09-03 — tap for who they are Same subject: A carmaker's claim to be the safest is mostly an artefact of comparing a new car against a fleet average twelve years old. — tap to centre the map on it A carmaker's claim to be the safestis mostly an artefact of comparing anew car against a fleet averagetwelve years old. Last stated 3 years ago 19 Dec 2023 PK Philip Koopman — holds since 2023-12-19 — tap for who they are Same subject: A model's capability is now a function of how much money you spend on it, so asking what a model can do means nothing until you name a budget. — tap to centre the map on it A model's capability is now afunction of how much money you spendon it, so asking what a model can domeans nothing until you name abudget. Last stated 3 months ago 26 Jun 2026 NB Noam Brown — holds since 2026-06-26 — tap for who they are Same subject: Language models perform better on benchmark problems released before their training data cutoff, indicating data contamination inflates scores. — tap to centre the map on it Language models perform better onbenchmark problems released beforetheir training data cutoff,indicating data contaminationinflates scores. Last stated 2 years ago 13 May 2024 SR Sebastian Ruder — holds since 2024-05-13 — tap for who they are Same subject: Model comparisons understate real progress, because benchmark tables do not control for how much test-time compute each answer used. — tap to centre the map on it Model comparisons understate realprogress, because benchmark tablesdo not control for how muchtest-time compute each answer used. Last stated 3 months ago 26 Jun 2026 NB Noam Brown — holds since 2026-06-26 — tap for who they are Same subject: The Abstraction and Reasoning Corpus can measure human-like general fluid intelligence and enable fair comparisons between AI systems and humans. — tap to centre the map on it The Abstraction and Reasoning Corpuscan measure human-like general fluidintelligence and enable faircomparisons between AI systems andhumans. Last stated 7 years ago 5 Nov 2019 FC François Chollet — holds since 2019-11-05 — tap for who they are Same subject: Combining an avocado and a chair into one image is evidence a model conceptually understands both, not that it memorised pictures of them. — tap to centre the map on it Combining an avocado and a chairinto one image is evidence a modelconceptually understands both, notthat it memorised pictures of them. Last stated 2 months ago 12 Aug 2026 CF Chelsea Finn — holds since 2026-08-12 — tap for who they are Same subject: Passing ARC-AGI does not amount to achieving AGI: o3 still fails on some very easy tasks, indicating fundamental differences from human intelligence. — tap to centre the map on it Passing ARC-AGI does not amount toachieving AGI: o3 still fails onsome very easy tasks, indicatingfundamental differences from humanintelligence. Last stated 2 years ago 20 Dec 2024 FC François Chollet — holds since 2024-12-20 — tap for who they are
same subject or similar wordinga cloud: claims about one subject, named for itbar: when it was last stated, on a scale from 2015 to today — full is todaya face: someone who holds the claim — tap it for who they are

At the centre Screen technology cannot yet match the dynamic range and contrast of the real world. Last stated 29 May 2026 · 4 months ago Holds John Siracusa Read this korrent →