korrents

On the map

Tap a claim on the ring to put it at the centre.

← AI learning to generate good conjectures will never show up as a…

17 connected korrents · 11 moments on record from 28 Oct 2011 to 3 Sept 2026. Nearly all of them are about benchmarks.

Everything filed under mathematics mathematics Everything filed under Anthropic Anthropic Everything filed under scaling laws scaling laws Everything filed under inflation inflation Everything filed under compilers compilers Everything filed under stock market stock market Everything filed under LLMs LLMs Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: AI learning to generate good conjectures will never show up as a benchmark being knocked down; it will show up as a shift in how mathematicians talk about the tools. AI learning to generate good conjectureswill never show up as a benchmark beingknocked down; it will show up as a shiftin how mathematicians talk about thetools. Last stated 2 months ago 30 Jun 2026 GS Grant Sanderson — holds since 2026-06-30 — tap for who they are Same subject: A benchmark result should be reported under a stated budget, or as a curve against test-time compute — never as a single number. — tap to centre the map on it A benchmark result should bereported under a stated budget, oras a curve against test-time compute— never as a single number. Last stated 2 months ago 26 Jun 2026 NB Noam Brown — holds since 2026-06-26 — tap for who they are Same subject: A benchmark that ranks Claude Code last while it stays first in use is measuring the wrong thing, and has been for a year. — tap to centre the map on it A benchmark that ranks Claude Codelast while it stays first in use ismeasuring the wrong thing, and hasbeen for a year. Last stated 4 days ago 3 Sept 2026 DR Dax Raad — holds since 2026-09-03 — tap for who they are Same subject: A company's staff-engineer bar should be set against the best companies in the industry rather than against its own history, which is what makes title inflation a real cost. — tap to centre the map on it A company's staff-engineer barshould be set against the bestcompanies in the industry ratherthan against its own history, whichis what makes title inflation a realcost. Last stated 5 months ago 1 Apr 2026 TP Thuan Pham — holds since 2026-04-01 — tap for who they are Same subject: A speed difference between languages that share the LLVM backend measures the benchmark author, not the languages. — tap to centre the map on it A speed difference between languagesthat share the LLVM backend measuresthe benchmark author, not thelanguages. Last stated a year ago 22 Mar 2025 TH ThePrimeagen — holds since 2025-03-22 — tap for who they are Same subject: Benchmarks only rise on problems somebody has already framed and scored, so saturating them does not mean senior engineers have been replaced. — tap to centre the map on it Benchmarks only rise on problemssomebody has already framed andscored, so saturating them does notmean senior engineers have beenreplaced. Last stated 3 months ago 24 May 2026 DS Dan Shipper — holds since 2026-05-24 — tap for who they are Same subject: Cash transfers are the index fund of development: the benchmark every actively managed aid programme should have to beat. — tap to centre the map on it Cash transfers are the index fund ofdevelopment: the benchmark everyactively managed aid programmeshould have to beat. Last stated 12 years ago 14 Mar 2014 CB Chris Blattman — holds since 2014-03-14 — tap for who they are Same subject: Every lab knows the benchmark grid is the wrong way to present a model, and publishes it anyway because everybody else does. — tap to centre the map on it Every lab knows the benchmark gridis the wrong way to present a model,and publishes it anyway becauseeverybody else does. Last stated 2 months ago 26 Jun 2026 NB Noam Brown — holds since 2026-06-26 — tap for who they are Same subject: Google currently has no leading frontier AI model and no agentic coding tool comparable to Codex or Claude Code — tap to centre the map on it Google currently has no leadingfrontier AI model and no agenticcoding tool comparable to Codex orClaude Code Last stated 2 months ago 23 Jul 2026 EM Ethan Mollick — holds since 2026-07-23 — tap for who they are Same subject: A proof and an explanation are different things, and a theorem can stay an unsolved expository problem long after it is proved. — tap to centre the map on it A proof and an explanation aredifferent things, and a theorem canstay an unsolved expository problemlong after it is proved. Last stated 2 months ago 30 Jun 2026 GS Grant Sanderson — holds since 2026-06-30 — tap for who they are Same subject: A stream of AI-written papers with any error rate at all becomes insufferable, because finding the error costs more than the paper is worth even at ninety-nine percent. — tap to centre the map on it A stream of AI-written papers withany error rate at all becomesinsufferable, because finding theerror costs more than the paper isworth even at ninety-nine percent. Last stated 2 months ago 30 Jun 2026 GS Grant Sanderson — holds since 2026-06-30 — tap for who they are Same subject: Academic credentials — grades, major, the prestige of the degree — barely matter to industry hiring. — tap to centre the map on it Academic credentials — grades,major, the prestige of the degree —barely matter to industry hiring. Last stated 15 years ago 28 Oct 2011 PM Patrick McKenzie — holds since 2011-10-28 — tap for who they are Same subject: AI could compete with human mathematicians once it acquires a mathematical sense of smell: knowing which way of splitting a problem makes it easier rather than harder. — tap to centre the map on it AI could compete with humanmathematicians once it acquires amathematical sense of smell: knowingwhich way of splitting a problemmakes it easier rather than harder. Last stated a year ago 14 Jun 2025 TT Terence Tao — holds since 2025-06-14 — tap for who they are Same subject: AI in chess and mathematics does not explain anything; it says which position is better, and humans build the theory from that. — tap to centre the map on it AI in chess and mathematics does notexplain anything; it says whichposition is better, and humans buildthe theory from that. Last stated a year ago 14 Jun 2025 LF Lex Fridman — holds since 2025-06-14 — tap for who they are Same subject: His prediction that research-level mathematics papers would be written in collaboration with AI by 2026 has already come true. — tap to centre the map on it His prediction that research-levelmathematics papers would be writtenin collaboration with AI by 2026 hasalready come true. Last stated a year ago 14 Jun 2025 TT Terence Tao — holds since 2025-06-14 — tap for who they are Same subject: Lean and tools like GitHub will let experimental mathematics scale far beyond what one mathematician's spaghetti code allows today. — tap to centre the map on it Lean and tools like GitHub will letexperimental mathematics scale farbeyond what one mathematician'sspaghetti code allows today. Last stated a year ago 14 Jun 2025 TT Terence Tao — holds since 2025-06-14 — tap for who they are Same subject: When formalising a proof costs no more than writing it, mathematics will flip: papers written in Lean first, and journals refereeing only for significance because correctness is certified. — tap to centre the map on it When formalising a proof costs nomore than writing it, mathematicswill flip: papers written in Leanfirst, and journals refereeing onlyfor significance because correctnessis certified. Last stated a year ago 14 Jun 2025 TT Terence Tao — holds since 2025-06-14 — tap for who they are Same subject: AI-written code makes formal proof necessary, because human review of all that generated code becomes the bottleneck. — tap to centre the map on it AI-written code makes formal proofnecessary, because human review ofall that generated code becomes thebottleneck. Last stated 5 months ago 22 Apr 2026 MK Martin Kleppmann — holds since 2026-04-22 — tap for who they are
same subject or similar wordinga cloud: claims about one subject, named for itbar: when it was last stated, on a scale from 2011 to today (stretched back to the oldest claim here) — full is todaya face: someone on record holding the claim — tap it for who they are

At the centre AI learning to generate good conjectures will never show up as a benchmark being knocked down; it will show up as a shift in how mathematicians talk about the tools. Last stated 30 Jun 2026 · 2 months ago Holds Grant Sanderson Read this korrent →