Tap a claim on the ring to put it at the centre.
← AI learning to generate good conjectures will never show up as a…
17 connected korrents · 11 moments on record from 28 Oct 2011 to 3 Sept 2026. Nearly all of them are about benchmarks .
Everything filed under mathematics
mathematics
Everything filed under Anthropic
Anthropic
Everything filed under scaling laws
scaling laws
Everything filed under inflation
inflation
Everything filed under compilers
compilers
Everything filed under stock market
stock market
Everything filed under LLMs
LLMs
Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject
Read this korrent: AI learning to generate good conjectures will never show up as a benchmark being knocked down; it will show up as a shift in how mathematicians talk about the tools.
AI learning to generate good conjectures will never show up as a benchmark being knocked down; it will show up as a shift in how mathematicians talk about the tools.
Last stated 2 months ago
30 Jun 2026
GS
Grant Sanderson — holds since 2026-06-30 — tap for who they are
Same subject: A benchmark result should be reported under a stated budget, or as a curve against test-time compute — never as a single number. — tap to centre the map on it
A benchmark result should be reported under a stated budget, or as a curve against test-time compute — never as a single number.
Last stated 2 months ago
26 Jun 2026
NB
Noam Brown — holds since 2026-06-26 — tap for who they are
Same subject: A benchmark that ranks Claude Code last while it stays first in use is measuring the wrong thing, and has been for a year. — tap to centre the map on it
A benchmark that ranks Claude Code last while it stays first in use is measuring the wrong thing, and has been for a year.
Last stated 4 days ago
3 Sept 2026
DR
Dax Raad — holds since 2026-09-03 — tap for who they are
Same subject: A company's staff-engineer bar should be set against the best companies in the industry rather than against its own history, which is what makes title inflation a real cost. — tap to centre the map on it
A company's staff-engineer bar should be set against the best companies in the industry rather than against its own history, which is what makes title inflation a real cost.
Last stated 5 months ago
1 Apr 2026
TP
Thuan Pham — holds since 2026-04-01 — tap for who they are
Same subject: A speed difference between languages that share the LLVM backend measures the benchmark author, not the languages. — tap to centre the map on it
A speed difference between languages that share the LLVM backend measures the benchmark author, not the languages.
Last stated a year ago
22 Mar 2025
TH
ThePrimeagen — holds since 2025-03-22 — tap for who they are
Same subject: Benchmarks only rise on problems somebody has already framed and scored, so saturating them does not mean senior engineers have been replaced. — tap to centre the map on it
Benchmarks only rise on problems somebody has already framed and scored, so saturating them does not mean senior engineers have been replaced.
Last stated 3 months ago
24 May 2026
DS
Dan Shipper — holds since 2026-05-24 — tap for who they are
Same subject: Cash transfers are the index fund of development: the benchmark every actively managed aid programme should have to beat. — tap to centre the map on it
Cash transfers are the index fund of development: the benchmark every actively managed aid programme should have to beat.
Last stated 12 years ago
14 Mar 2014
CB
Chris Blattman — holds since 2014-03-14 — tap for who they are
Same subject: Every lab knows the benchmark grid is the wrong way to present a model, and publishes it anyway because everybody else does. — tap to centre the map on it
Every lab knows the benchmark grid is the wrong way to present a model, and publishes it anyway because everybody else does.
Last stated 2 months ago
26 Jun 2026
NB
Noam Brown — holds since 2026-06-26 — tap for who they are
Same subject: Google currently has no leading frontier AI model and no agentic coding tool comparable to Codex or Claude Code — tap to centre the map on it
Google currently has no leading frontier AI model and no agentic coding tool comparable to Codex or Claude Code
Last stated 2 months ago
23 Jul 2026
EM
Ethan Mollick — holds since 2026-07-23 — tap for who they are
Same subject: A proof and an explanation are different things, and a theorem can stay an unsolved expository problem long after it is proved. — tap to centre the map on it
A proof and an explanation are different things, and a theorem can stay an unsolved expository problem long after it is proved.
Last stated 2 months ago
30 Jun 2026
GS
Grant Sanderson — holds since 2026-06-30 — tap for who they are
Same subject: A stream of AI-written papers with any error rate at all becomes insufferable, because finding the error costs more than the paper is worth even at ninety-nine percent. — tap to centre the map on it
A stream of AI-written papers with any error rate at all becomes insufferable, because finding the error costs more than the paper is worth even at ninety-nine percent.
Last stated 2 months ago
30 Jun 2026
GS
Grant Sanderson — holds since 2026-06-30 — tap for who they are
Same subject: Academic credentials — grades, major, the prestige of the degree — barely matter to industry hiring. — tap to centre the map on it
Academic credentials — grades, major, the prestige of the degree — barely matter to industry hiring.
Last stated 15 years ago
28 Oct 2011
PM
Patrick McKenzie — holds since 2011-10-28 — tap for who they are
Same subject: AI could compete with human mathematicians once it acquires a mathematical sense of smell: knowing which way of splitting a problem makes it easier rather than harder. — tap to centre the map on it
AI could compete with human mathematicians once it acquires a mathematical sense of smell: knowing which way of splitting a problem makes it easier rather than harder.
Last stated a year ago
14 Jun 2025
TT
Terence Tao — holds since 2025-06-14 — tap for who they are
Same subject: AI in chess and mathematics does not explain anything; it says which position is better, and humans build the theory from that. — tap to centre the map on it
AI in chess and mathematics does not explain anything; it says which position is better, and humans build the theory from that.
Last stated a year ago
14 Jun 2025
LF
Lex Fridman — holds since 2025-06-14 — tap for who they are
Same subject: His prediction that research-level mathematics papers would be written in collaboration with AI by 2026 has already come true. — tap to centre the map on it
His prediction that research-level mathematics papers would be written in collaboration with AI by 2026 has already come true.
Last stated a year ago
14 Jun 2025
TT
Terence Tao — holds since 2025-06-14 — tap for who they are
Same subject: Lean and tools like GitHub will let experimental mathematics scale far beyond what one mathematician's spaghetti code allows today. — tap to centre the map on it
Lean and tools like GitHub will let experimental mathematics scale far beyond what one mathematician's spaghetti code allows today.
Last stated a year ago
14 Jun 2025
TT
Terence Tao — holds since 2025-06-14 — tap for who they are
Same subject: When formalising a proof costs no more than writing it, mathematics will flip: papers written in Lean first, and journals refereeing only for significance because correctness is certified. — tap to centre the map on it
When formalising a proof costs no more than writing it, mathematics will flip: papers written in Lean first, and journals refereeing only for significance because correctness is certified.
Last stated a year ago
14 Jun 2025
TT
Terence Tao — holds since 2025-06-14 — tap for who they are
Same subject: AI-written code makes formal proof necessary, because human review of all that generated code becomes the bottleneck. — tap to centre the map on it
AI-written code makes formal proof necessary, because human review of all that generated code becomes the bottleneck.
Last stated 5 months ago
22 Apr 2026
MK
Martin Kleppmann — holds since 2026-04-22 — tap for who they are
same subject or similar wording a cloud: claims about one subject, named for it bar: when it was last stated, on a scale from 2011 to today (stretched back to the oldest claim here) — full is today a face: someone on record holding the claim — tap it for who they are
At the centre
AI learning to generate good conjectures will never show up as a benchmark being knocked down; it will show up as a shift in how mathematicians talk about the tools.
Last stated 30 Jun 2026 · 2 months ago
Holds GS Grant Sanderson
Read this korrent →
Same subject: benchmarks
A benchmark result should be reported under a stated budget, or as a curve against test-time compute — never as a single number.
Last stated 26 Jun 2026 · 2 months ago
Holds NB Noam Brown
Same subject: benchmarks
A benchmark that ranks Claude Code last while it stays first in use is measuring the wrong thing, and has been for a year.
Last stated 3 Sept 2026 · 4 days ago
Holds Dax Raad
Same subject: benchmarks
A company's staff-engineer bar should be set against the best companies in the industry rather than against its own history, which is what makes title inflation a real cost.
Last stated 1 Apr 2026 · 5 months ago
Holds TP Thuan Pham
Same subject: benchmarks
A speed difference between languages that share the LLVM backend measures the benchmark author, not the languages.
Last stated 22 Mar 2025 · a year ago
Holds TH ThePrimeagen
Same subject: benchmarks
Benchmarks only rise on problems somebody has already framed and scored, so saturating them does not mean senior engineers have been replaced.
Last stated 24 May 2026 · 3 months ago
Holds DS Dan Shipper
Same subject: benchmarks
Cash transfers are the index fund of development: the benchmark every actively managed aid programme should have to beat.
Last stated 14 Mar 2014 · 12 years ago
Holds CB Chris Blattman
Same subject: benchmarks
Every lab knows the benchmark grid is the wrong way to present a model, and publishes it anyway because everybody else does.
Last stated 26 Jun 2026 · 2 months ago
Holds NB Noam Brown
Same subject: benchmarks
Google currently has no leading frontier AI model and no agentic coding tool comparable to Codex or Claude Code
Last stated 23 Jul 2026 · 2 months ago
Holds EM Ethan Mollick
Same subject: mathematics
A proof and an explanation are different things, and a theorem can stay an unsolved expository problem long after it is proved.
Last stated 30 Jun 2026 · 2 months ago
Holds GS Grant Sanderson
Same subject: mathematics
A stream of AI-written papers with any error rate at all becomes insufferable, because finding the error costs more than the paper is worth even at ninety-nine percent.
Last stated 30 Jun 2026 · 2 months ago
Holds GS Grant Sanderson
Same subject: mathematics
Academic credentials — grades, major, the prestige of the degree — barely matter to industry hiring.
Last stated 28 Oct 2011 · 15 years ago
Holds PM Patrick McKenzie
Same subject: AI and science
AI could compete with human mathematicians once it acquires a mathematical sense of smell: knowing which way of splitting a problem makes it easier rather than harder.
Last stated 14 Jun 2025 · a year ago
Holds TT Terence Tao
Same subject: AI and science
AI in chess and mathematics does not explain anything; it says which position is better, and humans build the theory from that.
Last stated 14 Jun 2025 · a year ago
Holds LF Lex Fridman
Same subject: AI and science
His prediction that research-level mathematics papers would be written in collaboration with AI by 2026 has already come true.
Last stated 14 Jun 2025 · a year ago
Holds TT Terence Tao
Same subject: formal proof
Lean and tools like GitHub will let experimental mathematics scale far beyond what one mathematician's spaghetti code allows today.
Last stated 14 Jun 2025 · a year ago
Holds TT Terence Tao
Same subject: formal proof
When formalising a proof costs no more than writing it, mathematics will flip: papers written in Lean first, and journals refereeing only for significance because correctness is certified.
Last stated 14 Jun 2025 · a year ago
Holds TT Terence Tao
Same subject: formal proof
AI-written code makes formal proof necessary, because human review of all that generated code becomes the bottleneck.
Last stated 22 Apr 2026 · 5 months ago
Holds MK Martin Kleppmann