A korrentour readingWhat is a korrent?
Model comparisons understate real progress, because benchmark tables do not control for how much test-time compute each answer used.
Drawn from what Noam Brown said
What this subject means
scaling laws Whether more compute keeps buying more capability, and where the curve now bends -- pre-training, post-training, test-time.