A korrentour readingWhat is a korrent?
Models resemble each other because pre-training is the same everywhere; what differentiates labs now is RL and post-training.
Drawn from what Ilya Sutskever said
scaling laws Whether more compute keeps buying more capability, and where the curve now bends -- pre-training, post-training, test-time.