korrents

A korrentour readingWhat is a korrent?

Reinforcement learning leaves models sharply jagged: on the rails of a verifiable domain they are superintelligent, off them everything meanders.

Drawn from what Andrej Karpathy said

What this subject means

reinforcement learning Training by reward: environments, verifiable tasks, value functions, and whether it works or merely beats what came before.

A private bookmark. Not a position, and never counted.

What Andrej Karpathy actually said

Word for word, with the source under each one. They did not write this page.

  1. Andrej Karpathy

    Founding member of OpenAI and former director of AI at Tesla

    And so you're kind of like you're either on rails and you're part of the super intelligence circuits or you're not on rails and you're outside of the verifiable domains and suddenly everything kind of just like meanders.

Added to korrents 20 Mar 2026 · How quotes work · Something wrong? Tell us

Do you hold this korrent?Do you also believe this?

Sign in to record that you hold this, with a confidence number of your own.

Related korrents

Our reading — they may agree, disagree or merely touch the same thing. Closest first: a shared subject counts for most, then how near the wording is.

On the map

Loading the map… or open it on its own page

Open the map on its own page →