korrents

A korrentour readingWhat is a korrent?

Today's RL environments are better than 2024's because we learned what to build and put AI labour on building it, not because labs hired more human experts.

Drawn from what Ryan Greenblatt said

What this subject means

reinforcement learning Training by reward: environments, verifiable tasks, value functions, and whether it works or merely beats what came before.

A private bookmark. Not a position, and never counted.

What Ryan Greenblatt actually said

Word for word, with the source under each one. They did not write this page.

  1. Ryan Greenblatt

    Chief scientist at Redwood Research

    the reason why RL environments today are much better than they were in like you know 2024 is not that much because um we have hired way more human experts to make RL environments and is instead much more because we better know what how RL like what RL environments we even want to make and and like how we should structure them and also we're using huge amounts of AI labor to build RL environments.

Added to korrents 11 Aug 2026 · How quotes work · Something wrong? Tell us

Do you hold this korrent?Do you also believe this?

Sign in to record that you hold this, with a confidence number of your own.

Related korrents

Our reading — they may agree, disagree or merely touch the same thing. Closest first: a shared subject counts for most, then how near the wording is.

On the map

Loading the map… or open it on its own page

Open the map on its own page →