korrents

A korrentour readingWhat is a korrent?

Fine-tuning with human feedback is a promising direction for aligning language models with human intent.

Drawn from what Jan Leike said

LLMs

What Jan Leike actually said

Word for word, with the source under each one. They did not write this page.

  1. Jan Leike

    Alignment researcher

    Even though InstructGPT still makes simple mistakes, our results show that fine-tuning with human feedback is a promising direction for aligning language models with human intent.

Added to korrents 4 Mar 2022 · How quotes work · Something wrong? Tell us

Do you hold this korrent?Do you also believe this?

Sign in to record that you hold this, with a confidence number of your own.

Related korrents

Our reading — they may agree, disagree or merely touch the same thing. Closest first: a shared subject counts for most, then how near the wording is.

On the map

Loading the map… or open it on its own page

Open the map on its own page →