A korrentour readingWhat is a korrent?
Fine-tuning with human feedback is a promising direction for aligning language models with human intent.
Drawn from what Jan Leike said
SubjectLLMs
A korrentour readingWhat is a korrent?
Drawn from what Jan Leike said
Word for word, with the source under each one. They did not write this page.
Alignment researcher
Even though InstructGPT still makes simple mistakes, our results show that fine-tuning with human feedback is a promising direction for aligning language models with human intent.
Where this was said
↗Training language models to follow instructions with human feedback (with 19 co-authors)arxiv.org
All 4 korrents from this piecethis one is 3rd
Added to korrents 4 Mar 2022 · How quotes work · Something wrong? Tell us
Our reading — they may agree, disagree or merely touch the same thing. Closest first: a shared subject counts for most, then how near the wording is.