korrents

A korrentour readingWhat is a korrent?

Chain-of-thought explanations from language models can be biased because training does not explicitly reward faithful reasoning.

Drawn from what Lilian Weng said

LLMs

What Lilian Weng actually said

Word for word, with the source under each one. They did not write this page.

  1. Lilian Weng

    Machine-learning researcher

    model CoTs could be biased due to lack of explicit training objectives aimed at encouraging faithful reasoning

Added to korrents 1 May 2025 · How quotes work · Something wrong? Tell us

Do you hold this korrent?Do you also believe this?

Sign in to record that you hold this, with a confidence number of your own.

Related korrents

Our reading — they may agree, disagree or merely touch the same thing. Closest first: a shared subject counts for most, then how near the wording is.

On the map

Loading the map… or open it on its own page

Open the map on its own page →