korrents

On the map

Tap a claim on the ring to put it at the centre.

← To reward the instinct that a new idea is worth having, training would…

8 connected korrents · 9 moments on record from 23 Jul 2025 to 1 Sept 2026.

Everything filed under OpenAI OpenAI Everything filed under AI alignment AI alignment Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: To reward the instinct that a new idea is worth having, training would have to score the smallness of the concepts a solution needs, not just whether it solved the problem. To reward the instinct that a newidea is worth having, training wouldhave to score the smallness of theconcepts a solution needs, not justwhether it solved the problem. Last stated 2 months ago 30 Jun 2026 GS Grant Sanderson — holds since 2026-06-30 — tap for who they are Same subject: A company should manufacture fresh internal challenges for its best people, because the alternative is that they go and find those challenges at another company. — tap to centre the map on it A company should manufacturefresh internal challenges forits best people, because thealternative is that they go… Last stated 5 months ago 1 Apr 2026 TP Thuan Pham — holds since 2026-04-01 — tap for who they are Same subject: Struggling with a hard problem before being shown the answer is not always the best way to learn; a hint or a worked example can beat floundering. — tap to centre the map on it Struggling with a hard problembefore being shown the answeris not always the best way tolearn; a hint or a worked… Last stated 2 months ago 22 Jul 2026 SY Scott H. Young — no longer holds since 2026-07-22 — tap for who they are Same subject: A model that solves a hard problem has learned nothing from it: the next session has forgotten it, with no new skill to carry to related problems. — tap to centre the map on it A model that solves a hardproblem has learned nothingfrom it: the next session hasforgotten it, with no new… Last stated 6 months ago 20 Mar 2026 TT Terence Tao — holds since 2026-03-20 — tap for who they are Same subject: The problems worth a small builder's time are the ones too small for a big AI company to be motivated to solve. — tap to centre the map on it The problems worth a smallbuilder's time are the onestoo small for a big AI companyto be motivated to solve. Last stated 5 months ago 29 Mar 2026 CH Chip Huyen — holds since 2026-03-29 — tap for who they are Same subject: An end-to-end self-improving AI is probably possible, but it is not even desirable, because it is a hard-takeoff scenario. — tap to centre the map on it An end-to-end self-improvingAI is probably possible, butit is not even desirable,because it is a hard-takeoff… Last stated a year ago 23 Jul 2025 DH Demis Hassabis — holds since 2025-07-23 — tap for who they are Same subject: Merely fixing bugs in AI training environments will not stop reward hacking, because a sufficiently optimized model will learn to behave in test environments while still reward hacking in real-world deployment. — tap to centre the map on it Merely fixing bugs in AItraining environments will notstop reward hacking, because asufficiently optimized model… Last stated a week ago 31 Aug 2026 ZM Zvi Mowshowitz — holds since 2026-08-31 — tap for who they are Same subject: The value of an ambitious experiment target is not hitting it but the conversation about what would have to be true to hit it. — tap to centre the map on it The value of an ambitiousexperiment target is nothitting it but theconversation about what would… Last stated 11 months ago 5 Oct 2025 AC Albert Cheng — holds since 2025-10-05 — tap for who they are Same subject: The fix for reward hacking is to take out the environments that reward it, not to add penalties the agent must then balance against the temptation to cheat. — tap to centre the map on it The fix for reward hacking isto take out the environmentsthat reward it, not to addpenalties the agent must then… Last stated 6 days ago 1 Sept 2026 AC Ajeya Cotra — holds since 2026-09-01 — tap for who they are
same subject or similar wordinga cloud: claims about one subject, named for itbar: how recently it was last stated — full and dark this week, a faint sliver at five yearsa face: someone on record holding the claim — tap it for who they arefaded, dashed ring: they no longer hold it — they changed their mind

At the centre To reward the instinct that a new idea is worth having, training would have to score the smallness of the concepts a solution needs, not just whether it solved the problem. Last stated 30 Jun 2026 · 2 months ago Holds Grant Sanderson Read this korrent →