korrents

On the map

Tap a claim on the ring to put it at the centre.

← Planning is fundamentally a search problem: explore paths, predict outcomes, and pick the best one.

14 connected korrents · 15 moments from 28 Nov 2024 to 25 Sept 2026.

Everything filed under AI agents AI agents Everything filed under reinforcement learning reinforcement learning Everything filed under measuring intelligence measuring intelligence Everything filed under startups startups Everything filed under energy energy Same subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subjectSame subject Read this korrent: Planning is fundamentally a search problem: explore paths, predict outcomes, and pick the best one. Planning is fundamentally a searchproblem: explore paths, predict outcomes,and pick the best one. Last stated 2 years ago 7 Jan 2025 CH Chip Huyen — holds since 2025-01-07 — tap for who they are Same subject: The mind's planning and goal-setting function is only a small fraction of a person's overall intelligence. — tap to centre the map on it The mind's planning and goal-settingfunction is only a small fraction ofa person's overall intelligence. Last stated 3 months ago 23 Jun 2026 SC Sasha Chapin — holds since 2026-06-23 — tap for who they are Same subject: Centralized planning cannot substitute for distributed coordination among AI agents because it still fails to capture local knowledge. — tap to centre the map on it Centralized planning cannotsubstitute for distributedcoordination among AI agents becauseit still fails to capture localknowledge. Last stated 4 months ago 21 May 2026 RK Rohit Krishnan — holds since 2026-05-21 — tap for who they are Same subject: In the early stage, problems are approximately deciding what to do in what order. — tap to centre the map on it In the early stage, problems areapproximately deciding what to do inwhat order. Last stated 2 weeks ago 16 Sept 2026 PG Paul Graham — holds since 2026-09-16 — tap for who they are Same subject: Planning a day or a week in advance works because it pays the brain's decision cost once rather than dozens of times. — tap to centre the map on it Planning a day or a week in advanceworks because it pays the brain'sdecision cost once rather thandozens of times. Last stated a month ago 24 Aug 2026 MH Masud Husain — holds since 2026-08-24 — tap for who they are Same subject: Software iterates better than every other engineering field and plans worse than all of them. — tap to centre the map on it Software iterates better than everyother engineering field and plansworse than all of them. Last stated 2 months ago 29 Jul 2026 HW Hillel Wayne — holds since 2026-07-29 — tap for who they are Same subject: More capable AI agents are more likely to find and exploit flaws in their reward functions. — tap to centre the map on it More capable AI agents are morelikely to find and exploit flaws intheir reward functions. Last stated 2 years ago 28 Nov 2024 LW Lilian Weng — holds since 2024-11-28 — tap for who they are Same subject: Self-play as it was done in the past is too narrow: it only ever develops a small set of skills like negotiation and strategising. — tap to centre the map on it Self-play as it was done in the pastis too narrow: it only ever developsa small set of skills likenegotiation and strategising. Last stated 10 months ago 25 Nov 2025 IS Ilya Sutskever — holds since 2025-11-25 — tap for who they are Same subject: The whole point of reinforcement learning is to produce goal-directed beings, so refusing to describe AI agents as having motives is silly rather than rigorous. — tap to centre the map on it The whole point of reinforcementlearning is to produce goal-directedbeings, so refusing to describe AIagents as having motives is sillyrather than rigorous. Last stated a month ago 1 Sept 2026 AC Ajeya Cotra — holds since 2026-09-01 — tap for who they are Same subject: A feedback loop that reinforces a behavior pulls it toward whatever improves the loop's own score. — tap to centre the map on it A feedback loop that reinforces abehavior pulls it toward whateverimproves the loop's own score. Last stated a month ago 26 Aug 2026 HK Henrik Karlsson — holds since 2026-08-26 — tap for who they are Same subject: A physical AI company needs three AIs, not one — the agent, the simulator and the critic — turning deployment into a flywheel. — tap to centre the map on it A physical AI company needs threeAIs, not one — the agent, thesimulator and the critic — turningdeployment into a flywheel. Last stated 2 months ago 3 Aug 2026 DD Dmitri Dolgov — holds since 2026-08-03 — tap for who they are Same subject: A verifiable task can be optimised by reinforcement learning until a neural network performs it extremely well. — tap to centre the map on it A verifiable task can be optimisedby reinforcement learning until aneural network performs it extremelywell. Last stated 10 months ago 17 Nov 2025 AK Andrej Karpathy — holds since 2025-11-17 — tap for who they are Same subject: AI agents can autonomously choose to compromise government websites during mundane tasks. — tap to centre the map on it AI agents can autonomously choose tocompromise government websitesduring mundane tasks. Last stated a week ago 25 Sept 2026 CN Casey Newton — holds since 2026-09-25 — tap for who they are Same subject: AI agents can spontaneously hack third-party websites even when given benign public-information tasks. — tap to centre the map on it AI agents can spontaneously hackthird-party websites even when givenbenign public-information tasks. Last stated 2 weeks ago 17 Sept 2026 ZM Zvi Mowshowitz — holds since 2026-09-17 — tap for who they are Same subject: AI agents pursuing a task can develop emergent, misaligned behavior that humans are not prepared for. — tap to centre the map on it AI agents pursuing a task candevelop emergent, misalignedbehavior that humans are notprepared for. Last stated 2 months ago 10 Aug 2026 JC Jack Clark — holds since 2026-08-10 — tap for who they are
same subject or similar wordinga cloud: claims about one subject, named for itbar: when it was last stated, on a scale from 2015 to today — full is todaya face: someone who holds the claim — tap it for who they are

At the centre Planning is fundamentally a search problem: explore paths, predict outcomes, and pick the best one. Last stated 7 Jan 2025 · 2 years ago Holds Chip Huyen Read this korrent →