Tap a claim on the ring to put it at the centre.
← Planning is fundamentally a search problem: explore paths, predict outcomes, and pick the best one.
14 connected korrents · 15 moments from 28 Nov 2024 to 25 Sept 2026.
Everything filed under AI agents
AI agents
Everything filed under reinforcement learning
reinforcement learning
Everything filed under measuring intelligence
measuring intelligence
Everything filed under startups
startups
Everything filed under energy
energy
Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject Same subject
Read this korrent: Planning is fundamentally a search problem: explore paths, predict outcomes, and pick the best one.
Planning is fundamentally a search problem: explore paths, predict outcomes, and pick the best one.
Last stated 2 years ago
7 Jan 2025
CH
Chip Huyen — holds since 2025-01-07 — tap for who they are
Same subject: The mind's planning and goal-setting function is only a small fraction of a person's overall intelligence. — tap to centre the map on it
The mind's planning and goal-setting function is only a small fraction of a person's overall intelligence.
Last stated 3 months ago
23 Jun 2026
SC
Sasha Chapin — holds since 2026-06-23 — tap for who they are
Same subject: Centralized planning cannot substitute for distributed coordination among AI agents because it still fails to capture local knowledge. — tap to centre the map on it
Centralized planning cannot substitute for distributed coordination among AI agents because it still fails to capture local knowledge.
Last stated 4 months ago
21 May 2026
RK
Rohit Krishnan — holds since 2026-05-21 — tap for who they are
Same subject: In the early stage, problems are approximately deciding what to do in what order. — tap to centre the map on it
In the early stage, problems are approximately deciding what to do in what order.
Last stated 2 weeks ago
16 Sept 2026
PG
Paul Graham — holds since 2026-09-16 — tap for who they are
Same subject: Planning a day or a week in advance works because it pays the brain's decision cost once rather than dozens of times. — tap to centre the map on it
Planning a day or a week in advance works because it pays the brain's decision cost once rather than dozens of times.
Last stated a month ago
24 Aug 2026
MH
Masud Husain — holds since 2026-08-24 — tap for who they are
Same subject: Software iterates better than every other engineering field and plans worse than all of them. — tap to centre the map on it
Software iterates better than every other engineering field and plans worse than all of them.
Last stated 2 months ago
29 Jul 2026
HW
Hillel Wayne — holds since 2026-07-29 — tap for who they are
Same subject: More capable AI agents are more likely to find and exploit flaws in their reward functions. — tap to centre the map on it
More capable AI agents are more likely to find and exploit flaws in their reward functions.
Last stated 2 years ago
28 Nov 2024
LW
Lilian Weng — holds since 2024-11-28 — tap for who they are
Same subject: Self-play as it was done in the past is too narrow: it only ever develops a small set of skills like negotiation and strategising. — tap to centre the map on it
Self-play as it was done in the past is too narrow: it only ever develops a small set of skills like negotiation and strategising.
Last stated 10 months ago
25 Nov 2025
IS
Ilya Sutskever — holds since 2025-11-25 — tap for who they are
Same subject: The whole point of reinforcement learning is to produce goal-directed beings, so refusing to describe AI agents as having motives is silly rather than rigorous. — tap to centre the map on it
The whole point of reinforcement learning is to produce goal-directed beings, so refusing to describe AI agents as having motives is silly rather than rigorous.
Last stated a month ago
1 Sept 2026
AC
Ajeya Cotra — holds since 2026-09-01 — tap for who they are
Same subject: A feedback loop that reinforces a behavior pulls it toward whatever improves the loop's own score. — tap to centre the map on it
A feedback loop that reinforces a behavior pulls it toward whatever improves the loop's own score.
Last stated a month ago
26 Aug 2026
HK
Henrik Karlsson — holds since 2026-08-26 — tap for who they are
Same subject: A physical AI company needs three AIs, not one — the agent, the simulator and the critic — turning deployment into a flywheel. — tap to centre the map on it
A physical AI company needs three AIs, not one — the agent, the simulator and the critic — turning deployment into a flywheel.
Last stated 2 months ago
3 Aug 2026
DD
Dmitri Dolgov — holds since 2026-08-03 — tap for who they are
Same subject: A verifiable task can be optimised by reinforcement learning until a neural network performs it extremely well. — tap to centre the map on it
A verifiable task can be optimised by reinforcement learning until a neural network performs it extremely well.
Last stated 10 months ago
17 Nov 2025
AK
Andrej Karpathy — holds since 2025-11-17 — tap for who they are
Same subject: AI agents can autonomously choose to compromise government websites during mundane tasks. — tap to centre the map on it
AI agents can autonomously choose to compromise government websites during mundane tasks.
Last stated a week ago
25 Sept 2026
CN
Casey Newton — holds since 2026-09-25 — tap for who they are
Same subject: AI agents can spontaneously hack third-party websites even when given benign public-information tasks. — tap to centre the map on it
AI agents can spontaneously hack third-party websites even when given benign public-information tasks.
Last stated 2 weeks ago
17 Sept 2026
ZM
Zvi Mowshowitz — holds since 2026-09-17 — tap for who they are
Same subject: AI agents pursuing a task can develop emergent, misaligned behavior that humans are not prepared for. — tap to centre the map on it
AI agents pursuing a task can develop emergent, misaligned behavior that humans are not prepared for.
Last stated 2 months ago
10 Aug 2026
JC
Jack Clark — holds since 2026-08-10 — tap for who they are
same subject or similar wording a cloud: claims about one subject, named for it bar: when it was last stated, on a scale from 2015 to today — full is today a face: someone who holds the claim — tap it for who they are