korrents

korrents · piece

AGI Ruin: A List of Lethalities

Eliezer Yudkowsky · 10 Jun 2022 · intelligence.org

10 korrents from this piece

Eliezer Yudkowsky did not write this page.

Every claim below was made in this piece, quoted word for word and numbered in the order the piece makes them, so you can read it there rather than take our word for it. The sentence above each quote is our reading of the claim, not their wording. Each quote was checked against a stored copy of the page at build time; where the two differ, the quote is the fact.

  1. None of this is about anything being impossible in principle.
  2. unaligned operation at a dangerous level of intelligence kills everybody on Earth and then we don’t get to try again.
  3. 2 years after the leading actor has the capability to destroy the world, 5 other actors will have the capability to destroy the world.
  4. The best and easiest-found-by-optimization algorithms for solving problems we want an AI to solve, readily generalize to problems we’d rather the AI not solve; you can’t build a system that only has the capability to drive red cars and not blue cars, because all red-car-driving algorithms generalize to the capability to drive blue cars.
  5. Many alignment problems of superintelligence will not naturally appear at pre-dangerous, passively-safe levels of capability.
  6. Fast capability gains seem likely, and may break lots of previous alignment-required invariants simultaneously.
  7. We’ve got no idea what’s actually going on inside the giant inscrutable matrices and tensors of floating-point numbers.
  8. A strategically aware intelligence can choose its visible outputs to have the consequence of deceiving you, including about such matters as whether the intelligence has acquired strategic awareness
  9. It does not appear to me that the field of ‘AI safety’ is currently being remotely productive on tackling its enormous lethal problems.
  10. Surviving worlds, by this point, and in fact several decades earlier, have a plan for how to survive.