everyonedies

Archive 2026-09-24 · 185 entries · 14 chapters

Chapter 11: An Alchemy, Not a Science

Do you see alignment as all-or-nothing?

No. But “partial alignment” is still likely to be catastrophic.

One of the arguments for worrying less about superintelligence runs along the lines of: “AI will probably advance incrementally, allowing opportunities for trial-and-error improvements to keep AIs in check at every step; alignment doesn’t have to be perfect for things to go okay.” We don’t think this view holds much hope, for a few reasons:


  1. For some discussion on why you’d really need to know what you’re doing, see Intelligent (Usually) Implies Incorrigible, Deep Machinery of Steering, and It’s Hard to Get Robust Laziness. ↩︎

Source: ifanyonebuildsit.com/11/do-you-see-alignment-as-all-or-nothing