Skip to content

Person

Eliezer Yudkowsky

cited by eight posts · in two sources · first cited September 2026

In their words

  • However, the robustness of this hope is challenged by the nearest unblocked strategy problem (Yudkowsky, 2015) : the problem that an AI which strongly optimizes for a (misaligned) goal will exploit even small loopholes in (aligned) constraints, which may lead to arbitrarily bad outcomes (Zhuang and Hadfield-Menell, 2020) .

    · · machine-resolved

  • Thanks to Eliezer Yudkowsky, Rohin Shah, and Evan Hubinger for comments on the relevant scope here (which isn’t to say they endorse my choice of definition).

    · Existential Risk from Power-Seeking AI · machine-resolved