Person
Eliezer Yudkowsky
cited by eight posts · in two sources · first cited September 2026
In their words
“However, the robustness of this hope is challenged by the nearest unblocked strategy problem (Yudkowsky, 2015) : the problem that an AI which strongly optimizes for a (misaligned) goal will exploit even small loopholes in (aligned) constraints, which may lead to arbitrarily bad outcomes (Zhuang and Hadfield-Menell, 2020) .”
· · machine-resolved
“Thanks to Eliezer Yudkowsky, Rohin Shah, and Evan Hubinger for comments on the relevant scope here (which isn’t to say they endorse my choice of definition).”
· Existential Risk from Power-Seeking AI · machine-resolved