Person
David Krueger
Cited in two posts, AI Alignment with Changing and Influenceable… and Characterizing Manipulation from AI Systems, since September 2026
In the claim ledger
2 promoted claims about them. Assessments are the model’s knowledge, not verification.
Having an incentive is not the same as pursuing it.
consistent · Characterizing Manipulation from AI Systems
Company-level iteration over supervised systems can itself act as a long-horizon optimiser toward manipulation.
plausible · Characterizing Manipulation from AI Systems