Aligned With Whom?
What if an AI made a decision that most people disagreed with, and their lives became better because of it? Would we call that AI aligned? “Aligned with human interests” sounds straightforward until what we want and what benefits us point in different directions. Imagine a city banning cars from its center despite widespread opposition. Suppose the streets become safer, businesses thrive, and even the opponents later welcome the change. Does the better outcome justify overriding them?