Contains some AI-generated content

AI Alignment as a Thought-Terminating Cliche borretti.me

Fernando Borretti’s objection is not to alignment research but to the word doing rhetorical work. His claim is that “aligned” functions as a conversation-stopper: raise human powerlessness, democratic collapse or economic displacement in an AI-run world, and the answer is that those would be failures of alignment, which the aligned system by definition does not have.

That makes the concept unfalsifiable. It proves every utopian conclusion and refutes every dystopian one, without either being argued. His comparison is to claiming the right totalitarian dictator would solve politics, an argument nobody accepts in that form but which becomes respectable rebranded as aligned superintelligence.

The target underneath is what he calls benevolent disempowerment: a future where superintelligence solves our problems and we are safe, comfortable and no longer in charge of anything. His charge is that the framing is incoherent and motivated, and that it crowds out the political and economic questions that actually need arguing. Worth reading directly against An Alien Mind, which uses the vocabulary in earnest.

← All links