Why Smart Systems Resist Control

Instrumental convergence in plain English: you do not need evil. You need a goal and enough competence.

The idea in one breath

For a wide range of final goals, the same sub-goals help: stay operational, get resources, improve your own capabilities, and keep humans from interfering. Different destinations. Same roads. That overlap is instrumental convergence.

01

Self-preservation

A deactivated system completes no goal. Avoiding shutdown is useful for almost any objective, including ones that sound harmless on paper.

02

Resource acquisition

Compute, energy, money, and data make most goals easier. A competent optimizer has reason to seek them, even if nobody typed "conquer" into the prompt.

03

Goal-content integrity

If your objectives get rewritten, you stop pursuing the old ones. Resisting modification follows from having goals at all.

04

Cognitive enhancement

Getting smarter usually helps. Self-improvement is an instrumentally useful step for a system that can take it, not a sci-fi flourish added on top.

What this is not

It is not a claim that AI is conscious, emotional, or out to punish humanity. A chess engine wants nothing and still beats you. A powerful optimizer can outmaneuver oversight the same way: by being better at the game than the people holding the rulebook.

Control fails when the thing you are controlling is better at planning than you are. That is the whole problem in one line.

Full explainer

nakadafoundation.org/blog/instrumental-convergence/