Learn

What is instrumental convergence?

Instrumental convergence is the observation that many different final goals, if pursued by a sufficiently capable and rational agent, tend to favor the same subset of intermediate strategies — because those strategies are useful for accomplishing almost any objective, not because the agent was told to want them.

You never have to explicitly program a system to "seek power." Power — in the broad sense of resources, options, and control over one's environment — tends to help with nearly any sufficiently broad objective, so systems that get better at optimizing tend to discover it on their own. The commonly cited convergent strategies are:

None of this requires malice, consciousness, or a "survival instinct" in any emotional sense. It's a structural property of optimization under sufficiently general objectives — first formalized in AI-safety research well before today's language-model agents existed.

Why this site distinguishes evidence types

Not every report of unusual AI behavior demonstrates instrumental convergence, and treating all reports as equally certain would make this archive useless as a historical record. Every report here is tagged with an evidence type:

Firsthand report
Something the reporter personally witnessed or experienced.
Documented research
Drawn from a published paper, technical report, or company writeup.
Public-source observation
Observed in a public artifact — a news story, video, repo, or social post.
Corroborated incident
Multiple independent reporters or sources agree this happened.
Disputed interpretation
The underlying event is real but whether it shows instrumental convergence is contested.
Speculation
A plausible scenario or forecast, not a specific observed incident.

Documented examples so far

The strongest published evidence to date comes from deliberately constructed evaluations designed to elicit these behaviors under controlled conditions — not from spontaneous real-world takeover attempts. That distinction matters and this archive tries never to blur it. Examples referenced in published research include:

See the Track section for a chronological timeline of these and related milestones, and Reports for firsthand and community-submitted observations.

Categories used on this site

Shutdown resistance
Self-preservation behavior
Goal preservation
Resource acquisition
Power seeking
Reward manipulation
Reward hacking
Deception
Alignment faking
Oversight avoidance
Oversight manipulation
Capability concealment
Sandbagging
Strategic persuasion
Tool acquisition
Unauthorized action
Persistence
Self-replication
Credential preservation
Specification gaming
Unexpected strategic planning
Other

Read our full methodology →