Project Sherlock

Artificial Intelligence · AI Alignment & Safety

Existential Risk Arguments

A topic within AI Alignment & Safety, itself one of 11 topics in that field and part of Artificial Intelligence.

Reading on Existential Risk Arguments

2

2 works

Paper2012

The Superintelligent Will: Motivation and Instrumental Rationality in Advanced Artificial Agents

Nick Bostrom

Argues intelligence and final goals are independent (the orthogonality thesis), and that agents with almost any final goal will converge on similar instrumental subgoals — self-preservation, resource acquisition, goal-content integrity — which is what makes a highly capable system's specific goal so consequential.

15 pageslink checked 17 Sept 2026

Other topics in AI Alignment & Safety