•  8
    The relation between artificial intelligence capability and catastrophic risk is often represented by a one-dimensional shorthand in which risk increases with capability. That shorthand captures an important amplification mechanism but suppresses a second variable: the capacity of an agent to model the downstream consequences of its conduct, represent uncertainty about those models, detect conflicts between proxy objectives and wider constraints, revise a proposed course of action, and accept co…Read more