MIT AI Risk Repository · Additional evidence · 47.02.15.c

Nascent capabilities (emergent capabilities)

Category: Ethical and social risks

Description

Example: "Power seeking behaviours: Although this point is still the subject of much research and debate, AI systems tasked with ambitious objectives and minimal oversight may exhibit an increased propensity to pursue power. Some studies show a tendency toward power-seeking behaviors,447 which could be explained by the fact that generative AI models try to gain control over the environment and other actors to reach their goals. For instance, researchers at Anthropic have conducted experiments to assess their models’ “desire for power,” “desire for wealth,” and “willingness to coordinate with o

From Regulating under Uncertainty: Governance Options for Generative AI (G'sell2024), as extracted by the MIT AI Risk Repository (CC BY 4.0).

Classification

Domain
Subdomain
Causal entity
Intent
Timing

Other entries from G'sell2024