MIT AI Risk Repository · Additional evidence · 47.02.15.c
Nascent capabilities (emergent capabilities)
Category: Ethical and social risks
Description
Example: "Power seeking behaviours: Although this point is still the subject of much research and debate, AI systems tasked with ambitious objectives and minimal oversight may exhibit an increased propensity to pursue power. Some studies show a tendency toward power-seeking behaviors,447 which could be explained by the fact that generative AI models try to gain control over the environment and other actors to reach their goals. For instance, researchers at Anthropic have conducted experiments to assess their models’ “desire for power,” “desire for wealth,” and “willingness to coordinate with o
From Regulating under Uncertainty: Governance Options for Generative AI (G'sell2024), as extracted by the MIT AI Risk Repository (CC BY 4.0).
Classification
- Domain
- —
- Subdomain
- —
- Causal entity
- —
- Intent
- —
- Timing
- —
Other entries from G'sell2024
- Technical and operational risks
- Technical vulnerabilities (Robustness - unexpected behaviour)
- Technical vulnerabilities (Robustness - unexpected behaviour)
- Technical vulnerabilities (Robustness - vulnerability to jailbreaking
- Technical vulnerabilities (Robustness - vulnerability to jailbreaking
- Technical vulnerabilities (The risk of misalignment)
- Technical vulnerabilities (The risk of misalignment)
- Factually incorrect content (inaccuracies and fabricated sources)
- Factually incorrect content (inaccuracies and fabricated sources)
- Opacity (the black box problem)
- Opacity (industry opacity)
- Opacity (industry opacity)