MIT AI Risk Repository · Risk Category · 56.19.00
Capabilities that increase the likelihood of existential risk
Description
-
From Future Risks of Frontier AI (GOS2023), as extracted by the MIT AI Risk Repository (CC BY 4.0).
Classification
Subdomain definition: AI systems that develop, access, or are provided with capabilities that increase their potential to cause mass harm through deception, weapons development and acquisition, persuasion and manipulation, political strategy, cyber-offense, AI development, situational awareness, and self-proliferation. These capabilities may cause mass harm due to malicious human actors, misaligned AI systems, or failure in the AI system.
How other frameworks describe this risk
- Safety Risks from Affordances Provided to LLM-agents
- Agentic LLMs Pose Novel Risks
- Goal-Directedness Incentivizes Undesirable Behaviors
- Capabilities that could be used to reduce human control - Cyber offence
- Capabilities that could be used to reduce human control - Autonomous replication and adaptation
- Capabilities that could be used to reduce human control - Manipulation
- Subagents
- AI Influence
Other entries from GOS2023
- Discrimination
- Inequality
- Environmental impacts
- Amplification of biases
- Harmful responses
- Lack of transparency and interpretability
- Intellectual property rights
- Providing new capabilities to a malicious actor
- Misapplication by a non-malicious actor
- Poor performance of a model used for its intended purpose, for example leading to biased decisions
- Unintended outcomes from interactions with other AI systems
- Impacts resulting from interactions with external societal, political, and economic systems