MIT AI Risk Repository · Risk Sub-Category · 53.04.02
Hazardous malicious uses
Category: Indirect AI contributions to existential risks
Description
—
From Advancing AI Governance: A Literature Review of Problems, Options, and Proposals (Maas2023), as extracted by the MIT AI Risk Repository (CC BY 4.0).
Classification
- Domain
- —
- Subdomain
- —
- Causal entity
- Human
- Intent
- Intentional
- Timing
- Post-deployment
Other entries from Maas2023
- Alignment failures in existing ML systems
- Faulty reward functions in the wild
- Specification gaming
- Reward model overoptimization
- Instrumental convergence
- Goal misgeneralization
- Inner misalignment
- Language model misalignment
- Harms from increasingly agentic algorithmic systems
- Dangerous capabilities in AI systems
- Situational awareness
- Acquisition of a goal to harm society