MIT AI Risk Repository · Risk Sub-Category · 53.02.02

Acquisition of a goal to harm society

Category: Dangerous capabilities in AI systems

Description

"cases of AI systems being given the outright goal of harming humanity (ChaosGPT);"

From Advancing AI Governance: A Literature Review of Problems, Options, and Proposals (Maas2023), as extracted by the MIT AI Risk Repository (CC BY 4.0).

Classification

Causal entity
Human

Subdomain definition: Using AI systems to develop cyber weapons (e.g., coding cheaper, more effective malware), develop new or enhance existing weapons (e.g., Lethal Autonomous Weapons or CBRNE), or use weapons to cause mass harm.

Real-world incidents in this subdomain

Browse all incidents in this subdomain

How other frameworks describe this risk

Other entries from Maas2023