MIT AI Risk Repository · Risk Sub-Category · 22.01.02

Unleashing AI Agents

Category: Malicious Use (Intentional)

Description

"people could build AIs that pursue dangerous goals’"

From An Overview of Catastrophic AI Risks (Hendrycks2023), as extracted by the MIT AI Risk Repository (CC BY 4.0).

Classification

Causal entity
Human

Subdomain definition: Using AI systems to develop cyber weapons (e.g., coding cheaper, more effective malware), develop new or enhance existing weapons (e.g., Lethal Autonomous Weapons or CBRNE), or use weapons to cause mass harm.

Real-world incidents in this subdomain

Browse all incidents in this subdomain

How other frameworks describe this risk

Other entries from Hendrycks2023