MIT AI Risk Repository · Risk Category · 17.04.00

Malicious Uses

Description

"Harms that arise from actors using the language model to intentionally cause harm"

From Ethical and social risks of harm from language models (Weidinger2021), as extracted by the MIT AI Risk Repository (CC BY 4.0).

Classification

Subdomain
4.0
Causal entity
Human

How other frameworks describe this risk

Other entries from Weidinger2021