MIT AI Risk Repository · Risk Category · 17.04.00
Malicious Uses
Description
"Harms that arise from actors using the language model to intentionally cause harm"
From Ethical and social risks of harm from language models (Weidinger2021), as extracted by the MIT AI Risk Repository (CC BY 4.0).
Classification
- Domain
- 4. Malicious actors
- Subdomain
- 4.0
- Causal entity
- Human
- Intent
- Intentional
- Timing
- Post-deployment
How other frameworks describe this risk
Other entries from Weidinger2021
- Discrimination, Exclusion and Toxicity
- Social stereotypes and unfair discrmination
- Social stereotypes and unfair discrmination
- Exclusionary norms
- Exclusionary norms
- Exclusionary norms
- Exclusionary norms
- Toxic language
- Lower performance for some languages and social groups
- Lower performance for some languages and social groups
- Information Hazards
- Compromising privacy by leaking private infiormation